arXiv · 2509.08960
BRoverbs -- Measuring how much LLMs understand Portuguese proverbs
Abstract
Large Language Models (LLMs) exhibit significant performance variations depending on the linguistic and cultural context in which they are applied. This disparity signals the necessity of mature evaluation frameworks that can assess their capabilities in specific regional settings. In the case of Portuguese, existing evaluations remain limited, often relying on translated datasets that may not fully capture linguistic nuances or cultural references. Meanwhile, native Portuguese-language datasets predominantly focus on structured national exams or sentiment analysis of social media interactions, leaving gaps in evaluating broader linguistic understanding. To address this limitation, we introduce BRoverbs, a dataset specifically designed to assess LLM performance through Brazilian proverbs. Proverbs serve as a rich linguistic resource, encapsulating cultural wisdom, figurative expressions, and complex syntactic structures that challenge the model comprehension of regional expressions. BRoverbs aims to provide a new evaluation tool for Portuguese-language LLMs, contributing to advancing regionally informed benchmarking. The benchmark is available at https://huggingface.co/datasets/Tropic-AI/BRoverbs.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Thales Sales Almeida, Giovana Kerche Bonás, João Guilherme Alves Santos. 2025-09-10. BRoverbs -- Measuring how much LLMs understand Portuguese proverbs. https://arxiv.org/abs/2509.08960
Cite the original work for its findings. Save a collection to share your selection of sources.