Get Started
Research questionHow can we evaluate large language models’ understanding of Bangla idioms in a low-resource setting?Idioms carry culturally specific meanings that cannot always be inferred literally, while limited Bangla resources make model behavior difficult to measure. Assessment must distinguish paraphrasing, locating idiom spans, and identifying idiom meanings.
AI
Evaluation & Benchmarks
Natural Language Processing
Latest papersRecent research connected to this question, newest first.To What Extent Do Large Language Models Understand Bangla Idioms?The source provides a large-scale Bangla idiom benchmark and a synthetic multiple-choice dataset for meaning identification, along with evaluations of recent LLMs using zero-shot and few-shot prompts. Results vary by task, with no single model consistently leading across all three.research paper · Sep 3, 2026
Related questions
Can large language models reliably perform Arabic morphosyntactic tagging and dependency parsing despite morphological and orthographic ambiguity?How should low-resource LLM fine-tuning use task-level language priors with ambiguous or incomplete data?When do natural-language rules outperform examples for LLM in-context learning?How can we measure and control an LLM’s reliance on token-frequency priors when context is sparse?
Home
Topics
Search
Library