NeFut Logo NeFut
Admin Login

[CS.AI] VakyArth: Evaluating Pragmatic Competence in LLMs across Indic Languages

Published at: 2026-09-04 22:00 Last updated: 2026-09-05 12:23
#Machine Learning #LLM #Artificial Intelligence

VakyArth is the first pragmatic benchmark for Indic languages, covering Hindi, Punjabi, Tamil, and Malayalam. The benchmark diagnoses five pragmatic phenomena—deixis, speech acts, implicature, social pragmatics, and coherence—using multiple‑choice questions (MCQ), natural language inference (NLI), and translation tasks, all authored by native speakers.\ \ Experiments on multilingual large language models (LLMs) of various families and sizes reveal consistent failures when models must interpret meanings grounded in Indic linguistic and cultural conventions. Key observations include:\

Original Source: https://arxiv.org/abs/2609.01788

[h] Back to Home