Skip to content
AI Atlas
PaperActive

Multilingual in Name Only? Cultural and Linguistic Weaknesses of LLMs in Urdu

arxiv.org/abs/2609.10758

Updated 23 min ago · first seen 11 Sept 2026

paper_01M294FQ2NKDVQN2SXR2QHE8GM

Published
11 Sept 2026
T1 · 23 min ago
arXiv
2609.10758
T1 · 23 min ago
Category
cs.CL
T1 · 23 min ago

Abstract

Multilingual large language models (LLMs) are increasingly used for open-ended text generation, yet their behaviour in low-resource languages remains poorly understood. In this work, we question how correct and reliable is the generation of multilingual LLMs when used for the task of story generation. We consider Urdu language as a representative low-resource language. We generate Urdu-Stories, a corpus of 93 stories generated using three contemporary LLMs (GPT-5.1, Qwen-3-Max, DeepSeek-3.1). We manually annotate the errors present in them under a nine-label linguistic, semantic, and cultural taxonomy. Our notable findings suggest that LLMs often make basic errors of grammar and semantics. The stories lack coherence, have unnatural repetition and show pervasive cultural shallowness. We further show using few-shot prompting that the cultural and context errors largely remain unresolved. Our findings highlight the limitations of current LLMs as a reliable source of content generation and information retrieval for low-resource languages.

Authors 4

Farah Adeeba, Abdul Rafae Khan, Rajesh Bhatt, Hassan Sajjad

Specification

Official page

Source:arXiv (Atom API + RSS)T1observed 23 min agohigh

Arxiv announce type
new

Source:arXiv (Atom API + RSS)T1observed 23 min agohigh

arXiv id
2609.10758

Source:arXiv (Atom API + RSS)T1observed 23 min agohigh

Categories
cs.CL, cs.AI, cs.LG

Source:arXiv (Atom API + RSS)T1observed 23 min agohigh

PDF

Source:arXiv (Atom API + RSS)T1observed 23 min agohigh

Primary category
cs.CL

Source:arXiv (Atom API + RSS)T1observed 23 min agohigh

Published
11 Sept 2026

Source:arXiv (Atom API + RSS)T1observed 23 min agohigh

Each value shows its source, tier and observation time. Conflicting claims are kept side by side and flagged — never averaged. How AI Atlas records facts →

Provenance

Attributed facts

9

Source tiers

T19

Freshest observation

23 min ago

Conflicts

None