arxiv
PublishedSeptember 11, 2026 at 4:00 AM
—neutral
'Ghaib in Translation' aka Unseen Harm: Measuring Cross-Script Safety Inconsistency with 'Missed-in-Urdu' Scores in LLM Hate Speech Detection
Publisher summary· verbatim
arXiv:2608.24191v2 Announce Type: replace-cross Abstract: Urdu, the world's tenth most spoken language with 246 million speakers, remains almost entirely absent from mainstream LLM safety evaluation and nine years of WOAH proceedings. To investigate whether this absence has measurable consequences f
Stay posted· Newsletter
A 5-min weekly brief — top movers, price watch, story of the week.
Discussion
No replies yet. Be first.
Related coverage
More from ARXIV
arxivBringing Value Models Back: Generative Critics for Value Modeling in LLM Reinforcement Learning3harxivSubagents vs Agent Skills: Executing Reusable Knowledge for Long-Horizon Agentic Tasks3harxivDistribution-Consistent Inference for Dynamic Sparse Mixture-of-Experts3harxivIn RAG We Trust? Measuring Robustness of Retrieval-Augmented Generation Under Document Poisoning3hThe Bubble Brief
WEEKLYRead AI insights every Tuesday — top movers, new releases, story of the week.
Originally published on arxiv ↗