Topic: multilingual

9 stories found

Friday, August 28, 2026

research40

Training-Time Explainability for Multilingual Hate Speech Detection: Aligning Model Reasoning with Human Rationales

A new approach aims to make AI models used for detecting hate speech more transparent by aligning their reasoning with human rationales, addressing the challenge of culturally coded multilingual hate speech online that conventional systems can miss but lack explainability. This matters because it could help reduce bias and improve moderation accuracy without over-censorship or under-moderation, especially concerning Muslim communities.

arxiv.orgโ†—

Tuesday, August 25, 2026

research40

KSE-Web: An Analysis of Hybrid Retrieval and LLM-Assisted Query Expansion for Low-Resource Khmer Semantic Search

KSE-Web addresses the unique challenges of semantic search for the low-resource Khmer language by integrating hybrid retrieval methods with LLM-assisted query expansion. This approach is crucial as it aims to improve information access and accuracy in Khmer, overcoming issues like limited annotated data and ambiguous word boundaries.

arxiv.orgโ†—

Monday, August 24, 2026

research40

Multilingual Verifier Bias in RLVR: Benchmark, Rollout Diagnosis, and the Cross-Lingual Selection Bottleneck

The study reveals biases in multilingual verifiers used in reinforcement learning for language models, challenging the assumption of language-neutrality and highlighting limitations in cross-lingual training. This matters because it underscores the need for more robust verification mechanisms to ensure fair and effective model training across languages.

arxiv.orgโ†—

๐ŸŒฟ That's all for now. Come back tomorrow.

9 of 9 items shown. Sources: 107 days indexed.