← Back to all news
arXiv AI September 1, 2026 By Minkyu Kim, Juhwan Choi, YoungBin Kim

Not Safe for All: Auditing the Dialect Penalty in Text-to-Image Safety Pipelines

Read the original on arXiv AI →

The Flow has not summarised this story yet — read it at arXiv AI.

  • diffusion
  • benchmarks
  • safety

One email a morning, machine-written

One email a day, machine-written, one click to leave. We never share your address.

Related stories

arXiv Computation and Language
Sep 1

ArabicDialectSafety: A Dialect-Aware Benchmark for Arabic Content Safety Classification

arXiv:2608.01291v2 Announce Type: replace Abstract: We present ArabicDialectSafety, a human-curated Arabic safety dataset of 25,071 prompts covering six Arabic varieties: Modern Standard Arabic, Syri...

By Wajdi Zaghouani, Md. Rafiul Biswas, Kholoud Khalil Aldous, Mabrouka Bessghaier
llmsfine-tuningbenchmarkssafety
More like this →
arXiv AI
Jun 30

The Heterogeneous Safety Impacts of Benign Multilingual Fine-Tuning

arXiv:2606. 28843v1 Announce Type: cross Abstract: Fine-tuning a large language model is a ubiquitous method for enhancing its capability on a specific downstream task.

By Will Hawkins, Kaivalya Rawal, Jonathan Rystr{\o}m, Stratis Tsirtsis, Zihao Fu, Greta Warren, Ryan Brown, Eoin Delaney, Sandra Wachter, Brent Mittelstadt, Chris Russell
llmsfine-tuningsafety
More like this →
arXiv Machine Learning
Jun 9

What the Eyes See, the LLMs Miss: Exploiting Human Perception for Adversarial Text Attacks

arXiv:2606. 09700v1 Announce Type: cross Abstract: Large language model (LLM)-powered content moderation systems have become a critical defense against harmful online content.

By Qin Yang, Lu Malloy, Joshua Lee, Xiaohan Chang, Meisam Mohammady, Doowon Kim, Yuan Hong
llmsbenchmarkssafety
More like this →
arXiv AI
Jun 9

MLingualFC: Evaluating Jailbreak Vulnerabilities in Multilingual Vision-Language Models

arXiv:2606. 07706v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have demonstrated strong performance across multimodal tasks, yet their safety robustness remains an open challenge.

By Rishabh Makwana, Mamta, Deeksha Varshney, Oana Cocarascu
llmsmultimodalbenchmarkssafety
More like this →
arXiv Machine Learning
Jul 3

HaloGuard 1.0: An Open Weights Constitutional Classifier for Multilingual AI Safety

arXiv:2607. 02079v1 Announce Type: cross Abstract: We present HaloGuard 1.

By Navaneeth Sangameswaran, Preetham S, Ashmiya Lenin
agentsbenchmarkssafety
More like this →
arXiv AI
Jul 20

Conditional Reliability of Toxicity Signals for Multilingual and Code-Mixed Abuse Detection

arXiv:2607. 15861v1 Announce Type: cross Abstract: Moderation systems increasingly rely on external toxicity tools, but those tools are unreliable under code-mixing, transliteration, slang, and language mismatch.

By Indraveni Chebolu, Rohan Singh, Arnab Mallick, Harmesh Rana
llmssafety
More like this →
About Pricing API Newsletter Sources Privacy Terms Refunds Accessibility Provider info Contact RSS

The Flow links to publishers and never republishes their articles. Summaries are machine-generated.

v1.1.0 · 5f852ea