arXiv AI By Fatima Jahara, Mark Dredze, Sharon Levy

Evaluating Implicit Biases in LLM Reasoning through Logic Grid Puzzles

Read the original on arXiv AI →

arXiv:2511. 06160v2 Announce Type: replace Abstract: While recent safety guardrails effectively suppress overtly biased outputs, subtler forms of social bias emerge during complex logical reasoning tasks that evade current evaluation benchmarks.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.