Skip to content
arXiv cs.AI · Papers

Evaluating Implicit Biases in LLM Reasoning through Logic Grid Puzzles

arXiv:2511.06160v2 Announce Type: replace Abstract: While recent safety guardrails effectively suppress overtly biased outputs, subtler forms of social bias emerge during complex logical reasoning tasks that evade current evaluation benchmarks. To fill this gap, we introduce a new evaluation framework, PRIME (Puzzle Re