Skip to content
r/MachineLearning · Communities

Evaluating J-space entropy as an error predictor across 7 datasets on Qwen3-4B [R]

Anthropic’s Jacobian Lens work introduced a way to inspect verbalizable representations inside language models. Follow-up experiments suggested that entropy in this internal “workspace” might help identify confidently incorrect answers. I tested that hypothesis on Qwen3-4B across ~11,400 examples from seven distinct da