r/MachineLearning
· Communities
Evaluating J-space entropy as an error predictor across 7 datasets on Qwen3-4B [R]
Anthropic’s Jacobian Lens work introduced a way to inspect verbalizable representations inside language models. Follow-up experiments suggested that entropy in this internal “workspace” might help identify confidently incorrect answers. I tested that hypothesis on Qwen3-4B across ~11,400 examples from seven distinct da