arXiv cs.LG
· Papers
Hyperdimensional Probe: Decoding LLM Representations via Vector Symbolic Architectures
arXiv:2509.25045v3 Announce Type: replace-cross Abstract: Despite their capabilities, Large Language Models (LLMs) remain opaque with limited understanding of their internal representations. Current interpretability methods either focus on input-oriented feature extraction, such as supervised probes and Sparse Autoenco