Skip to content
arXiv cs.LG · Papers

Hyperdimensional Probe: Decoding LLM Representations via Vector Symbolic Architectures

arXiv:2509.25045v3 Announce Type: replace-cross Abstract: Despite their capabilities, Large Language Models (LLMs) remain opaque with limited understanding of their internal representations. Current interpretability methods either focus on input-oriented feature extraction, such as supervised probes and Sparse Autoenco