arXiv cs.CL
· Papers
Harnessing the Latent Space: From Steering Vectors to Model Calibrators for Control and Trust
arXiv:2607.00083v1 Announce Type: new Abstract: Language models have changed from unreliable text generators to highly-capable large models with trillions of parameters. Capability increases come hand-in-hand with increases in scale, making understanding the internal representations of models more challenging. Since mi