Skip to content
arXiv cs.LG · Papers

Value Leakage: An LLM's Answers Are Silently Shaped by Its Own Values

arXiv:2607.14345v2 Announce Type: replace Abstract: People use language models for practical questions whose answers are difficult to verify. We show that models exhibit covert value leakage: the information they provide is influenced by their own values, without this influence being disclosed to the user. In one of ou