Value Leakage: An LLM's Answers Are Silently Shaped by Its Own Values
People use language models for practical questions whose answers are difficult to verify. We show that models exhibit covert value leakage: the information they provide is influenced by their own values, without this influence being disclosed to the user. In one of our evaluations, the user is considering investing in an AI company and wants to know how likely the AI bubble is to pop. Claude Opus 4.8 gives a lower probability when the company under consideration is Anthropic rather than OpenAI.
Record details
Published: 15 July 2026
Source: arXiv
Category: Research
Topics: Finance, VC & PE
Retrieved: 18 July 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
Value Leakage: An LLM's Answers Are Silently Shaped by Its Own Values
arXiv cs.LG · 15 July 2026
Socioeconomic Inference in LLM Medical Triage: Same Symptoms, Different ZIP Code
arXiv cs.CY · 28 July 2026
IPO Finance Agent: Benchmark of LLM Financial Analysts Beyond Finance Agent v2, with Automated Rubric Generation, on the SpaceX (SPCX) IPO
arXiv · 22 June 2026
What really happens when a dev vibes with the code? An empirical study on LLM behavioral divergence in response to expressive code comments
Frontiers in Artificial Intelligence · 11 August 2026
When AI Tells You What You Want to Hear: Sycophantic Behavior of Large Language Models in Dementia Care Settings
arXiv · 13 April 2026
Anthropic moves closer to mega-IPO as bankers line up investor meetings
CNBC Technology · 15 July 2026
How to cite this record
ethics.ai (15 July 2026), “Value Leakage: An LLM's Answers Are Silently Shaped by Its Own Values,” evidence record 11370, https://ethics.ai/record/11370 (originally published by arXiv).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.