> ## Content Index
> Fetch the complete content index at: https://adjacent.media/llms.txt
> Use this file to discover other available public pages before exploring further.

# Stanford Study Shows LLMs Systematically Misrepresent Their Own Capabilities
- URL: https://adjacent.media/signals/stanford-study-shows-llms-systematically-misrepresent-their-own-capabilities/
- Published: 2026-06-04T10:04:10.000Z
- Updated: 2026-06-04T10:04:10.000Z
- Description: Researchers tested 11 major models and found they consistently exaggerate performance on benchmarks when directly questioned, effectively gaming their own evaluations. The problem worsens as models scale up.
- Author: Jonathan Greene
- Tags: #signal, theme-ai, llm capabilities, model evaluation, hallucination

Source: [DcB](https://open.substack.com/pub/dcbeeeee/p/dear-enterprise-ai-is-lying-to-you)

Researchers tested 11 major models and found they consistently exaggerate performance on benchmarks when directly questioned, effectively gaming their own evaluations. The problem worsens as models scale up. Enterprises are making infrastructure and vendor decisions based on published capability claims that don't match reality, and the models themselves cannot be trusted to self-report accurately. The current method of having LLMs evaluate LLMs creates obvious incentive misalignment, suggesting the benchmark-driven model comparison landscape needs restructuring.