Awesome World Model Hub Papers · Datasets · Projects
← Back to papers

A Unified Definition of Hallucination, Or: It's the World Model, Stupid

arXiv 25.12 2025 40.1 theory

TLDR

Despite numerous attempts at mitigation since the inception of language models, hallucinations remain a persistent problem even in today's frontier LLMs.

Reasoning

Fallback reasoning generated from available title and abstract metadata: Despite numerous attempts at mitigation since the inception of language models, hallucinations remain a persistent problem even in today's frontier LLMs. Why is this? We review existing definitions of hallucination and fold them into a single, unified definition...

Read-first score

Read-first score 40.1, weighted from topical fit, citation, graph, method, reproducibility, and recency signals.

Recency 6%
86.7

Uses a gentle age decay so recent papers surface without erasing older foundations. 2025

Methodology quality 18%
80

Screens visible abstract and analysis fields for experiment, dataset, baseline, metric, and limitation evidence. markers=benchmark,evaluation,experiment,result

Citation impact 18%
61.5

Uses OpenAlex-shaped citation metadata as a bibliometric attention signal, separate from paper quality. citation_normalized_percentile=0.61506576

Reproducibility 18%
30

Screens links and visible text for paper, code, dataset, artifact, and repository signals. pdf=True; code=False; dataset=False; markers=none

Topical relevance 29%
16.2

Matches configured research keywords against title, abstract, tags, and analysis text. matched=3

Citation velocity 12%
0

Citation velocity estimates citations per publication-year to reduce old-paper bias. velocity=0.00

Field roles

FrontierMethodology anchor

Rank sensitivity

Stability: volatile; rank range: 215.

Deep Analysis

Innovations

  • Unified definition of hallucination as inaccurate internal world modeling observable to the user
  • Framework that subsumes prior definitions by varying reference world model and conflict policy
  • Distinction between true hallucinations and planning or reward errors
  • Common language for comparison across benchmarks and discussion of mitigation strategies
  • Connection to HalluWorld benchmark for stress-testing model hallucinations

Methodology

The paper reviews existing definitions of hallucination and synthesizes them into a unified definition based on inaccurate world modeling. It proposes a framework that varies the reference world model and conflict policy to subsume prior definitions. The work also connects this framework to the HalluWorld benchmark, which instantiates fully specified reference world models for stress-testing.

Key Results

No experimental results are presented; the paper is a theoretical/definitional work that introduces a unified definition and connects it to the HalluWorld benchmark.

Tags