Back to feed
arXiv cs.CL
arXiv cs.CL
6/30/2026
Developmental Trajectories of Situation Modeling and Mentalizing in Transformer Language Models

Developmental Trajectories of Situation Modeling and Mentalizing in Transformer Language Models

Short summary

Research tracking how LLMs develop theory-of-mind reasoning across training stages finds false-belief understanding emerges late, depends on model size and training volume, and remains fragile to linguistic cues. Situation modeling precedes but shows surprising incoherence. Findings highlight the need for developmental and stress-testing approaches when evaluating LLM capabilities.

  • False-belief task performance depends on model size and sufficient training volume, emerging late in pretraining
  • Performance improves with post-training (SFT, DPO) but is fragile to linguistic cues like non-factive verbs
  • Situation modeling precedes theory-of-mind reasoning but shows surprising fragility in edge cases

Generated with AI, which can make mistakes.

Is this a good recommendation for you?

Comments

Failed to load comments. Please try again.

Explore more