ICML 2025 Workshop on Assessing World Models7/14/2025·Read paper
Abstract
Do LLMs have a consistent world model that is reflected in their responses? We study whether the behavior of gpt-4o reflects an underlying world model by measuring the consistency of its mistakes across different prompts and prompting strategies. We find that gpt-4o makes consistent mistakes regardless of the exact prompt phrasing or prompt language. However, substantially different prompts that rely on the same underlying information often yield inconsistent results, suggesting that gpt-4o's responses may not reflect a single universal world model.
Make this part of your paper trail
Save this paper to a shelf, write a review, and keep your own notes.