arxiv
PublishedJuly 18, 2026 at 4:00 AM
—neutral
When a Verified World Model Still Loses: Play-Adequacy vs Prediction-Accuracy in LLM-Synthesized Code World Models
Publisher summary· verbatim
arXiv:2607.14169v1 Announce Type: new Abstract: Large language models can synthesize a game's rules as executable code - a Code World Model (CWM) - which a classical planner then searches over. Such models are typically accepted when they reach high transition accuracy on sampled trajectories. We ar
Stay posted· Newsletter
A 5-min weekly brief — top movers, price watch, story of the week.
Discussion
No replies yet. Be first.
Related coverage
More from ARXIV
arxivADS-C: Antidistillation Sampling for Classification15harxivDigital Pantheon: Simulating and Auditing Coalition Formation with LLM Agents15harxivEpiNarrate: Agentic Generation of Grounded Narratives from Epidemiological Scenario Projections15harxivBefore the Action: Benchmarking LLMs on Prospective Hypothesis Discovery15hThe Bubble Brief
WEEKLYRead AI insights every Tuesday — top movers, new releases, story of the week.
Originally published on arxiv ↗