Get the latest tech news

Insights into Claude Opus 4.5 from Pokémon


You may be surprised to learn that ClaudePlaysPokemon is still running today, and that Claude still hasn't beaten Pokémon Red, more than half a year after Google proudly announced that Gemini 2.5 Pro beat Pokémon Blue. Indeed, since then, Google and OpenAI models have gone on to beat the longer and more complex Pokémon Crystal, yet Claude has made no real progress on Red since Claude 3.7 Sonnet![1] This is because ClaudePlaysPokemon is a purer test of LLM ability, thanks to its consistently simple agent harness[2] and the relatively hands-off approach of its creator, David Hershey of Anthropic.[3] When Claudes repeatedly hit brick walls in the form of the Team Rocket Hideout and Erika's Gym for months on end, nothing substantial was done to give Claude a leg up.

None

Get the Android app

Or read this on Hacker News

Read more on:

Photo of Pokémon

Pokémon

Photo of insights

insights

Photo of Claude Opus

Claude Opus

Related news:

News photo

Behind the scenes with Google's Gemini team - 3 insights that surprised me the most

News photo

Claude Opus 4.5

News photo

Our favorite 2025 advent calendars you can still get now: Top picks from Lego, Pokémon, Funko Pop and more