AI Plays Pokémon: Anthropic's Claude 3.7 Sonnet Beats Gym Leaders

Anthropic's AI, Claude 3.7 Sonnet, tested on Pokémon Red, shows "extended thinking" by defeating gym leaders, a leap from previous models.
Matilda
AI Plays Pokémon: Anthropic's Claude 3.7 Sonnet Beats Gym Leaders
Anthropic used Pokémon to benchmark its newest AI model. Yes, really.                                                      Image Credits:Pokémon In a blog post published Monday, Anthropic said that it tested its latest model, Claude 3.7 Sonnet, on the Game Boy classic Pokémon Red. The company equipped the model with basic memory, screen pixel input, and function calls to press buttons and navigate around the screen, allowing it to play Pokémon continuously. A unique feature of Claude 3.7 Sonnet is its ability to engage in “extended thinking.” Like OpenAI’s o3-mini and DeepSeek’s R1, Claude 3.7 Sonnet can “reason” through challenging problems by applying more computing — and taking more time. That came in handy in Pokémon Red, apparently. Compared to a previous version of Claude, Claude 3.0 Sonnet, which failed to leave the house in Pallet Town where the story begins, Claude 3.7 Sonnet successfully battled three Pokémon gym leaders and won their badges. Now, it’s not clear how much computing…