| 1. Claude Opus 5 (High) |
| 2. Claude Fable 5 (High) |
| 3. Claude Opus 5 (Max) |
| 4. GPT 5.6 Sol (xHigh) |
| 5. Kimi K3 (Max) |
The BriefAnthropic published a research note on Monday saying an unreleased version of Claude was asked to take a real run at the Riemann hypothesis, one of the oldest open problems in mathematics. It did not solve it. On the way it improved a related known result, raising the share of the zeta function's zeros proven to sit where the hypothesis says they should from 41.6 percent to 67.2 percent, checked by two of Anthropic's own mathematicians and by two outside reviewers, with a Lean formalization alongside. The method is the part worth your attention. Two Claude Code sessions, around 60 subagents, 31 million output tokens, and roughly 650 ideas that led nowhere before one landed. Anthropic says plainly that these techniques will not get anyone to a proof, and nobody outside the company can rerun the job, so treat the math as solid and the capability claim as unfinished.
Level UpClaude Code's auto mode becomes the default on Friday for Pro, Max and Team plans. It stops asking permission at every step and only pauses for actions it judges irreversible, destructive, or pointed outside your environment. Spend ten minutes on your deny rules before then. Anthropic's own testing found people approved 97 percent of the prompts they were shown, which tells you how much protection the prompt was providing. A written deny list for things like curl and .env files does more, and it lives in your project settings. Claude Code settings and permission rules