❯ Claude Opus 5.2 reportedly enters limited testing, internal research model figures remain unverified
User observationsAI publication Xinzhiyuan reported that some developers had noticed suspected Opus 5.2 routing in Claude Code, with faster responses, more concise output and more autonomous iteration on long tasks. At the time of checking, Anthropic’s news page contained no formal announcement of that version. The claims should currently be treated as rumors of limited testing.
Probe limitationsThe report suggested identifying the model by asking who Tibo is without web access. Answers can also vary because of context, system settings or sampling, so knowledge questions cannot authenticate a version. Whether the model skips 5.1, when it becomes widely available, and whether Claude Fable 5.2 arrives as early as the end of this month or next remain unannounced.
Internal claimsThe same article cited a purported internal report claiming that Model 2 scored 62.8% on CoBench v2 and was handling substantial company coding work. It also claimed that a recursive self-improvement model could perform 85% of research work and scored 22 points above Model 2. No independently verifiable original report, testing protocol or corresponding full results were obtained.
Capability boundariesThose figures cannot be treated as an achieved workforce replacement rate. Coding and repeated self-checking also do not, by themselves, establish recursive self-improvement. Users can record task completion and human rework to separate reproducible improvements from internal-model rumors while awaiting official releases and public evaluations.
▪ SIGNALRouting changes, better long-task performance and automated AI research involve three distinct levels of evidence; the current reports do not establish a verified leap connecting them.