Anthropic uses undisclosed source in its AI research capability evaluations

2 hours ago 3



Anthropic’s latest system card for Claude Opus 5.5 contains a detail that deserves more attention than it’s getting: the independent nonprofit METR, tasked with evaluating whether the model accelerates AI research to dangerous levels, relied on an undisclosed source of information to reach its conclusions. The evidence behind those conclusions can’t be shared publicly. What the system card actually says The Claude Opus 5.5 System Card, released on September 22, 2026, documents the pre-deployment evaluations Anthropic conducted before releasing its latest Opus-class model. Among the key findings: METR’s external testing did not identify a sustained doubling in the pace of AI development compared to Anthropic’s own internal assessments. But the methodology behind that conclusion is where things get murky. The METR team operated with restricted access, and their report was described as “highly experimental and preliminary.” The system card also notes that Anthropic itself utilized another undisclosed source of information in its evaluations, a detail buried in a document most people will never read in full. What Opus 5.5 actually brings to the table The Opus 5.5 model itself represent...

Read Entire Article