OpenAI and Anthropic did something unexpected this summer. They handed each other the keys to their best models and ran safety evaluations on their rival’s technology. The results, published between August 27 and 29, paint a nuanced picture: Anthropic’s Claude models are significantly more cautious, refusing roughly 70% of uncertain queries, while OpenAI’s o3 model matched or outperformed Claude Opus 4 on core alignment metrics. What the cross-lab tests actually measured The evaluation exercise took place in early summer 2025. OpenAI’s team examined Anthropic’s Claude Opus 4 and Sonnet 4 models, while Anthropic’s researchers got their hands on OpenAI’s GPT-4o, GPT-4.1, o3, and o4-mini. Both teams were granted public API access under relaxed external safeguards, meaning the models were tested closer to their raw capabilities rather than behind the usual guardrails consumers see. The tests focused on three key dimensions: how well models follow instruction hierarchy, resistance to jailbreak attempts, and propensity for hallucinations or engagement with harmful requests. Claude models posted a perfect 1.0 score on Password Protection tests, a metric for instruction hierarchy. That 70%...
OpenAI and Anthropic swap AI models in unprecedented safety stress test
3 days ago
3
Related
Visa Study Says Bank-Style Protections Could Push Stablecoin...
57 minutes ago
1
Blockchain.com And NYSE Plan 24/7 Tokenized Stock Access
2 hours ago
3
Tips
Online Tools
Site DoctorIcon Generator
Online Web Tools Collection 1
Online Web Tools Collection 2
Website Analysis
Website SEO
Domain Availability Check
Free videos download
Useful Information
Collection of Useful LinksListen to Free Radio
Listen to Free Music
Free Movie Information
IT Blog
IT News
IT Information
English Address Info
Global News Information
Global Bible Information
Global Book Information
Global Comic Book Information
Global Music Information
BTS, BlackPink Information
Cryptocurrency Information
Pet Dog Information
Overseas Real Estate Information
Cooking Information
Health Information
Overseas Travel Information
click
Popular
Starknet shields 45 assets with new privacy framework
2 weeks ago
74
© Clint's Cryto News 2026. All rights are reserved
















English (US) ·