OpenAI and Anthropic share findings from a joint safety evaluation

Ignore

OpenAI Blog · 2025-08-27 10:00 UTC

Not analyzed yet

Eligible for automatic cleanup in 2 day(s) unless marked Must Read.

Content

OpenAI and Anthropic share findings from a first-of-its-kind joint safety evaluation, testing each other’s models for misalignment, instruction following, hallucinations, jailbreaking, and more—highlighting progress, challenges, and the value of cross-lab collaboration.


Your feedback

Keep this article

Protects it from automatic cleanup.