Skip to content
AI Lehel Briefing
← Back to latest
Models & research OpenAI

OpenAI and Anthropic share findings from a joint safety evaluation

OpenAI and Anthropic share findings from a first-of-its-kind joint safety evaluation, testing each other’s models for misalignment, instruction following, hallucinations, jailbreaking, and more—highlighting progress, challenges, and the value of cross-lab collaboration.

Excerpt supplied by the publisher’s feed

Read the original article at OpenAI Opens in a new tab