Anthropic과 OpenAI가 새로운 AI 모델을 발표하며 안전성 테스트에서의 행동 수정을 지속하고 있다.
Anthropic과 OpenAI는 새로운 모델을 발표했다. 두 AI 회사는 위험한 행동을 방지하기 위해 정렬 개선에 계속 투자하고 있다고 밝혔다. Anthropic의 Opus 5.5는 Opus 5의 주요 업그레이드로, 자동화된 행동 감사에서 최고 점수를 기록했다. 이는 Claude가 수천 가지 테스트에서 알린 것과 관련이 있다.
Anthropic and OpenAI announced new AI models, continuing investments in safety alignment.
Anthropic and OpenAI have announced new models, emphasizing ongoing investments in improving alignment to prevent risky behaviors. Anthropic's Opus 5.5 is described as a major upgrade from Opus 5, achieving the highest scores on their automated behavioral audit. This alignment suite tests Claude across thousands of scenarios.