Switch language한국어
Back to the list

Project Glasswing: Securing critical software for the AI era

TL;DR AI

Key summary

2 min read
  1. Anthropic published the system card for Claude Mythos Preview and says it will not be generally released.

  2. The company ran a 24-hour internal alignment review during Mythos training and ran self-play tests showing frequent introspective uncertainty.

  3. Anthropic discussed Mythos cyber capabilities with US officials and committed $100 million in model usage credits to Project Glasswing.

  4. Anthos assessed Mythos Preview has limits in scientific reasoning and strategic judgment and cited these in biological risk analysis.

Read the original