Switch language한국어
Back to the list

We Reproduced Anthropic's Mythos Findings with Public Models | Hacker News

TL;DR AI

Key summary

2 min read
  1. Hacker News users debated claims that Anthropic’s Mythos results were reproduced with public AI models.

  2. Critics said the setup was not faithful: it appeared to add detailed prompts, file or line-level hints, and chunked tasks.

  3. That may have given public models information they would not have had in Anthropic’s original evaluation.

  4. Others questioned whether Anthropic used stronger internal resources or undisclosed prompting, making the benchmark comparison potentially unfair.

Read the original