Switch language한국어
Back to the list

Researchers let Claude Code discover AI scaling algorithms that humans probably wouldn't have designed

TL;DR AI

Key summary

2 min read
  1. Researchers used Claude Code to auto-discover a new test-time scaling algorithm in an offline simulation called AutoTTS.

  2. Instead of human-written branching and stopping rules, the agent learned to allocate compute based on confidence shifts during reasoning.

  3. The method used about 70% fewer tokens than standard self-consistency while matching or improving accuracy on math and other benchmarks.

  4. The algorithm also transferred to another model and task, suggesting AI agents can design better inference-control strategies than humans.

Read the original