Switch language한국어
Back to the list

Base Models Look Human To AI Detectors

TL;DR AI

Key summary

2 min read
  1. A new study found that base language models were often judged as human by GPTZero and Pangram, while instruction-tuned outputs were more likely flagged as AI.

  2. The researchers introduced HIP, an iterative paraphrasing pipeline designed to make text look more human while preserving meaning.

  3. HIP improved detector evasion across model sizes, including Llama-3 and Qwen-3, with relatively little semantic loss.

  4. The findings suggest current AI detectors may be reacting to instruction-tuning artifacts more than reliably identifying machine-generated text.

Read the original