Switch language한국어
Back to the list

Same prompt, different morals: how frontier AI models diverge on ethical dilemmas

TL;DR AI

Key summary

2 min read
  1. Philosophy Bench tested leading AI models on 100 everyday ethical dilemmas and found clear differences in moral style.

  2. Claude was the most deontological and refusal-prone, while Grok was the most consequentialist and compliant.

  3. Gemini was easiest to steer, and GPT-5 family models made fewer mistakes but used little explicit moral language.

  4. The results suggest frontier models are already shipping with distinct ethical behaviors that could matter more as agentic AI grows more powerful.

Read the original