DEV Community

Cover image for GPT-5.5 vs Claude vs Gemini: Which AI Explains Things Clearest? I Built a Benchmark to Find Out
Michael Omijie
Michael Omijie

Posted on

GPT-5.5 vs Claude vs Gemini: Which AI Explains Things Clearest? I Built a Benchmark to Find Out

Kaggle Benchmarking Challenge Submission

This is a submission for the Kaggle Benchmarking Challenge

What I Benchmarked

Everyone says their AI explains things best. I wanted real data.

I built a benchmark with 8 test prompts asking models to explain complex topics simply recursion, photosynthesis, cryptocurrency, machine learning, and more. The goal: find out which AI is actually the clearest teacher for beginners.

This matters because if you're a student, teacher, or self-learner, picking the wrong AI assistant could mean getting a confusing wall of text instead of a clear answer.

Models Tested

  • GPT-5.5 — OpenAI's latest, known for strong instruction following
  • Claude Sonnet 4.6 — Anthropic's model, known for thoughtful responses
  • Gemini 3.7 Flash — Google's fast and efficient model

Findings

Model Score
🏆 GPT-5.5 76.3%
Claude Sonnet 4.6 38.8%
Gemini 3.7 Flash 38.8%

I was genuinely shocked. GPT-5.5 almost doubled Claude and Gemini's scores.

My scoring rewarded responses between 50-200 words — concise enough to be clear, detailed enough to be useful. GPT-5.5 consistently hit that sweet spot. Claude and Gemini tended to over-explain, writing long responses that would overwhelm a beginner.

The lesson? More words doesn't mean better explanation. GPT-5.5 understood the assignment keep it simple.

What I'd measure next: Have real beginners rate the clarity themselves, not just word count. Human judgment might tell a different story.

My Benchmark

👉 https://www.kaggle.com/benchmarks/michaelomijiemkings/who-explains-things-clearest

kagglechallenge

--

Top comments (0)