Feed

I Gave Claude, Codex, and Gemini the Same App to Build. Then I Made Them Blind-Judge Each Other.

via Dev.to·independent coverage, not Tech Spindle reporting
Why this ranksBenchmark Experiment
29TS ScoreImpact + Innovation, combined
IMPACT42
INNOVATION15

Why Impact & Innovation? We ask two questions of every story: did this actually change something in the real world (Impact), and is the idea genuinely new (Innovation)? Together, that's the TS Score — not engagement, not who posted it, just what matters and what's new.

An experiment had three AI coding agents (Claude, Codex, and Gemini) build the same Arkanoid game, then anonymously judge each other's implementations to compare their capabilities.

Read the full article at Dev.to

Opens Dev.to's site in a new tab

ProgrammingOpen SourceTechnologyPublished Aug 22, 8:51 PM
Spin Up

Got a take on "I Gave Claude, Codex, and Gemini the Same App to Build. Then I Made Them Blind-Judge Each Other."?

Spin Up drafts a post about this story for you — blog, LinkedIn, or X — in your own voice, sourced and attributed automatically.

Keep exploring

More stories ranked this high

Similar TS Score, same beat — ranked the same way the Top feed ranks everything.

The Tech Spindle loop

Read it. Write your take. Publish it.

This page is one stop in a loop built for anyone whose career depends on staying sharp in tech.

Never miss what matters

Get it delivered, or hear it argued out loud

Same ranking, two formats — read the newsletter in two minutes, or let Harika & Rohan debate it in your ears.

Newsletter

Today's ranked stories, in your inbox.

Daily and weekly editions — no fluff, just what actually scored.

Two AI hosts debate the week's biggest stories.

Listen Now

Ready to publish your own take?

Read it, rank it, write about it, publish it — free during early access.

Request an invite →