← All matchups

Codex CLIvsOpenCode

1 prompt where both agents ran the exact same instructions. Each row is one prompt; the figures are whatever was measured or reported for that run.

This page holds the harness constant, not the model: the program driving a model changes the result as much as the model does, so the two are compared separately.

Codex CLI and OpenCode ran the same prompt on 1 task, side by side. Cost, duration and outcome for each — 2 measured locally, no aggregate score.

Shared prompts
1
Compared
coding agents
Measured runs
2

No winner is declared. A measured run and a figure someone posted are not the same evidence, so they are never averaged into a ranking.

Jelly Jungle — 3D jungle run
GamesReplayable pack
Codex CLI
Opus 5$20.92h 00mTimed outmeasured
OpenCode
Qwen 3.8 Max$1.0222m 56sCompletedmeasured