Isegoria

Software engineering · September 2026

Choose the model for the shape of the work.

Claude Opus 5, Claude Fable 5, and every GPT-5.6 tier and effort level - compared without pretending vendor benchmarks form a complete league table.

Measured evidence

Published Coding Agent Index

High-end configurations reported in OpenAI's shared Artificial Analysis table. Opus 5 was not included.

GPT-5.6 Sol80.0
GPT-5.6 Terra77.4
Claude Fable 577.2
GPT-5.576.4
GPT-5.6 Luna74.6

Operating guide

Reasoning effort changes the job profile.

low

Small fixes and scoped subagent leaves.

medium

Routine features and test repair.

high

Ambiguous bugs and multi-file work.

xhigh

Large migrations and deep verification.

max

Deepest single-agent reasoning.

ultra

Parallel research, implementation, and audit.

Sol and Terra support Ultra in the current Codex lineup. Luna, Opus 5, and Fable 5 stop at max.

Ultra under the microscope

More agents change the topology.

OpenAI reports Sol rising from 88.8% at max to 91.9% at Ultra on Terminal-Bench 2.1: a 3.1 percentage-point improvement.

88.8%Sol max+3.1 pp91.9%Sol Ultra

Use Ultra when branches can proceed independently and one integrator owns interfaces, conflicts, and final proof. Stay on max for tightly coupled hotspots.

Routing guide

What to start with

Fast loopLuna medium
Daily defaultTerra high
Hardest workSol max
Large projectFable high / xhigh
Download the six-page report

Primary sources

OpenAI GPT-5.6 · Anthropic Fable 5 · Anthropic Opus 5