Signal

GPT-5.6 Sol and Claude Opus 5 are now 0.4 points apart on the industry's toughest coding benchmark — the benchmark stopped being the decision

Independent benchmarking now has GPT-5.6 Sol and Claude Opus 5 within half a point of each other on Terminal-Bench 2.1 (89.5% vs 89.1%) — a gap small enough that raw model capability has stopped being a meaningful reason to pick one AI coding agent over another.