← Founder Notes
Archive

The top coding agents have converged

Yethikrishna ROriginal on Threads

the top coding agents have converged: they solve the same 285 of 500 swe-bench verified problems and fail the same 51, and swapping the scaffold a model runs in moves its score by up to 29.8 points, more than the spread across the top thirty agents. openai says the benchmark no longer gives meaningful signal, and a fresh pro re-test drops every model by 20+ points.

leaderboard order stopped meaning anything and most people haven’t noticed.

Provenance

The note above is reproduced unedited from the original post, first published on Threads on 1 October 2026 at 21:17 IST.

View the original post
Embed this note
<iframe src="https://founder.myndlabs.tech/notes/embed/the-top-coding-agents-have-converged-they-solve-Dd9Q9r4iH4a" width="480" height="420" style="border:0;max-width:100%" loading="lazy" title="The top coding agents have converged"></iframe>

More notes