GPT-5
4 items
4 posts
Blog
FrontierCode Benchmark Explained: Why AI Coding Quality Scores Are Wrong (And the Fix)SWE-Bench has an 81% false-positive problem. FrontierCode replaces it with mergeability as the metric - and the scores are sobering for every AI coding tool on the market.
Blog
OpenAI Codex: Terminal and Cloud AI Coding AgentCodex works from the terminal, cloud tasks, IDEs, GitHub, Slack, and Linear. Here is how to use it and how it compares to Claude Code.
Blog
GPT-5 Codex: OpenAI's Agentic Coding ModelOpenAI is drawing a line in the sand. GPT-5 Codex is not an API release.
Blog
GPT-5: OpenAI's Most Capable ModelGPT-5 introduces a fundamentally different approach to inference. Instead of forcing developers to manually configure reasoning parameters, the model operates as a unified system with real-time rou...

Get Smarter About AI Dev
New tutorials, open-source projects, and deep dives on coding agents - delivered weekly.
One email per weekReal code, not theoryFree forever