Skip to main content
Watch: I Asked Claude to Build Me a Business

GPT-5

4 items

4 posts

Blog
FrontierCode Benchmark Explained: Why AI Coding Quality Scores Are Wrong (And the Fix)

SWE-Bench has an 81% false-positive problem. FrontierCode replaces it with mergeability as the metric - and the scores are sobering for every AI coding tool on the market.

Blog
OpenAI Codex: Terminal and Cloud AI Coding Agent

Codex works from the terminal, cloud tasks, IDEs, GitHub, Slack, and Linear. Here is how to use it and how it compares to Claude Code.

Blog
GPT-5 Codex: OpenAI's Agentic Coding Model

OpenAI is drawing a line in the sand. GPT-5 Codex is not an API release.

Blog
GPT-5: OpenAI's Most Capable Model

GPT-5 introduces a fundamentally different approach to inference. Instead of forcing developers to manually configure reasoning parameters, the model operates as a unified system with real-time rou...

AI Development Stack

Get Smarter About AI Dev

New tutorials, open-source projects, and deep dives on coding agents - delivered weekly.

One email per weekReal code, not theoryFree forever