Skip to main content
Watch: I Asked Claude to Build Me a Business

OPENAI

98 items

92 posts, 6 tools

Blog
OpenAI's June API Updates Are Really a Control-Plane Upgrade

OpenAI's June 2026 API changelog looks like scattered platform plumbing. Read together, moderation scores, workload identity, Admin APIs, prompt-cache retention, container billing, and Secure MCP Tunnel are the pieces teams need to run agents with real controls.

Blog
Codex-Maxxing: How to Run Long-Running Codex Workflows Without Losing the Plot

Codex-Maxxing should mean bounded autonomy: AGENTS.md, small worktrees, explicit stop conditions, subagents only when work is separable, and review checkpoints that keep humans in control.

Blog
OpenAI Agent Builder and Evals Are Shutting Down: Move the Agent Stack Into Code

OpenAI's June deprecations put Agent Builder, hosted Evals, and reusable prompts on a November 30 shutdown path. Here is the practical migration plan: Agents SDK, repo-owned prompts, and eval receipts.

Blog
OpenAI Daybreak Shows the AppSec Bottleneck Is Patching, Not Finding

OpenAI's Daybreak and Patch the Planet point at the real agentic AppSec shift: security agents only matter when they produce validated, reviewable patches maintainers can actually merge.

Blog
Codex Logging Bug Can Write Terabytes to Your SSD

A Codex CLI SQLite logging bug showed how global TRACE logs can burn SSD write endurance. OpenAI has now merged fixes, but the incident is a useful local-agent operations lesson.

Blog
Noam Shazeer Joins OpenAI After Two Years Back at Google

The Transformer co-creator leaves Google DeepMind for OpenAI just two years after Google paid $2.7 billion to bring him back from Character.AI.

Blog
Codex Gets Computer Use in the EU - and a Clean Claude Code Import

OpenAI's mid-June 2026 Codex drop brings Computer Use to the EEA, UK, and Switzerland and adds selective Claude Code imports plus managed Bedrock auth to the CLI. Here is what actually shipped, verified against the changelog.

Blog
Frontier Model API Pricing, September 2026: Claude vs OpenAI vs Gemini vs DeepSeek

Same-day-verified llm api pricing september 2026: Claude Fable 5.1 and Opus 5.5, GPT-6 Astra/Sol/Luna, Grok 4.7, Claude Sonnet 5, and DeepSeek V4.1-Flash compared per million tokens, plus the caveats that change the math.

Blog
The Mid-Tier Shootout: GPT-5.4 vs Gemini 3.1 Pro vs DeepSeek V4 Pro

GPT-5.4 vs Gemini 3.1 Pro vs DeepSeek V4: pricing, benchmarks, context behavior, and license terms for the mid-tier models that carry most production traffic.

Blog
GPT-5.6 Sol vs Claude Opus 5: The $5 Workhorse Head-to-Head

GPT-5.6 Sol vs Claude Opus 5: both cost $5 per million input tokens, so the workhorse-tier decision comes down to output pricing, benchmarks, and tooling.

Blog
Migrating Off Retired GPT Models in 2026: A Working Checklist

Migrating off retired GPT models in 2026: the live retirement table, what maps to what, an eval-before-switch day plan, and when to jump providers.

Blog
Codex in June 2026: What Changed Since the Spring Wave

The Codex changelog from April through June 2026 covers GPT-5.5, Goal mode going stable, Sites, a Chrome extension, Amazon Bedrock support, and mobile access from iOS. Here is what actually shipped and what it means in practice.

Blog
Codex Exec in CI: The Practical Guide to Headless OpenAI Agents

codex exec is OpenAI's non-interactive mode for running Codex agents from scripts, CI pipelines, and GitHub Actions - here is how to set it up safely with real flags and working YAML.

Blog
OpenAI Agents SDK vs Claude Agent SDK: Building Agents on the Two Big Platforms

A practical comparison of OpenAI's Agents SDK and Anthropic's Claude Agent SDK - orchestration models, tool ecosystems, sandboxing, and how to choose the right platform for your team.

Blog
The TypeScript AI Agent Stack in Mid-2026: Mastra vs Vercel AI SDK vs OpenAI Agents SDK vs LangGraph.js

Four mature, production-ready TypeScript frameworks have made building agents genuinely enjoyable. Here is how to pick the right one - and how they fit together.

Blog
Spreadsheet Agents Need Permission Ledgers

The ChatGPT for Google Sheets exfiltration report is not just a spreadsheet bug. It is a warning about agentic office tools: permissions need to be action-scoped, logged, revocable, and visible.

Blog
Codex CLI Vim Mode Is an Ergonomics Signal

Codex CLI 0.129.0 added modal Vim editing in the composer. The feature is small, but it points at a bigger shift: terminal agents are becoming native engineering workbenches.

Blog
Codex Automations: Where Scheduled AI Agents Actually Help

Codex automations are useful when recurring engineering work has clear inputs, reviewable outputs, and safe boundaries. Here is the practical playbook.

Blog
Codex Is Becoming a General-Purpose AI Agent, Not Just a Coding Tool

OpenAI is turning Codex from a coding assistant into a broader agent workspace for files, apps, browser QA, images, automations, and repeatable knowledge work.

Blog
Codex SDK vs CLI vs GitHub Action: Which Surface Should You Build On?

Codex is no longer just a terminal agent. Here is when to use the Codex SDK, Codex CLI, or openai/codex-action, and how to avoid building the same agent loop three times.

PreviousPage 3 of 5Next
AI Development Stack

Get Smarter About AI Dev

New tutorials, open-source projects, and deep dives on coding agents - delivered weekly.

One email per weekReal code, not theoryFree forever