Skip to main content
Watch: I Asked Claude to Build Me a Business

CODEX

91 items

90 posts, 1 guide

Blog
Taste Skills Are Turning Agent Review Into Infrastructure

GitHub trending is full of anti-slop, taste, and compound-engineering skills. The real signal is not that agents need more prompts. It is that teams are trying to make subjective review criteria executable.

Blog
Local Code Graphs Are the Agent Context Layer

CodeGraph shows why coding agents need a local, queryable repo map. The win is not magic token savings. It is faster orientation, fewer wrong files, and better review receipts.

Blog
AI Code Review Is the New Bottleneck

Coding agents make code faster than teams can review it. The next advantage is not bigger prompts. It is review systems that force reproduction, small diffs, tests, and receipts.

Blog
AgentMemory Is Useful Only If You Audit What It Remembers

AgentMemory gives Claude Code, Codex, Cursor, and other agents persistent local memory. The real adoption question is not recall accuracy. It is whether your team can inspect, prune, and govern what gets remembered.

Blog
Codex CLI Vim Mode Is an Ergonomics Signal

Codex CLI 0.129.0 added modal Vim editing in the composer. The feature is small, but it points at a bigger shift: terminal agents are becoming native engineering workbenches.

Blog
Skills for Real Engineers Need Governance, Not Fandom

Matt Pocock's skills repo is a useful signal for AI coding teams. The next step is treating skills like governed production controls, not a folder of viral prompts.

Blog
Agent Memory Benchmarks Are Not Enough

Persistent memory for coding agents is trending because every session still starts too cold. The hard part is not saving facts. It is proving recall, freshness, deletion, and rollback under real development pressure.

Blog
Ruflo Is an Agent Meta-Harness. Treat the Star Count as a Warning Label.

Ruflo turns Claude Code and Codex into a larger agent harness with plugins, memory, swarms, MCP tools, and federation. The useful question is not the star count. It is how much harness you actually need.

Blog
Terminal Agents Are the New Developer Runtime

Terminal agents like Claude Code, Codex CLI, OpenCode, Copilot CLI, and DeepSeek-TUI are converging on the same runtime layer: permissions, sandboxing, rollback, diagnostics, subagents, receipts, and cost controls.

Blog
Codex Automations: Where Scheduled AI Agents Actually Help

Codex automations are useful when recurring engineering work has clear inputs, reviewable outputs, and safe boundaries. Here is the practical playbook.

Blog
Codex Is Becoming a General-Purpose AI Agent, Not Just a Coding Tool

OpenAI is turning Codex from a coding assistant into a broader agent workspace for files, apps, browser QA, images, automations, and repeatable knowledge work.

Blog
Codex Loops: What Boris Cherny Gets Right About Managing Agent Work

Boris Cherny's loop-heavy Claude Code workflow points at the next Codex content lane: recurring agents that babysit PRs, CI, deploys, and feedback streams.

Blog
Codex SDK vs CLI vs GitHub Action: Which Surface Should You Build On?

Codex is no longer just a terminal agent. Here is when to use the Codex SDK, Codex CLI, or openai/codex-action, and how to avoid building the same agent loop three times.

Blog
Karpathy's Loopy Era Is the Best Way to Understand Codex

Andrej Karpathy's loopy era frame explains why Codex is becoming less like a chatbot and more like an agent loop manager for real software work.

Blog
OpenAI's Codex Mac Certificate Deadline Is a Runbook Test

OpenAI's May 8 macOS certificate rotation for ChatGPT, Codex, Codex CLI, and Atlas is not just a one-off update. It is a useful test of how your team governs AI developer tools.

Blog
Parallel Coding Agents Need Merge Discipline

Parallel agents can move faster than one agent, but only when tasks have clean ownership, review receipts, and a merge path that does not turn speed into cleanup work.

Blog
Codex Changelog April 2026: Goals, Browser Use, GPT-5.5, and Safer Agents

OpenAI's April 2026 Codex changelog shows a clear product shift: Codex is becoming a full agent workspace with goals, browser verification, automatic approval reviews, plugins, and tighter permission profiles.

Blog
OpenAI Codex, Managed Agents, and AWS: What Developers Should Watch

OpenAI is moving Codex from a coding assistant into an enterprise agent platform. Here is what changed with Codex, Managed Agents, AWS, and the Responses API.

Blog
Codex Security Preview: AppSec Agent for Real Repos

OpenAI's Codex Security agent reviews app code for vulns. Here is what it caught and missed on three real production repos.

Blog
GPT-5.5-Codex in Production: What Actually Changes

GPT-5.5-Codex merges Codex and GPT-5 stacks. Here is what the unified model means for real coding agents - latency, costs, prompt rewrites.

PreviousPage 4 of 5Next
AI Development Stack

Get Smarter About AI Dev

New tutorials, open-source projects, and deep dives on coding agents - delivered weekly.

One email per weekReal code, not theoryFree forever