Skip to main content
Watch: Claude Opus 5.5 Built an Entire 3D World

DEVELOPER TOOLS

222 items

215 posts, 7 tools

Blog
Codex Record & Replay: Turn Screen Recordings Into Reusable Automation Skills

A companion guide to the Codex Record & Replay video: OpenAI Codex can now record a recurring computer task and replay it as a reusable automation skill. Here is what the feature is and where it fits.

Blog
GLM 5.2 in 9 Minutes: The Open-Weight Rival to GPT-5.5

A companion guide to the GLM 5.2 video: an open-weight model positioned against GPT-5.5, walked through with benchmarks, pricing, and a live OpenCode demo. Here is what the video covers and where to go deeper.

Blog
Godot Bans AI-Authored Code Contributions - What It Means for Open Source

The Godot Foundation has established a policy banning autonomous AI agent code and substantial AI-generated contributions, citing reviewer burnout and concerns about maintainer mentorship.

Blog
GPT-5.5 in 7 Minutes: Benchmarks, Codex Agents, Context Window, and Pricing

A companion guide to the GPT-5.5 video: OpenAI's newly released model rolling out to ChatGPT and Codex, reviewed through benchmarks, agent capabilities, context window, and pricing. Here is what the video covers and where to go deeper.

Blog
Loop Engineering in 9 Minutes: Stop Prompting, Start Building Loops

A companion guide to the Loop Engineering video: the shift from repeatedly prompting an LLM to building long-running loops, goals, and automations. Here is the core idea and where to go deeper.

Blog
OpenAI Codex in 7 Minutes: The Desktop App, Plan Modes, and Multi-Agent Workflows

A companion guide to the OpenAI Codex video: a tour of the Codex desktop app, its plan and goal modes, plugins, multi-agent workflows, and UI annotation. Here is what the video shows and where to go deeper.

Blog
Webernetes: Kubernetes Ported to the Browser in TypeScript

Ngrok engineer Sam Rose ported 100,000 lines of Kubernetes to TypeScript, creating a browser-based cluster for educational use - with 2,059 tests proving it behaves like real k8s.

Tool
AgentCanvas

A hosted infinite canvas your headless AI agents drive over MCP. Any MCP-speaking agent - Claude Code, Codex, Cursor, or a script - creates HTML docs, images, and video on a live canvas, streamed in as it builds.

Blog
Outer Shell: A Graphical Desktop for Your Remote Server via SSH

A new project proposes a graphical shell layer for SSH that turns remote servers into browsable desktops. The HN discussion digs into architecture choices, the terminology debate, and whether this solves a real problem.

Blog
LangSmith Fleet Turns Agent Ops Into On-Call Work

LangChain's June LangSmith updates point to a practical agent-ops pattern: Fleet templates, on-call triage, computer use, Slack interrupts, MCP auth, traces, and eval progress all belong in one operator loop.

Blog
OpenAI's June API Updates Are Really a Control-Plane Upgrade

OpenAI's June 2026 API changelog looks like scattered platform plumbing. Read together, moderation scores, workload identity, Admin APIs, prompt-cache retention, container billing, and Secure MCP Tunnel are the pieces teams need to run agents with real controls.

Blog
Grok Build Developer Guide: xAI's Terminal Coding Agent (June 2026)

Grok Build is xAI's agentic CLI with 8 parallel subagents, a plan-first workflow, and Arena Mode for competing outputs. Installation, pricing, real commands, and how it compares to Claude Code and Codex.

Blog
Perplexity Bumblebee: Developer Guide to the Open Source Supply Chain Scanner

Bumblebee is Perplexity's open source scanner for detecting compromised packages, extensions, and MCP configs on developer machines. A read-only Go binary that checks npm, PyPI, Go modules, and 10+ ecosystems against exposure catalogs - without running any install scripts. Here is how to set it up and use it.

Blog
Best AI Code Review Tools in 2026: CodeRabbit vs DeepSource vs Greptile Compared

AI-assisted development generates PRs faster than humans can review them. Here are the tools that help - CodeRabbit, DeepSource, Greptile, and others compared on pricing, platform support, and security capabilities.

Blog
AI's Affordability Crisis Is Really an Agent Cost Accounting Problem

A viral Hacker News thread about AI affordability points at the right problem, but developer teams need a more useful cost model: retries, cache misses, review time, routing, and failed loops.

Blog
Armin Ronacher on The Coming Loop and Why Agent-Driven Code Still Needs Human Comprehension

Armin Ronacher's new essay explores the tension between letting AI agents loop autonomously and maintaining the engineering comprehension that makes software maintainable. The Hacker News discussion adds practical caveats worth reading.

Blog
Anthropic Claude Tag Turns Slack Into a Shared Agent Workspace

Claude Tag is Anthropic's new Slack-based beta for Team and Enterprise users. The important shift is not chat convenience - it is shared agent identity, channel context, and team-visible work.

Blog
Envoy AI Gateway 1.0 Makes LLM Routing an Infrastructure Decision

Envoy AI Gateway 1.0 is production-ready. The useful question for builders is when an Envoy-based LLM gateway beats direct SDK calls, LiteLLM, OpenRouter, or a hosted AI gateway.

Blog
F3 Is a Reminder That File Formats Are Becoming Runtime Contracts

F3 is trending on Hacker News as a research prototype for a future-proof columnar file format. The useful takeaway is not to replace Parquet tomorrow. It is that data files are starting to carry more of their own runtime contract.

Blog
GitHub Copilot CLI, BYOK, and AI Credits: The New Cost-Control Stack

GitHub's June Copilot updates point beyond autocomplete: CLI access, bring-your-own-key model routing, AI credit metrics, and external agent providers make Copilot a governed agent platform.

PreviousPage 5 of 12Next
AI Development Stack

Get Smarter About AI Dev

New tutorials, open-source projects, and deep dives on coding agents - delivered weekly.

One email per weekReal code, not theoryFree forever