Skip to main content
Watch: Claude Opus 5.5 Built an Entire 3D World

DEVELOPER TOOLS

222 items

215 posts, 7 tools

Blog
How to Run an AI Agent Fleet on Herdr: Setup Guide

The hands-on guide to running a fleet of coding agents on Herdr: verified install and config steps, three fleet patterns pulled from real projects, the extension ecosystem, and the gaps nobody advertises.

Blog
Herdr vs pi vs tmux: Which Agent Harness Should You Run?

Herdr vs pi vs tmux compared against their own docs: which agent harness fits 2, 10, or 20 agents, and where Herdr genuinely loses.

Blog
Herdr Joined YC. Its Eight-Week Plugin Ecosystem Is the Signal

Within weeks of going public, Herdr collected policy gates, OS-level agent surfaces, editor bridges, a plugin marketplace, and a YC acceptance letter. We measured the ecosystem layer to test what that velocity actually proves about where agent tooling lands next.

Blog
pi Deep Dive: The Minimal Architecture Behind 95,000 GitHub Stars

How a one-developer protest against bloated coding harnesses became a 95,000-star agent toolkit: pi's five-package architecture, branching JSONL session trees, four run modes, and the philosophy that refuses to build sub-agents, plan mode, or MCP.

Blog
Hands-On With Pi: Run Modes and JSONL Session Trees

The practical guide to earendil-works/pi: verified install and auth steps, all four run modes from TUI to SDK, JSONL session trees with branch, fork and resume, and the rough edges nobody advertises.

Blog
DeepSeek V4 Flash Vision Exp: Experimental Vision, Limits, and How to Run It in OpenCode

DeepSeek shipped experimental vision for V4 Flash as deepseek-v4-flash-vision-exp. JPEG, PNG, GIF, and WebP; three input methods; 384 tokens per image. Here is the API contract and how to run it in OpenCode today.

Blog
Don't Paste the AI vs Vomit vs NoBuzz: AI Slop Tools Compared

dontpastetheai.com, Vomit, and NoBuzz hit Hacker News in the same week. A social contract, a local rewrite, and a second-model Claude Code skill - compared with a table and a when-to-stay guide.

Blog
Ox Alpha on OpenCode: The Free Stealth Model, Specs, Privacy Split, and How to Run It

OpenCode dropped Ox Alpha as a free stealth model on August 20, 2026: 1M context, multimodal, near-unlimited for about a week. Here is what is confirmed, where OpenCode and OpenRouter disagree on retention, and how to run it today.

Blog
Give Your Site a Voice: Build a Conversational Support Agent with ElevenLabs Agents

A support page nobody talks to is a support page doing half its job. ElevenLabs Agents gives you a two-way voice agent grounded on your own docs: ASR, LLM, TTS and turn-taking in one platform, a widget you embed in five lines, and CLI or MCP management so your coding agent can run it. The complete one-hour build.

Blog
Claude Code Cross-Session Messaging: Your Agents Can Now Talk to Each Other

Claude Code v2.1.224 lets one running session message another over a first-party channel - plain text, permission-aware, with approval dialogs when bypass-mode sessions talk to each other. Here is what ships, how delivery and inbound controls work, and where the feature stops.

Blog
GitHub Apps Can Now Be Installed at the Enterprise Level, Opening the Platform to Third-Party Integrators

GitHub now lets enterprise owners install third-party GitHub Apps on their enterprise account, and lets any user or organization create apps with enterprise permissions. This opens the enterprise management layer to the broader ecosystem - with a hard security boundary around the most powerful permission set.

Blog
Meta Ships Muse Code and Muse Spark 1.2: A Terminal Agent With a 12x Cheaper Contributor Tier

Meta released Muse Code, a terminal coding agent, and Muse Spark 1.2 on August 5, 2026. The model co-trains with the harness, logs every call to a replay-safe event log, and offers a $0.10/$0.20 contributor tier if Meta may train on your data.

Blog
The Harness Is the New Cost Lever: Databricks' Benchmark and Pi's Context Discipline

Databricks measured the same model through different coding harnesses and found cost per task varied more than 2x at identical quality. Pi's minimalism explains why: roughly 1k tokens of system prompt and 3x less context per turn.

Blog
Give Your Coding Agent a Voice: Dictate Prompts with Wispr Flow

The agent is only as good as the prompt, and the best prompts are the ones you would speak. How to dictate context-rich prompts into an agent CLI like OpenCode hands-free: hotkeys, snippets, dictionary, and Command Mode.

Blog
StateAct Shows Computer-Use Agents Need Program State, Not Just Pixels

Salesforce's StateAct paper argues that long-horizon computer-use agents should inspect files, DOM, and saved outputs directly instead of treating screenshots as the whole world.

Blog
Auto-Narrated Changelog Videos: Build the Pipeline in Under an Hour

Release notes nobody reads are a content problem with a mechanical fix: have a coding agent write the narration script from real git history, record the demo with Screen Studio, and let Descript narrate and edit it. A complete one-hour build.

Blog
Qwen-UI-Agent Points at the Next GUI Agent Runtime

Alibaba's Qwen-UI-Agent report is less interesting as a leaderboard and more interesting as a product spec: mobile, desktop, browser, CLI, and DeepSearch in one stateful agent runtime.

Blog
DeepSeek V4 Flash 0731: The Official Release, Benchmarks, and How to Run It in OpenCode

DeepSeek shipped the official V4 Flash release on July 31, 2026. The re-post-trained 0731 build beats V4-Pro-Preview on agent benchmarks at $0.14/$0.28 per million tokens. Here is what changed and how to run it through OpenCode today.

Blog
Put an AI Agent on a Cron Job: Automating Dev Chores with OpenCode

An agent CLI plus a cron schedule turns recurring dev chores into background work: dependency bumps, doc freshness checks, morning briefs. The pattern, the guardrails, and where to run it - your own hardware or a cloud host.

Blog
What Happens When Tokens Are Too Cheap to Meter: Five Scenarios for Developers and Knowledge Work

Model prices fell 80% in a single announcement this week. Run the trendline forward and the interesting question is not the price - it is what developers, teams, and the broader economy do when intelligence stops being the scarce input.

PreviousPage 2 of 12Next
AI Development Stack

Get Smarter About AI Dev

New tutorials, open-source projects, and deep dives on coding agents - delivered weekly.

One email per weekReal code, not theoryFree forever