Skip to main content
Watch: I Asked Claude to Build Me a Business

AUTOMATION

35 items

30 posts, 3 tools, 2 guides

Blog
Put an AI Reviewer on Every Pull Request: A Webhook Build with OpenCode and Railway

Review capacity is the real bottleneck now that agents ship pull requests faster than people can read them. A webhook service on Railway that runs OpenCode headless against every PR diff, posts findings as a review, and never touches the code: the complete one-hour build.

Blog
Put Your Coding Agent in Discord: A Team Ask-Bot Built in an Hour

Your team already lives in Discord. A slash command, a headless OpenCode agent, and a persistent Railway service add up to a bot that answers questions about your repository in the channel everyone already watches. The complete build, start to finish.

Blog
Pizza Bot Shows Why Background Agents Need Inboxes

AWS open-sourced Pizza Bot, a local-first inbox for long-running AI agent work. The useful lesson is not the brand. It is the queue, approval, checkpoint, and return-path pattern.

Blog
Podcast Your Release Notes: Two-Voice Audio from Git History with ElevenLabs

A changelog nobody reads is a story nobody heard. A coding agent reads your real git history and writes a two-speaker script, the ElevenLabs Text to Dialogue API turns it into a host-and-guest conversation, and ffmpeg stitches the episode. The complete one-hour build.

Blog
Codex Computer History Turns Repeated Work Into Reusable Skills

Codex Computer History gives agents a rolling view of work across apps. Here is how it works, where it helps, and the privacy boundaries developers should understand.

Blog
Grok Bot's Core Primitive: Every Bot Gets a Computer

Every Grok Bot works on a persistent cloud computer with browser and terminal access, and that single primitive explains everything else about the product. Here is why own-computer beats chat drafts, API integrations, and session-scoped agents.

Blog
Grok Bot's Meta Controls: Oversight as a Product Primitive

Grok Bot ships controls over agents rather than controls by agents: approval gates, a chief-of-staff structure, and escalation learning stand in for a settings page. Here is how that oversight model works, and the control questions xAI has not answered publicly.

Blog
Grok Bot Routines: Automations Without Automation-Building

Grok Bot's routines flip the automation playbook: do the job once while a Bot follows along, correct it in plain language, then let the Bot own the schedule. Here is how the mechanic works, where it fits, and how approval gates keep it safe.

Blog
Automate Video Editing with the Descript API: Raw Recording to Published Cut in One Script

The boring 80 percent of video editing is mechanical: cut the filler, clean the audio, add captions, export. The Descript API turns each of those into a scripted job, so a raw recording becomes a published, captioned cut without opening the editor once.

Blog
Make Your Coding Agent Talk: Audio Briefs from Agent Runs with ElevenLabs

The agent finishes, the summary scrolls past, and you will read it later. Build the fix: a coding agent that ends every run with a plain-language summary, piped into ElevenLabs text-to-speech and out as an MP3 you can listen to on the way to work. The complete one-hour build.

Blog
Put an AI Agent Behind a Webhook: Turn GitHub Issues into Pull Requests

The most common trigger for an AI coding agent is not a clock, it is an event. A GitHub webhook, a Railway service, and OpenCode headless add up to a repo where a labeled issue gets a real pull request without anyone at the keyboard. The full build, start to finish.

Blog
Self-Improving Applications Are Now Cheaper Than Hiring: The Claude Code and Codex Closed Loop

Self-improving applications shift the economics of maintenance. Instead of per-token pricing, you pay per closed issue - and the closed loop (user feedback to GitHub issue to Codex scheduled task to reviewed PR) means the cost is predictable, the fixes are testable, and the human is in the merge decision, not the implementation.

Blog
Auto-Narrated Changelog Videos: Build the Pipeline in Under an Hour

Release notes nobody reads are a content problem with a mechanical fix: have a coding agent write the narration script from real git history, record the demo with Screen Studio, and let Descript narrate and edit it. A complete one-hour build.

Blog
Put an AI Agent on a Cron Job: Automating Dev Chores with OpenCode

An agent CLI plus a cron schedule turns recurring dev chores into background work: dependency bumps, doc freshness checks, morning briefs. The pattern, the guardrails, and where to run it - your own hardware or a cloud host.

Blog
Spare Mac for Claude Code: The Remote Control Setup Guide

A guide to setting up an isolated spare Mac that Claude Code can control remotely over SSH, Remote Control from your phone, and Tailscale.

Blog
Loop Engineering: How to Design Agent Loops That Actually Converge

The architecture side of loop engineering: plan/act/verify cycles, convergence criteria, retry policies, budget-bounded loops, and the loop-until-dry pattern. Concrete TypeScript-shaped patterns for building agent loops that stop when they should.

Blog
GLM 5.2 Matches Human Bookkeeper Accuracy on UK VAT Returns - With Some Caveats

A new benchmark shows GLM 5.2 processing 59 transactions and producing VAT returns off by only 7 pence - at $2.73 versus typical accounting fees of $1,000+. Here is what the benchmark actually tested, where the model failed, and why the HN discussion focused on liability.

Blog
Codex Record & Replay: Turn Screen Recordings Into Reusable Automation Skills

A companion guide to the Codex Record & Replay video: OpenAI Codex can now record a recurring computer task and replay it as a reusable automation skill. Here is what the feature is and where it fits.

Blog
Loop Engineering in 9 Minutes: Stop Prompting, Start Building Loops

A companion guide to the Loop Engineering video: the shift from repeatedly prompting an LLM to building long-running loops, goals, and automations. Here is the core idea and where to go deeper.

Blog
Codex-Maxxing: How to Run Long-Running Codex Workflows Without Losing the Plot

Codex-Maxxing should mean bounded autonomy: AGENTS.md, small worktrees, explicit stop conditions, subagents only when work is separable, and review checkpoints that keep humans in control.

Page 1 of 2Next
AI Development Stack

Get Smarter About AI Dev

New tutorials, open-source projects, and deep dives on coding agents - delivered weekly.

One email per weekReal code, not theoryFree forever