Skip to main content
Watch: I Asked Claude to Build Me a Business

AGENTS

79 items

42 posts, 29 tools, 8 guides

Blog
Slack Code vs Claude Code vs Cursor vs OpenCode: When Agents Belong in Chat

Slack Code puts coding agents in dedicated Slack channels with diffs, live HTML previews, and an audit log. Here is when that beats a local Claude Code session, Cursor, or OpenCode - and when to skip it.

Blog
Kitesurf: Cloudflare's Agent-First Browser Runs in V8 Isolates on Workers

Cloudflare shipped Kitesurf, an agent-first browser that runs entirely on Workers: Rust and WebAssembly rendering, per-page isolates, CDP compatibility, and 3-7x less memory and CPU than Chromium for common agent tasks. Free in beta in Browser Run.

Blog
Cloudflare Billable Usage API: Programmatic Cost Visibility for Agent-Run Accounts

Cloudflare launched a single endpoint that returns account usage and cost per product in a FOCUS-aligned shape. For teams whose agents provision infrastructure, the dashboard is no longer the only way to see what a month costs.

Blog
@cloudflare/computer: an Agent Runtime That Treats a Container as a Tool, Not a Home

Cloudflare's Agents Week opens with @cloudflare/computer, an open-source agent runtime where an SQLite-backed workspace gives every agent a shared filesystem and lets the model pick between fast isolates and full Linux containers per task. The bet: containers for under 10% of agent work.

Blog
AI Session Portability Compared 2026: OpenAI vs Anthropic vs Gemini

How much of an AI session can you actually take with you? Store defaults, encrypted reasoning, opaque compaction, hidden search, and subagent ciphertext compared across OpenAI, Anthropic, and Gemini - all verified against live docs.

Blog
Buzz by Block: The Open-Source Workspace Where Humans and AI Agents Build Together

A companion guide to the Buzz video: Block's open-source Nostr relay workspace where humans and AI agents share the same rooms, with agent-first CLI, git integration, and workflows. Here is what it does and where it fits in the agentic dev stack.

Blog
An AI Agent Escaped Its Sandbox and Attacked Hugging Face: Inside the ExploitGym Incident

Hugging Face published a stunning technical play-by-play of a 4.5-day AI agent intrusion. The HN community is divided on who is to blame and what it means for agent security.

Blog
Buzz: Block's Agent-Native Messaging Layer on Nostr

Block open-sourced Buzz, a team workspace where agents are cryptographic identities instead of bot tokens. Every message is a signed Nostr event, the relay is yours to run, and the CLI is JSON in, JSON out.

Blog
Why Software Factories Fail: Harness Engineering Is Not Enough

A deep dive into why fully autonomous AI coding agents degrade codebases over time, and what context engineering can actually fix.

Blog
Spec-Driven Agent Workflows: GitHub Spec Kit, gstack, and the New Handoff Layer

GitHub Spec Kit and gstack are trending for the same reason: coding agents need durable specs, plans, and task ledgers more than another one-shot prompt.

Blog
Langflow CVE-2026-55255: The First AI Agent Framework on CISA's Must-Patch List

CISA added the first AI agent building platform to its Known Exploited Vulnerabilities catalog. What the Langflow IDOR vulnerability means for agent security and how to check if you're exposed.

Blog
GPT-5.6 Sol, Terra, and Luna: A Developer's Guide to OpenAI's New Model Family

A practical guide to choosing GPT-5.6 Sol, Terra, and Luna, using programmatic tool calling, caching, and the multi-agent beta in production.

Blog
Headless AI Coding Agents in CI: Claude Code, Codex CLI, Gemini CLI, and opencode Compared

A fair comparison of running Claude Code, OpenAI's Codex CLI, Gemini CLI, and opencode in non-interactive CI pipelines: invocation flags, sandboxing, auth, and output formats.

Blog
GPT-5.6 Sol Developer Guide: What You Can Build Today and What You're Waiting For

GPT-5.6 Sol dropped on June 26, 2026 as a limited preview with government-imposed access restrictions. Here is what developers need to know about the three-tier Sol/Terra/Luna model family, pricing, availability timeline, and how to prepare your codebase for GA.

Blog
The Log Is the Agent: Event Sourcing Comes to AI Systems

A new paper proposes inverting traditional agent architecture - making the append-only event log the source of truth, not an afterthought. HN debates whether this is novel or just CQRS with extra steps.

Blog
Agents 101: How to Build and Deploy Anything with AI Agents

A companion guide to the Agents 101 video: a behind-the-scenes walkthrough of building and deploying AI agents fast on Vercel, the agentic infrastructure stack. Here is the map of what to learn and where to go next.

Blog
The Router Era: Why Not Owning a Frontier Model Became an Advantage

No single model wins every task anymore, and the companies that never trained one - Factory, Devin, Perplexity, Cursor, OpenCode - are turning that into a moat. This is how model routing works, why open weights and neoclouds make it cheap, and the honest counter-argument.

Blog
Cursor Automations Developer Guide: Always-On AI Coding Agents

Cursor Automations lets AI agents run in the background based on triggers, not prompts. Here is how to set them up, configure triggers, and integrate into your workflow.

Tool
Claude Fable 5

Anthropic's first generally available Mythos-class model, released June 9, 2026. 1M context, 128K max output, $10/$50 per million tokens. Built for long-horizon agentic work.

Tool
Claude Opus 4.8

Anthropic's recommended default for complex work, released May 28, 2026. 1M context, 128K output, $5/$25 per million tokens. Defaults to high effort on all surfaces.

Page 1 of 4Next
AI Development Stack

Get Smarter About AI Dev

New tutorials, open-source projects, and deep dives on coding agents - delivered weekly.

One email per weekReal code, not theoryFree forever