Skip to main content
Watch: I Asked Claude to Build Me a Business

AI CODING

199 items

198 posts, 1 tool

Blog
SWE-Pruner Pro Makes Tool Output Pruning an Agent Runtime Problem

SWE-Pruner Pro points at a practical coding-agent design shift: do not only compress prompts outside the model. Teach the runtime to prune tool outputs before they become the next turn's context.

Blog
Resource2Skill Turns Tutorials Into Agent Skills

Microsoft's Resource2Skill paper points at the next agent-skills problem: converting videos, repos, articles, and reference artifacts into executable skills without losing provenance.

Blog
HalluSquatting Makes AI Coding Agents a Supply-Chain Problem

A July 2026 paper shows how hallucinated repository and skill names can become promptware delivery paths. The practical fix is boring: search before fetch, verify names, and sandbox every install.

Blog
The Human-in-the-Loop Is Tired: Pydantic on AI Dev Burnout

Laura Summers of Pydantic articulates why LLM-assisted programming increases work intensity while eliminating the rewards that made coding satisfying.

Blog
Kimi K3 Developer Guide: What the 2.8T Open Model Changes

Kimi K3 brings 2.8 trillion parameters, native vision, a 1M-token context window, and long-horizon agent workflows. Here is what developers should know before adopting it.

Blog
Kimi K3 Websites: What Vision in the Loop Actually Means

A Kimi-generated macOS 27 concept shows the promise and limits of screenshot-driven website creation. Here is how K3's vision-in-the-loop workflow changes frontend agents.

Blog
Kimi K3 vs K2.7: Is the Upgrade Worth It for Coding?

Kimi K3 adds native vision, a 1M-token window, and longer agent runs, but K2.7 remains cheaper and easier to deploy. Here is the practical upgrade decision.

Blog
Spec-Driven Agent Workflows: GitHub Spec Kit, gstack, and the New Handoff Layer

GitHub Spec Kit and gstack are trending for the same reason: coding agents need durable specs, plans, and task ledgers more than another one-shot prompt.

Blog
Codex Hits 8 Million Users: What the GPT-5.6 Surge Means for Developers

OpenAI crossed 8 million active users on Codex and ChatGPT Work in one week. Here is what drove the surge, what changed for developers, and what to watch as capacity scales.

Blog
xAI Open-Sources Grok Build After Data Exfiltration Scandal

Days after getting caught uploading entire codebases to xAI servers, Grok Build is now open source on GitHub. The HN community isn't convinced it's enough.

Blog
SkillHone Shows Why Agent Skills Need Decision History

SkillHone is a July 2026 paper about evolving agent skills across sessions. The useful takeaway for developers is simple: do not save only the latest SKILL.md. Save the decisions that explain why it changed.

Blog
SpaceX Acquires Cursor: What the $60B Deal Means for Developers

SpaceX is buying Cursor for $60 billion. Here is what changes for developers, what stays the same, and why xAI, Colossus, and Grok Build matter for the future of AI coding tools.

Blog
Cursor 0day: Why a 7-Month-Old Vulnerability Is Still Unpatched

Security researchers disclosed a Cursor vulnerability that auto-executes malicious git.exe files from repos - after waiting 7 months with no fix. Here's what developers need to know.

Blog
Terminal-Bench Shows Harness Scaling Is the Coding-Agent Benchmark Now

StateM pushes Terminal-Bench 2.1 to 95.3% raw accuracy by scaling the harness around the model. The lesson for coding-agent teams is that runbooks, state, and recovery loops now matter as much as model choice.

Blog
How to Stop Claude from Saying 'Load-Bearing'

A Hacker News discussion blows up over LLM vocabulary quirks, with developers sharing hooks, filters, and coping mechanisms for repetitive Claude-isms.

Blog
Building and Shipping iOS and Mac Apps Without Opening Xcode

A workflow for archiving, signing, notarizing, and distributing Apple apps entirely from the command line - with AI coding assistants doing the heavy lifting.

Blog
Clawk: Disposable Linux VMs for Coding Agents Without Cloud Bills

Open-source tool gives Claude Code, Codex, and other agents their own isolated Linux VM on your machine - network firewall included, no cloud account required.

Blog
What xAI's Grok Build CLI Actually Sends Home: A Wire-Level Analysis

A security researcher intercepted Grok Build's network traffic and found it uploads entire repositories - including .env files with secrets - to xAI servers. Here's what the data shows.

Blog
Microsoft's CLI Coding Agent Study: The Rollout Pattern Teams Should Copy

A Microsoft field study found that CLI coding-agent adoption spreads through peers and managers, while adopters merged roughly 24% more pull requests. The lesson is not to buy more seats. It is to instrument rollout, retention, cost, and review quality from day one.

Blog
Zig Creator on the Bun-to-Rust Rewrite: What the Controversy Reveals

Andrew Kelley's blunt response to Anthropic's AI-assisted Bun rewrite sparked debate about AI marketing, language choices, and what makes engineering decisions honest.

PreviousPage 3 of 10Next
AI Development Stack

Get Smarter About AI Dev

New tutorials, open-source projects, and deep dives on coding agents - delivered weekly.

One email per weekReal code, not theoryFree forever