Skip to main content
Watch: I Asked Claude to Build Me a Business

AI CODING

199 items

198 posts, 1 tool

Blog
Claude Code Sends 33k Tokens Before Your Prompt - OpenCode Sends 7k

New research shows Claude Code's system prompt and tool scaffolding consume 4.7x more tokens than OpenCode before processing user input. The HN thread debates whether that overhead buys better outcomes.

Blog
Dockerless Verification Is The Next Coding Agent Bottleneck

ByteDance's Dockerless paper asks whether coding-agent patches can be verified without spinning up per-repo environments. The practical answer is not replace CI. It is use cheaper evidence before CI.

Blog
GPT-5.6 vs Claude 5: What the New Tiers Mean for Choosing a Coding Model

OpenAI's GPT-5.6 Sol, Terra, and Luna tiers versus Anthropic's Claude Fable 5 and Mythos 5. Verified pricing, benchmarks, and a practical framework for picking a coding model in July 2026.

Blog
ChatGPT Work vs Claude Cowork 2026 - Complete Comparison

OpenAI launched ChatGPT Work to compete with Claude Cowork. Here is how they compare on features, pricing, integrations, and which workflow each handles best.

Blog
Cursor v3.11 Side Chats: Developer Guide for Parallel Agent Conversations

Cursor v3.11 introduces Side Chats for parallel agent conversations, Conversation Search across past sessions, and Cloud Agent Hooks for self-correcting loops. A practical guide to the new features released July 10, 2026.

Blog
Write Code Like a Human Will Maintain It - The AI Era Debate

A new essay argues that letting AI generate sloppy code creates a downward spiral where future AI absorbs those bad patterns. HN's 250+ comment thread is split between believers and pure vibe-coders.

Blog
Scarf Drops Haskell After 7 Years - LLMs Changed the Calculus

A Haskell Foundation board member explains why Scarf moved to Python after 7 years in production. The culprit: LLM-driven development made Haskell's compile times an unacceptable bottleneck.

Blog
Does Your Codebase Pattern Determine AI Output Quality? HN Debates the Economics of Rewrites

A viral post argues AI works better on standardized codebases, making rewrites economically sensible. HN pushes back with the Mythical Man-Month and maintainability concerns.

Blog
AI Test Generation Tools Compared 2026: Which One Actually Catches Bugs

A fair comparison of AI-assisted test generation tools for coding agents - what they generate, where they plug into your workflow, and which claims to verify yourself before trusting the output.

Blog
Bun Rewrites 535K Lines of Zig to Rust in 11 Days Using Claude

The Bun runtime completed an AI-assisted rewrite from Zig to Rust, fixing memory safety issues and improving performance. Here is what HN thinks and why it matters for LLM-assisted code migration.

Blog
ChatGPT Work and Codex Now Share One Desktop App: What Actually Changed

OpenAI is consolidating its desktop apps, not merging ChatGPT and Codex into one indistinguishable product. Here is how ChatGPT Work, Codex, and GPT-5.6 fit together.

Blog
Grok 4.5: xAI Releases Cursor-Trained Coding Model at $2/M Input Tokens

xAI launched Grok 4.5, trained on trillions of Cursor interaction tokens. At $2/M input pricing, it undercuts Claude and GPT while benchmarking near Opus 4.7 level.

Blog
VS Code 1.128 Multi-Chat Claude Sessions Developer Guide 2026

VS Code 1.128 shipped today with multi-chat support for Claude agent sessions. Run parallel conversations in one workspace, fork turns, compare approaches, and monitor subagents. Complete setup and workflow guide.

Blog
ZCode Developer Guide 2026: Z.ai's Agentic IDE for GLM-5.3

ZCode is Z.ai's free desktop agentic development environment tuned for GLM-5.3. Setup, Goal Mode, pricing, and how model connections work.

Blog
Clean Code Makes AI Agents 34% More Efficient - New Research

A controlled study of 660 Claude Code trials shows clean codebases reduce token usage by 7-8% and file revisitations by 34%, while pass rates stay the same. Traditional maintainability principles still matter in the age of AI coding.

Blog
Does Code Cleanliness Affect AI Coding Agents?

A new SonarSource study finds clean code doesn't boost agent pass rates - but it cuts token usage by 8% and file revisitations by 34%. Here's what that means for your codebase.

Blog
GPT-5.6 Sol Ultra Coming to Codex with Cooperative Subagents

OpenAI teases its most capable coding model yet - Sol Ultra uses trained subagents that communicate during tasks, reportedly hitting 91.9% on Terminal-Bench 2.1.

Blog
GPT-5.6 Sol Developer Guide: What You Can Build Today and What You're Waiting For

GPT-5.6 Sol dropped on June 26, 2026 as a limited preview with government-imposed access restrictions. Here is what developers need to know about the three-tier Sol/Terra/Luna model family, pricing, availability timeline, and how to prepare your codebase for GA.

Blog
Program-as-Weights Turns Prompts Into Local Fuzzy Functions

The Program-as-Weights paper is a useful signal for developers: some LLM calls may move from per-request API prompts into compact local artifacts that behave like reusable fuzzy functions.

Blog
Dan Luu's Agentic Coding Notes Point to the Real Bottleneck

Dan Luu's new agentic coding essay is not another vibe check. It is a useful reminder that coding agents only compound when the test loop, review loop, and task-selection loop are stronger than the code generator.

PreviousPage 4 of 10Next
AI Development Stack

Get Smarter About AI Dev

New tutorials, open-source projects, and deep dives on coding agents - delivered weekly.

One email per weekReal code, not theoryFree forever