Skip to main content
Watch: I Asked Claude to Build Me a Business

AI

176 items

103 posts, 62 tools, 11 guides

Blog
Weekly Highlights: The Price War Reached the Frontier

The 5-7 AI developer stories that actually mattered this week - ranked, linked, and cut for builders.

Blog
Weekly Highlights: Cheaper Tokens, Higher Stakes

The 7 AI developer stories that actually mattered this week - ranked, linked, and cut for builders.

Blog
Weekly Highlights: Cheaper Agents, Harder Questions

The 7 AI developer stories that actually mattered this week - ranked, linked, and cut for builders.

Blog
TutorMoments: AI2's New Benchmark Shows LLM Tutors Over-Help by Default

AI2 released TutorMoments, a replay-based benchmark that drops seven LLMs into real math tutoring transcripts and scores whether they scaffold when help is needed or push for rigor when the student can do more. The default finding: models over-help, and spelling out the trade-off in the prompt lifts every score but does not close the gap to a consistent human call.

Blog
Weekly Highlights: Agents Became the Attack Surface, Open Weights Took the Agentic Lead

The 7 AI developer stories that actually mattered this week - ranked, linked, and cut for builders.

Blog
EU Forces Google to Open 11 Android Features to Third-Party AI Assistants

A final Digital Markets Act decision requires Alphabet to give third-party AI assistants the same Android access Gemini has: DSP wake words, ambient sensors, screen automation, on-device models, and fair background execution. Home Assistant's three-year fight over the 'Okay Nabu' wake word shows exactly what the ruling unlocks.

Blog
GitHub Models Is Retired: What to Use for Model Access Now

GitHub Models is fully retired as of July 30, 2026. The playground, model catalog, inference API, and BYOK are gone for every customer. Here is the timeline and where to get model access instead.

Blog
OpenAI Disrupts a Cambodia Scam Network That Ran on ChatGPT

OpenAI took down a Cambodia-based operation that used ChatGPT for personas, translations, forged documents, and admin work. It is the clearest picture yet of how LLMs slot into organized fraud.

Blog
Weekly Highlights: Frontier AI Commoditized - Half-Price Opus 5, 3T Open Weights, and Agent Security Gets Real

The 7 AI developer stories that actually mattered this week - ranked, linked, and cut for builders.

Blog
Buzz by Block: The Open-Source Workspace Where Humans and AI Agents Build Together

A companion guide to the Buzz video: Block's open-source Nostr relay workspace where humans and AI agents share the same rooms, with agent-first CLI, git integration, and workflows. Here is what it does and where it fits in the agentic dev stack.

Blog
The New AI Superpowers: Focus and Followthrough

AI makes you 2-100x faster on every task. So why are developers burning out more than ever? The HN discussion on Rick Manelius's essay surfaces a hard truth about the gap between productivity and throughput.

Blog
Terence Tao Digests the Jacobian Conjecture Counterexample: How Claude Fable 5 Broke an 87-Year-Old Math Problem

Terence Tao published a deep mathematical digestion of the Jacobian conjecture counterexample discovered by Claude Fable 5. Here is what happened, what HN is saying, and what it means for AI-assisted research.

Blog
What AI Did to Stack Overflow, Visualized in One Graph

A Stack Exchange data query shows Stack Overflow's question volume dropped 65% since 2017, with a sharp acceleration after ChatGPT. HN debates whether AI killed the platform or just accelerated its decline.

Blog
Mozilla's State of Open Source AI Report: The Gap Is 3%, But Deployment Remains the Real Problem

Mozilla's inaugural report reveals open models now match closed AI on capability, but only 51% reach production. The harness layer and permission model gaps explain why.

Blog
AI Voice Fraud Needs Three Seconds of Your Voice

Voice cloning now requires just 3 seconds of audio to impersonate someone. With $893M in reported losses, detection has failed - here's what might actually work.

Blog
Codex Now Encrypts Multi-Agent Prompts, Breaking Local Auditability

OpenAI's Codex CLI now encrypts inter-agent communications for Sol and Terra models, leaving users unable to inspect what their agents are actually doing.

Blog
Apple SpeechAnalyzer vs Whisper: Independent Benchmark Shows Apple Winning on Accuracy

New benchmarks on 5,559 test utterances show Apple's iOS 26 SpeechAnalyzer API achieving 2.12% word error rate - beating all Whisper model sizes while running 3x faster.

Blog
AI Dev News: Week of July 12, 2026

Grok 4.5 lands at $2/$6, OpenAI splits GPT-5.6 into Sol, Terra, and Luna tiers, Anthropic ships the Claude 5 family, TypeScript 7 goes native, Bun gets rewritten in Rust, and a prompt injection hits GitHub agents.

Blog
Mesh LLM: Run 235B Models Across Your Home Lab with iroh

A new distributed inference system pools GPU resources across multiple machines and exposes them through a single OpenAI-compatible API. No RDMA, no NVLink - just QUIC and your existing hardware.

Blog
Terry Tao on Coding Agents: A Fields Medalist's Take on Vibe Coding

The world's most famous mathematician used AI coding agents to revive 25-year-old Java applets and build new visualization tools. His observations on risk, quality, and trust are worth reading.

Page 1 of 9Next
AI Development Stack

Get Smarter About AI Dev

New tutorials, open-source projects, and deep dives on coding agents - delivered weekly.

One email per weekReal code, not theoryFree forever