Skip to main content
Watch: I Asked Claude to Build Me a Business

AI

176 items

103 posts, 62 tools, 11 guides

Blog
AI 2040 Plan A: A Detailed Scenario for Navigating Superintelligence

Daniel Kokotajlo and the AI Futures Project released an ambitious 15-year roadmap for managing advanced AI development through international cooperation. Here's what HN thinks about it.

Blog
Ghost Font: Text That Humans Can Read But AI Cannot

A new experimental technology encodes messages in video using motion-based steganography, exploiting how AI models process video as individual frames rather than continuous motion.

Blog
Vector Database Comparison for RAG and AI Agents

pgvector, Pinecone, Qdrant, Weaviate, Chroma, Milvus, and Turbopuffer compared on hosting model, filtering, scale, and cost for RAG.

Blog
Kokoro: Local, CPU-Friendly TTS That Actually Sounds Good

An 82M parameter text-to-speech model that runs on CPU and produces high-quality speech across multiple languages - no cloud APIs or GPU required.

Blog
Mistral Releases Robostral Navigate: An 8B Robotics Navigation Model

Mistral's new 8B parameter model enables robots to navigate complex environments using only a camera and natural language commands. Here's what it does, how it works, and what the benchmarks actually mean.

Blog
Ilya Sutskever's 30 Papers: The Reading List That Covers 90% of What Matters

A CS student built 30papers.com to make Ilya's legendary ML reading list more accessible. HN has thoughts on the source, the format, and why compression equals intelligence.

Blog
Small AI Models Are Finding Real Users Where Networks Fail

IEEE Spectrum reports on pharmaceutical AI running on handheld devices. HN debates emergency kits, domain-specific models, and whether AGI will emerge from scaling or specialization.

Blog
AI Tutor Shows 0.71-1.30 SD Effect Size in Dartmouth Statistics Course

A new study from Dartmouth measures the impact of an AI tutoring platform on introductory statistics performance. Full engagement with the system correlated with significant exam score improvements, though selection bias remains a key limitation.

Blog
Why Price Per 1M Tokens Is a Misleading Metric for LLM Costs

Comparing LLMs by token pricing alone can lead you to choose worse, more expensive models. Cost per task tells the real story.

Blog
The Log Is the Agent: Event Sourcing Comes to AI Systems

A new paper proposes inverting traditional agent architecture - making the append-only event log the source of truth, not an afterthought. HN debates whether this is novel or just CQRS with extra steps.

Blog
Agents 101: How to Build and Deploy Anything with AI Agents

A companion guide to the Agents 101 video: a behind-the-scenes walkthrough of building and deploying AI agents fast on Vercel, the agentic infrastructure stack. Here is the map of what to learn and where to go next.

Blog
Ornith-1.0: What an Open Source Self-Improving Coding Model Actually Means

DeepReinforce AI released Ornith-1.0, a family of open-source coding models claiming self-improvement. The HN thread reveals a mix of skepticism and genuine interest - here is what the model actually does and whether the hype holds up.

Blog
Using Claude Code for a Second Opinion on MRI Scans - What Actually Happened

A developer fed 266MB of DICOM MRI data to Claude Code Opus for a second opinion on a shoulder diagnosis. The AI disagreed with the doctor. HN radiologists weighed in.

Blog
GLM 5.2 Outperforms Claude Code on Semgrep's IDOR Vulnerability Benchmarks

Semgrep's security research team benchmarked LLMs on IDOR vulnerability detection. The open-weight GLM 5.2 beat Claude Code by 7 points at roughly one-sixth the cost.

Blog
Vulnerability Reports Are Not Special Anymore

Filippo Valsorda argues that LLMs have ended the era of treating security researchers with kid gloves. When anyone can discover vulnerabilities with an AI, the old coordinated disclosure model breaks down.

Blog
Unlimited OCR: Baidu's Open-Source Solution for Long Document Parsing

Baidu releases Unlimited OCR, an open-source vision-language model that parses 100+ page documents in a single pass without memory blowup. Here's what developers need to know.

Blog
Cloudflare Now Lets AI Agents Deploy Workers Without Signup

The new wrangler deploy --temporary flag creates ephemeral Cloudflare accounts for AI agents. 60-minute deployments, no OAuth, no browser - just deploy and claim later.

Blog
LLM Architectures Got Complicated Fast

Modern LLMs now use MoE routing, mixed attention variants, and fused vision encoders. The simple transformer stack is gone - here's what replaced it and why it matters for developers.

Blog
Noam Shazeer Joins OpenAI After Two Years Back at Google

The Transformer co-creator leaves Google DeepMind for OpenAI just two years after Google paid $2.7 billion to bring him back from Character.AI.

Tool
Claude Fable 5

Anthropic's first generally available Mythos-class model, released June 9, 2026. 1M context, 128K max output, $10/$50 per million tokens. Built for long-horizon agentic work.

PreviousPage 2 of 9Next
AI Development Stack

Get Smarter About AI Dev

New tutorials, open-source projects, and deep dives on coding agents - delivered weekly.

One email per weekReal code, not theoryFree forever