Skip to main content
Watch: I Asked Claude to Build Me a Business

AI MODELS

112 items

111 posts, 1 guide

Blog
Gemini Robotics 2: Google DeepMind Brings Whole-Body Intelligence to Humanoid Robots

Google DeepMind's Gemini Robotics 2 family gives humanoid robots whole-body control, dexterous hands, and multi-robot teamwork - with an ER 2 model devs can try today. The HN thread (575 points, 459 comments) debated how real the progress is.

Blog
OpenAI Cuts GPT-5.6 Luna by 80%: The Price-Performance Frontier Just Shifted

OpenAI slashes GPT-5.6 Luna by 80% to $0.20/M input tokens, cuts Terra by 20%, adds Sol Fast mode at 2.5x speed, and reveals Sol autonomously optimized its own production kernels.

Blog
Grok 4.5 in 10 Minutes: xAI's Fastest Model, 500K Context, and Build-Mode Integration

A companion guide to the Grok 4.5 video: xAI's most intelligent model with a 500K context window, function calling, structured outputs, and a build-mode agent workflow for developers.

Blog
Fable 5 Effort Levels vs Switching Models: When to Dial and When to Change

Effort levels and model choice both cost more for more capability, but they are not interchangeable. Here is when to move the effort dial and when to switch models instead.

Blog
Kimi K3 Weights Land on HuggingFace: 2.8T Open Frontier Model You Can Actually Download

Moonshot AI released the full Kimi K3 weights on HuggingFace today - 2.8T parameters, 1M context, native MXFP4 quantization, ~1.63TB download. The HN community reaction, what the license really says, and why this matters for the open-weights AI market.

Blog
Claude Opus 5: Near-Fable Intelligence at Half the Cost

Anthropic released Opus 5 on July 24, 2026 - same price as Opus 4.8, within 0.5% of Fable 5 on CursorBench, and the new #1 on Artificial Analysis. We break down the benchmarks, HN reaction, and what it means for every developer choosing a daily-driver model.

Blog
Claude Opus 5 in 8 Minutes: What Developers Need to Know

Claude Opus 5 ships today with Frontier-Bench SOTA, near-Fable-5 coding at half the price, and self-verification that catches its own bugs. Here is what changed, what to migrate, and when the price-performance curve makes Opus 5 the right default.

Blog
Echo Claims Fable-Level Results at One-Third the Cost Using Open-Weight Models

A new multi-model orchestration system routes requests across open-weight models to match frontier performance at reduced inference cost. Here is what we know.

Blog
FLUX 3: Black Forest Labs Ships a Unified Multimodal Foundation Model for Image, Video, Audio, and Robotics

Black Forest Labs released FLUX 3, a single multimodal model trained jointly on images, video, and audio that also drives robots on Audi production lines. Here is what it does, how it works, and how to try it.

Blog
Kimi K3 in 10 Minutes: Moonshot AI's 2.8T Open Model, API Setup, Pricing, and Benchmarks

Kimi K3 is the first open-source 3T-class model with a 1M-token context window, native vision, and OpenAI-compatible API. Here is what it does, how to call it, what it costs, and how it benchmarks against Fable 5 and GPT-5.6 Sol.

Blog
Where to Access Kimi K3: Every Provider and Price Compared (2026)

Where to access Kimi K3: Moonshot's API at $3 input and $15 output per million tokens, OpenRouter hosts, inference clouds, and open weights.

Blog
Kimi K3 Developer Guide: What the 2.8T Open Model Changes

Kimi K3 brings 2.8 trillion parameters, native vision, a 1M-token context window, and long-horizon agent workflows. Here is what developers should know before adopting it.

Blog
Kimi K3 vs K2.7: Is the Upgrade Worth It for Coding?

Kimi K3 adds native vision, a 1M-token window, and longer agent runs, but K2.7 remains cheaper and easier to deploy. Here is the practical upgrade decision.

Blog
Kimi K3 Drops: Moonshot's 2.8T Parameter Frontier Model Takes on GPT-5.6 and Fable 5

Moonshot AI releases Kimi K3 with 2.8 trillion parameters, 1M context window, and Delta Attention architecture. Here's what developers need to know about pricing, performance, and where it fits in the frontier model landscape.

Blog
Inkling: Thinking Machines Lab Drops a 975B Open-Weights Model

A new American open-weights frontier model with multimodal capabilities, 1M token context, and competitive benchmarks. Here's what the HN community thinks.

Blog
Bonsai 27B: How PrismML Fit a 27 Billion Parameter Model on Your Phone

PrismML's Bonsai 27B uses 1-bit quantization to compress a 27B model to 3.9GB - small enough to run on an iPhone. Here's how it works and what HN thinks.

Blog
Claude Fable 5 in 7 Minutes: Benchmarks, Pricing, Availability, and Real-World Examples

A companion guide to the Claude Fable 5 video: what the first general-use Mythos class model is, the walkthrough beats from the review, hands-on developer takeaways, and the pricing and context specs from primary sources.

Blog
GPT-5.6 vs Claude 5: What the New Tiers Mean for Choosing a Coding Model

OpenAI's GPT-5.6 Sol, Terra, and Luna tiers versus Anthropic's Claude Fable 5 and Mythos 5. Verified pricing, benchmarks, and a practical framework for picking a coding model in July 2026.

Blog
Grok 4.5 for Developers: What Changed and When to Pick It

xAI's Grok 4.5 ships at $2/$6 per million tokens with 80 TPS speeds, a 500k context window, and benchmark results that put it in the Opus and GPT 5.5 tier. What actually shipped, how the pricing compares, and when it makes sense over Claude, GPT, or Gemini.

Blog
Tencent Hy3: A 295B Open MoE That Punches Above Its Weight

Tencent's Hy3 ships 295B parameters but activates only 21B per token, matching flagship performance at flash-tier pricing under Apache 2.0.

PreviousPage 3 of 6Next
AI Development Stack

Get Smarter About AI Dev

New tutorials, open-source projects, and deep dives on coding agents - delivered weekly.

One email per weekReal code, not theoryFree forever