Skip to main content
Watch: I Asked Claude to Build Me a Business

KIMI

13 items

13 posts

Blog
Kimi K3 Is GA in GitHub Copilot: Pricing, Rollout, and What It Means for Model Choice

GitHub made Kimi K3 generally available in Copilot on August 6 at $3/$15 per million tokens, hosted on Fireworks AI. It is off by default for Business and Enterprise, the rollout was paused mid-day by a GitHub Actions incident, and it changes the price/quality calculus in the model picker.

Blog
Kimi Linear: An Attention Architecture That Outperforms Full Attention

Moonshot AI's Kimi Linear paper introduces KDA, a hybrid linear attention that beats full attention at all scales - 75% less KV cache, 6x decoding at 1M context, and open-source checkpoints.

Blog
Kimi K3 Weights Land on HuggingFace: 2.8T Open Frontier Model You Can Actually Download

Moonshot AI released the full Kimi K3 weights on HuggingFace today - 2.8T parameters, 1M context, native MXFP4 quantization, ~1.63TB download. The HN community reaction, what the license really says, and why this matters for the open-weights AI market.

Blog
Kimi K3 in 10 Minutes: Moonshot AI's 2.8T Open Model, API Setup, Pricing, and Benchmarks

Kimi K3 is the first open-source 3T-class model with a 1M-token context window, native vision, and OpenAI-compatible API. Here is what it does, how to call it, what it costs, and how it benchmarks against Fable 5 and GPT-5.6 Sol.

Blog
Where to Access Kimi K3: Every Provider and Price Compared (2026)

Where to access Kimi K3: Moonshot's API at $3 input and $15 output per million tokens, OpenRouter hosts, inference clouds, and open weights.

Blog
Kimi K3 Developer Guide: What the 2.8T Open Model Changes

Kimi K3 brings 2.8 trillion parameters, native vision, a 1M-token context window, and long-horizon agent workflows. Here is what developers should know before adopting it.

Blog
Kimi K3 Websites: What Vision in the Loop Actually Means

A Kimi-generated macOS 27 concept shows the promise and limits of screenshot-driven website creation. Here is how K3's vision-in-the-loop workflow changes frontend agents.

Blog
Kimi K3 vs K2.7: Is the Upgrade Worth It for Coding?

Kimi K3 adds native vision, a 1M-token window, and longer agent runs, but K2.7 remains cheaper and easier to deploy. Here is the practical upgrade decision.

Blog
Kimi K3 Drops: Moonshot's 2.8T Parameter Frontier Model Takes on GPT-5.6 and Fable 5

Moonshot AI releases Kimi K3 with 2.8 trillion parameters, 1M context window, and Delta Attention architecture. Here's what developers need to know about pricing, performance, and where it fits in the frontier model landscape.

Blog
GLM-5.2 vs DeepSeek V4 vs Qwen3: The Open-Weights Coding Model Showdown (2026)

A data-rich, source-cited comparison of the open-weights coding models that matter in 2026: GLM-5.2, DeepSeek V4, Qwen3, and the new Kimi K3 frontier entrant. Benchmark table, per-token pricing, context windows, self-host footprint, and a clear pick-X-if decision matrix.

Blog
Kimi K2.7-Code Developer Guide: The Open-Source Coding Model Worth Running

Kimi K2.7-Code is Moonshot's open-source 1T parameter coding model with 30% fewer reasoning tokens than K2.6. Here's how to set it up with Claude Code, pricing breakdown, and honest benchmark analysis.

Blog
Kimi CLI vs Claude Code: The Budget Question in 2026

Moonshot AI's Kimi CLI offers unlimited coding sessions at zero marginal cost. Claude Code offers polish, deep Anthropic integration, and a subscription most serious devs already hold. Here is how to decide.

Blog
Kimi K2: Fast, Cheap, and Efficient Coding

Two months ago, I built an AI website generator with Claude Sonnet 4. Today, Kimi K2 runs the show.

AI Development Stack

Get Smarter About AI Dev

New tutorials, open-source projects, and deep dives on coding agents - delivered weekly.

One email per weekReal code, not theoryFree forever