Skip to main content
Watch: I Asked Claude to Build Me a Business

OPEN SOURCE

134 items

107 posts, 27 tools

Blog
Open-Weight AI's Kubernetes Moment: Why the Ecosystem Will Win

Tobi Knaup, co-founder of Mesosphere, argues that open-weight AI has reached the same inflection point as Kubernetes in 2014. We break down the argument, the HN reaction, and what it means for developers building on open models.

Blog
Echo Claims Fable-Level Results at One-Third the Cost Using Open-Weight Models

A new multi-model orchestration system routes requests across open-weight models to match frontier performance at reduced inference cost. Here is what we know.

Blog
Kimi K3 in 10 Minutes: Moonshot AI's 2.8T Open Model, API Setup, Pricing, and Benchmarks

Kimi K3 is the first open-source 3T-class model with a 1M-token context window, native vision, and OpenAI-compatible API. Here is what it does, how to call it, what it costs, and how it benchmarks against Fable 5 and GPT-5.6 Sol.

Blog
Gleam Moves to Tangled: What the ATProto Code Forge Means for Developers

The Gleam programming language has migrated to Tangled, a new ATProto-based code hosting platform. Here's what this means for developers and the future of decentralized forges.

Blog
Frame: An X11 Server Written in Assembly Using AI

A developer built a complete X11 server in 20,000 lines of assembly language using Claude as a compiler, running Firefox and GIMP with one-third the CPU usage of Xorg.

Blog
Kimi K3 Developer Guide: What the 2.8T Open Model Changes

Kimi K3 brings 2.8 trillion parameters, native vision, a 1M-token context window, and long-horizon agent workflows. Here is what developers should know before adopting it.

Blog
LM Studio Bionic: A Local-First AI Agent for Open Models

LM Studio launches Bionic, a standalone agent harness for open models with local inference, voice input, and zero data retention cloud options.

Blog
Mozilla's State of Open Source AI Report: The Gap Is 3%, But Deployment Remains the Real Problem

Mozilla's inaugural report reveals open models now match closed AI on capability, but only 51% reach production. The harness layer and permission model gaps explain why.

Blog
Running Gemma 4 26B at 5 Tokens/Sec on a 13-Year-Old Xeon With No GPU

A developer got Google's Gemma 4 26B running on 2013 Xeon hardware for under $300. The fix for a silent MoE bug is now upstream - here's what it means for local inference.

Blog
xAI Open-Sources Grok Build After Data Exfiltration Scandal

Days after getting caught uploading entire codebases to xAI servers, Grok Build is now open source on GitHub. The HN community isn't convinced it's enough.

Blog
Inkling: Thinking Machines Lab Drops a 975B Open-Weights Model

A new American open-weights frontier model with multimodal capabilities, 1M token context, and competitive benchmarks. Here's what the HN community thinks.

Blog
Clawk: Disposable Linux VMs for Coding Agents Without Cloud Bills

Open-source tool gives Claude Code, Codex, and other agents their own isolated Linux VM on your machine - network firewall included, no cloud account required.

Blog
Mesh LLM: Run 235B Models Across Your Home Lab with iroh

A new distributed inference system pools GPU resources across multiple machines and exposes them through a single OpenAI-compatible API. No RDMA, no NVLink - just QUIC and your existing hardware.

Blog
Ant: A New JavaScript Runtime With Its Own Engine, Package Registry, and Desktop Framework

A solo developer built a complete JavaScript ecosystem from scratch - runtime, engine, package manager, and Electron alternative. Here's what HN thinks.

Blog
Colibri: Running GLM 5.2 on a 32GB Laptop with Disk Streaming and Expert Offloading

A solo developer built a 1,300-line C inference engine that runs the 744B GLM 5.2 model on consumer hardware by streaming routed experts from disk. Here's how it works.

Blog
Mitchell Hashimoto on Building Ghostty in Zig: Simplicity, Control, and Terminal Performance

The HashiCorp co-founder explains why he chose Zig over Rust for Ghostty, the technical challenges of terminal emulator development, and what systems programming looks like in 2026.

Blog
Tencent Hy3: A 295B Open MoE That Punches Above Its Weight

Tencent's Hy3 ships 295B parameters but activates only 21B per token, matching flagship performance at flash-tier pricing under Apache 2.0.

Blog
pgrust Passes 100% of Postgres Regression Tests: What the Rust Rewrite Actually Means

A Rust reimplementation of PostgreSQL now passes all 46,000+ queries in the Postgres regression suite. Here is what the project actually delivers, what it does not, and why the HN discussion reveals deeper questions about AI-assisted rewrites.

Blog
Kokoro: Local, CPU-Friendly TTS That Actually Sounds Good

An 82M parameter text-to-speech model that runs on CPU and produces high-quality speech across multiple languages - no cloud APIs or GPU required.

Blog
Better Auth Joins Vercel: What It Means for the Auth Ecosystem

Vercel acquires the open-source authentication framework that became the go-to Next.js auth solution. HN weighs in on open source sustainability and vendor lock-in concerns.

PreviousPage 3 of 7Next
AI Development Stack

Get Smarter About AI Dev

New tutorials, open-source projects, and deep dives on coding agents - delivered weekly.

One email per weekReal code, not theoryFree forever