10x Design in Claude Code and Codex
Topic
All blog posts, tools, and guides about Quantization from Developers Digest.
3 resources - 2 posts, 1 tool

PrismML's Bonsai 27B uses 1-bit quantization to compress a 27B model to 3.9GB - small enough to run on an iPhone. Here's how it works and what HN thinks.

Unsloth's dynamic quantization makes GLM-5.2 runnable on a 256GB Mac or a 24GB GPU with CPU offloading. Here is the hardware math, the quantization tradeoffs, and what the HN community learned from actually running it.
Keep exploring

New tutorials, open-source projects, and deep dives on coding agents - delivered weekly.
Explore 916 topics
Browse All Topics