← All writing

Tag

Quantization

1 post tagged Quantization.

BitNet b1.58: What the 1-bit LLM Paper Actually Says

A 70B BitNet model fits in 7GB instead of 140GB — and the math says output quality matches FP16 at scale. The catch: you can't convert existing models. Here's what the paper actually proves, and why the hardware story matters more than the math.