I packed 16 GB of GGUF quants into 1.8 GB, losslessly
adamdanielsson1
11 points
6 comments
July 06, 2026
Related Discussions
Found 5 related stories in 297.1ms across 14,238 title embeddings via pgvector HNSW
- What's in a GGUF, besides the weights – and what's still missing? bashbjorn · 120 pts · May 14, 2026 · 56% similar
- Show HN: Quant Picker – which GGUF file fits your model and machine ermantrout · 15 pts · June 13, 2026 · 51% similar
- Lossless model compression experiment: GLM-5.2 in 25% less memory hambandit · 16 pts · July 20, 2026 · 49% similar
- NanoGPT Slowrun: 10x Data Efficiency with Infinite Compute sdpmas · 122 pts · March 19, 2026 · 46% similar
- Lossless GIF recompression via exhaustive search ZacnyLos · 61 pts · June 23, 2026 · 46% similar
Discussion Highlights (3 comments)
v3ss0n
for disk only
sarjann
Seems pretty useful, often hopping between different quants. I wonder if this would work for different "branches" of a model. E.g. qwen 3.6 35b regular vs abliterated.
VASTL
No offense, but do you really you "cracked the code" if multi-billion/trillion companies cant achieve that? I heavily doubt, you have losless quality with those 25 commits on GH, most likely vibecoded.