Running a 180B-Parameter LLM on a Laptop Without a GPU: VIDRAFT's POCKET-Darwin-180B
Running a 180B-Parameter LLM on a Laptop Without a GPU: VIDRAFT's POCKET-Darwin-180B TL;DR: VIDRAFT has released POCKET-Darwin-180B, a 4-bit GGUF-quantized, llama.cpp-compatible build of their Darwin-180B-RSI frontier model that runs on consumer hardware — including CPU-only laptops and mini PCs — without requiring enterprise GPU clusters. It achieves this through sparse Mixture-of-Experts routing and graft quantization, shrinking a 360 GB BF16 model to 111 GB across just 4 files while maintaining identical MMLU-Pro scores. For engineers priced out of H100 clusters, this represents a meaningfu
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in