VIDRAFT releases POCKET-Darwin-180B for CPU-only machines
10/03/2026 — 10/03, 07:11·1 sources·1 reports
Story overview
On October 3, 2026, DEV Community's AI section reported that VIDRAFT had released POCKET-Darwin-180B, a 4-bit GGUF-quantized build of the company's Darwin-180B-RSI frontier model. According to the report, the release is compatible with llama.cpp and runs on consumer hardware, including CPU-only laptops and mini PCs, rather than the enterprise GPU clusters usually associated with models of this size. The article frames the news around running a 180B-parameter LLM on a laptop without a GPU.
The compression is attributed to two techniques working together: sparse MoE routing and graft quantization. Together they shrink the original BF16 model from 360 GB to 111 GB, packaged as only 4 files. The parameter count is 180B.
Beyond those points the material is thin. The report does not list minimum hardware requirements, expected inference throughput, memory needs, or any benchmark results for the quantized build. It also does not indicate whether POCKET-Darwin-180B arrived before or after the original Darwin-180B-RSI, and it gives no download or licensing details. Only this single report dated October 3, 2026 is available here, so the account of the announcement rests on it alone.
AI-generated from 1 reports · updated 2 hours ago
Latest turnVIDRAFT has released POCKET-Darwin-180B, a 4-bit GGUF quantized, llama.cpp-compatible build of its Darwin-180B-RSI model that runs on CPU-only laptops and mini PCs without GPU clusters. Sparse Mixture-of-Experts routing and graft quantization shrink the 360 GB BF16 model to 111 GB across four files.

Reports on this story headlines open the original
VIDRAFT has released POCKET-Darwin-180B, a 4-bit GGUF quantized, llama.cpp-compatible build of its Darwin-180B-RSI model that runs on CPU-only laptops and mini PCs without GPU clusters. Sparse Mixture-of-Experts routing and graft quantization shrink the 360 GB BF16 model to 111 GB across four files.
DEV Community · AIAI score 76
Other stories people are talking about
- 640RisingNVIDIA launches 64GB DGX Spark desktop AI computer at $4,9998 sources
- 623SurgeApple tightens macOS Full Disk Access over AI agent risks7 sources
- 463NewMeta open-sources Muse Gadgets firmware and SDK5 sources
- 270Google releases Gemini 4 Argon as next-gen frontier AI model7 sources
- 263SurgeAnthropic launches Claude Frontier Academy with $100M3 sources
- 241SurgeOpenAI DevDay 2026 launches Dots, Decisions API, and more3 sources
How is heat calculated?About the methodHide
Heat counts how many independent sources covered a story in the last 48 hours: one source counts once no matter how many posts it published, decaying with a 24-hour half-life. What ranks first is what many people are talking about.
This page aggregates public feeds. Headlines and summaries are machine-organized and remain the property of the original authors; verify important facts at the source.
- Surge
- Discussion rising fast
- New
- First report within 6 hours
- Rising
- Still gathering discussion
