Skip to content
UtilityHub Logo
UtilityHub
Model Launch 5 min read

I trained a 44M parameter quantized LLM from scratch on 45B tokens. It ships in 19.8 MB and runs at ~1,900 tok/s on CPU.

Source: r/MachineLearning
September 15, 2026 · 8h ago
Visual for I trained a 44M parameter quantized LLM from scratch on 45B tokens. It ships in 19.8 MB and runs at ~1,900 tok/s on CPU.

Summary

Three weeks back , i posted SHADOW-250M here. It got 360 upvotes, 293 on r/LocalLLaMA and 94 GitHub stars.

Why It Matters

New model releases reshape what developers can build. As capabilities advance, existing applications can be upgraded and entirely new use cases become feasible — making it critical for engineers to evaluate these releases early.

#llm

Related News