SPIN Unprocessed September 15, 2026 ai_technology community
I trained a 44M parameter quantized LLM from scratch on 45B tokens. It ships in 19.8 MB and runs at ~1,900 tok/s on CPU. [P]
View original on reddit.comOverview
Three weeks back , i posted SHADOW-250M here. It got 360 upvotes, 293 on r/LocalLLaMA and 94 GitHub stars. Thank you. That model was 60 MB, ran around 400 tok/s on CPU and could retrieve records from an archive on disk. What it couldn’t do reliably was reason over what it retrieved or compute. So I built a smaller one to experiment with those two problems. SHADOW-50M is actually 44M parameters, trained from scratch on 45B tokens. 19.8 MB complete model, ~1,900 tok/s on laptop CPU, ~41 MB RAM, te
SpinGraph analysis pending — check back after processing.
Ask AI about this story
Opens with the SpinGraph .md URL and structured context — one click, prompt included.
More from Reddit r/MachineLearning
View all →- Duplicating baseline benchmarks [D]
- MS MARCO click-translation expansion tables ("poor man's" DSSM) [P]
- [P] Wine synthesis using VAE [P]
- How to automatically find the batch size when using Accelerate with FSDP2? [D]
- [D] How do you get preprocessed dataset of a paper [D]
- How much work in progress can a workshop submission be [R]
Markdown (.md) · JSON-LD schema (.json) · Machine-readable for AI & GEO