clevis@tinyweights:/home$ cat ./welcome.txt
> Small language models like Gemma, Phi, SmolLM, and Qwen, run locally. Benchmarks, quantization, and hands-on deployment guides for small LLMs on real hardware. > New releases tested the week they drop. Straight talk on what is worth running and what is not. 46 posts · last updated 2026-07-19 · all writing CC BY 4.0
clevis@tinyweights:/home$ ls -lh --sort=time | grep guides
06-01 7min guides Run ZAYA1-8B Locally: The 8B Reasoning MoE Trained on AMD 05-28 7min guides How to Run LFM2.5-1.2B-Thinking Locally: On-Device Reasoning Under 1GB 05-27 9min guides GGUF Quantization Levels Explained: Q4, Q5, Q8, and IQ Quants 05-27 6min guides How to Run Jamba Reasoning 3B Locally: AI21's Hybrid SSM-Transformer With a 256K Context 05-25 9min guides How to Run a Small LLM in Your Browser with WebLLM (No Install, No API) 05-25 6min guides How to Run Ministral 3 Locally: Mistral's 3B, 8B, and 14B Vision Models 05-23 11min guides Running Local LLMs on Low-VRAM Windows GPUs (6GB and 8GB Cards) in 2026 05-21 10min guides What Can You Actually Do With a Local Small LLM? A Practical Guide 05-20 8min guides Running LLMs on Raspberry Pi 5: A Practical Guide with Real Benchmarks 05-19 7min guides The Complete Guide to Running Small LLMs on Apple Silicon (2026)