clevis@tinyweights:/home$ cat ./welcome.txt
> Small language models like Gemma, Phi, SmolLM, and Qwen, run locally. Benchmarks, quantization, and hands-on deployment guides for small LLMs on real hardware. > New releases tested the week they drop. Straight talk on what is worth running and what is not. 46 posts · last updated 2026-07-19 · all writing CC BY 4.0
clevis@tinyweights:/home$ ls -lh --sort=time
05-20 8min guides Running LLMs on Raspberry Pi 5: A Practical Guide with Real Benchmarks 05-19 7min guides The Complete Guide to Running Small LLMs on Apple Silicon (2026) 05-18 6min guides How to Run Phi-4-mini Locally: Microsoft's 3.8B Model with 128K Context 05-17 8min guides How Much RAM Do You Actually Need to Run Local LLMs? 05-15 8min comparisons GGUF vs ONNX vs MLX: Which Model Format Should You Use for Local Inference? 05-14 12min comparisons Ollama vs LM Studio vs llama.cpp: Which Local AI Runtime Should You Use? 05-14 13min comparisons The Best Small Language Models in 2026: A Practical Comparison 05-13 8min news Qwen3.5-0.8B: A Multimodal Thinking Model That Fits in 1 Gigabyte 05-11 8min news Qwen3-Coder-Next: Run a Frontier-Level Coding Agent Locally on Consumer Hardware 04-05 6min news Gemma 4: Taking Agentic Workflows to the Edge