clevis@tinyweights:/home$ cat ./welcome.txt
> Small language models like Gemma, Phi, SmolLM, and Qwen, run locally. Benchmarks, quantization, and hands-on deployment guides for small LLMs on real hardware.
> New releases tested the week they drop. Straight talk on what is worth running and what is not.
46 posts · last updated 2026-07-19 · all writing CC BY 4.0
clevis@tinyweights:/home$ ls -lh --sort=time
05-20
8min
guides
Running LLMs on Raspberry Pi 5: A Practical Guide with Real Benchmarks
05-19
7min
guides
The Complete Guide to Running Small LLMs on Apple Silicon (2026)
05-18
6min
guides
How to Run Phi-4-mini Locally: Microsoft's 3.8B Model with 128K Context
05-17
8min
guides
How Much RAM Do You Actually Need to Run Local LLMs?
05-15
8min
comparisons
GGUF vs ONNX vs MLX: Which Model Format Should You Use for Local Inference?
05-14
12min
comparisons
Ollama vs LM Studio vs llama.cpp: Which Local AI Runtime Should You Use?
05-14
13min
comparisons
The Best Small Language Models in 2026: A Practical Comparison
05-13
8min
news
Qwen3.5-0.8B: A Multimodal Thinking Model That Fits in 1 Gigabyte
05-11
8min
news
Qwen3-Coder-Next: Run a Frontier-Level Coding Agent Locally on Consumer Hardware
04-05
6min
news
Gemma 4: Taking Agentic Workflows to the Edge