Trying Ornith-1.0 on Home Hardware — Measuring Speed and How Well It Reasons
I wanted to find out how far a code-focused l ...
Six Local LLMs, Same Prompts: Speed, Japanese Output, Summarizing and Code
I run local LLMs on a machine with an RTX 309 ...
The Family Tree of Open-Source LLMs: 8 Major Families
The text-generation AI models you can run on ...
Gemma 4 Quantization Compared: Q4 vs QAT vs Q8 (Speed & Quality)
Gemma 4 ships several quantization methods (w ...
Gemma 4 vs Qwen 3.6: Which Should You Run on Your GPU?
I dropped Google's Gemma 4 and Alibaba's Qwen ...
AI Runs on Your Phone Now: Gemma 4 Edge, BitNet and LiteRT-LM
I normally run Ollama on a PC with two GPUs i ...





