Parsing Web Pages: Sending the Whole Thing to an LLM Gets Expensive — I Measured What Moving It Local Saved
When I started pulling specific facts out of ...
I Set It, So Why Is Nothing Happening? — The KV Cache Compression That Was Being Ignored
Running an LLM (a large language model, the t ...
Can You Keep MCP Entirely on Your Own Machine? I Counted 2,500 Servers | MCP Part 3
I run AI on my own machine less because it is ...
Running a 235B Model on a Mini PC — The Settings Traps on the GMKtec EVO-X2
Once a mini PC with a lot of memory arrives, ...
Local AI Speed Is Decided by Memory Bandwidth — Four Machines Measured
When you are thinking about running local AI ...
Trying Ornith-1.0 on Home Hardware — Measuring Speed and How Well It Reasons
I wanted to find out how far a code-focused l ...
OCuLink or Thunderbolt for an External GPU? I Measured Both With the Same Card | Mini PC eGPU Part 1
When you hang an external graphics card off a ...
What the Intel Arc B580 Could and Could Not Do | Intel Arc B580 Local LLM Part 7
I put a 12GB Intel Arc B580 in next to the NV ...
What Happens If You Mix GPUs From Different Vendors? RTX 3090 + Intel Arc B580 Measured | Intel Arc B580 Local LLM Part 6
This desktop has an NVIDIA GeForce RTX 3090 ( ...
Six Local LLMs, Same Prompts: Speed, Japanese Output, Summarizing and Code
I run local LLMs on a machine with an RTX 309 ...









