120 agentic runs: the 35B coder post-trains, and the model that knew it was lying27 July 2026·11 minsLocal-Ai Qwen Llama-Cpp Benchmarking Agents Homelab
90 agentic runs, zero failures, and one invented person26 July 2026·15 minsLocal-Ai Qwen Llama-Cpp Benchmarking Agents Homelab
Strix Halo at Full Context — Why Your Decode Drops 64% and What Actually Fixes It16 May 2026·8 minsStrix-Halo Benchmarks Rocm Vulkan Llama.cpp Qwen Mtp Inference
Strix Halo LLM Serving: 25 tok/s at 151k Context Under 100W26 April 2026·7 minsAi Homelab Llm-Inference Strix-Halo Llama-Cpp Qwen Local-Ai