NAME
Posts tagged "local-llm" - wada-dev
APROPOS
- 2026-08-15 Qwen3.8 27B Measured on an RTX 5090 — and Why vLLM Didn't Work
- 2026-08-15 Only the Design of My Eval Harness Survived
- 2026-07-25 A 3x Faster Local LLM That Can't Finish the Job — Two Days of MoE vs Dense on Real Tasks
- 2026-07-08 My Home Server Kept Freezing (Part 5) — The Spill Cliff: tok/s Is Set by What Overflows
- 2026-06-21 Running DiffusionGemma Locally on an RTX 5090 with vLLM: The One-Time Pioneer Cost of Bleeding-Edge Model × Bleeding-Edge GPU