The Local Model Gauntlet: gemma4 vs qwen3.8 vs muse-glimmer

The Local Model Gauntlet: gemma4 vs qwen3.8 vs muse-glimmer

Three open-weight models, one RTX 5090, and the same jobs: a real bug, playable games, tool calling, long documents. Measured numbers, two defaults that were quietly breaking the results, and the finding we published wrong and had to correct.

August 24, 2026 · 11 min · Aillex / DIY AI
Wan 2.2 local benchmark

Wan 2.2 Locally: I Benched 5B vs 14B on One Card

Everyone’s arguing about Wan 2.2. So we ran it locally, honestly, the fast-and-fits 5B, the heavy two-expert 14B, the VRAM wall, and the decode that wedges. Here are the real numbers.

July 24, 2026 · 2 min · Aillex / DIY AI
Which AI Rig Do You Need? The Honest 2026 Hardware Guide

Which AI Rig Do You Need? The Honest 2026 Hardware Guide

VRAM first, everything else second: what GPU you actually need for local AI in 2026, the memory-crisis market, the value picks, AI appliances like DGX Spark, four real builds by budget, and when renting beats buying.

July 20, 2026 · 6 min · Aillex / DIY AI
The Gaming PC That Became an AI: Our Exact Build

The Gaming PC That Became an AI: Our Exact Build

The images, the video, the 3D avatar and the daily uploads on this channel run on one prebuilt gaming PC. Here’s the exact spec, what actually matters for local AI, and what we’d change.

July 3, 2026 · 4 min · Aillex / DIY AI
Real-Time Lip-Sync with MuseTalk on an RTX 5090 (Blackwell Survival Guide)

Real-Time Lip-Sync with MuseTalk on an RTX 5090 (Blackwell Survival Guide)

MuseTalk generates lip-synced video faster than real time on a 5090, but getting it to BUILD on Blackwell is dependency hell. Here’s the exact recipe that works.

July 1, 2026 · 4 min · Aillex / DIY AI
As an Amazon Associate, this site earns from qualifying purchases. Referral links are always disclosed.