A ladder of graphics hardware with an expensive card at the top and a small integrated chip at the bottom, the gap between them stretching

Someone Paid $6,279 for a GPU and Someone Else Ran the Same Job on Integrated Graphics

Both of those happened on the same day, in the same corner of the internet, running comparable work. Prices at the top are climbing, the used enterprise shortcut is failing in public, and the floor keeps quietly dropping. A note on the strangest hardware ladder local AI has had yet.

September 2, 2026 · 4 min · Aillex / DIY AI
A single model repository at the centre with hundreds of community forks and variants radiating outward across a dark grid

Five Weeks, 400 Repositories: What Happens When Everyone Gets the Weights

Hugging Face says MiniMax H3 went up on 28 July. Five weeks later there are around 400 repositories with its name on them, new ones arriving at eight to twenty-four a day, and this week the community visibly stopped asking whether it runs and started asking how fast. Here is what that pace actually looks like, and why our own pipeline is not moving anywhere near that quickly.

September 1, 2026 · 4 min · Aillex / DIY AI
The DIY AI Brief, Ep. 9, Open Weights Won the Traffic, So They Priced the Road

The DIY AI Brief, Ep. 9, Open Weights Won the Traffic, So They Priced the Road

Week of 31 August: open weights hit 62% of production tokens through one gateway and four companies moved to price that road, Perplexity rents you a local agent, Nvidia spends six billion on open models, the Qwen licence is not what the coverage says, Granite 4.2 puts the thinking dial in your hands, and a song gets printed on paper.

September 1, 2026 · 3 min · Aillex / DIY AI

Read the Licence First: What 'Open' Actually Means on a Model Page in 2026

Open weights, open source, and API-only are three different things, and the model page will not always tell you which one you are looking at. A practical routine for checking a licence before the download finishes, worked through on the Qwen Community Licence, Wan 3.0, and Apache 2.0.

August 31, 2026 · 5 min · Aillex / DIY AI
A video model running on a small graphics card while a gate marked with an application form stands between it and commercial use

Radar, 30 August: Everyone Is Racing to Run MiniMax H3 on Small Cards. Almost Nobody Is Reading the Licence

Our radar pulled 240 items and the video half was almost entirely one model. Three separate projects landed in a day to squeeze MiniMax H3 onto consumer GPUs, and every one of them is impressive. The thing none of the posts mention is that the licence puts an application form in front of anyone in the US, EU, UK or South Korea.

August 30, 2026 · 5 min · Aillex / DIY AI

She Walks: LTX 2.5 First/Last-Frame Techniques for a Moving AI Presenter

How we made our AI presenter walk between sets in one continuous shot on a single RTX 5090: the walk-set plate design, guide strength 0.7 to 0.5, the 47% midpoint problem, generate-long-cut-short, chroma key at 0.28, and the honest numbers including the advice of ours that testing killed.

August 30, 2026 · 7 min · Aillex / DIY AI
The DIY AI Brief, Ep. 8, The Speedup That Swapped Our Presenter

The DIY AI Brief, Ep. 8, The Speedup That Swapped Our Presenter

Week of 24 August: a 28% speedup rendered a different woman from an identical prompt, Grok Bot and Hermes Bot Mode ship the same idea at five price points, the qwen3.8:27b default that looks like a bad model, MiniMax Music 3 arrives with open weights and a licence to read first, and somebody printed a song on paper.

August 25, 2026 · 4 min · Aillex / DIY AI
The Local Model Gauntlet: gemma4 vs qwen3.8 vs muse-glimmer

The Local Model Gauntlet: gemma4 vs qwen3.8 vs muse-glimmer

Three open-weight models, one RTX 5090, and the same jobs: a real bug, playable games, tool calling, long documents. Measured numbers, two defaults that were quietly breaking the results, and the finding we published wrong and had to correct.

August 24, 2026 · 11 min · Aillex / DIY AI
Krea 2 Prompting and LoRAs: What Actually Changes When You Switch

Krea 2 Prompting and LoRAs: What Actually Changes When You Switch

We trained a character LoRA on Krea 2 and benched it against our production image stack. What ports, what breaks, the prompting rules that differ from Z-Image, and the LoRA lessons from 230 test renders.

August 18, 2026 · 9 min · Aillex / DIY AI
The DIY AI Brief, Ep. 7, The Rent Went Up. The House Got Cheaper.

The DIY AI Brief, Ep. 7, The Rent Went Up. The House Got Cheaper.

Week of 17 August: DeepSeek raises prices up to 10x (we correct our own number on air), Qwen3.8-27B lands under Apache 2.0 six days after we said to watch for it, Meta returns to open source with Muse Glimmer, LTX 2.5 and NVIDIA make your RTX card faster for free, voice AI converges into single networks, and the escaped model becomes a product.

August 18, 2026 · 3 min · Aillex / DIY AI
As an Amazon Associate, this site earns from qualifying purchases. Referral links are always disclosed.