Ciru/CrownINFERENCE LAB

DUAL STRIX HALO EXPERIMENT 24 SEP 2026 EDT

64 builds.
Two AMD desktops.
One live build storm.

Thirty-two independent apps and games per machine, generating at the same time on local AMD Ryzen AI MAX+ 395 systems. This is the complete wall: every finished artifact, ready to open and play.

POWERED BY 2 × AMD STRIX HALO CIRU-OPTIMIZED GEMMA 4 26B vLLM
THE OUTPUT WALL64 LIVE ARTIFACTS
SOZO 32CIRU 32 LOCAL INFERENCE
01 / THE RUN

A whole wall in 5:32.

One first generation pass. Two machines. Sixty-four concurrent requests.

First-pass build time 15:32Excludes repair passes
Combined peak 21,166output tokens / sec
Combined average 3662output tokens / sec
First-pass output219,988provider-reported tokens
Concurrent builds64/6432 requests on each AMD system
LIVE TELEMETRY / FIRST PASS

Two machines. One output stream.

OUTPUT TOKENS / SECOND · ROLLING 5S
CombinedSozo / 32 buildsCiru / 32 builds
MACHINE 01 / SOZO32 SLOTS

Ryzen AI MAX+ 395

Peak 2569tok/s
Average 3339tok/s

112,713 first-pass output tokens

MACHINE 02 / CIRU32 SLOTS

Ryzen AI MAX+ 395

Peak 2663tok/s
Average 3323tok/s

107,275 first-pass output tokens

1 332.44 seconds from launch through the first generation pass. 2 Highest five-second rolling output rate; machine peaks occurred at different times and should not be added. 3 First-pass output tokens divided by the same 332.44-second wall time. Rates exclude repair work. One of 64 first-pass outputs hit its token limit; the final gallery includes the later repair and publication fixes.

03 / UNDER THE HOOD

Local from prompt to pixels.

01

The challenge

Sixty-four distinct briefs launched together: games, utilities, and mini-sites. Each worker produced its own self-contained HTML, CSS, and JavaScript, with no external runtime dependencies.

02

The stack

Two AMD Ryzen AI MAX+ 395 Strix Halo desktops, 32 concurrent requests on each. Both ran Ciru-optimized Gemma 4 26B through vLLM. The live wall displayed generation and output as the run progressed.

03

What the numbers mean

The speed and 5:32 time describe only the original generation pass. Sixty-three outputs ended normally; one hit its output limit. Repairs and six small publication fixes were completed afterward and are reflected in the playable gallery, not the benchmark figures.

DISCLOSURE

AMD provided one of the two Ryzen AI Halo systems used for this content. The run, measurements, apps, and showcase were produced by Ciru/Crown.