Measured training tokens/sec for LoRA and full-SFT on the AMD Ryzen AI Max+ 395 “Strix Halo” iGPU (Radeon 8060S, gfx1151, 124 GB unified memory) under ROCm. Four model arms, five research questions — every number on this site is derived from the raw run log, not hand-typed.
Five research questions, each on its own page with the charts and the raw numbers behind it.
Chosen to span method (LoRA vs full-SFT), dense vs MoE, and a deliberate bf16/NaN probe.
Curious how a run is measured, or want to reproduce it? See Methods & reproducibility.