SecondSourceJudgment rebuilt from primary sources
Product & chips · Aug 16, 2026

AMD's attack on the CUDA moat has moved forward, from persuading data centres to switch to making sure developers start on its side.

Product moves · Trend watch (event dates 07-20 and 07-23)

From the Aug 16, 2026 daily brief

Evidence on AMD's software ecosystem used to cluster around large inference frameworks; the past two weeks produced two independent sources pointing somewhere else. From the edge segment of AMD's July 23 annual event: the next-generation Halo development platform will offer 192GB of memory in a very small box and has secured direct support from the open model hub Hugging Face, and every Halo machine ships with a year of Hugging Face Pro (Ian Cutress's live relay). On July 20, the open-source fine-tuning toolchain Unsloth announced an official collaboration with AMD, bringing training and inference to more than 500 models and claiming training twice as fast on 70% less memory (Unsloth co-founder Daniel Han). Those last two numbers are vendor-reported, with no baseline attached to either. Read together, they say this: what decides an ecosystem is often not the large deployments but the long-tail entry point — the first notebook a graduate student opens when they want to fine-tune a model. What to watch to test this is whether AMD hardware's coverage list on open-source fine-tuning toolchains and model hubs keeps getting longer, rather than which new data centre customers AMD announces. ⚠️ Neither number has a comparison baseline (twice as fast as what, 70% less memory relative to which configuration), and all of this sits in inference and fine-tuning, not large-scale pre-training. On the moat around core training workloads, these two data points say nothing at all.

Subscribe free — first issue lands tomorrow morning

Just an email address, unsubscribe anytime. This is the only thing we ask of you.

More in this section