SecondSourceJudgment rebuilt from primary sources
Research · Aug 30, 2026

Today's entry turns on a gap.

Model watch

From the Aug 30, 2026 daily brief

One user reported that Qwen3.8-Flash broke down on multi-turn tracking at low precision, and that switching the key-value cache back to higher precision fixed it (QuixiAI, 08-29). The key-value cache holds the vectors for every preceding token, and multi-turn conversation leans on it to hold state; compressing it to low precision accumulates error, and the error stays invisible in single-turn inference and blows up in multi-turn tracking. ⚠️ This is one user's deployment experience, a single case. What it points at is general: the specs and prices at release are all single-turn measures; deployment failures show up in holding state; and no lab publishes a multi-turn stability reading at release. Anyone buying on the release numbers is missing exactly that. Specs and prices line up side by side; reliability in holding state does not. Three labs matching on specs does not mean a buyer's risk has matched too.

Subscribe free — first issue lands tomorrow morning

Just an email address, unsubscribe anytime. This is the only thing we ask of you.

More in this section