Papers and technical work — only the ones that move an industry judgment.
19 pieces
The current mainstay of post-training (the stage after base training where a model is taug…
In the same AMD assessment, SemiAnalysis reports firsthand engineering observations: 2.5 e…
The Jacobian conjecture, posed in 1939, is a famous problem in algebraic geometry. It says…
Our July 21 Research Notes covered this empirical study (gains from optimizing an agent pi…
Alex Zhang, author of the RLM framework paper, argues that carefully designed task orchest…
Raia Hadsell, VP of Research at Google DeepMind, argued in a talk that the recipe that bui…
An empirical study of 25,264 pull requests (code-change proposals) opened by AI agents acr…
Three mainstream "don't touch the model, optimize the agent workflow" methods went through…
François Chollet, creator of the Keras framework, once gave the claim "large models memori…
Gwern, the anonymous independent researcher known for his early systematic case for the sc…
The conclusion first: evaluating AI with AI-generated data has a structural blind spot — t…
For problems with no answer key, sample many solutions from the model, take the majority-c…
Yesterday's July 18 issue introduced MemCon from a UCLA-affiliated team in these research …
A Meta FAIR-affiliated team's March paper, Principia, argues that current math benchmarks …
Oracle, in a 13-author technical report on July 14, builds agent memory as a database-nati…
A paper from the team that includes Turing Award winner Yoshua Bengio shows that an AI age…
Practitioners reading Thinking Machines' official model card for Inkling flagged a rare se…
Databricks benchmarked models on engineering tasks against its own multi-million-line code…
The honest note on the academic side: this week's academic main course is the five peer-re…
Just an email address, unsubscribe anytime. This is the only thing we ask of you.