All Posts
/
Aug 25, 2026
plus more about SPADE, skill issue, and Q-learning with World Models
Aug 19, 2026
Plus more about Full-Bandwidth Transformer, Small Scaling Law, and SFT Conflicts RL Coexists
Aug 11, 2026
plus more about Leanstral, Towards Physics of Multimodal Pretraining, and loss does not see the basis
Aug 4, 2026
Plus more about Memory Foundation Model, Weak-to-Strong OPD, and LeRoPE
Jul 29, 2026
plus more about Hilbert Operator for Progressive Encoding, RLVR-Native Optimization Stack, Soap & Muon At Scale, and Measuring Reward-Seeking
Jul 21, 2026
plus more about Concurrent Image Understanding and Generation, Latent and Explicit Reasoning with Looped Transformers and more
Jul 14, 2026
Plus more about Single Async Opt for Agentic RL, Remember When It Matters, and Sparse Delta Memory
Jul 7, 2026
plus more about AdaJEPA, Program-as-Weights, The World Is In Your Mind, and Dual On-policy Distillation
Jun 30, 2026
plus more about Tapered LMs, Improved LLDMs, AutoData, and You Don't Need To Run Every Eval