/
Aug 11, 2026
plus more about Leanstral, Towards Physics of Multimodal Pretraining, and loss does not see the basis
Aug 4, 2026
Plus more about Memory Foundation Model, Weak-to-Strong OPD, and LeRoPE
Aug 19, 2026
Plus more about Full-Bandwidth Transformer, Small Scaling Law, and SFT Conflicts RL Coexists
Subscribe to our newsletter
Jul 29, 2026
plus more about Hilbert Operator for Progressive Encoding, RLVR-Native Optimization Stack, Soap & Muon At Scale, and Measuring Reward-Seeking
Jun 17, 2025
Plus more about "The Diffusion Duality" and "Reinforcement Pre-Training"
Jun 3, 2025
Read about "Rethinking Training Signals in RLVR", why LLMs are headless chickens, and "Learning to Reason without External Rewards"
Premium Insights
Jun 6, 2025
Premium Insights: A recap of popular AI research papers and research trends in May 2025
May 19, 2025
Premium Insights: A closer look into the DeepSeek Prover series
Jan 28, 2025
Plus more about Transformer2 and Kimi k1.5
Jul 21, 2026
plus more about Concurrent Image Understanding and Generation, Latent and Explicit Reasoning with Looped Transformers and more
Jul 14, 2026
Plus more about Single Async Opt for Agentic RL, Remember When It Matters, and Sparse Delta Memory
Jul 7, 2026
plus more about AdaJEPA, Program-as-Weights, The World Is In Your Mind, and Dual On-policy Distillation
Jun 30, 2026
plus more about Tapered LMs, Improved LLDMs, AutoData, and You Don't Need To Run Every Eval