/
Aug 4, 2026
Plus more about Memory Foundation Model, Weak-to-Strong OPD, and LeRoPE
Jul 29, 2026
plus more about Hilbert Operator for Progressive Encoding, RLVR-Native Optimization Stack, Soap & Muon At Scale, and Measuring Reward-Seeking
Aug 11, 2026
plus more about Leanstral, Towards Physics of Multimodal Pretraining, and loss does not see the basis
Subscribe to our newsletter
Jul 21, 2026
plus more about Concurrent Image Understanding and Generation, Latent and Explicit Reasoning with Looped Transformers and more
Jun 17, 2025
Plus more about "The Diffusion Duality" and "Reinforcement Pre-Training"
Jun 3, 2025
Read about "Rethinking Training Signals in RLVR", why LLMs are headless chickens, and "Learning to Reason without External Rewards"
Premium Insights
Jun 6, 2025
Premium Insights: A recap of popular AI research papers and research trends in May 2025
May 19, 2025
Premium Insights: A closer look into the DeepSeek Prover series
Jan 28, 2025
Plus more about Transformer2 and Kimi k1.5
Jul 14, 2026
Plus more about Single Async Opt for Agentic RL, Remember When It Matters, and Sparse Delta Memory
Jul 7, 2026
plus more about AdaJEPA, Program-as-Weights, The World Is In Your Mind, and Dual On-policy Distillation
Jun 30, 2026
plus more about Tapered LMs, Improved LLDMs, AutoData, and You Don't Need To Run Every Eval
Jun 23, 2026
plus more about Looped World Models, Fixed-Point Reasoners, and ExpRL