/
Aug 19, 2026
Plus more about Full-Bandwidth Transformer, Small Scaling Law, and SFT Conflicts RL Coexists
Aug 11, 2026
plus more about Leanstral, Towards Physics of Multimodal Pretraining, and loss does not see the basis
Aug 25, 2026
plus more about SPADE, skill issue, and Q-learning with World Models
Subscribe to our newsletter
Aug 4, 2026
Plus more about Memory Foundation Model, Weak-to-Strong OPD, and LeRoPE
Jun 17, 2025
Plus more about "The Diffusion Duality" and "Reinforcement Pre-Training"
Jun 3, 2025
Read about "Rethinking Training Signals in RLVR", why LLMs are headless chickens, and "Learning to Reason without External Rewards"
Premium Insights
Jun 6, 2025
Premium Insights: A recap of popular AI research papers and research trends in May 2025
May 19, 2025
Premium Insights: A closer look into the DeepSeek Prover series
Jan 28, 2025
Plus more about Transformer2 and Kimi k1.5
Jul 29, 2026
plus more about Hilbert Operator for Progressive Encoding, RLVR-Native Optimization Stack, Soap & Muon At Scale, and Measuring Reward-Seeking
Jul 21, 2026
plus more about Concurrent Image Understanding and Generation, Latent and Explicit Reasoning with Looped Transformers and more
Jul 14, 2026
Plus more about Single Async Opt for Agentic RL, Remember When It Matters, and Sparse Delta Memory
Jul 7, 2026
plus more about AdaJEPA, Program-as-Weights, The World Is In Your Mind, and Dual On-policy Distillation