All Posts
/
Nov 27, 2025
Plus more on Seer, Virtual Width Networks, SAM 3, and Evolution Strategies at the Hyperscale
Nov 18, 2025
LeJEPA, The Path Not Taken, and more
Nov 11, 2025
From Memorization to Reasoning in the Spectrum of Loss Curvature and Introducing Nested Learning: A new ML paradigm for continual learning
Nov 4, 2025
and more on Kimi Linear, Looped Transformer, How FP16 fixes RL...
Oct 29, 2025
How to Compress Long Text into Images To Reduce LLM Tokens and more
Oct 21, 2025
RLM, RAE, Reasoning with Sampling, and more
Oct 14, 2025
Plus more about Moloch's Bargain: Emergent Misalignment When LLMs Compete for Audiences and LLM Fine-Tuning Beyond Reinforcement Learning
Oct 7, 2025
Plus more about Polychromic Objectives for Reinforcement Learning and Stochastic activations
Sep 30, 2025
Plus more about Thinking Augmented Pre-training and Reinforcement Learning on Pre-Training Data