All Posts
/
Aug 4, 2026
Plus more about Memory Foundation Model, Weak-to-Strong OPD, and LeRoPE
Jul 29, 2026
plus more about Hilbert Operator for Progressive Encoding, RLVR-Native Optimization Stack, Soap & Muon At Scale, and Measuring Reward-Seeking
Jul 21, 2026
plus more about Concurrent Image Understanding and Generation, Latent and Explicit Reasoning with Looped Transformers and more
Jul 14, 2026
Plus more about Single Async Opt for Agentic RL, Remember When It Matters, and Sparse Delta Memory
Jul 7, 2026
plus more about AdaJEPA, Program-as-Weights, The World Is In Your Mind, and Dual On-policy Distillation
Jun 30, 2026
plus more about Tapered LMs, Improved LLDMs, AutoData, and You Don't Need To Run Every Eval
Jun 23, 2026
plus more about Looped World Models, Fixed-Point Reasoners, and ExpRL
Jun 16, 2026
plus more about FlashMemory-DeepSeek-V4, Trajectory-Refined Distillation, Test-Time Gradient Guidance, and End-to-End Context Compression at Scale
Jun 9, 2026
plus more about If LLMs Have Human-Like Attributes, Then So Does Age of Empires II, Cosmos 3, and Robots Need More than VLA and World Models