/
Aug 4, 2026
Plus more about Memory Foundation Model, Weak-to-Strong OPD, and LeRoPE
Jul 29, 2026
plus more about Hilbert Operator for Progressive Encoding, RLVR-Native Optimization Stack, Soap & Muon At Scale, and Measuring Reward-Seeking
Jul 21, 2026
plus more about Concurrent Image Understanding and Generation, Latent and Explicit Reasoning with Looped Transformers and more
Jul 14, 2026
Plus more about Single Async Opt for Agentic RL, Remember When It Matters, and Sparse Delta Memory
Jul 7, 2026
plus more about AdaJEPA, Program-as-Weights, The World Is In Your Mind, and Dual On-policy Distillation
Jun 30, 2026
plus more about Tapered LMs, Improved LLDMs, AutoData, and You Don't Need To Run Every Eval
Jun 23, 2026
plus more about Looped World Models, Fixed-Point Reasoners, and ExpRL
Jun 16, 2026
plus more about FlashMemory-DeepSeek-V4, Trajectory-Refined Distillation, Test-Time Gradient Guidance, and End-to-End Context Compression at Scale
Jun 9, 2026
plus more about If LLMs Have Human-Like Attributes, Then So Does Age of Empires II, Cosmos 3, and Robots Need More than VLA and World Models
Jun 2, 2026
plus more about Bitter Lesson in Data Filtering, Do Language Models Need Sleep, and Neural Weight Norm.
May 26, 2026
plus more on the Benefits of Subword Tokenization, HRM-Text, Probabilistic Tiny Recursive Model, and Vector Policy Optimization
May 19, 2026
plus more about Self-distilled Agentic RL, Embedded Language Flows, and Negation Neglect
May 12, 2026
plus more on Sparser, Faster, Lighter Transformer LMs, Manifold Steering, and Teaching Claude Why
May 5, 2026
can't believe they removed this paper unknowningly
Apr 21, 2026
plus more about Looped Transformers, Nexus, RNN with Memory, and more
Apr 14, 2026
plus more about In-Place TTT, TriAttention, and Interleaved Head Attention.
Apr 7, 2026
plus more on Path-Constrained MoE, HISA, and Screening is not enough
weekly papers recap
Mar 31, 2026
plus more on Claudini, Composer 2, and self-distillation
Mar 25, 2026
plus more about V-JEPA 2.1, Mamba 3, and latent planning
Mar 17, 2026
and more about GLM-OCR, pre-pre-training on NCA, IndexCache, and neural thickets
Mar 10, 2026
and more about Speculative Speculative Decoding, SWE-CI, and Beyond Language Modeling
Mar 4, 2026
plus more on Learning Without Training and The Geometry of Noise
Feb 24, 2026
plus more about Experiential RL, GLM-5 Report, and Attention Matching
Feb 17, 2026
plus more on Evolving Agents via Recursive Skill-Augmented RL and Low Hanging Fruits in Vision Transformers
Feb 10, 2026
an insane big week in AI reseasrch