Logo
Search
The AI Timeline
LOG IN
HOME
ARCHIVE
TAGS
AUTHORS
UPGRADE
Logo
Explorative Modeling: Third Pre-training Axis?

/

Aug 4, 2026

Explorative Modeling: Third Pre-training Axis?

Plus more about Memory Foundation Model, Weak-to-Strong OPD, and LeRoPE

by cloud
by cloud
Kimi K3 Technical Report

/

Jul 29, 2026

Kimi K3 Technical Report

plus more about Hilbert Operator for Progressive Encoding, RLVR-Native Optimization Stack, Soap & Muon At Scale, and Measuring Reward-Seeking

by cloud
by cloud
On-Policy Self-Distillation without Any Supervision

/

Aug 11, 2026

On-Policy Self-Distillation without Any Supervision

plus more about Leanstral, Towards Physics of Multimodal Pretraining, and loss does not see the basis

by cloud
by cloud

Subscribe to our newsletter

Get all the latest news delivered straight to your inbox

Kimi K3 Technical Report

/

Jul 29, 2026

Kimi K3 Technical Report

plus more about Hilbert Operator for Progressive Encoding, RLVR-Native Optimization Stack, Soap & Muon At Scale, and Measuring Reward-Seeking

by cloud
by cloud
On-Policy Delta Distillation

/

Jul 21, 2026

On-Policy Delta Distillation

plus more about Concurrent Image Understanding and Generation, Latent and Explicit Reasoning with Looped Transformers and more

by cloud
by cloud
LLM That Can Modify Itself?

/

Jun 17, 2025

LLM That Can Modify Itself?

Plus more about "The Diffusion Duality" and "Reinforcement Pre-Training"

by cloud
by cloud
A Shocking RLVR Revelation For LLM Just Dropped

/

Jun 3, 2025

A Shocking RLVR Revelation For LLM Just Dropped

Read about "Rethinking Training Signals in RLVR", why LLMs are headless chickens, and "Learning to Reason without External Rewards"

by cloud
by cloud
May 2025 Research Trend Report

Premium Insights

/

Jun 6, 2025

May 2025 Research Trend Report

Premium Insights: A recap of popular AI research papers and research trends in May 2025

by cloud
by cloud
How DeepSeek Made The Best Math Prover Ever (+500% vs prev. SoTA)

Premium Insights

/

May 19, 2025

How DeepSeek Made The Best Math Prover Ever (+500% vs prev. SoTA)

Premium Insights: A closer look into the DeepSeek Prover series

by cloud
himanshu
by cloud, +1
DeepSeek-R1 Explained

/

Jan 28, 2025

DeepSeek-R1 Explained

Plus more about Transformer2 and Kimi k1.5

by cloud
by cloud

Featured Posts

Latest Videos

Archive

On-Policy Self-Distillation without Any Supervision

/

Aug 11, 2026

On-Policy Self-Distillation without Any Supervision

plus more about Leanstral, Towards Physics of Multimodal Pretraining, and loss does not see the basis

by cloud
by cloud
Explorative Modeling: Third Pre-training Axis?

/

Aug 4, 2026

Explorative Modeling: Third Pre-training Axis?

Plus more about Memory Foundation Model, Weak-to-Strong OPD, and LeRoPE

by cloud
by cloud
Kimi K3 Technical Report

/

Jul 29, 2026

Kimi K3 Technical Report

plus more about Hilbert Operator for Progressive Encoding, RLVR-Native Optimization Stack, Soap & Muon At Scale, and Measuring Reward-Seeking

by cloud
by cloud
On-Policy Delta Distillation

/

Jul 21, 2026

On-Policy Delta Distillation

plus more about Concurrent Image Understanding and Generation, Latent and Explicit Reasoning with Looped Transformers and more

by cloud
by cloud
Why Memorized Knowledge Fails to Generalize in LLM Finetuning

/

Jul 14, 2026

Why Memorized Knowledge Fails to Generalize in LLM Finetuning

Plus more about Single Async Opt for Agentic RL, Remember When It Matters, and Sparse Delta Memory

by cloud
by cloud
You Only Need 1 Layer for RLVR?

/

Jul 7, 2026

You Only Need 1 Layer for RLVR?

plus more about AdaJEPA, Program-as-Weights, The World Is In Your Mind, and Dual On-policy Distillation

by cloud
by cloud
DeepSeek Just dropped a new speculative decoding method!

/

Jun 30, 2026

DeepSeek Just dropped a new speculative decoding method!

plus more about Tapered LMs, Improved LLDMs, AutoData, and You Don't Need To Run Every Eval

by cloud
by cloud
What even is a >< former (yes >< former)

/

Jun 23, 2026

What even is a >< former (yes >< former)

plus more about Looped World Models, Fixed-Point Reasoners, and ExpRL

by cloud
by cloud
Load more

The AI Timeline

Follow The Latest Cutting Edge AI Research in 5 minutes a week.

© 2026 bycloudai.
beehiivPowered by beehiiv