Dwarkesh Podcast ·

Analysis: from 2019 to 2025, gains in pretraining compute efficiency came mostly from data improvements rather than model improvements

Breaking down 6 years of pretraining progress into data vs model improvements

Lead Source

How this story grew

Coverage · 0 Discussion · 0
Sep 9Sep 10

Discussion

Related stories