← Back to Model Beat
Research·2d ago·all news from August 20, 2026

Scaling Laws for Mixture Pretraining Under Data Constraints

Apple researchers have analyzed how language models should balance limited, high-quality data with abundant generic datasets during pretraining. Their findings provide a framework for optimizing model performance when scaling in fields with constrained information, such as specialized technical domains or languages with fewer available resources.

Covered by 1 source

Related stories

ResearchAirTag reveals how Amazon destroys rare books for AI trainingAug 17 · 5 sourcesResearchGRPO Beyond English: A Large-Scale Study of GRPO in Non-English and Multilingual SettingsAug 17 · 3 sourcesResearchEconomic ResearchAug 20ResearchMicron unveils $10 billion AI memory research lab in BoiseAug 20 · 3 sources