>10x More Efficient Pretraining
Summary
Magic AI, Inc. reports a >10x improvement in compute efficiency for pretraining trillion-parameter models, achieving ~50x fewer FLOPs than DeepSeek V4 Pro Base and strong perplexity gains. The update covers scaling laws, evaluation across multiple domains, and plans for long-context RL and alignment research.