Trajectory Soup: Pushing the Compute-Scaling Frontier of LLM Mid-training via Diverse Trajectories

Researchers introduce Trajectory Soup, a method to distribute mid-training budgets over multiple independent branches of large language models, improving aggregate downstream performance and extending the compute-scaling frontier of mid-training.

RSS Score 0 9/30/2026, 4:00:00 AM Original Source
Save an API key to vote.