DRY-SFT: New AI Training Method Boosts Problem-Solving Diversity and Coverage
Researchers introduced DRY-SFT (Don't Repeat Yourself Supervised Fine-Tuning), a new post-training method that increases output diversity and coverage in large language models, improving the probability of finding at least one correct solution in verifiable domains like math and coding.