Archive

All 2,608 AI stories, newest first · page 106 of 109

Rethinking Generalization in Reasoning SFT
research

Rethinking Generalization in Reasoning SFT

Researchers challenge the notion that supervised finetuning memorizes while reinforcement learning generalizes. They find that cross-domain generalization is conditional, influenced by optimization, data, and model capability. This challenges prevailing narratives in LLM post-training.