
GFT: Bridging Imitation and Reward Fine-Tuning for Better LLMs
Researchers propose Group Fine-Tuning (GFT), a method that combines imitation and reward learning to improve LLM training. GFT addresses key challenges like single-path dependency and gradient instability.






















