reinforcement fine-tuning

A technique where a pre-trained model is subsequently refined through reinforcement learning to improve its performance on a specific task. This often combines the advantages of supervised learning with the adaptability of reinforcement learning.

15 papers