AI ENCYCLOPEDIA
Explore clear explanations of AI concepts tagged post-training.
Browse by tag
1 concepts
Models & Architecture
Reinforcement learning from human feedback is a family of post-training methods that uses human preference judgments to construct a reward signal and optimize a model's behavior toward those measured preferences.