HomeAI glossaryReinforcement Learning from Human Feedback (RLHF)

Modified

Reinforcement Learning from Human Feedback (RLHF)

In Swedish: RLHF

Reinforcement Learning from Human Feedback: a method where human feedback is used to train the model to give better answers. It's often used to make models more helpful and safer.

Domain:Modellträning

Browse the full glossary · 243 terms in Swedish and English