Modified
AI glossary
Reinforcement Learning from Human Feedback (RLHF)
In Swedish: RLHF
Reinforcement Learning from Human Feedback: a method where human feedback is used to train the model to give better answers. It's often used to make models more helpful and safer.
Domain:Modellträning
Browse the full glossary · 243 terms in Swedish and English