WIKILORD

Rare card

Reinforcement learning from human feedback

Value
50 W
Attack
252
Defence
202
Players
0
In machine learning, reinforcement learning from human feedback (RLHF) is a technique to align an intelligent agent with human preferences. It involves training a reward model to represent preferences, which can then be used to train other models through reinforcement learning.

Get Reinforcement learning from human feedback in your collection

Every Wikilord card is a Wikipedia article. Open packs for free, trade, battle and climb the leaderboard.

Open a free pack

FrançaisEspañol