본문으로 건너뛰기

RLHF (Reinforcement Learning from Human Feedback)

A training technique that aligns models with human preferences using a reward model learned from human ratings.

관련 리소스