you are viewing a single comment's thread
view the rest of the comments
[–] 8 points 7 months ago

Reinforcement Learning from Human Feedback

It's a method of fine-tuning and aligning LLMs which requires active human input

  • source
  • parent