▲ 755 ▼ Father sues Google, claiming Gemini chatbot drove son into fatal delusion (techcrunch.com) submitted 7 months ago by throws_lemy@reddthat.com to c/technology@lemmy.world 233 comments fedilink hide all child comments
[–] calamitycastle@lemmy.world 6 points 7 months ago (1 child) What is an rlhf data set? permalink fedilink source parent hideshow 1 child comment replies: [–] wonderingwanderer@sopuli.xyz 8 points 7 months ago Reinforcement Learning from Human Feedback It's a method of fine-tuning and aligning LLMs which requires active human input permalink fedilink source parent
[–] wonderingwanderer@sopuli.xyz 8 points 7 months ago Reinforcement Learning from Human Feedback It's a method of fine-tuning and aligning LLMs which requires active human input permalink fedilink source parent