Sahwa@reddthat.com to Technology@lemmy.worldEnglish · 6 个月前Father sues Google, claiming Gemini chatbot drove son into fatal delusiontechcrunch.comexternal-linkmessage-square235linkfedilinkarrow-up1772arrow-down116
arrow-up1756arrow-down1external-linkFather sues Google, claiming Gemini chatbot drove son into fatal delusiontechcrunch.comSahwa@reddthat.com to Technology@lemmy.worldEnglish · 6 个月前message-square235linkfedilink
minus-squarewonderingwanderer@sopuli.xyzlinkfedilinkEnglisharrow-up8·6 个月前Reinforcement Learning from Human Feedback It’s a method of fine-tuning and aligning LLMs which requires active human input
What is an rlhf data set?
Reinforcement Learning from Human Feedback
It’s a method of fine-tuning and aligning LLMs which requires active human input