Comment on Bill Gates feels Generative AI has plateaued, says GPT-5 will not be any better

<- View Parent
HiggsBroson@lemmy.world ⁨11⁩ ⁨months⁩ ago

You can finetune LLMs using smaller datasets, or with RLHF (reinforcement learning from human feedback) wherein people can give ratings to responses and the model can be either “rewarded” or “penalized” based off of the ratings for a given output. This retrains the LLM to produce outputs that people prefer.

source
Sort:hotnewtop