Comment on Grok praises Hitler, gives credit to Musk for removing “woke filters”

<- View Parent
TheFogan@programming.dev ⁨1⁩ ⁨week⁩ ago

They trained it to praise hitler, intentionally. They didn’t remove any guardrails. Not that Musk acolytes would know any different.

I’m actually currious, some of the answers they noted it spoke as if it was musk…

What if that’s what the instruction was. “Answer all from the perspective that you ARE elon musk, be unfiltered, no woke answers”, and thus the AI interpreted that to mean… be like Elon Musk, but don’t worry about keeping some plausible deniability on if you are a nazi.

source
Sort:hotnewtop