Comment on Grok praises Hitler, gives credit to Musk for removing “woke filters”

<- View Parent
brucethemoose@lemmy.world ⁨17⁩ ⁨hours⁩ ago

DeepSeek, now that is a filtered LLM.

The web version has a strict filter that cuts it off. Not sure about API access, but raw Deepseek 671B is actually pretty open. Especially with the right prompting.

There are also finetunes that specifically remove China-specific refusals:

huggingface.co/microsoft/MAI-DS-R1

huggingface.co/perplexity-ai/r1-1776

Note that Microsoft actually added saftey training to “improve its risk profile”

Grok losing the guardrails means it will be distilled internet speech deprived of decency and empathy.

Instruct LLMs aren’t trained on raw data.

It wouldn’t be talking like this if it was just trained on randomized, augmented conversations, or even mostly Twitter data. They cherry picked “anti woke” data to do this real quick, and the result effectively drove the model crazy. It has all the signatures of a bad finetune: specific overused phrases.

source
Sort:hotnewtop