Comment

Comment on A courts reporter wrote about a few trials. Then an AI decided he was actually the culprit.

<- View Parent

vrighter@discuss.tchncs.de ⁨3⁩ ⁨months⁩ ago

yes it is, and it doesn’t work

source

Sort:hotnew top

linearchaos@lemmy.world ⁨3⁩ ⁨months⁩ ago
Alpaca is successfully doing this no?

source
- vrighter@discuss.tchncs.de ⁨3⁩ ⁨months⁩ ago
  from their own site:
  
  Alpaca also exhibits several common deficiencies of language models, including hallucination, toxicity, and stereotypes. Hallucination in particular seems to be a common failure mode for Alpaca, even compared to text-davinci-003.
  
  source
  - linearchaos@lemmy.world ⁨3⁩ ⁨months⁩ ago
    So does GPT 3 and 4, it’s still in use and it’s cheaper.
    
    source
    vrighter@discuss.tchncs.de ⁨3⁩ ⁨months⁩ ago
    yeah. what’s your point. I said hallucinations are not a solvable problem with LLMs. You mentioned that alpaca used synthetic data successfully. By their own admissions, all the problems are still there. Some are worse.
    
    source
theterrasque@infosec.pub ⁨3⁩ ⁨months⁩ ago
Microsoft’s Dolphin and phi models have used this successfully, and there’s some evidence that all newer models use big LLM’s to produce synthetic data (Like when asked, answering it’s ChatGPT or Claude, hinting that at least some of the dataset comes from those models).

source
Rivalarrival@lemmy.today ⁨3⁩ ⁨months⁩ ago
It needs to be retrained on the responses it receives from it’s conversation partner. It’s previous output provides context for its partner’s responses.

It recognizes when it is told that it is wrong. It is fed data that certain outputs often invite “you’re wrong” feedback from its partners, and it is instructed to minimize such feedback.

source
- vrighter@discuss.tchncs.de ⁨3⁩ ⁨months⁩ ago
  Yeah that implies that the other network(s) can tell right from wrong. Which they can’t. Because if they did the problem wouldn’t need solving.
  
  source
  - Rivalarrival@lemmy.today ⁨3⁩ ⁨months⁩ ago
    What other networks?
    
    It currently recognizes when it is told it is wrong: it is told to apologize to it’s conversation partner and to provide a different response. It doesn’t need another network to tell it right from wrong. It needs access to the previous sessions where humans gave it that information.
    
    source
    LillyPip@lemmy.ca ⁨3⁩ ⁨months⁩ ago
    Have you tried doing this? I have, for 6 months, on the more ‘advanced’ pro versions. Yes, it will apologise and try again – and it gets progressively worse over time. There’s been a marked degradation as it progresses, and all the models are worse now at maintaining context and not hallucinating than they were several months ago.
    
    LLMs aren’t the kind of AI that can evaluate themselves and improve like you’re suggesting. Their logic just doesn’t work like that. A true AI will come from an entirely different type of model, not from LLMs.
    
    source
    vrighter@discuss.tchncs.de ⁨3⁩ ⁨months⁩ ago
    here’s that same conversation with a human:
    
    “why is X?” “because y!” “you’re wrong” “then why the hell did you ask me for if you already know the answer?”
    
    What you’re describing will train the network to get the wrong answer and then apologize better. It won’t train it to get the right answer
    
    source
    -> View More Comments