Yeah if it’s just a basic input with a function call for spam or not spam the risk is low. What’s the worst case outcome, it tricks it into saying no it isn’t a scam and you have to delete it manually? Hahaha
Comment on Phishing / spam detection with local LLMs?
avidamoeba@lemmy.ca 1 week ago Good point. It’ll have to have no access to the internet or anything local outside of its container. Just text in, text out.
Zikeji@programming.dev 1 week ago
Dran_Arcana@lemmy.world 1 week ago
You could (and probably should) use a system-one style inference system for spam classification. Much cheaper and the structured output means it’s impossible to go rogue and curl some malware or whatever. It can absolutely misclassify but its output is programmatically structured and just ranks a pre-selected set of output tokens.
In your case that’s
Spam
Not_spam