I’ve tried giving screenshots of phishing emails to a local Qwen instance and so far it always correctly detected it as scam, even points out the exact elements that it based its judgement on. Sending screenshots to it ad-hoc isn’t too scalable for family and friends. I’d like to be able to either forward emails for screening, or perhaps have it screen everything from a mailbox.

Has anyone done anything like this? Is there anything self-hostable that does this?

  • KairuByte@lemmy.dbzer0.com
    link
    fedilink
    English
    arrow-up
    2
    ·
    8 hours ago

    The ability to do it aside, if feel it should be brought up that handing off things like this can lead to giving anything that gets through an “unofficial seal of approval” so to speak.

    “This email looks phishy, but AI says it’s safe so I’ll ignore my initial concerns” is going to inevitably happen.

    There’s also the fact that using a model to detect scams, means you can use that same model to train a different model on how to build scams that bypass the detection.

    This also feels a little like a hammer looking for a nail. There have been tools to handle this without the insane computational requirements of an LLM for many years.