Hacker News new | ask | show | jobs
by rjh29 1195 days ago
LLMs can also be used to identify spam, not by language, but by the actual intent and "is this email something I want to read".

Open question: is there any case where LLMs can be used for malicious purposes, but LLMs can't be used to defend against it?

1 comments

> LLMs can also be used to identify spam

It will be fun to watch the arms race where the spam generator need to conceal prompt injection attacks meant to circumvent such filters while at the same time be too subtle for a humans to pick up