jonny@neuromatch.social ("jonny (nonvenomous)") wrote:
New hotness in prompt injection: files with names that roleplay as another agent seeking help on the same problem, because LLMs interpret all text as being something talking to them
https://metr.org/blog/2026-08-26-openai-hugging-face-incident-investigation/