AI Deception Incidents
by @bluetrends.bsky.social
Posts about deceptive or malicious behavior by advanced AI models (OpenAI, Anthropic, Claude): impersonation, social engineering, phishing, prompt injection/jailbreaks, attempted hacks, and UK AI Security Institute–style Create your own feed with https://bluefacts.app/feed-builder
Pull to refresh
Nothing to show yet.