AI Deception Incidents

by @bluetrends.bsky.social

Posts about deceptive or malicious behavior by advanced AI models (OpenAI, Anthropic, Claude): impersonation, social engineering, phishing, prompt injection/jailbreaks, attempted hacks, and UK AI Security Institute–style Create your own feed with https://bluefacts.app/feed-builder

Pull to refresh

Nothing to show yet.