Javier Rando @javirandor.com Red-Teaming LLMs / PhD student at ETH Zurich / Prev. research intern at Meta / People call me Javi / Vegan 🌱
Website: javirando.com
@floriantramer.bsky.social @floriantramer.bsky.social Assistant professor of computer science at ETH Zürich. Interested in Security, Privacy and Machine Learning.
https://floriantramer.com
https://spylab.ai
Peter Hase @peterbhase.bsky.social Visiting Scientist at Schmidt Sciences. Visiting Researcher at Stanford NLP Group
Interested in AI safety and interpretability
Previously: Anthropic, AI2, Google, Meta, UNC Chapel Hill
David Lindner @davidlindner.bsky.social Making AI safer at Google DeepMind
davidlindner.me
@goodside.bsky.social @goodside.bsky.social
Peter Henderson @peterhenderson.bsky.social Assistant Professor, leading the Polaris Lab @ Princeton (https://www.polarislab.org/); Researching: RL, Strategic Decision-Making+Exploration; Law
Dylan Hadfield-Menell @dhadfieldmenell.bsky.social Assistant Prof of AI & Decision-Making @MIT EECS
I run the Algorithmic Alignment Group (https://algorithmicalignment.csail.mit.edu/) in CSAIL.
I work on value (mis)alignment in AI systems.
https://people.csail.mit.edu/dhm/
Sam Bowman @sleepinyourhat.bsky.social AI safety at Anthropic, on leave from a faculty job at NYU.
Views not employers'.
I think you should join Giving What We Can.
cims.nyu.edu/~sbowman
Norman Mu @normanmu.com prev: safety lead xAI, Berkeley EECS PhD
Jonathan Hayase @jon.jon.ke 5th year PhD student at UW CSE, working on Security and Privacy for ML
Edoardo Debenedetti @edebenedetti.bsky.social PhD student at ETH Zurich | Student Researcher at Google | Agents Security and more in general ML Security and Privacy
edoardo.science
spylab.ai
Boyi Wei @boyiwei.bsky.social PhD Student @Princeton
Michael Aerni @aemai.bsky.social AI privacy and security | PhD student in the SPY Lab at ETH Zurich | Ask me about coffee ☕️
Daniel Paleka @dpaleka.bsky.social ai safety researcher | phd ETH Zurich | https://danielpaleka.com
Tinghao Xie (✈️ Neurips) @tinghaox.bsky.social 3rd year Phd candidate @ Princeton ECE
Maksym Andriushchenko @maksym-andr.bsky.social Faculty at the ELLIS Institute Tübingen and Max Planck Institute for Intelligent Systems. Leading the AI Safety and Alignment group. PhD from EPFL supported by Google & OpenPhil PhD fellowships.
More details: https://www.andriushchenko.me/
Hannah Rose Kirk @hannahrosekirk.bsky.social
Ben Edelman @benedelman.bsky.social Thinking about how/why AI works/doesn't, and how to make it go well for us.
Currently: AI Agent Security @ US AI Safety Institute
benjaminedelman.com
Sean O hEigeartaigh @sean-o-h.bsky.social Academic, AI nerd and science nerd more broadly. Currently obsessed with stravinsky (not sure how that happened).
Kristina Nikolić @nkristina.bsky.social PhD student at ETH Zurich, working on AI safety. Cambridge MPhil in ML graduate | Alumnus of Mathematical Grammar School | from Serbia
Evžen @evzen.bsky.social Data Science student, ETH Zurich 🇨🇭 MATS alumnus, Talos AI Governance Fellow 🇪🇺 Based in London ☔
Leo Schwinn @NeurIPS @schwinnl.bsky.social Father of two :-), Working on LLM robustness @TU_Muenchen
David Rein @idavidrein.bsky.social sentio ergo sum. developing the science of evals at METR. prev NYU, cohere- N Nitarshan Rajkumar @nitarshan.bsky.social Co-founded UK AI Safety Institute. Co-created AI Safety Summit and UK AI Research Resource. PhD @cambridge_cl. ex @mila_quebec, @airbnb, @uwaterloo
@eliotkjones.bsky.social @eliotkjones.bsky.social AI Safety + Security @ Gray Swan AI
Formerly PleIAs + Stanford
Jie Zhang @jiezhang-ethz.bsky.social PhD student at ETH Zurich, working on ML privacy and security
https://zj-jayzhang.github.io/
Jan Kulveit @kulveit.bsky.social Researching x-risks, AI alignment, complex systems, rational decision making
@asacoopstick.bsky.social @asacoopstick.bsky.social I test language models @ the UK AI Safety Institute
@usmananwar.bsky.social @usmananwar.bsky.social
Anka Reuel ➡️ NeurIPS @ankareuel.bsky.social Computer Science PhD Student @ Stanford | Geopolitics & Technology Fellow @ Harvard Kennedy School/Belfer | Vice Chair EU AI Code of Practice | Views are my own
derek guy @dieworkwear.bsky.social Menswear writer. Editor at Put This On. Words at The New York Times, The Washington Post, The Financial Times, Esquire, and Mr. Porter.
If you have a style question, search:
https://dieworkwear.com/ | https://putthison.com/start-here/
@alpindale.bsky.social @alpindale.bsky.social
deepfates @deepfates.com.deepfates.com.deepfates.com.deepfates.com.deepfates.com cofounder Grove Research @grove-research.bsky.social
Joey Politano🏳️🌈 @josephpolitano.bsky.social Writing a data-driven newsletter about economics @ apricitas.io
Nuance? In this Economy
Full Employment Stan, Brazilian Coffee Tariff Victim |
Paul Crowley @ciphergoth.org I don't really come here any more sorry!
OSINTtechnical @osinttechnical.bsky.social PAI enjoyer, OSINT guy @truth.bsky.social , my views/freezing cold takes are my own. Standard spiel about not endorsing retweets, likes, and comments.
Daniel Lowd @dlowd.com CS Prof at the University of Oregon, studying adversarial machine learning, data poisoning, interpretable AI, probabilistic and relational models, and more. Avid unicyclist and occasional singer-songwriter. He/him
pranav @pranav.bsky.social Research Scientist at Google DeepMind. ಕನ್ನಡಿಗ.
Past: Researchoor, Algorithms team at OpenAI & with Juergen Schmidhuber.
🇺🇦 Alex Polozov @alexpolozov.com Sr. Staff Research Scientist @ Google DeepMind • previously Google X, Microsoft Research, UW • program synthesis, AI for Code and SWE • he/him • alexpolozov.com
Frank Sauer @drfranksauer.bsky.social 🎙️ Co-Host @sicherheitspod.de
🤓 Head of Research @metis.unibw.de
🐦 https://twitter.com/drfranksauer/ 👈 inactive
🦣 http://mastodontech.de/@drfranksauer 👈 bridged
Miles Brundage @milesbrundage.bsky.social AI policy researcher, wife guy, fan of cute animals and sci-fi, executive director of AVERI, Substacker
wint @dril.bsky.social Never Bullshit
I challenge any and every one who wants to kick my ass to a debate .
https://www.patreon.com/dril
https://www.instagram.com/dril
https://linktr.ee/drilreal
Bella Rudd @bellarudd.bsky.social imagine a pronoun to start,
a quote that will make me sound smart,
a cause that's supported,
emojis assorted,
disclaimers i wish to impart
𝛙ꙮ @eigenrobot.bsky.social real live human boy