MIRI @intelligence.org For over two decades, the Machine Intelligence Research Institute (MIRI) has worked to understand and prepare for the critical challenges that humanity will face as it transitions to a world with artificial superintelligence.
Misha Ahrens @mishaahrens.bsky.social Neuroscientist @Janelia, HHMI
www.ahrenslab.org
Liv @livgorton.bsky.social ✨ mechanistic interpretability research scientist @ Goodfire | deep learning, math, biology | creating a more beautiful future
Jannik Brinkmann @jannikbrinkmann.bsky.social
Francesco Ortu @francescortu.bsky.social NLP & Interpretability | PhD Student @ University of Trieste & Laboratory of Data Engineering of Area Science Park | Prev MPI-IS
Kaiser Sun @kaiserwholearns.bsky.social Ph.D. student at @jhuclsp, human LM that hallucinates. Formerly @MetaAI, @uwnlp, and @AWS they/them🏳️🌈 #NLProc #NLP Crossposting on X.
Natalie Shapira @natalieshapira.bsky.social Tell me about challenges, the unbelievable, the human mind and artificial intelligence, thoughts, social life, family life, science and philosophy.
Carl Allen @carl-allen.bsky.social Laplace Junior Chair, Machine Learning
ENS Paris. (prev ETH Zurich, Edinburgh, Oxford..)
Working on mathematical foundations/probabilistic interpretability of ML (what NNs learn🤷♂️, disentanglement🤔, king-man+woman=queen?👌…)
Tiago Pimentel @tpimentel.bsky.social Postdoc at ETH. Formerly, PhD student at the University of Cambridge :)
Alessandro Stolfo @alestolfo.bsky.social PhD @ ETHZ - LLM Interpretability
alestolfo.github.io
Bart Bussmann @bartbussmann.bsky.social Independent Mechanistic Interpretability Researcher
Michael Hanna @michaelwhanna.bsky.social PhD Student at the ILLC / UvA doing work at the intersection of (mechanistic) interpretability and cognitive science. Current Anthropic Fellow.
hannamw.github.io
David Atkinson @diatkinson.bsky.social PhD student at Northeastern, previously at EpochAI. Doing AI interpretability.
diatkinson.github.io
@kevdududu.bsky.social @kevdududu.bsky.social
Vaidehi Patil @vaidehipatil.bsky.social Ph.D. Student at UNC NLP | Prev: Apple, Amazon, Adobe (Intern) vaidehi99.github.io | Undergrad @IITBombay
Neel Rajani @neelrajani.bsky.social PhD student in Responsible NLP at the University of Edinburgh, curious about interpretability and alignment
Koyena Pal @koyena.bsky.social CS Ph.D. Candidate @ Northeastern | Interpretability + Data Science | BS/MS @ Brown
koyenapal.github.io
Jasmijn Bastings @jasmijn.bastings.me Researcher, writer, photographer.
📸 bastings.art
📃 transponder.blog
🌐 jasmijn.bastings.me
📍 Amsterdam
Taufeeque @taufeeque.bsky.social Research Engineer @ FAR.AI
taufeeque9.github.io
Gonçalo Paulo @goncalo-paulo.bsky.social Interpretability researcher at @eleutherai.bsky.social
Shivam Raval @sraval.bsky.social Physics, Visualization and AI PhD @ Harvard | Embedding visualization and LLM interpretability | Love pretty visuals, math, physics and pets | Currently into manifolds
Wanna meet and chat? Book a meeting here: https://zcal.co/shivam-raval
Javier Ferrando @javifer.bsky.social Interpretability
@woog0.bsky.social @woog0.bsky.social
@neelnanda.bsky.social @neelnanda.bsky.social
Marianne de Heer Kloots @mdhk.net Linguist in AI & CogSci 🧠👩💻🤖
PhD student @illc-uva.bsky.social
🌐 https://mdhk.net/
🐘 https://scholar.social/@mdhk
🐦 https://twitter.com/mariannedhk
Abhilasha Ravichander @lasha.bsky.social Tenure-track faculty at the Max Planck Institute for Software Systems
Previously postdoc at UW and AI2, working on Natural Language Processing
Recruiting PhD students!
🌐 https://lasharavichander.github.io/
Max Müller-Eberstein @mxij.me Postdoc AI Researcher (NLP) @ ITU Copenhagen
🧭 https://mxij.me
Anne Oeldorf-Hirsch @anneo.bsky.social Comm tech & social media research professor by day, symphony violinist by night, outside as much as possible otherwise. German/American Pacific Northwestern New Englander, #firstgen academic, she/her, 🏳️🌈
https://anne-oeldorf-hirsch.uconn.edu
Alicia Curth @aliciacurth.bsky.social Machine Learner by day, 🦮 Statistician at ❤️
In search of statistical intuition for modern ML & simple explanations for complex things👀
Interested in the mysteries of modern ML, causality & all of stats. Opinions my own.
https://aliciacurth.github.io
Oskar van der Wal @ovdw.bsky.social Technology specialist at the EU AI Office / AI Safety / Prev: University of Amsterdam, EleutherAI, BigScience
Thoughts & opinions are my own and do not necessarily represent my employer.
Eliana Pastor @elianapastor.bsky.social Assistant Professor at PoliTo 🇮🇹 |
Former Visiting scholar at UCSC 🇺🇸 |
she/her | TrustworthyAI, XAI, Fairness in AI
https://elianap.github.io/
Dilyara Bareeva @dilya.bsky.social PhD Candidate in Interpretability @FraunhoferHHI | 📍Berlin, Germany
dilyabareeva.github.io
Rachel Lawrence @rachel-law.bsky.social Organic machine turning tea into theorems ☕️
AI @ Microsoft Research ➡️ Goal: Teach models (and humans) to reason better
Let’s connect re: AI for social good, graphs & network dynamics, discrete math, logic 🧩, 🥾,🎨, protecting democracy🗽
www.rlaw.me
Eric Todd @ericwtodd.bsky.social CS PhD Student, Northeastern University - Machine Learning, Interpretability https://ericwtodd.github.io
Aryaman Arora @aryaman.io member of technical staff @stanfordnlp.bsky.social
Chris Wendler @wendlerc.bsky.social Postdoc at the interpretable deep learning lab at Northeastern University, deep learning, LLMs, mechanistic interpretability
vedang @vedanglad.bsky.social ai interpretability research and running • thinking about how models think • prev @MIT cs + physics
Jonathan Ling @jonling.bsky.social Assistant Professor @HopkinsMedicine @JHUPath
https://scholar.google.com/citations?user=dGBD72YAAAAJ
Cristina @cristinaml.bsky.social ML/AI researcher @JohnsHopkins
Shan Chen @shan23chen.bsky.social PhDing @Harvard @MassGenBrigham|PhD Fellow @Google | Previously @Bos_CHIP @BrandeisU
More robustness and explainabilities 🧐 for Health AI.
shanchen.dev
José Oramas @jaom7.bsky.social Associate Professor @UAntwerp, sqIRL/IDLab, imec.
#RepresentationLearning, #Model #Interpretability & #Explainability
A guy who plays with toy bricks, enjoys research and gaming.
Opinions are my own
idlab.uantwerpen.be/~joramasmogrovejo
@michael-pearce.bsky.social @michael-pearce.bsky.social
Nishant Subramani @ ACL @nsubramani23.bsky.social PhD student @CMU LTI - working on model #interpretability, student researcher @google; prev predoc @ai2; intern @MSFT
nishantsubramani.github.io
Julian Minder @jkminder.bsky.social PhD at EPFL with Robert West, Master at ETHZ
Mainly interested in Language Model Interpretability and Model Diffing.
MATS 7.0 Winter 2025 Scholar w/ Neel Nanda
jkminder.ch
Arthur Conmy @arthurconmy.bsky.social Aspiring 10x reverse engineer at Google DeepMind
Kayo Yin @kayoyin.bsky.social PhD student at UC Berkeley. NLP for signed languages and LLM interpretability. kayoyin.github.io
🏂🎹🚵♀️🥋
Martin Wattenberg @wattenberg.bsky.social Human/AI interaction. ML interpretability. Visualization as design, science, art. Professor at Harvard, and part-time at Google DeepMind.
Dr. Kareem Carr, Ph.D. @kareemcarr.bsky.social Statistician. PhD @Harvard • Masters degree in pure math • Follow me for fun, nerdy content.
Sign up to my newsletter: kareemcarr.substack.com
Tomer Ullman @tomerullman.bsky.social Associate Professor, Department of Psychology, Harvard University. Computation, cognition, development.
Sam Gershman @gershbrain.bsky.social Professor, Department of Psychology and Center for Brain Science, Harvard University
https://gershmanlab.com/
Roger Levy @rplevy.bsky.social Director, MIT Computational Psycholinguistics Lab. President, Cognitive Science Society. Chair of the MIT Faculty. Open access & open science advocate. He.
Lab webpage: http://cpl.mit.edu/
Personal webpage: https://www.mit.edu/~rplevy
David Smith @dasmiq.bsky.social Associate professor of computer science at Northeastern University. Natural language processing, digital humanities, OCR, computational bibliography, and computational social sciences. Artificial intelligence is an archival science.
Jennifer Hu @jennhu.bsky.social Asst Prof at Johns Hopkins Cognitive Science • Director of the Group for Language and Intelligence (glint) ✨• Interested in all things language, cognition, and AI
jennhu.github.io
Kanaka Rajan @kanakarajanphd.bsky.social Associate Professor at Harvard & Kempner Institute. Applying computational frameworks & machine learning to decode multi-scale neural processes. Marathoner. Rescue dog mom. https://www.rajanlab.com/
Leshem (Legend) Choshen @EMNLP @lchoshen.bsky.social 🥇 LLMs together (co-created model merging, BabyLM, textArena.ai)
🥈 Spreading science over hype in #ML & #NLP
Proud shareLM💬 Donor
@IBMResearch & @MIT_CSAIL
Ilenna Jones @ilennaj.bsky.social | Cellular/Molecular-turned-Computational Neuroscientist |
| What do neurons even do?? | Neural Computation with Dendrites |
| Biophysical Optimization | AI <-> Neuro |
| Postdoctoral Research Fellow at the Harvard Kempner Institute |
| www.ilenna.com
Ella Batty @ellabatty.bsky.social Senior Machine Learning Researcher, Kempner Institute for the Study of Natural and Artificial Intelligence, Harvard. Board Member, Neuromatch. she/her. Views are my own.
Harvard Brain Science Initiative @harvardbrainsci.bsky.social HBI brings together neuroscience researchers from different parts of Harvard and its affiliated hospitals.
Throughout all that we do, we aspire to build and nurture a scientific community that is diverse, inclusive, and welcoming.
https://brain.harvard.edu
@aaronwalsman.bsky.social @aaronwalsman.bsky.social
Ekdeep Singh @ ICML @ekdeepl.bsky.social Postdoc at CBS, Harvard University
(New around here)
Zelda Mariet @zzzelda.bsky.social
Isabel Papadimitriou @isabelpapad.bsky.social (jolly good) Fellow at the Kempner Institute @kempnerinstitute.bsky.social, incoming assistant professor at UBC Linguistics (and by courtesy CS, Sept 2025). PhD @stanfordnlp.bsky.social with the lovely @jurafsky.bsky.social
isabelpapad.com
Venki Murthy @neurovenki.bsky.social Neuroscience Professor at Harvard University. Personal account and posts here. Research group website: https://vnmurthylab.org.
Brandon Rohrer @brandonrohrer.com Robotics and Reinforcement Learning tinkerer.
brandonrohrer.org
Wrangler of algorithms.
Eater of bread. Sipper of whisky.
Reports to a Shih Tzu.
Roma Patel @romapatel.bsky.social research scientist @deepmind. language & multi-agent rl & interpretability. phd @BrownUniversity '22 under ellie pavlick (she/her)
https://roma-patel.github.io
Stella Biderman @stellaathena.bsky.social I make sure that OpenAI et al. aren't the only people who are able to study large scale AI systems.
Christoph Molnar @christophmolnar.bsky.social Author of Interpretable Machine Learning and other books
Newsletter: https://mindfulmodeler.substack.com/
Website: https://christophmolnar.com/
Mimansa Jaiswal @mimansaj.bsky.social Robustness, Data & Annotations, Evaluation & Interpretability in LLMs
http://mimansajaiswal.github.io/
Christina (Chrisy) Bornberg @variint.bsky.social Enjoy not enjoying ideals | Interpretability of modular convnets applied to 👁️ and 🛰️🐝 | she/her 🦒💕
variint.github.io
Miryam de Lhoneux @mdlhx.bsky.social NLP assistant prof at KU Leuven, PI @lagom-nlp.bsky.social. I like syntax more than most people. Also multilingual NLP, interpretability, mountains and beer. (She/her)
Stephanie Brandl @stephaniebrandl.bsky.social Assistant Professor in NLP (Fairness, Interpretability and lately interested in Political Science) at the University of Copenhagen ✨
Before: PostDoc in NLP at Uni of CPH, PhD student in ML at TU Berlin
Joao Barbosa @jbarbosa.org INSERM group leader @ Neuromodulation Institute and NeuroSpin (Paris) in computational neuroscience.
How and why are computations enabling cognition distributed across the brain?
Expect neuroscience and ML content.
jbarbosa.org
parisneuro.fr
Kyle Morgenstein @kylem.bsky.social Full of childlike wonder. Building friendly robots. UT Austin PhD student, MIT ‘20.
Bharath Radhakrishnan @bharathr98.com
Anna Rogers @annarogers.bsky.social Associate professor at IT University of Copenhagen: NLP, language models, interpretability, AI & society. Co-editor-in-chief of ACL Rolling Review. #NLProc #NLP
Jenny Kunz @jeku.bsky.social Postdoc at Linköping University🇸🇪. Doing NLP, particularly explainability, language adaptation, modular LLMs. I‘m also into🌋🏕️🚴.
Mark Riedl @markriedl.bsky.social AI for storytelling, games, explainability, safety, ethics. Professor at Georgia Tech. Director of ML Center at GT. Time travel expert. Geek. Dad. he/him
Dino Sejdinovic @sejdino.bsky.social Professor of Statistical Machine Learning at the University of Adelaide.
https://sejdino.github.io/
Thomas Fel @thomasfel.bsky.social Explainability, Computer Vision, Neuro-AI.🪴 Kempner Fellow @Harvard.
Prev. PhD @Brown, @Google, @GoPro. Crêpe lover.
📍 Boston | 🔗 thomasfel.me
André Panisson @panisson.bsky.social Principal Researcher @ CENTAI.eu | Leading the Responsible AI Team. Building Responsible AI through Explainable AI, Fairness, and Transparency. Researching Graph Machine Learning, Data Science, and Complex Systems to understand collective human behavior.
Sarah Wiegreffe @sarah-nlp.bsky.social Research in NLP (mostly LM interpretability & explainability).
Assistant prof at UMD CS + CLIP.
Previously @ai2.bsky.social @uwnlp.bsky.social
Views my own.
sarahwie.github.io
Chris Olah @colah.bsky.social Reverse engineering neural networks at Anthropic. Previously Distill, OpenAI, Google Brain.Personal account.
Clément Dumas @butanium.bsky.social Master student at ENS Paris-Saclay / aspiring AI safety researcher / improviser
Prev research intern @ EPFL w/ wendlerc.bsky.social and Robert West
MATS Winter 7.0 Scholar w/ neelnanda.bsky.social
https://butanium.github.io
Aaron Mueller @amuuueller.bsky.social Postdoc at Northeastern and incoming Asst. Prof. at Boston U. Working on NLP, interpretability, causality. Previously: JHU, Meta, AWS- D David Bau @davidbau.bsky.social Interpretable Deep Networks. http://baulab.info/ @davidbau
Mor Geva @megamor2.bsky.social https://mega002.github.io
Niklas Stoehr @niklasstoehr.bsky.social Gemini Post-Training ⚫️ Research Scientist at Google DeepMind ⚫️ PhD from ETH Zurich
Nina Rimsky @ninarimsky.bsky.social AI Safety Research // Software Engineering
Gabriele Sarti @gsarti.com Open-source interpretability to seize the means of prediction. Postdoc @ Northeastern, @ndif-team.bsky.social w/ @davidbau.bsky.social.
gsarti.com
Naomi Saphra @nsaphra.bsky.social Waiting on a robot body. All opinions are universal and held by both employers and family. ML/NLP professor.
nsaphra.net
Dashiell @dashiells.bsky.social Machine learning haruspex
Joe Stacey @joestacey.bsky.social NLP PhD student at Imperial College London and Apple AI/ML Scholar.
Sweta Karlekar @swetakar.bsky.social Machine learning PhD student @ Blei Lab in Columbia University
Working in mechanistic interpretability, nlp, causal inference, and probabilistic modeling!
Previously at Meta for ~3 years on the Bayesian Modeling & Generative AI teams.
🔗 www.sweta.dev
Nicolas Beltran-Velez @velezbeltran.bsky.social Machine Learning PhD Student
@ Blei Lab & Columbia University.
Working on probabilistic ML | uncertainty quantification | LLM interpretability.
Excited about everything ML, AI and engineering!
Daniel Johnson @ddjohnson.bsky.social PhD student at Vector Institute / University of Toronto. Building tools to study neural nets and find out what they know. He/him.
www.danieldjohnson.com
Alex Makelov @amakelov.bsky.social Mechanistic interpretability
Creator of https://github.com/amakelov/mandala
prev. Harvard/MIT
machine learning, theoretical computer science, competition math.
Andrew Lee @ajyl.bsky.social Post-doc @ Harvard. PhD UMich. Spent time at FAIR and MSR. ML/NLP/Interpretability
Martina G. Vilas @martinagvilas.bsky.social AI Evaluation & Interpretability @NVIDIA
Isabelle Lee @wordscompute.bsky.social ml/nlp phding @ usc, currently visiting harvard;
training & interpretability & reasoning
iglee.me