Gary Marcus @garymarcus.bsky.social AI and cognitive science, Founder and CEO (Geometric Intelligence, acquired by Uber). 8 books including Guitar Zero, Rebooting AI and Taming Silicon Valley.
Newsletter (100k subscribers): garymarcus.substack.com
Suresh Venkatasubramanian @geomblog.bsky.social Director, Center for Tech Responsibility@Brown. FAccT OG. AI Bill of Rights coauthor. Former tech advisor to President Biden @WHOSTP. He/him/his. Posts my own.
Eugene Vinitsky 🍒 @eugenevinitsky.bsky.social Anti-cynic. Towards a weirder future. Reinforcement Learning, Autonomous Vehicles, transportation systems, the works. Asst. Prof at NYU. Founding research scientist at Percepta.
https://emerge-lab.github.io
https://www.admonymous.co/eugenevinitsky
Pranav Rajan @pranavrajan.bsky.social digital signals, aiml, generative modeling, music. bach enthusiast
Machine Learning Student @ KTH
Martin Wattenberg @wattenberg.bsky.social Human/AI interaction. ML interpretability. Visualization as design, science, art. Professor at Harvard, and part-time at Google DeepMind.
Vishal Verma @vermavishal.bsky.social Research @carnegiemellon. || Building Causal Agent and ArbitrageX ||
Interested in reinforcement learning, alignment, birds, jazz music ||
Clément Dumas @butanium.bsky.social Master student at ENS Paris-Saclay / aspiring AI safety researcher / improviser
Prev research intern @ EPFL w/ wendlerc.bsky.social and Robert West
MATS Winter 7.0 Scholar w/ neelnanda.bsky.social
https://butanium.github.io
Aaron Mueller @amuuueller.bsky.social Postdoc at Northeastern and incoming Asst. Prof. at Boston U. Working on NLP, interpretability, causality. Previously: JHU, Meta, AWS- D David Bau @davidbau.bsky.social Interpretable Deep Networks. http://baulab.info/ @davidbau
Mor Geva @megamor2.bsky.social https://mega002.github.io
Niklas Stoehr @niklasstoehr.bsky.social Gemini Post-Training ⚫️ Research Scientist at Google DeepMind ⚫️ PhD from ETH Zurich
Nina Rimsky @ninarimsky.bsky.social AI Safety Research // Software Engineering
Gabriele Sarti @gsarti.com Open-source interpretability to seize the means of prediction. Postdoc @ Northeastern, @ndif-team.bsky.social w/ @davidbau.bsky.social.
gsarti.com
Dashiell @dashiells.bsky.social Machine learning haruspex
Joe Stacey @joestacey.bsky.social NLP PhD student at Imperial College London and Apple AI/ML Scholar.
Sweta Karlekar @swetakar.bsky.social Machine learning PhD student @ Blei Lab in Columbia University
Working in mechanistic interpretability, nlp, causal inference, and probabilistic modeling!
Previously at Meta for ~3 years on the Bayesian Modeling & Generative AI teams.
🔗 www.sweta.dev
Nicolas Beltran-Velez @velezbeltran.bsky.social Machine Learning PhD Student
@ Blei Lab & Columbia University.
Working on probabilistic ML | uncertainty quantification | LLM interpretability.
Excited about everything ML, AI and engineering!
Daniel Johnson @ddjohnson.bsky.social PhD student at Vector Institute / University of Toronto. Building tools to study neural nets and find out what they know. He/him.
www.danieldjohnson.com
Alex Makelov @amakelov.bsky.social Mechanistic interpretability
Creator of https://github.com/amakelov/mandala
prev. Harvard/MIT
machine learning, theoretical computer science, competition math.
Andrew Lee @ajyl.bsky.social Post-doc @ Harvard. PhD UMich. Spent time at FAIR and MSR. ML/NLP/Interpretability
Martina G. Vilas @martinagvilas.bsky.social AI Evaluation & Interpretability @NVIDIA
Isabelle Lee @wordscompute.bsky.social ml/nlp phding @ usc, currently visiting harvard;
training & interpretability & reasoning
iglee.me
Pepa Atanasova @apepa.bsky.social Assistant Professor, University of Copenhagen; interpretability, xAI, factuality, accountability, xAI diagnostics https://apepa.github.io/
Chris Olah @colah.bsky.social Reverse engineering neural networks at Anthropic. Previously Distill, OpenAI, Google Brain.Personal account.
Lee Sharkey @leesharkey.bsky.social Scruting matrices @ Apollo Research
Kayo Yin @kayoyin.bsky.social PhD student at UC Berkeley. NLP for signed languages and LLM interpretability. kayoyin.github.io
🏂🎹🚵♀️🥋
Arthur Conmy @arthurconmy.bsky.social Aspiring 10x reverse engineer at Google DeepMind
Julian Minder @jkminder.bsky.social PhD at EPFL with Robert West, Master at ETHZ
Mainly interested in Language Model Interpretability and Model Diffing.
MATS 7.0 Winter 2025 Scholar w/ Neel Nanda
jkminder.ch
Nishant Subramani @ ACL @nsubramani23.bsky.social PhD student @CMU LTI - working on model #interpretability, student researcher @google; prev predoc @ai2; intern @MSFT
nishantsubramani.github.io
Eric Todd @ericwtodd.bsky.social CS PhD Student, Northeastern University - Machine Learning, Interpretability https://ericwtodd.github.io
Aryaman Arora @aryaman.io member of technical staff @stanfordnlp.bsky.social
Chris Wendler @wendlerc.bsky.social Postdoc at the interpretable deep learning lab at Northeastern University, deep learning, LLMs, mechanistic interpretability
vedang @vedanglad.bsky.social ai interpretability research and running • thinking about how models think • prev @MIT cs + physics
Jonathan Ling @jonling.bsky.social Assistant Professor @HopkinsMedicine @JHUPath
https://scholar.google.com/citations?user=dGBD72YAAAAJ
Cristina @cristinaml.bsky.social ML/AI researcher @JohnsHopkins
Shan Chen @shan23chen.bsky.social PhDing @Harvard @MassGenBrigham|PhD Fellow @Google | Previously @Bos_CHIP @BrandeisU
More robustness and explainabilities 🧐 for Health AI.
shanchen.dev
José Oramas @jaom7.bsky.social Associate Professor @UAntwerp, sqIRL/IDLab, imec.
#RepresentationLearning, #Model #Interpretability & #Explainability
A guy who plays with toy bricks, enjoys research and gaming.
Opinions are my own
idlab.uantwerpen.be/~joramasmogrovejo
Jannik Brinkmann @jannikbrinkmann.bsky.social
Francesco Ortu @francescortu.bsky.social NLP & Interpretability | PhD Student @ University of Trieste & Laboratory of Data Engineering of Area Science Park | Prev MPI-IS
Kaiser Sun @kaiserwholearns.bsky.social Ph.D. student at @jhuclsp, human LM that hallucinates. Formerly @MetaAI, @uwnlp, and @AWS they/them🏳️🌈 #NLProc #NLP Crossposting on X.
Natalie Shapira @natalieshapira.bsky.social Tell me about challenges, the unbelievable, the human mind and artificial intelligence, thoughts, social life, family life, science and philosophy.
Carl Allen @carl-allen.bsky.social Laplace Junior Chair, Machine Learning
ENS Paris. (prev ETH Zurich, Edinburgh, Oxford..)
Working on mathematical foundations/probabilistic interpretability of ML (what NNs learn🤷♂️, disentanglement🤔, king-man+woman=queen?👌…)
Tiago Pimentel @tpimentel.bsky.social Postdoc at ETH. Formerly, PhD student at the University of Cambridge :)
@michael-pearce.bsky.social @michael-pearce.bsky.social
Dilyara Bareeva @dilya.bsky.social PhD Candidate in Interpretability @FraunhoferHHI | 📍Berlin, Germany
dilyabareeva.github.io
Alessandro Stolfo @alestolfo.bsky.social PhD @ ETHZ - LLM Interpretability
alestolfo.github.io
Thomas Fel @thomasfel.bsky.social Explainability, Computer Vision, Neuro-AI.🪴 Kempner Fellow @Harvard.
Prev. PhD @Brown, @Google, @GoPro. Crêpe lover.
📍 Boston | 🔗 thomasfel.me
Bart Bussmann @bartbussmann.bsky.social Independent Mechanistic Interpretability Researcher
Michael Hanna @michaelwhanna.bsky.social PhD Student at the ILLC / UvA doing work at the intersection of (mechanistic) interpretability and cognitive science. Current Anthropic Fellow.
hannamw.github.io
David Atkinson @diatkinson.bsky.social PhD student at Northeastern, previously at EpochAI. Doing AI interpretability.
diatkinson.github.io
@kevdududu.bsky.social @kevdududu.bsky.social
Vaidehi Patil @vaidehipatil.bsky.social Ph.D. Student at UNC NLP | Prev: Apple, Amazon, Adobe (Intern) vaidehi99.github.io | Undergrad @IITBombay
Neel Rajani @neelrajani.bsky.social PhD student in Responsible NLP at the University of Edinburgh, curious about interpretability and alignment
Koyena Pal @koyena.bsky.social CS Ph.D. Candidate @ Northeastern | Interpretability + Data Science | BS/MS @ Brown
koyenapal.github.io
@woog0.bsky.social @woog0.bsky.social
@neelnanda.bsky.social @neelnanda.bsky.social
Jasmijn Bastings @jasmijn.bastings.me Researcher, writer, photographer.
📸 bastings.art
📃 transponder.blog
🌐 jasmijn.bastings.me
📍 Amsterdam
Javier Ferrando @javifer.bsky.social Interpretability
Taufeeque @taufeeque.bsky.social Research Engineer @ FAR.AI
taufeeque9.github.io
Gonçalo Paulo @goncalo-paulo.bsky.social Interpretability researcher at @eleutherai.bsky.social
Kunvar Thaman @firstuserhere.bsky.social
dribnet @drib.net creations with code and networks
Caden @cadentj.bsky.social
@joshengels.bsky.social @joshengels.bsky.social PhD student at MIT.
Working on mechanistic interpretability and AI safety.
Ana Lučić @a-lucic.bsky.social Assistant professor at the University of Amsterdam. Previously at Microsoft Research, Partnership on AI.
Yoann Poupart @xmaster6y.bsky.social XAI PhD Student & Entrepreneur
claudia shi @claudiashi.bsky.social machine learning, causal inference, science of llm, ai safety, phd student @bleilab, keen bean
https://www.claudiashi.com/
Ruizhe Li @ruizheli.bsky.social Assistant Professor at University of Aberdeen | Postdoc at UCL | PhD at University of Sheffield | mechanistic interpretability & multimodal LLMs | https://www.ruizhe.space
Laura Kopf @lkopf.bsky.social PhD student in Interpretable Machine Learning at @tuberlin.bsky.social & @bifold.berlin
https://web.ml.tu-berlin.de/author/laura-kopf/
NDIF Team @ndif-team.bsky.social The National Deep Inference Fabric, an NSF-funded computational infrastructure to enable research on large-scale Artificial Intelligence.
🔗 NDIF: https://ndif.us
🧰 NNsight API: https://nnsight.net
😸 GitHub: https://github.com/ndif-team/nnsight
Chandan Singh @csinva.bsky.social Seeking superhuman explanations.
Senior researcher at Microsoft Research, PhD from UC Berkeley, https://csinva.io/
Sophie Hao @cinnamonlab.ai Assistant professor of Linguistics and Data Science at Boston University. NLP, computational linguistics, interpretability, social bias and fairness. she/her. https://www.notaphonologist.com/
Hidenori Tanaka @hidenori8tanaka.bsky.social Group Leader, CBS-NTT "Physics of Intelligence" Program at Harvard
website: https://sites.google.com/view/htanaka/home
Shivam Raval @sraval.bsky.social Physics, Visualization and AI PhD @ Harvard | Embedding visualization and LLM interpretability | Love pretty visuals, math, physics and pets | Currently into manifolds
Wanna meet and chat? Book a meeting here: https://zcal.co/shivam-raval
Naomi Saphra @nsaphra.bsky.social Waiting on a robot body. All opinions are universal and held by both employers and family. ML/NLP professor.
nsaphra.net
Bluesky @bsky.app official Bluesky account (check username👆)
Bugs, feature requests, feedback: support@bsky.app