Jascha Sohl-Dickstein @jascha.sohldickstein.com Recently a principal scientist at Google DeepMind. Joining Anthropic. Most (in)famous for inventing diffusion models. AI + physics + neuroscience + dynamical systems.
Ben Poole @benmpoole.bsky.social research scientist at google deepmind.
phd in neural nonsense from stanford.
poolio.github.io
Tanishq Mathew Abraham @iscienceluvr.bsky.social PhD at 19 |
Founder and CEO at @MedARC_AI |
Research Director at @StabilityAI |
@kaggle Notebooks GM |
Biomed. engineer @ 14 |
TEDx talk➡https://bit.ly/3tpAuan- T Tuomas Virtanen @tuomasvirtanen.bsky.social
Yoshua Bengio @yoshuabengio.bsky.social Working towards the safe development of AI for the benefit of all at Université de Montréal, LawZero and Mila.
A.M. Turing Award Recipient and most-cited AI researcher.
https://lawzero.org/en
https://yoshuabengio.org/profile/
karpathy @karpathy.bsky.social AI @ OpenAI, Tesla, Stanford
Francis Bach @bachfrancis.bsky.social Researcher in machine learning
Yann LeCun @yann-lecun.bsky.social Professor a NYU; Executive Chairman at AMI Labs..
Researcher in AI, Machine Learning, Robotics, etc.
ACM Turing Award Laureate.
http://yann.lecun.com
Alex Alemi @alexalemi.bsky.social Machine Learning Researcher
https://alexalemi.com
https://blog.alexalemi.com
Arnaud Doucet @arnauddoucet.bsky.social Senior Staff Research Scientist @Google DeepMind
https://arnauddoucet.github.io/
Eric Fosler-Lussier @ericfos.bsky.social Professor/Admin @ Ohio State. All opinions expressed on this channel are my personal opinions and do not represent that of my employer.
Odette Scharenborg @odettes.bsky.social Full professor of inclusive speech communication at TU Delft, The Netherlands. Former president of the International Speech Communication Association (ISCA). Mother of 3🌈
Ramon Astudillo @ramon-astudillo.bsky.social Principal Research Scientist at IBM Research AI in New York. Speech, Formal/Natural Language Processing. Currently LLM post-training, structured SDG and RL. Opinions my own and non stationary.
ramon.astudillo.com
Catherine Breslin @catherinebreslin.bsky.social AI scientist & consultant :: prev Amazon Alexa, Toshiba, Cam Uni :: voice & language tech :: powered by coffee :: photographer :: Cambridge UK
https://www.catherinebreslin.co.uk
Greta Tuckute @gretatuckute.bsky.social Studying language in biological brains and artificial ones at the Kempner Institute at Harvard University.
www.tuckute.com
Jesse Engel @jesseengel.bsky.social Guitarist, Researcher Google DeepMind. Opinions are my own.- M @markbcartwright.bsky.social @markbcartwright.bsky.social
Prem Seetharaman @pseeth.bsky.social Researcher in computer audition, machine learning, and HCI. Sr. Research Scientist, @AdobeResearch. Previously @DescriptApp, @Northwestern.
https://pseeth.github.io/
Christian Steinmetz @csteinmetz1.bsky.social AI for Music • Research Scientist @ Suno
Hervé Bredin (a.k.a. the pyannote guy) @hbredin.bsky.social I created pyannote open source toolkit.
Co-founder and CSO at pyannoteAI
Keisuke Imoto @keisukeimoto.bsky.social
Vincent Lostanlen @lostanlen.bsky.social Scientist at CNRS.
https://audio.ls2n.fr
Johanna Devaney @jcdevaney.bsky.social Canadian in NYC (she/her) teaching music and data analysis at Brooklyn College and the Graduate Center, CUNY. Co-Editor-in-Chief of Journal of New Music Research.
Steve Renals @srenals.bsky.social Once was speech technologist - Water of Leith, Edinburgh - Born 320.23 ppm
Heiga Zen (全 炳河) @heigazen.bsky.social Principal Scientist (Director) at Google DeepMind in Japan. 波瀬小⇒一志中⇒鈴鹿高専⇒名工大 (IBM T.J. Watson Research intern)⇒東芝欧州研究所⇒Google (Speech🇬🇧⇒Brain🇯🇵) ⇒Google DeepMind. 3rd generation Korean in Japan.
Ricard Marxer @ricard.bsky.social ml, audio, cv, nlp, speech, bioacoustics // Assoc. Prof. at Université de Toulon, researcher at LIS CNRS UMR 7020, director of http://www.master-mir.eu in marine robotics and AI
@imotts.bsky.social @imotts.bsky.social KU←田辺坂←R←SOKENDAI←通信会社N
🐿🦋 @sythonuk.bsky.social 🐿earcher
Akira Tamamori @ballforest.bsky.social Hopfield Networks / Associative Memory / Kernel methods / Dynamical Systems / Information Geometry / Outlier Detection
yamakatz @kyama0321.bsky.social Auditory Signal Processing/Objective Metrics/Hearing Assistive Technologies.
Twitter: @kyama0321
WEB: https://sites.google.com/site/kyama0321/en
Muramasa @muramasa2.bsky.social 音声の研究をしています
Kentaro Seki @trgkpc.bsky.social 1st-year doctoral student @ Univ. Tokyo | audio signal processing, speech synthesis, machine learning
https://trgkpc.github.io/
Marianne de Heer Kloots @mdhk.net Linguist in AI & CogSci 🧠👩💻🤖
PhD student @illc-uva.bsky.social
🌐 https://mdhk.net/
🐘 https://scholar.social/@mdhk
🐦 https://twitter.com/mariannedhk
Ilyass Moummad @ilyassmoummad.bsky.social Postdoctoral Researcher @ Inria Montpellier (IROKO, Pl@ntNet)
SSL for Plant Image Recognition
Interested in Computer Vision, Natural Language Processing, Machine Listening, and Biodiversity Monitoring
Website: ilyassmoummad.github.io
ErlendA @froskekongen.bsky.social AI researcher.
CTO at HANCE, Associate Professor at NTNU.
Compression, generative, audio, time series
Romain Serizel @rserizel.bsky.social Professor at Université de Lorraine/Loria/Mines Nancy. Doing research is speech and audio processing.
Andrew Owens @andrewowens.bsky.social Associate professor @ Cornell Tech
@keunwoochoi.bsky.social @keunwoochoi.bsky.social AI researcher in music, audio, LLMs.
Desh Raj @rdesh26.bsky.social Research Scientist @ Meta GenAI in NYC.
Working on audio/speech for LLaMA.
Previously: PhD @ JHU CLSP
desh2608.github.io
Emmanouil Benetos @emmanouilb.bsky.social Reader (Associate Professor), @qmuleecs.bsky.social Queen Mary University of London - research on AI for audio. Website: https://www.seresearch.qmul.ac.uk/cmai/people/ebenetos/
Michel Olvera @michelolzam.bsky.social 🎧 Machine Listening Researcher
Grzegorz Chrupała @grzegorz.chrupala.me Speech • Language • Learning
https://grzegorz.chrupala.me
@ Tilburg University
Gasper Begus @begus.bsky.social Assoc. Professor at UC Berkeley
Artificial and biological intelligence and language
Linguistics Lead at Project CETI 🐳
PI Berkeley Biological and Artificial Language Lab 🗣️
College Principal of Bowles Hall 🏰
https://www.gasperbegus.com
Daan van Esch @daanvanesch.nl I work on speech and language technologies at Google. I like languages, history, maps, traveling, cycling, and buying way too many books.
Catherine Lai @catlai.bsky.social Lecturer in speech and language technology, CSTR, University of Edinburgh. She/her.
https://homepages.inf.ed.ac.uk/clai/
Maureen de Seyssel @maureendeseyssel.bsky.social machine learning researcher @Apple | PhD from @CoML_ENS | speech, ml and cognition.
Yossi Keshet @keshet.bsky.social Speech, language, and deep learning at the Technion. But also psychology, philosophy, and history. And Jazz improv.
François Grondin @francoisgrondin.bsky.social Assistant professor at USherbrooke. Creator of the ODAS framework. Research in speech, multichannel audio processing, robot audition, embedded AI.
francoisgrondin.com
@naoyukikandaslp.bsky.social @naoyukikandaslp.bsky.social
Antoine Deleforge @adeleforge.bsky.social Research scientist at #Inria. Audio signal processing, Acoustics, Machine Learning, Bicycle Riding, Lindy Hop Dancing.
Julian Lenz @jlenzyy.bsky.social Audio AI research engineer w/ Lemonaide. prev. Neutone, Okio. MSc in Audio Computation at UPF. I also fly planes and play the cello sometimes!
Joan Serrà @serrjoa.bsky.social Does research on machine learning at Sony AI, Barcelona. Works on audio analysis, synthesis, and retrieval. Likes tennis, music, and wine.
https://serrjoa.github.io/
Julius Richter @julius-richter.bsky.social Postdoctoral researcher at Meta
Oriol (Uri) Nieto @urinieto.bsky.social Researcher at Adobe Research. Machine learning on audio. Screamer. Oaklander born in Barcelona. Titan. He/they 🌈
www.urinieto.com
Hao Tang @larryniven4.bsky.social Lecturer at the University of Edinburgh. Member of Centre of Speech Technology Research (CSTR).
Justin Salamon @justinsalamon.bsky.social Head of Sound Design AI Research at Adobe. Machine learning and signal processing for audio & video. Musician. He/him.
www.justinsalamon.com
Fernando Espinosa Iñiguez @neuralvocoder.bsky.social Audio ML Research @ Auto-Tune 🎤🎵
Bay Area SSBM & RL gamer
Love to talk Cognitive Science, Linguistics, Bio-inspired Learning, Topological Signal Processing & TDA
SeungHeon Doh @seungheon-doh.bsky.social research on llm + music (https://seungheondoh.github.io/).
PhD Candidate @ Music and Audio Computing Lab, KAIST. Previously an intern @Adobe, @BytedanceTalk, @Naver, @Chartmetric.
Zachary Novack @zacknovack.bsky.social Efficient+Controllable Audio Generation @ UCSD | Interning Stability AI, Adobe | Teaching drums @ POW Percussion
Lancelot @lancelotblanchard.bsky.social Musician, Engineer, AI Researcher - @mitofficial.bsky.social @medialab.bsky.social
@yoshiki-masuyama.bsky.social @yoshiki-masuyama.bsky.social
@gtzan.bsky.social @gtzan.bsky.social
Francesco Paissan @fpaissan.bsky.social research in ML at MERL and Mila
francescopaissan.it
Luca Comanducci @lucacomanducci.bsky.social Fixed-term researcher (RTDA) @polimi
working on audio signal processing, music informatics, spatial audio and generative models (https://lucacoma.github.io/)
Simon Leglaive @sleglaive.bsky.social Tenured Assistant Professor at CentraleSupélec.
Signal processing and machine learning for speech and audio.
sleglaive.github.io
@kzmolikova.bsky.social @kzmolikova.bsky.social
Kyle Kastner @kastnerkyle.bsky.social computers and music are (still) fun
Martijn Bartelds @mbartelds.bsky.social Researcher at Together AI | Formerly at Stanford NLP, University of Groningen, TU Delft and UPenn
Interspeech 2026 @interspeech.bsky.social interspeech2026.org
27 September – 1 October, ICC, Sydney, Australia
'Speaking Together'
Proudly hosted by the Australasian Speech Science and Technology Association (ASSTA) and the International Speech Communication Association (ISCA).
DCASE Challenge @dcase-challenge.bsky.social Challenge on Detection and Classification of Acoustic Scenes and Events.
https://dcase.community/
Bernardo Torres @bernardo-torres.bsky.social PhD Student @ Telecom Paris, ADASP team. Previously research scientist intern @ Deezer, Sony CSL (Music Team).
AI/ML for audio and music signal processing and synthesis.
hugofloresgarcía @hugofloresgarcia.bsky.social human computer musical instruments
https://hugofloresgarcia.art/
phd candidate @northwestern
research intern @adobe
prev @spotify, @descript
chicago // honduras- A Albert Zeyer @albertzeyer.bsky.social Deep Learning, speech recognition, language modeling, https://scholar.google.com/citations?user=qrh5CBEAAAAJ&hl=en
Open source, https://github.com/albertz/
robinsch @fakufaku.bsky.social Farming chili peppers for fun and hot sauce 🌶️
Zhaoheng Ni @nateanl.bsky.social Researcher@Meta Reality Labs, working on generative models, speech enhancement, speech recognition, TTS, etc.
https://nateanl.github.io/
Jordi Pons @jordiponsdotme.bsky.social Music and artificial intelligence.
Researcher at Stability AI.
Musician at BRNRT Collective.
Previously at Dolby and Universitat Pompeu Fabra.
artintech.substack.com
www.jordipons.me
Hao-Wen (Herman) Dong 董皓文 @hermandong.bsky.social Assistant Professor at University of Michigan | PhD from UC San Diego | Human-Centered Generative AI for Content Creation- Y Yoshiaki Bando @yoshipon0520.bsky.social
Matthias Mauch @matthiasmauch.bsky.social I lead music ML research for Music. Flexitalian.
Faro Stöter @faroit.bsky.social AudioML research scientist at https://audioshake.ai, before: post-doc @inria@social.numerique.gouv.fr, Editor at https://bsky.app/profile/joss-openjournals.bsky.social
All in 17.68% of grey, located in Frankfurt (Germany)
Carl Thomé @carlthome.bsky.social Music machine learning, MIR, ML, DSP
Haven Kim @havenkim.bsky.social 1st-year CS PhD student at UCSD
I work on music and ML.
havenpersona.github.io
Saurjya Sarkar @dinosaurjya.bsky.social Ph.D. in Artificial Intelligence and Music, C4DM
https://saurjya.github.io/
Minje Kim @minjekim.bsky.social Audio and AI researcher. Faculty in Siebel School at UIUC and Visiting Academic at Amazon Lab126. A working dad. Some obsolete hobbies: music, photography, drawing, and writing. Still active interests: cooking.
🏠 https://minjekim.com
DJ 🐳 @djain.bsky.social HCI Assistant Professor at UMich researching accessibility, audio AI, sound interaction, XR, and health. Director, Soundability Lab. Previously, Google, Apple, Microsoft, UW, and MIT Media Lab.
https://dhruv-jain.com
Shinji Watanabe @shinjiw.bsky.social I'm working at CMU (2021-). I was working at NTT (2001-2011), MERL (2012-2017), and JHU (2017-2020). Speech and Audio Processing is my main research topic.
Vivek Kumar @v1vekkumar.bsky.social Senior Manager, Foundational Research , @GoogleDeepMind
Googler, Ex @Dolby & @Broadcom
Talks and Investments 👉🏽 http://portfolio.v1vek.com
Zhongweiyang Xu @zhongweiyangxu.bsky.social I’m a PhD student in University of Illinois Urbana-Champaign working on audio inverse problems.
My website: https://xzwy.github.io/alanweiyang.github.io/
Speech and Audio in the Northeast (SANE) @saneworkshop.org Official account for the SANE series of workshops. The one-day events annually gather researchers and students in speech and audio from the Northeast of the American continent, alternately in Boston and NYC.
🌐 saneworkshop.org
Daniele Giacobello @dgiacobello.bsky.social Milanese-Californian Digital Speech and Audio Processing Technologist @ Apple
Constantinos Dimitriou @cnstntns.bsky.social audio & ml research at apple’s music creation apps. previously autotune, audioshake.ai, gracenote.com. also interested in photography, bicycles, and beer.
Mathieu Fontaine @mathfontaine.bsky.social Associate professor at Télécom Paris in machine listening and audio applied to extended reality
Josh McDermott @joshhmcdermott.bsky.social Working to understand how humans and machines hear. Prof at MIT; director of Lab for Computational Audition. https://mcdermottlab.mit.edu/
Gautham Mysore @gauthamjmysore.bsky.social Head of Audio and Video AI Research at Adobe Research
@geoffroypeeters.bsky.social @geoffroypeeters.bsky.social
Hugo Malard @hugomlrd.bsky.social PhD student in multimodal learning for audio understanding at telecom-paris
Ashvala Vinay @handle.invalid Current: Co-founder, NoneType (https://nonety.pe). Prev: PhD and Masters (music tech) @ Georgia Tech, Bachelors (EPD) @ Berklee, Keyo and LiveAds. Knower of guitars.
Xiaoyu Bie @xiaoyubie.bsky.social Postdoc Researcher @Télécom Paris, Institut Polytechnique de Paris
Prev PhD @INRIA Prev intern @Meta @Baidu
I work on generative models and audio applications
https://xiaoyubie1994.github.io/
@danstowell.bsky.social @danstowell.bsky.social Associate Professor of Computational Biodiversity.
Leiden University (LIACS) and
Naturalis Biodiversity Centre, the Netherlands