Scott Wisdom @ScottTWisdom
Research scientist at @GoogleAI working on sound separation Joined July 2011-
Tweets20
-
Followers186
-
Following121
-
Likes49
Veo 3 is here, and in addition to better visuals, it makes noises and speaks! This was a massive effort made possible by incredible passion from the whole Veo team and the many other team enabling it to launch today. Looking forward to seeing what others do with it! #veo3
Veo 3, our SOTA video generation model, has native audio generation and is absolutely mindblowing. For filmmakers + creatives, we’re combining the best of Veo, Imagen and Gemini into a new filmmaking tool called Flow. Ready today for Google AI Pro and Ultra plan subscribers.
We're sharing progress on our video-to-audio (V2A) generative technology. 🎥 It can add sound to silent clips that match the acoustics of the scene, accompany on-screen action, and more. Here are 4 examples - turn your sound on. 🧵🔊 dpmd.ai/v2a
It's so awesome to see the impact of the computational audio capabilities we developed featured in @madebygoogle 🎉 🎉 🎉 Congrats to John Hershey, @ScottTWisdom, @PGetreuer & everyone who contributed for pioneering new computational audio capabilities in Pixel8 #MadeByGoogle
Check out the 4 new Google Photos features coming first to Pixel 8 and 8 Pro ↓ Whether it’s noise from wind, traffic, or barking dogs, Audio Magic Eraser in Google Photos reduces distracting sounds in your video in just a few taps! 🪄
Sorry it took forever (I did the editing this year...): videos of all #SANE2022 talks by @TweetRupal @mhnt1580 @ScottTWisdom @tnsainath @shinjiw_at_cmu @anoopcherian @gan_chuang are finally available! Here's the essential binge-watching YouTube playlist👇 youtube.com/playlist?list=…
Strong showing at #SANE2022 to learn about the latest and greatest in speech and audio research from a stellar lineup!
Here is a short presentation of AudioScopeV2!📢 @ScottTWisdom and I are looking forward to discussing further about open-domain on-screen sound separation and meeting you in #ECCV2022! webpage:google-research.github.io/sou... arxiv:arxiv.org/abs/2207.10141 video:youtu.be/6UgcS3NdPn8
Full list of speakers and talk details for #SANE2022 (Thursday 10/6, Cambridge, MA) now available! @anoopcherian @gan_chuang @mhnt1580 @TweetRupal @tnsainath @shinjiw_at_cmu @ScottTWisdom Poster & demo submissions due 9/21. Registration/Details: saneworkshop.org
I am 😃 that we will present AudioScopeV2 at #ECCV2022! If you want to learn about improved audio-visual attention models and calibration for on-screen sound separation check our paper w. @ScottTWisdom! project-page: google-research.github.io/sound-separati… new dataset: github.com/google-researc…
``AudioScopeV2: Audio-Visual Attention Architectures for Calibrated Open-Domain On-Screen Sound Separation. (arXiv:2207.10141v1 [cs.SD]),'' Efthymios Tzinis, Scott Wisdom, Tal Remez, John R. Hershey, ift.tt/jOrEQWR
Distance-Based Sound Separation abs: arxiv.org/abs/2207.00562 project page: google-research.github.io/sound-separati… With a single nearby speaker and four distant speakers, the model improves scale-invariant signal to noise ratio by 4.4 dB for near sounds and 6.8 dB for far sounds
SANE is back! Thursday, Oct. 6 in Kendall Square, Cambridge, MA. Confirmed speakers: A. Cherian @anoopcherian, C. Gan @gan_chuang, W.-N. Hsu @mhnt1580, T. Sainath @tnsainath, S. Watanabe @shinjiw_at_cmu, S. Wisdom @ScottTWisdom. More details: saneworkshop.org
Happy to see my summer work with @ScottTWisdom, Hakan Erdogan, and John Hershey was accepted for presentation at @ieeeICASSP 2022 😊 My first ICASSP paper in the books! Immensely thankful for their mentorship. Our first version can be found on arXiv at: arxiv.org/abs/2110.10739
We can learn a lot about our environment just by listening to the birds. New #GoogleAI approaches can help isolate and identify birdsongs, helping ecologists better understand food systems and forest health. 🐦 ai.googleblog.com/2022/01/separa…
Our paper received a #WASPAA2021 special award for *Best Audio Representation Learning Paper*: "Self-Supervised Learning from Automatically Separated Sound Scenes". 🎉🚀 paper: arxiv.org/abs/2105.02132 talk: youtu.be/Tts5vYmGwUY slides: bit.ly/3lBjAnr 👇
Our DF-Conformer paper has received the “Best Speech Enhancement Paper Award” from #WASPAA2021! Yay!!
🔊Here's the video presentation of our WASPAA21 paper: "Self-Supervised Learning from Automatically Separated Sound Scenes". Work done during an internship at Google Research. paper: arxiv.org/abs/2105.02132 video: youtu.be/Tts5vYmGwUY slides: bit.ly/3lBjAnr
🔊Happy to announce FSD50K: the new open dataset of human-labeled sound events! Over 51k Freesound audio clips, totalling over 100h of audio manually labeled using 200 classes drawn from the AudioSet Ontology. Paper: arxiv.org/pdf/2010.00475… Dataset: doi.org/10.5281/zenodo…
I am thrilled to announce that our paper "Unsupervised Sound Separation using Mixtures of Mixtures" got accepted to #NeurIPS2020 as a #Spotlight paper!! 📢📢 All kudos to @ScottTWisdom and the rest of the Google guys! arxiv.org/pdf/2006.12701…
We are very happy to announce that all the videos of our recent #ICML2020 workshop on self-supervised learning are now publicly available at slideslive.com/icml-2020/self… Thanks #ICML2020 and @SlidesLive for that! @MILAMontreal #DeepLearning #AI #Speech #MachineLearning
``AudioScopeV2: Audio-Visual Attention Architectures for Calibrated Open-Domain On-Screen Sound Separation. (arXiv:2207.10141v1 [cs.SD]),'' Efthymios Tzinis, Scott Wisdom, Tal Remez, John R. Hershey, ift.tt/jOrEQWR
Glad you like it, thanks for the nice summary!
I'm a bit late posting this, but a very cool paper from Scott Wisdom and collaborators (including @ETzinis) out of Google introducing "MixIT": arxiv.org/abs/2006.12701 They tackle the problem of *unsupervised* source separation! 1/10
Yi Zhong @yiz_be_building
483 Followers 1K Following Building @besimple_ai (YC P25) 🍊 to make AI listen to you, ex Meta / Dropbox / MSFT Product Lead from MIT
Johnathan Lyon (they/... @boxofbox
341 Followers 2K Following hybrid engineer/scientist/artist -- into: software//sonics//noise//weird//absurd//electricity//lexicon
super intelligence @eacc72
15 Followers 2K Following
zuoli @zuoliao11
70 Followers 123 Following 音声強調・分離など / Kaggle Competitions Master🥇1🥈2🥉1 / 🇺🇿 / SB Intuitions
Yassine El Kheir @YassineElkheir
62 Followers 606 Following PhD Student at DFKI & Technical University of Berlin
David Braun @DoItRealTime
2K Followers 446 Following PhD candidate @PrincetonCS audio ML. @CCRMA/@Stanford, @BrownUniversity
Salah Zaiem @salah_zaiem
772 Followers 2K Following Research Scientist @GoogleDeepMind working on audio-visual generation. Veo/Omni
Kranti Kumar Parida @KrantiParida
64 Followers 156 Following
Rotnodip Sarkar @rotnodip
11 Followers 794 Following
Yoram Bachrach @yorambac
4K Followers 7K Following Research Scientist at Meta (prev Google DeepMind and Microsoft Research). Working on LLM Agents and Multi-Agent Systems.
Teetaj Pavaritpong @TeetajP
32 Followers 1K Following Software Engineer | AI/ML | UIUC ‘24 B.S. in CS & Stats @SiebelSchool
Şükrü Fırat Sarp @FtrtS
64 Followers 3K Following
Enric Corona @enric_corona
1K Followers 1K Following Generative models at @GoogleDeepMind - Video/Audio generation Projects: Gemini #Omni, #veo3, Veo 2 Capabilities
JAEHYEONG_KIM @jhk40160806
830 Followers 7K Following I'm a technical imagineer: If MyBrain ideas + great Scientists meet, I & Scientists make New things, if MyBrain ideas can link(implant) Neuro into AI quantum 📩
Siddhant Arora @Sid_Arora_18
935 Followers 803 Following Research Scientist @Meta, PhD @LTIatCMU; Silver Medalist @iitdelhi
Music, Mind and Brain... @musicmindbrain
681 Followers 1K Following Official Twitter of the Music, Mind and Brain (MMB) Group at @GoldsmithsUoL, University of London. Tweets by @DianaOmigie, @AngladaTort, and group members.
Sasha @Nechorsm8DK6
30 Followers 890 Following Originally from Malaysia, now living in the UK, runs a clothing shop
Matan Gover @matangover
540 Followers 1K Following Music/audio ML researcher, musician, software engineer
Janek Ebbers @EbbersJanek
20 Followers 36 Following
Mirco Ravanelli @mirco_ravanelli
4K Followers 2K Following Deep learning for Conversational AI. Creator of SpeechBrain.
Rob Smith @robmsmt
69 Followers 1K Following Interested in ML, Engineering and Science. Like to write code. 🇬🇧 🇺🇸 https://t.co/kVQuiGMlja
Arda Senocak @ardasnck
228 Followers 449 Following Assistant Professor, UNIST https://t.co/zewMlmFRZ0
francesco paissan @fpaissan_
463 Followers 283 Following student of science @unitrento // ml @fbk_research @speechbrain1, @mila_quebec, @infn_
Jade Martin @Jade6467
268 Followers 4K Following Investment Analyst,Choice is more important than effort, and the direction of steps is more important than .Love fitness, travel, reading,Like energy issues.
Taishi Nakashima (中... @_tai_shi
591 Followers 863 Following Assistant Professor at Tokyo Metropolitan University, 東京都立大学 システムデザイン学部 助教, 音源分離の研究してます🎛️
Roshan Sharma @RoshanSSharma2
411 Followers 372 Following Research Scientist @GoogleDeepMind | PhD @CMU_ECE | #SpeechProc #NLProc | Previously @AIatMeta @Qualcomm
Satyajeet Prabhu @satyajeetprabhu
4 Followers 245 Following
Phyra @PhyraUK
978 Followers 6K Following Music, Ai, art. Warner Music - Emerging Tech Strategy. (opinions expressed my own)
SGM @234Sagyboy
177 Followers 3K Following
Adrián Aquino @aquino4A
39 Followers 546 Following DE Graduate Student in Acoustics (Penn State). NVH Simulation Engineer. Opinions my own. He/him. 4A.
Brian Hamilton @b_hamilton2
358 Followers 351 Following Acoustics. previously @AAG_Edinburgh. All views are my own. He/him. https://t.co/D5Wh97HBcK.
Edvin LZ @eduniw
381 Followers 2K Following ML Research Scientist, PhD from KTH Royal Institute of Technology. Previously at RISE Research Institutes of Sweden.
Hanoi Hantrakul @yaboihanoi
1K Followers 311 Following Music AI Research Scientist @tiktok_us | AI Song Contest 2022 Winner | @yaboihanoi on all socials | Previously @GoogleAI and @GoogleMagenta Opinions are my own.
Satwik Dutta @the_satwikdutta
100 Followers 925 Following @QuadFellowship Cohort 2023 | Ph.D. Student @UT_Dallas | Speech Processing for Early Childhood | @ASAStudents 22-24 | @iscaSAC 23-25 | Policy @scientistsorg
liang wen @liang001_wen
41 Followers 1K Following
Saurabh Kataria @saukataria14
50 Followers 505 Following Postdoc @EmoryUniversity. Prev: ML Audio PhD @JohnsHopkins, EE @IITKanpur. Personal account. RT,like,casual cmnt!=endorse
Tomas Gajarsky @TomasGajarsky
87 Followers 5K Following
Salah Zaiem @salah_zaiem
772 Followers 2K Following Research Scientist @GoogleDeepMind working on audio-visual generation. Veo/Omni
Haohe Liu @LiuHaohe
2K Followers 497 Following Research Scientist at @Meta SuperIntelligence Lab (FAIR); Speech LLMs and Conversational AI. https://t.co/qmZe2lwexv
IEEE WASPAA 2025 @IEEE_WASPAA
258 Followers 82 Following IEEE Workshop on Applications of Signal Processing to Audio and Acoustics
DailyAudioPapers @mlsp4audio
765 Followers 650 Following Daily tweets on selected arXiv papers on audio (eess․AS/cs․SD) | Brief reviews of interesting papers | Machine learning | Signal processing
Christian Steinmetz @csteinmetz1
6K Followers 2K Following Research Scientist @ Suno // working on generative music • audio fidelity • signal processing • ML
Nikhil Singh @nikhilsinghmus
564 Followers 492 Following Assistant Professor @DartmouthCS. Human-AI Systems (https://t.co/X8LV4tiS1S). Prev: PhD @MIT. @allen_ai @NetflixResearch @berkleecollege.
Andrew Rouditchenko �... @arouditchenko
469 Followers 572 Following Research scientist at NVIDIA working on speech and multi-modal LLMs. Previously PhD student at MIT CSAIL and intern at @AIatMeta and @Apple MLR.
Jenelle Feather @jenellefeather
615 Followers 124 Following Flatiron Research Fellow @FlatironCCN. PhD from @mitbrainandcog. Incoming Asst Prof @CarnegieMellon in Fall 2025. I study how humans and computers hear and see.
Minje Kim @minje_research
423 Followers 244 Following Associate Professor at CS@UIUC; Visitic Academic at Amazon Lab126; Want to share my thoughts on audio & AI research, graduate studies, and life.
Donald S. Williamson @TheASPIREGrpOSU
112 Followers 134 Following I am an associate professor in the Department of Computer Science and Engineering at The Ohio State University. I lead The ASPIRE Group.
Xuanjun (Victor) Chen @xjchen_ntu
43 Followers 322 Following PhD Candidate, NTU EECS (@NTU_TW). Now: @ntu_spml, advised by @HungyiLee2.
Chang-Bin Jeon @jeonchangbin49
330 Followers 362 Following Staff Engineer at Samsung Electronics, PhD from MARG Seoul National University. Previous intern @merl_news @Gaudiolab.
Muqiao Yang @muqiaoy
75 Followers 114 Following Research Scientist @GoogleDeepMind. Past: PhD @CarnegieMellon, BEng @HongKongPolyU
Ankit Shah @ankits0052
2K Followers 8K Following Full Stack LLM Associate Director. Ph.D. @LTIatCMU. Sharing insights about AI research, LLMs, multimodal AI, coding & tech. 🚀 Views are my own
HEAR Benchmark @hearbenchmark
443 Followers 94 Following HEAR Benchmark. Holistic Evaluation of Audio Representations.
Nauman Dawalatabad @NaumanDawalatab
459 Followers 976 Following Voice AI @Zoom | @MIT_CSAIL | @SpeechBrain1 @Mila_Quebec | @SamsungResearch | IIT Madras @iitmcse @iitmadras
Michael I Mandel @asterix77
196 Followers 295 Following
Xuankai @kaikai0019
191 Followers 253 Following PhD student @WavLab @LTIatCMU @SCSatCMU, working on speech processing
Puyuan Peng @PuyuanPeng
2K Followers 522 Following Research Scientist @Meta Superintelligence Lab. Speech & Audio. Previously @utaustin @uchicago @bnu_1902
Aswin Subramanian @S_Aswin19
267 Followers 757 Following Sr Applied Scientist @Microsoft, PhD from JHU @jhuclsp @jhuece
Nick Braun @njbraun
51 Followers 46 Following
Aswin Sivaraman @actuallyaswin
670 Followers 3K Following PhD @IULuddy. BS @ECEIllinois. Tweets about machine learning, music, movies, and video games. Sometimes politics. Hot takes are my own. (He/Him)
Desh Raj @rdesh26
4K Followers 2K Following Speech + LLMs @nvidia | Previously: @Meta MSL, @jhuclsp, @IITGuwahati
Francesca Ronchini @FraRonchini
173 Followers 390 Following PhD student @IsplPolimi - @polimi Former AI Research Intern @BoschResearch Exploring text-based generative models for audio and music applications
no context memes @nocontextmemes
2.9M Followers 607 Following memes | dm for promo | @dailyhopecores
Gautham Mysore @GauthamMysore
777 Followers 304 Following Head of Audio and Video AI Research @AdobeResearch
IEEE ICASSP @ieeeICASSP
5K Followers 1 Following IEEE International Conference on Acoustics, Speech, and Signal Processing. #ICASSP2027 will be held 16-21 May 2027 in Toronto, Canada.
IEEE Signal Processin... @IEEEsps
5K Followers 218 Following IEEE's first society, the Signal Processing Society is the world’s premier professional society for signal processing scientists and professionals since 1948.
Shinji Watanabe @shinjiw_at_cmu
5K Followers 371 Following I'm working at CMU (2021-). I was working at NTT (2001-2011), MERL (2012-2017), and JHU (2017-2020). Speech and Audio Processing is my main research topic.
AI Conference Deadlin... @AiDeadline
1K Followers 1 Following Reminding you of the submission deadline for the closest AI Conferences each day.
Qiuqiang Kong @QiuqiangK
1K Followers 245 Following Assistant Professor at @CUHKofficial, previously at @ByteDanceTalk, Ph.D. at @UniOfSurrey
Francois Grondin @fgrondinmit
719 Followers 951 Following Assistant prof. at USherbrooke, former postdoc at MIT-CSAIL. Creator of ODAS (https://t.co/rSXpnhrn5c). Contributor to SpeechBrain (https://t.co/IozUkIxhwV).
Jonah Casebeer @CasebeerJonah
252 Followers 423 Following Senior Researcher @PhysicalAI | Ph.D. @IllinoisCS | personal account
Samuele Cornell @SamueleCornell
982 Followers 524 Following Post-doc @ CMU LTI. Audio and speech researcher.
Tejas Kulkarni @tejasdkulkarni
22K Followers 2K Following Solving World Models @GoogleDeepMind. ex CEO @CSM_ai. Interested in AGI, Brain and AI creativity. PhD @mitbrainandcog
Ben Poole @poolio
23K Followers 2K Following ex-google brain and deepmind. phd in neural nonsense from stanford.
Zhepei Wang @zhepeiw03
20 Followers 57 Following
Félix de Chaumont Qu... @FelixCQ
113 Followers 382 Following Research engineer @GoogleDeepMind. Large language models for audio generation.
Yuxuan Wang @log_pie
193 Followers 2K Following
Soroosh Mariooryad @sorooshooryad
234 Followers 411 Following Staff Research Scientist @GoogleDeepMind Gemini Audio core team ♊🌊 Generative speech and language research at GDM foundational research unit
Eric Battenberg @EricBattenberg
786 Followers 133 Following Generative Models of Speech and Audio @GoogleDeepMind Frontier AI https://t.co/BrshqMKe25
NeurIPS Conference @NeurIPSConf
162K Followers 41 Following Sydney Dec 6-12, 26, Paris and Atlanta. Tweets to this account are not monitored. Please send feedback to [email protected].
ICLR @iclr_conf
61K Followers 57 Following International Conference on Learning Representations #ICLR2027. SPC is @jacobandreas and GC is @BharathHarihar3
Nicholas J. Bryan @NicholasJBryan
1K Followers 520 Following Head of Music AI / Principal Scientist, Adobe Research (personal account)




























