#SpeechTech
📣 #SpeechTech & #SpeechScience people

We are organizing a special session at #Interspeech2025 on: Interpretability in Audio & Speech Technology

Check out the special session website: sites.google.com/view/intersp...

Paper submission deadline 📆 12 February 2025
December 6, 2024 at 9:30 PM
Come work with us! Our department is hiring an associate/assistant prof in language and speech technology www.ru.nl/en/working-a...
#interspeech #speech #SpeechTech #SpeechScience
Associate/Assistant Professor: Language and Speech Technology | Radboud University
Do you want to work as a Associate/Assistant Professor: Language and Speech Technology at the Faculty of Arts? Check our vacancy!
www.ru.nl
February 19, 2025 at 12:26 PM
Today's task: model compression!!
🆕 New at IWSLT! But no less exciting 🔥

🎯 Goal: Compress a large, general-purpose multimodal model, making speech translation more efficient ⚡️, deployable 📲, and sustainable ♻️, while preserving translation quality ⭐️
#AI #SpeechTech #ModelCompression #LLMcompression
January 29, 2025 at 4:47 PM
Our pick of the week by @sarapapi.bsky.social: "Step-Audio: Unified Understanding and Generation in Intelligent Speech Interaction" by StepFun (2025).

#speech #speechtech #audio
February 20, 2025 at 3:17 PM
🚀 I can’t resist adding these two ‘Realistic AI’ voices to PNL Reader. Hey Swedes, how does this sound to you?

👉 github.com/pnlpal/pnl-r...

#TTS #VoiceAI #PNLReader #Sverige #svenska #Sweden #SpeechTech #webdev #devlog #buildinpublic #indiedev

www.youtube.com/watch?v=7nV0...
Realistic AI voices on PNL Reader
YouTube video by Programming N' Language
www.youtube.com
November 15, 2025 at 3:10 PM
Yay! KB-Whisper launched today - a freely available speech-to-text service able to transcribe many varieties of Swedish. Available now at Hugginface! #NLP #speechtech www.dn.se/kultur/kungl...
Kungliga biblioteket lanserar AI som transkriberar tal
Slutet är nära för roliga och pinsamma feltranskriberingar signerad AI. Modellen KB-Whisper beskrivs som en milstolpe för taligenkänning på svenska.
www.dn.se
February 21, 2025 at 1:04 PM
🎙️ Can spontaneous speech improve voice pathology detection?

Moving beyond sustained vowels & read speech, our study uses CNNs on MFCC features from natural, spontaneous speech—capturing real-world acoustic nuances & reaching ~92% eval ACC.

link.springer.com/article/10.1...

#SpeechTech #HealthAI
Voice pathology detection on spontaneous speech data using deep learning models - International Journal of Speech Technology
Speech problems are a common issue that affects people everywhere and can affect the quality of their lives. The human speech production system involves various components. Dysfunction of any of these...
link.springer.com
September 25, 2026 at 8:05 AM
🔥 Is your real-time SimulST system REAL?

Our TACL paper analyzes 110 works and reveals:
🚫 Overreliance on short-form speech
🌀 Terminology chaos
📉 Real-world deployment gaps
We bring order-New taxonomy, trends & recommendations!

📍#ACL2025 Poster: Monday 11-12:30, Hall 4/5

#Speech #SpeechTech
July 27, 2025 at 1:18 PM
🎙️ Challenge Alert: The Speech Accessibility Project Challenge at #Interspeech2025 focuses on advancing dysarthric speech recognition! Compete to build the best ASR using a 290-hour dataset. 🏆 Prizes for the lowest WER & highest semantic score. Details: eval.ai/web/challeng...
#SpeechTech #AIForGood
November 21, 2024 at 4:23 PM
Heureux d'accueillir le prof. Renauld Govain dans le cadre du projet #ANR CREAM sur les langues créoles !
Le #LLL a remis un #corpus unique: 1400 heures de données orales en #kreyòl, transcrites et alignées automatiquement au caractère près 🎧✨
@univorleans.bsky.social
#créolehaïtien #speechtech #NLP
December 8, 2025 at 11:09 AM
📣 Call for Tutorials: #Interspeech2025! 🧠 Share your expertise on speech science & technology. Proposals due Feb 1, 2025. Tutorials will be held on Aug 17, 2025 in Rotterdam. 🌍 Details & submission info: www.interspeech2025.org/call-for-tut... 🎙️ #SpeechScience #AI #SpeechTech
November 19, 2024 at 5:08 PM
Our pick of the week by @mgaido91.bsky.social: "OpenOmni: Advancing Open-Source Omnimodal Large Language Models with Progressive Multimodal Alignment and Real-Time Self-Aware Emotional Speech Synthesis" by Luo et al. (2025)

#SpeechProcessing #LLM #SFM #NLProc #speechtech #audio
Interesting to see multimodal LLM built by combining modality encoders and LLM with adapters, as in the SFM+LLM paradigm, independently for each modality. This modularity may ease the creation of more MLMs from collaborations of single-modality experts. arxiv.org/abs/2501.04561
https://arxiv.org/abs/2501.04561
t.co
April 16, 2025 at 1:28 PM
Mistral Challenges OpenAI and Google with New Voxtral Open-Source Voice AI Model

#AI #MistralAI #Voxtral #OpenSource #VoiceAI #GenerativeAI #SpeechTech

winbuzzer.com/2025/07/15/m...
July 15, 2025 at 5:07 PM
Today's task is an IWSLT mainstay: Offline ST!

🎯 Goal – Provide a stable & shared evaluation framework to track advances in Spoken Language Translation from English into multiple languages & domains. 🌍🎙️

#AI #SpeechTech #SpeechTranslation #IWSLT2025
January 31, 2025 at 5:32 PM
Today is #Interspeech2025 deadline

Don't forgot to submit your work to the special session on Interpretability in Audio & Speech Technology, if it fits the theme

We are looking forward to see exciting submissions ✨

#SpeechTech #SpeechScience
📣 #SpeechTech & #SpeechScience people

We are organizing a special session at #Interspeech2025 on: Interpretability in Audio & Speech Technology

Check out the special session website: sites.google.com/view/intersp...

Paper submission deadline 📆 12 February 2025
February 12, 2025 at 8:50 AM
Our last task: Subtitling!!

🎯 Goal: The IWSLT 2025 Subtitling Track challenges participants to generate accurate Arabic & German subtitles for English audiovisual recordings, bridging language gaps in media! 🌍📺
#AI #SpeechTech #Subtitling #IWSLT2025

🔗: iwslt.org/2025/subtitl...
Subtitling track
Home of the IWSLT conference and SIGSLT.
iwslt.org
February 5, 2025 at 4:52 PM
📢 #SpeechTech & #SpeechScience researchers!
We are thrilled to announce that Prof. Karen Livescu will keynote our Special Session on Interpretable Audio and Speech Models at #Interspeech2025:
"What can interpretability do for us (and what can it not)?"
🗓️ Aug 18, 11:00
@interspeech.bsky.social
Announcements
Keynote Speaker Announcement 🔊 30.07.2025 We are delighted to announce the keynote speech t`hat will happen at the special session! Speaker: Prof. Karen Livescu, Toyota Technological Institute at Ch...
sites.google.com
July 30, 2025 at 6:25 PM
Our pick of the week by @mgaido91.bsky.social: "AlignFormer: Modality Matching Can Achieve Better Zero-shot Instruction-Following Speech-LLM" by Ruchao Fan, Bo Ren, Yuxuan Hu, Rui Zhao, Shujie Liu, Jinyu Li (2024).

#NLProc #Speech #instructionfollowing #zeroshot #speechtech #speechllm
February 14, 2025 at 10:38 AM
Microsoft's expressive MAI-Voice-1 model now live in Copilot Daily & Podcasts enterprise AI expands! #AI #Microsoft #Copilot #SpeechTech
September 3, 2025 at 8:31 AM
Great turnout at #DutchSpeechTechDay! Moderated a panel on inclusive speech tech w leaders from academia & industry. Highlight: seeing how Dutch research addresses accessible speech solutions. The future of #SpeechTech is diverse 🎯
voicetechnology.substack.com/p/dutch-spee...
#AcademicSky
Dutch Speech Tech Day 2025
Speech in Context: Advancing Technology for Diverse Human Communication
voicetechnology.substack.com
February 13, 2025 at 12:13 PM
We invite applications for a #PhDposition on #SpeechTech to support Dutch farmers at the University of Groningen (Campus Fryslân) -- more info below. #academicsky

voicetechnology.substack.com/p/phd-on-speech-technology-for-frisian?r=2fcmj9&utm_campaign=post&utm_medium=web&triedredirect=true
PhD on speech technology for Frisian farmers
PhD oer spraaktechnology foar Nederlânske boeren
voicetechnology.substack.com
November 26, 2024 at 1:35 PM
🎓 Open PhD Position: Speech Tech for Minority Languages

Working on Frisian-Dutch bilingual speech + AI at Fryske Akademy/Campus Fryslân. Fully funded, 4 years, starts Sept 2025.

More info ⬇️
open.substack.com/pub/voicetec...

#SpeechTech #PhD #LowResourceLanguages #AcademicJobs #AcademicSky
Innovating Speech Tech for Smaller Languages
New PhD position combining AI and bilingual speech research in the Netherlands
open.substack.com
February 16, 2025 at 11:47 AM
Marco Gaido and Roldano Cattoni presenting our SimulStream Demo at the DI Center Demo Day at FBK!

The open-source tool, which is going to be released soon, natively supports any speech-to-text #HuggingFace models! 🤖

#SpeechTech #Translation
October 10, 2025 at 8:39 AM