We are organizing a special session at #Interspeech2025 on: Interpretability in Audio & Speech Technology
Check out the special session website: sites.google.com/view/intersp...
Paper submission deadline 📆 12 February 2025
We are organizing a special session at #Interspeech2025 on: Interpretability in Audio & Speech Technology
Check out the special session website: sites.google.com/view/intersp...
Paper submission deadline 📆 12 February 2025
#interspeech #speech #SpeechTech #SpeechScience
#interspeech #speech #SpeechTech #SpeechScience
🆕 New at IWSLT! But no less exciting 🔥
🎯 Goal: Compress a large, general-purpose multimodal model, making speech translation more efficient ⚡️, deployable 📲, and sustainable ♻️, while preserving translation quality ⭐️
#AI #SpeechTech #ModelCompression #LLMcompression
🆕 New at IWSLT! But no less exciting 🔥
🎯 Goal: Compress a large, general-purpose multimodal model, making speech translation more efficient ⚡️, deployable 📲, and sustainable ♻️, while preserving translation quality ⭐️
#AI #SpeechTech #ModelCompression #LLMcompression
#speech #speechtech #audio
arxiv.org/abs/2502.11946
#speech #speechtech #audio
👉 github.com/pnlpal/pnl-r...
#TTS #VoiceAI #PNLReader #Sverige #svenska #Sweden #SpeechTech #webdev #devlog #buildinpublic #indiedev
www.youtube.com/watch?v=7nV0...
👉 github.com/pnlpal/pnl-r...
#TTS #VoiceAI #PNLReader #Sverige #svenska #Sweden #SpeechTech #webdev #devlog #buildinpublic #indiedev
www.youtube.com/watch?v=7nV0...
Moving beyond sustained vowels & read speech, our study uses CNNs on MFCC features from natural, spontaneous speech—capturing real-world acoustic nuances & reaching ~92% eval ACC.
link.springer.com/article/10.1...
#SpeechTech #HealthAI
Moving beyond sustained vowels & read speech, our study uses CNNs on MFCC features from natural, spontaneous speech—capturing real-world acoustic nuances & reaching ~92% eval ACC.
link.springer.com/article/10.1...
#SpeechTech #HealthAI
Our TACL paper analyzes 110 works and reveals:
🚫 Overreliance on short-form speech
🌀 Terminology chaos
📉 Real-world deployment gaps
We bring order-New taxonomy, trends & recommendations!
📍#ACL2025 Poster: Monday 11-12:30, Hall 4/5
#Speech #SpeechTech
Our TACL paper analyzes 110 works and reveals:
🚫 Overreliance on short-form speech
🌀 Terminology chaos
📉 Real-world deployment gaps
We bring order-New taxonomy, trends & recommendations!
📍#ACL2025 Poster: Monday 11-12:30, Hall 4/5
#Speech #SpeechTech
#SpeechTech #AIForGood
#SpeechTech #AIForGood
Le #LLL a remis un #corpus unique: 1400 heures de données orales en #kreyòl, transcrites et alignées automatiquement au caractère près 🎧✨
@univorleans.bsky.social
#créolehaïtien #speechtech #NLP
Le #LLL a remis un #corpus unique: 1400 heures de données orales en #kreyòl, transcrites et alignées automatiquement au caractère près 🎧✨
@univorleans.bsky.social
#créolehaïtien #speechtech #NLP
#SpeechProcessing #LLM #SFM #NLProc #speechtech #audio
#SpeechProcessing #LLM #SFM #NLProc #speechtech #audio
#AI #MistralAI #Voxtral #OpenSource #VoiceAI #GenerativeAI #SpeechTech
winbuzzer.com/2025/07/15/m...
#AI #MistralAI #Voxtral #OpenSource #VoiceAI #GenerativeAI #SpeechTech
winbuzzer.com/2025/07/15/m...
🎯 Goal – Provide a stable & shared evaluation framework to track advances in Spoken Language Translation from English into multiple languages & domains. 🌍🎙️
#AI #SpeechTech #SpeechTranslation #IWSLT2025
🎯 Goal – Provide a stable & shared evaluation framework to track advances in Spoken Language Translation from English into multiple languages & domains. 🌍🎙️
#AI #SpeechTech #SpeechTranslation #IWSLT2025
Don't forgot to submit your work to the special session on Interpretability in Audio & Speech Technology, if it fits the theme
We are looking forward to see exciting submissions ✨
#SpeechTech #SpeechScience
We are organizing a special session at #Interspeech2025 on: Interpretability in Audio & Speech Technology
Check out the special session website: sites.google.com/view/intersp...
Paper submission deadline 📆 12 February 2025
Don't forgot to submit your work to the special session on Interpretability in Audio & Speech Technology, if it fits the theme
We are looking forward to see exciting submissions ✨
#SpeechTech #SpeechScience
🎯 Goal: The IWSLT 2025 Subtitling Track challenges participants to generate accurate Arabic & German subtitles for English audiovisual recordings, bridging language gaps in media! 🌍📺
#AI #SpeechTech #Subtitling #IWSLT2025
🔗: iwslt.org/2025/subtitl...
🎯 Goal: The IWSLT 2025 Subtitling Track challenges participants to generate accurate Arabic & German subtitles for English audiovisual recordings, bridging language gaps in media! 🌍📺
#AI #SpeechTech #Subtitling #IWSLT2025
🔗: iwslt.org/2025/subtitl...
We are thrilled to announce that Prof. Karen Livescu will keynote our Special Session on Interpretable Audio and Speech Models at #Interspeech2025:
"What can interpretability do for us (and what can it not)?"
🗓️ Aug 18, 11:00
@interspeech.bsky.social
We are thrilled to announce that Prof. Karen Livescu will keynote our Special Session on Interpretable Audio and Speech Models at #Interspeech2025:
"What can interpretability do for us (and what can it not)?"
🗓️ Aug 18, 11:00
@interspeech.bsky.social
#NLProc #Speech #instructionfollowing #zeroshot #speechtech #speechllm
arxiv.org/abs/2412.01145
#NLProc #Speech #instructionfollowing #zeroshot #speechtech #speechllm
voicetechnology.substack.com/p/dutch-spee...
#AcademicSky
voicetechnology.substack.com/p/dutch-spee...
#AcademicSky
voicetechnology.substack.com/p/phd-on-speech-technology-for-frisian?r=2fcmj9&utm_campaign=post&utm_medium=web&triedredirect=true
voicetechnology.substack.com/p/phd-on-speech-technology-for-frisian?r=2fcmj9&utm_campaign=post&utm_medium=web&triedredirect=true
Working on Frisian-Dutch bilingual speech + AI at Fryske Akademy/Campus Fryslân. Fully funded, 4 years, starts Sept 2025.
More info ⬇️
open.substack.com/pub/voicetec...
#SpeechTech #PhD #LowResourceLanguages #AcademicJobs #AcademicSky
Working on Frisian-Dutch bilingual speech + AI at Fryske Akademy/Campus Fryslân. Fully funded, 4 years, starts Sept 2025.
More info ⬇️
open.substack.com/pub/voicetec...
#SpeechTech #PhD #LowResourceLanguages #AcademicJobs #AcademicSky
The open-source tool, which is going to be released soon, natively supports any speech-to-text #HuggingFace models! 🤖
#SpeechTech #Translation
The open-source tool, which is going to be released soon, natively supports any speech-to-text #HuggingFace models! 🤖
#SpeechTech #Translation