Do AI audio embeddings *hear* timbre like we do?
➡️ Benchmarked 18 reps vs 2.6 K human ratings (21 datasets)
🏅 Style embeddings from CLAP & our sound-matching model are best aligned!
Paper: arxiv.org/abs/2507.07764
#ISMIR2025 #MIR #AudioAI #SonyCSLMusic
Do AI audio embeddings *hear* timbre like we do?
➡️ Benchmarked 18 reps vs 2.6 K human ratings (21 datasets)
🏅 Style embeddings from CLAP & our sound-matching model are best aligned!
Paper: arxiv.org/abs/2507.07764
#ISMIR2025 #MIR #AudioAI #SonyCSLMusic
📜 Paper: arxiv.org/pdf/2411.19806
Thx to my colleagues Alain Riou, Geoffroy Peeters, Gaetan Hadjeres and Antonin Gagneré!
🎶 SonyCSLMusic 🎶
📜 Paper: arxiv.org/pdf/2411.19806
Thx to my colleagues Alain Riou, Geoffroy Peeters, Gaetan Hadjeres and Antonin Gagneré!
🎶 SonyCSLMusic 🎶
Surprisal can be used for segment boundary detection and to simulate the information processing of a listener. 🎶 🧠
📜 Link to the paper: arxiv.org/pdf/2501.07474
Model weights are soon to come! 🏋️
💫✨ #SonyCSLMusic 💫✨
Surprisal can be used for segment boundary detection and to simulate the information processing of a listener. 🎶 🧠
📜 Link to the paper: arxiv.org/pdf/2501.07474
Model weights are soon to come! 🏋️
💫✨ #SonyCSLMusic 💫✨