mlscientist.com/phd-vision-l...
#PhDposition #VisionLanguageActionmodels #videounderstanding
mlscientist.com/phd-vision-l...
#PhDposition #VisionLanguageActionmodels #videounderstanding
<a href="https://elephantsinkroom.com/elections-and-media-in-the-modern-era-part-1-of-3-elections-behind-the-ballot-box-video/" class="hover:underline text-blue-600 dark:text-sky-400 no-card-link" target="_blank" rel="noopener" data-link="bsky">https://elephantsinkroom.com/elections-and-media-in-the-modern-era-part-1-of-3-elections-behind-the-ballot-box-video/
<a href="https://elephantsinkroom.com/elections-and-media-in-the-modern-era-part-1-of-3-elections-behind-the-ballot-box-video/" class="hover:underline text-blue-600 dark:text-sky-400 no-card-link" target="_blank" rel="noopener" data-link="bsky">https://elephantsinkroom.com/elections-and-media-in-the-modern-era-part-1-of-3-elections-behind-the-ballot-box-video/
We have an incredible lineup of speakers: Prof. Cees Snoek, Dr @antoninofurnari.it , and Dr Laura Sevilla.
🕸️ github.com/yunlong10/Aw...
🕸️ github.com/yunlong10/Aw...
🗓️ Fri Jun 13, 4PM-6PM
📍 ExHall D Poster #306
🔗 Paper: arxiv.org/abs/2504.02259
🌐 Website: longvideohaystack.github.io
💻 Code: github.com/LongVideoHay...
📊 Data: huggingface.co/datasets/LVH...
#VideoUnderstanding
🗓️ Fri Jun 13, 4PM-6PM
📍 ExHall D Poster #306
🔗 Paper: arxiv.org/abs/2504.02259
🌐 Website: longvideohaystack.github.io
💻 Code: github.com/LongVideoHay...
📊 Data: huggingface.co/datasets/LVH...
#VideoUnderstanding
ReVisionLLM — by MCML Members Tanveer Hannan, Thomas Seidl & team — learns to scan like we do: look wide, zoom in and spot interesting segments.
🔗 mcml.ai/news/2025-06...
#AI #LLM #VideoUnderstanding #MCML
ReVisionLLM — by MCML Members Tanveer Hannan, Thomas Seidl & team — learns to scan like we do: look wide, zoom in and spot interesting segments.
🔗 mcml.ai/news/2025-06...
#AI #LLM #VideoUnderstanding #MCML
Marlin-2B análisis de videos: modelo VLM abierto que crea subtítulos timestampeados. Explorá la revolución de Hugging Face en visión por computadora.
#marlin2b #videounderstanding #vlm #huggingface #modelosopensource
Marlin-2B análisis de videos: modelo VLM abierto que crea subtítulos timestampeados. Explorá la revolución de Hugging Face en visión por computadora.
#marlin2b #videounderstanding #vlm #huggingface #modelosopensource
#TwelveLabs #AmazonBedrock #VideoUnderstanding #AIModels #VideoSearch
#TwelveLabs #AmazonBedrock #VideoUnderstanding #AIModels #VideoSearch
Understanding in MLLMs
Baoqi Pei, Guo Chen et al.
Paper
Details
#EgoExoBench #MLLMs #VideoUnderstanding
Understanding in MLLMs
Baoqi Pei, Guo Chen et al.
Paper
Details
#EgoExoBench #MLLMs #VideoUnderstanding