#VideoUnderstanding
The University of Technology Nuremberg's FunAI Lab is offering a PhD position in Vision-Language-Action models, focusing on video...

mlscientist.com/phd-vision-l...

#PhDposition #VisionLanguageActionmodels #videounderstanding
September 22, 2026 at 10:46 PM
Elections and Media in the Modern Era - Part 1 of 3: Elections – Behind the Ballot Box - VideoUnderstanding why every vote matters is only the first part of the story.
<a href="https://elephantsinkroom.com/elections-and-media-in-the-modern-era-part-1-of-3-elections-behind-the-ballot-box-video/" class="hover:underline text-blue-600 dark:text-sky-400 no-card-link" target="_blank" rel="noopener" data-link="bsky">https://elephantsinkroom.com/elections-and-media-in-the-modern-era-part-1-of-3-elections-behind-the-ballot-box-video/
September 23, 2026 at 2:36 PM
Really interesting workshop by my colleagues at Surrey, don't miss it if you're at #BMVC in Glasgow #BMVC2024 #videounderstanding #computervision
If you're attending BMVC this year and are passionate about Video Understanding, make sure to join our workshop, VUA, on November 28th.
We have an incredible lineup of speakers: Prof. Cees Snoek, Dr @antoninofurnari.it , and Dr Laura Sevilla.
November 27, 2024 at 5:43 PM
Check out our pioneering paper on Video and Audiovisual Understanding with LLMs! Dive into the future of AI with us: #VideoUnderstanding #LargeLanguageModels #AIResearch

🕸️ github.com/yunlong10/Aw...
GitHub - yunlong10/Awesome-LLMs-for-Video-Understanding: 🔥🔥🔥Latest Papers, Codes and Datasets on Vid-LLMs.
🔥🔥🔥Latest Papers, Codes and Datasets on Vid-LLMs. Contribute to yunlong10/Awesome-LLMs-for-Video-Understanding development by creating an account on GitHub.
github.com
January 13, 2025 at 9:23 PM
Next, "Re-thinking Temporal Search for Long-Form Video Understanding" #CVPR2025

🗓️ Fri Jun 13, 4PM-6PM
📍 ExHall D Poster #306
🔗 Paper: arxiv.org/abs/2504.02259
🌐 Website: longvideohaystack.github.io
💻 Code: github.com/LongVideoHay...
📊 Data: huggingface.co/datasets/LVH...

#VideoUnderstanding
Re-thinking Temporal Search for Long-Form Video Understanding
Efficiently understanding long-form videos remains a significant challenge in computer vision. In this work, we revisit temporal search paradigms for long-form video understanding and address a fundam...
arxiv.org
June 10, 2025 at 5:37 AM
Efficiency meets transparency. 🎥 VideoChat3, a 4B-parameter model, is beating GPT-5 and Gemini 2.5 Flash in video grounding tasks. By releasing the full training stack—code, weights, and data—it's setting a new standard for open-source AI. #AI #OpenSource #VideoUnderstanding #MachineLearning #Com...
VideoChat3 outperforms GPT-5 in video grounding with fully open stack
Efficiency meets transparency. 🎥 VideoChat3, a 4B-parameter model, is beating GPT-5 and Gemini 2.5 Flash in video grounding tasks. By releasing the full training stack—code, weights, and data—it's set
www.alextech.ai
July 20, 2026 at 6:34 AM
𝗠𝗖𝗠𝗟 𝗕𝗹𝗼𝗴: Finding moments in long videos? Easy for humans, tough for AI.

ReVisionLLM — by MCML Members Tanveer Hannan, Thomas Seidl & team — learns to scan like we do: look wide, zoom in and spot interesting segments.

🔗 mcml.ai/news/2025-06...

#AI #LLM #VideoUnderstanding #MCML
June 20, 2025 at 7:32 AM
Marlin-2B: análisis de video con timestamps exactos

Marlin-2B análisis de videos: modelo VLM abierto que crea subtítulos timestampeados. Explorá la revolución de Hugging Face en visión por computadora.

#marlin2b #videounderstanding #vlm #huggingface #modelosopensource
Marlin-2B análisis de videos con timestamps
Marlin-2B análisis de videos: modelo VLM abierto que crea subtítulos timestampeados. Explorá la revolución de Hugging Face en visión por computadora.
blog.donweb.com
May 27, 2026 at 10:46 AM
📰🚨TwelveLabs video understanding models are now available in Amazon Bedrock by Channy Yun (윤석찬)

#TwelveLabs #AmazonBedrock #VideoUnderstanding #AIModels #VideoSearch
TwelveLabs video understanding models are now available in Amazon Bedrock | Amazon Web Services
TwelveLabs video understanding models are now available on Amazon Bedrock and enable customers to search through videos, classify scenes, summarize content, and extract insights with precision and reliability.
ift.tt
July 16, 2025 at 6:33 AM
EgoExoBench: A Benchmark for First- and Third-person View Video
Understanding in MLLMs
Baoqi Pei, Guo Chen et al.
Paper
Details
#EgoExoBench #MLLMs #VideoUnderstanding
July 27, 2025 at 4:06 PM