Thinking on Shots: Consistent Multi-Shot Video Editing with Agentic Reasoning
https://arxiv.org/abs/2608.26809
Thinking on Shots: Consistent Multi-Shot Video Editing with Agentic Reasoning
https://arxiv.org/abs/2608.26809
Prefix Sliding for efficient test-time scaling
https://arxiv.org/abs/2608.26070
Prefix Sliding for efficient test-time scaling
https://arxiv.org/abs/2608.26070
“Multi-scale greenhouse gas emissions accounting in developing countries”
🔗 More info:
www.rug.nl/about-ug/lat...
#PhDDefense #ClimateAction #Sustainability
“Multi-scale greenhouse gas emissions accounting in developing countries”
🔗 More info:
www.rug.nl/about-ug/lat...
#PhDDefense #ClimateAction #Sustainability
Rethinking Position Embedding as a Context Controller for Multi-Reference and Multi-Shot Video Generation
https://arxiv.org/abs/2604.03738
Rethinking Position Embedding as a Context Controller for Multi-Reference and Multi-Shot Video Generation
https://arxiv.org/abs/2604.03738
https://arxiv.org/pdf/2603.11483
Yuping Tian, Chen-Hao Zhao, Chao-Bo Wang, Binyuan Zhang, Xiangru Kong, Wei-Jiang Gong.
https://arxiv.org/pdf/2603.11483
Yuping Tian, Chen-Hao Zhao, Chao-Bo Wang, Binyuan Zhang, Xiangru Kong, Wei-Jiang Gong.
Alex Volkov / @altryne:
Looks like @Ali_TongyiLab @Alibaba_Qwen is going through a change of leadership! Both Binyuan and Junyang who were very active on here, interacting with the community, releasing great models have departed [image]
Alex Volkov / @altryne:
Looks like @Ali_TongyiLab @Alibaba_Qwen is going through a change of leadership! Both Binyuan and Junyang who were very active on here, interacting with the community, releasing great models have departed [image]
SWE-RM: Execution-free Feedback For Software Engineering Agents
https://arxiv.org/abs/2512.21919
SWE-RM: Execution-free Feedback For Software Engineering Agents
https://arxiv.org/abs/2512.21919
Qwen3-VL Technical Report
https://arxiv.org/abs/2511.21631
Qwen3-VL Technical Report
https://arxiv.org/abs/2511.21631
PlotCraft: Pushing the Limits of LLMs for Complex and Interactive Data Visualization
https://arxiv.org/abs/2511.00010
PlotCraft: Pushing the Limits of LLMs for Complex and Interactive Data Visualization
https://arxiv.org/abs/2511.00010
VideoAgentTrek: Computer Use Pretraining from Unlabeled Videos
https://arxiv.org/abs/2510.19488
VideoAgentTrek: Computer Use Pretraining from Unlabeled Videos
https://arxiv.org/abs/2510.19488
MoGA: Mixture-of-Groups Attention for End-to-End Long Video Generation
https://arxiv.org/abs/2510.18692
MoGA: Mixture-of-Groups Attention for End-to-End Long Video Generation
https://arxiv.org/abs/2510.18692
Watch Where You Move: Region-aware Dynamic Aggregation and Excitation for Gait Recognition
https://arxiv.org/abs/2510.16541
Watch Where You Move: Region-aware Dynamic Aggregation and Excitation for Gait Recognition
https://arxiv.org/abs/2510.16541
Binyuan Hui / @huybery:
🚀 Qwen-Max has successfully scaled to 1T parameters, and we're still pushing further. Hopefully this giant will bring some surprises, see you next week!
Binyuan Hui / @huybery:
🚀 Qwen-Max has successfully scaled to 1T parameters, and we're still pushing further. Hopefully this giant will bring some surprises, see you next week!