Manyi Wang, Junjielong Xu, Pinjia He
#arXiv #cs.SE #cs.AI
Dual-Actor Fine-Tuning of VLA Models: A Talk-and-Tweak Human-in-the-Loop Approach
https://arxiv.org/abs/2509.13774
Dual-Actor Fine-Tuning of VLA Models: A Talk-and-Tweak Human-in-the-Loop Approach
https://arxiv.org/abs/2509.13774
Scalable Supervising Software Agents with Patch Reasoner
https://arxiv.org/abs/2510.22775
Scalable Supervising Software Agents with Patch Reasoner
https://arxiv.org/abs/2510.22775
Curing Miracle Steps in LLM Mathematical Reasoning with Rubric Rewards
https://arxiv.org/abs/2510.07774
Curing Miracle Steps in LLM Mathematical Reasoning with Rubric Rewards
https://arxiv.org/abs/2510.07774
Towards Evaluating Proactive Risk Awareness of Multimodal Language Models
https://arxiv.org/abs/2505.17455
Towards Evaluating Proactive Risk Awareness of Multimodal Language Models
https://arxiv.org/abs/2505.17455
Can't See the Forest for the Trees: Benchmarking Multimodal Safety Awareness for Multimodal LLMs
https://arxiv.org/abs/2502.11184
Can't See the Forest for the Trees: Benchmarking Multimodal Safety Awareness for Multimodal LLMs
https://arxiv.org/abs/2502.11184
Difficult Task Yes but Simple Task No: Unveiling the Laziness in Multimodal LLMs
https://arxiv.org/abs/2410.11437
Difficult Task Yes but Simple Task No: Unveiling the Laziness in Multimodal LLMs
https://arxiv.org/abs/2410.11437
Insight Over Sight? Exploring the Vision-Knowledge Conflicts in Multimodal LLMs
https://arxiv.org/abs/2410.08145
Insight Over Sight? Exploring the Vision-Knowledge Conflicts in Multimodal LLMs
https://arxiv.org/abs/2410.08145
Refuse Whenever You Feel Unsafe: Improving Safety in LLMs via Decoupled Refusal Training
https://arxiv.org/abs/2407.09121
Refuse Whenever You Feel Unsafe: Improving Safety in LLMs via Decoupled Refusal Training
https://arxiv.org/abs/2407.09121