English
Related papers

Related papers: SaFeR-ToolKit: Structured Reasoning via Virtual To…

200 papers

As Multimodal Large Language Models (MLLMs) acquire stronger reasoning capabilities to handle complex, multi-image instructions, this advancement may pose new safety risks. We study this problem by introducing MIR-SafetyBench, the first…

Computer Vision and Pattern Recognition · Computer Science 2026-01-21 Renmiao Chen , Yida Lu , Shiyao Cui , Xuan Ouyang , Victor Shea-Jay Huang , Shumin Zhang , Chengwei Pan , Han Qiu , Minlie Huang

Multimodal LLMs are turning their focus to video benchmarks, however most video benchmarks only provide outcome supervision, with no intermediate or interpretable reasoning steps. This makes it challenging to assess if models are truly able…

Multi-modal Large Language Models (MLLMs) are increasingly deployed in interactive applications. However, their safety vulnerabilities become pronounced in multi-turn multi-modal scenarios, where harmful intent can be gradually…

Computation and Language · Computer Science 2026-01-09 Han Zhu , Jiale Chen , Chengkun Cai , Shengjie Sun , Haoran Li , Yujin Zhou , Chi-Min Chan , Pengcheng Wen , Lei Li , Sirui Han , Yike Guo

Effectiveness and interpretability are two essential properties for trustworthy AI systems. Most recent studies in visual reasoning are dedicated to improving the accuracy of predicted answers, and less attention is paid to explaining the…

Computer Vision and Pattern Recognition · Computer Science 2022-03-14 Shi Chen , Qi Zhao

Large Vision-Language Models (VLMs) have achieved remarkable performance across a wide range of tasks. However, their deployment in safety-critical domains poses significant challenges. Existing safety fine-tuning methods, which focus on…

Computer Vision and Pattern Recognition · Computer Science 2026-02-04 Yi Ding , Lijun Li , Bing Cao , Jing Shao

Traditional workflow-based agents exhibit limited intelligence when addressing real-world problems requiring tool invocation. Tool-integrated reasoning (TIR) agents capable of autonomous reasoning and tool invocation are rapidly emerging as…

While the safety risks of image-based large language models (Image LLMs) have been extensively studied, their video-based counterparts (Video LLMs) remain critically under-examined. To systematically study this problem, we introduce…

Computer Vision and Pattern Recognition · Computer Science 2026-03-17 Yiwei Sun , Peiqi Jiang , Chuanbin Liu , Luohao Lin , Zhiying Lu , Hongtao Xie

Recent advances in large language models have significantly improved textual reasoning through the effective use of Chain-of-Thought (CoT) and reinforcement learning. However, extending these successes to vision-language tasks remains…

Computer Vision and Pattern Recognition · Computer Science 2025-05-27 Minheng Ni , Zhengyuan Yang , Linjie Li , Chung-Ching Lin , Kevin Lin , Wangmeng Zuo , Lijuan Wang

We introduce V-tableR1, a process-supervised reinforcement learning framework that elicits rigorous, verifiable reasoning from multimodal large language models (MLLMs). Current MLLMs trained solely on final outcomes often treat visual…

Artificial Intelligence · Computer Science 2026-04-23 Yubo Jiang , Yitong An , Xin Yang , Abudukelimu Wuerkaixi , Xuxin Cheng , Fengying Xie , Zhiguo Jiang , Cao Liu , Ke Zeng , Haopeng Zhang

The growing complexity of factual claims in real-world scenarios presents significant challenges for automated fact verification systems, particularly in accurately aggregating and reasoning over multi-hop evidence. Existing approaches…

Artificial Intelligence · Computer Science 2025-06-10 Liwen Zheng , Chaozhuo Li , Haoran Jia , Xi Zhang

Recent advancements in large language models (LLMs) have accelerated progress toward artificial general intelligence, yet their potential to generate harmful content poses critical safety challenges. Existing alignment methods often…

Computation and Language · Computer Science 2025-10-08 Kehua Feng , Keyan Ding , Yuhao Wang , Menghan Li , Fanjunduo Wei , Xinda Wang , Qiang Zhang , Huajun Chen

MLLMs are increasingly deployed in multi-turn settings, where attackers can escalate unsafe intent through the evolving visual-text history and exploit long-context safety decay. Yet safety alignment is still dominated by single-turn data…

Machine Learning · Computer Science 2026-05-28 Haolong Hu , Hanyu Li , Tiancheng He , Huahui Yi , An Zhang , Qiankun Li , Kun Wang , Yang Liu , Zhigang Zeng

Large language models exhibit safety degradation in non-English languages. Standard evaluation relies on Jailbreak Success Rate (JSR), which confounds several safety-driving factors into one, obscuring the specific cause(s) of safety…

Computation and Language · Computer Science 2026-05-19 Max Zhang , Ameen Patel , Sang T. Truong , Sanmi Koyejo

Safe and feasible trajectory planning is critical for real-world autonomous driving systems. However, existing learning-based planners rely heavily on expert demonstrations, which not only lack explicit safety awareness but also risk…

Robotics · Computer Science 2025-09-29 Xiaolong Tang , Meina Kan , Shiguang Shan , Xilin Chen

Large reasoning models (LRMs) increasingly expose chain-of-thought-like reasoning for transparency, verification, and deliberate problem solving. This creates a safety blind spot: harmful or policy-violating content may appear in reasoning…

Artificial Intelligence · Computer Science 2026-05-08 Xiaomin Li , Jianheng Hou , Zheyuan Deng , Zhiwei Zhang , Taoran Li , Binghang Lu , Bing Hu , Yunhan Zhao , Yuexing Hao

Safety evaluation of multimodal foundation models often treats vision and language inputs separately, missing risks from joint interpretation where benign content becomes harmful in combination. Existing approaches also fail to distinguish…

Computer Vision and Pattern Recognition · Computer Science 2025-12-04 Shruti Palaskar , Leon Gatys , Mona Abdelrahman , Mar Jacobo , Larry Lindsey , Rutika Moharir , Gunnar Lund , Yang Xu , Navid Shiee , Jeffrey Bigham , Charles Maalouf , Joseph Yitan Cheng

Visual reasoning models (VRMs) have recently shown strong cross-modal reasoning capabilities by integrating visual perception with language reasoning. However, they often suffer from overthinking, producing unnecessarily long reasoning…

Computer Vision and Pattern Recognition · Computer Science 2026-04-17 Yixu Huang , Tinghui Zhu , Muhao Chen

We introduce SafeWork-R1, a cutting-edge multimodal reasoning model that demonstrates the coevolution of capabilities and safety. It is developed by our proposed SafeLadder framework, which incorporates large-scale, progressive,…

Artificial Intelligence · Computer Science 2025-08-08 Shanghai AI Lab , : , Yicheng Bao , Guanxu Chen , Mingkang Chen , Yunhao Chen , Chiyu Chen , Lingjie Chen , Sirui Chen , Xinquan Chen , Jie Cheng , Yu Cheng , Dengke Deng , Yizhuo Ding , Dan Ding , Xiaoshan Ding , Yi Ding , Zhichen Dong , Lingxiao Du , Yuyu Fan , Xinshun Feng , Yanwei Fu , Yuxuan Gao , Ruijun Ge , Tianle Gu , Lujun Gui , Jiaxuan Guo , Qianxi He , Yuenan Hou , Xuhao Hu , Hong Huang , Kaichen Huang , Shiyang Huang , Yuxian Jiang , Shanzhe Lei , Jie Li , Lijun Li , Hao Li , Juncheng Li , Xiangtian Li , Yafu Li , Lingyu Li , Xueyan Li , Haotian Liang , Dongrui Liu , Qihua Liu , Zhixuan Liu , Bangwei Liu , Huacan Liu , Yuexiao Liu , Zongkai Liu , Chaochao Lu , Yudong Lu , Xiaoya Lu , Zhenghao Lu , Qitan Lv , Caoyuan Ma , Jiachen Ma , Xiaoya Ma , Zhongtian Ma , Lingyu Meng , Ziqi Miao , Yazhe Niu , Yuezhang Peng , Yuan Pu , Han Qi , Chen Qian , Xingge Qiao , Jingjing Qu , Jiashu Qu , Wanying Qu , Wenwen Qu , Xiaoye Qu , Qihan Ren , Qingnan Ren , Qingyu Ren , Jing Shao , Wenqi Shao , Shuai Shao , Dongxing Shi , Xin Song , Xinhao Song , Yan Teng , Xuan Tong , Yingchun Wang , Xuhong Wang , Shujie Wang , Xin Wang , Yige Wang , Yixu Wang , Yuanfu Wang , Futing Wang , Ruofan Wang , Wenjie Wang , Yajie Wang , Muhao Wei , Xiaoyu Wen , Fenghua Weng , Yuqi Wu , Yingtong Xiong , Xingcheng Xu , Chao Yang , Yue Yang , Yang Yao , Yulei Ye , Zhenyun Yin , Yi Yu , Bo Zhang , Qiaosheng Zhang , Jinxuan Zhang , Yexin Zhang , Yinqiang Zheng , Hefeng Zhou , Zhanhui Zhou , Pengyu Zhu , Qingzi Zhu , Yubo Zhu , Bowen Zhou

Visual Chain-of-Thought (VCoT) has emerged as a promising paradigm for enhancing multimodal reasoning by integrating visual perception into intermediate reasoning steps. However, existing VCoT approaches are largely confined to static…

Computer Vision and Pattern Recognition · Computer Science 2026-02-12 Junhua Liu , Zhangcheng Wang , Zhike Han , Ningli Wang , Guotao Liang , Kun Kuang

Vision-language models (VLMs) are increasingly deployed in real-world and embodied settings where safety decisions depend on visual context. However, it remains unclear which visual evidence drives these judgments. We study whether…

Computer Vision and Pattern Recognition · Computer Science 2026-03-20 Carlos Hinojosa , Clemens Grange , Bernard Ghanem