English
Related papers

Related papers: COMPASS: A Formal Framework and Aggregate Dataset …

200 papers

Due to the noises in crowdsourced labels, label aggregation (LA) has emerged as a standard procedure to post-process crowdsourced labels. LA methods estimate true labels from crowdsourced labels by modeling worker qualities. Most existing…

Human-Computer Interaction · Computer Science 2022-12-02 Yi Yang , Zhong-Qiu Zhao , Quan Bai , Qing Liu , Weihua Li

Accurate medical image segmentation is essential for clinical diagnosis and treatment planning. While recent interactive foundation models (e.g., nnInteractive) enhance generalization through large-scale multimodal pretraining, they still…

Computer Vision and Pattern Recognition · Computer Science 2025-12-23 Ziyu Zhang , Yi Yu , Simeng Zhu , Ahmed Aly , Yunhe Gao , Ning Gu , Yuan Xue

One central goal of robotics is to enable robots to interact with the physical world. Traditional manipulation studies primarily focus on single robots and relatively small objects. However, factory and domestic environments often require…

Robotics · Computer Science 2026-05-26 Kun Song , Gaoming Chen , Shentao Ma , Ninglong Jin , Guangbao Zhao , Mingyu Ding , Zhenhua Xiong , Jia Pan

Medical imaging provides essential visual insights for diagnosis, and multimodal large language models (MLLMs) are increasingly utilized for its analysis due to their strong generalization capabilities; however, the underlying factors…

Computer Vision and Pattern Recognition · Computer Science 2025-06-03 Zhenyang Cai , Junying Chen , Rongsheng Wang , Weihong Wang , Yonglin Deng , Dingjie Song , Yize Chen , Zixu Zhang , Benyou Wang

Although the Segment Anything Model (SAM) is highly effective in natural image segmentation, it requires dependencies on prompts, which limits its applicability to medical imaging where manual prompts are often unavailable. Existing efforts…

Computer Vision and Pattern Recognition · Computer Science 2025-06-04 Mengmeng Zhang , Xingyuan Dai , Yicheng Sun , Jing Wang , Yueyang Yao , Xiaoyan Gong , Fuze Cong , Feiyue Wang , Yisheng Lv

Parallel cross-lingual summarization data is scarce, requiring models to better use the limited available cross-lingual resources. Existing methods to do so often adopt sequence-to-sequence networks with multi-task frameworks. Such…

Computation and Language · Computer Science 2021-06-15 Yu Bai , Yang Gao , Heyan Huang

Video-language foundation models have proven to be highly effective in zero-shot applications across a wide range of tasks. A particularly challenging area is the intraoperative surgical procedure domain, where labeled data is scarce, and…

Computer Vision and Pattern Recognition · Computer Science 2026-03-26 Florian Stilz , Vinkle Srivastav , Nassir Navab , Nicolas Padoy

Automated medical image segmentation has achieved remarkable progress with fully labeled data. However, site-specific clinical priorities and the high cost of manual annotation often yield scans with only a subset of organs labeled, leading…

Computer Vision and Pattern Recognition · Computer Science 2026-04-02 Qiaochu Zhao , Wei Wei , David Horowitz , Richard Bakst , Yading Yuan

Understanding surgical tasks represents an important challenge for autonomy in surgical robotic systems. To achieve this, we propose an online task segmentation framework that uses hierarchical transition state clustering to activate…

Robotics · Computer Science 2024-06-17 Yutaro Yamada , Jacinto Colan , Ana Davila , Yasuhisa Hasegawa

The limited availability of labeled data has driven advancements in semi-supervised learning for medical image segmentation. Modern large-scale models tailored for general segmentation, such as the Segment Anything Model (SAM), have…

Computer Vision and Pattern Recognition · Computer Science 2024-12-19 Kaiwen Huang , Tao Zhou , Huazhu Fu , Yizhe Zhang , Yi Zhou , Chen Gong , Dong Liang

Automated video-based assessment of surgical skills is a promising task in assisting young surgical trainees, especially in poor-resource areas. Existing works often resort to a CNN-LSTM joint framework that models long-term relationships…

Computer Vision and Pattern Recognition · Computer Science 2022-08-05 Zhenqiang Li , Lin Gu , Weimin Wang , Ryosuke Nakamura , Yoichi Sato

Collaborative robotic industrial cells are workspaces where robots collaborate with human operators. In this context, safety is paramount, and for that a complete perception of the space where the collaborative robot is inserted is…

Robotics · Computer Science 2022-10-20 Daniela Rato , Miguel Oliveira , Vítor Santos , Manuel Gomes , Angel Sappa

Segment Anything Model (SAM) is an advanced foundational model for image segmentation, which is gradually being applied to remote sensing images (RSIs). Due to the domain gap between RSIs and natural images, traditional methods typically…

Computer Vision and Pattern Recognition · Computer Science 2025-01-14 Nanqing Liu , Xun Xu , Yongyi Su , Haojie Zhang , Heng-Chao Li

The current landscape of scientific research is widely based on modeling and simulation, typically with complexity in the simulation's flow of execution and parameterization properties. Execution flows are not necessarily straightforward…

Distributed, Parallel, and Cluster Computing · Computer Science 2018-07-26 Eduardo Ponce , Brittany Stephenson , Suzanne Lenhart , Judy Day , Gregory D. Peterson

In computer vision, object detection is an important task that finds its application in many scenarios. However, obtaining extensive labels can be challenging, especially in crowded scenes. Recently, the Segment Anything Model (SAM) has…

Computer Vision and Pattern Recognition · Computer Science 2024-07-22 Zhi Cai , Yingjie Gao , Yaoyan Zheng , Nan Zhou , Di Huang

Spectral-based subspace clustering methods have proved successful in many challenging applications such as gene sequencing, image recognition, and motion segmentation. In this work, we first propose a novel spectral-based subspace…

Machine Learning · Statistics 2021-06-09 Hankui Peng , Nicos G. Pavlidis

Event-based cameras are bio-inspired sensors with pixels that independently and asynchronously respond to brightness changes at microsecond resolution, offering the potential to handle visual tasks in challenging scenarios. However, due to…

Computer Vision and Pattern Recognition · Computer Science 2026-02-25 Sheng Zhong , Zhongyang Ren , Xiya Zhu , Dehao Yuan , Cornelia Fermuller , Yi Zhou

Supervised image segmentation assigns image voxels to a set of labels, as defined by a specific labeling protocol. In this paper, we decompose segmentation into two steps. The first step is what we call "primitive segmentation", where…

Image and Video Processing · Electrical Eng. & Systems 2018-09-07 Sundaresh Ram , Mert R. Sabuncu

Foundation models have demonstrated remarkable success across diverse domains and tasks, primarily due to the thrive of large-scale, diverse, and high-quality datasets. However, in the field of medical imaging, the curation and assembling…

Computer Vision and Pattern Recognition · Computer Science 2026-03-31 Zhongying Deng , Cheng Tang , Ziyan Huang , Jiashi Lin , Ying Chen , Junzhi Ning , Chenglong Ma , Jiyao Liu , Wei Li , Yinghao Zhu , Shujian Gao , Yanyan Huang , Sibo Ju , Yanzhou Su , Pengcheng Chen , Wenhao Tang , Tianbin Li , Haoyu Wang , Yuanfeng Ji , Hui Sun , Shaobo Min , Liang Peng , Feilong Tang , Haochen Xue , Rulin Zhou , Chaoyang Zhang , Wenjie Li , Shaohao Rui , Weijie Ma , Xingyue Zhao , Yibin Wang , Kun Yuan , Zhaohui Lu , Shujun Wang , Jinjie Wei , Lihao Liu , Dingkang Yang , Lin Wang , Yulong Li , Haolin Yang , Yiqing Shen , Lequan Yu , Xiaowei Hu , Yun Gu , Yicheng Wu , Benyou Wang , Minghui Zhang , Angelica I. Aviles-Rivero , Qi Gao , Hongming Shan , Xiaoyu Ren , Fang Yan , Hongyu Zhou , Haodong Duan , Maosong Cao , Shanshan Wang , Bin Fu , Xiaomeng Li , Zhi Hou , Chunfeng Song , Lei Bai , Yuan Cheng , Yuandong Pu , Xiang Li , Wenhai Wang , Hao Chen , Jiaxin Zhuang , Songyang Zhang , Huiguang He , Mengzhang Li , Bohan Zhuang , Zhian Bai , Rongshan Yu , Liansheng Wang , Yukun Zhou , Xiaosong Wang , Xin Guo , Guanbin Li , Xiangru Lin , Dakai Jin , Mianxin Liu , Wenlong Zhang , Qi Qin , Conghui He , Yuqiang Li , Ye Luo , Nanqing Dong , Jie Xu , Wenqi Shao , Bo Zhang , Qiujuan Yan , Yihao Liu , Jun Ma , Zhi Lu , Yuewen Cao , Zongwei Zhou , Jianming Liang , Shixiang Tang , Qi Duan , Dongzhan Zhou , Chen Jiang , Yuyin Zhou , Yanwu Xu , Jiancheng Yang , Shaoting Zhang , Xiaohong Liu , Siqi Luo , Yi Xin , Chaoyu Liu , Haochen Wen , Xin Chen , Alejandro Lozano , Min Woo Sun , Yuhui Zhang , Yue Yao , Xiaoxiao Sun , Serena Yeung-Levy , Xia Li , Jing Ke , Chunhui Zhang , Zongyuan Ge , Ming Hu , Jin Ye , Zhifeng Li , Yirong Chen , Yu Qiao , Junjun He

Model merging has recently emerged as a lightweight alternative to ensembling, combining multiple fine-tuned models into a single set of parameters with no additional training overhead. Yet, existing merging methods fall short of matching…