English
Related papers

Related papers: ProstaTD: Bridging Surgical Triplet from Classific…

200 papers

Reliable 4D object detection, which refers to 3D object detection in streaming video, is crucial for perceiving and understanding the real world. Existing open-set 4D object detection methods typically make predictions on a frame-by-frame…

Computer Vision and Pattern Recognition · Computer Science 2025-11-25 Jiawei Hou , Shenghao Zhang , Can Wang , Zheng Gu , Yonggen Ling , Taiping Zeng , Xiangyang Xue , Jingbo Zhang

Accurate real-time object detection is vital across numerous industrial applications, from safety monitoring to quality control. Traditional approaches, however, are hindered by arduous manual annotation and data collection, struggling to…

Computer Vision and Pattern Recognition · Computer Science 2025-06-24 Chen Xin , Andreas Hartel , Enkelejda Kasneci

We propose a new method that employs transfer learning techniques to effectively correct sampling selection errors introduced by sparse annotations during supervised learning for automated tumor segmentation. The practicality of current…

360 video captures the complete surrounding scenes with the ultra-large field of view of 360X180. This makes 360 scene understanding tasks, eg, segmentation and tracking, crucial for appications, such as autonomous driving, robotics. With…

Computer Vision and Pattern Recognition · Computer Science 2025-06-18 Weiming Zhang , Dingwen Xiao , Aobotao Dai , Yexin Liu , Tianbo Pan , Shiqi Wen , Lei Chen , Lin Wang

Polyp segmentation is a crucial step towards computer-aided diagnosis of colorectal cancer. However, most of the polyp segmentation methods require pixel-wise annotated datasets. Annotated datasets are tedious and time-consuming to produce,…

Computer Vision and Pattern Recognition · Computer Science 2022-11-23 Guangyu Ren , Michalis Lazarou , Jing Yuan , Tania Stathaki

Accurate segmentation for medical images is important for clinical diagnosis. Existing automatic segmentation methods are mainly based on fully supervised learning and have an extremely high demand for precise annotations, which are very…

Computer Vision and Pattern Recognition · Computer Science 2021-06-10 Yuanpeng Liu , Qinglei Hui , Zhiyi Peng , Shaolin Gong , Dexing Kong

Deep learning-based diagnostic performance increases with more annotated data, but large-scale manual annotations are expensive and labour-intensive. Experts evaluate diagnostic images during clinical routine, and write their findings in…

Image and Video Processing · Electrical Eng. & Systems 2024-07-01 Joeran S. Bosma , Anindo Saha , Matin Hosseinzadeh , Ilse Slootweg , Maarten de Rooij , Henkjan Huisman

Surgical scene understanding is a cornerstone of computer-assisted intervention. While recent advances, particularly in surgical image segmentation, have driven progress, real-world clinical applications require a more holistic…

Computer Vision and Pattern Recognition · Computer Science 2026-05-14 Jincai Huang , Shihao Zou , Yuchen Guo , Jingjing Li , Wei Ji , Kai Wang , Shanshan Wang , Weixin Si

Obtaining large-scale medical data, annotated or unannotated, is challenging due to stringent privacy regulations and data protection policies. In addition, annotating medical images requires that domain experts manually delineate…

Computer Vision and Pattern Recognition · Computer Science 2025-05-16 Tushar Kataria , Shireen Y. Elhabian

Breast lesion segmentation in ultrasound (US) videos is essential for diagnosing and treating axillary lymph node metastasis. However, the lack of a well-established and large-scale ultrasound video dataset with high-quality annotations has…

Image and Video Processing · Electrical Eng. & Systems 2023-10-04 Junhao Lin , Qian Dai , Lei Zhu , Huazhu Fu , Qiong Wang , Weibin Li , Wenhao Rao , Xiaoyang Huang , Liansheng Wang

Stuttering is a complex disorder that requires specialized expertise for effective assessment and treatment. This paper presents an effort to enhance the FluencyBank dataset with a new stuttering annotation scheme based on established…

Recent advances of 3D acquisition devices have enabled large-scale acquisition of 3D scene data. Such data, if completely and well annotated, can serve as useful ingredients for a wide spectrum of computer vision and graphics works such as…

Computer Vision and Pattern Recognition · Computer Science 2016-10-20 Duc Thanh Nguyen , Binh-Son Hua , Lap-Fai Yu , Sai-Kit Yeung

The application of deep learning to nursing procedure activity understanding has the potential to greatly enhance the quality and safety of nurse-patient interactions. By utilizing the technique, we can facilitate training and education,…

Computer Vision and Pattern Recognition · Computer Science 2023-10-23 Ming Hu , Lin Wang , Siyuan Yan , Don Ma , Qingli Ren , Peng Xia , Wei Feng , Peibo Duan , Lie Ju , Zongyuan Ge

Computer-assisted surgery has been developed to enhance surgery correctness and safety. However, researchers and engineers suffer from limited annotated data to develop and train better algorithms. Consequently, the development of…

Computer Vision and Pattern Recognition · Computer Science 2020-12-24 W. -Y. Hong , C. -L. Kao , Y. -H. Kuo , J. -R. Wang , W. -L. Chang , C. -S. Shih

Volumetric magnetic resonance (MR) image segmentation plays an important role in many clinical applications. Deep learning (DL) has recently achieved state-of-the-art or even human-level performance on various image segmentation tasks.…

Computer Vision and Pattern Recognition · Computer Science 2022-11-17 Yousuf Babiker M. Osman , Cheng Li , Weijian Huang , Nazik Elsayed , Zhenzhen Xue , Hairong Zheng , Shanshan Wang

Surgical scene perception via videos is critical for advancing robotic surgery, telesurgery, and AI-assisted surgery, particularly in ophthalmology. However, the scarcity of diverse and richly annotated video datasets has hindered the…

Computer Vision and Pattern Recognition · Computer Science 2024-07-22 Ming Hu , Peng Xia , Lin Wang , Siyuan Yan , Feilong Tang , Zhongxing Xu , Yimin Luo , Kaimin Song , Jurgen Leitner , Xuelian Cheng , Jun Cheng , Chi Liu , Kaijing Zhou , Zongyuan Ge

Fine-grained spatiotemporal reasoning on surgical videos is critical, yet the capabilities of Multi-modal Large Language Models (MLLMs) in this domain remain largely unexplored. To bridge this gap, we introduce SurgCoT, a unified benchmark…

Computer Vision and Pattern Recognition · Computer Science 2026-04-23 Gui Wang , YongSong Zhou , Kaijun Deng , Wooi Ping Cheah , Rong Qu , Jianfeng Ren , Linlin Shen

Due to the intensive cost of labor and expertise in annotating 3D medical images at a voxel level, most benchmark datasets are equipped with the annotations of only one type of organs and/or tumors, resulting in the so-called partially…

Computer Vision and Pattern Recognition · Computer Science 2020-11-23 Jianpeng Zhang , Yutong Xie , Yong Xia , Chunhua Shen

Understanding objects at the level of their constituent parts is fundamental to advancing computer vision, graphics, and robotics. While datasets like PartNet have driven progress in 3D part understanding, their reliance on untextured…

Computer Vision and Pattern Recognition · Computer Science 2026-04-01 Penghao Wang , Yiyang He , Xin Lv , Yukai Zhou , Lan Xu , Jingyi Yu , Jiayuan Gu

In recent decades, the vision community has witnessed remarkable progress in visual recognition, partially owing to advancements in dataset benchmarks. Notably, the established COCO benchmark has propelled the development of modern…

Computer Vision and Pattern Recognition · Computer Science 2024-04-15 Xueqing Deng , Qihang Yu , Peng Wang , Xiaohui Shen , Liang-Chieh Chen