中文
相关论文

相关论文: VISEM-Tracking, a human spermatozoa tracking datas…

200 篇论文

Automated surgical workflow analysis is crucial for education, research, and clinical decision-making, but the lack of annotated datasets hinders the development of accurate and comprehensive workflow analysis solutions. We introduce a…

计算机视觉与模式识别 · 计算机科学 2025-03-17 David Gastager , Ghazal Ghazaei , Constantin Patsch

Computer-assisted surgery research requires large, deeply annotated video datasets that capture clinical and technical variability. Existing cataract surgery resources lack the diversity and annotation depth required to train generalizable…

Global warming is predicted to profoundly impact ocean ecosystems. Fish behavior is an important indicator of changes in such marine environments. Thus, the automatic identification of key fish behavior in videos represents a much needed…

计算机视觉与模式识别 · 计算机科学 2020-12-01 Declan McIntosh , Tunai Porto Marques , Alexandra Branzan Albu , Rodney Rountree , Fabio De Leo

Quantifying behavior is crucial for many applications in neuroscience. Videography provides easy methods for the observation and recording of animal behavior in diverse settings, yet extracting particular aspects of a behavior for further…

计算机视觉与模式识别 · 计算机科学 2018-08-24 Alexander Mathis , Pranav Mamidanna , Taiga Abe , Kevin M. Cury , Venkatesh N. Murthy , Mackenzie W. Mathis , Matthias Bethge

Existing image/video datasets for cattle behavior recognition are mostly small, lack well-defined labels, or are collected in unrealistic controlled environments. This limits the utility of machine learning (ML) models learned from them.…

计算机视觉与模式识别 · 计算机科学 2023-07-04 Ali Zia , Renuka Sharma , Reza Arablouei , Greg Bishop-Hurley , Jody McNally , Neil Bagnall , Vivien Rolland , Brano Kusy , Lars Petersson , Aaron Ingham

Collecting high-quality data for training large-scale robotic models typically relies on real robot platforms, which is labor-intensive and costly, whether via teleoperation or scripted demonstrations. To scale data collection, many…

机器人学 · 计算机科学 2025-12-02 X. Hu , G. Ye

Robust behaviour recognition in real-world farm environments remains challenging due to several data-related limitations, including the scarcity of well-annotated livestock video datasets and the substantial domain gap between large-scale…

计算机视觉与模式识别 · 计算机科学 2026-04-08 Huimin Liu , Jing Gao , Daria Baran , AxelX Montout , Neill W Campbell , Andrew W Dowsey

Recognizing various surgical tools, actions and phases from surgery videos is an important problem in computer vision with exciting clinical applications. Existing deep-learning-based methods for this problem either process each surgical…

计算机视觉与模式识别 · 计算机科学 2024-04-09 Haifeng Wang , Hao Xu , Jun Wang , Jian Zhou , Ke Deng

Laparoscopic surgery is a complex surgical technique that requires extensive training. Recent advances in deep learning have shown promise in supporting this training by enabling automatic video-based assessment of surgical skills. However,…

The accurate tracking of live cells using video microscopy recordings remains a challenging task for popular state-of-the-art image processing based object tracking methods. In recent years, several existing and new applications have…

图像与视频处理 · 电气工程与系统科学 2025-02-03 Gergely Szabó , Paolo Bonaiuti , Andrea Ciliberto , András Horváth

Time-lapse is a technology used to record the development of embryos during in-vitro fertilization (IVF). Accurate classification of embryo early development stages can provide embryologists valuable information for assessing the embryo…

图像与视频处理 · 电气工程与系统科学 2019-08-27 Zihan Liu , Bo Huang , Yuqi Cui , Yifan Xu , Bo Zhang , Lixia Zhu , Yang Wang , Lei Jin , Dongrui Wu

Conventional sleep monitoring is time-consuming, expensive and uncomfortable, requiring a large number of contact sensors to be attached to the patient. Video data is commonly recorded as part of a sleep laboratory assessment. If accurate…

计算机视觉与模式识别 · 计算机科学 2023-06-07 Jonathan Carter , João Jorge , Bindia Venugopal , Oliver Gibson , Lionel Tarassenko

Large vision-language models have achieved remarkable capabilities by training on massive internet-scale data, yet a fundamental asymmetry persists: while LLMs can leverage self-supervised pretraining on abundant text and image data, the…

计算机视觉与模式识别 · 计算机科学 2026-04-08 Kidus Zewde , Yuchen Zhou , Dennis Ng , Neo Tiangratanakul , Tommy Duong , Ankit Raj , Yuxin Zhang , Xingyu Shen , Simiao Ren

Interpretation of giga-pixel whole-slide images (WSIs) is an important but difficult task for pathologists. Their diagnostic accuracy is estimated to average around 70%. Adding a second pathologist does not substantially improve decision…

计算机视觉与模式识别 · 计算机科学 2025-10-29 Veronica Thai , Rui Li , Meng Ling , Shuning Jiang , Jeremy Wolfe , Raghu Machiraju , Yan Hu , Zaibo Li , Anil Parwani , Jian Chen

We present the first deep learning model for the analysis of intracytoplasmic sperm injection (ICSI) procedures. Using a dataset of ICSI procedure videos, we train a deep neural network to segment key objects in the videos achieving a mean…

计算机视觉与模式识别 · 计算机科学 2023-03-03 Chloe He , Raksha Jain , Jérôme Chambost , Céline Jacques , Cristina Hickman

Tool tracking in surgical videos is essential for advancing computer-assisted interventions, such as skill assessment, safety zone estimation, and human-machine collaboration. However, the lack of context-rich datasets limits AI…

计算机视觉与模式识别 · 计算机科学 2025-03-25 Chinedu Innocent Nwoye , Kareem Elgohary , Anvita Srinivas , Fauzan Zaid , Joël L. Lavanchy , Nicolas Padoy

Deep learning methods have recently exhibited impressive performance in object detection. However, such methods needed much training data to achieve high recognition accuracy, which was time-consuming and required considerable manual work…

计算机视觉与模式识别 · 计算机科学 2023-01-05 Hao Chen , Weiwei Wan , Masaki Matsushita , Takeyuki Kotaka , Kensuke Harada

We introduce a new high resolution, high frame rate stereo video dataset, which we call SPIN, for tracking and action recognition in the game of ping pong. The corpus consists of ping pong play with three main annotation streams that can be…

计算机视觉与模式识别 · 计算机科学 2019-12-16 Steven Schwarcz , Peng Xu , David D'Ambrosio , Juhana Kangaspunta , Anelia Angelova , Huong Phan , Navdeep Jaitly

In recent years many different deep neural networks were developed, but due to a large number of layers in deep networks, their training requires a long time and a large number of datasets. Today is popular to use trained deep neural…

计算机视觉与模式识别 · 计算机科学 2021-06-15 R. Ildar

Technological advancements have spurred the usage of machine learning based applications in sports science. Physiotherapists, sports coaches and athletes actively look to incorporate the latest technologies in order to further improve…

计算机视觉与模式识别 · 计算机科学 2022-10-04 Ashish Singh , Antonio Bevilacqua , Thach Le Nguyen , Feiyan Hu , Kevin McGuinness , Martin OReilly , Darragh Whelan , Brian Caulfield , Georgiana Ifrim