English
Related papers

Related papers: Learning from Single Timestamps: Complexity Estima…

200 papers

Spatial transcriptomics (ST) technologies enable gene expression profiling with spatial resolution, offering unprecedented insights into tissue organization and disease heterogeneity. However, current analysis methods often struggle with…

Adapting large-scale image-text pre-training models, e.g., CLIP, to the video domain represents the current state-of-the-art for text-video retrieval. The primary approaches involve transferring text-video pairs to a common embedding space…

Computer Vision and Pattern Recognition · Computer Science 2024-07-17 Haonan Zhang , Pengpeng Zeng , Lianli Gao , Jingkuan Song , Yihang Duan , Xinyu Lyu , Hengtao Shen

Background: The critical view of safety (CVS) is poorly adopted in surgical practices although it is ubiquitously recommended to prevent major bile duct injuries during laparoscopic cholecystectomy (LC). This study aims to determine whether…

Sparse-View Computed Tomography (SVCT) offers low-dose and fast imaging but suffers from severe artifacts. Optimizing the sampling strategy is an essential approach to improving the imaging quality of SVCT. However, current methods…

Image and Video Processing · Electrical Eng. & Systems 2024-09-04 Liutao Yang , Jiahao Huang , Yingying Fang , Angelica I Aviles-Rivero , Carola-Bibiane Schonlieb , Daoqiang Zhang , Guang Yang

Colorectal diseases, including inflammatory conditions and neoplasms, require quick, accurate care to be effectively treated. Traditional diagnostic pipelines require extensive preparation and rely on separate, individual evaluations on…

Computer Vision and Pattern Recognition · Computer Science 2025-09-09 Krithik Ramesh , Ritvik Koneru

Small intestinal capsule endoscopy is the mainstream method for inspecting small intestinal lesions,but a single small intestinal capsule endoscopy will produce 60,000 - 120,000 images, the majority of which are similar and have no…

Image and Video Processing · Electrical Eng. & Systems 2020-04-07 Rui Nie , Huan Yang , Hejuan Peng , Wenbin Luo , Weiya Fan , Jie Zhang , Jing Liao , Fang Huang , Yufeng Xiao

Token pruning is essential for enhancing the computational efficiency of vision-language models (VLMs), particularly for video-based tasks where temporal redundancy is prevalent. Prior approaches typically prune tokens either (1) within the…

Computer Vision and Pattern Recognition · Computer Science 2026-03-19 Jianrui Zhang , Yue Yang , Rohun Tripathi , Winson Han , Ranjay Krishna , Christopher Clark , Yong Jae Lee , Sangho Lee

Self-supervised learning (SSL) has emerged as a powerful paradigm for Chest X-ray (CXR) analysis under limited annotations. Yet, existing SSL strategies remain suboptimal for medical imaging. Masked image modeling allocates substantial…

Computer Vision and Pattern Recognition · Computer Science 2026-03-25 Wangyu Feng , Shawn Young , Lijian Xu

Timeseries classification as stochastic (noise-like) or non-stochastic (structured), helps understand the underlying dynamics, in several domains. Here we propose a two-legged matrix decomposition-based algorithm utilizing two complementary…

Machine Learning · Computer Science 2023-07-18 Sai Pradeep Chakka , Sunil Kumar Vengalil , Neelam Sinha

Self-supervised learning (SSL) has emerged as a promising paradigm for addressing the annotation bottleneck in medical imaging by learning representations from unlabeled data. However, its effectiveness depends heavily on the design of the…

Computer Vision and Pattern Recognition · Computer Science 2026-05-26 Chathura Wimalasiri

The performance of Video Instance Segmentation (VIS) methods has improved significantly with the advent of transformer networks. However, these networks often face challenges in training due to the high annotation cost. To address this,…

Computer Vision and Pattern Recognition · Computer Science 2024-11-26 Farnoosh Arefi , Amir M. Mansourian , Shohreh Kasaei

Multi-class colorectal tissue classification is a challenging problem that is typically addressed in a setting, where it is assumed that ample amounts of training data is available. However, manual annotation of fine-grained colorectal…

Computer Vision and Pattern Recognition · Computer Science 2024-01-03 Dmitry Demidov , Roba Al Majzoub , Amandeep Kumar , Fahad Khan

Online local-life service platforms provide services like nearby daily essentials and food delivery for hundreds of millions of users. Different from other types of recommender systems, local-life service recommendation has the following…

Information Retrieval · Computer Science 2023-09-25 Huixuan Chi , Hao Xu , Mengya Liu , Yuanchen Bei , Sheng Zhou , Danyang Liu , Mengdi Zhang

Computer-aided diagnosis via deep learning relies on large-scale annotated data sets, which can be costly when involving expert knowledge. Semi-supervised learning (SSL) mitigates this challenge by leveraging unlabeled data. One effective…

Machine Learning · Computer Science 2020-05-25 Prashnna Kumar Gyawali , Sandesh Ghimire , Pradeep Bajracharya , Zhiyuan Li , Linwei Wang

We develop an algorithm which can learn from partially labeled and unsegmented sequential data. Most sequential loss functions, such as Connectionist Temporal Classification (CTC), break down when many labels are missing. We address this…

Machine Learning · Computer Science 2022-03-07 Vineel Pratap , Awni Hannun , Gabriel Synnaeve , Ronan Collobert

Prostate cancer is one of the main diseases affecting men worldwide. The gold standard for diagnosis and prognosis is the Gleason grading system. In this process, pathologists manually analyze prostate histology slides under microscope, in…

Image and Video Processing · Electrical Eng. & Systems 2021-05-24 Julio Silva-Rodríguez , Adrián Colomer , Jose Dolz , Valery Naranjo

Text-guided medical segmentation enhances segmentation accuracy by utilizing clinical reports as auxiliary information. However, existing methods typically rely on unaligned image and text encoders, which necessitate complex interaction…

Computer Vision and Pattern Recognition · Computer Science 2025-12-25 Gaoren Lin , Huangxuan Zhao , Yuan Xiong , Lefei Zhang , Bo Du , Wentao Zhu

Video large language models (Video-LLMs) face high computational costs due to large volumes of visual tokens. Existing token compression methods typically adopt a two-stage spatiotemporal compression strategy, relying on stage-specific…

Computer Vision and Pattern Recognition · Computer Science 2026-03-24 Junhao Du , Jialong Xue , Anqi Li , Jincheng Dai , Guo Lu

Computer-assisted surgery research requires large, deeply annotated video datasets that capture clinical and technical variability. Existing cataract surgery resources lack the diversity and annotation depth required to train generalizable…

Early-stage scoliosis is often difficult to detect, particularly in adolescents, where delayed diagnosis can lead to serious health issues. Traditional X-ray-based methods carry radiation risks and rely heavily on clinical expertise,…

Computer Vision and Pattern Recognition · Computer Science 2025-10-07 Haiqing Li , Yuzhi Guo , Feng Jiang , Thao M. Dang , Hehuan Ma , Qifeng Zhou , Jean Gao , Junzhou Huang