English
Related papers

Related papers: Adaptive Texture-aware Masking for Self-Supervised…

200 papers

The human voice is a promising non-invasive digital biomarker, yet deep learning for voice-based health analysis is hindered by data scarcity and domain mismatch, where models pre-trained on general audio fail to capture the subtle…

Audio and Speech Processing · Electrical Eng. & Systems 2026-02-02 Weixin Liu , Bowen Qu , Matthew Pontell , Maria Powell , Bradley Malin , Zhijun Yin

Accurate and automatic segmentation of three-dimensional (3D) individual teeth from cone-beam computerized tomography (CBCT) images is a challenging problem because of the difficulty in separating an individual tooth from adjacent teeth and…

Computer Vision and Pattern Recognition · Computer Science 2021-12-06 Tae Jun Jang , Kang Cheol Kim , Hyun Cheol Cho , Jin Keun Seo

3D scene reconstruction from 2D images has been a long-standing task. Instead of estimating per-frame depth maps and fusing them in 3D, recent research leverages the neural implicit surface as a unified representation for 3D reconstruction.…

Computer Vision and Pattern Recognition · Computer Science 2023-09-19 Xinyi Yu , Liqin Lu , Jintao Rong , Guangkai Xu , Linlin Ou

Spinal surgery planning necessitates automatic segmentation of vertebrae in cone-beam computed tomography (CBCT), an intraoperative imaging modality that is widely used in intervention. However, CBCT images are of low-quality and…

Image and Video Processing · Electrical Eng. & Systems 2021-03-10 Yuanyuan Lyu , Haofu Liao , Heqin Zhu , S. Kevin Zhou

Accurate semantic segmentation of 3D dental models is essential for digital dentistry applications such as orthodontics and dental implants. However, due to complex tooth arrangements and similarities in shape among adjacent teeth, existing…

Computer Vision and Pattern Recognition · Computer Science 2026-03-18 Qiang He , Wentian Qu , Jiajia Dai , Changsong Lei , Shaofeng Wang , Feifei Zuo , Yajie Wang , Yaqian Liang , Xiaoming Deng , Cuixia Ma , Yong-Jin Liu , Hongan Wang

Self-supervised learning (SSL) has emerged as a powerful technique for improving the efficiency and effectiveness of deep learning models. Contrastive methods are a prominent family of SSL that extract similar representations of two…

Compressive sensing (CS) reconstructs images from sub-Nyquist measurements by solving a sparsity-regularized inverse problem. Traditional CS solvers use iterative optimizers with hand crafted sparsifiers, while early data-driven methods…

Computer Vision and Pattern Recognition · Computer Science 2023-06-02 Pamuditha Somarathne , Tharindu Wickremasinghe , Amashi Niwarthana , A. Thieshanthan , Chamira U. S. Edussooriya , Dushan N. Wadduwage

Concept Bottleneck Models (CBMs) and other concept-based interpretable models show great promise for making AI applications more transparent, which is essential in fields like medicine. Despite their success, we demonstrate that CBMs…

Computer Vision and Pattern Recognition · Computer Science 2025-08-01 Jessica Bader , Leander Girrbach , Stephan Alaniz , Zeynep Akata

Context-aware methods have achieved remarkable advancements in supervised scene text recognition by leveraging semantic priors from words. Considering the heterogeneity of text and background in STR, we propose that such contextual priors…

Computer Vision and Pattern Recognition · Computer Science 2024-11-20 Tiancheng Lin , Jinglei Zhang , Yi Xu , Kai Chen , Rui Zhang , Chang-Wen Chen

Cone-beam computed tomography (CBCT) is widely used in interventional surgeries and radiation oncology. Due to the limited size of flat-panel detectors, anatomical structures might be missing outside the limited field-of-view (FOV), which…

Computer Vision and Pattern Recognition · Computer Science 2024-09-16 Yixing Huang , Fuxin Fan , Ahmed Gomaa , Andreas Maier , Rainer Fietkau , Christoph Bert , Florian Putz

We present MInDI-3D (Medical Inversion by Direct Iteration in 3D), the first 3D conditional diffusion-based model for real-world sparse-view Cone Beam Computed Tomography (CBCT) artefact removal, aiming to reduce imaging radiation exposure.…

Computer Vision and Pattern Recognition · Computer Science 2025-10-10 Daniel Barco , Marc Stadelmann , Martin Oswald , Ivo Herzig , Lukas Lichtensteiger , Pascal Paysan , Igor Peterlik , Michal Walczak , Bjoern Menze , Frank-Peter Schilling

We propose a novel continual self-supervised learning (CSSL) framework for simultaneously learning diverse features from multi-window-obtained chest computed tomography (CT) images and ensuring data privacy. Achieving a robust and highly…

Computer Vision and Pattern Recognition · Computer Science 2025-11-03 Ren Tasai , Guang Li , Ren Togo , Takahiro Ogawa , Kenji Hirata , Minghui Tang , Takaaki Yoshimura , Hiroyuki Sugimori , Noriko Nishioka , Yukie Shimizu , Kohsuke Kudo , Miki Haseyama

Since the development of self-supervised visual representation learning from contrastive learning to masked image modeling (MIM), there is no significant difference in essence, that is, how to design proper pretext tasks for vision…

Computer Vision and Pattern Recognition · Computer Science 2023-01-31 Kun Yi , Yixiao Ge , Xiaotong Li , Shusheng Yang , Dian Li , Jianping Wu , Ying Shan , Xiaohu Qie

Purpose: Paranasal anomalies, frequently identified in routine radiological screenings, exhibit diverse morphological characteristics. Due to the diversity of anomalies, supervised learning methods require large labelled dataset exhibiting…

Metal artifacts in computed tomography (CT) arise from a mismatch between physics of image formation and idealized assumptions during tomographic reconstruction. These artifacts are particularly strong around metal implants, inhibiting…

Image and Video Processing · Electrical Eng. & Systems 2019-09-20 Jan-Nico Zaech , Cong Gao , Bastian Bier , Russell Taylor , Andreas Maier , Nassir Navab , Mathias Unberath

The rapid growth of dermatological imaging and mobile diagnostic tools calls for systems that not only demonstrate empirical performance but also provide strong theoretical guarantees. Deep learning models have shown high predictive…

Machine Learning · Computer Science 2026-01-07 Rohit Kaushik , Eva Kaushik

In this work, we benchmark with different backbones and study their impact for self-supervised learning (SSL) as an auxiliary task to blend texture-based local descriptors into feature modelling for efficient face analysis. It is…

Computer Vision and Pattern Recognition · Computer Science 2026-03-24 Shukesh Reddy , Abhijit Das

Coronary calcification creates blooming artifacts in Computed Tomography Angiography (CTA), severely hampering the diagnosis of lumen stenosis. While Deep Convolutional Neural Networks (DCNNs) like Dense-Unet have shown promise in removing…

Computer Vision and Pattern Recognition · Computer Science 2026-01-07 Mo Chen

This paper explores improvements to the masked image modeling (MIM) paradigm. The MIM paradigm enables the model to learn the main object features of the image by masking the input image and predicting the masked part by the unmasked part.…

Computer Vision and Pattern Recognition · Computer Science 2022-05-24 Jiawei Mao , Xuesong Yin , Yuanqi Chang , Honggu Zhou

The past year has witnessed a rapid development of masked image modeling (MIM). MIM is mostly built upon the vision transformers, which suggests that self-supervised visual representations can be done by masking input image parts while…

Computer Vision and Pattern Recognition · Computer Science 2022-03-29 Yunjie Tian , Lingxi Xie , Jiemin Fang , Mengnan Shi , Junran Peng , Xiaopeng Zhang , Jianbin Jiao , Qi Tian , Qixiang Ye