English
Related papers

Related papers: Transformer-Based Tooth Alignment Prediction With …

200 papers

Deep Neural Networks have shown promising classification performance when predicting certain biomarkers from Whole Slide Images in digital pathology. However, the calibration of the networks' output probabilities is often not evaluated.…

Image and Video Processing · Electrical Eng. & Systems 2023-12-18 Alexander Kurz , Hendrik A. Mehrtens , Tabea-Clara Bucher , Titus J. Brinker

Occlusion is one of the challenging issues when estimating 3D hand pose. This problem becomes more prominent when hand interacts with an object or two hands are involved. In the past works, much attention has not been given to these…

Computer Vision and Pattern Recognition · Computer Science 2025-03-28 Mallika Garg , Debashis Ghosh , Pyari Mohan Pradhan

Most existing approaches for point cloud normal estimation aim to locally fit a geometric surface and calculate the normal from the fitted surface. Recently, learning-based methods have adopted a routine of predicting point-wise weights to…

Computer Vision and Pattern Recognition · Computer Science 2023-03-31 Hang Du , Xuejun Yan , Jingjing Wang , Di Xie , Shiliang Pu

Diffusion models have recently shown promise in offline RL. However, these methods often suffer from high training costs and slow convergence, particularly when using transformer-based denoising backbones. While several optimization…

Machine Learning · Computer Science 2025-06-23 Zhiying Qiu , Tao Lin

Tooth segmentation from intraoral scans is a crucial part of digital dentistry. Many Deep Learning based tooth segmentation algorithms have been developed for this task. In most of the cases, high accuracy has been achieved, although, most…

Computer Vision and Pattern Recognition · Computer Science 2023-05-02 Ananya Jana , Aniruddha Maiti , Dimitris N. Metaxas

In the field of 3D Human Pose Estimation from monocular videos, the presence of diverse occlusion types presents a formidable challenge. Prior research has made progress by harnessing spatial and temporal cues to infer 3D poses from 2D…

Computer Vision and Pattern Recognition · Computer Science 2024-10-08 Mehwish Ghafoor , Arif Mahmood , Muhammad Bilal

State-of-the-art language models are becoming increasingly large in an effort to achieve the highest performance on large corpora of available textual data. However, the sheer size of the Transformer architectures makes it difficult to…

Machine Learning · Computer Science 2024-03-22 Tycho F. A. van der Ouderaa , Markus Nagel , Mart van Baalen , Yuki M. Asano , Tijmen Blankevoort

Background and Aim: Over-fitting issue has been the reason behind deep learning technology not being successfully implemented in oral cancer images classification. The aims of this research were reducing overfitting for accurately producing…

Image and Video Processing · Electrical Eng. & Systems 2022-08-17 Prakrit Joshi , Omar Hisham Alsadoon , Abeer Alsadoon , Nada AlSallami , Tarik A. Rashid , P. W. C. Prasad , Sami Haddad

We introduce a novel framework for multiway point cloud mosaicking (named Wednesday), designed to co-align sets of partially overlapping point clouds -- typically obtained from 3D scanners or moving RGB-D cameras -- into a unified…

Computer Vision and Pattern Recognition · Computer Science 2024-04-02 Shengze Jin , Iro Armeni , Marc Pollefeys , Daniel Barath

Large occlusions result in a significant decline in image classification accuracy. During inference, diverse types of unseen occlusions introduce out-of-distribution data to the classification model, leading to accuracy dropping as low as…

Computer Vision and Pattern Recognition · Computer Science 2024-02-13 Ketan Kotwal , Tanay Deshmukh , Preeti Gopal

Image-guided mouse irradiation is essential to understand interventions involving radiation prior to human studies. Our objective is to employ Swin UNEt Transformers (Swin UNETR) to segment native micro-CT and contrast-enhanced micro-CT…

Medical Physics · Physics 2024-05-30 Lu Jiang , Di Xu , Qifan Xu , Arion Chatziioannou , Keisuke S. Iwamoto , Susanta Hui , Ke Sheng

Estimating the layout of a room from a single-shot panoramic image is important in virtual/augmented reality and furniture layout simulation. This involves identifying three-dimensional (3D) geometry, such as the location of corners and…

Computer Vision and Pattern Recognition · Computer Science 2023-04-26 Mizuki Tabata , Kana Kurata , Junichiro Tamamatsu

This paper proposes the first pure Transformer structure inversion network called SwinStyleformer, which can compensate for the shortcomings of the CNNs inversion framework by handling long-range dependencies and learning the global…

Computer Vision and Pattern Recognition · Computer Science 2024-06-21 Jiawei Mao , Guangyi Zhao , Xuesong Yin , Yuanqi Chang

Medical image segmentation is a crucial task in the field of medical image analysis. Harmonizing the convolution and multi-head self-attention mechanism is a recent research focus in this field, with various combination methods proposed.…

Image and Video Processing · Electrical Eng. & Systems 2023-09-28 Lichao Wang , Jiahao Huang , Xiaodan Xing , Guang Yang

Optical Intra-oral Scanners (IOS) are widely used in digital dentistry, providing 3-Dimensional (3D) and high-resolution geometrical information of dental crowns and the gingiva. Accurate 3D tooth segmentation, which aims to precisely…

Computer Vision and Pattern Recognition · Computer Science 2022-11-01 Huimin Xiong , Kunle Li , Kaiyuan Tan , Yang Feng , Joey Tianyi Zhou , Jin Hao , Zuozhu Liu

Neural networks with self-attention (a.k.a. Transformers) like ViT and Swin have emerged as a better alternative to traditional convolutional neural networks (CNNs). However, our understanding of how the new architecture works is still…

Computer Vision and Pattern Recognition · Computer Science 2023-12-15 Juyeop Kim , Junha Park , Songkuk Kim , Jong-Seok Lee

Teeth segmentation is an essential task in dental image analysis for accurate diagnosis and treatment planning. While supervised deep learning methods can be utilized for teeth segmentation, they often require extensive manual annotation of…

Computer Vision and Pattern Recognition · Computer Science 2024-02-27 Tomáš Kunzo , Viktor Kocur , Lukáš Gajdošech , Martin Madaras

Learning and predicting the pose parameters of a 3D hand model given an image, such as locations of hand joints, is challenging due to large viewpoint changes and articulations, and severe self-occlusions exhibited particularly in…

Computer Vision and Pattern Recognition · Computer Science 2018-05-23 Qi Ye , Tae-Kyun Kim

In the medical domain, acquiring large datasets poses significant challenges due to privacy concerns. Nonetheless, the development of a robust deep-learning model for retinal disease diagnosis necessitates a substantial dataset for…

Computer Vision and Pattern Recognition · Computer Science 2024-09-18 Fatema-E- Jannat , Sina Gholami , Jennifer I. Lim , Theodore Leng , Minhaj Nur Alam , Hamed Tabkhi

The vision community is witnessing a modeling shift from CNNs to Transformers, where pure Transformer architectures have attained top accuracy on the major video recognition benchmarks. These video models are all built on Transformer layers…

Computer Vision and Pattern Recognition · Computer Science 2021-06-25 Ze Liu , Jia Ning , Yue Cao , Yixuan Wei , Zheng Zhang , Stephen Lin , Han Hu