English
Related papers

Related papers: Anatomical Positional Embeddings

200 papers

3D shape analysis is an important research topic in computer vision and graphics. While existing methods have generalized image-based deep learning to meshes using graph-based convolutions, the lack of an effective pooling operation…

Graphics · Computer Science 2019-08-08 Yu-Jie Yuan , Yu-Kun Lai , Jie Yang , Hongbo Fu , Lin Gao

3D reconstruction from 2D inputs, especially for non-rigid objects like humans, presents unique challenges due to the significant range of possible deformations. Traditional methods often struggle with non-rigid shapes, which require…

Computer Vision and Pattern Recognition · Computer Science 2025-05-26 Fahd Alhamazani , Yu-Kun Lai , Paul L. Rosin

Industrial CAD workflows require robust, generalizable 3D geometric representations supporting accuracy and explainability. We introduce Shape, a self-supervised foundation model converting surface meshes into dense per-token embeddings.…

Computer Vision and Pattern Recognition · Computer Science 2026-04-28 Bayangmbe Mounmo , Sam Chien , Mile Mitrovic

This paper proposes an efficient and probabilistic adaptive voxel mapping method for LiDAR odometry. The map is a collection of voxels; each contains one plane (or edge) feature that enables the probabilistic representation of the…

Robotics · Computer Science 2022-07-11 Chongjian Yuan , Wei xu , Xiyuan Liu , Xiaoping Hong , Fu Zhang

This paper addresses the task of detecting and localising fetal anatomical regions in 2D ultrasound images, where only image-level labels are present at training, i.e. without any localisation or segmentation information. We examine the use…

Computer Vision and Pattern Recognition · Computer Science 2018-08-17 Nicolas Toussaint , Bishesh Khanal , Matthew Sinclair , Alberto Gomez , Emily Skelton , Jacqueline Matthew , Julia A. Schnabel

Deep neural networks are powerful tools for biomedical image segmentation. These models are often trained with heavy supervision, relying on pairs of images and corresponding voxel-level labels. However, obtaining segmentations of…

Image and Video Processing · Electrical Eng. & Systems 2020-04-30 Evan M. Yu , Juan Eugenio Iglesias , Adrian V. Dalca , Mert R. Sabuncu

This paper presents CAPE, a method to extract planes and cylinder segments from organized point clouds, which processes 640x480 depth images on a single CPU core at an average of 300 Hz, by operating on a grid of planar cells. While,…

Computer Vision and Pattern Recognition · Computer Science 2018-07-06 Pedro F. Proença , Yang Gao

Health professionals extensively use Two- Dimensional (2D) Ultrasound (US) videos and images to visualize and measure internal organs for various purposes including evaluation of muscle architectural changes. US images can be used to…

Image and Video Processing · Electrical Eng. & Systems 2021-10-01 Alzayat Saleh , Issam H. Laradji , Corey Lammie , David Vazquez , Carol A Flavell , Mostafa Rahimi Azghadi

This paper presents a method for automatic segmentation, localization, and identification of vertebrae in arbitrary 3D CT images. Many previous works do not perform the three tasks simultaneously even though requiring a priori knowledge of…

Image and Video Processing · Electrical Eng. & Systems 2020-10-01 Naoto Masuzawa , Yoshiro Kitamura , Keigo Nakamura , Satoshi Iizuka , Edgar Simo-Serra

Active learning (AL) has the potential to drastically reduce annotation costs in 3D biomedical image segmentation, where expert labeling of volumetric data is both time-consuming and expensive. Yet, existing AL methods are unable to…

Computer Vision and Pattern Recognition · Computer Science 2026-01-21 Carsten T. Lüth , Jeremias Traub , Kim-Celine Kahl , Till J. Bungert , Lukas Klein , Lars Krämer , Paul F. Jäger , Klaus Maier-Hein , Fabian Isensee

One of the fundamental challenges in supervised learning for multimodal image registration is the lack of ground-truth for voxel-level spatial correspondence. This work describes a method to infer voxel-level transformation from…

Automatic delineation and measurement of main organs such as liver is one of the critical steps for assessment of hepatic diseases, planning and postoperative or treatment follow-up. However, addressing this problem typically requires…

Computer Vision and Pattern Recognition · Computer Science 2019-04-02 Elena Balashova , Jiangping Wang , Vivek Singh , Bogdan Georgescu , Brian Teixeira , Ankur Kapoor

Transformer architecture has enabled recent progress in speech enhancement. Since Transformers are position-agostic, positional encoding is the de facto standard component used to enable Transformers to distinguish the order of elements in…

Audio and Speech Processing · Electrical Eng. & Systems 2024-02-15 Qiquan Zhang , Meng Ge , Hongxu Zhu , Eliathamby Ambikairajah , Qi Song , Zhaoheng Ni , Haizhou Li

Volumetric image compression has become an urgent task to effectively transmit and store images produced in biological research and clinical practice. At present, the most commonly used volumetric image compression methods are based on…

Image and Video Processing · Electrical Eng. & Systems 2022-10-19 Dongmei Xue , Haichuan Ma , Li Li , Dong Liu , Zhiwei Xiong

Accurate 3D geometric perception is an important prerequisite for a wide range of spatial AI systems. While state-of-the-art methods depend on large-scale training data, acquiring consistent and precise 3D annotations from in-the-wild…

3D Large Vision-Language Models (3D LVLMs) built upon Large Language Models (LLMs) have achieved remarkable progress across various multimodal tasks. However, their inherited position-dependent modeling mechanism, Rotary Position Embedding…

Computer Vision and Pattern Recognition · Computer Science 2026-02-27 Guanting Ye , Qiyan Zhao , Wenhao Yu , Liangyu Yuan , Mingkai Li , Xiaofeng Zhang , Jianmin Ji , Yanyong Zhang , Qing Jiang , Ka-Veng Yuen

Sizing and fitting of Personal Protective Equipment (PPE) is a critical part of the product creation process; however, traditional methods to do this type of work can be labor intensive and based on limited or non-representative…

Machine Learning · Computer Science 2021-05-24 Jacob A. Searcy , Susan L. Sokolowski

Estimating displacement vector field via a cost volume computed in the feature space has shown great success in image registration, but it suffers excessive computation burdens. Moreover, existing feature descriptors only extract local…

Computer Vision and Pattern Recognition · Computer Science 2025-06-30 Zi Li , Lin Tian , Tony C. W. Mok , Xiaoyu Bai , Puyang Wang , Jia Ge , Jingren Zhou , Le Lu , Xianghua Ye , Ke Yan , Dakai Jin

Fully Convolutional Neural Networks (F-CNNs) achieve state-of-the-art performance for image segmentation in medical imaging. Recently, squeeze and excitation (SE) modules and variations thereof have been introduced to recalibrate feature…

Image and Video Processing · Electrical Eng. & Systems 2019-06-13 Anne-Marie Rickmann , Abhijit Guha Roy , Ignacio Sarasua , Nassir Navab , Christian Wachinger

This paper presents Volumetric Transformer Pose estimator (VTP), the first 3D volumetric transformer framework for multi-view multi-person 3D human pose estimation. VTP aggregates features from 2D keypoints in all camera views and directly…

Computer Vision and Pattern Recognition · Computer Science 2023-08-07 Yuxing Chen , Renshu Gu , Ouhan Huang , Gangyong Jia