English
Related papers

Related papers: Morphology-Guided Cross-Task Coupling for Joint Bu…

200 papers

Total-body PET/CT enables system-wide molecular imaging, but heterogeneous anatomical and metabolic signals, approximately 2 m axial coverage, and structured radiology semantics challenge existing medical AI models that assume…

Computer Vision and Pattern Recognition · Computer Science 2026-01-21 Wei Chen , Liang Wu , Shuyi Lu , Yuanyuan Sun , Wenkai Bi , Zilong Yuan , Yaoyao He , Feng Wang , Junchi Ma , Shuyong Liu , Zhaoping Cheng , Xiaoyan Hu , Jianfeng Qiu

Deep learning-based methods have been extensively explored for automatic building mapping from high-resolution remote sensing images over recent years. While most building mapping models produce vector polygons of buildings for geographic…

Computer Vision and Pattern Recognition · Computer Science 2024-01-11 Mingming Zhang , Qingjie Liu , Yunhong Wang

Encoder-decoder networks become a popular choice for various medical image segmentation tasks. When they are trained with a standard loss function, these networks are not explicitly enforced to preserve the shape integrity of an object in…

Computer Vision and Pattern Recognition · Computer Science 2023-09-22 Mehmet Bahadir Erden , Selahattin Cansiz , Onur Caki , Haya Khattak , Durmus Etiz , Melek Cosar Yakar , Kerem Duruer , Berke Barut , Cigdem Gunduz-Demir

Hybrid simulation (HS) is a widely used structural testing method that combines a computational substructure with a numerical model for well-understood components and an experimental substructure for other parts of the structure that are…

Machine Learning · Computer Science 2020-04-07 Elif Ecem Bas , Mohamed A. Moustafa , David Feil-Seifer , Janelle Blankenburg

Recently, deep learning based facial landmark detection (FLD) methods have achieved considerable success. However, in challenging scenarios such as large pose variations, illumination changes, and facial expression variations, they still…

Computer Vision and Pattern Recognition · Computer Science 2026-01-21 Jun Wan , Xinyu Xiong , Ning Chen , Zhihui Lai , Jie Zhou , Wenwen Min

We describe a novel metric-based learning approach that introduces a multimodal framework and uses deep audio and geophone encoders in siamese configuration to design an adaptable and lightweight supervised model. This framework eliminates…

Sound · Computer Science 2021-11-16 Muhammad Shakeel , Katsutoshi Itoyama , Kenji Nishida , Kazuhiro Nakadai

Transformer-based approaches have been successfully proposed for 3D human pose estimation (HPE) from 2D pose sequence and achieved state-of-the-art (SOTA) performance. However, current SOTAs have difficulties in modeling spatial-temporal…

Computer Vision and Pattern Recognition · Computer Science 2023-01-19 Xiaoye Qian , Youbao Tang , Ning Zhang , Mei Han , Jing Xiao , Ming-Chun Huang , Ruei-Sung Lin

Geometric differences between cross-view images, such as drone and satellite views, significantly increase the challenge of Cross-View Geo-Localization (CVGL), which aims to acquire the geolocation of images by image retrieval. To further…

Computer Vision and Pattern Recognition · Computer Science 2026-05-12 Wei Wang , Dou Quan , Ning Huyan , Shuang Wang , Yi Li , Pei He , Licheng Jiao

Change detection in remote sensing imagery is essential for a variety of applications such as urban planning, disaster management, and climate research. However, existing methods for identifying semantically changed areas overlook the…

Computer Vision and Pattern Recognition · Computer Science 2023-12-08 Maximilian Bernhard , Niklas Strauß , Matthias Schubert

Foreground segmentation algorithms aim segmenting moving objects from the background in a robust way under various challenging scenarios. Encoder-decoder type deep neural networks that are used in this domain recently perform impressive…

Computer Vision and Pattern Recognition · Computer Science 2019-09-04 Long Ang Lim , Hacer Yalim Keles

Wireless communications at high-frequency bands with large antenna arrays face challenges in beam management, which can potentially be improved by multimodality sensing information from cameras, LiDAR, radar, and GPS. In this paper, we…

Signal Processing · Electrical Eng. & Systems 2023-09-22 Yu Tian , Qiyang Zhao , Zine el abidine Kherroubi , Fouzi Boukhalfa , Kebin Wu , Faouzi Bader

We propose a pipeline for combined multi-class object geolocation and height estimation from street level RGB imagery, which is considered as a single available input data modality. Our solution is formulated via Markov Random Field…

Computer Vision and Pattern Recognition · Computer Science 2023-05-16 Matej Ulicny , Vladimir A. Krylov , Julie Connelly , Rozenn Dahyot

Human motion transfer aims to transfer motions from a target dynamic person to a source static one for motion synthesis. An accurate matching between the source person and the target motion in both large and subtle motion changes is vital…

Computer Vision and Pattern Recognition · Computer Science 2023-02-28 Hongyu Liu , Xintong Han , Chengbin Jin , Lihui Qian , Huawei Wei , Zhe Lin , Faqiang Wang , Haoye Dong , Yibing Song , Jia Xu , Qifeng Chen

Seismology faces fundamental challenges in state forecasting and reconstruction (e.g., earthquake early warning and ground motion prediction) and managing the parametric variability of source locations, mechanisms, and Earth models (e.g.,…

Camera calibration consists of estimating camera parameters such as the zenith vanishing point and horizon line. Estimating the camera parameters allows other tasks like 3D rendering, artificial reality effects, and object insertion in an…

Computer Vision and Pattern Recognition · Computer Science 2024-09-25 Sebastian Janampa , Marios Pattichis

The existing binary foreground map (FM) measures to address various types of errors in either pixel-wise or structural ways. These measures consider pixel-level match or image-level information independently, while cognitive vision studies…

Computer Vision and Pattern Recognition · Computer Science 2019-09-04 Deng-Ping Fan , Cheng Gong , Yang Cao , Bo Ren , Ming-Ming Cheng , Ali Borji

This research paper presents an innovative multi-task learning framework that allows concurrent depth estimation and semantic segmentation using a single camera. The proposed approach is based on a shared encoder-decoder architecture, which…

Computer Vision and Pattern Recognition · Computer Science 2024-03-19 Pardis Taghavi , Reza Langari , Gaurav Pandey

High-fidelity personalized human musculoskeletal models are crucial for simulating realistic behavior of physically coupled human-robot interactive systems and verifying their safety-critical applications in simulations before actual…

Robotics · Computer Science 2025-08-20 Yingfan Zhou , Philip Sanderink , Sigurd Jager Lemming , Cheng Fang

Multimodal image registration is a fundamental task and a prerequisite for downstream cross-modal analysis. Despite recent progress in shared feature extraction and multi-scale architectures, two key limitations remain. First, some methods…

Computer Vision and Pattern Recognition · Computer Science 2026-03-23 Chunlei Zhang , Jiahao Xia , Yun Xiao , Bo Jiang , Jian Zhang

Wearable sensors enable the continuous acquisition of high-resolution physiological waveforms, such as photoplethysmography and accelerometry, under free-living conditions. However, inferring health-related phenotypes from these signals…