English
Related papers

Related papers: MinD-3D: Reconstruct High-quality 3D objects in Hu…

200 papers

Neural reconstruction and rendering strategies have demonstrated state-of-the-art performances due, in part, to their ability to preserve high level shape details. Existing approaches, however, either represent objects as implicit surface…

Computer Vision and Pattern Recognition · Computer Science 2023-12-29 Angtian Wang , Yuanlu Xu , Nikolaos Sarafianos , Robert Maier , Edmond Boyer , Alan Yuille , Tony Tung

3D shape reconstruction is essential in the navigation of minimally-invasive and auto robot-guided surgeries whose operating environments are indirect and narrow, and there have been some works that focused on reconstructing the 3D shape of…

Image and Video Processing · Electrical Eng. & Systems 2021-10-13 Bowen Hu , Baiying Lei , Shuqiang Wang , Yong Liu , Bingchuan Wang , Min Gan , Yanyan Shen

Decoding visual experiences from brain activity is a significant challenge. Existing fMRI-to-video methods often focus on semantic content while overlooking spatial and motion information. However, these aspects are all essential and are…

Computer Vision and Pattern Recognition · Computer Science 2025-04-02 Chong Li , Jingyang Huo , Weikang Gong , Yanwei Fu , Xiangyang Xue , Jianfeng Feng

Longitudinal brain MRI is essential for characterizing the progression of neurological diseases such as Alzheimer's disease assessment. However, current deep-learning tools fragment this process: classifiers reduce a scan to a label,…

Computer Vision and Pattern Recognition · Computer Science 2026-03-17 Zhaoyang Jiang , Zhizhong Fu , David McAllister , Yunsoo Kim , Honghan Wu

Humanoid motion control has witnessed significant breakthroughs in recent years, with deep reinforcement learning (RL) emerging as a primary catalyst for achieving complex, human-like behaviors. However, the high dimensionality and…

Simultaneous imaging of fluorescence-labeled and label-free phase objects in the same sample provides distinct and complementary information. Most multimodal fluorescence-phase imaging operates in transmission mode, capturing fluorescence…

Optics · Physics 2024-08-20 Renzhi He , Yucheng Li , Junjie Chen , Yi Xue

Generating high-quality 3D content from text, single images, or sparse view images remains a challenging task with broad applications. Existing methods typically employ multi-view diffusion models to synthesize multi-view images, followed…

Computer Vision and Pattern Recognition · Computer Science 2025-06-03 Junlin Han , Jianyuan Wang , Andrea Vedaldi , Philip Torr , Filippos Kokkinos

Mental rotation -- the ability to compare objects seen from different viewpoints -- is a fundamental example of mental simulation and spatial world modeling in humans. Here we propose a mechanistic model of human mental rotation, leveraging…

Neurons and Cognition · Quantitative Biology 2026-05-29 Raymond Khazoum , Daniela Fernandes , Aleksandr Krylov , Qin Li , Stephane Deny

Recent research has shown that mmWave radar sensing is effective for object detection in low visibility environments, which makes it an ideal technique in autonomous navigation systems such as autonomous vehicles. However, due to the…

Image and Video Processing · Electrical Eng. & Systems 2021-09-21 Yue Sun , Honggang Zhang , Zhuoming Huang , Benyuan Liu

Though recent advances in vision-language models (VLMs) have achieved remarkable progress across a wide range of multimodal tasks, understanding 3D spatial relationships from limited views remains a significant challenge. Previous reasoning…

Computer Vision and Pattern Recognition · Computer Science 2026-03-16 Zhangquan Chen , Manyuan Zhang , Xinlei Yu , Xufang Luo , Mingze Sun , Zihao Pan , Xiang An , Yan Feng , Peng Pei , Xunliang Cai , Ruqi Huang

Prior works for reconstructing hand-held objects from a single image train models on images paired with 3D shapes. Such data is challenging to gather in the real world at scale. Consequently, these approaches do not generalize well when…

Computer Vision and Pattern Recognition · Computer Science 2024-09-24 Aditya Prakash , Matthew Chang , Matthew Jin , Ruisen Tu , Saurabh Gupta

Magnetic resonance imaging (MRI) reconstruction is a fundamental task aimed at recovering high-quality images from undersampled or low-quality MRI data. This process enhances diagnostic accuracy and optimizes clinical applications. In…

Image and Video Processing · Electrical Eng. & Systems 2025-03-11 Xiaoyan Kui , Zijie Fan , Zexin Ji , Qinsong Li , Chengtao Liu , Weixin Si , Beiji Zou

Precise volumetric delineation of hippocampal structures is essential for quantifying neurodevelopmental trajectories in pre-term and term infants, where subtle morphological variations may carry prognostic significance. While foundation…

Image and Video Processing · Electrical Eng. & Systems 2026-03-02 Annayah Usman , Behraj Khan , Tahir Qasim Syed

This work addresses the problem of recovering complete, simulatable object geometry from reconstructed real-world scenes, enabling physics-based interaction with objects embedded in the scene. While modern multi-view reconstruction methods…

Computer Vision and Pattern Recognition · Computer Science 2026-05-29 Xin Dong , Weijian Deng , Lihan Zhang , Tianru Dai , Wenfeng Deng , Yansong Tang

Functional magnetic resonance imaging (fMRI) is a neuroimaging technique that records neural activations in the brain by capturing the blood oxygen level in different regions based on the task performed by a subject. Given fMRI data, the…

Computer Vision and Pattern Recognition · Computer Science 2021-09-21 Ashish Jaiswal , Ashwin Ramesh Babu , Mohammad Zaki Zadeh , Fillia Makedon , Glenn Wylie

Fully convolutional networks have become the backbone of modern medical imaging due to their ability to learn multi-scale representations and perform end-to-end inference. Yet their potential for slice-to-volume reconstruction (SVR), the…

Image and Video Processing · Electrical Eng. & Systems 2026-01-13 Margherita Firenze , Sean I. Young , Clinton J. Wang , Hyuk Jin Yun , Elfar Adalsteinsson , Kiho Im , P. Ellen Grant , Polina Golland

In this work, we explore the decoding of mental imagery from subjects using their fMRI measurements. In order to achieve this decoding, we first created a mapping between a subject's fMRI signals elicited by the videos the subjects watched.…

Image and Video Processing · Electrical Eng. & Systems 2024-10-02 Arman Afrasiyabi , Erica Busch , Rahul Singh , Dhananjay Bhaskar , Laurent Caplette , Nicholas Turk-Browne , Smita Krishnaswamy

In this paper, we investigate an open research task of cross-modal retrieval between 3D shapes and textual descriptions. Previous approaches mainly rely on point cloud encoders for feature extraction, which may ignore key inherent features…

Computer Vision and Pattern Recognition · Computer Science 2024-05-08 Hao Wu , Ruochong LI , Hao Wang , Hui Xiong

3D imaging enables accurate diagnosis by providing spatial information about organ anatomy. However, using 3D images to train AI models is computationally challenging because they consist of 10x or 100x more pixels than their 2D…

Generating realistic MRIs to accurately predict future changes in the structure of brain is an invaluable tool for clinicians in assessing clinical outcomes and analysing the disease progression at the patient level. However, current…

Computer Vision and Pattern Recognition · Computer Science 2025-09-04 Mattia Litrico , Francesco Guarnera , Mario Valerio Giuffrida , Daniele Ravì , Sebastiano Battiato