中文
相关论文

相关论文: Deep Volumetric Ambient Occlusion

200 篇论文

Deformable image registration estimates voxel-wise correspondences between images through spatial transformations, and plays a key role in medical imaging. While deep learning methods have significantly reduced runtime, efficiently handling…

计算机视觉与模式识别 · 计算机科学 2025-10-28 Tianran Li , Marius Staring , Yuchuan Qiao

In this paper, we propose a deep learning based multi-speaker direction of arrival (DOA) estimation with audio and visual signals by using permutation-free loss function. We first collect a data set for multi-modal sound source localization…

音频与语音处理 · 电气工程与系统科学 2022-10-27 Qing Wang , Hang Chen , Ya Jiang , Zhe Wang , Yuyang Wang , Jun Du , Chin-Hui Lee

Moving around in the world is naturally a multisensory experience, but today's embodied agents are deaf---restricted to solely their visual perception of the environment. We introduce audio-visual navigation for complex, acoustically and…

Multiple object tracking (MOT) tends to become more challenging when severe occlusions occur. In this paper, we analyze the limitations of traditional Convolutional Neural Network-based methods and Transformer-based methods in handling…

计算机视觉与模式识别 · 计算机科学 2023-09-12 Teng Fu , Xiaocong Wang , Haiyang Yu , Ke Niu , Bin Li , Xiangyang Xue

Monocular omnidirectional visual odometry (OVO) systems leverage 360-degree cameras to overcome field-of-view limitations of perspective VO systems. However, existing methods, reliant on handcrafted features or photometric objectives, often…

计算机视觉与模式识别 · 计算机科学 2026-01-12 Xiaopeng Guo , Yinzhe Xu , Huajian Huang , Sai-Kit Yeung

Odometry is of key importance for localization in the absence of a map. There is considerable work in the area of visual odometry (VO), and recent advances in deep learning have brought novel approaches to VO, which directly learn salient…

计算机视觉与模式识别 · 计算机科学 2020-03-06 Wei Wang , Muhamad Risqi U. Saputra , Peijun Zhao , Pedro Gusmao , Bo Yang , Changhao Chen , Andrew Markham , Niki Trigoni

In this work we present a monocular visual odometry (VO) algorithm which leverages geometry-based methods and deep learning. Most existing VO/SLAM systems with superior performance are based on geometry and have to be carefully designed for…

计算机视觉与模式识别 · 计算机科学 2020-02-19 Huangying Zhan , Chamara Saroj Weerasekera , Jiawang Bian , Ian Reid

Neural volumetric representations have become a widely adopted model for radiance fields in 3D scenes. These representations are fully implicit or hybrid function approximators of the instantaneous volumetric radiance in a scene, which are…

计算机视觉与模式识别 · 计算机科学 2023-05-09 Yuval Bahat , Yuxuan Zhang , Hendrik Sommerhoff , Andreas Kolb , Felix Heide

The Variational Autoencoder (VAE) is a powerful deep generative model that is now extensively used to represent high-dimensional complex data via a low-dimensional latent space learned in an unsupervised manner. In the original VAE model,…

声音 · 计算机科学 2021-06-15 Xiaoyu Bie , Laurent Girin , Simon Leglaive , Thomas Hueber , Xavier Alameda-Pineda

We propose a deep learning strategy to estimate the mean curvature of two-dimensional implicit interfaces in the level-set method. Our approach is based on fitting feed-forward neural networks to synthetic data sets constructed from…

数值分析 · 数学 2022-09-29 Luis Ángel Larios-Cárdenas , Frederic Gibou

Wireless-connected Virtual Reality (VR) provides immersive experience for VR users from any-where at anytime. However, providing wireless VR users with seamless connectivity and real-time VR video with high quality is challenging due to its…

信号处理 · 电气工程与系统科学 2020-05-19 Xiaonan Liu , Yansha Deng

Autonomous navigation in crowded spaces poses a challenge for mobile robots due to the highly dynamic, partially observable environment. Occlusions are highly prevalent in such settings due to a limited sensor field of view and obstructing…

机器人学 · 计算机科学 2023-05-02 Ye-Ji Mun , Masha Itkina , Shuijing Liu , Katherine Driggs-Campbell

Face recognition remains a challenging task in unconstrained scenarios, especially when faces are partially occluded. To improve the robustness against occlusion, augmenting the training images with artificial occlusions has been proved as…

计算机视觉与模式识别 · 计算机科学 2022-01-19 Mingjie He , Jie Zhang , Shiguang Shan , Xiao Liu , Zhongqin Wu , Xilin Chen

Volumetric imaging by fluorescence microscopy is often limited by anisotropic spatial resolution from inferior axial resolution compared to the lateral resolution. To address this problem, here we present a deep-learning-enabled…

计算机视觉与模式识别 · 计算机科学 2022-07-06 Hyoungjun Park , Myeongsu Na , Bumju Kim , Soohyun Park , Ki Hean Kim , Sunghoe Chang , Jong Chul Ye

This paper explores how deep learning techniques can improve visual-based SLAM performance in challenging environments. By combining deep feature extraction and deep matching methods, we introduce a versatile hybrid visual SLAM system…

机器人学 · 计算机科学 2024-06-05 Zhang Xiao , Shuaixin Li

The near-field effect of short-range multiple-input multiple-output (MIMO) systems imposes many challenges on direction-of-arrival (DoA) estimation. Most conventional scenarios assume that the far-field planar wavefronts hold. In this…

信号处理 · 电气工程与系统科学 2020-07-22 Yashuai Cao , Tiejun Lv , Zhipeng Lin , Pingmu Huang , Fuhong Lin

Creating realistic virtual assets is a time-consuming process: it usually involves an artist designing the object, then spending a lot of effort on tweaking its appearance. Intricate details and certain effects, such as subsurface…

计算机视觉与模式识别 · 计算机科学 2022-12-13 Aljaž Božič , Denis Gladkov , Luke Doukakis , Christoph Lassner

This paper proposes a deconvolution-based network (DCNN) model for DOA estimation of direct source and early reflections under reverberant scenarios. Considering that the first-order reflections of the sound source also contain spatial…

音频与语音处理 · 电气工程与系统科学 2021-10-25 Shan Gao , Xihong Wu , Tianshu Qu

We present a method for estimating intravoxel parameters from a DW-MRI based on deep learning techniques. We show that neural networks (DNNs) have the potential to extract information from diffusion-weighted signals to reconstruct cerebral…

图像与视频处理 · 电气工程与系统科学 2022-01-02 Hanna Ehrlich , Mariano Rivera

We propose a transfer deep learning (TDL) framework that can transfer the knowledge obtained from a single-modal neural network to a network with a different modality. Specifically, we show that we can leverage speech data to fine-tune the…

神经与进化计算 · 计算机科学 2016-02-19 Seungwhan Moon , Suyoun Kim , Haohan Wang
‹ 上一页 1 8 9 10 下一页 ›