中文
相关论文

相关论文: Computer Vision and Deep Learning for 4D Augmented…

200 篇论文

Interactive volume visualization using a mixed reality (MR) system helps provide users with an intuitive spatial perception of volumetric data. Due to sophisticated requirements of user interaction and vision when using MR head-mounted…

图形学 · 计算机科学 2023-09-06 Haojie Cheng , Chunxiao Xu , Xujing Chen , Zhenxin Chen , Jiajun Wang , Lingxiao Zhao

The availability of affordable 3D full body reconstruction systems has given rise to free-viewpoint video (FVV) of human shapes. Most existing solutions produce temporally uncorrelated point clouds or meshes with unknown point/vertex…

计算机视觉与模式识别 · 计算机科学 2018-10-15 Zhong Li , Minye Wu , Wangyiteng Zhou , Jingyi Yu

While deep learning methods have achieved impressive success in many vision benchmarks, it remains difficult to understand and explain the representations and decisions of these models. Though vision models are typically trained on 2D…

计算机视觉与模式识别 · 计算机科学 2026-02-24 Benjamin Beilharz , Thomas S. A. Wallis

Human performance capture is a highly important computer vision problem with many applications in movie production and virtual/augmented reality. Many previous performance capture approaches either required expensive multi-view setups or…

计算机视觉与模式识别 · 计算机科学 2021-11-23 Marc Habermann , Weipeng Xu , Michael Zollhoefer , Gerard Pons-Moll , Christian Theobalt

Convolutional neural networks (CNNs) have been extensively applied for image recognition problems giving state-of-the-art results on recognition, detection, segmentation and retrieval. In this work we propose and evaluate several deep…

计算机视觉与模式识别 · 计算机科学 2015-04-14 Joe Yue-Hei Ng , Matthew Hausknecht , Sudheendra Vijayanarasimhan , Oriol Vinyals , Rajat Monga , George Toderici

In this paper, we present preliminary results on the use of deep learning techniques to integrate the users self-body and other participants into a head-mounted video see-through augmented virtuality scenario. It has been previously shown…

计算机视觉与模式识别 · 计算机科学 2020-01-03 Pierre-Olivier Pigny , Lionel Dominjon

Augmented Reality is a topic of foremost interest nowadays. Its main goal is to seamlessly blend virtual content in real-world scenes. Due to the lack of computational power in mobile devices, rendering a virtual object with high-quality,…

计算机视觉与模式识别 · 计算机科学 2018-09-24 Rafael Monroy , Matis Hudon , Aljosa Smolic

We introduce 3D4D, an interactive 4D visualization framework that integrates WebGL with Supersplat rendering. It transforms static images and text into coherent 4D scenes through four core modules and employs a foveated rendering strategy…

计算机视觉与模式识别 · 计算机科学 2025-11-12 Yunhong He , Zhengqing Yuan , Zhengzhong Tu , Yanfang Ye , Lichao Sun

Learning high-quality video representation has shown significant applications in computer vision and remains challenging. Previous work based on mask autoencoders such as ImageMAE and VideoMAE has proven the effectiveness of learning…

计算机视觉与模式识别 · 计算机科学 2023-12-22 Xingjian Diao , Ming Cheng , Shitong Cheng

Complex-valued data is ubiquitous in signal and image processing applications, and complex-valued representations in deep learning have appealing theoretical properties. While these aspects have long been recognized, complex-valued deep…

计算机视觉与模式识别 · 计算机科学 2020-11-10 Rudrasis Chakraborty , Yifei Xing , Stella Yu

Virtual Reality, and Extended Reality in general, connect the physical body with the virtual world. Movement of our body translates to interactions with this virtual world. Only by moving our head will we see a different perspective. By…

多媒体 · 计算机科学 2022-12-09 Glenn Van Wallendael , Lucas Liegeois , Julie Artois , Peter Lambert

Recent advances in 3D perception have shown impressive progress in understanding geometric structures of 3Dshapes and even scenes. Inspired by these advances in geometric understanding, we aim to imbue image-based perception with…

计算机视觉与模式识别 · 计算机科学 2021-12-21 Ji Hou , Saining Xie , Benjamin Graham , Angela Dai , Matthias Nießner

Rendering an accurate image of an isosurface in a volumetric field typically requires large numbers of data samples. Reducing the number of required samples lies at the core of research in volume rendering. With the advent of deep learning…

图形学 · 计算机科学 2022-05-31 Sebastian Weiss , Mengyu Chu , Nils Thuerey , Rüdiger Westermann

We propose a solution to the novel task of rendering sharp videos from new viewpoints from a single motion-blurred image of a face. Our method handles the complexity of face blur by implicitly learning the geometry and motion of faces…

计算机视觉与模式识别 · 计算机科学 2021-12-15 Givi Meishvili , Attila Szabó , Simon Jenni , Paolo Favaro

Augmented Reality has been subject to various integration efforts within industries due to its ability to enhance human machine interaction and understanding. Neural networks have achieved remarkable results in areas of computer vision,…

计算机视觉与模式识别 · 计算机科学 2020-01-17 Linh Kästner , Daniel Dimitrov , Jens Lambrecht

Recognizing human actions in untrimmed videos is an important challenging task. An effective 3D motion representation and a powerful learning model are two key factors influencing recognition performance. In this paper we introduce a new…

计算机视觉与模式识别 · 计算机科学 2018-12-31 Huy-Hieu Pham , Louahdi Khoudour , Alain Crouzil , Pablo Zegers , Sergio A. Velastin

Videos are inherently multimodal. This paper studies the problem of how to fully exploit the abundant multimodal clues for improved video categorization. We introduce a hybrid deep learning framework that integrates useful clues from…

多媒体 · 计算机科学 2017-06-15 Yu-Gang Jiang , Zuxuan Wu , Jinhui Tang , Zechao Li , Xiangyang Xue , Shih-Fu Chang

We introduce a deep multitask architecture to integrate multityped representations of multimodal objects. This multitype exposition is less abstract than the multimodal characterization, but more machine-friendly, and thus is more precise…

机器学习 · 统计学 2016-03-07 Truyen Tran , Dinh Phung , Svetha Venkatesh

The emergence of neural representations has revolutionized our means for digitally viewing a wide range of 3D scenes, enabling the synthesis of photorealistic images rendered from novel views. Recently, several techniques have been proposed…

计算机视觉与模式识别 · 计算机科学 2025-02-14 Gal Fiebelman , Tamir Cohen , Ayellet Morgenstern , Peter Hedman , Hadar Averbuch-Elor

Extended reality (XR) technologies-encompassing virtual reality (VR), augmented reality (AR), and mixed reality (MR) are transforming cognitive assessment and training by offering immersive, interactive environments that simulate real-world…

人机交互 · 计算机科学 2025-01-15 Palmira Victoria González-Erena , Sara Fernández-Guinea , Panagiotis Kourtesis