中文
相关论文

相关论文: G2D: from GTA to Data

200 篇论文

Reconstructing 3D clothed human avatars from single images is a challenging task, especially when encountering complex poses and loose clothing. Current methods exhibit limitations in performance, largely attributable to their dependence on…

计算机视觉与模式识别 · 计算机科学 2023-10-24 Zechuan Zhang , Li Sun , Zongxin Yang , Ling Chen , Yi Yang

Simultaneously localizing camera poses and constructing Gaussian radiance fields in dynamic scenes establish a crucial bridge between 2D images and the 4D real world. Instead of removing dynamic objects as distractors and reconstructing…

计算机视觉与模式识别 · 计算机科学 2025-03-24 Yanyan Li , Youxu Fang , Zunjie Zhu , Kunyi Li , Yong Ding , Federico Tombari

Cognitive radars are systems that rely on learning through interactions of the radar with the surrounding environment. To realize this, radar transmit parameters can be adapted such that they facilitate some downstream task. This paper…

信号处理 · 电气工程与系统科学 2021-12-15 Tristan S. W. Stevens , R. Firat Tigrek , Eric S. Tammam , Ruud J. G. van Sloun

We propose a system that uses video as the input to track the position of objects relative to their surrounding environment in real-time. The neural network employed is trained on a 100% synthetic dataset coming from our own automated…

计算机视觉与模式识别 · 计算机科学 2020-10-30 David Albarracín , Jesús Hormigo , José David Fernández

We seek to learn a generalizable goal-conditioned policy that enables zero-shot robot manipulation: interacting with unseen objects in novel scenes without test-time adaptation. While typical approaches rely on a large amount of…

机器人学 · 计算机科学 2024-08-12 Homanga Bharadhwaj , Roozbeh Mottaghi , Abhinav Gupta , Shubham Tulsiani

In this work, we construct a large-scale dataset for Ground-to-Aerial Person Search, named G2APS, which contains 31,770 images of 260,559 annotated bounding boxes for 2,644 identities appearing in both of the UAVs and ground surveillance…

计算机视觉与模式识别 · 计算机科学 2023-08-25 Shizhou Zhang , Qingchun Yang , De Cheng , Yinghui Xing , Guoqiang Liang , Peng Wang , Yanning Zhang

We introduce a new RGB-D object dataset captured in the wild called WildRGB-D. Unlike most existing real-world object-centric datasets which only come with RGB capturing, the direct capture of the depth channel allows better 3D annotations…

计算机视觉与模式识别 · 计算机科学 2024-07-30 Hongchi Xia , Yang Fu , Sifei Liu , Xiaolong Wang

This paper aims to investigate representation learning for large scale visual place recognition, which consists of determining the location depicted in a query image by referring to a database of reference images. This is a challenging task…

计算机视觉与模式识别 · 计算机科学 2022-10-20 Amar Ali-bey , Brahim Chaib-draa , Philippe Giguère

Appearance-based gaze estimation, which uses only a regular camera to estimate human gaze, is important in various application fields. While the technique faces data bias issues, data collection protocol is often demanding, and collecting…

人机交互 · 计算机科学 2024-09-04 Mingtao Yue , Tomomi Sayuda , Miles Pennington , Yusuke Sugano

Head avatar reconstruction, crucial for applications in virtual reality, online meetings, gaming, and film industries, has garnered substantial attention within the computer vision community. The fundamental objective of this field is to…

计算机视觉与模式识别 · 计算机科学 2024-01-19 Xuangeng Chu , Yu Li , Ailing Zeng , Tianyu Yang , Lijian Lin , Yunfei Liu , Tatsuya Harada

The metaverse is a virtual space that combines physical and digital elements, creating immersive and connected digital worlds. For autonomous mobility, it enables new possibilities with edge computing and digital twins (DTs) that offer…

计算机视觉与模式识别 · 计算机科学 2025-04-25 Eugen Šlapak , Matúš Dopiriak , Mohammad Abdullah Al Faruque , Juraj Gazda , Marco Levorato

We present VGGT, a feed-forward neural network that directly infers all key 3D attributes of a scene, including camera parameters, point maps, depth maps, and 3D point tracks, from one, a few, or hundreds of its views. This approach is a…

计算机视觉与模式识别 · 计算机科学 2025-03-17 Jianyuan Wang , Minghao Chen , Nikita Karaev , Andrea Vedaldi , Christian Rupprecht , David Novotny

Digital Twins (DTs) for physical wireless environments have been recently proposed as accurate virtual representations of the propagation environment that can enable multi-layer decisions at the physical communication equipment. At…

信号处理 · 电气工程与系统科学 2024-07-18 Lorenzo Cazzella , Francesco Linsalata , Maurizio Magarini , Matteo Matteucci , Umberto Spagnolini

We present a system to capture video footage of human subjects in the real world. Our system leverages a quadrotor camera to automatically capture well-composed video of two subjects. Subjects are tracked in a large-scale outdoor…

Graph Neural Networks (GNNs) achieve high performance but can be opaque to humans, making it difficult to understand and compare the many proposed architectures. While existing explainability methods attribute individual predictions to…

机器学习 · 计算机科学 2026-05-11 Debolina Halder Lina , Arlei Silva

Multi-view data capture permits free-viewpoint video (FVV) content creation. To this end, several users must capture video streams, calibrated in both time and pose, framing the same object/scene, from different viewpoints. New-generation…

多媒体 · 计算机科学 2020-05-08 Matteo Bortolon , Paul Chippendale , Stefano Messelodi , Fabio Poiesi

With increasing automation in passenger vehicles, the study of safe and smooth occupant-vehicle interaction and control transitions is key. In this study, we focus on the development of contextual, semantically meaningful representations of…

机器人学 · 计算机科学 2021-07-26 Akshay Rangesh , Nachiket Deo , Ross Greer , Pujitha Gunaratne , Mohan M. Trivedi

Recent advancements in computer graphics technology allow more realistic ren-dering of car driving environments. They have enabled self-driving car simulators such as DeepGTA-V and CARLA (Car Learning to Act) to generate large amounts of…

计算机视觉与模式识别 · 计算机科学 2022-07-04 Minh Cao , Ramin Ramezani

Radar-based human activity recognition (HAR) still lacks a comprehensive simulation method. Existing software is developed based on models or motion-captured data, resulting in limited flexibility. To address this issue, a simulator that…

信号处理 · 电气工程与系统科学 2025-11-13 Weicheng Gao

Synthesizing free-view photo-realistic images is an important task in multimedia. With the development of advanced driver assistance systems~(ADAS) and their applications in autonomous vehicles, experimenting with different scenarios…

计算机视觉与模式识别 · 计算机科学 2022-05-13 Zhuopeng Li , Lu Li , Zeyu Ma , Ping Zhang , Junbo Chen , Jianke Zhu