中文
相关论文

相关论文: Learning-based Monocular 3D Reconstruction of Bird…

200 篇论文

Monocular 3D object detection (Mono3D) is a fundamental computer vision task that estimates an object's class, 3D position, dimensions, and orientation from a single image. Its applications, including autonomous driving, augmented reality,…

计算机视觉与模式识别 · 计算机科学 2025-08-28 Abhinav Kumar

Monocular estimation of three dimensional human self-contact is fundamental for detailed scene analysis including body language understanding and behaviour modeling. Existing 3d reconstruction methods do not focus on body regions in…

计算机视觉与模式识别 · 计算机科学 2020-12-21 Mihai Fieraru , Mihai Zanfir , Elisabeta Oneata , Alin-Ionut Popa , Vlad Olaru , Cristian Sminchisescu

One major challenge for monocular 3D human pose estimation in-the-wild is the acquisition of training data that contains unconstrained images annotated with accurate 3D poses. In this paper, we address this challenge by proposing a…

计算机视觉与模式识别 · 计算机科学 2020-03-18 Umar Iqbal , Pavlo Molchanov , Jan Kautz

The assessment of laboratory animal behavior is of central interest in modern neuroscience research. Behavior is typically studied in terms of pose changes, which are ideally captured in three dimensions. This requires triangulation over a…

计算机视觉与模式识别 · 计算机科学 2021-06-25 Indrani Sarkar , Indranil Maji , Charitha Omprakash , Sebastian Stober , Sanja Mikulovic , Pavol Bauer

We present a method to infer the 3D pose of mice, including the limbs and feet, from monocular videos. Many human clinical conditions and their corresponding animal models result in abnormal motion, and accurately measuring 3D motion at…

计算机视觉与模式识别 · 计算机科学 2021-06-18 Bo Hu , Bryan Seybold , Shan Yang , David Ross , Avneesh Sud , Graham Ruby , Yi Liu

Convolutions on monocular dash cam videos capture spatial invariances in the image plane but do not explicitly reason about distances and depth. We propose a simple transformation of observations into a bird's eye view, also known as plan…

计算机视觉与模式识别 · 计算机科学 2019-05-17 Dequan Wang , Coline Devin , Qi-Zhi Cai , Philipp Krähenbühl , Trevor Darrell

We present a system to recover the 3D shape and motion of a wide variety of quadrupeds from video. The system comprises a machine learning front-end which predicts candidate 2D joint positions, a discrete optimization which finds…

计算机视觉与模式识别 · 计算机科学 2018-11-15 Benjamin Biggs , Thomas Roddick , Andrew Fitzgibbon , Roberto Cipolla

Deep neural networks have achieved great progress in single-image 3D human reconstruction. However, existing methods still fall short in predicting rare poses. The reason is that most of the current models perform regression based on a…

计算机视觉与模式识别 · 计算机科学 2021-01-01 Yu Rong , Ziwei Liu , Chen Change Loy

We present PlanarRecon -- a novel framework for globally coherent detection and reconstruction of 3D planes from a posed monocular video. Unlike previous works that detect planes in 2D from a single image, PlanarRecon incrementally detects…

计算机视觉与模式识别 · 计算机科学 2022-06-16 Yiming Xie , Matheus Gadelha , Fengting Yang , Xiaowei Zhou , Huaizu Jiang

Learning deformable 3D objects from 2D images is often an ill-posed problem. Existing methods rely on explicit supervision to establish multi-view correspondences, such as template shape models and keypoint annotations, which restricts…

计算机视觉与模式识别 · 计算机科学 2022-06-30 Shangzhe Wu , Tomas Jakab , Christian Rupprecht , Andrea Vedaldi

We focus on the task of estimating a physically plausible articulated human motion from monocular video. Existing approaches that do not consider physics often produce temporally inconsistent output with motion artifacts, while…

计算机视觉与模式识别 · 计算机科学 2022-05-26 Erik Gärtner , Mykhaylo Andriluka , Hongyi Xu , Cristian Sminchisescu

Biodiversity loss poses a significant threat to humanity, making wildlife monitoring essential for assessing ecosystem health. Avian species are ideal subjects for this due to their popularity and the ease of identifying them through their…

机器学习 · 计算机科学 2026-02-23 Nina Brolich , Simon Geis , Maximilian Kasper , Alexander Barnhill , Axel Plinge , Dominik Seuß

Animals that travel together in groups display a variety of fascinating motion patterns thought to be the result of delicate local interactions among group members. Although the most informative way of investigating and interpreting…

生物物理 · 物理学 2010-10-27 Mate Nagy , Zsuzsa Akos , Dora Biro , Tamas Vicsek

Human pose estimation from single images is a challenging problem in computer vision that requires large amounts of labeled training data to be solved accurately. Unfortunately, for many human activities (\eg outdoor sports) such training…

计算机视觉与模式识别 · 计算机科学 2020-12-01 Bastian Wandt , Marco Rudolph , Petrissa Zell , Helge Rhodin , Bodo Rosenhahn

Ecological and conservation studies monitoring bird communities typically rely on species classification based on bird vocalizations. Historically, this has been based on expert volunteers going into the field and making lists of the bird…

统计方法学 · 统计学 2026-05-29 Haoxuan Wang , Patrik Lauha , David B. Dunson

We present a new method for reconstructing the appearance properties of human faces from a lightweight capture procedure in an unconstrained environment. Our method recovers the surface geometry, diffuse albedo, specular intensity and…

计算机视觉与模式识别 · 计算机科学 2025-09-03 Yingyan Xu , Kate Gadola , Prashanth Chandran , Sebastian Weiss , Markus Gross , Gaspard Zoss , Derek Bradley

We present the first method for real-time full body capture that estimates shape and motion of body and hands together with a dynamic 3D face model from a single color image. Our approach uses a new neural network architecture that exploits…

计算机视觉与模式识别 · 计算机科学 2021-04-16 Yuxiao Zhou , Marc Habermann , Ikhsanul Habibie , Ayush Tewari , Christian Theobalt , Feng Xu

1. Animal movement patterns contribute to our understanding of variation in breeding success and survival of individuals, and the implications for population dynamics. 2. Over time, sensor technology for measuring movement patterns has…

定量方法 · 定量生物学 2018-01-11 Leah R. Johnson , Philipp H. Boersch-Supan , Richard A. Phillips , Sadie J. Ryan

We propose a technique for learning single-view 3D object pose estimation models by utilizing a new source of data -- in-the-wild videos where objects turn. Such videos are prevalent in practice (e.g., cars in roundabouts, airplanes near…

计算机视觉与模式识别 · 计算机科学 2022-12-14 Zezhou Cheng , Matheus Gadelha , Subhransu Maji

Existing works on motion deblurring either ignore the effects of depth-dependent blur or work with the assumption of a multi-layered scene wherein each layer is modeled in the form of fronto-parallel plane. In this work, we consider the…

计算机视觉与模式识别 · 计算机科学 2022-02-08 Kuldeep Purohit , Subeesh Vasu , M. Purnachandra Rao , A. N. Rajagopalan