中文
相关论文

相关论文: 3-D Rigid Models from Partial Views - Global Facto…

200 篇论文

We present a new approach to rigid-body motion segmentation from two views. We use a previously developed nonlinear embedding of two-view point correspondences into a 9-dimensional space and identify the different motions by segmenting…

计算机视觉与模式识别 · 计算机科学 2014-06-25 Bryan Poling , Gilad Lerman

Despite much recent progress in video-based person re-identification (re-ID), the current state-of-the-art still suffers from common real-world challenges such as appearance similarity among various people, occlusions, and frame…

计算机视觉与模式识别 · 计算机科学 2021-08-17 Abhishek Aich , Meng Zheng , Srikrishna Karanam , Terrence Chen , Amit K. Roy-Chowdhury , Ziyan Wu

Face frontalization consists of synthesizing a frontally-viewed face from an arbitrarily-viewed one. The main contribution of this paper is a robust face alignment method that enables pixel-to-pixel warping. The method simultaneously…

计算机视觉与模式识别 · 计算机科学 2021-03-11 Zhiqi Kang , Mostafa Sadeghi , Radu Horaud

Active stereo vision is important in reconstructing objects without obvious textures. However, it is still very challenging to extract and match the projected patterns from two camera views automatically and robustly. In this paper, we…

计算机视觉与模式识别 · 计算机科学 2020-04-01 Yongcan Shuang , Zhenzhou Wang

Understanding and predicting dynamics of the physical world can enhance a robot's ability to plan and interact effectively in complex environments. While recent video generation models have shown strong potential in modeling dynamic scenes,…

计算机视觉与模式识别 · 计算机科学 2026-05-19 Zeyi Liu , Shuang Li , Eric Cousineau , Siyuan Feng , Benjamin Burchfiel , Shuran Song

When viewing a 3D Gaussian Splatting (3DGS) model from camera positions significantly outside the training data distribution, substantial visual noise commonly occurs. These artifacts result from the lack of training data in these…

计算机视觉与模式识别 · 计算机科学 2025-10-24 Damian Bowness , Charalambos Poullis

In recent years, there has been an explosion of 2D vision models for numerous tasks such as semantic segmentation, style transfer or scene editing, enabled by large-scale 2D image datasets. At the same time, there has been renewed interest…

计算机视觉与模式识别 · 计算机科学 2024-03-29 Mukund Varma T , Peihao Wang , Zhiwen Fan , Zhangyang Wang , Hao Su , Ravi Ramamoorthi

Recent advances in visual generative models have highlighted the promise of learning generative world models. However, most existing approaches frame world modeling as novel-view synthesis or future-frame prediction, emphasizing visual…

计算机视觉与模式识别 · 计算机科学 2026-05-13 Yifan Yin , Zehao Wen , Jieneng Chen , Zehan Zheng , Nanru Dai , Haojun Shi , Suyu Ye , Aydan Huang , Zheyuan Zhang , Alan Yuille , Jianwen Xie , Ayush Tewari , Tianmin Shu

We present a novel method to infer, in closed-form, a general 3D spatial occupancy and orientation of a collection of rigid objects given 2D image detections from a sequence of images. In particular, starting from 2D ellipses fitted to…

计算机视觉与模式识别 · 计算机科学 2015-07-21 Cosimo Rubino , Marco Crocco , Alessandro Perina , Vittorio Murino , Alessio Del Bue

LiDAR-based place recognition serves as a crucial enabler for long-term autonomy in robotics and autonomous driving systems. Yet, prevailing methodologies relying on handcrafted feature extraction face dual challenges: (1) Inconsistent…

计算机视觉与模式识别 · 计算机科学 2025-08-28 Xiaohui Jiang , Haijiang Zhu , Chade Li , Fulin Tang , Ning An

Video classification researches that have recently attracted attention are the fields of temporal modeling and 3D efficient architecture. However, the temporal modeling methods are not efficient or the 3D efficient architecture is less…

计算机视觉与模式识别 · 计算机科学 2021-04-23 Youngwan Lee , Hyung-Il Kim , Kimin Yun , Jinyoung Moon

Multiview 3D face modeling has attracted increasing attention recently and has become one of the potential avenues in future video systems. We aim to make more reliable and robust automatic feature extraction and natural 3D feature…

计算机视觉与模式识别 · 计算机科学 2009-12-04 Alireza Ghahari , Reza Aghaeizadeh Zoroofi

We propose PartField, a feedforward approach for learning part-based 3D features, which captures the general concept of parts and their hierarchy without relying on predefined templates or text-based names, and can be applied to open-world…

计算机视觉与模式识别 · 计算机科学 2025-04-16 Minghua Liu , Mikaela Angelina Uy , Donglai Xiang , Hao Su , Sanja Fidler , Nicholas Sharp , Jun Gao

Person re-identification (ReID) is aimed at identifying the same person across videos captured from different cameras. In the view that networks extracting global features using ordinary network architectures are difficult to extract local…

计算机视觉与模式识别 · 计算机科学 2018-11-21 Pengyuan Ren , Jianmin Li

Learning to understand dynamic 3D scenes from imagery is crucial for applications ranging from robotics to scene reconstruction. Yet, unlike other problems where large-scale supervised training has enabled rapid progress, directly…

计算机视觉与模式识别 · 计算机科学 2025-05-01 Linyi Jin , Richard Tucker , Zhengqi Li , David Fouhey , Noah Snavely , Aleksander Holynski

Visual repetition is ubiquitous in our world. It appears in human activity (sports, cooking), animal behavior (a bee's waggle dance), natural phenomena (leaves in the wind) and in urban environments (flashing lights). Estimating visual…

计算机视觉与模式识别 · 计算机科学 2018-06-20 Tom F. H. Runia , Cees G. M. Snoek , Arnold W. M. Smeulders

Visual object recognition systems need to generalize from a set of 2D training views to novel views. The question of how the human visual system can generalize to novel views has been studied and modeled in psychology, computer vision, and…

计算机视觉与模式识别 · 计算机科学 2023-04-20 Shoaib Ahmed Siddiqui , David Krueger , Thomas Breuel

Recent camera-based 3D object detection methods have introduced sequential frames to improve the detection performance hoping that multiple frames would mitigate the large depth estimation error. Despite improved detection performance,…

计算机视觉与模式识别 · 计算机科学 2023-09-06 Sanmin Kim , Youngseok Kim , In-Jae Lee , Dongsuk Kum

Compositional video generation aims to synthesize multiple instances with diverse appearance and motion. However, current approaches mainly focus on binding semantics, neglecting to understand diverse motion categories specified in prompts.…

计算机视觉与模式识别 · 计算机科学 2026-03-31 Zixuan Wang , Ziqin Zhou , Feng Chen , Duo Peng , Yixin Hu , Changsheng Li , Yinjie Lei

Appearance-based gaze estimation has been actively studied in recent years. However, its generalization performance for unseen head poses is still a significant limitation for existing methods. This work proposes a generalizable multi-view…

计算机视觉与模式识别 · 计算机科学 2023-11-16 Yoichiro Hisadome , Tianyi Wu , Jiawei Qin , Yusuke Sugano