English
Related papers

Related papers: PlaneCycle: Training-Free 2D-to-3D Lifting of Foun…

200 papers

Existing networks directly learn feature representations on 3D point clouds for shape analysis. We argue that 3D point clouds are highly redundant and hold irregular (permutation-invariant) structure, which makes it difficult to achieve…

Machine Learning · Computer Science 2020-07-21 Sameera Ramasinghe , Salman Khan , Nick Barnes , Stephen Gould

This paper shows the effectiveness of 2D backbone scaling and pretraining for pillar-based 3D object detectors. Pillar-based methods mainly employ randomly initialized 2D convolution neural network (ConvNet) for feature extraction and fail…

Computer Vision and Pattern Recognition · Computer Science 2023-11-30 Weixin Mao , Tiancai Wang , Diankun Zhang , Junjie Yan , Osamu Yoshie

Plane instance segmentation from RGB-D data is a crucial research topic for many downstream tasks. However, most existing deep-learning-based methods utilize only information within the RGB bands, neglecting the important role of the depth…

Computer Vision and Pattern Recognition · Computer Science 2024-10-23 Zhongchen Deng , Zhechen Yang , Chi Chen , Cheng Zeng , Yan Meng , Bisheng Yang

Recent advances in 2D-to-3D perception have enabled the recovery of 3D scene semantics from unposed images. However, prevailing methods often suffer from limited generalization, reliance on per-scene optimization, and semantic…

Computer Vision and Pattern Recognition · Computer Science 2026-03-27 Jie Hu , Shizun Wang , Xinchao Wang

Predictive modeling in engineering applications has long been dominated by bespoke models and small, siloed tabular datasets, limiting the applicability of large-scale learning approaches. Despite recent progress in tabular foundation…

Machine Learning · Computer Science 2026-03-06 Lyle Regenwetter , Rosen Yu , Cyril Picard , Faez Ahmed

Extracting planes from a 3D scene is useful for downstream tasks in robotics and augmented reality. In this paper we tackle the problem of estimating the planar surfaces in a scene from posed images. Our first finding is that a surprisingly…

Computer Vision and Pattern Recognition · Computer Science 2024-06-14 Jamie Watson , Filippo Aleotti , Mohamed Sayed , Zawar Qureshi , Oisin Mac Aodha , Gabriel Brostow , Michael Firman , Sara Vicente

Transfer learning is important for foundation models to adapt to downstream tasks. However, many foundation models are proprietary, so users must share their data with model owners to fine-tune the models, which is costly and raise privacy…

Computation and Language · Computer Science 2023-02-10 Guangxuan Xiao , Ji Lin , Song Han

We present FaceLift, a novel feed-forward approach for generalizable high-quality 360-degree 3D head reconstruction from a single image. Our pipeline first employs a multi-view latent diffusion model to generate consistent side and back…

Computer Vision and Pattern Recognition · Computer Science 2025-08-04 Weijie Lyu , Yi Zhou , Ming-Hsuan Yang , Zhixin Shu

Diffeomorphic deformable image registration is one of the crucial tasks in medical image analysis, which aims to find a unique transformation while preserving the topology and invertibility of the transformation. Deep convolutional neural…

Image and Video Processing · Electrical Eng. & Systems 2022-02-09 Ameneh Sheikhjafari , Michelle Noga , Kumaradevan Punithakumar , Nilanjan Ray

Reconstructing 2D freehand Ultrasound (US) frames into 3D space without using a tracker has recently seen advances with deep learning. Predicting good frame-to-frame rigid transformations is often accepted as the learning objective,…

Image and Video Processing · Electrical Eng. & Systems 2024-10-22 Qi Li , Ziyi Shen , Qianye Yang , Dean C. Barratt , Matthew J. Clarkson , Tom Vercauteren , Yipeng Hu

With the overwhelming trend of mask image modeling led by MAE, generative pre-training has shown a remarkable potential to boost the performance of fundamental models in 2D vision. However, in 3D vision, the over-reliance on…

Computer Vision and Pattern Recognition · Computer Science 2023-09-08 Ziyi Wang , Xumin Yu , Yongming Rao , Jie Zhou , Jiwen Lu

Given a designer created free-form surface in 3d space, our method computes a grid composed of elastic elements which are completely planar and straight. Only by fixing the ends of the planar elements to appropriate locations, the 2d grid…

Graphics · Computer Science 2021-11-18 Stefan Pillwein , Przemyslaw Musialski

We introduce DreamCraft3D++, an extension of DreamCraft3D that enables efficient high-quality generation of complex 3D assets. DreamCraft3D++ inherits the multi-stage generation process of DreamCraft3D, but replaces the time-consuming…

Computer Vision and Pattern Recognition · Computer Science 2024-10-18 Jingxiang Sun , Cheng Peng , Ruizhi Shao , Yuan-Chen Guo , Xiaochen Zhao , Yangguang Li , Yanpei Cao , Bo Zhang , Yebin Liu

Multi-view 3D reconstruction has remained an essential yet challenging problem in the field of computer vision. While DUSt3R and its successors have achieved breakthroughs in 3D reconstruction from unposed images, these methods exhibit…

Image and Video Processing · Electrical Eng. & Systems 2025-09-16 Sidun Liu , Wenyu Li , Peng Qiao , Yong Dou

We tackle the problem of monocular 3D reconstruction of articulated objects like humans and animals. We contribute DensePose 3D, a method that can learn such reconstructions in a weakly supervised fashion from 2D image annotations only.…

Computer Vision and Pattern Recognition · Computer Science 2021-09-02 Roman Shapovalov , David Novotny , Benjamin Graham , Patrick Labatut , Andrea Vedaldi

Recent feed-forward 3D reconstruction transformers have scaled to over a billion parameters, following the broader trend of increasing model capacity in computer vision. Yet emerging evidence suggests that contiguous transformer layers…

Robotic grasping is an essential and fundamental task and has been studied extensively over the past several decades. Traditional work analyzes physical models of the objects and computes force-closure grasps. Such methods require…

Robotics · Computer Science 2023-05-25 Yuwei Wu , Weixiao Liu , Zhiyang Liu , Gregory S. Chirikjian

Robots operating in human environments must be able to rearrange objects into semantically-meaningful configurations, even if these objects are previously unseen. In this work, we focus on the problem of building physically-valid structures…

Robotics · Computer Science 2023-04-26 Weiyu Liu , Yilun Du , Tucker Hermans , Sonia Chernova , Chris Paxton

Learning to manipulate cloth is both a paradigmatic problem for robotic research and a problem of immediate relevance to a variety of applications ranging from assistive care to the service industry. The complex physics of the deformable…

Robotics · Computer Science 2026-02-19 Jack Rome , Stephen James , Subramanian Ramamoorthy

We present Farm3D, a method for learning category-specific 3D reconstructors for articulated objects, relying solely on "free" virtual supervision from a pre-trained 2D diffusion-based image generator. Recent approaches can learn a…

Computer Vision and Pattern Recognition · Computer Science 2024-05-15 Tomas Jakab , Ruining Li , Shangzhe Wu , Christian Rupprecht , Andrea Vedaldi
‹ Prev 1 8 9 10 Next ›