中文
相关论文

相关论文: PhyCAGE: Physically Plausible Compositional 3D Ass…

200 篇论文

Envisioning physically plausible outcomes from a single image requires a deep understanding of the world's dynamics. To address this, we introduce PhysGen3D, a novel framework that transforms a single image into an amodal, camera-centric,…

计算机视觉与模式识别 · 计算机科学 2025-03-27 Boyuan Chen , Hanxiao Jiang , Shaowei Liu , Saurabh Gupta , Yunzhu Li , Hao Zhao , Shenlong Wang

Recently, 3D generative models have made impressive progress, enabling the generation of almost arbitrary 3D assets from text or image inputs. However, these approaches generate objects in isolation without any consideration for the scene…

计算机视觉与模式识别 · 计算机科学 2025-05-02 Jinghao Zhou , Tomas Jakab , Philip Torr , Christian Rupprecht

We present a novel approach for enhancing the resolution and geometric fidelity of 3D Gaussian Splatting (3DGS) beyond native training resolution. Current 3DGS methods are fundamentally limited by their input resolution, producing…

图形学 · 计算机科学 2025-06-10 Shuja Khalid , Mohamed Ibrahim , Yang Liu

Recent advances in 3D content creation mostly leverage optimization-based 3D generation via score distillation sampling (SDS). Though promising results have been exhibited, these methods often suffer from slow per-sample optimization,…

计算机视觉与模式识别 · 计算机科学 2024-04-01 Jiaxiang Tang , Jiawei Ren , Hang Zhou , Ziwei Liu , Gang Zeng

Despite the ability of text-to-image models to generate high-quality, realistic, and diverse images, they face challenges in compositional generation, often struggling to accurately represent details specified in the input prompt. A…

计算机视觉与模式识别 · 计算机科学 2025-07-01 Parham Rezaei , Arash Marioriyad , Mahdieh Soleymani Baghshah , Mohammad Hossein Rohban

Automatic text-to-3D generation that combines Score Distillation Sampling (SDS) with the optimization of volume rendering has achieved remarkable progress in synthesizing realistic 3D objects. Yet most existing text-to-3D methods by SDS and…

计算机视觉与模式识别 · 计算机科学 2024-04-03 Zilong Chen , Feng Wang , Yikai Wang , Huaping Liu

4D content generation aims to create dynamically evolving 3D content that responds to specific input objects such as images or 3D representations. Current approaches typically incorporate physical priors to animate 3D representations, but…

计算机视觉与模式识别 · 计算机科学 2025-11-04 Jiajing Lin , Zhenzhong Wang , Dejun Xu , Shu Jiang , YunPeng Gong , Min Jiang

Gaussian Splatting (GS) has recently emerged as an efficient representation for rendering 3D scenes from 2D images and has been extended to images, videos, and dynamic 4D content. However, applying style transfer to GS-based…

计算机视觉与模式识别 · 计算机科学 2025-10-27 Kornel Howil , Joanna Waczyńska , Piotr Borycki , Tadeusz Dziarmaga , Marcin Mazur , Przemysław Spurek

Computer vision-based technologies significantly enhance surgical automation by advancing tool tracking, detection, and localization. However, Current data-driven approaches are data-voracious, requiring large, high-quality labeled image…

计算机视觉与模式识别 · 计算机科学 2025-08-12 Tianle Zeng , Junlei Hu , Gerardo Loza Galindo , Sharib Ali , Duygu Sarikaya , Pietro Valdastri , Dominic Jones

This paper studies the problem of estimating physical properties (system identification) through visual observations. To facilitate geometry-aware guidance in physical property estimation, we introduce a novel hybrid framework that…

计算机视觉与模式识别 · 计算机科学 2024-11-01 Junhao Cai , Yuji Yang , Weihao Yuan , Yisheng He , Zilong Dong , Liefeng Bo , Hui Cheng , Qifeng Chen

3D Gaussian splatting (3DGS) has recently emerged as an alternative representation that leverages a 3D Gaussian-based representation and introduces an approximated volumetric rendering, achieving very fast rendering speed and promising…

计算机视觉与模式识别 · 计算机科学 2024-08-08 Joo Chan Lee , Daniel Rho , Xiangyu Sun , Jong Hwan Ko , Eunbyung Park

Animatable 3D reconstruction has significant applications across various fields, primarily relying on artists' handcraft creation. Recently, some studies have successfully constructed animatable 3D models from monocular videos. However,…

计算机视觉与模式识别 · 计算机科学 2024-03-19 Tingyang Zhang , Qingzhe Gao , Weiyu Li , Libin Liu , Baoquan Chen

3D Gaussian Splatting (3DGS) has emerged as a prominent 3D representation for high-fidelity and real-time rendering. Prior work has coupled physics simulation with Gaussians, but predominantly targets soft, deformable materials, leaving…

计算机视觉与模式识别 · 计算机科学 2026-01-15 Bei Huang , Yixin Chen , Ruijie Lu , Gang Zeng , Hongbin Zha , Yuru Pei , Siyuan Huang

Gaussian Splatting (GS) has gained attention as a fast and effective method for novel view synthesis. It has also been applied to 3D reconstruction using multi-view images and can achieve fast and accurate 3D reconstruction. However, GS…

计算机视觉与模式识别 · 计算机科学 2025-05-30 Natsuki Takama , Shintaro Ito , Koichi Ito , Hwann-Tzong Chen , Takafumi Aoki

Creating and editing high-quality 3D content remains a central challenge in computer graphics. We address this challenge by introducing CompoSE, a novel method for Compositional Synthesis and Editing of 3D shapes via part-aware control. Our…

图形学 · 计算机科学 2026-05-20 Habib Slim , Shariq Farooq Bhat , Mohamed Elhoseiny , Yifan Wang , Mike Roberts

Personalizing 3D scenes from a single reference image enables intuitive user-guided editing, which requires achieving both multi-view consistency across perspectives and referential consistency with the input image. However, these goals are…

计算机视觉与模式识别 · 计算机科学 2025-12-29 Yuxuan Wang , Xuanyu Yi , Qingshan Xu , Yuan Zhou , Long Chen , Hanwang Zhang

Gaussian splatting, renowned for its exceptional rendering quality and efficiency, has emerged as a prominent technique in 3D scene representation. However, the substantial data volume of Gaussian splatting impedes its practical utility in…

计算机视觉与模式识别 · 计算机科学 2024-04-16 Xiangrui Liu , Xinju Wu , Pingping Zhang , Shiqi Wang , Zhu Li , Sam Kwong

Physical adversarial camouflage poses a severe security threat to autonomous driving systems by mapping adversarial textures onto 3D objects. Nevertheless, current methods remain brittle in complex dynamic scenarios, failing to generalize…

计算机视觉与模式识别 · 计算机科学 2026-03-30 Tianrui Lou , Siyuan Liang , Jiawei Liang , Yuze Gao , Xiaochun Cao

We propose PixelGaussian, an efficient feed-forward framework for learning generalizable 3D Gaussian reconstruction from arbitrary views. Most existing methods rely on uniform pixel-wise Gaussian representations, which learn a fixed number…

计算机视觉与模式识别 · 计算机科学 2024-10-25 Xin Fei , Wenzhao Zheng , Yueqi Duan , Wei Zhan , Masayoshi Tomizuka , Kurt Keutzer , Jiwen Lu

Physical adversarial attack methods expose the vulnerabilities of deep neural networks and pose a significant threat to safety-critical scenarios such as autonomous driving. Camouflage-based physical attack is a more promising approach…

计算机视觉与模式识别 · 计算机科学 2025-08-14 Tianrui Lou , Xiaojun Jia , Siyuan Liang , Jiawei Liang , Ming Zhang , Yanjun Xiao , Xiaochun Cao