中文
相关论文

相关论文: DM-CFO: A Diffusion Model for Compositional 3D Too…

200 篇论文

Recent advances in generative modeling, namely Diffusion models, have revolutionized generative modeling, enabling high-quality image generation tailored to user needs. This paper proposes a framework for the generative design of structural…

Video composition is the core task of video editing. Although image composition based on diffusion models has been highly successful, it is not straightforward to extend the achievement to video object composition tasks, which not only…

计算机视觉与模式识别 · 计算机科学 2024-06-25 Wei Wang , Yaosen Chen , Yuegen Liu , Qi Yuan , Shubin Yang , Yanru Zhang

This Letter introduces an approach for precisely designing surface friction properties using a conditional generative machine learning model, specifically a diffusion denoising probabilistic model (DDPM). We created a dataset of synthetic…

Recent advances in talking face generation have significantly improved facial animation synthesis. However, existing approaches face fundamental limitations: 3DMM-based methods maintain temporal consistency but lack fine-grained regional…

计算机视觉与模式识别 · 计算机科学 2025-03-26 Kangwei Liu , Junwu Liu , Yun Cao , Jinlin Guo , Xiaowei Yi

In aerodynamic shape optimization, the convergence and computational cost are greatly affected by the representation capacity and compactness of the design space. Previous research has demonstrated that using a deep generative model to…

机器学习 · 计算机科学 2021-01-11 Wei Chen , Arun Ramamurthy

We present DC-Gaussian, a new method for generating novel views from in-vehicle dash cam videos. While neural rendering techniques have made significant strides in driving scenarios, existing methods are primarily designed for videos…

计算机视觉与模式识别 · 计算机科学 2024-11-07 Linhan Wang , Kai Cheng , Shuo Lei , Shengkun Wang , Wei Yin , Chenyang Lei , Xiaoxiao Long , Chang-Tien Lu

A single-pass driving clip frequently results in incomplete scanning of the road structure, making reconstructed scene expanding a critical requirement for sensor simulators to effectively regress driving actions. Although contemporary 3D…

计算机视觉与模式识别 · 计算机科学 2025-07-28 Sicong Du , Jiarun Liu , Qifeng Chen , Hao-Xiang Chen , Tai-Jiang Mu , Sheng Yang

In recent years, the denoising diffusion model has achieved remarkable success in image segmentation modeling. With its powerful nonlinear modeling capabilities and superior generalization performance, denoising diffusion models have…

计算机视觉与模式识别 · 计算机科学 2024-08-06 Weiping Ding , Sheng Geng , Haipeng Wang , Jiashuang Huang , Tianyi Zhou

Recent advancements in Text-to-3D generation have yielded remarkable progress, particularly through methods that rely on Score Distillation Sampling (SDS). While SDS exhibits the capability to create impressive 3D assets, it is hindered by…

机器学习 · 计算机科学 2024-07-30 Runjie Yan , Kailu Wu , Kaisheng Ma

Animating virtual avatars to make co-speech gestures facilitates various applications in human-machine interaction. The existing methods mainly rely on generative adversarial networks (GANs), which typically suffer from notorious mode…

计算机视觉与模式识别 · 计算机科学 2023-03-21 Lingting Zhu , Xian Liu , Xuanyu Liu , Rui Qian , Ziwei Liu , Lequan Yu

Diffusion models are powerful generative models that map noise to data using stochastic processes. However, for many applications such as image editing, the model input comes from a distribution that is not random noise. As such, diffusion…

计算机视觉与模式识别 · 计算机科学 2023-12-06 Linqi Zhou , Aaron Lou , Samar Khanna , Stefano Ermon

We develop a generalized 3D shape generation prior model, tailored for multiple 3D tasks including unconditional shape generation, point cloud completion, and cross-modality shape generation, etc. On one hand, to precisely capture local…

计算机视觉与模式识别 · 计算机科学 2023-03-21 Yuhan Li , Yishun Dou , Xuanhong Chen , Bingbing Ni , Yilin Sun , Yutian Liu , Fuzhen Wang

Mammography is the most commonly used imaging modality for breast cancer screening, driving an increasing demand for deep-learning techniques to support large-scale analysis. However, the development of accurate and robust methods is often…

图像与视频处理 · 电气工程与系统科学 2025-07-28 Xin Li , Kaixiang Yang , Qiang Li , Zhiwei Wang

While diffusion models have demonstrated remarkable progress in 2D image generation and editing, extending these capabilities to 3D editing remains challenging, particularly in maintaining multi-view consistency. Classical approaches…

计算机视觉与模式识别 · 计算机科学 2025-08-05 Yufeng Chi , Huimin Ma , Kafeng Wang , Jianmin Li

Autonomous driving needs fast, scalable 4D reconstruction and re-simulation for training and evaluation, yet most methods for dynamic driving scenes still rely on per-scene optimization, known camera calibration, or short frame windows,…

计算机视觉与模式识别 · 计算机科学 2025-12-03 Xiaoxue Chen , Ziyi Xiong , Yuantao Chen , Gen Li , Nan Wang , Hongcheng Luo , Long Chen , Haiyang Sun , Bing Wang , Guang Chen , Hangjun Ye , Hongyang Li , Ya-Qin Zhang , Hao Zhao

3D semantic occupancy prediction is essential for achieving safe, reliable autonomous driving and robotic navigation. Compared to camera-only perception systems, multi-modal pipelines, especially LiDAR-camera fusion methods, can produce…

计算机视觉与模式识别 · 计算机科学 2026-02-17 Lingjun Zhao , Sizhe Wei , James Hays , Lu Gan

Diffusion models (DM) can gradually learn to remove noise, which have been widely used in artificial intelligence generated content (AIGC) in recent years. The property of DM for removing noise leads us to wonder whether DM can be applied…

信息论 · 计算机科学 2023-05-17 Tong Wu , Zhiyong Chen , Dazhi He , Liang Qian , Yin Xu , Meixia Tao , Wenjun Zhang

The assembly of virus capsids from free coat proteins proceeds by a complicated cascade of association and dissociation steps, the great majority of which cannot be directly experimentally observed. This has made capsid assembly a rich…

定量方法 · 定量生物学 2015-07-09 Lu Xie , Gregory R. Smith , Russell Schwartz

In orthodontic treatment, particularly within telemedicine contexts, observing patients' dental occlusion from multiple viewpoints facilitates timely clinical decision-making. Recent advances in 3D Gaussian Splatting (3DGS) have shown…

计算机视觉与模式识别 · 计算机科学 2025-11-06 Yiyi Miao , Taoyu Wu , Tong Chen , Sihao Li , Ji Jiang , Youpeng Yang , Angelos Stefanidis , Limin Yu , Jionglong Su

Recent 3D reconstruction methods achieve impressive results with dense multi-view imagery but struggle when only a few views are available. Various approaches, including regularization techniques, semantic priors, and geometric constraints,…

计算机视觉与模式识别 · 计算机科学 2026-04-07 Yi-Chuan Huang , Hao-Jen Chien , Chin-Yang Lin , Ying-Huan Chen , Yu-Lun Liu
‹ 上一页 1 8 9 10 下一页 ›