English
Related papers

Related papers: MatE: Material Extraction from Single-Image via Ge…

200 papers

Human matting is a foundation task in image and video processing, where human foreground pixels are extracted from the input. Prior works either improve the accuracy by additional guidance or improve the temporal consistency of a single…

Computer Vision and Pattern Recognition · Computer Science 2024-04-25 Chuong Huynh , Seoung Wug Oh , Abhinav Shrivastava , Joon-Young Lee

In recent years, the underwater image formation model has found extensive use in the generation of synthetic underwater data. Although many approaches focus on scenes primarily affected by discoloration, they often overlook the model's…

Computer Vision and Pattern Recognition · Computer Science 2025-09-22 Vasiliki Ismiroglou , Malte Pedersen , Stefan H. Bengtson , Andreas Aakerberg , Thomas B. Moeslund

In the portrait matting, the goal is to predict an alpha matte that identifies the effect of each pixel on the foreground subject. Traditional approaches and most of the existing works utilized an additional input, e.g., trimap, background…

Computer Vision and Pattern Recognition · Computer Science 2022-04-26 Dogucan Yaman , Hazım Kemal Ekenel , Alexander Waibel

There has been significant progress in Masked Image Modeling (MIM). Existing MIM methods can be broadly categorized into two groups based on the reconstruction target: pixel-based and tokenizer-based approaches. The former offers a simpler…

Computer Vision and Pattern Recognition · Computer Science 2023-08-02 Yuan Liu , Songyang Zhang , Jiacheng Chen , Zhaohui Yu , Kai Chen , Dahua Lin

High-quality material generation is key for virtual environment authoring and inverse rendering. We propose MaterialPicker, a multi-modal material generator leveraging a Diffusion Transformer (DiT) architecture, improving and simplifying…

Computer Vision and Pattern Recognition · Computer Science 2025-07-29 Xiaohe Ma , Valentin Deschaintre , Miloš Hašan , Fujun Luan , Kun Zhou , Hongzhi Wu , Yiwei Hu

The reconstruction of indoor scenes from multi-view RGB images is challenging due to the coexistence of flat and texture-less regions alongside delicate and fine-grained regions. Recent methods leverage neural radiance fields aided by…

Computer Vision and Pattern Recognition · Computer Science 2024-08-14 Sheng Ye , Yubin Hu , Matthieu Lin , Yu-Hui Wen , Wang Zhao , Yong-Jin Liu , Wenping Wang

Text-to-3D generation by distilling pretrained large-scale text-to-image diffusion models has shown great promise but still suffers from inconsistent 3D geometric structures (Janus problems) and severe artifacts. The aforementioned problems…

Computer Vision and Pattern Recognition · Computer Science 2023-12-05 Baorui Ma , Haoge Deng , Junsheng Zhou , Yu-Shen Liu , Tiejun Huang , Xinlong Wang

Despite remarkable advancements in video depth estimation, existing methods exhibit inherent limitations in achieving geometric fidelity through the affine-invariant predictions, limiting their applicability in reconstruction and other…

Graphics · Computer Science 2025-04-02 Tian-Xing Xu , Xiangjun Gao , Wenbo Hu , Xiaoyu Li , Song-Hai Zhang , Ying Shan

Image matting is a long-standing problem in computer graphics and vision, mostly identified as the accurate estimation of the foreground in input images. We argue that the foreground objects can be represented by different-level…

Computer Vision and Pattern Recognition · Computer Science 2021-03-04 Yu Qiao , Yuhao Liu , Qiang Zhu , Xin Yang , Yuxin Wang , Qiang Zhang , Xiaopeng Wei

Video diffusion models have recently enabled high-quality video generation with ViT-based architectures, but remain computationally intensive because generation requires attention computation over long spatiotemporal sequences. Token…

Computer Vision and Pattern Recognition · Computer Science 2026-05-22 Sheng Li , Yang Sui , Junhao Ran , Bo Yuan , Yue Dai , Xulong Tang

We propose DiMeR, a novel geometry-texture disentangled feed-forward model with 3D supervision for sparse-view mesh reconstruction. Existing methods confront two persistent obstacles: (i) textures can conceal geometric errors, i.e.,…

Computer Vision and Pattern Recognition · Computer Science 2025-05-27 Lutao Jiang , Jiantao Lin , Kanghao Chen , Wenhang Ge , Xin Yang , Yifan Jiang , Yuanhuiyi Lyu , Xu Zheng , Yinchuan Li , Yingcong Chen

Variational Autoencoders (VAE) are probabilistic deep generative models underpinned by elegant theory, stable training processes, and meaningful manifold representations. However, they produce blurry images due to a lack of explicit…

Computer Vision and Pattern Recognition · Computer Science 2019-11-15 Prashnna K Gyawali , Rudra Saha , Linwei Wang , VSR Veeravasarapu , Maneesh Singh

Image matting aims to obtain an alpha matte that separates foreground objects from the background accurately. Recently, trimap-free matting has been well studied because it requires only the original image without any extra input. Such…

Computer Vision and Pattern Recognition · Computer Science 2024-05-29 Leo Shan Wenzhang Zhou Grace Zhao

Inferring full-body poses from Head Mounted Devices, which capture only 3-joint observations from the head and wrists, is a challenging task with wide AR/VR applications. Previous attempts focus on learning one-stage motion mapping and thus…

Computer Vision and Pattern Recognition · Computer Science 2025-05-13 Fangyu Du , Yang Yang , Xuehao Gao , Hongye Hou

Reconstructing object geometry and material from multiple views typically requires optimization. Differentiable path tracing is an appealing framework as it can reproduce complex appearance effects. However, it is difficult to use due to…

Computer Vision and Pattern Recognition · Computer Science 2020-12-09 Purvi Goel , Loudon Cohen , James Guesman , Vikas Thamizharasan , James Tompkin , Daniel Ritchie

We present PIXLRelight, a feed-forward approach for physically controllable single-image relighting. Existing methods either provide limited lighting control (e.g. through text or environment maps), accumulate errors when chaining inverse…

Computer Vision and Pattern Recognition · Computer Science 2026-05-19 Miguel Farinha , Ronald Clark

We present a novel approach to leverage prior knowledge encapsulated in pre-trained text-to-image diffusion models for blind super-resolution (SR). Specifically, by employing our time-aware encoder, we can achieve promising restoration…

Computer Vision and Pattern Recognition · Computer Science 2024-07-01 Jianyi Wang , Zongsheng Yue , Shangchen Zhou , Kelvin C. K. Chan , Chen Change Loy

Real-world image super-resolution (RealSR) aims to enhance the visual quality of in-the-wild images, such as those captured by mobile phones. While existing methods leveraging large generative models demonstrate impressive results, the high…

Computer Vision and Pattern Recognition · Computer Science 2025-10-06 Haoze Sun , Linfeng Jiang , Fan Li , Renjing Pei , Zhixin Wang , Yong Guo , Jiaqi Xu , Haoyu Chen , Jin Han , Fenglong Song , Yujiu Yang , Wenbo Li

This paper proposes a deep learning based method for colored transparent object matting from a single image. Existing approaches for transparent object matting often require multiple images and long processing times, which greatly hinder…

Computer Vision and Pattern Recognition · Computer Science 2019-10-08 Jamal Ahmed Rahim , Kwan-Yee Kenneth Wong

We introduce an inversion based method, denoted as IMAge-Guided model INvErsion (IMAGINE), to generate high-quality and diverse images from only a single training sample. We leverage the knowledge of image semantics from a pre-trained…

Computer Vision and Pattern Recognition · Computer Science 2021-04-14 Pei Wang , Yijun Li , Krishna Kumar Singh , Jingwan Lu , Nuno Vasconcelos
‹ Prev 1 4 5 6 7 8 10 Next ›