中文
相关论文

相关论文: NILUT: Conditional Neural Implicit 3D Lookup Table…

200 篇论文

Diffusion models have shown great promise for image generation, beating GANs in terms of generation diversity, with comparable image quality. However, their application to 3D shapes has been limited to point or voxel representations that…

计算机视觉与模式识别 · 计算机科学 2022-12-16 Gimin Nam , Mariem Khlifi , Andrew Rodriguez , Alberto Tono , Linqi Zhou , Paul Guerrero

Recently, deep learning-based pan-sharpening algorithms have achieved notable advancements over traditional methods. However, deep learning-based methods incur substantial computational overhead during inference, especially with large…

计算机视觉与模式识别 · 计算机科学 2025-12-04 Zhongnan Cai , Yingying Wang , Hui Zheng , Panwang Pan , ZiXu Lin , Ge Meng , Chenxin Li , Chunming He , Jiaxin Xie , Yunlong Lin , Junbin Lu , Yue Huang , Xinghao Ding

Training and testing supervised object detection models require a large collection of images with ground truth labels. Labels define object classes in the image, as well as their locations, shape, and possibly other information such as…

计算机视觉与模式识别 · 计算机科学 2022-07-26 John Rachwan , Charbel Zalaket

We propose a novel unsupervised backlit image enhancement method, abbreviated as CLIP-LIT, by exploring the potential of Contrastive Language-Image Pre-Training (CLIP) for pixel-level image enhancement. We show that the open-world CLIP…

计算机视觉与模式识别 · 计算机科学 2023-10-02 Zhexin Liang , Chongyi Li , Shangchen Zhou , Ruicheng Feng , Chen Change Loy

In this paper, we introduce a new approach for high-quality multi-exposure image fusion (MEF). We show that the fusion weights of an exposure can be encoded into a 1D lookup table (LUT), which takes pixel intensity value as input and…

计算机视觉与模式识别 · 计算机科学 2023-09-22 Ting Jiang , Chuan Wang , Xinpeng Li , Ru Li , Haoqiang Fan , Shuaicheng Liu

Video portraits relighting is critical in user-facing human photography, especially for immersive VR/AR experience. Recent advances still fail to recover consistent relit result under dynamic illuminations from monocular RGB stream,…

计算机视觉与模式识别 · 计算机科学 2021-04-02 Longwen Zhang , Qixuan Zhang , Minye Wu , Jingyi Yu , Lan Xu

Implicit representations of 3D objects have recently achieved impressive results on learning-based 3D reconstruction tasks. While existing works use simple texture models to represent object appearance, photo-realistic image synthesis…

计算机视觉与模式识别 · 计算机科学 2020-03-30 Michael Oechsle , Michael Niemeyer , Lars Mescheder , Thilo Strauss , Andreas Geiger

Understanding 3D medical image volumes is a critical task in the medical domain. However, existing 3D convolution and transformer-based methods have limited semantic understanding of an image volume and also need a large set of volumes for…

计算机视觉与模式识别 · 计算机科学 2024-03-11 Qiuhui Chen , Huping Ye , Yi Hong

Implicit neural representation has recently shown a promising ability in representing images with arbitrary resolutions. In this paper, we present a Local Implicit Transformer (LIT), which integrates the attention mechanism and frequency…

计算机视觉与模式识别 · 计算机科学 2023-03-30 Hao-Wei Chen , Yu-Syuan Xu , Min-Fong Hong , Yi-Min Tsai , Hsien-Kai Kuo , Chun-Yi Lee

Existing Vision Language Models (VLMs) often struggle to preserve logic, entity identity, and artistic style during extended, interleaved image-text interactions. We identify this limitation as "Multimodal Context Drift", which stems from…

计算机视觉与模式识别 · 计算机科学 2026-01-01 Zeteng Lin , Xingxing Li , Wen You , Xiaoyang Li , Zehan Lu , Yujun Cai , Jing Tang

Neural implicit surfaces have become an important technique for multi-view 3D reconstruction but their accuracy remains limited. In this paper, we argue that this comes from the difficulty to learn and render high frequency textures with…

计算机视觉与模式识别 · 计算机科学 2022-05-10 François Darmon , Bénédicte Bascle , Jean-Clément Devaux , Pascal Monasse , Mathieu Aubry

This paper presents a collaborative implicit neural simultaneous localization and mapping (SLAM) system with RGB-D image sequences, which consists of complete front-end and back-end modules including odometry, loop detection, sub-map…

计算机视觉与模式识别 · 计算机科学 2023-11-15 Jiarui Hu , Mao Mao , Hujun Bao , Guofeng Zhang , Zhaopeng Cui

We present SILT, a Self-supervised Implicit Lighting Transfer method. Unlike previous research on scene relighting, we do not seek to apply arbitrary new lighting configurations to a given scene. Instead, we wish to transfer the lighting…

计算机视觉与模式识别 · 计算机科学 2022-03-16 Nikolina Kubiak , Armin Mustafa , Graeme Phillipson , Stephen Jolly , Simon Hadfield

FPGAs have distinct advantages as a technology for deploying deep neural networks (DNNs) at the edge. Lookup Table (LUT) based networks, where neurons are directly modeled using LUTs, help maximize this promise of offering ultra-low latency…

机器学习 · 计算机科学 2024-09-17 Binglei Lou , Richard Rademacher , David Boland , Philip H. W. Leong

Efficient neural networks (NNs) leveraging lookup tables (LUTs) have demonstrated significant potential for emerging AI applications, particularly when deployed on field-programmable gate arrays (FPGAs) for edge computing. These…

机器学习 · 计算机科学 2025-04-02 Marta Andronic , George A. Constantinides

In this paper, we present Neural Adaptive Tomography (NeAT), the first adaptive, hierarchical neural rendering pipeline for multi-view inverse rendering. Through a combination of neural features with an adaptive explicit representation, we…

计算机视觉与模式识别 · 计算机科学 2022-02-07 Darius Rückert , Yuanhao Wang , Rui Li , Ramzi Idoughi , Wolfgang Heidrich

Tone mapping aims to convert high dynamic range (HDR) images to low dynamic range (LDR) representations, a critical task in the camera imaging pipeline. In recent years, 3-Dimensional LookUp Table (3D LUT) based methods have gained…

计算机视觉与模式识别 · 计算机科学 2024-01-04 Feng Zhang , Ming Tian , Zhiqiang Li , Bin Xu , Qingbo Lu , Changxin Gao , Nong Sang

This work has been submitted to the IEEE for possible publication. Copyright may be transferred without notice, after which this version may no longer be accessible. Computational imaging, especially non-line-of-sight (NLOS) imaging, the…

计算机视觉与模式识别 · 计算机科学 2026-04-01 Lianfang Wang , Kuilin Qin , Xueying Liu , Huibin Chang , Yong Wang , Yuping Duan

Identifying changes in a pair of 3D aerial LiDAR point clouds, obtained during two distinct time periods over the same geographic region presents a significant challenge due to the disparities in spatial coverage and the presence of noise…

计算机视觉与模式识别 · 计算机科学 2023-08-31 Peter Naylor , Diego Di Carlo , Arianna Traviglia , Makoto Yamada , Marco Fiorucci

MixUp is a computer vision data augmentation technique that uses convex interpolations of input data and their labels to enhance model generalization during training. However, the application of MixUp to the natural language understanding…

计算与语言 · 计算机科学 2021-02-24 Wancong Zhang , Ieshan Vaidya