English
Related papers

Related papers: Predicting Chroma from Luma with Frequency Domain …

200 papers

We explore frame-level audio feature learning for chord recognition using artificial neural networks. We present the argument that chroma vectors potentially hold enough information to model harmonic content of audio for chord recognition,…

Sound · Computer Science 2016-12-16 Filip Korzeniowski , Gerhard Widmer

We propose a real-time method to estimate spatiallyvarying indoor lighting from a single RGB image. Given an image and a 2D location in that image, our CNN estimates a 5th order spherical harmonic representation of the lighting at the given…

Computer Vision and Pattern Recognition · Computer Science 2019-06-11 Mathieu Garon , Kalyan Sunkavalli , Sunil Hadap , Nathan Carr , Jean-François Lalonde

Practical applications of fragment embedding and closely related local correlation methods critically depend on a judicious choice of a low-level theory to define the local embedding subspace and to capture long-range electrostatic and…

Chemical Physics · Physics 2026-05-14 Ruiheng Song , Xiliang Gong , Aamy Bakry , Hong-Zhou Ye

Fourier-encoded implicit neural representations (INRs) have shown strong capability in modeling continuous signals from discrete samples. However, conventional Fourier feature mappings use a fixed set of frequencies over the entire spatial…

Computer Vision and Pattern Recognition · Computer Science 2026-04-21 Ligen Shi , Jun Qiu , Yuhang Zheng , Zengyu Pang , Chang Liu

We propose a novel multi-stage depth super-resolution network, which progressively reconstructs high-resolution depth maps from explicit and implicit high-frequency features. The former are extracted by an efficient transformer processing…

Computer Vision and Pattern Recognition · Computer Science 2023-05-31 Xin Qiao , Chenyang Ge , Youmin Zhang , Yanhui Zhou , Fabio Tosi , Matteo Poggi , Stefano Mattoccia

Autonomous driving algorithms usually employ sRGB images as model input due to their compatibility with the human visual system. However, visually pleasing sRGB images are possibly sub-optimal for downstream tasks when compared to RAW…

Image and Video Processing · Electrical Eng. & Systems 2024-09-05 Anqi Liu , Shiyi Mu , Shugong Xu

In this work, we present a comparison between color spaces namely YUV, LAB, RGB and their effect on learned image compression. For this we use the structure and color based learned image codec (SLIC) from our prior work, which consists of…

Image and Video Processing · Electrical Eng. & Systems 2024-06-21 Srivatsa Prativadibhayankaram , Mahadev Prasad Panda , Jürgen Seiler , Thomas Richter , Heiko Sparenberg , Siegfried Fößel , André Kaup

The standard in LLM-based prediction is to use the final-layer representation as the input to a downstream predictor. However, intermediate layers may encode complementary task-relevant signals. Existing approaches therefore either search…

Computation and Language · Computer Science 2026-05-14 Tom Ulanovski , Eyal Blyachman , Maya Bechler-Speicher

Prediction problems from spectra are largely encountered in chemometry. In addition to accurate predictions, it is often needed to extract information about which wavelengths in the spectra contribute in an effective way to the quality of…

Neural and Evolutionary Computing · Computer Science 2008-02-05 Catherine Krier , Fabrice Rossi , Damien François , Michel Verleysen

Modern camera pipelines apply extensive on-device processing, such as exposure adjustment, white balance, and color correction, which, while beneficial individually, often introduce photometric inconsistencies across views. These appearance…

Computer Vision and Pattern Recognition · Computer Science 2025-10-01 Jisu Shin , Richard Shaw , Seunghyun Shin , Zhensong Zhang , Hae-Gon Jeon , Eduardo Perez-Pellitero

This paper presents a novel network structure with illumination-aware gamma correction and complete image modelling to solve the low-light image enhancement problem. Low-light environments usually lead to less informative large-scale dark…

Computer Vision and Pattern Recognition · Computer Science 2023-08-17 Yinglong Wang , Zhen Liu , Jianzhuang Liu , Songcen Xu , Shuaicheng Liu

Computational color constancy refers to the estimation of the scene illumination and makes the perceived color relatively stable under varying illumination. In the past few years, deep Convolutional Neural Networks (CNNs) have delivered…

Computer Vision and Pattern Recognition · Computer Science 2019-07-12 Jun Zhang , Tong Zheng , Shengping Zhang , Meng Wang

Traditional intra prediction usually utilizes the nearest reference line to generate the predicted block when considering strong spatial correlation. However, this kind of single line-based method does not always work well due to at least…

Multimedia · Computer Science 2016-12-05 Jiahao Li , Bin Li , Jizheng Xu , Ruiqin Xiong

Unsupervised domain adaptation (UDA) aims to learn a model trained on source domain and performs well on unlabeled target domain. In medical image segmentation field, most existing UDA methods depend on adversarial learning to address the…

Computer Vision and Pattern Recognition · Computer Science 2022-11-29 Shaolei Liu , Siqi Yin , Linhao Qu , Manning Wang

Iterative phase retrieval algorithms typically employ projections onto constraint subspaces to recover the unknown phases in the Fourier transform of an image, or, in the case of x-ray crystallography, the electron density of a molecule.…

Numerical Analysis · Mathematics 2025-10-20 Veit Elser

While Low-Rank Adaptation (LoRA) has proven beneficial for efficiently fine-tuning large models, LoRA fine-tuned text-to-image diffusion models lack diversity in the generated images, as the model tends to copy data from the observed…

Most of the existing deep learning based end-to-end image/video coding (DLEC) architectures are designed for non-subsampled RGB color format. However, in order to achieve a superior coding performance, many state-of-the-art block-based…

Image and Video Processing · Electrical Eng. & Systems 2021-08-30 Hilmi E. Egilmez , Ankitesh K. Singh , Muhammed Coban , Marta Karczewicz , Yinhao Zhu , Yang Yang , Amir Said , Taco S. Cohen

In contrast to traditional compression techniques performing linear transforms, the latent space of popular compressive autoencoders is obtained from a learned nonlinear mapping and hard to interpret. In this paper, we explore a promising…

Image and Video Processing · Electrical Eng. & Systems 2023-03-10 Anna Meyer , André Kaup

Image forensics, aiming to ensure the authenticity of the image, has made great progress in dealing with common image manipulation such as copy-move, splicing, and inpainting in the past decades. However, only a few researchers pay…

Computer Vision and Pattern Recognition · Computer Science 2022-04-26 Yushu Zhang , Nuo Chen , Shuren Qi , Mingfu Xue , Xiaochun Cao

Image inpainting plays a vital role in restoring missing image regions and supporting high-level vision tasks, but traditional methods struggle with complex textures and large occlusions. Although Transformer-based approaches have…

Computer Vision and Pattern Recognition · Computer Science 2025-06-24 Sijin He , Guangfeng Lin , Tao Li , Yajun Chen
‹ Prev 1 4 5 6 7 8 10 Next ›