English
Related papers

Related papers: Movement-induced Priors for Deep Stereo

200 papers

Gaussian processes provide a flexible, non-parametric framework for the approximation of functions in high-dimensional spaces. The covariance kernel is the main engine of Gaussian processes, incorporating correlations that underpin the…

Machine Learning · Statistics 2024-03-20 Dionissios T. Hristopulos

We present a stereo-matching method for depth estimation from high-resolution images using visual hulls as priors, and a memory-efficient technique for the correlation computation. Our method uses object masks extracted from supplementary…

Computer Vision and Pattern Recognition · Computer Science 2024-06-05 Markus Plack , Hannah Dröge , Leif Van Holland , Matthias B. Hullin

This paper presents a novel incremental learning algorithm for pedestrian motion prediction, with the ability to improve the learned model over time when data is incrementally available. In this setup, trajectories are modeled as simple…

Robotics · Computer Science 2019-11-22 Golnaz Habibi , Nikita Japuria , Jonathan P. How

Automatic instrument segmentation in video is an essentially fundamental yet challenging problem for robot-assisted minimally invasive surgery. In this paper, we propose a novel framework to leverage instrument motion information, by…

Computer Vision and Pattern Recognition · Computer Science 2019-07-19 Yueming Jin , Keyun Cheng , Qi Dou , Pheng-Ann Heng

Generating non-existing frames from a consecutive video sequence has been an interesting and challenging problem in the video processing field. Typical kernel-based interpolation methods predict pixels with a single convolution process that…

Computer Vision and Pattern Recognition · Computer Science 2021-03-05 Xianhang Cheng , Zhenzhong Chen

Deep Gaussian processes have recently been proposed as natural objects to fit, similarly to deep neural networks, possibly complex features present in modern data samples, such as compositional structures. Adopting a Bayesian nonparametric…

Statistics Theory · Mathematics 2025-02-04 Ismaël Castillo , Thibault Randrianarisoa

Stereo is a prominent technique to infer dense depth maps from images, and deep learning further pushed forward the state-of-the-art, making end-to-end architectures unrivaled when enough data is available for training. However, deep…

Computer Vision and Pattern Recognition · Computer Science 2019-05-27 Matteo Poggi , Davide Pallotti , Fabio Tosi , Stefano Mattoccia

This paper investigates continuous representations of steering vectors over frequency and microphone/source positions for augmented listening (e.g., spatial filtering and binaural rendering), enabling user-parameterized control of the…

Audio and Speech Processing · Electrical Eng. & Systems 2026-04-17 Diego Di Carlo , Shoichi Koyama , Nugraha Aditya Arie , Fontaine Mathieu , Bando Yoshiaki , Yoshii Kazuyoshi

The deep image prior was recently introduced as a prior for natural images. It represents images as the output of a convolutional network with random inputs. For "inference", gradient descent is performed to adjust network parameters to…

Computer Vision and Pattern Recognition · Computer Science 2019-04-17 Zezhou Cheng , Matheus Gadelha , Subhransu Maji , Daniel Sheldon

We analyze the prior that a Deep Gaussian Process with polynomial kernels induces. We observe that, even for relatively small depths, averaging effects occur within such a Deep Gaussian Process and that the prior can be analyzed and…

Machine Learning · Statistics 2025-03-18 Daryna Chernobrovkina , Steffen Grünewälder

Motion detection in video is important for a number of applications and fields. In video surveillance, motion detection is an essential accompaniment to activity recognition for early warning systems. Robotics also has much to gain from…

Computer Vision and Pattern Recognition · Computer Science 2017-02-20 Peter Henderson , Matthew Vertescher

Deep learning-based methods have achieved significant successes on solving the blind super-resolution (BSR) problem. However, most of them request supervised pre-training on labelled datasets. This paper proposes an unsupervised kernel…

Image and Video Processing · Electrical Eng. & Systems 2024-04-29 Zhixiong Yang , Jingyuan Xia , Shengxi Li , Xinghua Huang , Shuanghui Zhang , Zhen Liu , Yaowen Fu , Yongxiang Liu

Recent advances in dance generation have enabled the automatic synthesis of 3D dance motions. However, existing methods still face significant challenges in simultaneously achieving high realism, precise dance-music synchronization, diverse…

A deep image compression scheme is proposed in this paper, offering the state-of-the-art compression efficiency, against the traditional JPEG, JPEG2000, BPG and those popular learning based methodologies. This is achieved by a novel…

Image and Video Processing · Electrical Eng. & Systems 2019-02-28 Haojie Liu , Tong Chen , Peiyao Guo , Qiu Shen , Zhan Ma

Learning and inference movement is a very challenging problem due to its high dimensionality and dependency to varied environments or tasks. In this paper, we propose an effective probabilistic method for learning and inference of basic…

Machine Learning · Computer Science 2018-10-30 Mingxuan Jing , Xiaojian Ma , Fuchun Sun , Huaping Liu

Event cameras are novel bio-inspired vision sensors that output pixel-level intensity changes in microsecond accuracy with a high dynamic range and low power consumption. Despite these advantages, event cameras cannot be directly applied to…

Computer Vision and Pattern Recognition · Computer Science 2022-11-02 Jinjin Gu , Jinan Zhou , Ringo Sai Wo Chu , Yan Chen , Jiawei Zhang , Xuanye Cheng , Song Zhang , Jimmy S. Ren

Gaussian processes are arguably the most important class of spatiotemporal models within machine learning. They encode prior information about the modeled function and can be used for exact or approximate Bayesian learning. In many…

A large number of cameras embedded on smart-phones, drones or inside cars have a direct access to external motion sensing from gyroscopes and accelerometers. On these power-limited devices, video compression must be of low-complexity. For…

Image and Video Processing · Electrical Eng. & Systems 2020-02-03 Karim El Khoury , Pascal Pellegrin , Antonin Descampe , Sébastien Lugan , Benoit Macq

Recent progress in image-to-video (I2V) diffusion models has significantly advanced the field of generative inbetweening, which aims to generate semantically plausible frames between two keyframes. In particular, inference-time sampling…

Computer Vision and Pattern Recognition · Computer Science 2026-02-20 Wooseok Jeon , Seunghyun Shin , Dongmin Shin , Hae-Gon Jeon

Depth sensing is an important problem for 3D vision-based robotics. Yet, a real-world active stereo or ToF depth camera often produces noisy and incomplete depth which bottlenecks robot performances. In this work, we propose D3RoMa, a…