中文
相关论文

相关论文: A learning-based view extrapolation method for axi…

200 篇论文

Large diffusion models demonstrate remarkable zero-shot capabilities in novel view synthesis from a single image. However, these models often face challenges in maintaining consistency across novel and reference views. A crucial factor…

计算机视觉与模式识别 · 计算机科学 2025-02-26 Botao Ye , Sifei Liu , Xueting Li , Marc Pollefeys , Ming-Hsuan Yang

Linear position interpolation helps pre-trained models using rotary position embeddings (RoPE) to extrapolate to longer sequence lengths. We propose using linear position interpolation to extend the extrapolation range of models using…

计算与语言 · 计算机科学 2023-10-23 Faisal Al-Khateeb , Nolan Dey , Daria Soboleva , Joel Hestness

We demonstrate that a deep neural network can significantly improve optical microscopy, enhancing its spatial resolution over a large field-of-view and depth-of-field. After its training, the only input to this network is an image acquired…

机器学习 · 计算机科学 2017-11-21 Yair Rivenson , Zoltan Gorocs , Harun Gunaydin , Yibo Zhang , Hongda Wang , Aydogan Ozcan

eXplanation Based Learning (XBL) is an interactive learning approach that provides a transparent method of training deep learning models by interacting with their explanations. XBL augments loss functions to penalize a model based on…

计算机视觉与模式识别 · 计算机科学 2023-09-12 Misgina Tsighe Hagos , Niamh Belton , Kathleen M. Curran , Brian Mac Namee

Diffusion transformers (DiTs) struggle to generate images at resolutions higher than their training resolutions. The primary obstacle is that the explicit positional encodings(PE), such as RoPE, need extrapolating to unseen positions which…

计算机视觉与模式识别 · 计算机科学 2025-09-25 Shen Zhang , Siyuan Liang , Yaning Tan , Zhaowei Chen , Linze Li , Ge Wu , Yuhao Chen , Shuheng Li , Zhenyu Zhao , Caihua Chen , Jiajun Liang , Yao Tang

Light field reconstruction from images captured by focal plane sweeping, such as light field moment imaging (LFMI) and light field reconstruction with back projection (LFBP), can achieve high lateral resolution comparable to the modern…

光学 · 物理学 2018-05-23 Haichao Wang , Ni Chen , Jingdan Liu , Guohai Situ

Classical approximation and learning methods are typically optimized for interpolation over a sampled domain {\Omega}, with no guarantees on their behavior in an extrapolation region {\Xi}, where small in-domain errors may amplify. We…

数值分析 · 数学 2026-03-11 Guy Hay , Nir Sharon

In this paper, a novel convolutional neural network (CNN)-based framework is developed for light field reconstruction from a sparse set of views. We indicate that the reconstruction can be efficiently modeled as angular restoration on an…

图像与视频处理 · 电气工程与系统科学 2021-03-25 Gaochang Wu , Yebin Liu , Lu Fang , Qionghai Dai , Tianyou Chai

In this paper, we demonstrate light field triangulation to determine depth distances and baselines in a plenoptic camera. Advances in micro lenses and image sensors have enabled plenoptic cameras to capture a scene from different viewpoints…

信息检索 · 计算机科学 2021-01-21 Christopher Hahne , Amar Aggoun , Vladan Velisavljevic , Susanne Fiebig , Matthias Pesch

Convolutional layers are an integral part of many deep neural network solutions in computer vision. Recent work shows that replacing the standard convolution operation with mechanisms based on self-attention leads to improved performance on…

计算机视觉与模式识别 · 计算机科学 2020-12-21 Souvik Kundu , Hesham Mostafa , Sharath Nittur Sridhar , Sairam Sundaresan

Artificial neural networks have revolutionized fields from computer vision to natural language processing, yet their growing energy and computational demands threaten future progress. Optical neural networks promise greater speed,…

光学 · 物理学 2025-08-18 Bofeng Liu , Xu Mei , Sadman Shafi , Tunan Xia , Iam-Choon Khoo , Zhiwen Liu , Xingjie Ni

We consider the problem of high-dimensional light field reconstruction and develop a learning-based framework for spatial and angular super-resolution. Many current approaches either require disparity clues or restore the spatial and…

图像与视频处理 · 电气工程与系统科学 2020-09-18 Nan Meng , Hayden K. -H. So , Xing Sun , Edmund Y. Lam

We suggest representing light field (LF) videos as "one-off" neural networks (NN), i.e., a learned mapping from view-plus-time coordinates to high-resolution color values, trained on sparse views. Initially, this sounds like a bad idea for…

图形学 · 计算机科学 2020-04-23 Mojtaba Bemana , Karol Myszkowski , Hans-Peter Seidel , Tobias Ritschel

Autofocus is an important task for digital cameras, yet current approaches often exhibit poor performance. We propose a learning-based approach to this problem, and provide a realistic dataset of sufficient size for effective learning. Our…

计算机视觉与模式识别 · 计算机科学 2020-05-05 Charles Herrmann , Richard Strong Bowen , Neal Wadhwa , Rahul Garg , Qiurui He , Jonathan T. Barron , Ramin Zabih

Image Completion refers to the task of filling in the missing regions of an image and Image Extrapolation refers to the task of extending an image at its boundaries while keeping it coherent. Many recent works based on GAN have shown…

计算机视觉与模式识别 · 计算机科学 2020-06-05 Sai Hemanth Kasaraneni , Abhishek Mishra

Existing image-based rendering methods usually adopt depth-based image warping operation to synthesize novel views. In this paper, we reason the essential limitations of the traditional warping operation to be the limited neighborhood and…

计算机视觉与模式识别 · 计算机科学 2023-01-31 Mantang Guo , Junhui Hou , Jing Jin , Hui Liu , Huanqiang Zeng , Jiwen Lu

We analyze the far field resolution of apertures which are illuminated by a point dipole located at subwavelength distances. It is well known that radiation emitted by a localized source can be considered a combination of travelling and…

光学 · 物理学 2024-03-25 Aziz Kolkiran , G. S. Agarwal

Learning approaches have shown great success in the task of super-resolving an image given a low resolution input. Video super-resolution aims for exploiting additionally the information from multiple images. Typically, the images are…

计算机视觉与模式识别 · 计算机科学 2017-07-04 Osama Makansi , Eddy Ilg , Thomas Brox

We predict future video frames from complex dynamic scenes, using an invertible neural network as the encoder of a nonlinear dynamic system with latent linear state evolution. Our invertible linear embedding (ILE) demonstrates successful…

计算机视觉与模式识别 · 计算机科学 2019-03-04 Robert Pottorff , Jared Nielsen , David Wingate

Light field disparity estimation is an essential task in computer vision with various applications. Although supervised learning-based methods have achieved both higher accuracy and efficiency than traditional optimization-based methods,…

计算机视觉与模式识别 · 计算机科学 2022-04-01 Peng Li , Jiayin Zhao , Jingyao Wu , Chao Deng , Haoqian Wang , Tao Yu