中文
相关论文

相关论文: Physically-Guided Optical Inversion Enable Non-Con…

200 篇论文

Projecting the point cloud on the 2D spherical range image transforms the LiDAR semantic segmentation to a 2D segmentation task on the range image. However, the LiDAR range image is still naturally different from the regular 2D RGB image;…

计算机视觉与模式识别 · 计算机科学 2021-09-09 Yiming Zhao , Lin Bai , Xinming Huang

We develop a reduced-order operator-learning framework for forward and inverse band-structure design of two-dimensional photonic crystals with binary, pixel-based $p4m$-symmetric unit cells. We construct a POD--DeepONet surrogate for the…

光学 · 物理学 2026-01-05 Yueqi Wang , Guanglian Li , Guang Lin

Current methods for restoring underexposed images typically rely on supervised learning with paired underexposed and well-illuminated images. However, collecting such datasets is often impractical in real-world scenarios. Moreover, these…

计算机视觉与模式识别 · 计算机科学 2025-07-04 Hailong Yan , Junjian Huang , Tingwen Huang

Super-resolution of LiDAR range images is crucial to improving many downstream tasks such as object detection, recognition, and tracking. While deep learning has made a remarkable advances in super-resolution techniques, typical…

机器人学 · 计算机科学 2022-03-15 Youngsun Kwon , Minhyuk Sung , Sung-Eui Yoon

3D LiDARs and 2D cameras are increasingly being used alongside each other in sensor rigs for perception tasks. Before these sensors can be used to gather meaningful data, however, their extrinsics (and intrinsics) need to be accurately…

机器人学 · 计算机科学 2019-08-06 Ganesh Iyer , R. Karnik Ram. , J. Krishna Murthy , K. Madhava Krishna

We consider a class of inverse problems characterized by forward operators that are partially specified, non-smooth, and non-differentiable. Although generative inverse solvers have made significant progress, we find that these forward…

计算机视觉与模式识别 · 计算机科学 2026-03-13 Sattwik Basu , Chaitanya Amballa , Zhongweiyang Xu , Jorge Vančo Sampedro , Srihari Nelakuditi , Romit Roy Choudhury

Optical Coherence Tomography (OCT) and Optical Coherence Tomography Angiography (OCTA) are key diagnostic tools for clinical evaluation and management of retinal diseases. Compared to traditional OCT, OCTA provides richer microvascular…

图像与视频处理 · 电气工程与系统科学 2025-04-01 Renzhi Tian , Jinjie Wang , Wei Yang , Weizhen Li , Haoran Chen , Yiran Zhu , Chengchang Pan , Honggang Qi

Generalizable 3D object reconstruction from single-view RGB-D images remains a challenging task, particularly with real-world data. Current state-of-the-art methods develop Transformer-based implicit field learning, necessitating an…

计算机视觉与模式识别 · 计算机科学 2024-04-02 Yushuang Wu , Luyue Shi , Junhao Cai , Weihao Yuan , Lingteng Qiu , Zilong Dong , Liefeng Bo , Shuguang Cui , Xiaoguang Han

Deep convolutional neural networks have revolutionized many machine learning and computer vision tasks, however, some remaining key challenges limit their wider use. These challenges include improving the network's robustness to…

计算机视觉与模式识别 · 计算机科学 2019-05-21 Eldad Haber , Keegan Lensink , Eran Treister , Lars Ruthotto

This paper addresses the problem of photometric stereo, in both calibrated and uncalibrated scenarios, for non-Lambertian surfaces based on deep learning. We first introduce a fully convolutional deep network for calibrated photometric…

计算机视觉与模式识别 · 计算机科学 2020-07-28 Guanying Chen , Kai Han , Boxin Shi , Yasuyuki Matsushita , Kwan-Yee K. Wong

Recent progress in computational photography has shown that we can acquire near-infrared (NIR) information in addition to the normal visible (RGB) band, with only slight modifications to standard digital cameras. Due to the proximity of the…

计算机视觉与模式识别 · 计算机科学 2014-06-25 Neda Salamati , Diane Larlus , Gabriela Csurka , Sabine Süsstrunk

Temporal drift of sensory data is a severe problem impacting the data quality of wireless sensor networks (WSNs). With the proliferation of large-scale and long-term WSNs, it is becoming more important to calibrate sensors when the ground…

机器学习 · 计算机科学 2017-07-13 Yuzhi Wang , Anqi Yang , Xiaoming Chen , Pengjun Wang , Yu Wang , Huazhong Yang

Optical super-oscillation enables far-field super-resolution imaging beyond diffraction limits. However, the existing super-oscillatory lens for the spatial super-resolution imaging system still confronts critical limitations in performance…

Transferring optical information through random diffusers is a critical yet challenging task. In this work, we introduce a cascaded diffractive optical network for information transfer through random and unknown diffusers, achieved through…

光学 · 物理学 2026-03-10 Yuhang Li , Yiyang Wu , Shiqi Chen , Xilin Yang , Aydogan Ozcan

360{\deg} images are usually represented in either equirectangular projection (ERP) or multiple perspective projections. Different from the flat 2D images, the detection task is challenging for 360{\deg} images due to the distortion of ERP…

计算机视觉与模式识别 · 计算机科学 2019-07-30 Pengyu Zhao , Ansheng You , Yuanxing Zhang , Jiaying Liu , Kaigui Bian , Yunhai Tong

We present a photonics integrated circuit on silicon substrate withreconfigurable nonreciprocal transmission that exhibits a large isolationratio and low insertion loss. It also offers ability for all-optical function-alities, like optical…

光学 · 物理学 2019-05-01 Ang Li , Wim Bogaerts

Adversarial attacks can mislead deep learning models to make false predictions by implanting small perturbations to the original input that are imperceptible to the human eye, which poses a huge security threat to the computer vision…

计算机视觉与模式识别 · 计算机科学 2024-10-28 Junbin Fang , You Jiang , Canjian Jiang , Zoe L. Jiang , Siu-Ming Yiu , Chuanyi Liu

The quality of the prompts provided to text-to-image diffusion models determines how faithful the generated content is to the user's intent, often requiring `prompt engineering'. To harness visual concepts from target images without prompt…

计算机视觉与模式识别 · 计算机科学 2023-12-20 Shweta Mahajan , Tanzila Rahman , Kwang Moo Yi , Leonid Sigal

Reconstruction of 3D open surfaces (e.g., non-watertight meshes) is an underexplored area of computer vision. Recent learning-based implicit techniques have removed previous barriers by enabling reconstruction in arbitrary resolutions. Yet,…

计算机视觉与模式识别 · 计算机科学 2023-11-07 Mohammad Samiul Arshad , William J. Beksi

We propose Intra and Inter Parser-Prompted Transformers (PPTformer) that explore useful features from visual foundation models for image restoration. Specifically, PPTformer contains two parts: an Image Restoration Network (IRNet) for…

计算机视觉与模式识别 · 计算机科学 2025-03-19 Cong Wang , Jinshan Pan , Liyan Wang , Wei Wang