中文
相关论文

相关论文: Fast Autofocusing using Tiny Transformer Networks …

200 篇论文

Digital breast tomosynthesis (DBT) exams should utilize the lowest possible radiation dose while maintaining sufficiently good image quality for accurate medical diagnosis. In this work, we propose a convolution neural network (CNN) to…

图像与视频处理 · 电气工程与系统科学 2022-03-23 Rodrigo de Barros Vimieiro , Chuang Niu , Hongming Shan , Lucas Rodrigues Borges , Ge Wang , Marcelo Andrade da Costa Vieira

This paper aims to address a common challenge in deep learning-based image transformation methods, such as image enhancement and super-resolution, which heavily rely on precisely aligned paired datasets with pixel-level alignments. However,…

计算机视觉与模式识别 · 计算机科学 2024-02-29 Zhangkai Ni , Juncheng Wu , Zian Wang , Wenhan Yang , Hanli Wang , Lin Ma

Rectifying the orientation of images represents a daily task for every photographer. This task may be complicated even for the human eye, especially when the horizon or other horizontal and vertical lines in the image are missing. In this…

计算机视觉与模式识别 · 计算机科学 2021-05-13 Ionut Mironica , Andrei Zugravu

Terahertz (THz) communications is considered as one of key solutions to support extremely high data demand in 6G. One main difficulty of the THz communication is the severe signal attenuation caused by the foliage loss, oxygen/atmospheric…

信号处理 · 电气工程与系统科学 2024-05-14 Jinhong Kim , Yongjun Ahn , Seungnyun Kim , Byonghyo Shim

Vision Transformers have attracted a lot of attention recently since the successful implementation of Vision Transformer (ViT) on vision tasks. With vision Transformers, specifically the multi-head self-attention modules, networks can…

计算机视觉与模式识别 · 计算机科学 2022-10-27 Xiangyu Chen , Ying Qin , Wenju Xu , Andrés M. Bur , Cuncong Zhong , Guanghui Wang

In the emerging high mobility Vehicle-to-Everything (V2X) communications using millimeter Wave (mmWave) and sub-THz, Multiple-Input Multiple-Output (MIMO) channel estimation is an extremely challenging task. At mmWaves/sub-THz frequencies,…

Vision transformers (ViTs) have become the popular structures and outperformed convolutional neural networks (CNNs) on various vision tasks. However, such powerful transformers bring a huge computation burden, because of the exhausting…

计算机视觉与模式识别 · 计算机科学 2022-09-13 Zhuofan Zong , Kunchang Li , Guanglu Song , Yali Wang , Yu Qiao , Biao Leng , Yu Liu

Recovering ghost-free High Dynamic Range (HDR) images from multiple Low Dynamic Range (LDR) images becomes challenging when the LDR images exhibit saturation and significant motion. Recent Diffusion Models (DMs) have been introduced in HDR…

计算机视觉与模式识别 · 计算机科学 2024-04-02 Tao Hu , Qingsen Yan , Yuankai Qi , Yanning Zhang

Vision Transformer (ViT) self-attention mechanism is characterized by feature collapse in deeper layers, resulting in the vanishing of low-level visual features. However, such features can be helpful to accurately represent and identify…

计算机视觉与模式识别 · 计算机科学 2024-08-06 Anxhelo Diko , Danilo Avola , Marco Cascio , Luigi Cinque

Depth from focus (DFF) is one of the classical ill-posed inverse problems in computer vision. Most approaches recover the depth at each pixel based on the focal setting which exhibits maximal sharpness. Yet, it is not obvious how to…

计算机视觉与模式识别 · 计算机科学 2018-10-30 Caner Hazirbas , Sebastian Georg Soyer , Maximilian Christian Staab , Laura Leal-Taixé , Daniel Cremers

Deep-learning (DL)-based image deconvolution (ID) has exhibited remarkable recovery performance, surpassing traditional linear methods. However, unlike traditional ID approaches that rely on analytical properties of the point spread…

图像与视频处理 · 电气工程与系统科学 2025-01-28 Romario Gualdrón-Hurtado , Roman Jacome , Sergio Urrea , Henry Arguello , Luis Gonzalez

X-ray Computed Tomography (CT) is widely used in clinical applications such as diagnosis and image-guided interventions. In this paper, we propose a new deep learning based model for CT image reconstruction with the backbone network…

图像与视频处理 · 电气工程与系统科学 2020-09-21 Haimiao Zhang , Baodong Liu , Hengyong Yu , Bin Dong

Deep learning (DL) is a powerful tool in computational imaging for many applications. A common strategy is to reconstruct a preliminary image as the input of a neural network to achieve an optimized image. Usually, the preliminary image is…

图像与视频处理 · 电气工程与系统科学 2021-05-12 Ruibo Shang , Kevin Hoffer-Hawlik , Geoffrey P. Luke

We introduce a novel technique for designing color filter metasurfaces using a data-driven approach based on deep learning. Our innovative approach employs inverse design principles to identify highly efficient designs that outperform all…

Ocular Toxoplasmosis (OT), is a common eye infection caused by T. gondii that can cause vision problems. Diagnosis is typically done through a clinical examination and imaging, but these methods can be complicated and costly, requiring…

图像与视频处理 · 电气工程与系统科学 2023-05-19 Syed Samiul Alam , Samiul Based Shuvo , Shams Nafisa Ali , Fardeen Ahmed , Arbil Chakma , Yeong Min Jang

Current video deblurring methods have limitations in recovering high-frequency information since the regression losses are conservative with high-frequency details. Since Diffusion Models (DMs) have strong capabilities in generating…

计算机视觉与模式识别 · 计算机科学 2024-08-27 Chen Rao , Guangyuan Li , Zehua Lan , Jiakai Sun , Junsheng Luan , Wei Xing , Lei Zhao , Huaizhong Lin , Jianfeng Dong , Dalong Zhang

Half of wavefunction information is undetected by conventional transmission electron microscopy (CTEM) as only the intensity, and not the phase, of an image is recorded. Following successful applications of deep learning to optical hologram…

图像与视频处理 · 电气工程与系统科学 2020-01-31 Jeffrey M. Ede , Jonathan J. P. Peters , Jeremy Sloan , Richard Beanland

In recent years, deep learning has made brilliant achievements in Environmental Microorganism (EM) image classification. However, image classification of small EM datasets has still not obtained good research results. Therefore, researchers…

计算机视觉与模式识别 · 计算机科学 2022-02-04 Peng Zhao , Chen Li , Md Mamunur Rahaman , Hao Xu , Hechen Yang , Hongzan Sun , Tao Jiang , Marcin Grzegorzek

Diffusion language models (D-LLMs) offer parallel denoising and bidirectional context, but hallucination detection for D-LLMs remains underexplored. Prior detectors developed for auto-regressive LLMs typically rely on single-pass cues and…

计算与语言 · 计算机科学 2026-02-10 Arshia Hemmat , Philip Torr , Yongqiang Chen , Junchi Yu

In-line holographic microscopy provides an unparalleled wealth of information about the properties of colloidal dispersions. Analyzing one colloidal particle's hologram with the Lorenz-Mie theory of light scattering yields the particle's…

软凝聚态物质 · 物理学 2020-02-25 Lauren E. Altman , David G. Grier
‹ 上一页 1 8 9 10 下一页 ›