中文
相关论文

相关论文: Audio Fingerprinting with Holographic Reduced Repr…

200 篇论文

As artificial intelligence and digital medicine increasingly permeate healthcare systems, robust governance frameworks are essential to ensure ethical, secure, and effective implementation. In this context, medical image retrieval becomes a…

计算机视觉与模式识别 · 计算机科学 2025-06-30 Yang Nan , Huichi Zhou , Xiaodan Xing , Giorgos Papanastasiou , Lei Zhu , Zhifan Gao , Alejandro F Fangi , Guang Yang

Dynamic Photoacoustic Computed Tomography (PACT) is an important imaging technique for monitoring physiological processes, capable of providing high-contrast images of optical absorption at much greater depths than traditional optical…

图像与视频处理 · 电气工程与系统科学 2025-06-05 Youshen Xiao , Yiling Shi , Ruixi Sun , Hongjiang Wei , Fei Gao , Yuyao Zhang

The speaker extraction algorithm extracts the target speech from a mixture speech containing interference speech and background noise. The extraction process sometimes over-suppresses the extracted target speech, which not only creates…

音频与语音处理 · 电气工程与系统科学 2022-06-22 Zexu Pan , Meng Ge , Haizhou Li

As speech generation technologies advance, so do risks of impersonation, misinformation, and spoofing. We present a lightweight, training-free approach for detecting synthetic speech and attributing it to its source model. Our method…

音频与语音处理 · 电气工程与系统科学 2025-12-12 Matías Pizarro , Mike Laszkiewicz , Dorothea Kolossa , Asja Fischer

Recent advances in machine learning techniques are enabling Automated Speech Recognition (ASR) more accurate and practical. The evidence of this can be seen in the rising number of smart devices with voice processing capabilities. More and…

密码学与安全 · 计算机科学 2022-03-15 Yogachandran Rahulamathavan

Audio fingerprinting converts audio to much lower-dimensional representations, allowing distorted recordings to still be recognized as their originals through similar fingerprints. Existing deep learning approaches rigidly fingerprint…

声音 · 计算机科学 2026-03-26 Hongjie Chen , Hanyu Meng , Huimin Zeng , Ryan A. Rossi , Lie Lu , Josh Kimball

Efficient face detection is critical to provide natural human-robot interactions. However, computer vision tends to involve a large computational load due to the amount of data (i.e. pixels) that needs to be processed in a short amount of…

音频与语音处理 · 电气工程与系统科学 2024-03-19 William Aris , François Grondin

This paper presents a novel technique to recover audio from sonorines, an early 20th century form of analogue sound storage. Our method uses high resolution photographs of sonorines under different lighting conditions to observe the change…

声音 · 计算机科学 2020-06-05 Kevin Feng

Beamforming in ultrasound imaging has significant impact on the quality of the final image, controlling its resolution and contrast. Despite its low spatial resolution and contrast, delay-and-sum is still extensively used nowadays in…

计算机视觉与模式识别 · 计算机科学 2016-05-20 Teodora Szasz , Adrian Basarab , Denis Kouamé

Acoustic-Resolution Photoacoustic Microscopy (AR-PAM) is promising for subcutaneous vascular imaging, but its spatial resolution is constrained by the Point Spread Function (PSF). Traditional deconvolution methods like Richardson-Lucy and…

计算机视觉与模式识别 · 计算机科学 2024-10-29 Youshen Xiao , Sheng Liao , Xuanyang Tian , Fan Zhang , Xinlong Dong , Yunhui Jiang , Xiyu Chen , Ruixi Sun , Yuyao Zhang , Fei Gao

Magnetic resonance imaging (MRI) is crucial for enhancing diagnostic accuracy in clinical settings. However, the inherent long scan time of MRI restricts its widespread applicability. Deep learning-based image super-resolution (SR) methods…

图像与视频处理 · 电气工程与系统科学 2024-02-19 Hao Li , Quanwei Liu , Jianan Liu , Xiling Liu , Yanni Dong , Tao Huang , Zhihan Lv

Automatic Speech Recognition (ASR) is an imperfect process that results in certain mismatches in ASR output text when compared to plain written text or transcriptions. When plain text data is to be used to train systems for spoken language…

计算与语言 · 计算机科学 2021-04-02 Prashant Serai , Vishal Sunder , Eric Fosler-Lussier

Using solely the information retrieved by audio fingerprinting techniques, we propose methods to treat a possibly large dataset of user-generated audio content, that (1) enable the grouping of several audio files that contain a common audio…

音频与语音处理 · 电气工程与系统科学 2017-09-18 Gonçalo Mordido , João Magalhães , Sofia Cavaco

We propose a novel low-rank tensor method for respiratory motion-resolved multi-echo image reconstruction. The key idea is to construct a 3-way image tensor (space $\times$ echo $\times$ motion state) from the conventional gridding…

图像与视频处理 · 电气工程与系统科学 2023-05-02 Seongho Jeong , MungSoo Kang , Gerald Behr , Heechul Jeong , Youngwook Kee

Cardiac pulsation is a physiological confound of functional magnetic resonance imaging (fMRI) time-series that introduces spurious signal fluctuations in proximity to blood vessels. fMRI alone is not sufficiently fast to resolve cardiac…

Super-resolution reconstruction (SRR) is a process aimed at enhancing spatial resolution of images, either from a single observation, based on the learned relation between low and high resolution, or from multiple images presenting the same…

计算机视觉与模式识别 · 计算机科学 2020-06-24 Michal Kawulok , Pawel Benecki , Szymon Piechaczek , Krzysztof Hrynczenko , Daniel Kostrzewa , Jakub Nalepa

The method of superposition is proposed in combination with a sparse $\ell_1$ optimisation algorithm with the aim of finding a sparse basis to accurately reconstruct the structural vibrations of a radiating object from a set of acoustic…

计算物理 · 物理学 2016-01-20 Nadia M. Abusag , David J. Chappell

Photoacoustic computed tomography (PACT) is a non-invasive imaging modality, similar to ultrasound, with wide-ranging medical applications. Conventional PACT images are degraded by wavefront distortion caused by the heterogeneous speed of…

图像与视频处理 · 电气工程与系统科学 2025-08-05 Tianao Li , Manxiu Cui , Cheng Ma , Emma Alexander

The technology of optical coherence tomography (OCT) to fingerprint imaging opens up a new research potential for fingerprint recognition owing to its ability to capture depth information of the skin layers. Developing robust and high…

计算机视觉与模式识别 · 计算机科学 2022-09-27 Wentian Zhang , Haozhe Liu , Feng Liu , Raghavendra Ramachandra

Natural images tend to mostly consist of smooth regions with individual pixels having highly correlated spectra. This information can be exploited to recover hyperspectral images of natural scenes from their incomplete and noisy…

计算机视觉与模式识别 · 计算机科学 2016-11-03 Reza Arablouei , Frank de Hoog