中文
相关论文

相关论文: SAMA-IR: comprehensive input refinement methodolog…

200 篇论文

Network quantization is a dominant paradigm of model compression. However, the abrupt changes in quantized weights during training often lead to severe loss fluctuations and result in a sharp loss landscape, making the gradients unstable…

计算机视觉与模式识别 · 计算机科学 2023-03-22 Jing Liu , Jianfei Cai , Bohan Zhuang

The current state-of-the-art No-Reference Image Quality Assessment (NR-IQA) methods typically rely on feature extraction from upstream semantic backbone networks, assuming that all extracted features are relevant. However, we make a key…

计算机视觉与模式识别 · 计算机科学 2025-06-17 Xudong Li , Timin Gao , Runze Hu , Yan Zhang , Shengchuan Zhang , Xiawu Zheng , Jingyuan Zheng , Yunhang Shen , Ke Li , Yutao Liu , Pingyang Dai , Rongrong Ji

Reinforcement fine-tuning (RFT) is a proliferating paradigm for LMM training. Analogous to high-level reasoning tasks, RFT is similarly applicable to low-level vision domains, including image quality assessment (IQA). Existing RFT-based IQA…

计算机视觉与模式识别 · 计算机科学 2025-08-18 Ziheng Jia , Jiaying Qian , Zicheng Zhang , Zijian Chen , Xiongkuo Min

Networks are widely used in many fields for their powerful ability to provide vivid representations of relationships between variables. However, many of them may be corrupted by experimental noise or inappropriate network inference methods…

分子网络 · 定量生物学 2021-09-21 Jiating Yu , Jiacheng Leng , Ling-Yun Wu

Sharpness-Aware Minimization (SAM) was introduced to improve generalization by seeking flat minima, yet it also exhibits robustness to label noise, a phenomenon that remains only partially understood. Prior work has mainly attributed this…

机器学习 · 计算机科学 2026-03-31 Hoang-Chau Luong , Quang-Thuc Nguyen , Dat Ba Tran , Minh-Triet Tran

Deep learning approaches have shown promising performance for compressed sensing-based Magnetic Resonance Imaging. While deep neural networks trained with mean squared error (MSE) loss functions can achieve high peak signal to noise ratio,…

Speaker extraction aims to extract target speech signal from a multi-talker environment with interference speakers and surrounding noise, given the target speaker's reference information. Most speaker extraction systems achieve satisfactory…

音频与语音处理 · 电气工程与系统科学 2022-08-12 Chengyun Deng , Shiqian Ma , Yi Zhang , Yongtao Sha , Hui Zhang , Hui Song , Xiangang Li

A machine learning method for prediction of Raman gain and noise spectra is presented: it guarantees high-accuracy (RMSE < 0.4 dB) and low computational complexity making it suitable for real-time implementation in future optical networks…

信号处理 · 电气工程与系统科学 2019-05-03 Ann Margareth Rosa Brusin , Vittorio Curri , Darko Zibar , Andrea Carena

Deep learning approaches to optical flow estimation have seen rapid progress over the recent years. One common trait of many networks is that they refine an initial flow estimate either through multiple stages or across the levels of a…

计算机视觉与模式识别 · 计算机科学 2019-04-11 Junhwa Hur , Stefan Roth

Model compression by way of parameter pruning, quantization, or distillation has recently gained popularity as an approach for reducing the computational requirements of modern deep neural network models for NLP. Inspired by prior works…

计算与语言 · 计算机科学 2023-10-10 Clara Na , Sanket Vaibhav Mehta , Emma Strubell

Existing methods to recover model accuracy on analog-digital hardware in the presence of quantization and analog noise include noise-injection training. However, it can be slow in practice, incurring high computational costs, even when…

机器学习 · 计算机科学 2023-06-06 Lakshmi Nair , Darius Bunandar

Data augmentation is vital to the generalization ability and robustness of deep neural networks (DNNs) models. Existing augmentation methods for speaker verification manipulate the raw signal, which are time-consuming and the augmented…

音频与语音处理 · 电气工程与系统科学 2023-10-19 Yuanyuan Wang , Yang Zhang , Zhiyong Wu , Zhihan Yang , Tao Wei , Kun Zou , Helen Meng

The importance of Image quality assessment (IQA) is ever increasing due to the fast paced advances in imaging technology and computer vision. Among the numerous IQA methods, Structural SIMilarity (SSIM) index and its variants are better…

图像与视频处理 · 电气工程与系统科学 2022-12-06 X. Li , W. Armour

In inference problems, we often have domain knowledge which allows us to define summary statistics that capture most of the information content in a dataset. In this paper, we present a hybrid approach, where such physics-based summaries…

宇宙学与河外天体物理 · 物理学 2025-04-24 T. Lucas Makinen , Alan Heavens , Natalia Porqueres , Tom Charnock , Axel Lapel , Benjamin D. Wandelt

Probabilistic shaping of quadrature amplitude modulation (QAM) is used to enhance the sensitivity of an optical communication system. Sensitivity gains of 0.43 dB and 0.8 dB are demonstrated in back-to-back experiments by shaping of 16QAM…

信息论 · 计算机科学 2016-01-19 Tobias Fehenberger , Domaniç Lavery , Robert Maher , Alex Alvarado , Polina Bayvel , Norbert Hanik

Noise power estimation is a key issue in modern wireless communication systems. It allows resource allocation by detecting white spectral spaces effectively, and gives control over the communication process by adjusting transmission power.…

信息论 · 计算机科学 2017-11-16 Jakub Nikonowicz , Aamir Mahmood , Emiliano Sisinni , Mikael Gidlund

In recent years, the joint training of speech enhancement front-end and automatic speech recognition (ASR) back-end has been widely used to improve the robustness of ASR systems. Traditional joint training methods only use enhanced speech…

声音 · 计算机科学 2023-05-31 Haoyu Lu , Nan Li , Tongtong Song , Longbiao Wang , Jianwu Dang , Xiaobao Wang , Shiliang Zhang

We propose a n input parameter refinement scheme for the physics-based Raman amplifier model. Experiments over C+L band are conducted. Results show the scheme can lower the physical model's maximum estimation error by 2.13 dB.

光学 · 物理学 2024-05-31 Yihao Zhang , Xiaomin Liu , Qizhi Qiu , Yichen Liu , Lilin Yi , Weisheng Hu , Qunbi Zhuge

Supervised fine-tuning (SFT) on visual instruction data often improves perceptual capabilities in vision-language models (VLMs) while degrading reasoning performance, creating a persistent reasoning tax during post-training. We investigate…

计算机视觉与模式识别 · 计算机科学 2026-03-30 Yiming Ren , Yujiu Yang , Junjie Wang

In today's heavily overparameterized models, the value of the training loss provides few guarantees on model generalization ability. Indeed, optimizing only the training loss value, as is commonly done, can easily lead to suboptimal model…

机器学习 · 计算机科学 2021-04-30 Pierre Foret , Ariel Kleiner , Hossein Mobahi , Behnam Neyshabur
‹ 上一页 1 2 3 10 下一页 ›