English
Related papers

Related papers: Dual Prompting for Diverse Count-level PET Denoisi…

200 papers

Prompt learning is one of the most effective paradigms for adapting pre-trained vision-language models (VLMs) to the biomedical image classification tasks in few shot scenarios. However, most of the current prompt learning methods only used…

Computer Vision and Pattern Recognition · Computer Science 2025-05-09 Wei Peng , Kang Liu , Jianchen Hu , Meng Zhang

Current multi-modal object re-identification approaches based on large-scale pre-trained backbones (i.e., ViT) have displayed remarkable progress and achieved excellent performance. However, these methods usually adopt the standard full…

Computer Vision and Pattern Recognition · Computer Science 2025-04-16 Minghui Lin , Shu Wang , Xiang Wang , Jianhua Tang , Longbin Fu , Zhengrong Zuo , Nong Sang

A flexible discriminative image denoiser is introduced in which multi-task learning methods are applied to a densoising FCN based on U-Net. The activations of the U-Net model are modified by affine transforms that are a learned function of…

Image and Video Processing · Electrical Eng. & Systems 2020-11-26 Anthony Kelly

We present a deep neural network to reduce coherent noise in three-dimensional quantitative phase imaging. Inspired by the cycle generative adversarial network, the denoising network was trained to learn a transform between two image…

Prompt learning methods are gaining increasing attention due to their ability to customize large vision-language models to new domains using pre-trained contextual knowledge and minimal training data. However, existing works typically rely…

Computer Vision and Pattern Recognition · Computer Science 2024-07-08 Duy M. H. Nguyen , An T. Le , Trung Q. Nguyen , Nghiem T. Diep , Tai Nguyen , Duy Duong-Tran , Jan Peters , Li Shen , Mathias Niepert , Daniel Sonntag

Prompt-based fine-tuning has boosted the performance of Pre-trained Language Models (PLMs) on few-shot text classification by employing task-specific prompts. Yet, PLMs are unfamiliar with prompt-style expressions during pre-training, which…

Computation and Language · Computer Science 2022-05-12 Jianing Wang , Chengyu Wang , Fuli Luo , Chuanqi Tan , Minghui Qiu , Fei Yang , Qiuhui Shi , Songfang Huang , Ming Gao

Diffusion models are well known for their ability to generate a high-fidelity image for an input prompt through an iterative denoising process. Unfortunately, the high fidelity also comes at a high computational cost due the inherently…

Computer Vision and Pattern Recognition · Computer Science 2025-06-24 Qinchan Li , Kenneth Chen , Changyue Su , Wittawat Jitkrittum , Qi Sun , Patsorn Sangkloy

Score-based generative models have demonstrated highly promising results for medical image reconstruction tasks in magnetic resonance imaging or computed tomography. However, their application to Positron Emission Tomography (PET) is still…

Image and Video Processing · Electrical Eng. & Systems 2024-01-24 Imraj RD Singh , Alexander Denker , Riccardo Barbano , Željko Kereta , Bangti Jin , Kris Thielemans , Peter Maass , Simon Arridge

Low-dose computed tomography (CT) denoising algorithms aim to enable reduced patient dose in routine CT acquisitions while maintaining high image quality. Recently, deep learning~(DL)-based methods were introduced, outperforming…

Image and Video Processing · Electrical Eng. & Systems 2022-10-21 Fabian Wagner , Mareike Thies , Felix Denzinger , Mingxuan Gu , Mayank Patwari , Stefan Ploner , Noah Maul , Laura Pfaff , Yixing Huang , Andreas Maier

Deep learning-based positron emission tomography (PET) image denoising offers the potential to reduce radiation exposure and scanning time by transforming low-count images into high-count equivalents. However, existing methods typically…

Image and Video Processing · Electrical Eng. & Systems 2025-03-06 Menghua Xia , Huidong Xie , Qiong Liu , Bo Zhou , Hanzhong Wang , Biao Li , Axel Rominger , Quanzheng Li , Ramsey D. Badawi , Kuangyu Shi , Georges El Fakhri , Chi Liu

Anomaly detection (AD) in 3D point clouds is crucial in a wide range of industrial applications, especially in various forms of precision manufacturing. Considering the industrial demand for reliable 3D AD, several methods have been…

Computer Vision and Pattern Recognition · Computer Science 2025-02-18 Jiaxiang Wang , Haote Xu , Xiaolu Chen , Haodi Xu , Yue Huang , Xinghao Ding , Xiaotong Tu

Learned denoisers play a fundamental role in various signal generation (e.g., diffusion models) and reconstruction (e.g., compressed sensing) architectures, whose success derives from their ability to leverage low-dimensional structure in…

Machine Learning · Computer Science 2025-08-14 Shiyu Wang , Mariam Avagyan , Yihan Shen , Arnaud Lamy , Tingran Wang , Szabolcs Márka , Zsuzsa Márka , John Wright

Dual-energy X-ray Computed Tomography (DECT) constitutes an advanced technology which enables automatic decomposition of materials in clinical images without manual segmentation using the dependency of the X-ray linear attenuation with…

Image and Video Processing · Electrical Eng. & Systems 2025-07-25 Hang Xu , Alexandre Bousse , Alessandro Perelli

Multi-rater annotations commonly occur when medical images are independently annotated by multiple experts (raters). In this paper, we tackle two challenges arisen in multi-rater annotations for medical image segmentation (called ambiguous…

Computer Vision and Pattern Recognition · Computer Science 2024-08-26 Jinhong Wang , Yi Cheng , Jintai Chen , Hongxia Xu , Danny Chen , Jian Wu

Foundation models like the segment anything model require high-quality manual prompts for medical image segmentation, which is time-consuming and requires expertise. SAM and its variants often fail to segment structures in ultrasound (US)…

Computer Vision and Pattern Recognition · Computer Science 2025-06-24 Assefa Seyoum Wahd , Banafshe Felfeliyan , Yuyue Zhou , Shrimanti Ghosh , Adam McArthur , Jiechen Zhang , Jacob L. Jaremko , Abhilash Hareendranathan

The pre-trained foundation models (PFMs) have become essential for facilitating large-scale multimodal learning. Researchers have effectively employed the ``pre-train, prompt, and predict'' paradigm through prompt learning to induce…

Computation and Language · Computer Science 2025-12-24 Xiang Chen , Yixin Ou , Quan Feng , Lei Li , Piji Li , Haibo Ye , Sheng-Jun Huang , Shuofei Qiao , Shumin Deng , Huajun Chen , Ningyu Zhang

Much of named entity recognition (NER) research focuses on developing dataset-specific models based on data from the domain of interest, and a limited set of related entity types. This is frustrating as each new dataset requires a new model…

Computation and Language · Computer Science 2023-02-23 Jinghui Lu , Rui Zhao , Brian Mac Namee , Fei Tan

We continue studies of the uncertainty quantification problem in emission tomographies such as PET or SPECT when additional multimodal data (e.g., anatomical MRI images) are available. To solve the aforementioned problem we adapt the…

Machine Learning · Statistics 2021-12-03 Fedor Goncharov , Éric Barat , Thomas Dautremer

Latest diffusion models have shown promising results in category-level 6D object pose estimation by modeling the conditional pose distribution with depth image input. The existing methods, however, suffer from slow convergence during…

Computer Vision and Pattern Recognition · Computer Science 2025-10-07 Seunghyun Lee , Tae-Kyun Kim

This paper presents a method of decoupled pronunciation and prosody modeling to improve the performance of meta-learning-based multilingual speech synthesis. The baseline meta-learning synthesis method adopts a single text encoder with a…

Audio and Speech Processing · Electrical Eng. & Systems 2022-09-15 Yukun Peng , Zhenhua Ling
‹ Prev 1 4 5 6 7 8 10 Next ›