English
Related papers

Related papers: CHEM: Estimating and Understanding Hallucinations …

200 papers

Deep learning based approaches to Computer Aided Diagnosis (CAD) typically pose the problem as an image classification (Normal or Abnormal) problem. These systems achieve high to very high accuracy in specific disease detection for which…

Computer Vision and Pattern Recognition · Computer Science 2020-10-06 Aniket Joshi , Gaurav Mishra , Jayanthi Sivaswamy

Dense reconstructions often contain errors that prior work has so far minimised using high quality sensors and regularising the output. Nevertheless, errors still persist. This paper proposes a machine learning technique to identify errors…

Computer Vision and Pattern Recognition · Computer Science 2018-01-31 Michael Tanner , Stefan Saftescu , Alex Bewley , Paul Newman

While deep learning-based image reconstruction methods have shown significant success in removing objects from pictures, they have yet to achieve acceptable results for attributing consistency to gender, ethnicity, expression, and other…

Computer Vision and Pattern Recognition · Computer Science 2022-04-04 Gourango Modak , Shuvra Smaran Das , Md. Ajharul Islam Miraj , Md. Kishor Morol

Deep learning models have witnessed depth and pose estimation framework on unannotated datasets as a effective pathway to succeed in endoscopic navigation. Most current techniques are dedicated to developing more advanced neural networks to…

Computer Vision and Pattern Recognition · Computer Science 2023-09-15 Junyang Wu , Yun Gu

Large Language Models (LLMs) are powerful linguistic engines but remain susceptible to hallucinations: plausible-sounding outputs that are factually incorrect or unsupported. In this work, we present a mathematically grounded framework to…

Computation and Language · Computer Science 2025-11-20 Moses Kiprono

High-quality image inpainting requires filling missing regions in a damaged image with plausible content. Existing works either fill the regions by copying image patches or generating semantically-coherent patches from region context, while…

Computer Vision and Pattern Recognition · Computer Science 2019-07-12 Yanhong Zeng , Jianlong Fu , Hongyang Chao , Baining Guo

Large Visual Language Models (LVLMs) struggle with hallucinations in visual instruction following task(s), limiting their trustworthiness and real-world applicability. We propose Pelican -- a novel framework designed to detect and mitigate…

Computation and Language · Computer Science 2024-10-30 Pritish Sahu , Karan Sikka , Ajay Divakaran

We propose the first general framework to automatically correct different types of geometric distortion in a single input image. Our proposed method employs convolutional neural networks (CNNs) trained by using a large synthetic distortion…

Computer Vision and Pattern Recognition · Computer Science 2019-09-10 Xiaoyu Li , Bo Zhang , Pedro V. Sander , Jing Liao

One primary technical challenge in photoacoustic microscopy (PAM) is the necessary compromise between spatial resolution and imaging speed. In this study, we propose a novel application of deep learning principles to reconstruct…

Image and Video Processing · Electrical Eng. & Systems 2020-06-02 Anthony DiSpirito , Daiwei Li , Tri Vu , Maomao Chen , Dong Zhang , Jianwen Luo , Roarke Horstmeyer , Junjie Yao

Whole slide image (WSI) normalization remains a vital preprocessing step in computational pathology. Increasingly driven by deep learning, these models learn to approximate data distributions from training examples. This often results in…

Computer Vision and Pattern Recognition · Computer Science 2025-12-09 Karel Moens , Matthew B. Blaschko , Tinne Tuytelaars , Bart Diricx , Jonas De Vylder , Mustafa Yousif

Recent advances in photoacoustic (PA) imaging have enabled detailed images of microvascular structure and quantitative measurement of blood oxygenation or perfusion. Standard reconstruction methods for PA imaging are based on solving an…

Signal Processing · Electrical Eng. & Systems 2020-04-17 MinWoo Kim , Geng-Shi Jeng , Ivan Pelivanov , Matthew O'Donnell

Skin cancer is one of the most common forms of cancer and its incidence is projected to rise over the next decade. Artificial intelligence is a viable solution to the issue of providing quality care to patients in areas lacking access to…

Computer Vision and Pattern Recognition · Computer Science 2018-08-07 Nithin D Reddy

Magnetic Resonance Imaging generally requires long exposure times, while being sensitive to patient motion, resulting in artifacts in the acquired images, which may hinder their diagnostic relevance. Despite research efforts to decrease the…

Image and Video Processing · Electrical Eng. & Systems 2025-02-04 Paolo Angella , Vito Paolo Pastore , Matteo Santacesaria

Homography estimation is a basic image alignment method in many applications. It is usually conducted by extracting and matching sparse feature points, which are error-prone in low-light and low-texture images. On the other hand, previous…

Computer Vision and Pattern Recognition · Computer Science 2020-07-21 Jirong Zhang , Chuan Wang , Shuaicheng Liu , Lanpeng Jia , Nianjin Ye , Jue Wang , Ji Zhou , Jian Sun

With the effective application of deep learning in computer vision, breakthroughs have been made in the research of super-resolution images reconstruction. However, many researches have pointed out that the insufficiency of the neural…

Image and Video Processing · Electrical Eng. & Systems 2021-06-11 Yibo Guo , Haidi Wang , Yiming Fan , Shunyao Li , Mingliang Xu

Automated brain lesions detection is an important and very challenging clinical diagnostic task because the lesions have different sizes, shapes, contrasts, and locations. Deep Learning recently has shown promising progress in many…

Computer Vision and Pattern Recognition · Computer Science 2018-01-08 Mina Rezaei , Haojin Yang , Christoph Meinel

3D reconstruction is a longstanding ill-posed problem, which has been explored for decades by the computer vision, computer graphics, and machine learning communities. Since 2015, image-based 3D reconstruction using convolutional neural…

Computer Vision and Pattern Recognition · Computer Science 2019-11-28 Xian-Feng Han , Hamid Laga , Mohammed Bennamoun

Video language models (Video-LLMs) are prone to hallucinations, often generating plausible but ungrounded content when visual evidence is weak, ambiguous, or biased. Existing decoding methods, such as contrastive decoding (CD), rely on…

Artificial Intelligence · Computer Science 2026-02-10 Qixin Xiao

In real-world clinical practice, overlooking unanticipated findings can result in serious consequences. However, supervised learning, which is the foundation for the current success of deep learning, only encourages models to identify…

Large Language Models (LLMs) have the tendency to hallucinate, i.e., to sporadically generate false or fabricated information. This presents a major challenge, as hallucinations often appear highly convincing and users generally lack the…

‹ Prev 1 8 9 10 Next ›