English
Related papers

Related papers: Enhancing the machine vision performance with mult…

200 papers

Large datasets often contain multiple distinct feature sets, or views, that offer complementary information that can be exploited by multi-view learning methods to improve results. We investigate anatomical multi-view data, where each brain…

Quantitative Methods · Quantitative Biology 2024-01-17 Yuxiang Wei , Yuqian Chen , Tengfei Xue , Leo Zekelman , Nikos Makris , Yogesh Rathi , Weidong Cai , Fan Zhang , Lauren J. O' Donnell

Frontier multimodal large language models (MLLMs) have been reported to achieve over 90% accuracy on fine-grained perception benchmarks. However, such scores do not necessarily imply faithful use of visual evidence. Prior studies have…

Computer Vision and Pattern Recognition · Computer Science 2026-05-27 Jingru Chen , Yiming Liu , Mingtao Chen , Sijie Chen , Richeng Xuan , Liang Yang , Zhichao Hu , Fanyang Lu

Machine-learning (ML) algorithms will play a crucial role in studying the large datasets delivered by new facilities over the next decade and beyond. Here, we investigate the capabilities and limits of such methods in finding galaxies with…

Instrumentation and Methods for Astrophysics · Physics 2019-08-22 Andreas L. Faisst , Abhishek Prakash , Peter L. Capak , Bomee Lee

We estimate the incidence of multiply-imaged AGNs among the optical counterparts of X-ray selected point-like sources in the XXL field. We also derive the expected statistical properties of this sample, such as the redshift distribution of…

Cosmology and Nongalactic Astrophysics · Physics 2015-08-06 F. Finet , A. Elyiv , O. Melnyk , O. Wertz , C. Horellou , J. Surdej

Lately, researchers in artificial intelligence have been really interested in how language and vision come together, giving rise to the development of multimodal models that aim to seamlessly integrate textual and visual information.…

Computer Vision and Pattern Recognition · Computer Science 2024-10-29 Rajat Chawla , Arkajit Datta , Tushar Verma , Adarsh Jha , Anmol Gautam , Ayush Vatsal , Sukrit Chaterjee , Mukunda NS , Ishaan Bhola

Vehicle information recognition is crucial in various practical domains, particularly in criminal investigations. Vehicle Color Recognition (VCR) has garnered significant research interest because color is a visually distinguishable…

Computer Vision and Pattern Recognition · Computer Science 2024-10-22 Gabriel E. Lima , Rayson Laroca , Eduardo Santos , Eduil Nascimento , David Menotti

Contemporary approaches frame the color constancy problem as learning camera specific illuminant mappings. While high accuracy can be achieved on camera specific data, these models depend on camera spectral sensitivity and typically exhibit…

Computer Vision and Pattern Recognition · Computer Science 2020-03-03 Daniel Hernandez-Juarez , Sarah Parisot , Benjamin Busam , Ales Leonardis , Gregory Slabaugh , Steven McDonagh

This paper presents a proposed AI Deep Learning model that addresses common challenges encountered in Visible Light Communication (VLC) systems. In this work, we run a Python simulation that models a basic VLC system primarily affected by…

Signal Processing · Electrical Eng. & Systems 2025-07-14 A. A. Nutfaji , Moustafa Hassan Elmallah

Scene text instances found in natural images carry explicit semantic information that can provide important cues to solve a wide array of computer vision problems. In this paper, we focus on leveraging multi-modal content in the form of…

Computer Vision and Pattern Recognition · Computer Science 2020-09-22 Andres Mafla , Sounak Dey , Ali Furkan Biten , Lluis Gomez , Dimosthenis Karatzas

The growing availability of digitized art collections has created the need to manage, analyze and categorize large amounts of data related to abstract concepts, highlighting a demanding problem of computer science and leading to new…

Computer Vision and Pattern Recognition · Computer Science 2023-05-01 Vassilis Lyberatos , Paraskevi-Antonia Theofilou , Jason Liartis , Georgios Siolas

We address the problem of keypoint selection, and find that the performance of 6DoF pose estimation methods can be improved when pre-defined keypoint locations are learned, rather than being heuristically selected as has been the standard…

Computer Vision and Pattern Recognition · Computer Science 2023-11-13 Yangzheng Wu , Michael Greenspan

This paper presents a novel multi-fake evolutionary generative adversarial network(MFEGAN) for handling imbalance hyperspectral image classification. It is an end-to-end approach in which different generative objective losses are considered…

Image and Video Processing · Electrical Eng. & Systems 2024-09-04 Tanmoy Dam , Nidhi Swami , Sreenatha G. Anavatti , Hussein A. Abbass

Current synoptic sky surveys monitor large areas of the sky to find variable and transient astronomical sources. As the number of detections per night at a single telescope easily exceeds several thousand, current detection pipelines make…

We present a multispectral extension to 3D Gaussian Splatting (3DGS) for wavelength-aware view synthesis. Each Gaussian is augmented with spectral radiance, represented via per-band spherical harmonics, and optimized under a dual-loss…

Computer Vision and Pattern Recognition · Computer Science 2026-04-16 Iris Zheng , Guojun Tang , Alexander Doronin , Paul Teal , Fang-Lue Zhang

When taking images against strong light sources, the resulting images often contain heterogeneous flare artifacts. These artifacts can importantly affect image visual quality and downstream computer vision tasks. While collecting real data…

Image and Video Processing · Electrical Eng. & Systems 2023-09-01 Yuyan Zhou , Dong Liang , Songcan Chen , Sheng-Jun Huang , Shuo Yang , Chongyi Li

Forthcoming imaging surveys will potentially increase the number of known galaxy-scale strong lenses by several orders of magnitude. For this to happen, images of tens of millions of galaxies will have to be inspected to identify potential…

Astrophysics of Galaxies · Physics 2024-01-29 Euclid Collaboration , L. Leuzzi , M. Meneghetti , G. Angora , R. B. Metcalf , L. Moscardini , P. Rosati , P. Bergamini , F. Calura , B. Clément , R. Gavazzi , F. Gentile , M. Lochner , C. Grillo , G. Vernardos , N. Aghanim , A. Amara , L. Amendola , S. Andreon , N. Auricchio , S. Bardelli , C. Bodendorf , D. Bonino , E. Branchini , M. Brescia , J. Brinchmann , S. Camera , V. Capobianco , C. Carbone , J. Carretero , S. Casas , M. Castellano , S. Cavuoti , A. Cimatti , R. Cledassou , G. Congedo , C. J. Conselice , L. Conversi , Y. Copin , L. Corcione , F. Courbin , H. M. Courtois , M. Cropper , A. Da Silva , H. Degaudenzi , J. Dinis , F. Dubath , X. Dupac , S. Dusini , M. Farina , S. Farrens , S. Ferriol , M. Frailis , E. Franceschi , M. Fumana , S. Galeotta , B. Gillis , C. Giocoli , A. Grazian , F. Grupp , L. Guzzo , S. V. H. Haugan , W. Holmes , I. Hook , F. Hormuth , A. Hornstrup , P. Hudelot , K. Jahnke , B. Joachimi , M. Kümmel , E. Keihänen , S. Kermiche , A. Kiessling , T. Kitching , M. Kunz , H. Kurki-Suonio , P. B. Lilje , V. Lindholm , I. Lloro , D. Maino , E. Maiorano , O. Mansutti , O. Marggraf , K. Markovic , N. Martinet , F. Marulli , R. Massey , E. Medinaceli , S. Mei , M. Melchior , Y. Mellier , E. Merlin , G. Meylan , M. Moresco , E. Munari , S. -M. Niemi , J. W. Nightingale , T. Nutma , C. Padilla , S. Paltani , F. Pasian , K. Pedersen , V. Pettorino , S. Pires , G. Polenta , M. Poncet , F. Raison , A. Renzi , J. Rhodes , G. Riccio , E. Romelli , M. Roncarelli , E. Rossetti , R. Saglia , D. Sapone , B. Sartoris , M. Schirmer , P. Schneider , A. Secroun , G. Seidel , S. Serrano , C. Sirignano , G. Sirri , L. Stanco , P. Tallada-Crespí , A. N. Taylor , I. Tereno , R. Toledo-Moreo , F. Torradeflot , I. Tutusaus , L. Valenziano , T. Vassallo , A. Veropalumbo , Y. Wang , J. Weller , G. Zamorani , J. Zoubian , E. Zucca , A. Boucaud , E. Bozzo , C. Colodro-Conde , D. Di Ferdinando , R. Farinelli , J. Graciá-Carpio , N. Mauri , C. Neissner , V. Scottez , M. Tenti , A. Tramacere , Y. Akrami , V. Allevato , C. Baccigalupi , M. Ballardini , F. Bernardeau , A. Biviano , S. Borgani , A. S. Borlaff , H. Bretonnière , C. Burigana , R. Cabanac , A. Cappi , C. S. Carvalho , G. Castignani , T. Castro , K. C. Chambers , A. R. Cooray , J. Coupon , S. Davini , S. de la Torre , G. De Lucia , G. Desprez , S. Di Domizio , H. Dole , J. A. Escartin Vigo , S. Escoffier , I. Ferrero , L. Gabarra , K. Ganga , J. Garcia-Bellido , E. Gaztanaga , K. George , G. Gozaliasl , H. Hildebrandt , M. Huertas-Company , J. J. E. Kajava , V. Kansal , C. C. Kirkpatrick , L. Legrand , A. Loureiro , M. Magliocchetti , G. Mainetti , R. Maoli , M. Martinelli , C. J. A. P. Martins , S. Matthew , L. Maurin , P. Monaco , G. Morgante , S. Nadathur , A. A. Nucita , M. Pöntinen , L. Patrizii , V. Popa , C. Porciani , D. Potter , P. Reimberg , A. G. Sánchez , Z. Sakr , A. Schneider , M. Sereno , P. Simon , A. Spurio Mancini , J. Stadel , J. Steinwagner , R. Teyssier , J. Valiviita , M. Viel , I. A. Zinchenko , H. Domínguez Sánchez

Low-light image enhancement is challenging due to complex degradations, including amplified noise, artifacts, and color distortion. While Retinex-based deep learning methods have achieved promising results, they primarily rely on…

Computer Vision and Pattern Recognition · Computer Science 2026-05-14 Youssef Aboelwafa , Hicham G. Elmongui , Marwan Torki

Recently, the attention-enhanced multi-layer encoder, such as Transformer, has been extensively studied in Machine Reading Comprehension (MRC). To predict the answer, it is common practice to employ a predictor to draw information only from…

Computation and Language · Computer Science 2021-02-03 Nuo Chen , Fenglin Liu , Chenyu You , Peilin Zhou , Yuexian Zou

Nowadays, many applications rely on images of high quality to ensure good performance in conducting their tasks. However, noise goes against this objective as it is an unavoidable issue in most applications. Therefore, it is essential to…

Computer Vision and Pattern Recognition · Computer Science 2017-04-20 Ahmed Ben Said , Rachid Hadjidj , Kamel Eddine Melkemi , Sebti Foufou

Deep generative models have been applied to multiple applications in image-to-image translation. Generative Adversarial Networks and Diffusion Models have presented impressive results, setting new state-of-the-art results on these tasks.…

Computer Vision and Pattern Recognition · Computer Science 2024-02-26 Sagar Saxena , Mohammad Nayeem Teli