English
Related papers

Related papers: Image Realness Assessment and Localization with Mu…

200 papers

Computational models have emerged as powerful tools for multi-scale energy modeling research at the building and urban scale, supporting data-driven analysis across building and urban energy systems. However, these models require large…

Artificial Intelligence · Computer Science 2026-04-09 Jackson Eshbaugh , Chetan Tiwari , Jorge Silveyra

The misuse of AI imagery can have harmful societal effects, prompting the creation of detectors to combat issues like the spread of fake news. Existing methods can effectively detect images generated by seen generators, but it is…

Computer Vision and Pattern Recognition · Computer Science 2023-12-15 Mingjian Zhu , Hanting Chen , Mouxiao Huang , Wei Li , Hailin Hu , Jie Hu , Yunhe Wang

This paper investigates the inverse capabilities and broader utility of multimodal latent spaces within task-specific AI (Artificial Intelligence) models. While these models excel at their designed forward tasks (e.g., text-to-image…

Machine Learning · Computer Science 2025-08-01 Siwoo Park

Product images strongly influence consumer decision-making in online marketplaces. Empowered by multimodal contrastive learning, generative AI can output images that closely align with text prompts. Yet existing generative AI models do not…

Artificial Intelligence · Computer Science 2026-05-28 Xiaohang Feng , Yiling Xie

Purpose: Optical imaging is evolving as a key technique for advanced sensing in the operating room. Recent research has shown that machine learning algorithms can be used to address the inverse problem of converting pixel-wise multispectral…

Artificial Intelligence (AI)-generated images have become increasingly realistic and readily adaptable to concrete real-world claims, creating new challenges for verifying visual evidence. A concrete emerging risk is AI-generated refund…

Computer Vision and Pattern Recognition · Computer Science 2026-05-12 Xinyu Yan , Boyang Chen , Jiaming Zhang , Tiantong Wu , Hong Xi Tae , Yichen He , Tiantong Wang , Yachun Mi , Yurong Hao , Yilei Zhao , Lei Xiao , Longtao Huang , Pengjun Xie , Wei Liu , Wei Yang Bryan Lim

Multi-modal medical image segmentation plays an essential role in clinical diagnosis. It remains challenging as the input modalities are often not well-aligned spatially. Existing learning-based methods mainly consider sharing trainable…

Computer Vision and Pattern Recognition · Computer Science 2021-01-06 Jingkun Chen , Wenqi Li , Hongwei Li , Jianguo Zhang

Generative AI has increasingly been used for artistic creation, but little work has explored how it shapes the experiential meaning of practice. We consider how generative AI might transform the embodied and tangible process of instant…

Human-Computer Interaction · Computer Science 2026-05-05 Michael Yin , Angela Chiang , Robert Xiao

Non-photorealistic rendering techniques work on image features and often manipulate a set of characteristics such as edges and texture to achieve a desired depiction of the scene. Most computational photography methods decompose an image…

Computer Vision and Pattern Recognition · Computer Science 2016-04-20 Akshay Gadi Patil , Shanmuganathan Raman

Our goal is to develop stable, accurate, and robust semantic scene understanding methods for wide-area scene perception and understanding, especially in challenging outdoor environments. To achieve this, we are exploring and evaluating a…

Computer Vision and Pattern Recognition · Computer Science 2021-04-01 Jiesi Hu , Ganning Zhao , Suya You , C. C. Jay Kuo

The recent development of generative models unleashes the potential of generating hyper-realistic fake images. To prevent the malicious usage of fake images, AI-generated image detection aims to distinguish fake images from real images.…

Computer Vision and Pattern Recognition · Computer Science 2024-04-23 Jiaxuan Chen , Jieteng Yao , Li Niu

Generating natural questions from an image is a semantic task that requires using visual and language modality to learn multimodal representations. Images can have multiple visual and language contexts that are relevant for generating…

Computation and Language · Computer Science 2019-10-18 Badri N. Patro , Sandeep Kumar , Vinod K. Kurmi , Vinay P. Namboodiri

Fine-grained image-text alignment is a pivotal challenge in multimodal learning, underpinning key applications such as visual question answering, image captioning, and vision-language navigation. Unlike global alignment, fine-grained…

Computer Vision and Pattern Recognition · Computer Science 2025-12-02 Jiale Liu , Haoming Zhou , Yishu Liu , Bingzhi Chen , Yuncheng Jiang

One fundamental task of multimodal models is to translate referred image regions to human preferred language descriptions. Existing methods, however, ignore the resolution adaptability needs of different tasks, which hinders them to find…

Computer Vision and Pattern Recognition · Computer Science 2025-03-04 Yuzhong Zhao , Feng Liu , Yue Liu , Mingxiang Liao , Chen Gong , Qixiang Ye , Fang Wan

Advances in generative models have created Artificial Intelligence-Generated Images (AIGIs) nearly indistinguishable from real photographs. Leveraging a large corpus of 30,824 AIGIs collected from Instagram and Twitter, and combining…

Computers and Society · Computer Science 2025-03-19 Qiyao Peng , Yingdan Lu , Yilang Peng , Sijia Qian , Xinyi Liu , Cuihua Shen

Progress in image generation raises significant public security concerns. We argue that fake image detection should not operate as a "black box". Instead, an ideal approach must ensure both strong generalization and transparency. Recent…

Computer Vision and Pattern Recognition · Computer Science 2025-11-10 Yikun Ji , Yan Hong , Jiahui Zhan , Haoxing Chen , jun lan , Huijia Zhu , Weiqiang Wang , Liqing Zhang , Jianfu Zhang

Omnidirectional and 360{\deg} images are becoming widespread in industry and in consumer society, causing omnidirectional computer vision to gain attention. Their wide field of view allows the gathering of a great amount of information…

Computer Vision and Pattern Recognition · Computer Science 2024-01-31 Bruno Berenguel-Baeta , Jesus Bermudez-Cameo , Jose J. Guerrero

Cross-resolution image alignment is a key problem in multiscale gigapixel photography, which requires to estimate homography matrix using images with large resolution gap. Existing deep homography methods concatenate the input images or…

Computer Vision and Pattern Recognition · Computer Science 2021-06-15 Ruizhi Shao , Gaochang Wu , Yuemei Zhou , Ying Fu , Yebin Liu

Current AI-Generated Image (AIGI) detection approaches predominantly rely on binary classification to distinguish real from synthetic images, often lacking interpretable or convincing evidence to substantiate their decisions. This…

Computer Vision and Pattern Recognition · Computer Science 2026-01-28 Yao Xiao , Weiyan Chen , Jiahao Chen , Zijie Cao , Weijian Deng , Binbin Yang , Ziyi Dong , Xiangyang Ji , Wei Ke , Pengxu Wei , Liang Lin

Supervised and unsupervised homography estimation methods depend on image pairs tailored to specific modalities to achieve high accuracy. However, their performance deteriorates substantially when applied to unseen modalities. To address…

Computer Vision and Pattern Recognition · Computer Science 2026-03-05 Jinkun You , Jiaxin Cheng , Jie Zhang , Yicong Zhou