English
Related papers

Related papers: Is Image-to-Image Translation the Panacea for Mult…

200 papers

Image-to-image translation is a general name for a task where an image from one domain is converted to a corresponding image in another domain, given sufficient training data. Traditionally different approaches have been proposed depending…

Computer Vision and Pattern Recognition · Computer Science 2018-05-09 Soumya Tripathy , Juho Kannala , Esa Rahtu

Unsupervised image-to-image (I2I) translation learns cross-domain image mapping that transfers input from the source domain to output in the target domain while preserving its semantics. One challenge is that different semantic statistics…

Computer Vision and Pattern Recognition · Computer Science 2023-10-10 Ganning Zhao , Wenhui Cui , Suya You , C. -C. Jay Kuo

Since thermal imagery offers a unique modality to investigate pain, the U.S. National Institutes of Health (NIH) has collected a large and diverse set of cancer patient facial thermograms for AI-based pain research. However, differing…

Computer Vision and Pattern Recognition · Computer Science 2023-08-24 Catherine Ordun , Alexandra Cha , Edward Raff , Sanjay Purushotham , Karen Kwok , Mason Rule , James Gulley

On shopping websites, product images of low quality negatively affect customer experience. Although there are plenty of work in detecting images with different defects, few efforts have been dedicated to correct those defects at scale. A…

Computer Vision and Pattern Recognition · Computer Science 2023-09-13 Moyan Li , Jinmiao Fu , Shaoyuan Xu , Huidong Liu , Jia Liu , Bryan Wang

Image registration and in particular deformable registration methods are pillars of medical imaging. Inspired by the recent advances in deep learning, we propose in this paper, a novel convolutional neural network architecture that couples…

Computer Vision and Pattern Recognition · Computer Science 2018-09-18 Stergios Christodoulidis , Mihir Sahasrabudhe , Maria Vakalopoulou , Guillaume Chassagnon , Marie-Pierre Revel , Stavroula Mougiakakou , Nikos Paragios

Recently, creative generative artificial intelligence software has emerged as a pivotal assistant, enabling users to generate content and seek inspiration rapidly. Text-to-Image (T2I) software, one of the most widely used, synthesizes…

Software Engineering · Computer Science 2025-01-14 Siqi Gu , Chunrong Fang , Quanjun Zhang , Zhenyu Chen

Deep generative models have been applied to multiple applications in image-to-image translation. Generative Adversarial Networks and Diffusion Models have presented impressive results, setting new state-of-the-art results on these tasks.…

Computer Vision and Pattern Recognition · Computer Science 2024-02-26 Sagar Saxena , Mohammad Nayeem Teli

Image Registration (IR) is the process of aligning two (or more) images of the same scene taken at different times, different viewpoints and/or by different sensors. It is an important, crucial step in various image analysis tasks where…

Computer Vision and Pattern Recognition · Computer Science 2017-11-21 Sarit Chicotay , Eli David , Nathan S. Netanyahu

The dominant probing approaches rely on the zero-shot performance of image-text matching tasks to gain a finer-grained understanding of the representations learned by recent multimodal image-language transformer models. The evaluation is…

Computation and Language · Computer Science 2024-01-31 Ivana Beňová , Jana Košecká , Michal Gregor , Martin Tamajka , Marcel Veselý , Marián Šimko

Image-to-image translation (I2IT) refers to the process of transforming images from a source domain to a target domain while maintaining a fundamental connection in terms of image content. In the past few years, remarkable advancements in…

Computer Vision and Pattern Recognition · Computer Science 2023-12-04 Or Greenberg , Eran Kishon , Dani Lischinski

Image-to-point cloud (I2P) registration is a fundamental task for robots and autonomous vehicles to achieve cross-modality data fusion and localization. Current I2P registration methods primarily focus on estimating correspondences at the…

Computer Vision and Pattern Recognition · Computer Science 2024-09-13 Shuhao Kang , Youqi Liao , Jianping Li , Fuxun Liang , Yuhao Li , Xianghong Zou , Fangning Li , Xieyuanli Chen , Zhen Dong , Bisheng Yang

Healthcare applications are inherently multimodal, benefiting greatly from the integration of diverse data sources. However, the modalities available in clinical settings can vary across different locations and patients. A key area that…

Computer Vision and Pattern Recognition · Computer Science 2025-09-04 Mohammed Amer , Mohamed A. Suliman , Tu Bui , Nuria Garcia , Serban Georgescu

Robust and accurate alignment of multimodal medical images is a very challenging task, which however is very useful for many clinical applications. For example, magnetic resonance (MR) and transrectal ultrasound (TRUS) image registration is…

Computer Vision and Pattern Recognition · Computer Science 2018-10-03 Pingkun Yan , Sheng Xu , Ardeshir R. Rastinehad , Brad J. Wood

Deep Learning has implemented a wide range of applications and has become increasingly popular in recent years. The goal of multimodal deep learning is to create models that can process and link information using various modalities. Despite…

Computer Vision and Pattern Recognition · Computer Science 2021-05-25 Jabeen Summaira , Xi Li , Amin Muhammad Shoib , Songyuan Li , Jabbar Abdul

Conventional deformable registration methods aim at solving an optimization model carefully designed on image pairs and their computational costs are exceptionally high. In contrast, recent deep learning based approaches can provide fast…

Computer Vision and Pattern Recognition · Computer Science 2021-10-01 Risheng Liu , Zi Li , Xin Fan , Chenying Zhao , Hao Huang , Zhongxuan Luo

Interactive machine learning (IML) allows users to build their custom machine learning models without expert knowledge. While most existing IML systems are designed with classification algorithms, they sometimes oversimplify the…

Human-Computer Interaction · Computer Science 2024-04-16 Wataru Kawabe , Yusuke Sugano

We introduce a deep encoder-decoder architecture for image deformation prediction from multimodal images. Specifically, we design an image-patch-based deep network that jointly (i) learns an image similarity measure and (ii) the…

Computer Vision and Pattern Recognition · Computer Science 2017-04-03 Xiao Yang , Roland Kwitt , Martin Styner , Marc Niethammer

In this study, we aim to enhance the capabilities of diffusion-based text-to-image (T2I) generation models by integrating diverse modalities beyond textual descriptions within a unified framework. To this end, we categorize widely used…

Computer Vision and Pattern Recognition · Computer Science 2025-08-27 Sungnyun Kim , Junsoo Lee , Kibeom Hong , Daesik Kim , Namhyuk Ahn

Multimodal MR image synthesis aims to generate missing modality images by effectively fusing and mapping from a subset of available MRI modalities. Most existing methods adopt an image-to-image translation paradigm, treating multiple…

Image and Video Processing · Electrical Eng. & Systems 2025-04-29 Tao Song , Yicheng Wu , Minhao Hu , Xiangde Luo , Linda Wei , Guotai Wang , Yi Guo , Feng Xu , Shaoting Zhang

Recently image-to-image translation has received increasing attention, which aims to map images in one domain to another specific one. Existing methods mainly solve this task via a deep generative model, and focus on exploring the…

Computer Vision and Pattern Recognition · Computer Science 2019-01-24 Songyao Jiang , Zhiqiang Tao , Yun Fu
‹ Prev 1 8 9 10 Next ›