中文
相关论文

相关论文: NTIRE 2025 challenge on Text to Image Generation M…

200 篇论文

With the evolution of Text-to-Image (T2I) models, the quality defects of AI-Generated Images (AIGIs) pose a significant barrier to their widespread adoption. In terms of both perception and alignment, existing models cannot always guarantee…

This paper presents a comprehensive review of the NTIRE 2025 Low-Light Image Enhancement (LLIE) Challenge, highlighting the proposed solutions and final outcomes. The objective of the challenge is to identify effective networks capable of…

计算机视觉与模式识别 · 计算机科学 2025-10-16 Xiaoning Liu , Zongwei Wu , Florin-Alexandru Vasluianu , Hailong Yan , Bin Ren , Yulun Zhang , Shuhang Gu , Le Zhang , Ce Zhu , Radu Timofte , Kangbiao Shi , Yixu Feng , Tao Hu , Yu Cao , Peng Wu , Yijin Liang , Yanning Zhang , Qingsen Yan , Han Zhou , Wei Dong , Yan Min , Mohab Kishawy , Jun Chen , Pengpeng Yu , Anjin Park , Seung-Soo Lee , Young-Joon Park , Zixiao Hu , Junyv Liu , Huilin Zhang , Jun Zhang , Fei Wan , Bingxin Xu , Hongzhe Liu , Cheng Xu , Weiguo Pan , Songyin Dai , Xunpeng Yi , Qinglong Yan , Yibing Zhang , Jiayi Ma , Changhui Hu , Kerui Hu , Donghang Jing , Tiesheng Chen , Zhi Jin , Hongjun Wu , Biao Huang , Haitao Ling , Jiahao Wu , Dandan Zhan , G Gyaneshwar Rao , Vijayalaxmi Ashok Aralikatti , Nikhil Akalwadi , Ramesh Ashok Tabib , Uma Mudenagudi , Ruirui Lin , Guoxi Huang , Nantheera Anantrasirichai , Qirui Yang , Alexandru Brateanu , Ciprian Orhei , Cosmin Ancuti , Daniel Feijoo , Juan C. Benito , Álvaro García , Marcos V. Conde , Yang Qin , Raul Balmez , Anas M. Ali , Bilel Benjdira , Wadii Boulila , Tianyi Mao , Huan Zheng , Yanyan Wei , Shengeng Tang , Dan Guo , Zhao Zhang , Sabari Nathan , K Uma , A Sasithradevi , B Sathya Bama , S. Mohamed Mansoor Roomi , Ao Li , Xiangtao Zhang , Zhe Liu , Yijie Tang , Jialong Tang , Zhicheng Fu , Gong Chen , Joe Nasti , John Nicholson , Zeyu Xiao , Zhuoyuan Li , Ashutosh Kulkarni , Prashant W. Patil , Santosh Kumar Vipparthi , Subrahmanyam Murala , Duan Liu , Weile Li , Hangyuan Lu , Rixian Liu , Tengfeng Wang , Jinxing Liang , Chenxin Yu

This paper reviews the NTIRE 2024 Portrait Quality Assessment Challenge, highlighting the proposed solutions and results. This challenge aims to obtain an efficient deep neural network capable of estimating the perceptual quality of real…

计算机视觉与模式识别 · 计算机科学 2024-04-18 Nicolas Chahine , Marcos V. Conde , Daniela Carfora , Gabriel Pacianotto , Benoit Pochon , Sira Ferradans , Radu Timofte

This paper presents an overview of the NTIRE 2025 Challenge on UGC Video Enhancement. The challenge constructed a set of 150 user-generated content videos without reference ground truth, which suffer from real-world degradations such as…

We provide a new multi-task benchmark for evaluating text-to-image models. We perform a human evaluation comparing the most common open-source (Stable Diffusion) and commercial (DALL-E 2) models. Twenty computer science AI graduate students…

Text-and-Image-To-Image (TI2I), an extension of Text-To-Image (T2I), integrates image inputs with textual instructions to enhance image generation. Existing methods often partially utilize image inputs, focusing on specific elements like…

计算机视觉与模式识别 · 计算机科学 2025-03-20 Teng-Fang Hsiao , Bo-Kai Ruan , Yi-Lun Wu , Tzu-Ling Lin , Hong-Han Shuai

Text-to-image generation and text-guided image manipulation have received considerable attention in the field of image generation tasks. However, the mainstream evaluation methods for these tasks have difficulty in evaluating whether all…

计算机视觉与模式识别 · 计算机科学 2024-11-18 Mizuki Miyamoto , Ryugo Morita , Jinjia Zhou

In this paper, we present a comprehensive overview of the NTIRE 2026 3rd Restore Any Image Model (RAIM) challenge, with a specific focus on Track 3: AI Flash Portrait. Despite significant advancements in deep learning for image restoration,…

This paper provides a review of the NTIRE 2026 challenge on mobile real-world image super-resolution, highlighting the proposed solutions and the resulting outcomes. The challenge aims to recover high-resolution (HR) images from…

This paper presents a comprehensive review of the 1st Challenge on Video Quality Enhancement for Video Conferencing held at the NTIRE workshop at CVPR 2025, and highlights the problem statement, datasets, proposed solutions, and results.…

计算机视觉与模式识别 · 计算机科学 2025-05-27 Varun Jain , Zongwei Wu , Quan Zou , Louis Florentin , Henrik Turbell , Sandeep Siddhartha , Radu Timofte , others

Despite remarkable progress in Text-to-Image models, many real-world applications require generating coherent image sets with diverse consistency requirements. Existing consistent methods often focus on a specific domain with specific…

计算机视觉与模式识别 · 计算机科学 2025-09-26 Chengyou Jia , Xin Shen , Zhuohang Dang , Zhuohang Dang , Changliang Xia , Weijia Wu , Xinyu Zhang , Hangwei Qian , Ivor W. Tsang , Minnan Luo

In this paper, we present an overview of the NTIRE 2026 challenge on the 3rd Restore Any Image Model in the Wild, specifically focusing on Track 1: Professional Image Quality Assessment. Conventional Image Quality Assessment (IQA) typically…

Text-to-image generative models excel in creating images from text but struggle with ensuring alignment and consistency between outputs and prompts. This paper introduces TextMatch, a novel framework that leverages multimodal optimization…

计算机视觉与模式识别 · 计算机科学 2025-01-28 Yucong Luo , Mingyue Cheng , Jie Ouyang , Xiaoyu Tao , Qi Liu

Diffusion models have revitalized the image generation domain, playing crucial roles in both academic research and artistic expression. With the emergence of new diffusion models, assessing the performance of text-to-image models has become…

计算机视觉与模式识别 · 计算机科学 2024-11-15 Chutian Meng , Fan Ma , Jiaxu Miao , Chi Zhang , Yi Yang , Yueting Zhuang

Despite significant progress in generative AI, comprehensive evaluation remains challenging because of the lack of effective metrics and standardized benchmarks. For instance, the widely-used CLIPScore measures the alignment between a…

计算机视觉与模式识别 · 计算机科学 2024-06-19 Zhiqiu Lin , Deepak Pathak , Baiqi Li , Jiayao Li , Xide Xia , Graham Neubig , Pengchuan Zhang , Deva Ramanan

As large language models have demonstrated impressive performance in many domains, recent works have adopted language models (LMs) as controllers of visual modules for vision-and-language tasks. While existing work focuses on equipping LMs…

计算机视觉与模式识别 · 计算机科学 2023-10-30 Jaemin Cho , Abhay Zala , Mohit Bansal

This paper reviews the NTIRE 2025 RAW Image Restoration and Super-Resolution Challenge, highlighting the proposed solutions and results. New methods for RAW Restoration and Super-Resolution could be essential in modern Image Signal…

Motion blur is a common photography artifact in dynamic environments that typically comes jointly with the other types of degradation. This paper reviews the NTIRE 2021 Challenge on Image Deblurring. In this challenge report, we describe…

计算机视觉与模式识别 · 计算机科学 2021-05-03 Seungjun Nah , Sanghyun Son , Suyoung Lee , Radu Timofte , Kyoung Mu Lee

This paper reviews the NTIRE 2020 challenge on real image denoising with focus on the newly introduced dataset, the proposed methods and their results. The challenge is a new version of the previous NTIRE 2019 challenge on real image…

With the rapid advancement of large multimodal models (LMMs), recent text-to-image (T2I) models can generate high-quality images and demonstrate great alignment to short prompts. However, they still struggle to effectively understand and…

计算机视觉与模式识别 · 计算机科学 2025-10-06 Juntong Wang , Huiyu Duan , Jiarui Wang , Ziheng Jia , Guangtao Zhai , Xiongkuo Min