中文
相关论文

相关论文: Room for improvement in automatic image descriptio…

200 篇论文

Automatic description generation from natural images is a challenging problem that has recently received a large amount of interest from the computer vision and natural language processing communities. In this survey, we classify the…

Automatic image aesthetics assessment is a computer vision problem dealing with categorizing images into different aesthetic levels. The categorization is usually done by analyzing an input image and computing some measure of the degree to…

计算机视觉与模式识别 · 计算机科学 2022-02-08 Abbas Anwar , Saira Kanwal , Muhammad Tahir , Muhammad Saqib , Muhammad Uzair , Mohammad Khalid Imam Rahmani , Habib Ullah

Automatic image description systems are commonly trained and evaluated using crowdsourced, human-generated image descriptions. The best-performing system is then determined using some measure of similarity to the reference data (BLEU,…

计算与语言 · 计算机科学 2020-06-17 Emiel van Miltenburg

Automatically generating a human-like description for a given image is a potential research in artificial intelligence, which has attracted a great of attention recently. Most of the existing attention methods explore the mapping…

计算机视觉与模式识别 · 计算机科学 2020-11-03 Feicheng Huang , Zhixin Li , Haiyang Wei , Canlong Zhang , Huifang Ma

Text-to-image models often struggle to generate images that precisely match textual prompts. Prior research has extensively studied the evaluation of image-text alignment in text-to-image generation. However, existing evaluations primarily…

计算与语言 · 计算机科学 2025-06-11 Huixuan Zhang , Xiaojun Wan

Attention mechanisms have recently been introduced in deep learning for various tasks in natural language processing and computer vision. But despite their popularity, the "correctness" of the implicitly-learned attention maps has only been…

计算机视觉与模式识别 · 计算机科学 2016-11-24 Chenxi Liu , Junhua Mao , Fei Sha , Alan Yuille

A major challenge in the field of Text Generation is evaluation: Human evaluations are cost-intensive, and automated metrics often display considerable disagreement with human judgments. In this paper, we propose a statistical model of Text…

计算与语言 · 计算机科学 2023-06-07 Jan Deriu , Pius von Däniken , Don Tuggener , Mark Cieliebak

Image captioning involves generating textual descriptions from input images, bridging the gap between computer vision and natural language processing. Recent advancements in transformer-based models have significantly improved caption…

计算机视觉与模式识别 · 计算机科学 2025-06-09 Israa A. Albadarneh , Bassam H. Hammo , Omar S. Al-Kadi

Illustrations are widely used in education, and sometimes, alternatives are not available for visually impaired students. Therefore, those students would benefit greatly from an automatic illustration description system, but only if those…

人机交互 · 计算机科学 2021-06-30 Anett Hoppe , David Morris , Ralph Ewerth

While the ImageNet dataset has been driving computer vision research over the past decade, significant label noise and ambiguity have made top-1 accuracy an insufficient measure of further progress. To address this, new label-sets and…

计算机视觉与模式识别 · 计算机科学 2024-01-08 Momchil Peychev , Mark Niklas Müller , Marc Fischer , Martin Vechev

For some images, descriptions written by multiple people are consistent with each other. But for other images, descriptions across people vary considerably. In other words, some images are specific $-$ they elicit consistent descriptions…

计算机视觉与模式识别 · 计算机科学 2015-04-17 Mainak Jas , Devi Parikh

Automatically generating a natural language description of an image has attracted interests recently both because of its importance in practical applications and because it connects two major artificial intelligence fields: computer vision…

计算机视觉与模式识别 · 计算机科学 2016-03-15 Quanzeng You , Hailin Jin , Zhaowen Wang , Chen Fang , Jiebo Luo

The widespread adoption of automatic sentiment and emotion classifiers makes it important to ensure that these tools perform reliably across different populations. Yet their reliability is typically assessed using benchmarks that rely on…

计算与语言 · 计算机科学 2026-01-09 Ivan Smirnov , Segun T. Aroyehun , Paul Plener , David Garcia

Recently, a myriad of conditional image generation and editing models have been developed to serve different downstream tasks, including text-to-image generation, text-guided image editing, subject-driven image generation, control-guided…

计算机视觉与模式识别 · 计算机科学 2024-03-12 Max Ku , Tianle Li , Kai Zhang , Yujie Lu , Xingyu Fu , Wenwen Zhuang , Wenhu Chen

Rapid progress in text-to-image generative models coupled with their deployment for visual content creation has magnified the importance of thoroughly evaluating their performance and identifying potential biases. In pursuit of models that…

计算机视觉与模式识别 · 计算机科学 2024-05-08 Melissa Hall , Samuel J. Bell , Candace Ross , Adina Williams , Michal Drozdzal , Adriana Romero Soriano

While the importance of automatic image analysis is continuously increasing, recent meta-research revealed major flaws with respect to algorithm validation. Performance metrics are particularly key for meaningful, objective, and transparent…

图像与视频处理 · 电气工程与系统科学 2023-12-08 Annika Reinke , Minu D. Tizabi , Carole H. Sudre , Matthias Eisenmann , Tim Rädsch , Michael Baumgartner , Laura Acion , Michela Antonelli , Tal Arbel , Spyridon Bakas , Peter Bankhead , Arriel Benis , Matthew Blaschko , Florian Buettner , M. Jorge Cardoso , Jianxu Chen , Veronika Cheplygina , Evangelia Christodoulou , Beth Cimini , Gary S. Collins , Sandy Engelhardt , Keyvan Farahani , Luciana Ferrer , Adrian Galdran , Bram van Ginneken , Ben Glocker , Patrick Godau , Robert Haase , Fred Hamprecht , Daniel A. Hashimoto , Doreen Heckmann-Nötzel , Peter Hirsch , Michael M. Hoffman , Merel Huisman , Fabian Isensee , Pierre Jannin , Charles E. Kahn , Dagmar Kainmueller , Bernhard Kainz , Alexandros Karargyris , Alan Karthikesalingam , A. Emre Kavur , Hannes Kenngott , Jens Kleesiek , Andreas Kleppe , Sven Kohler , Florian Kofler , Annette Kopp-Schneider , Thijs Kooi , Michal Kozubek , Anna Kreshuk , Tahsin Kurc , Bennett A. Landman , Geert Litjens , Amin Madani , Klaus Maier-Hein , Anne L. Martel , Peter Mattson , Erik Meijering , Bjoern Menze , David Moher , Karel G. M. Moons , Henning Müller , Brennan Nichyporuk , Felix Nickel , M. Alican Noyan , Jens Petersen , Gorkem Polat , Susanne M. Rafelski , Nasir Rajpoot , Mauricio Reyes , Nicola Rieke , Michael Riegler , Hassan Rivaz , Julio Saez-Rodriguez , Clara I. Sánchez , Julien Schroeter , Anindo Saha , M. Alper Selver , Lalith Sharan , Shravya Shetty , Maarten van Smeden , Bram Stieltjes , Ronald M. Summers , Abdel A. Taha , Aleksei Tiulpin , Sotirios A. Tsaftaris , Ben Van Calster , Gaël Varoquaux , Manuel Wiesenfarth , Ziv R. Yaniv , Paul Jäger , Lena Maier-Hein

Visual aesthetic assessment has been an active research field for decades. Although latest methods have achieved promising performance on benchmark datasets, they typically rely on a large number of manual annotations including both…

计算机视觉与模式识别 · 计算机科学 2019-12-04 Kekai Sheng , Weiming Dong , Menglei Chai , Guohui Wang , Peng Zhou , Feiyue Huang , Bao-Gang Hu , Rongrong Ji , Chongyang Ma

Personalized image generation via text prompts has great potential to improve daily life and professional work by facilitating the creation of customized visual content. The aim of image personalization is to create images based on a…

计算机视觉与模式识别 · 计算机科学 2025-03-11 Mingxiao Li , Tingyu Qu , Tinne Tuytelaars , Marie-Francine Moens

Text-to-image models, which can generate high-quality images based on textual input, have recently enabled various content-creation tools. Despite significantly affecting a wide range of downstream applications, the distributions of these…

计算机视觉与模式识别 · 计算机科学 2025-05-27 Yanzhe Zhang , Lu Jiang , Greg Turk , Diyi Yang

Editing images using natural language instructions has become a natural and expressive way to modify visual content; yet, evaluating the performance of such models remains challenging. Existing evaluation approaches often rely on image-text…

计算机视觉与模式识别 · 计算机科学 2025-07-28 Yusu Qian , Jiasen Lu , Tsu-Jui Fu , Xinze Wang , Chen Chen , Yinfei Yang , Wenze Hu , Zhe Gan
‹ 上一页 1 2 3 10 下一页 ›