中文
相关论文

相关论文: Position: Evaluation of Visual Processing Should B…

200 篇论文

Automatic image description systems are commonly trained and evaluated using crowdsourced, human-generated image descriptions. The best-performing system is then determined using some measure of similarity to the reference data (BLEU,…

计算与语言 · 计算机科学 2020-06-17 Emiel van Miltenburg

Human pose estimation, with its broad applications in action recognition and motion capture, has experienced significant advancements. However, current Transformer-based methods for video pose estimation often face challenges in managing…

计算机视觉与模式识别 · 计算机科学 2025-03-10 Zhigang Wang , Shaojing Fan , Zhenguang Liu , Zheqi Wu , Sifan Wu , Yingying Jiao

The rapid advancement of Generative Artificial Intelligence (AI), such as Large Language Models (LLMs) and Multimodal Large Language Models (MLLM), has the potential to revolutionize the way we work and interact with digital systems across…

人机交互 · 计算机科学 2024-05-28 Carlos Toxtli

We propose a task we name Portrait Interpretation and construct a dataset named Portrait250K for it. Current researches on portraits such as human attribute recognition and person re-identification have achieved many successes, but…

计算机视觉与模式识别 · 计算机科学 2022-07-28 Yixuan Fan , Zhaopeng Dou , Yali Li , Shengjin Wang

In the rapidly advancing field of artificial intelligence, machine perception is becoming paramount to achieving increased performance. Image classification systems are becoming increasingly integral to various applications, ranging from…

计算机视觉与模式识别 · 计算机科学 2024-12-18 Javon Hickmon

Human visual reasoning is characterized by an ability to identify abstract patterns from only a small number of examples, and to systematically generalize those patterns to novel inputs. This capacity depends in large part on our ability to…

计算机视觉与模式识别 · 计算机科学 2023-11-14 Taylor W. Webb , Shanka Subhra Mondal , Jonathan D. Cohen

On the way towards general Visual Question Answering (VQA) systems that are able to answer arbitrary questions, the need arises for evaluation beyond single-metric leaderboards for specific datasets. To this end, we propose a browser-based…

计算机视觉与模式识别 · 计算机科学 2021-10-12 Dirk Väth , Pascal Tilli , Ngoc Thang Vu

Machine learning systems are increasingly deployed in high-stakes domains, yet they remain vulnerable to bias systematic disparities that disproportionately impact specific demographic groups. Traditional bias detection methods often depend…

机器学习 · 计算机科学 2025-06-16 Chirudeep Tupakula , Rittika Shamsuddin

Although we have seen a proliferation of algorithms for recommending visualizations, these algorithms are rarely compared with one another, making it difficult to ascertain which algorithm is best for a given visual analysis scenario.…

人机交互 · 计算机科学 2021-09-08 Zehua Zeng , Phoebe Moh , Fan Du , Jane Hoffswell , Tak Yeon Lee , Sana Malik , Eunyee Koh , Leilani Battle

The proliferation of DeepFake technology is a rising challenge in today's society, owing to more powerful and accessible generation methods. To counter this, the research community has developed detectors of ever-increasing accuracy.…

计算机视觉与模式识别 · 计算机科学 2022-10-10 Federico Baldassarre , Quentin Debard , Gonzalo Fiz Pontiveros , Tri Kurniawan Wijaya

This paper revisits the role of quantitative and qualitative methods in visualization research in the context of advancements in artificial intelligence (AI). The focus is on how we can bridge between the different methods in an integrated…

人机交互 · 计算机科学 2024-09-12 Daniel Weiskopf

While the importance of automatic image analysis is continuously increasing, recent meta-research revealed major flaws with respect to algorithm validation. Performance metrics are particularly key for meaningful, objective, and transparent…

图像与视频处理 · 电气工程与系统科学 2023-12-08 Annika Reinke , Minu D. Tizabi , Carole H. Sudre , Matthias Eisenmann , Tim Rädsch , Michael Baumgartner , Laura Acion , Michela Antonelli , Tal Arbel , Spyridon Bakas , Peter Bankhead , Arriel Benis , Matthew Blaschko , Florian Buettner , M. Jorge Cardoso , Jianxu Chen , Veronika Cheplygina , Evangelia Christodoulou , Beth Cimini , Gary S. Collins , Sandy Engelhardt , Keyvan Farahani , Luciana Ferrer , Adrian Galdran , Bram van Ginneken , Ben Glocker , Patrick Godau , Robert Haase , Fred Hamprecht , Daniel A. Hashimoto , Doreen Heckmann-Nötzel , Peter Hirsch , Michael M. Hoffman , Merel Huisman , Fabian Isensee , Pierre Jannin , Charles E. Kahn , Dagmar Kainmueller , Bernhard Kainz , Alexandros Karargyris , Alan Karthikesalingam , A. Emre Kavur , Hannes Kenngott , Jens Kleesiek , Andreas Kleppe , Sven Kohler , Florian Kofler , Annette Kopp-Schneider , Thijs Kooi , Michal Kozubek , Anna Kreshuk , Tahsin Kurc , Bennett A. Landman , Geert Litjens , Amin Madani , Klaus Maier-Hein , Anne L. Martel , Peter Mattson , Erik Meijering , Bjoern Menze , David Moher , Karel G. M. Moons , Henning Müller , Brennan Nichyporuk , Felix Nickel , M. Alican Noyan , Jens Petersen , Gorkem Polat , Susanne M. Rafelski , Nasir Rajpoot , Mauricio Reyes , Nicola Rieke , Michael Riegler , Hassan Rivaz , Julio Saez-Rodriguez , Clara I. Sánchez , Julien Schroeter , Anindo Saha , M. Alper Selver , Lalith Sharan , Shravya Shetty , Maarten van Smeden , Bram Stieltjes , Ronald M. Summers , Abdel A. Taha , Aleksei Tiulpin , Sotirios A. Tsaftaris , Ben Van Calster , Gaël Varoquaux , Manuel Wiesenfarth , Ziv R. Yaniv , Paul Jäger , Lena Maier-Hein

Existing Visual Question Answering (VQA) methods tend to exploit dataset biases and spurious statistical correlations, instead of producing right answers for the right reasons. To address this issue, recent bias mitigation methods for VQA…

计算机视觉与模式识别 · 计算机科学 2024-04-24 Robik Shrestha , Kushal Kafle , Christopher Kanan

Classical clustering methods do not provide users with direct control of the clustering results, and the clustering results may not be consistent with the relevant criterion that a user has in mind. In this work, we present a new…

计算机视觉与模式识别 · 计算机科学 2024-02-23 Sehyun Kwon , Jaeseung Park , Minkyu Kim , Jaewoong Cho , Ernest K. Ryu , Kangwook Lee

The predominant approach to Visual Question Answering (VQA) demands that the model represents within its weights all of the information required to answer any question about any image. Learning this information from any real training set…

计算机视觉与模式识别 · 计算机科学 2017-11-23 Damien Teney , Anton van den Hengel

Image editing and compositing have become ubiquitous in entertainment, from digital art to AR and VR experiences. To produce beautiful composites, the camera needs to be geometrically calibrated, which can be tedious and requires a physical…

Augmented and Mixed Reality are emerging as likely successors to the mobile internet. However, many technical challenges remain. One of the key requirements of these systems is the ability to create a continuity between physical and virtual…

计算机视觉与模式识别 · 计算机科学 2022-02-18 Scott Y. L. Chin , Bradley R. Quinton

Current research on visual analytics systems largely follows the research paradigm of interactive system design in the field of Human-Computer Interaction (HCI), and includes key methodologies including design requirement development based…

人机交互 · 计算机科学 2026-03-02 Xiaolong Zhang

Problems at the intersection of vision and language are of significant importance both as challenging research questions and for the rich set of applications they enable. However, inherent structure in our world and bias in our language…

计算机视觉与模式识别 · 计算机科学 2017-05-16 Yash Goyal , Tejas Khot , Douglas Summers-Stay , Dhruv Batra , Devi Parikh

In question-answering scenarios, humans can assess whether the available information is sufficient and seek additional information if necessary, rather than providing a forced answer. In contrast, Vision Language Models (VLMs) typically…

计算机视觉与模式识别 · 计算机科学 2024-11-04 Li Liu , Diji Yang , Sijia Zhong , Kalyana Suma Sree Tholeti , Lei Ding , Yi Zhang , Leilani H. Gilpin
‹ 上一页 1 8 9 10 下一页 ›