中文
相关论文

相关论文: Bongard-LOGO: A New Benchmark for Human-Level Conc…

200 篇论文

A significant gap remains between today's visual pattern recognition models and human-level visual cognition especially when it comes to few-shot learning and compositional reasoning of novel concepts. We introduce Bongard-HOI, a new visual…

计算机视觉与模式识别 · 计算机科学 2023-04-14 Huaizu Jiang , Xiaojian Ma , Weili Nie , Zhiding Yu , Yuke Zhu , Song-Chun Zhu , Anima Anandkumar

We introduce Bongard-OpenWorld, a new benchmark for evaluating real-world few-shot reasoning for machine vision. It originates from the classical Bongard Problems (BPs): Given two sets of images (positive and negative), the model needs to…

机器学习 · 计算机科学 2025-01-08 Rujie Wu , Xiaojian Ma , Zhenliang Zhang , Wei Wang , Qing Li , Song-Chun Zhu , Yizhou Wang

Vision--language models (VLMs) often fail on abstract visual reasoning benchmarks such as Bongard problems, raising the question of whether the main bottleneck lies in reasoning or representation. We study this on Bongard-LOGO, a synthetic…

人工智能 · 计算机科学 2026-04-24 Mohit Vaishnav , Tanel Tammet

More than 50 years ago Bongard introduced 100 visual concept learning problems as a testbed for intelligent vision systems. These problems are now known as Bongard problems. Although they are well known in the cognitive science and AI…

机器学习 · 统计学 2018-04-13 Stefan Depeweg , Constantin A. Rothkopf , Frank Jäkel

Abstract visual reasoning (AVR) involves discovering shared concepts across images through analogy, akin to solving IQ test problems. Bongard Problems (BPs) remain a key challenge in AVR, requiring both visual reasoning and verbal…

人工智能 · 计算机科学 2025-06-24 Mikołaj Małkiński , Szymon Pawlonka , Jacek Mańdziuk

Current machine learning methods struggle to solve Bongard problems, which are a type of IQ test that requires deriving an abstract "concept" from a set of positive and negative "support" images, and then classifying whether or not a new…

计算机视觉与模式识别 · 计算机科学 2024-12-03 Nikhil Raghuraman , Adam W. Harley , Leonidas Guibas

Even though AI has advanced rapidly in recent years displaying success in solving highly complex problems, the class of Bongard Problems (BPs) yet remain largely unsolved by modern ML techniques. In this paper, we propose a new approach in…

机器学习 · 计算机科学 2022-12-26 Salahedine Youssef , Matej Zečević , Devendra Singh Dhami , Kristian Kersting

Recently, newly developed Vision-Language Models (VLMs), such as OpenAI's o1, have emerged, seemingly demonstrating advanced reasoning capabilities across text and image modalities. However, the depth of these advances in language-guided…

Bongard Problems (BPs) provide a challenging testbed for abstract visual reasoning (AVR), requiring models to identify visual concepts fromjust a few examples and describe them in natural language. Early BP benchmarks featured synthetic…

人工智能 · 计算机科学 2026-02-20 Szymon Pawlonka , Mikołaj Małkiński , Jacek Mańdziuk

Vision-Language Models (VLMs) have made great strides in everyday visual tasks, such as captioning a natural image, or answering commonsense questions about such images. But humans possess the puzzling ability to deploy their visual…

计算机视觉与模式识别 · 计算机科学 2026-02-04 Cassidy Langenfeld , Claas Beger , Gloria Geng , Wasu Top Piriyakulkij , Keya Hu , Yewen Pu , Kevin Ellis

A fundamental challenge in artificial intelligence involves understanding the cognitive mechanisms underlying visual reasoning in sophisticated models like Vision-Language Models (VLMs). How do these models integrate visual perception with…

计算机视觉与模式识别 · 计算机科学 2025-05-07 Mohit Vaishnav , Tanel Tammet

Visual abstract reasoning problems pose significant challenges to the perception and cognition abilities of artificial intelligence algorithms, demanding deeper pattern recognition and inductive reasoning beyond mere identification of…

计算机视觉与模式识别 · 计算机科学 2025-05-27 Ruizhuo Song , Beiming Yuan

One of the primary challenges faced by deep learning is the degree to which current methods exploit superficial statistics and dataset bias, rather than learning to generalise over the specific representations they have experienced. This is…

计算机视觉与模式识别 · 计算机科学 2019-07-30 Damien Teney , Peng Wang , Jiewei Cao , Lingqiao Liu , Chunhua Shen , Anton van den Hengel

While Multimodal Large Language Models (MLLMs) are adept at answering what is in an image-identifying objects and describing scenes-they often lack the ability to understand how an image feels to a human observer. This gap is most evident…

计算机视觉与模式识别 · 计算机科学 2025-12-01 Yiming Chen , Junlin Han , Tianyi Bai , Shengbang Tong , Filippos Kokkinos , Philip Torr

Understanding how humans and AI systems interpret ambiguous visual stimuli offers critical insight into the nature of perception, reasoning, and decision-making. This paper examines image labeling performance across human participants and…

人工智能 · 计算机科学 2025-12-11 Chethana Prasad Kabgere

Reasoning benchmarks such as the Abstraction and Reasoning Corpus (ARC) and ARC-AGI are widely used to assess progress in artificial intelligence and are often interpreted as probes of core, so-called ``fluid'' reasoning abilities. Despite…

计算与语言 · 计算机科学 2026-01-12 Xinhe Wang , Jin Huang , Xingjian Zhang , Tianhao Wang , Jiaqi W. Ma

A fundamental component of human vision is our ability to parse complex visual scenes and judge the relations between their constituent objects. AI benchmarks for visual reasoning have driven rapid progress in recent years with…

计算机视觉与模式识别 · 计算机科学 2022-06-14 Aimen Zerroug , Mohit Vaishnav , Julien Colin , Sebastian Musslick , Thomas Serre

The abilities to form and abstract concepts is key to human intelligence, but such abilities remain lacking in state-of-the-art AI systems. There has been substantial research on conceptual abstraction in AI, particularly using idealized…

机器学习 · 计算机科学 2023-08-09 Arseny Moskvichev , Victor Vikram Odouard , Melanie Mitchell

Recent advances in multimodal large language models (MLLMs) have been primarily evaluated on general-purpose benchmarks, while their applications in domain-specific scenarios, such as intelligent product moderation, remain underexplored. To…

计算机视觉与模式识别 · 计算机科学 2025-10-01 Zichen Liang , Jingjing Fei , Jie Wang , Zheming Yang , Changqing Li , Pei Wu , Minghui Qiu , Fei Yang , Xialei Liu

The ability to recognise and make analogies is often used as a measure or test of human intelligence. The ability to solve Bongard problems is an example of such a test. It has also been postulated that the ability to rapidly construct…

机器学习 · 计算机科学 2021-10-20 Atharv Sonwane , Sharad Chitlangia , Tirtharaj Dash , Lovekesh Vig , Gautam Shroff , Ashwin Srinivasan
‹ 上一页 1 2 3 10 下一页 ›