中文
相关论文

相关论文: ChatGPT and general-purpose AI count fruits in pic…

200 篇论文

Tree fruit breeding is a long-term activity involving repeated measurements of various fruit quality traits on a large number of samples. These traits are traditionally measured by manually counting the fruits, weighing to indirectly…

计算机视觉与模式识别 · 计算机科学 2023-02-15 Ritayu Nagpal , Sam Long , Shahid Jahagirdar , Weiwei Liu , Scott Fazackerley , Ramon Lawrence , Amritpal Singh

The challenge of formal proof generation has a rich history, but with modern techniques, we may finally be at the stage of making actual progress in real-life mathematical problems. This paper explores the integration of ChatGPT and basic…

计算机科学中的逻辑 · 计算机科学 2025-02-20 Sangjun Han , Taeil Hur , Youngmi Hur , Kathy Sangkyung Lee , Myungyoon Lee , Hyojae Lim

We propose the use of conversational GPT models for easy and quick few-shot text classification in the financial domain using the Banking77 dataset. Our approach involves in-context learning with GPT-3.5 and GPT-4, which minimizes the…

计算与语言 · 计算机科学 2023-08-29 Lefteris Loukas , Ilias Stogiannidis , Prodromos Malakasiotis , Stavros Vassos

Existing works on visual counting primarily focus on one specific category at a time, such as people, animals, and cells. In this paper, we are interested in counting everything, that is to count objects from any category given only a few…

计算机视觉与模式识别 · 计算机科学 2021-04-20 Viresh Ranjan , Udbhav Sharma , Thu Nguyen , Minh Hoai

With the rise of foundation models, a new artificial intelligence paradigm has emerged, by simply using general purpose foundation models with prompting to solve problems instead of training a separate machine learning model for each…

人工智能 · 计算机科学 2023-08-29 Mostafa M. Amin , Rui Mao , Erik Cambria , Björn W. Schuller

This study highlights the potential of ChatGPT (specifically GPT-4o) as a competitive alternative for Face Presentation Attack Detection (PAD), outperforming several PAD models, including commercial solutions, in specific scenarios. Our…

计算机视觉与模式识别 · 计算机科学 2025-01-16 Alain Komaty , Hatef Otroshi Shahreza , Anjith George , Sebastien Marcel

The proliferation of AI models in everyday devices has highlighted a critical challenge: prediction errors that degrade user experience. While existing solutions focus on error detection, they rarely provide efficient correction mechanisms,…

In this paper, different techniques of few-shot, zero-shot, and regular object detection have been investigated. The need for few-shot learning and zero-shot learning techniques is crucial and arises from the limitations and challenges in…

计算机视觉与模式识别 · 计算机科学 2024-06-25 Maged Badawi , Mohammedyahia Abushanab , Sheethal Bhat , Andreas Maier

Deep learning models are transforming agricultural applications by enabling automated phenotyping, monitoring, and yield estimation. However, their effectiveness heavily depends on large amounts of annotated training data, which can be…

计算机视觉与模式识别 · 计算机科学 2025-04-11 Rajhans Singh , Rafael Bidese Puhl , Kshitiz Dhakal , Sudhir Sornapudi

ChatGPT has shown the potential of emerging general artificial intelligence capabilities, as it has demonstrated competent performance across many natural language processing tasks. In this work, we evaluate the capabilities of ChatGPT to…

计算与语言 · 计算机科学 2023-03-07 Mostafa M. Amin , Erik Cambria , Björn W. Schuller

Large-scale generative language models such as GPT-3 are competitive few-shot learners. While these models are known to be able to jointly represent many different languages, their training data is dominated by English, potentially limiting…

Conventional approaches to dietary assessment are primarily grounded in self-reporting methods or structured interviews conducted under the supervision of dietitians. These methods, however, are often subjective, potentially inaccurate, and…

计算机视觉与模式识别 · 计算机科学 2023-12-15 Frank P. -W. Lo , Jianing Qiu , Zeyu Wang , Junhong Chen , Bo Xiao , Wu Yuan , Stamatia Giannarou , Gary Frost , Benny Lo

Foundation models have had a big impact in recent years and billions of dollars are being invested in them in the current AI boom. The more popular ones, such as Chat-GPT, are trained on large amounts of Internet data. However, it is…

人工智能 · 计算机科学 2024-08-07 Dionis Barcari , David Gamez , Aliya Grig

ChatGPT is attracting a cross-field interest as it provides a language interface with remarkable conversational competency and reasoning capabilities across many domains. However, since ChatGPT is trained with languages, it is currently not…

计算机视觉与模式识别 · 计算机科学 2023-03-09 Chenfei Wu , Shengming Yin , Weizhen Qi , Xiaodong Wang , Zecheng Tang , Nan Duan

In this paper, we consider the problem of generalised visual object counting, with the goal of developing a computational model for counting the number of objects from arbitrary semantic categories, using arbitrary number of "exemplars",…

计算机视觉与模式识别 · 计算机科学 2023-06-05 Chang Liu , Yujie Zhong , Andrew Zisserman , Weidi Xie

Generalizable object fetching in cluttered scenes remains a fundamental and application-critical challenge in embodied AI. Closely packed objects cause inevitable occlusions, making safe action generation particularly difficult. Under such…

机器人学 · 计算机科学 2025-08-26 Weiheng Liu , Yuxuan Wan , Jilong Wang , Yuxuan Kuang , Wenbo Cui , Xuesong Shi , Haoran Li , Dongbin Zhao , Zhizheng Zhang , He Wang

Recent research on dialogue state tracking (DST) focuses on methods that allow few- and zero-shot transfer to new domains or schemas. However, performance gains heavily depend on aggressive data augmentation and fine-tuning of ever larger…

Few-shot learning-the ability to train models with access to limited data-has become increasingly popular in the natural language processing (NLP) domain, as large language models such as GPT and T0 have been empirically shown to achieve…

软件工程 · 计算机科学 2023-06-16 Robert Kraig Helmeczi , Mucahit Cevik , Savas Yıldırım

Zero-shot object counting attempts to estimate the number of object instances belonging to novel categories that the vision model performing the counting has never encountered during training. Existing methods typically require large amount…

计算机视觉与模式识别 · 计算机科学 2025-12-04 Richard Füzesséry , Kaziwa Saleh , Sándor Szénási , Zoltán Vámossy

The emergence of Large Language Models (LLMs) and multimodal foundation models (FMs) has generated heightened interest in their applications that integrate vision and language. This paper investigates the capabilities of ChatGPT-4V and…

计算机视觉与模式识别 · 计算机科学 2024-08-26 Zhenyuan Yang , Xuhui Lin , Qinyi He , Ziye Huang , Zhengliang Liu , Hanqi Jiang , Peng Shu , Zihao Wu , Yiwei Li , Stephen Law , Gengchen Mai , Tianming Liu , Tao Yang
‹ 上一页 1 2 3 10 下一页 ›