中文
相关论文

相关论文: ARChef: An iOS-Based Augmented Reality Cooking Ass…

200 篇论文

Food image segmentation is a critical and indispensible task for developing health-related applications such as estimating food calories and nutrients. Existing food image segmentation models are underperforming due to two reasons: (1)…

计算机视觉与模式识别 · 计算机科学 2021-05-13 Xiongwei Wu , Xin Fu , Ying Liu , Ee-Peng Lim , Steven C. H. Hoi , Qianru Sun

Conventional approaches to dietary assessment are primarily grounded in self-reporting methods or structured interviews conducted under the supervision of dietitians. These methods, however, are often subjective, potentially inaccurate, and…

计算机视觉与模式识别 · 计算机科学 2023-12-15 Frank P. -W. Lo , Jianing Qiu , Zeyu Wang , Junhong Chen , Bo Xiao , Wu Yuan , Stamatia Giannarou , Gary Frost , Benny Lo

Understanding electronics is a critical area in the maker scene. Many of the makers' projects require electronics knowledge to connect microcontrollers with sensors and actuators. Yet, learning electronics is challenging, as internal…

人机交互 · 计算机科学 2022-10-26 Thomas Kosch , Julian Rasch , Albrecht Schmidt , Sebastian Feger

As augmented reality (AR) becomes increasingly integrated into everyday life, ensuring the safety and trustworthiness of its virtual content is critical. Our research addresses the risks of task-detrimental AR content, particularly that…

计算机视觉与模式识别 · 计算机科学 2025-08-01 Yanming Xiu

This paper introduces an augmented reality (AR) captioning framework designed to support Deaf and Hard of Hearing (DHH) learners in STEM classrooms by integrating non-verbal emotional cues into live transcriptions. Unlike conventional…

人机交互 · 计算机科学 2025-04-29 Sunday David Ubur

The increasing trust in large language models (LLMs), especially in the form of chatbots, is often undermined by the lack of their extrinsic evaluation. This holds particularly true in nutrition, where randomised controlled trials (RCTs)…

人机交互 · 计算机科学 2025-11-27 Karen Jia-Hui Li , Simone Balloccu , Ondrej Dusek , Ehud Reiter

Smartphones have significantly enhanced our daily learning, communication, and entertainment, becoming an essential component of modern life. However, certain populations, including the elderly and individuals with disabilities, encounter…

机器人学 · 计算机科学 2024-09-17 Kelin Fu , Yang Tian , Kaigui Bian

Augmented Reality (AR) systems are increasingly integrating foundation models, such as Multimodal Large Language Models (MLLMs), to provide more context-aware and adaptive user experiences. This integration has led to the development of AR…

人工智能 · 计算机科学 2025-08-13 Dongwook Choi , Taeyoon Kwon , Dongil Yang , Hyojun Kim , Jinyoung Yeo

This paper introduces Teachable Reality, an augmented reality (AR) prototyping tool for creating interactive tangible AR applications with arbitrary everyday objects. Teachable Reality leverages vision-based interactive machine teaching…

人机交互 · 计算机科学 2023-02-23 Kyzyl Monteiro , Ritik Vatsal , Neil Chulpongsatorn , Aman Parnami , Ryo Suzuki

Worldwide, in 2014, more than 1.9 billion adults, 18 years and older, were overweight. Of these, over 600 million were obese. Accurately documenting dietary caloric intake is crucial to manage weight loss, but also presents challenges…

计算机视觉与模式识别 · 计算机科学 2016-06-21 Chang Liu , Yu Cao , Yan Luo , Guanling Chen , Vinod Vokkarane , Yunsheng Ma

Robot-assisted feeding has the potential to improve the quality of life for individuals with mobility limitations who are unable to feed themselves independently. However, there exists a large gap between the homogeneous, curated plates…

机器人学 · 计算机科学 2024-07-11 Rajat Kumar Jenamani , Priya Sundaresan , Maram Sakr , Tapomayukh Bhattacharjee , Dorsa Sadigh

Image-to-recipe retrieval is a challenging vision-to-language task of significant practical value. The main challenge of the task lies in the ultra-high redundancy in the long recipe and the large variation reflected in both food item…

计算机视觉与模式识别 · 计算机科学 2023-05-22 Bhanu Prakash Voutharoja , Peng Wang , Lei Wang , Vivienne Guan

In this paper, we study the novel problem of not only predicting ingredients from a food image, but also predicting the relative amounts of the detected ingredients. We propose two prediction-based models using deep learning that output…

机器学习 · 计算机科学 2019-10-02 Jiatong Li , Ricardo Guerrero , Vladimir Pavlovic

Augmented Reality (AR) is a major immersive media technology that enriches our perception of reality by overlaying digital content (the foreground) onto physical environments (the background). It has far-reaching applications, from…

计算机视觉与模式识别 · 计算机科学 2024-12-10 Aymen Sekhri , Seyed Ali Amirshahi , Mohamed-Chaker Larabi

Mild Cognitive Impairment (MCI) affects 15-20% of adults aged 65 and older, often making kitchen navigation and independent living difficult, particularly in lower-income communities with limited access to professional design help. This…

人机交互 · 计算机科学 2026-04-16 Ibrahim Bilau , Nicole Li , Terrence Malayvong , Eunhwa Yang

This study explores the capabilities of multimodal large language models (LLMs) in handling challenging multistep tasks that integrate language and vision, focusing on model steerability, composability, and the application of long-term…

人工智能 · 计算机科学 2023-12-20 David Noever , Samantha Elizabeth Miller Noever

In augmented reality (AR), where digital content is overlaid onto the real world, realistic thermal feedback has been shown to enhance immersion. Yet current thermal feedback devices, heavily influenced by the needs of virtual reality,…

人机交互 · 计算机科学 2025-03-28 Alexandra Watkins , Ritam Ghosh , Evan Chow , Nilanjan Sarkar

Search engines enable the retrieval of unknown information with texts. However, traditional methods fall short when it comes to understanding unfamiliar visual content, such as identifying an object that the model has never seen before.…

计算机视觉与模式识别 · 计算机科学 2024-10-29 Zhixin Zhang , Yiyuan Zhang , Xiaohan Ding , Xiangyu Yue

Multimodal Large Language Models (MLLMs) are evolving from passive observers into active agents, solving problems through Visual Expansion (invoking visual tools) and Knowledge Expansion (open-web search). However, existing evaluations fall…

Argumentative LLMs (ArgLLMs) are an existing approach leveraging Large Language Models (LLMs) and computational argumentation for decision-making, with the aim of making the resulting decisions faithfully explainable to and contestable by…

计算与语言 · 计算机科学 2026-03-02 Adam Dejl , Deniz Gorur , Francesca Toni