中文
相关论文

相关论文: ARChef: An iOS-Based Augmented Reality Cooking Ass…

200 篇论文

Accurate and reliable search on online healthcare platforms is critical for user safety and service efficacy. Traditional methods, however, often fail to comprehend complex and nuanced user queries, limiting their effectiveness. Large…

计算与语言 · 计算机科学 2025-12-04 Chuyue Wang , Jie Feng , Yuxi Wu , Hang Zhang , Zhiguo Fan , Bing Cheng , Wei Lin

Large Multimodal Models (LMMs) have shown strong potential for assisting users in tasks, such as programming, content creation, and information access, yet their interaction remains largely limited to traditional interfaces such as desktops…

人机交互 · 计算机科学 2026-02-12 Liuchuan Yu , Yongqi Zhang , Lap-Fai Yu

One of the main challenges of advancing task-oriented learning such as visual task planning and reinforcement learning is the lack of realistic and standardized environments for training and testing AI agents. Previously, researchers often…

人机交互 · 计算机科学 2019-03-15 Xiaofeng Gao , Ran Gong , Tianmin Shu , Xu Xie , Shu Wang , Song-Chun Zhu

Cooking requires not only following instructions but also understanding, executing, and monitoring each step - a process that can be challenging without visual guidance. Although recipe images and videos offer helpful cues, they often lack…

计算机视觉与模式识别 · 计算机科学 2025-10-07 Oleh Kuzyk , Zuoyue Li , Marc Pollefeys , Xi Wang

Implementing visual and audio notifications on augmented reality devices is a crucial element of intuitive and easy-to-use interfaces. In this paper, we explored creating intuitive interfaces through visual and audio notifications. The…

人机交互 · 计算机科学 2024-09-18 Aditya Raikwar , Lucas Plabst , Anil Ufuk Batmaz , Florian Niebling , Francisco R. Ortega

In this work we propose a new computational framework, based on generative deep models, for synthesis of photo-realistic food meal images from textual descriptions of its ingredients. Previous works on synthesis of images from text…

计算机视觉与模式识别 · 计算机科学 2019-05-31 Fangda Han , Ricardo Guerrero , Vladimir Pavlovic

Food is not only a basic human necessity but also a key factor driving a society's health and economic well-being. As a result, the cooking domain is a popular use-case to demonstrate decision-support (AI) capabilities in service of…

Augmented Reality (AR) solutions are providing tools that could improve applications in the medical and industrial fields. Augmentation can provide additional information in training, visualization, and work scenarios, to increase…

人机交互 · 计算机科学 2023-06-30 Akos Nagy , Thomas Lagkas , Panagiotis Sarigiannidis , Vasileios Argyriou

Accurate nutrient estimation from unstructured recipe text is an important yet challenging problem in dietary monitoring, due to ambiguous ingredient terminology and highly variable quantity expressions. We systematically evaluate models…

计算与语言 · 计算机科学 2026-05-14 Wei-Chun Chen , Yu-Xuan Chen , I-Fang Chung , Ying-Jia Lin

Food computing has emerged as a prominent multidisciplinary field of research in recent years. An ambitious goal of food computing is to develop end-to-end intelligent systems capable of autonomously producing recipe information for a food…

计算机视觉与模式识别 · 计算机科学 2024-05-14 Prateek Chhikara , Dhiraj Chaurasia , Yifan Jiang , Omkar Masur , Filip Ilievski

Recent advances in large language models (LLMs) and the abundance of food data have resulted in studies to improve food understanding using LLMs. Despite several recommendation systems utilizing LLMs and Knowledge Graphs (KGs), there has…

机器学习 · 计算机科学 2025-05-21 Fnu Mohbat , Mohammed J Zaki

Large language models (LLMs) and large multimodal models (LMMs) promise to accelerate biomedical discovery, yet their reliability remains unclear. We introduce ARIEL (AI Research Assistant for Expert-in-the-Loop Learning), an open-source…

Large Language Models (LLMs) excel at many tasks, yet they struggle to produce truly creative, diverse ideas. In this paper, we introduce a novel approach that enhances LLM creativity. We apply LLMs for translating between natural language…

计算与语言 · 计算机科学 2025-09-30 Moran Mizrahi , Chen Shani , Gabriel Stanovsky , Dan Jurafsky , Dafna Shahaf

The integration of augmented reality (AR), extended reality (XR), and virtual reality (VR) technologies in agriculture has shown significant promise in enhancing various agricultural practices. Mobile robots have also been adopted as…

机器人学 · 计算机科学 2024-11-07 Caio Mucchiani , Dimitrios Chatziparaschis , Konstantinos Karydis

Cooking recipes are challenging to translate to robot plans as they feature rich linguistic complexity, temporally-extended interconnected tasks, and an almost infinite space of possible actions. Our key insight is that combining a source…

机器人学 · 计算机科学 2024-03-08 Angelos Mavrogiannis , Christoforos Mavrogiannis , Yiannis Aloimonos

Mixed Reality (MR) and Artificial Intelligence (AI) are increasingly becoming integral parts of our daily lives. Their applications range in fields from healthcare to education to entertainment. MR has opened a new frontier for such fields…

人机交互 · 计算机科学 2024-02-14 Ramez Yousri , Zeyad Essam , Yehia Kareem , Youstina Sherief , Sherry Gamil , Soha Safwat

Automatic dietary assessment based on food images remains a challenge, requiring precise food detection, segmentation, and classification. Vision-Language Models (VLMs) offer new possibilities by integrating visual and textual reasoning. In…

Individuals with vision impairments employ a variety of strategies for object identification, such as pans or soy sauce, in the culinary process. In addition, they often rely on contextual details about objects, such as location,…

人机交互 · 计算机科学 2024-02-26 Franklin Mingzhe Li , Michael Xieyang Liu , Shaun K. Kane , Patrick Carrington

AI has driven significant progress in the nutrition field, especially through multimedia-based automatic dietary assessment. However, existing automatic dietary assessment systems often overlook critical non-visual factors, such as…

人工智能 · 计算机科学 2025-07-15 Lubnaa Abdur Rahman , Ioannis Papathanail , Stavroula Mougiakakou

Synthesizing realistic cooked food images from raw inputs on edge devices is a challenging generative task, requiring models to capture complex changes in texture, color and structure during cooking. Existing image-to-image generation…

计算机视觉与模式识别 · 计算机科学 2025-11-24 Jigyasa Gupta , Soumya Goyal , Anil Kumar , Ishan Jindal