English
Related papers

Related papers: TripTailor: A Real-World Benchmark for Personalize…

200 papers

Recent advancements in probing Large Language Models (LLMs) have explored their latent potential as personalized travel planning agents, yet existing benchmarks remain limited in real world applicability. Existing datasets, such as…

Computation and Language · Computer Science 2025-03-03 Soumyabrata Chaudhuri , Pranav Purkar , Ritwik Raghav , Shubhojit Mallick , Manish Gupta , Abhik Jana , Shreya Ghosh

Travel planning is a valuable yet complex task that poses significant challenges even for advanced large language models (LLMs). While recent benchmarks have advanced in evaluating LLMs' planning capabilities, they often fall short in…

Artificial Intelligence · Computer Science 2025-10-17 Yincen Qu , Huan Xiao , Feng Li , Gregory Li , Hui Zhou , Xiangying Dai , Xiaoru Dai

Recent efforts like TripCraft and TravelPlanner have advanced the use of Large Language Models ( LLMs) for personalized, constraint aware travel itinerary generation. Yet, real travel often faces disruptions. To address this, we present…

Computation and Language · Computer Science 2025-10-27 Priyanshu Karmakar , Soumyabrata Chaudhuri , Shubhojit Mallick , Manish Gupta , Abhik Jana , Shreya Ghosh

Planning has been part of the core pursuit for artificial intelligence since its conception, but earlier AI agents mostly focused on constrained settings because many of the cognitive substrates necessary for human-level planning have been…

Computation and Language · Computer Science 2024-10-24 Jian Xie , Kai Zhang , Jiangjie Chen , Tinghui Zhu , Renze Lou , Yuandong Tian , Yanghua Xiao , Yu Su

As global tourism expands and artificial intelligence technology advances, intelligent travel planning services have emerged as a significant research focus. Within dynamic real-world travel scenarios with multi-dimensional constraints,…

Artificial Intelligence · Computer Science 2024-09-13 Aili Chen , Xuyang Ge , Ziquan Fu , Yanghua Xiao , Jiangjie Chen

Travel planning is a realistic task for evaluating the planning and tool-use abilities of LLM agents. However, existing benchmarks typically assume only a single user, thereby avoiding one of the most challenging aspects of real-world…

Computation and Language · Computer Science 2026-05-26 Xiang Cheng , Yulan Hu , Lulu Zheng , Zheng Pan , Xin Li , Yong Liu

Travel planning is a natural real-world task to test large language models' (LLMs) planning and tool-use abilities. Although prior work has studied LLM performance on travel planning, existing settings still differ from real-world needs,…

Artificial Intelligence · Computer Science 2026-04-22 Xiang Cheng , Yulan Hu , Xiangwen Zhang , Lu Xu , Lide Tan , Zheng Pan , Xin Li , Yong Liu

Travel planning is a complex task that involves generating a sequence of actions related to visiting places subject to constraints and maximizing some user satisfaction criteria. Traditional approaches rely on problem formulation in a given…

Artificial Intelligence · Computer Science 2024-06-17 Tomas de la Rosa , Sriram Gopalakrishnan , Alberto Pozanco , Zhen Zeng , Daniel Borrajo

Large Language Models (LLMs) struggle to directly generate correct plans for complex multi-constraint planning problems, even with self-verification and self-critique. For example, a U.S. domestic travel planning benchmark TravelPlanner was…

Artificial Intelligence · Computer Science 2025-01-30 Yilun Hao , Yongchao Chen , Yang Zhang , Chuchu Fan

Real-world planning problems require constant adaptation to changing requirements and balancing of competing constraints. However, current benchmarks for evaluating LLMs' planning capabilities primarily focus on static, single-turn…

Computation and Language · Computer Science 2025-06-06 Juhyun Oh , Eunsu Kim , Alice Oh

Although large language models have enhanced automated travel planning abilities, current systems remain misaligned with real-world scenarios. First, they assume users provide explicit queries, while in reality requirements are often…

Artificial Intelligence · Computer Science 2025-08-22 Bin Deng , Yizhe Feng , Zeming Liu , Qing Wei , Xiangrong Zhu , Shuai Chen , Yuanfang Guo , Yunhong Wang

Real-world autonomous planning requires coordinating tightly coupled constraints where a single decision dictates the feasibility of all subsequent actions. However, existing benchmarks predominantly feature loosely coupled constraints…

Real-world trip planning requires transforming open-ended user requests into executable itineraries under strict spatial, temporal, and budgetary constraints while aligning with user preferences. Existing LLM-based agents struggle with…

Artificial Intelligence · Computer Science 2025-12-15 Yuxing Chen , Basem Suleiman , Qifan Chen

We introduce NATURAL PLAN, a realistic planning benchmark in natural language containing 3 key tasks: Trip Planning, Meeting Planning, and Calendar Scheduling. We focus our evaluation on the planning capabilities of LLMs with full…

Computation and Language · Computer Science 2024-06-10 Huaixiu Steven Zheng , Swaroop Mishra , Hugh Zhang , Xinyun Chen , Minmin Chen , Azade Nova , Le Hou , Heng-Tze Cheng , Quoc V. Le , Ed H. Chi , Denny Zhou

Route-planning agents powered by large language models (LLMs) have emerged as a promising paradigm for supporting everyday human mobility through natural language interaction and tool-mediated decision making. However, systematic evaluation…

Artificial Intelligence · Computer Science 2026-02-27 Zhiheng Song , Jingshuai Zhang , Chuan Qin , Chao Wang , Chao Chen , Longfei Xu , Kaikui Liu , Xiangxiang Chu , Hengshu Zhu

The recent trend of using Large Language Models (LLMs) as tool agents in real-world applications underscores the necessity for comprehensive evaluations of their capabilities, particularly in complex scenarios involving planning, creating,…

Computation and Language · Computer Science 2024-06-04 Shijue Huang , Wanjun Zhong , Jianqiao Lu , Qi Zhu , Jiahui Gao , Weiwen Liu , Yutai Hou , Xingshan Zeng , Yasheng Wang , Lifeng Shang , Xin Jiang , Ruifeng Xu , Qun Liu

We present PricingLogic, the first benchmark that probes whether Large Language Models(LLMs) can reliably automate tourism-related prices when multiple, overlapping fare rules apply. Travel agencies are eager to offload this error-prone…

Artificial Intelligence · Computer Science 2025-10-15 Yunuo Liu , Dawei Zhu , Zena Al-Khalili , Dai Cheng , Yanjun Chen , Dietrich Klakow , Wei Zhang , Xiaoyu Shen

Large language models (LLMs) have brought autonomous agents closer to artificial general intelligence (AGI) due to their promising generalization and emergent capabilities. There is, however, a lack of studies on how LLM-based agents…

Artificial Intelligence · Computer Science 2024-08-13 Yanan Chen , Ali Pesaranghader , Tanmana Sadhu , Dong Hoon Yi

Addressing itinerary modification is crucial for enhancing the travel experience as it is a frequent requirement during traveling. However, existing research mainly focuses on fixed itinerary planning, leaving modification underexplored due…

Information Retrieval · Computer Science 2026-02-23 Zhuoxuan Huang , Yunshan Ma , Hongyu Zhang , Hua Ma , Zhu Sun

Planning trips is a cognitively intensive task involving conflicting user preferences, dynamic external information, and multi-step temporal-spatial optimization. Traditional platforms often fall short - they provide static results, lack…

Multiagent Systems · Computer Science 2025-05-19 Binwen Liu , Jiexi Ge , Jiamin Wang
‹ Prev 1 2 3 10 Next ›