中文
相关论文

相关论文: Diverse Yet Consistent: Context-Guided Diffusion w…

200 篇论文

Interactive decision-making is essential in applications such as autonomous driving, where the agent must infer the behavior of nearby human drivers while planning in real-time. Traditional predict-then-act frameworks are often insufficient…

To accurately predict trajectories in multi-agent settings, e.g. team games, it is important to effectively model the interactions among agents. Whereas a number of methods have been developed for this purpose, existing methods implicitly…

计算机视觉与模式识别 · 计算机科学 2022-10-25 Zikai Wei , Xinge Zhu , Bo Dai , Dahua Lin

Diffusion and flow-based models have enabled significant progress in generation tasks across various modalities and have recently found applications in predictive learning. However, unlike typical generation tasks that encourage sample…

计算机视觉与模式识别 · 计算机科学 2026-03-24 Yu Zhang , Xingzhuo Guo , Haoran Xu , Jialong Wu , Mingsheng Long

Generative models such as Variational Autoencoders (VAEs) and Generative Adversarial Networks (GANs) have shown promise in sequential recommendation tasks. However, they face challenges, including posterior collapse and limited…

机器学习 · 计算机科学 2024-10-28 Sharare Zolghadr , Ole Winther , Paul Jeha

Due to the complex and changing interactions in dynamic scenarios, motion forecasting is a challenging problem in autonomous driving. Most existing works exploit static road graphs to characterize scenarios and are limited in modeling…

人工智能 · 计算机科学 2023-03-09 Xing Gao , Xiaogang Jia , Yikang Li , Hongkai Xiong

We present a new predictor combination algorithm that improves a given task predictor based on potentially relevant reference predictors. Existing approaches are limited in that, to discover the underlying task dependence, they either…

计算机视觉与模式识别 · 计算机科学 2019-04-11 Kwang In Kim , Hyung Jin Chang

Motion forecasting plays a significant role in various domains (e.g., autonomous driving, human-robot interaction), which aims to predict future motion sequences given a set of historical observations. However, the observed elements may be…

计算机视觉与模式识别 · 计算机科学 2021-08-04 Jiachen Li , Fan Yang , Hengbo Ma , Srikanth Malla , Masayoshi Tomizuka , Chiho Choi

Diffusion models have emerged as a widely utilized and successful methodology in human motion synthesis. Task-oriented diffusion models have significantly advanced action-to-motion, text-to-motion, and audio-to-motion applications. In this…

计算机视觉与模式识别 · 计算机科学 2025-12-05 Yuduo Jin , Brandon Haworth

Most recommender systems research focuses on binary historical user-item interaction encodings to predict future interactions. User features, item features, and interaction strengths remain largely under-utilized in this space or only…

信息检索 · 计算机科学 2024-09-24 Utkarsh Priyam , Hemit Shah , Edoardo Botta

Reinforcement Learning (RL)-based motion planning has recently shown the potential to outperform traditional approaches from autonomous navigation to robot manipulation. In this work, we focus on a motion planning task for an evasive target…

机器人学 · 计算机科学 2025-05-12 Zixuan Wu , Sean Ye , Manisha Natarajan , Matthew C. Gombolay

Automated parking is a challenging operational domain for advanced driver assistance systems, requiring robust scene understanding and interaction reasoning. The key challenge is twofold: (i) predict multiple plausible ego intentions…

机器人学 · 计算机科学 2026-02-25 Jiarong Wei , Anna Rehr , Christian Feist , Abhinav Valada

Recent text-driven motion generation methods span both discrete token-based approaches and continuous-latent formulations. MotionGPT3 exemplifies the latter paradigm, combining a learned continuous motion latent space with a diffusion-based…

计算机视觉与模式识别 · 计算机科学 2026-04-23 Jaymin Ban , JiHong Jeon , SangYeop Jeong

There is a gap in risk assessment of trajectories between the trajectory information coming from a traffic motion prediction module and what is actually needed. Closing this gap necessitates advancements in prediction beyond current…

机器学习 · 计算机科学 2025-12-04 Marlon Steiner , Marvin Klemp , Christoph Stiller

Motivated by the problem of pursuit-evasion, we present a motion planning framework that combines energy-based diffusion models with artificial potential fields for robust real time trajectory generation in complex environments. Our…

机器人学 · 计算机科学 2025-10-17 Wondmgezahu Teshome , Kian Behzad , Octavia Camps , Michael Everett , Milad Siami , Mario Sznaier

Numerous models have been developed for scanpath and saliency prediction, which are typically trained on scanpaths, which model eye movement as a sequence of discrete fixation points connected by saccades, while the rich information…

计算机视觉与模式识别 · 计算机科学 2025-10-10 Ozgur Kara , Harris Nisar , James M. Rehg

While modern diffusion models excel at generating diverse single images, extending this to sequential generation reveals a fundamental challenge: balancing narrative dynamism with multi-character coherence. Existing methods often falter at…

计算机视觉与模式识别 · 计算机科学 2026-05-13 Qi Zhao , Jun Chen , Ivor Tsang , Guang Dai

Natural and expressive human motion generation is the holy grail of computer animation. It is a challenging task, due to the diversity of possible motion, human perceptual sensitivity to it, and the difficulty of accurately describing it.…

计算机视觉与模式识别 · 计算机科学 2022-10-04 Guy Tevet , Sigal Raab , Brian Gordon , Yonatan Shafir , Daniel Cohen-Or , Amit H. Bermano

Dexterous manipulation with contact-rich interactions is crucial for advanced robotics. While recent diffusion-based planning approaches show promise for simple manipulation tasks, they often produce unrealistic ghost states (e.g., the…

机器人学 · 计算机科学 2025-06-18 Zhixuan Liang , Yao Mu , Yixiao Wang , Tianxing Chen , Wenqi Shao , Wei Zhan , Masayoshi Tomizuka , Ping Luo , Mingyu Ding

Text-driven human motion generation is a multimodal task that synthesizes human motion sequences conditioned on natural language. It requires the model to satisfy textual descriptions under varying conditional inputs, while generating…

计算机视觉与模式识别 · 计算机科学 2024-10-01 Xingyu Chen

Estimating the joint distribution of on-road agents' future trajectories is essential for autonomous driving. In this technical report, we propose a next-generation framework for joint multi-agent trajectory prediction called QCNeXt. First,…

计算机视觉与模式识别 · 计算机科学 2023-06-21 Zikang Zhou , Zihao Wen , Jianping Wang , Yung-Hui Li , Yu-Kai Huang