中文
相关论文

相关论文: FantasyHSI: Video-Generation-Centric 4D Human Synt…

200 篇论文

Reconstructing metrically accurate humans and their surrounding scenes from a single image is crucial for virtual reality, robotics, and comprehensive 3D scene understanding. However, existing methods struggle with depth ambiguity,…

计算机视觉与模式识别 · 计算机科学 2025-10-14 Pradyumna Yalandur Muralidhar , Yuxuan Xue , Xianghui Xie , Margaret Kostyrko , Gerard Pons-Moll

Current 3D scene graph generation (3DSGG) approaches heavily rely on a single-agent assumption and small-scale environments, exhibiting limited scalability to real-world scenarios. In this work, we introduce Multi-Agent 3D Scene Graph…

机器人学 · 计算机科学 2026-02-05 Yirum Kim , Jaewoo Kim , Ue-Hwan Kim

This paper introduces the first text-guided work for generating the sequence of hand-object interaction in 3D. The main challenge arises from the lack of labeled data where existing ground-truth datasets are nowhere near generalizable in…

计算机视觉与模式识别 · 计算机科学 2024-04-03 Junuk Cha , Jihyeon Kim , Jae Shin Yoon , Seungryul Baek

At-home physiotherapy compliance remains critically low due to a lack of personalized supervision and dynamic feedback. Existing digital health solutions rely on static, pre-recorded video libraries or generic 3D avatars that fail to…

人工智能 · 计算机科学 2026-04-24 Abhishek Dharmaratnakar , Srivaths Ranganathan , Anushree Sinha , Debanshu Das

We propose a new task to benchmark human-in-scene understanding for embodied agents: Human-In-Scene Question Answering (HIS-QA). Given a human motion within a 3D scene, HIS-QA requires the agent to comprehend human states and behaviors,…

计算机视觉与模式识别 · 计算机科学 2025-07-17 Jiahe Zhao , Ruibing Hou , Zejie Tian , Hong Chang , Shiguang Shan

Intelligent agents must autonomously interact with the environments to perform daily tasks based on human-level instructions. They need a foundational understanding of the world to accurately interpret these instructions, along with precise…

人工智能 · 计算机科学 2025-08-22 Zhen Wu , Jiaman Li , Pei Xu , C. Karen Liu

ComfyUI is a popular workflow-based interface that allows users to customize image generation tasks through an intuitive node-based system. However, the complexity of managing node connections and diverse modules can be challenging for…

多智能体系统 · 计算机科学 2025-09-18 Oucheng Huang , Yuhang Ma , Zeng Zhao , Mingrui Wu , Jiayi Ji , Rongsheng Zhang , Zhipeng Hu , Xiaoshuai Sun , Rongrong Ji

Recent advancements in large language models (LLMs) have significantly improved the capabilities of web agents. However, effectively navigating complex and dynamic web environments still requires more advanced trajectory-level planning and…

人工智能 · 计算机科学 2025-07-08 Yifei Gao , Junhong Ye , Jiaqi Wang , Jitao Sang

Recent advances in 3D human-aware generation have made significant progress. However, existing methods still struggle with generating novel Human Object Interaction (HOI) from text, particularly for open-set objects. We identify three main…

计算机视觉与模式识别 · 计算机科学 2025-06-02 Jinlu Zhang , Yixin Chen , Zan Wang , Jie Yang , Yizhou Wang , Siyuan Huang

We present a novel framework for animating humans in 3D scenes using 3D Gaussian Splatting (3DGS), a neural scene representation that has recently achieved state-of-the-art photorealistic results for novel-view synthesis but remains…

计算机视觉与模式识别 · 计算机科学 2026-01-06 Aymen Mir , Jian Wang , Riza Alp Guler , Chuan Guo , Gerard Pons-Moll , Bing Zhou

We address hyperspectral image (HSI) synthesis, a problem that has garnered growing interest yet remains constrained by the conditional generative paradigms that limit sample diversity. While diffusion models have emerged as a…

计算机视觉与模式识别 · 计算机科学 2025-08-20 Shiyu Shen , Bin Pan , Ziye Zhang , Zhenwei Shi

Recent advancements in human motion synthesis have focused on specific types of motions, such as human-scene interaction, locomotion or human-human interaction, however, there is a lack of a unified system capable of generating a diverse…

计算机视觉与模式识别 · 计算机科学 2025-02-14 Jianqi Chen , Panwen Hu , Xiaojun Chang , Zhenwei Shi , Michael Kampffmeyer , Xiaodan Liang

Recent advancements in diffusion models have led to significant improvements in the generation and animation of 4D full-body human-object interactions (HOI). Nevertheless, existing methods primarily focus on SMPL-based motion generation,…

计算机视觉与模式识别 · 计算机科学 2024-10-10 Yukang Cao , Liang Pan , Kai Han , Kwan-Yee K. Wong , Ziwei Liu

Virtual film production requires intricate decision-making processes, including scriptwriting, virtual cinematography, and precise actor positioning and actions. Motivated by recent advances in automated decision-making with language…

计算与语言 · 计算机科学 2025-01-23 Zhenran Xu , Longyue Wang , Jifang Wang , Zhouyi Li , Senbao Shi , Xue Yang , Yiyu Wang , Baotian Hu , Jun Yu , Min Zhang

Hyperspectral image (HSI) plays a vital role in various fields such as agriculture and environmental monitoring. However, due to the expensive acquisition cost, the number of hyperspectral images is limited, degenerating the performance of…

计算机视觉与模式识别 · 计算机科学 2024-11-04 Li Pang , Xiangyong Cao , Datao Tang , Shuang Xu , Xueru Bai , Feng Zhou , Deyu Meng

We present HOI-PAGE, a new approach that prioritizes part-level affordance reasoning to generate high-fidelity 4D human-object interactions (HOIs) from text prompts in a zero-shot fashion. In contrast to prior works that focus on global,…

图形学 · 计算机科学 2026-05-20 Lei Li , Angela Dai

Humans interact with an object in many different ways by making contact at different locations, creating a highly complex motion space that can be difficult to learn, particularly when synthesizing such human interactions in a controllable…

计算机视觉与模式识别 · 计算机科学 2022-05-03 Xiaohan Zhang , Bharat Lal Bhatnagar , Vladimir Guzov , Sebastian Starke , Gerard Pons-Moll

Digital human motion synthesis is a vibrant research field with applications in movies, AR/VR, and video games. Whereas methods were proposed to generate natural and realistic human motions, most only focus on modeling humans and largely…

计算机视觉与模式识别 · 计算机科学 2023-11-07 Quanzhou Li , Jingbo Wang , Chen Change Loy , Bo Dai

Humans interact with objects all the time. Enabling a humanoid to learn human-object interaction (HOI) is a key step for future smart animation and intelligent robotics systems. However, recent progress in physics-based HOI requires…

计算机视觉与模式识别 · 计算机科学 2023-12-08 Yinhuai Wang , Jing Lin , Ailing Zeng , Zhengyi Luo , Jian Zhang , Lei Zhang

In this study, we tackle the complex task of generating 3D human-object interactions (HOI) from textual descriptions in a zero-shot text-to-3D manner. We identify and address two key challenges: the unsatisfactory outcomes of direct…

计算机视觉与模式识别 · 计算机科学 2024-07-17 Sisi Dai , Wenhao Li , Haowen Sun , Haibin Huang , Chongyang Ma , Hui Huang , Kai Xu , Ruizhen Hu