English
Related papers

Related papers: MemoryDiorama: Generating Dynamic 3D Diorama from …

200 papers

AI has been increasingly integrated into screenwriting practice. In refinement, screenwriters expect AI to provide feedback that supports reflection across the internal perspective of characters and the external perspective of the overall…

Human-Computer Interaction · Computer Science 2026-02-06 Yuying Tang , Xinyi Chen , Haotian Li , Xing Xie , Xiaojuan Ma , Huamin Qu

We propose a dual-domain generative model to estimate a texture map from a single image for colorizing a 3D human model. When estimating a texture map, a single image is insufficient as it reveals only one facet of a 3D object. To provide…

Computer Vision and Pattern Recognition · Computer Science 2022-03-15 Seunggyu Chang , Jungchan Cho , Songhwai Oh

Dynamical systems are ubiquitous within science and engineering, from turbulent flow across aircraft wings to structural variability of proteins. Although some systems are well understood and simulated, scientific imaging often confronts…

Computer Vision and Pattern Recognition · Computer Science 2025-09-03 Ali SaraerToosi , Renbo Tu , Kamyar Azizzadenesheli , Aviad Levis

Older adults have increasing difficulty with retrospective memory, hindering their abilities to perform daily activities and posing stress on caregivers to ensure their wellbeing. Recent developments in Artificial Intelligence (AI) and…

Human-Computer Interaction · Computer Science 2025-02-05 Natasha Maniar , Samantha W. T. Chan , Wazeer Zulfikar , Scott Ren , Christine Xu , Pattie Maes

Humans learn to recognize and manipulate new objects in lifelong settings without forgetting the previously gained knowledge under non-stationary and sequential conditions. In autonomous systems, the agents also need to mitigate similar…

Robotics · Computer Science 2022-01-25 Krishnakumar Santhakumar , Hamidreza Kasaei

Using image as prompts for 3D generation demonstrate particularly strong performances compared to using text prompts alone, for images provide a more intuitive guidance for the 3D generation process. In this work, we delve into the…

Computer Vision and Pattern Recognition · Computer Science 2024-04-29 Seungwook Kim , Yichun Shi , Kejie Li , Minsu Cho , Peng Wang

3D human reconstruction and animation are long-standing topics in computer graphics and vision. However, existing methods typically rely on sophisticated dense-view capture and/or time-consuming per-subject optimization procedures. To…

Graphics · Computer Science 2025-06-04 Zhiyuan Yu , Zhe Li , Hujun Bao , Can Yang , Xiaowei Zhou

Simultaneous localization and mapping (SLAM) with implicit neural representations has received extensive attention due to the expressive representation power and the innovative paradigm of continual learning. However, deploying such a…

Computer Vision and Pattern Recognition · Computer Science 2024-07-19 Baicheng Li , Zike Yan , Dong Wu , Hanqing Jiang , Hongbin Zha

Human memory has notable limitations (e.g., forgetting) which have necessitated a variety of memory aids (e.g., calendars). As we grow closer to mass adoption of everyday Extended Reality (XR), which is frequently leveraging perceptual…

Human-Computer Interaction · Computer Science 2023-04-06 Elise Bonnail , Eric Lecolinet , Wen-Jie Tseng , Samuel Huron , Mark Mcgill , Jan Gugenheimer

RAM incorporates a motion-aware semantic tracker with adaptive Kalman filtering to achieve robust identity association under severe occlusions and dynamic interactions. A memory-augmented Temporal HMR module further enhances human motion…

Computer Vision and Pattern Recognition · Computer Science 2026-04-13 Sen Jia , Ning Zhu , Jinqin Zhong , Jiale Zhou , Huaping Zhang , Jenq-Neng Hwang , Lei Li

Human memory exhibits significant vulnerability in cognitive tasks and daily life. Comparisons between visual working memory and new perceptual input (e.g., during cognitive tasks) can lead to unintended memory distortions. Previous studies…

Neurons and Cognition · Quantitative Biology 2025-07-31 Yuang Cao , Jiachen Zou , Chen Wei , Quanying Liu

This research study proposes using Generative Adversarial Networks (GAN) that incorporate a two-dimensional measure of human memorability to generate memorable or non-memorable images of scenes. The memorability of the generated images is…

Computer Vision and Pattern Recognition · Computer Science 2020-05-07 Cameron Kyle-Davidson , Adrian G. Bors , Karla K. Evans

Dynamic imaging is essential for analyzing various biological systems and behaviors but faces two main challenges: data incompleteness and computational burden. For many imaging systems, high frame rates and short acquisition times require…

Image and Video Processing · Electrical Eng. & Systems 2024-06-12 Luke Lozenski , Mark A. Anastasio , Umberto Villa

Existing image-to-video generation methods often produce physically implausible motions and lack precise control over object dynamics. While prior approaches have incorporated physics simulators, they remain confined to 2D planar motions…

Computer Vision and Pattern Recognition · Computer Science 2026-04-28 Tianyidan Xie , Zhentao Huang , Mingjie Wang , Xin Huang , Jun Zhou , Minglun Gong , Zili Yi

To enable machines to understand the way humans interact with the physical world in daily life, 3D interaction signals should be captured in natural settings, allowing people to engage with multiple objects in a range of sequential and…

Computer Vision and Pattern Recognition · Computer Science 2025-01-23 Jeonghwan Kim , Jisoo Kim , Jeonghyeon Na , Hanbyul Joo

Constructing memory from users' long-term conversations overcomes LLMs' contextual limitations and enables personalized interactions. Recent studies focus on hierarchical memory to model users' multi-granular behavioral patterns via…

Multiagent Systems · Computer Science 2026-01-13 Wenyu Mao , Haosong Tan , Shuchang Liu , Haoyang Liu , Yifan Xu , Huaxiang Ji , Xiang Wang

Memory plays a foundational role in augmenting the reasoning, adaptability, and contextual fidelity of modern Large Language Models and Multi-Modal LLMs. As these models transition from static predictors to interactive systems capable of…

Artificial Intelligence · Computer Science 2026-01-15 Zixia Jia , Jiaqi Li , Yipeng Kang , Yuxuan Wang , Tong Wu , Quansen Wang , Xiaobo Wang , Shuyi Zhang , Junzhe Shen , Qing Li , Siyuan Qi , Yitao Liang , Di He , Zilong Zheng , Song-Chun Zhu

Embodied reasoning is inherently viewpoint-dependent: what is visible, occluded, or reachable depends critically on where the agent stands. However, existing spatial memory systems for embodied agents typically store either multi-view…

Artificial Intelligence · Computer Science 2026-03-17 JooHyun Park , HyeongYeop Kang

Leveraging Large Language Models (LLMs) to harness user-item interaction histories for item generation has emerged as a promising paradigm in generative recommendation. However, the limited context window of LLMs often restricts them to…

Information Retrieval · Computer Science 2025-04-30 Chengbing Wang , Yang Zhang , Fengbin Zhu , Jizhi Zhang , Tianhao Shi , Fuli Feng

Holography is capable of rendering three-dimensional scenes with full-depth control, and delivering transformative experiences across numerous domains, including virtual and augmented reality, education, and communication. However,…

Optics · Physics 2024-09-12 Zhenxing Dong , Yuye Ling , Yan Li , Yikai Su