English
Related papers

Related papers: Calipso: Physics-based Image and Video Editing thr…

200 papers

We present Kamino, a GPU-based physics solver for massively parallel simulations of heterogeneous highly-coupled mechanical systems. Implemented in Python using NVIDIA Warp and integrated into the Newton framework, it enables the…

This paper presents \emph{ControlVideo} for text-driven video editing -- generating a video that aligns with a given text while preserving the structure of the source video. Building on a pre-trained text-to-image diffusion model,…

Computer Vision and Pattern Recognition · Computer Science 2023-11-29 Min Zhao , Rongzhen Wang , Fan Bao , Chongxuan Li , Jun Zhu

Character video synthesis aims to produce realistic videos of animatable characters within lifelike scenes. As a fundamental problem in the computer vision and graphics community, 3D works typically require multi-view captures for per-case…

Computer Vision and Pattern Recognition · Computer Science 2025-06-12 Yifang Men , Yuan Yao , Miaomiao Cui , Liefeng Bo

While language-guided image manipulation has made remarkable progress, the challenge of how to instruct the manipulation process faithfully reflecting human intentions persists. An accurate and comprehensive description of a manipulation…

Computer Vision and Pattern Recognition · Computer Science 2023-08-03 Yasheng Sun , Yifan Yang , Houwen Peng , Yifei Shen , Yuqing Yang , Han Hu , Lili Qiu , Hideki Koike

We introduce CRISP, a method that recovers simulatable human motion and scene geometry from monocular video. Prior work on joint human-scene reconstruction relies on data-driven priors and joint optimization with no physics in the loop, or…

Computer Vision and Pattern Recognition · Computer Science 2026-03-03 Zihan Wang , Jiashun Wang , Jeff Tan , Yiwen Zhao , Jessica Hodgins , Shubham Tulsiani , Deva Ramanan

This paper proposes a novel and physically interpretable method for face editing based on arbitrary text prompts. Different from previous GAN-inversion-based face editing methods that manipulate the latent space of GANs, or diffusion-based…

Computer Vision and Pattern Recognition · Computer Science 2023-08-14 Yapeng Meng , Songru Yang , Xu Hu , Rui Zhao , Lincheng Li , Zhenwei Shi , Zhengxia Zou

Existing deep models predict 2D and 3D kinematic poses from video that are approximately accurate, but contain visible errors that violate physical constraints, such as feet penetrating the ground and bodies leaning at extreme angles. In…

Computer Vision and Pattern Recognition · Computer Science 2020-07-27 Davis Rempe , Leonidas J. Guibas , Aaron Hertzmann , Bryan Russell , Ruben Villegas , Jimei Yang

Editing faces in videos is a popular yet challenging aspect of computer vision and graphics, which encompasses several applications including facial attractiveness enhancement, makeup transfer, face replacement, and expression manipulation.…

Computer Vision and Pattern Recognition · Computer Science 2014-08-12 Xiaoyan Li , Dacheng Tao

We propose a spatial-constraint approach for modeling spatial-based interactions and enabling interactive visualizations, which involves the manipulation of visualizations through selection, filtering, navigation, arrangement, and…

Human-Computer Interaction · Computer Science 2024-03-21 Can Liu , Yu Zhang , Cong Wu , Chen Li , Xiaoru Yuan

We consider the problem of estimating an object's physical properties such as mass, friction, and elasticity directly from video sequences. Such a system identification problem is fundamentally ill-posed due to the loss of information…

Humans have the remarkable ability to use held objects as tools to interact with their environment. For this to occur, humans internally estimate how hand movements affect the object's movement. We wish to endow robots with this capability.…

Robotics · Computer Science 2024-07-16 Weiming Zhi , Haozhan Tang , Tianyi Zhang , Matthew Johnson-Roberson

Due to visual ambiguities and inter-person occlusions, existing human pose estimation methods cannot recover plausible close interactions from in-the-wild videos. Even state-of-the-art large foundation models~(\eg, SAM) cannot accurately…

Computer Vision and Pattern Recognition · Computer Science 2025-07-04 Buzhen Huang , Chen Li , Chongyang Xu , Dongyue Lu , Jinnan Chen , Yangang Wang , Gim Hee Lee

Even though large-scale text-to-image generative models show promising performance in synthesizing high-quality images, applying these models directly to image editing remains a significant challenge. This challenge is further amplified in…

Computer Vision and Pattern Recognition · Computer Science 2025-02-03 Shutong Jin , Ruiyu Wang , Florian T. Pokorny

We propose a general, prior-free approach for the uncalibrated non-rigid structure-from-motion problem for modelling and analysis of non-rigid objects such as human faces. The word general refers to an approach that recovers the non-rigid…

Computer Vision and Pattern Recognition · Computer Science 2018-11-26 Sami Sebastian Brandt , Hanno Ackermann , Stella Grasshof

Extending 3D Gaussian Splatting (3DGS) to 4D physical simulation remains challenging. Based on the Material Point Method (MPM), existing methods either rely on manual parameter tuning or distill dynamics from video diffusion models,…

Computer Vision and Pattern Recognition · Computer Science 2026-02-03 Yikun Ma , Yiqing Li , Jingwen Ye , Zhongkai Wu , Weidong Zhang , Lin Gao , Zhi Jin

We propose a generative approach to physics-based motion capture. Unlike prior attempts to incorporate physics into tracking that assume the subject and scene geometry are calibrated and known a priori, our approach is automatic and online.…

Computer Vision and Pattern Recognition · Computer Science 2018-12-05 Micha Livne , Leonid Sigal , Marcus A. Brubaker , David J. Fleet

Human shape editing enables controllable transformation of a person's body shape, such as thin, muscular, or overweight, while preserving pose, identity, clothing, and background. Unlike human pose editing, which has advanced rapidly, shape…

Computer Vision and Pattern Recognition · Computer Science 2025-09-26 Siddharth Khandelwal , Sridhar Kamath , Arjun Jain

In this work, we introduce a novel approach for creating controllable dynamics in 3D-generated Gaussians using casually captured reference videos. Our method transfers the motion of objects from reference videos to a variety of generated 3D…

Computer Vision and Pattern Recognition · Computer Science 2024-07-09 Zhoujie Fu , Jiacheng Wei , Wenhao Shen , Chaoyue Song , Xiaofeng Yang , Fayao Liu , Xulei Yang , Guosheng Lin

We present Pippo, a generative model capable of producing 1K resolution dense turnaround videos of a person from a single casually clicked photo. Pippo is a multi-view diffusion transformer and does not require any additional inputs - e.g.,…

Computer Vision and Pattern Recognition · Computer Science 2025-02-12 Yash Kant , Ethan Weber , Jin Kyu Kim , Rawal Khirodkar , Su Zhaoen , Julieta Martinez , Igor Gilitschenski , Shunsuke Saito , Timur Bagautdinov

Force estimation in human-object interactions is crucial for various fields like ergonomics, physical therapy, and sports science. Traditional methods depend on specialized equipment such as force plates and sensors, which makes accurate…

Computer Vision and Pattern Recognition · Computer Science 2025-03-31 Nandakishor M , Vrinda Govind , Anuradha Puthalath , Anzy L , Swathi P S , Aswathi R , Devaprabha A R , Varsha Raj , Midhuna Krishnan K , Akhila Anilkumar T , Yamuna P
‹ Prev 1 8 9 10 Next ›