English
Related papers

Related papers: Moir\'e Video Authentication: A Physical Signature…

200 papers

The rapid progress of image-to-video (I2V) generation models has introduced significant risks by enabling deceptive or malicious video synthesis from a single image. Prior defenses such as I2VGuard attempt to immunize images by inducing…

Computer Vision and Pattern Recognition · Computer Science 2026-04-07 Rohit Chowdhury , Aniruddha Bala , Rohan Jaiswal , Siddharth Roheda

Distinguishing between real and AI-generated images, commonly referred to as 'image detection', presents a timely and significant challenge. Despite extensive research in the (semi-)supervised regime, zero-shot and few-shot solutions have…

Computer Vision and Pattern Recognition · Computer Science 2025-04-23 Jonathan Brokman , Amit Giloni , Omer Hofman , Roman Vainshtein , Hisashi Kojima , Guy Gilboa

Video generators are increasingly evaluated as potential world models, which requires them to encode and understand physical laws. We investigate their representation of a fundamental law: gravity. Out-of-the-box video generators…

Computer Vision and Pattern Recognition · Computer Science 2025-12-02 Varun Varma Thozhiyoor , Shivam Tripathi , Venkatesh Babu Radhakrishnan , Anand Bhattad

Existing person video generation methods either lack the flexibility in controlling both the appearance and motion, or fail to preserve detailed appearance and temporal consistency. In this paper, we tackle the problem of motion transfer…

Computer Vision and Pattern Recognition · Computer Science 2019-08-13 Kun Cheng , Hao-Zhi Huang , Chun Yuan , Lingyiqing Zhou , Wei Liu

Generating realistic human videos remains a challenging task, with the most effective methods currently relying on a human motion sequence as a control signal. Existing approaches often use existing motion extracted from other videos, which…

Computer Vision and Pattern Recognition · Computer Science 2024-12-18 Hsin-Ping Huang , Yang Zhou , Jui-Hsien Wang , Difan Liu , Feng Liu , Ming-Hsuan Yang , Zhan Xu

Artificial Intelligence (AI) software based on transformer model is developed to automatically design gratings for possible integrations in ion traps to perform optical addressing on ions. From the user-defined (x,z) coordinates and…

Optics · Physics 2025-11-27 Yu Dian Lim , Chuan Seng Tan

Recent advancements in AI-generated content have significantly improved the realism of 3D and 4D generation. However, most existing methods prioritize appearance consistency while neglecting underlying physical principles, leading to…

Computer Vision and Pattern Recognition · Computer Science 2025-07-02 Siwei Meng , Yawei Luo , Ping Liu

We present Emu Video, a text-to-video generation model that factorizes the generation into two steps: first generating an image conditioned on the text, and then generating a video conditioned on the text and the generated image. We…

Computer Vision and Pattern Recognition · Computer Science 2024-08-06 Rohit Girdhar , Mannat Singh , Andrew Brown , Quentin Duval , Samaneh Azadi , Sai Saketh Rambhatla , Akbar Shah , Xi Yin , Devi Parikh , Ishan Misra

Producing prompt-faithful videos that preserve a user-specified identity remains challenging: models need to extrapolate facial dynamics from sparse reference while balancing the tension between identity preservation and motion naturalness.…

Computer Vision and Pattern Recognition · Computer Science 2026-01-06 Yixuan Lai , He Wang , Kun Zhou , Tianjia Shao

The field of automatic video generation has received a boost thanks to the recent Generative Adversarial Networks (GANs). However, most existing methods cannot control the contents of the generated video using a text caption, losing their…

Computer Vision and Pattern Recognition · Computer Science 2018-12-06 Shohei Yamamoto , Antonio Tejero-de-Pablos , Yoshitaka Ushiku , Tatsuya Harada

Single-view reference-to-video methods often struggle to preserve identity consistency under large facial-angle variations. This limitation naturally motivates the incorporation of multi-view facial references. However, simply introducing…

Computer Vision and Pattern Recognition · Computer Science 2026-03-24 Bin Hu , Zipeng Qi , Guoxi Huang , Zunnan Xu , Ruicheng Zhang , Chongjie Ye , Jun Zhou , Xiu Li , Jingdong Wang

Inferring full-body poses from Head Mounted Devices, which capture only 3-joint observations from the head and wrists, is a challenging task with wide AR/VR applications. Previous attempts focus on learning one-stage motion mapping and thus…

Computer Vision and Pattern Recognition · Computer Science 2025-05-13 Fangyu Du , Yang Yang , Xuehao Gao , Hongye Hou

The rapid progress of text-to-image models has made AI-generated images increasingly realistic, posing significant challenges for accurate detection of generated content. While training-based detectors often suffer from limited…

Computer Vision and Pattern Recognition · Computer Science 2026-03-24 Ryosuke Sonoda , Ramya Srinivasan

While recent video generation models have achieved significant visual fidelity, they often suffer from the lack of explicit physical controllability and plausibility. To address this, some recent studies attempted to guide the video…

Computer Vision and Pattern Recognition · Computer Science 2025-11-26 Haoze Zhang , Tianyu Huang , Zichen Wan , Xiaowei Jin , Hongzhi Zhang , Hui Li , Wangmeng Zuo

Moir\'e physics plays an important role for the characterization of functional materials and the engineering of physical properties in general, ranging from strain-driven transport phenomena to superconductivity. Here, we report the…

Materials Science · Physics 2023-05-03 L. Richarz , J. He , U. Ludacka , E. Bourret , Z. Yan , A. T. J. van Helvoort , D. Meier

Nowadays, mobile devices have become the natural substitute for the digital camera, as they capture everyday situations easily and quickly, encouraging users to express themselves through images and videos. These videos can be shared across…

Cryptography and Security · Computer Science 2024-02-14 Carlos Quinto Huamán , Ana Lucila Sandoval Orozco , Luis Javier García Villalba

We study the ongoing debate regarding the statistical fidelity of AI-generated data compared to human-generated data in the context of non-verbal communication using full body motion. Concretely, we ask if contemporary generative models…

Computer Vision and Pattern Recognition · Computer Science 2026-03-03 Dragos Costea , Alina Marcu , Cristina Lazar , Marius Leordeanu

The application of video captioning models aims at translating the content of videos by using accurate natural language. Due to the complex nature inbetween object interaction in the video, the comprehensive understanding of spatio-temporal…

Computer Vision and Pattern Recognition · Computer Science 2023-08-15 Yutao Jin , Bin Liu , Jing Wang

This paper presents a novel approach to the digital signing of electronic documents through the use of a camera-based interaction system, single-finger tracking for sign recognition, and multi commands executing hand gestures. The proposed…

Computer Vision and Pattern Recognition · Computer Science 2024-05-20 P. Sarveswarasarma , T. Sathulakjan , V. J. V. Godfrey , Thanuja D. Ambegoda

Generative AI (GenAI) models have revolutionized animation, enabling the synthesis of humans and motion patterns with remarkable visual fidelity. However, generating truly realistic human animation remains a formidable challenge, where even…

Computer Vision and Pattern Recognition · Computer Science 2025-12-23 Ivan DeAndres-Tame , Chengwei Ye , Ruben Tolosana , Ruben Vera-Rodriguez , Shiqi Yu