中文
相关论文

相关论文: Geometry-Corrected Geodesic Motion Modeling with P…

200 篇论文

Selective segmentation is an important application of image processing. In contrast to global segmentation in which all objects are segmented, selective segmentation is used to isolate specific objects in an image and is of particular…

数值分析 · 数学 2019-07-08 Michael Roberts , Ke Chen , Klaus L. Irion

Swept volume computation, the determination of regions occupied by moving objects, is essential in graphics, robotics, and manufacturing. Existing approaches either explicitly track surfaces, suffering from robustness issues under complex…

计算几何 · 计算机科学 2025-09-12 Pengfei Wang , Yuexin Yang , Shuangmin Chen , Shiqing Xin , Changhe Tu , Wenping Wang

We propose a lightweight compressed-domain tracking model that operates directly on video streams, without requiring full RGB video decoding. Using motion vectors and transform coefficients from compressed data, our deep model propagates…

计算机视觉与模式识别 · 计算机科学 2026-02-03 Axel Duché , Clément Chatelain , Gilles Gasso

High-quality scene reconstruction and novel view synthesis based on Gaussian Splatting (3DGS) typically require steady, high-quality photographs, often impractical to capture with handheld cameras. We present a method that adapts to camera…

计算机视觉与模式识别 · 计算机科学 2024-07-18 Otto Seiskari , Jerry Ylilammi , Valtteri Kaatrasalo , Pekka Rantalankila , Matias Turkulainen , Juho Kannala , Esa Rahtu , Arno Solin

Although fisheye cameras are in high demand in many application areas due to their large field of view, many image and video signal processing tasks such as motion compensation suffer from the introduced strong radial distortions. A…

计算机视觉与模式识别 · 计算机科学 2022-03-01 Andy Regensky , Christian Herglotz , André Kaup

Efficient point cloud compression is essential for applications like virtual and mixed reality, autonomous driving, and cultural heritage. This paper proposes a deep learning-based inter-frame encoding scheme for dynamic point cloud…

计算机视觉与模式识别 · 计算机科学 2024-09-04 Anique Akhtar , Zhu Li , Geert Van der Auwera

Medical applications like Computed Tomography (CT) or Magnetic Resonance Tomography (MRT) often require an efficient scalable representation of their huge output volumes in the further processing chain of medical routine. A downscaled…

图像与视频处理 · 电气工程与系统科学 2023-01-13 Daniela Lanz , André Kaup

The minimal geodesic models based on the Eikonal equations are capable of finding suitable solutions in various image segmentation scenarios. Existing geodesic-based segmentation approaches usually exploit image features in conjunction with…

计算机视觉与模式识别 · 计算机科学 2022-11-28 Da Chen , Jean-Marie Mirebeau , Minglei Shu , Xuecheng Tai , Laurent D. Cohen

Recent advances in video generation have enabled the synthesis of high-quality and visually realistic clips using diffusion transformer models. However, most existing approaches operate purely in the 2D pixel space and lack explicit…

计算机视觉与模式识别 · 计算机科学 2025-12-04 Yunpeng Bai , Shaoheng Fang , Chaohui Yu , Fan Wang , Qixing Huang

Learned video compression has recently emerged as an essential research topic in developing advanced video compression technologies, where motion compensation is considered one of the most challenging issues. In this paper, we propose a…

图像与视频处理 · 电气工程与系统科学 2023-06-30 Huairui Wang , Zhenzhong Chen , Chang Wen Chen

Deep video compression has made remarkable process in recent years, with the majority of advancements concentrated on P-frame coding. Although efforts to enhance B-frame coding are ongoing, their compression performance is still far behind…

计算机视觉与模式识别 · 计算机科学 2025-01-22 Xihua Sheng , Li Li , Dong Liu , Shiqi Wang

Video-based human pose transfer is a video-to-video generation task that animates a plain source human image based on a series of target human poses. Considering the difficulties in transferring highly structural patterns on the garments…

计算机视觉与模式识别 · 计算机科学 2023-07-19 Wing-Yin Yu , Lai-Man Po , Ray C. C. Cheung , Yuzhi Zhao , Yu Xue , Kun Li

Efficient point cloud compression is fundamental to enable the deployment of virtual and mixed reality applications, since the number of points to code can range in the order of millions. In this paper, we present a novel data-driven…

计算机视觉与模式识别 · 计算机科学 2020-02-19 Maurice Quach , Giuseppe Valenzise , Frederic Dufaux

We introduce the method of compressed dynamic mode decomposition (cDMD) for background modeling. The dynamic mode decomposition (DMD) is a regression technique that integrates two of the leading data analysis methods in use today: Fourier…

计算机视觉与模式识别 · 计算机科学 2016-12-13 N. Benjamin Erichson , Steven L. Brunton , J. Nathan Kutz

This paper shows that motion vectors representing the true motion of an object in a scene can be exploited to improve the encoding process of computer generated video sequences. Therefore, a set of sequences is presented for which the true…

图像与视频处理 · 电气工程与系统科学 2023-09-14 Christian Herglotz , David Müller , Andreas Weinlich , Frank Bauer , Michael Ortner , Marc Stamminger , André Kaup

Video generation models have progressed tremendously through large latent diffusion transformers trained with rectified flow techniques. Yet these models still struggle with geometric inconsistencies, unstable motion, and visual artifacts…

计算机视觉与模式识别 · 计算机科学 2025-10-27 Orest Kupyn , Fabian Manhardt , Federico Tombari , Christian Rupprecht

Human motion transfer aims at animating a static source image with a driving video. While recent advances in one-shot human motion transfer have led to significant improvement in results, it remains challenging for methods with 2D body…

计算机视觉与模式识别 · 计算机科学 2024-12-10 Yuzhu Ji , Chuanxia Zheng , Tat-Jen Cham

The compression of geometric structures is a relatively new field of data compression. Since about 1995, several articles have dealt with the coding of meshes, using for most of them the following approach: the vertices of the mesh are…

计算几何 · 计算机科学 2007-05-23 Olivier Devillers , Pierre-Maris Gandoin

Encoding textural content remains a challenge for current standardised video codecs. It is therefore beneficial to understand video textures in terms of both their spatio-temporal characteristics and their encoding statistics in order to…

图像与视频处理 · 电气工程与系统科学 2021-02-09 Angeliki V. Katsenou , Mariana Afonso , David R. Bull

Detection of moving objects in videos is a crucial step towards successful surveillance and monitoring applications. A key component for such tasks is called background subtraction and tries to extract regions of interest from the image…

计算机视觉与模式识别 · 计算机科学 2017-10-30 Konstantinos Makantasis , Antonis Nikitakis , Anastasios Doulamis , Nikolaos Doulamis , Yannis Papaefstathiou