中文
相关论文

相关论文: HaSPeR: An Image Repository for Hand Shadow Puppet…

200 篇论文

Handwritten Mathematical Expression Recognition (HMER) has wide applications in human-machine interaction scenarios, such as digitized education and automated offices. Recently, sequence-based models with encoder-decoder architectures have…

计算机视觉与模式识别 · 计算机科学 2024-07-11 Tongkun Guan , Chengyu Lin , Wei Shen , Xiaokang Yang

We present HiFiHR, a high-fidelity hand reconstruction approach that utilizes render-and-compare in the learning-based framework from a single image, capable of generating visually plausible and accurate 3D hand meshes while recovering…

计算机视觉与模式识别 · 计算机科学 2023-08-29 Jiayin Zhu , Zhuoran Zhao , Linlin Yang , Angela Yao

3D scene graphs have recently emerged as a powerful high-level representation of 3D environments. A 3D scene graph describes the environment as a layered graph where nodes represent spatial concepts at multiple levels of abstraction and…

机器人学 · 计算机科学 2022-06-22 Nathan Hughes , Yun Chang , Luca Carlone

Humans can leverage physical interaction to teach robot arms. As the human kinesthetically guides the robot through demonstrations, the robot learns the desired task. While prior works focus on how the robot learns, it is equally important…

Superpixels offer a compact image representation by grouping pixels into coherent regions. Recent methods have reached a plateau in terms of segmentation accuracy by generating noisy superpixel shapes. Moreover, most existing approaches…

计算机视觉与模式识别 · 计算机科学 2026-04-14 Julien Walther , Rémi Giraud , Michaël Clément

A key challenge in the task of human pose and shape estimation is occlusion, including self-occlusions, object-human occlusions, and inter-person occlusions. The lack of diverse and accurate pose and shape training data becomes a major…

计算机视觉与模式识别 · 计算机科学 2022-03-02 Kaibing Yang , Renshu Gu , Maoyu Wang , Masahiro Toyoura , Gang Xu

Despite recent advancements, text-to-image generation models often produce images containing artifacts, especially in human figures. These artifacts appear as poorly generated human bodies, including distorted, missing, or extra body parts,…

计算机视觉与模式识别 · 计算机科学 2025-03-11 Kaihong Wang , Lingzhi Zhang , Jianming Zhang

This paper presents an approach to forecast future presence and location of human hands and objects. Given an image frame, the goal is to predict what objects will appear in the future frame (e.g., 5 seconds later) and where they will be…

计算机视觉与模式识别 · 计算机科学 2018-08-24 Chenyou Fan , Jangwon Lee , Michael S. Ryoo

Projections, or dimensionality reduction methods, are techniques of choice for the visual exploration of high-dimensional data. Many such techniques exist, each one of them having a distinct visual signature - i.e., a recognizable way to…

人机交互 · 计算机科学 2026-02-25 Alister Machado , Alexandru Telea , Michael Behrisch

The human hand is the main medium through which we interact with our surroundings, making its digitization an important problem. While there are several works modeling the geometry of hands, little attention has been paid to capturing…

Reconstructing soft tissues from stereo endoscope videos is an essential prerequisite for many medical applications. Previous methods struggle to produce high-quality geometry and appearance due to their inadequate representations of 3D…

计算机视觉与模式识别 · 计算机科学 2023-09-06 Ruyi Zha , Xuelian Cheng , Hongdong Li , Mehrtash Harandi , Zongyuan Ge

Advances in 3D reconstruction using neural rendering have enabled high-quality 3D capture. However, they often fail when the input imagery is corrupted by motion blur, due to fast motion of the camera or the objects in the scene. This work…

图像与视频处理 · 电气工程与系统科学 2025-06-30 Sai Sri Teja , Sreevidya Chintalapati , Vinayak Gupta , Mukund Varma T , Haejoon Lee , Aswin Sankaranarayanan , Kaushik Mitra

We present NBAvatar - a method for realistic rendering of head avatars handling non-rigid deformations caused by hand-face interaction. We introduce a novel representation for animated avatars by combining the training of oriented planar…

计算机视觉与模式识别 · 计算机科学 2026-03-13 David Svitov , Mahtab Dahaghin

For robot manipulation, a complete and accurate object shape is desirable. Here, we present a method that combines visual and haptic reconstruction in a closed-loop pipeline. From an initial viewpoint, the object shape is reconstructed…

机器人学 · 计算机科学 2024-09-11 Lukas Rustler , Jiri Matas , Matej Hoffmann

The advanced role-playing capabilities of Large Language Models (LLMs) have enabled rich interactive scenarios, yet existing research in social interactions neglects hallucination while struggling with poor generalizability and implicit…

计算与语言 · 计算机科学 2025-06-04 Chuyi Kong , Ziyang Luo , Hongzhan Lin , Zhiyuan Fan , Yaxin Fan , Yuxi Sun , Jing Ma

When humans and robotic agents coexist in an environment, scene understanding becomes crucial for the agents to carry out various downstream tasks like navigation and planning. Hence, an agent must be capable of localizing and identifying…

计算机视觉与模式识别 · 计算机科学 2025-06-25 Mrunmai Vivek Phatak , Julian Lorenz , Nico Hörmann , Jörg Hähner , Rainer Lienhart

Video compression technology is essential for transmitting and storing videos. Many video compression methods reduce information in videos by removing high-frequency components and utilizing similarities between frames. Alternatively, the…

计算机视觉与模式识别 · 计算机科学 2024-07-09 Taiga Hayami , Hiroshi Watanabe

Graph Transformers have recently achieved remarkable progress in graph representation learning by capturing long-range dependencies through self-attention. However, their quadratic computational complexity and inability to effectively model…

Predicting user influence in social networks is a critical problem, and hypergraphs, as a prevalent higher-order modeling approach, provide new perspectives for this task. However, the absence of explicit cascade or infection probability…

社会与信息网络 · 计算机科学 2025-08-22 Su-Su Zhang , JinFeng Xie , Yang Chen , Min Gao , Cong Li , Chuang Liu , Xiu-Xiu Zhan

Spoof detectors are classifiers that are trained to distinguish spoof fingerprints from bonafide ones. However, state of the art spoof detectors do not generalize well on unseen spoof materials. This study proposes a style transfer based…

计算机视觉与模式识别 · 计算机科学 2019-12-10 Rohit Gajawada , Additya Popli , Tarang Chugh , Anoop Namboodiri , Anil K. Jain