中文
相关论文

相关论文: Real-Time ESFP: Estimating, Smoothing, Filtering, …

200 篇论文

Purpose: To introduce a combined machine learning (ML) and physics-based image reconstruction framework that enables navigator-free, highly accelerated multishot echo planar imaging (msEPI), and demonstrate its application in…

Standard Vision Transformers flatten 2D images into 1D sequences, disrupting the natural spatial topology. While Rotary Positional Embedding (RoPE) excels in 1D, it inherits this limitation, often treating spatially distant patches (e.g.,…

计算机视觉与模式识别 · 计算机科学 2025-12-05 Yupu Yao , Bowen Yang

We present a new method, called MEsh TRansfOrmer (METRO), to reconstruct 3D human pose and mesh vertices from a single image. Our method uses a transformer encoder to jointly model vertex-vertex and vertex-joint interactions, and outputs 3D…

计算机视觉与模式识别 · 计算机科学 2021-06-16 Kevin Lin , Lijuan Wang , Zicheng Liu

A representation gap exists between grasp synthesis for rigid and soft grippers. Anygrasp [1] and many other grasp synthesis methods are designed for rigid parallel grippers, and adapting them to soft grippers often fails to capture their…

机器人学 · 计算机科学 2026-02-20 Tanisha Parulekar , Ge Shi , Josh Pinskier , David Howard , Jen Jen Chung

3D hand pose estimation based on RGB images has been studied for a long time. Most of the studies, however, have performed frame-by-frame estimation based on independent static images. In this paper, we attempt to not only consider the…

计算机视觉与模式识别 · 计算机科学 2020-07-13 John Yang , Hyung Jin Chang , Seungeui Lee , Nojun Kwak

This paper introduces a novel framework for continuous 3D trajectory optimization in cluttered environments, leveraging online neural Euclidean Signed Distance Fields (ESDFs). Unlike prior approaches that rely on discretized ESDF grids with…

机器人学 · 计算机科学 2025-09-25 Guillermo Gil , Jose Antonio Cobano , Luis Merino , Fernando Caballero

Monocular depth estimation is still an open challenge due to the ill-posed nature of the problem at hand. Deep learning based techniques have been extensively studied and proved capable of producing acceptable depth estimation accuracy even…

图像与视频处理 · 电气工程与系统科学 2022-04-15 Mazen Mel , Muhammad Siddiqui , Pietro Zanuttigh

We present the first real-time method to capture the full global 3D skeletal pose of a human in a stable, temporally consistent manner using a single RGB camera. Our method combines a new convolutional neural network (CNN) based pose…

Recent years have witnessed tremendous progress in the 3D reconstruction of dynamic humans from a monocular video with the advent of neural rendering techniques. This task has a wide range of applications, including the creation of virtual…

计算机视觉与模式识别 · 计算机科学 2024-09-24 Kanghao Chen , Zeyu Wang , Lin Wang

Segmenting thin structures like infrastructure cracks and anatomical vessels is a task hampered by topology-sensitive geometry, high annotation costs, and poor generalization across domains. Existing methods address these challenges in…

计算机视觉与模式识别 · 计算机科学 2026-03-17 Babak Asadi , Peiyang Wu , Mani Golparvar-Fard , Viraj Shah , Ramez Hajj

We describe a new spatio-temporal video autoencoder, based on a classic spatial image autoencoder and a novel nested temporal autoencoder. The temporal encoder is represented by a differentiable visual memory composed of convolutional long…

机器学习 · 计算机科学 2016-09-02 Viorica Patraucean , Ankur Handa , Roberto Cipolla

Rectified Flow (RF) models achieve state-of-the-art generation quality, yet controlling them for precise tasks -- such as semantic editing or blind image recovery -- remains a challenge. Current approaches bifurcate into inversion-based…

机器学习 · 计算机科学 2026-03-09 Vansh Bansal , James G Scott

Video-based human motion transfer creates video animations of humans following a source motion. Current methods show remarkable results for tightly-clad subjects. However, the lack of temporally consistent handling of plausible clothing…

We present Reusable Motion prior (ReMP), an effective motion prior that can accurately track the temporal evolution of motion in various downstream tasks. Inspired by the success of foundation models, we argue that a robust spatio-temporal…

计算机视觉与模式识别 · 计算机科学 2024-11-15 Hojun Jang , Young Min Kim

Learning based 6D object pose estimation methods rely on computing large intermediate pose representations and/or iteratively refining an initial estimation with a slow render-compare pipeline. This paper introduces a novel method we call…

计算机视觉与模式识别 · 计算机科学 2022-10-24 Pedro Castro , Tae-Kyun Kim

In this work, we propose a new head-tracking solution for human-machine real-time interaction with virtual 3D environments. This solution leverages RGBD data to compute virtual camera pose according to the movements of the user's head. The…

计算机视觉与模式识别 · 计算机科学 2021-10-28 Abdenour Amamra

4D face reconstruction from a single camera is a challenging task, especially when it is required to be performed in real time. We demonstrate a system of our own implementation that solves this task accurately and runs in real time on a…

计算机视觉与模式识别 · 计算机科学 2020-06-19 Mohammad Rami Koujan , Nikolai Dochev , Anastasios Roussos

A compact large-range six-degrees-of-freedom (six-DOF) parallel positioning system with high resolution, high resonant frequency, and high repeatability was proposed. It mainly consists of three identical kinematic sections. Each kinematic…

机器人学 · 计算机科学 2024-10-28 Mohammadali Ghafarian , Bijan Shirinzadeh , Ammar Al-Jodah

Modeling deformable objects - especially continuum materials - in a way that is physically plausible, generalizable, and data-efficient remains challenging across 3D vision, graphics, and robotic manipulation. Many existing methods…

机器人学 · 计算机科学 2026-01-27 Yunuo Chen , Yafei Hu , Lingfeng Sun , Tushar Kusnur , Laura Herlant , Chenfanfu Jiang

Pipe inspection is a critical task for many industries and infrastructure of a city. The 3D information of a pipe can be used for revealing the deformation of the pipe surface and position of the camera during the inspection. In this paper,…

计算机视觉与模式识别 · 计算机科学 2020-07-06 Sho kagami , Hajime Taira , Naoyuki Miyashita , Akihiko Torii , Masatoshi Okutomi