中文
相关论文

相关论文: ARMO: Autoregressive Rigging for Multi-Category Ob…

200 篇论文

Mesh reconstruction is a cornerstone process across various applications, including in-silico trials, digital twins, surgical planning, and navigation. Recent advancements in deep learning have notably enhanced mesh reconstruction speeds.…

图像与视频处理 · 电气工程与系统科学 2025-05-22 Fengting Zhang , Boxu Liang , Qinghao Liu , Min Liu , Xiang Chen , Yaonan Wang

While large reasoning models demonstrate strong performance on complex tasks, they lack the ability to adjust reasoning token usage based on task difficulty. This often leads to the "overthinking" problem -- excessive and unnecessary…

计算与语言 · 计算机科学 2025-10-14 Siye Wu , Jian Xie , Yikai Zhang , Aili Chen , Kai Zhang , Yu Su , Yanghua Xiao

AutoRegressive (AR) models have made notable progress in image generation, with Masked AutoRegressive (MAR) models gaining attention for their efficient parallel decoding. However, MAR models have traditionally underperformed when compared…

计算机视觉与模式识别 · 计算机科学 2025-07-18 Yi Xin , Le Zhuo , Qi Qin , Siqi Luo , Yuewen Cao , Bin Fu , Yangfan He , Hongsheng Li , Guangtao Zhai , Xiaohong Liu , Peng Gao

We present MoRig, a method that automatically rigs character meshes driven by single-view point cloud streams capturing the motion of performing characters. Our method is also able to animate the 3D meshes according to the captured point…

图形学 · 计算机科学 2022-10-19 Zhan Xu , Yang Zhou , Li Yi , Evangelos Kalogerakis

Existing end-to-end approaches of robotic manipulation often lack generalization to unseen objects or tasks due to limited data and poor interpretability. While recent Multimodal Large Language Models (MLLMs) demonstrate strong commonsense…

机器人学 · 计算机科学 2026-03-03 Zilong Xie , Jingyu Gong , Xin Tan , Zhizhong Zhang , Yuan Xie

Estimating agent pose and 3D scene structure from multi-camera rigs is a central task in embodied AI applications such as autonomous driving. Recent learned approaches such as DUSt3R have shown impressive performance in multiview settings.…

计算机视觉与模式识别 · 计算机科学 2025-06-04 Samuel Li , Pujith Kachana , Prajwal Chidananda , Saurabh Nair , Yasutaka Furukawa , Matthew Brown

Automated patient positioning can improve radiology workflow efficiency by reducing the time required for manual table adjustments and scout-based scan planning. We propose a learning-based framework that predicts 3D organ locations and…

计算机视觉与模式识别 · 计算机科学 2026-03-17 Eytan Kats , Kai Geissler , Daniel Mensing , Julien Senegas , Jochen G. Hirsch , Stefan Heldman , Mattias P. Heinrich

Human motion prediction is a challenging and important task in many computer vision application domains. Existing work only implicitly models the spatial structure of the human skeleton. In this paper, we propose a novel approach that…

计算机视觉与模式识别 · 计算机科学 2019-10-22 Emre Aksan , Manuel Kaufmann , Otmar Hilliges

We present RASO, a foundation model designed to Recognize Any Surgical Object, offering robust open-set recognition capabilities across a broad range of surgical procedures and object classes, in both surgical images and videos. RASO…

计算机视觉与模式识别 · 计算机科学 2025-05-07 Jiajie Li , Brian R Quaranto , Chenhui Xu , Ishan Mishra , Ruiyang Qin , Dancheng Liu , Peter C W Kim , Jinjun Xiong

Matrix-based optimizers have attracted growing interest for improving LLM training efficiency, with significant progress centered on orthogonalization/whitening based methods. While yielding substantial performance gains, a fundamental…

机器学习 · 计算机科学 2026-02-10 Wenbo Gong , Javier Zazo , Qijun Luo , Puqian Wang , James Hensman , Chao Ma

Medical image registration is crucial for various clinical and research applications including disease diagnosis or treatment planning which require alignment of images from different modalities, time points, or subjects. Traditional…

计算机视觉与模式识别 · 计算机科学 2025-12-02 Ahsan Raza Siyal , Markus Haltmeier , Ruth Steiger , Malik Galijasevic , Elke Ruth Gizewski , Astrid Ellen Grams

We propose a general algorithm for non-conforming adaptive mesh refinement (AMR) of unstructured meshes in high-order finite element codes. Our focus is on h-refinement with a fixed polynomial order. The algorithm handles triangular,…

数值分析 · 计算机科学 2019-05-13 Jakub Červený , Veselin Dobrev , Tzanio Kolev

Online micro gesture recognition from hand skeletons is critical for VR/AR interaction but faces challenges due to limited public datasets and task-specific algorithms. Micro gestures involve subtle motion patterns, which make constructing…

计算机视觉与模式识别 · 计算机科学 2026-03-31 Haochen Chang , Pengfei Ren , Buyuan Zhang , Da Li , Tianhao Han , Haoyang Zhang , Liang Xie , Hongbo Chen , Erwei Yin

Most existing 3D assembly methods treat the problem as pure pose estimation, rearranging observed parts via rigid transformations. In contrast, human assembly naturally couples structural reasoning with holistic shape inference. Inspired by…

计算机视觉与模式识别 · 计算机科学 2026-02-27 Zeyu Jiang , Sihang Li , Siqi Tan , Chenyang Xu , Juexiao Zhang , Julia Galway-Witham , Xue Wang , Scott A. Williams , Radu Iovita , Chen Feng , Jing Zhang

The integration of large language models (LLMs) into robotic systems has accelerated progress in embodied artificial intelligence, yet current approaches remain constrained by existing robotic architectures, particularly serial mechanisms.…

机器人学 · 计算机科学 2025-10-07 Guanglu Jia , Ceng Zhang , Gregory S. Chirikjian

Although commercial and open-source software exist to reconstruct a static object from a sequence recorded with an RGB-D sensor, there is a lack of tools that build rigged models of articulated objects that deform realistically and can be…

计算机视觉与模式识别 · 计算机科学 2016-09-12 Dimitrios Tzionas , Juergen Gall

Medical image registration is critical for aligning anatomical structures across imaging modalities such as computed tomography (CT), magnetic resonance imaging (MRI), and ultrasound. Among existing techniques, non-rigid registration (NRR)…

图像与视频处理 · 电气工程与系统科学 2025-12-17 Sneha Sree C. , Dattesh Shanbhag , Sudhanya Chatterjee

Time consumption and the complexity of manual layout design make automated layout generation a critical task, especially for multiple applications across different mobile devices. Existing graph-based layout generation approaches suffer…

计算机视觉与模式识别 · 计算机科学 2025-05-27 Jiongchao Jin , Shengchu Zhao , Dajun Chen , Wei Jiang , Yong Li

Current auto-regressive mesh generation methods suffer from issues such as incompleteness, insufficient detail, and poor generalization. In this paper, we propose an Auto-regressive Auto-encoder (ArAE) model capable of generating…

计算机视觉与模式识别 · 计算机科学 2024-09-27 Jiaxiang Tang , Zhaoshuo Li , Zekun Hao , Xian Liu , Gang Zeng , Ming-Yu Liu , Qinsheng Zhang

Robots excel at avoiding obstacles but struggle to traverse complex 3-D terrain with cluttered large obstacles. By contrast, insects like cockroaches excel at doing so. Recent research in our lab elucidated how locomotor transitions emerge…

机器人学 · 计算机科学 2025-09-26 Jonathan Mi , Yaqing Wang , Chen Li