中文
相关论文

相关论文: Agentic Pipeline for Self-Synchronized Multiview J…

200 篇论文

Recovering 3D human pose from 2D joints is still a challenging problem, especially without any 3D annotation, video information, or multi-view information. In this paper, we present an unsupervised GAN-based model consisting of multiple…

计算机视觉与模式识别 · 计算机科学 2022-04-14 Yicheng Deng , Cheng Sun , Jiahui Zhu , Yongqi Sun

Purpose: Accurate 3D hand pose estimation supports surgical applications such as skill assessment, robot-assisted interventions, and geometry-aware workflow analysis. However, surgical environments pose severe challenges, including intense…

At-home physiotherapy compliance remains critically low due to a lack of personalized supervision and dynamic feedback. Existing digital health solutions rely on static, pre-recorded video libraries or generic 3D avatars that fail to…

人工智能 · 计算机科学 2026-04-24 Abhishek Dharmaratnakar , Srivaths Ranganathan , Anushree Sinha , Debanshu Das

Positioning patients for scanning and interventional procedures is a critical task that requires high precision and accuracy. The conventional workflow involves manually adjusting the patient support to align the center of the target body…

计算机视觉与模式识别 · 计算机科学 2024-12-10 Zhongpai Gao , Abhishek Sharma , Meng Zheng , Benjamin Planche , Terrence Chen , Ziyan Wu

Purpose: Accurate detection and 6D pose estimation of surgical instruments are crucial for many computer-assisted interventions. However, supervised methods lack flexibility for new or unseen tools and require extensive annotated data. This…

计算机视觉与模式识别 · 计算机科学 2026-03-27 Jonas Hein , Lilian Calvet , Matthias Seibold , Siyu Tang , Marc Pollefeys , Philipp Fürnstahl

Recently, through development of several 3d vision systems, widely used in various applications, medical and biometric fields. Microsoft kinect sensor have been most of used camera among 3d vision systems. Microsoft kinect sensor can obtain…

计算机视觉与模式识别 · 计算机科学 2024-09-16 M. S. Gokmen , M. Akbaba , O. Findik

Manual assembly workers face increasing complexity in their work. Human-centered assistance systems could help, but object recognition as an enabling technology hinders sophisticated human-centered design of these systems. At the same time,…

计算机视觉与模式识别 · 计算机科学 2024-02-15 Christian Jauch , Timo Leitritz , Marco F. Huber

Advances in machine learning and wearable sensors offer new opportunities for capturing and analyzing human movement outside specialized laboratories. Accurate assessment of human movement under real-world conditions is essential for…

Continuous physiological monitoring is central to emergency care, yet deploying trustworthy AI is challenging. While LLMs can translate complex physiological signals into clinical narratives, it is unclear how agentic systems perform…

机器学习 · 计算机科学 2026-03-05 Davide Gabrielli , Paola Velardi , Stefano Faralli , Bardh Prenkaj

Multi-agent systems, e.g., automobiles and UAVs (Unmanned Ariel Vehicles), rely on the precision of onboard sensors to accurately perceive their environment, which in turn depends on the precision of onboard sensors and reliable in-field…

信号处理 · 电气工程与系统科学 2026-04-23 Bichi Zhang , Holger Caesar , Raj Thilak Rajan

This paper addresses the distributed attitude synchronization problem for a network of rigid-body systems on the special orthogonal group SO(3). Each agent measures, in its body frame, its own angular velocity and a set of vectors whose…

系统与控制 · 电气工程与系统科学 2026-04-07 Mouaad Boughellaba , Soulaimane Berkane , Abdelhamid Tayebi

Multimodal Large Language Models (MLLMs) are evolving from passive observers into active agents, solving problems through Visual Expansion (invoking visual tools) and Knowledge Expansion (open-web search). However, existing evaluations fall…

This paper presents a novel framework for real-time human action recognition in industrial contexts, using standard 2D cameras. We introduce a complete pipeline for robust and real-time estimation of human joint kinematics, input to a…

There has been significant progress in machine learning algorithms for human pose estimation that may provide immense value in rehabilitation and movement sciences. However, there remain several challenges to routine use of these tools for…

计算机视觉与模式识别 · 计算机科学 2022-03-17 R. James Cotton

We introduce a method for manifold alignment of different modalities (or domains) of remote sensing images. The problem is recurrent when a set of multitemporal, multisource, multisensor and multiangular images is available. In these…

计算机视觉与模式识别 · 计算机科学 2021-04-19 Devis Tuia , Michele Volpi , Maxime Trolliet , Gustau Camps-Valls

To date, little attention has been given to multi-view 3D human mesh estimation, despite real-life applicability (e.g., motion capture, sport analysis) and robustness to single-view ambiguities. Existing solutions typically suffer from poor…

计算机视觉与模式识别 · 计算机科学 2022-12-13 Xuan Gong , Liangchen Song , Meng Zheng , Benjamin Planche , Terrence Chen , Junsong Yuan , David Doermann , Ziyan Wu

Diagnosing a whole-slide image is an interactive, multi-stage process of changing magnification and moving between fields. Although recent pathology foundation models demonstrated superior performances, practical agentic systems that decide…

计算机视觉与模式识别 · 计算机科学 2025-10-15 Sheng Wang , Ruiming Wu , Charles Herndon , Yihang Liu , Shunsuke Koga , Jeanne Shen , Zhi Huang

In this paper we contribute a simple yet effective approach for estimating 3D poses of multiple people from multi-view images. Our proposed coarse-to-fine pipeline first aggregates noisy 2D observations from multiple camera views into 3D…

计算机视觉与模式识别 · 计算机科学 2021-10-07 Zijian Dong , Jie Song , Xu Chen , Chen Guo , Otmar Hilliges

Active Multi-Object Tracking (AMOT) is a task where cameras are controlled by a centralized system to adjust their poses automatically and collaboratively so as to maximize the coverage of targets in their shared visual field. In AMOT, each…

计算机视觉与模式识别 · 计算机科学 2022-02-23 Zeyu Fang , Jian Zhao , Mingyu Yang , Wengang Zhou , Zhenbo Lu , Houqiang Li

Over the past one hundred years, the classic teaching methodology of "see one, do one, teach one" has governed the surgical education systems worldwide. With the advent of Operation Room 2.0, recording video, kinematic and many other types…

计算机视觉与模式识别 · 计算机科学 2019-07-23 Hassan Ismail Fawaz , Germain Forestier , Jonathan Weber , François Petitjean , Lhassane Idoumghar , Pierre-Alain Muller
‹ 上一页 1 2 3 10 下一页 ›