中文
相关论文

相关论文: POE: Acoustic Soft Robotic Proprioception for Omni…

200 篇论文

We propose a self-supervised model producing 3D anatomical positional embeddings (APE) of individual medical image voxels. APE encodes voxels' anatomical closeness, i.e., voxels of the same organ or nearby organs always have closer…

计算机视觉与模式识别 · 计算机科学 2024-09-17 Mikhail Goncharov , Valentin Samokhin , Eugenia Soboleva , Roman Sokolov , Boris Shirokikh , Mikhail Belyaev , Anvar Kurmukov , Ivan Oseledets

Applying Transformers to irregular time-series typically requires specializations to their baseline architecture, which can result in additional computational overhead and increased method complexity. We present the Rotary Masked…

Orthogonal-strip planar high-purity germanium (HPGe) detectors can reconstruct three-dimensional (3D) positions of photon interactions through analysis of parameters extracted from multiple charge signals. The conventional method…

仪器与探测器 · 物理学 2025-07-25 Qiuli Zhang , Peng Zhang , Wenhhan Dai , Mingxin Yang , Yang Tian , Ming Zeng , Hao Ma , Zhi Zeng

The potential of facial expression reconstruction technology is significant, with applications in various fields such as human-computer interaction, affective computing, and virtual reality. Recent studies have proposed using ear-worn…

人机交互 · 计算机科学 2025-11-20 Xianrong Yao , Lingde Hu , Dong She , Yincheng Jin , Yang Gao , Zhanpeng Jin

Locomotive soft robots (SoRos) have gained prominence due to their adaptability. Traditional locomotive SoRo design is based on limb structures inspired by biological organisms and requires human intervention. Evolutionary robotics,…

Recent general-purpose audio representations show state-of-the-art performance on various audio tasks. These representations are pre-trained by self-supervised learning methods that create training signals from the input. For example,…

音频与语音处理 · 电气工程与系统科学 2023-03-09 Daisuke Niizumi , Daiki Takeuchi , Yasunori Ohishi , Noboru Harada , Kunio Kashino

This paper is concerned with the problem of estimating (interpolating and smoothing) the shape (pose and the six modes of deformation) of a slender flexible body from multiple camera measurements. This problem is important in both biology,…

Accurate orientation estimation is a crucial component of 3D molecular structure reconstruction, both in single-particle cryo-electron microscopy (cryo-EM) and in the increasingly popular field of cryo-electron tomography (cryo-ET). The…

应用统计 · 统计学 2026-02-25 Sheng Xu , Amnon Balanov , Amit Singer , Tamir Bendory

Continuum robots have emerged as a promising technology in the medical field due to their potential of accessing deep sited locations of the human body with low surgical trauma. When deriving physics-based models for these robots,…

机器人学 · 计算机科学 2024-05-27 Matthias K. Hoffmann , Julian Mühlenhoff , Zhaoheng Ding , Thomas Sattel , Kathrin Flaßkamp

Multimodal generative models should be able to learn a meaningful latent representation that enables a coherent joint generation of all modalities (e.g., images and text). Many applications also require the ability to accurately sample…

机器学习 · 计算机科学 2021-08-02 Svetlana Kutuzova , Oswin Krause , Douglas McCloskey , Mads Nielsen , Christian Igel

Personalizing medical devices such as lower limb wearable robots is challenging. While the initial feasibility of automating the process of knee prosthesis control parameter tuning has been demonstrated in a principled way, the next…

系统与控制 · 电气工程与系统科学 2021-06-08 Minhan Li , Yue Wen , Xiang Gao , Jennie Si , He Helen Huang

Visual odometry (VO) aims to estimate camera poses from visual inputs -- a fundamental building block for many applications such as VR/AR and robotics. This work focuses on monocular RGB VO where the input is a monocular RGB video without…

计算机视觉与模式识别 · 计算机科学 2025-04-09 Junda Cheng , Zhipeng Cai , Zhaoxing Zhang , Wei Yin , Matthias Muller , Michael Paulitsch , Xin Yang

Mixture-of-Experts (MoE) architectures are evolving towards finer granularity to improve parameter efficiency. However, existing MoE designs face an inherent trade-off between the granularity of expert specialization and hardware execution…

计算与语言 · 计算机科学 2026-02-06 Jingze Shi , Zhangyang Peng , Yizhang Zhu , Yifan Wu , Guang Liu , Yuyu Luo

Recent talking head synthesis works typically adopt speech features extracted from large-scale pre-trained acoustic models. However, the intrinsic many-to-many relationship between speech and lip motion causes phoneme-viseme alignment…

图形学 · 计算机科学 2025-10-16 Yihuan Huang , Jiajun Liu , Yanzhen Ren , Jun Xue , Wuyang Liu , Zongkun Sun

Enable neural networks to capture 3D geometrical-aware features is essential in multi-view based vision tasks. Previous methods usually encode the 3D information of multi-view stereo into the 2D features. In contrast, we present a novel…

计算机视觉与模式识别 · 计算机科学 2023-05-25 Lixin Yang , Jian Xu , Licheng Zhong , Xinyu Zhan , Zhicheng Wang , Kejian Wu , Cewu Lu

We introduce Post-DAE, a post-processing method based on denoising autoencoders (DAE) to improve the anatomical plausibility of arbitrary biomedical image segmentation algorithms. Some of the most popular segmentation methods (e.g. based on…

计算机视觉与模式识别 · 计算机科学 2020-06-25 Agostina J Larrazabal , César Martínez , Ben Glocker , Enzo Ferrante

Continuum robots are typically slender and flexible with infinite freedoms in theory, which poses a challenge for their control and application. The shape sensing of continuum robots is vital to realise accuracy control. This letter…

机器人学 · 计算机科学 2021-03-10 Hao Cheng , Hongji Shang , Bin Lan , Houde Liu , Xueqian Wang , Bin Liang

Achieving a balance between lightweight design and high performance remains a significant challenge for speech enhancement (SE) tasks on resource-constrained devices. Existing state-of-the-art methods, such as MUSE, have established a…

声音 · 计算机科学 2025-12-02 Xinxin Tang , Bin Qin , Yufang Li

Personalized speech enhancement (PSE) models can improve the audio quality of teleconferencing systems by adapting to the characteristics of a speaker's voice. However, most existing methods require a separate speaker embedding model to…

声音 · 计算机科学 2024-06-17 Tanel Pärnamaa , Ando Saabas

Accurate and real-time three-dimensional (3D) pose estimation is challenging in resource-constrained and dynamic environments owing to its high computational complexity. To address this issue, this study proposes a novel cooperative…

计算机视觉与模式识别 · 计算机科学 2025-04-07 Hyun-Ho Choi , Kangsoo Kim , Ki-Ho Lee , Kisong Lee