English
Related papers

Related papers: POE: Acoustic Soft Robotic Proprioception for Omni…

200 papers

We propose a self-supervised model producing 3D anatomical positional embeddings (APE) of individual medical image voxels. APE encodes voxels' anatomical closeness, i.e., voxels of the same organ or nearby organs always have closer…

Computer Vision and Pattern Recognition · Computer Science 2024-09-17 Mikhail Goncharov , Valentin Samokhin , Eugenia Soboleva , Roman Sokolov , Boris Shirokikh , Mikhail Belyaev , Anvar Kurmukov , Ivan Oseledets

Applying Transformers to irregular time-series typically requires specializations to their baseline architecture, which can result in additional computational overhead and increased method complexity. We present the Rotary Masked…

Machine Learning · Computer Science 2026-05-13 Uros Zivanovic , Serafina Di Gioia , Andre Scaffidi , Martín de los Rios , Gabriella Contardo , Roberto Trotta

Orthogonal-strip planar high-purity germanium (HPGe) detectors can reconstruct three-dimensional (3D) positions of photon interactions through analysis of parameters extracted from multiple charge signals. The conventional method…

Instrumentation and Detectors · Physics 2025-07-25 Qiuli Zhang , Peng Zhang , Wenhhan Dai , Mingxin Yang , Yang Tian , Ming Zeng , Hao Ma , Zhi Zeng

The potential of facial expression reconstruction technology is significant, with applications in various fields such as human-computer interaction, affective computing, and virtual reality. Recent studies have proposed using ear-worn…

Human-Computer Interaction · Computer Science 2025-11-20 Xianrong Yao , Lingde Hu , Dong She , Yincheng Jin , Yang Gao , Zhanpeng Jin

Locomotive soft robots (SoRos) have gained prominence due to their adaptability. Traditional locomotive SoRo design is based on limb structures inspired by biological organisms and requires human intervention. Evolutionary robotics,…

Computational Engineering, Finance, and Science · Computer Science 2024-07-26 Hiroki Kobayashi , Farzad Gholami , S. Macrae Montgomery , Masato Tanaka , Liang Yue , Changyoung Yuhn , Yuki Sato , Atsushi Kawamoto , H. Jerry Qi , Tsuyoshi Nomura

Recent general-purpose audio representations show state-of-the-art performance on various audio tasks. These representations are pre-trained by self-supervised learning methods that create training signals from the input. For example,…

Audio and Speech Processing · Electrical Eng. & Systems 2023-03-09 Daisuke Niizumi , Daiki Takeuchi , Yasunori Ohishi , Noboru Harada , Kunio Kashino

This paper is concerned with the problem of estimating (interpolating and smoothing) the shape (pose and the six modes of deformation) of a slender flexible body from multiple camera measurements. This problem is important in both biology,…

Accurate orientation estimation is a crucial component of 3D molecular structure reconstruction, both in single-particle cryo-electron microscopy (cryo-EM) and in the increasingly popular field of cryo-electron tomography (cryo-ET). The…

Applications · Statistics 2026-02-25 Sheng Xu , Amnon Balanov , Amit Singer , Tamir Bendory

Continuum robots have emerged as a promising technology in the medical field due to their potential of accessing deep sited locations of the human body with low surgical trauma. When deriving physics-based models for these robots,…

Robotics · Computer Science 2024-05-27 Matthias K. Hoffmann , Julian Mühlenhoff , Zhaoheng Ding , Thomas Sattel , Kathrin Flaßkamp

Multimodal generative models should be able to learn a meaningful latent representation that enables a coherent joint generation of all modalities (e.g., images and text). Many applications also require the ability to accurately sample…

Machine Learning · Computer Science 2021-08-02 Svetlana Kutuzova , Oswin Krause , Douglas McCloskey , Mads Nielsen , Christian Igel

Personalizing medical devices such as lower limb wearable robots is challenging. While the initial feasibility of automating the process of knee prosthesis control parameter tuning has been demonstrated in a principled way, the next…

Systems and Control · Electrical Eng. & Systems 2021-06-08 Minhan Li , Yue Wen , Xiang Gao , Jennie Si , He Helen Huang

Visual odometry (VO) aims to estimate camera poses from visual inputs -- a fundamental building block for many applications such as VR/AR and robotics. This work focuses on monocular RGB VO where the input is a monocular RGB video without…

Computer Vision and Pattern Recognition · Computer Science 2025-04-09 Junda Cheng , Zhipeng Cai , Zhaoxing Zhang , Wei Yin , Matthias Muller , Michael Paulitsch , Xin Yang

Mixture-of-Experts (MoE) architectures are evolving towards finer granularity to improve parameter efficiency. However, existing MoE designs face an inherent trade-off between the granularity of expert specialization and hardware execution…

Computation and Language · Computer Science 2026-02-06 Jingze Shi , Zhangyang Peng , Yizhang Zhu , Yifan Wu , Guang Liu , Yuyu Luo

Recent talking head synthesis works typically adopt speech features extracted from large-scale pre-trained acoustic models. However, the intrinsic many-to-many relationship between speech and lip motion causes phoneme-viseme alignment…

Graphics · Computer Science 2025-10-16 Yihuan Huang , Jiajun Liu , Yanzhen Ren , Jun Xue , Wuyang Liu , Zongkun Sun

Enable neural networks to capture 3D geometrical-aware features is essential in multi-view based vision tasks. Previous methods usually encode the 3D information of multi-view stereo into the 2D features. In contrast, we present a novel…

Computer Vision and Pattern Recognition · Computer Science 2023-05-25 Lixin Yang , Jian Xu , Licheng Zhong , Xinyu Zhan , Zhicheng Wang , Kejian Wu , Cewu Lu

We introduce Post-DAE, a post-processing method based on denoising autoencoders (DAE) to improve the anatomical plausibility of arbitrary biomedical image segmentation algorithms. Some of the most popular segmentation methods (e.g. based on…

Computer Vision and Pattern Recognition · Computer Science 2020-06-25 Agostina J Larrazabal , César Martínez , Ben Glocker , Enzo Ferrante

Continuum robots are typically slender and flexible with infinite freedoms in theory, which poses a challenge for their control and application. The shape sensing of continuum robots is vital to realise accuracy control. This letter…

Robotics · Computer Science 2021-03-10 Hao Cheng , Hongji Shang , Bin Lan , Houde Liu , Xueqian Wang , Bin Liang

Achieving a balance between lightweight design and high performance remains a significant challenge for speech enhancement (SE) tasks on resource-constrained devices. Existing state-of-the-art methods, such as MUSE, have established a…

Sound · Computer Science 2025-12-02 Xinxin Tang , Bin Qin , Yufang Li

Personalized speech enhancement (PSE) models can improve the audio quality of teleconferencing systems by adapting to the characteristics of a speaker's voice. However, most existing methods require a separate speaker embedding model to…

Sound · Computer Science 2024-06-17 Tanel Pärnamaa , Ando Saabas

Accurate and real-time three-dimensional (3D) pose estimation is challenging in resource-constrained and dynamic environments owing to its high computational complexity. To address this issue, this study proposes a novel cooperative…

Computer Vision and Pattern Recognition · Computer Science 2025-04-07 Hyun-Ho Choi , Kangsoo Kim , Ki-Ho Lee , Kisong Lee