English
Related papers

Related papers: New starting point registration method for tagged …

200 papers

Automatic robotic facial expression generation is crucial for human-robot interaction, as handcrafted methods based on fixed joint configurations often yield rigid and unnatural behaviors. Although recent automated techniques reduce the…

Robotics · Computer Science 2025-07-04 Dongsheng Yang , Qianying Liu , Wataru Sato , Takashi Minato , Chaoran Liu , Shin'ya Nishida

Capturing voxel-wise spatial correspondence across distinct modalities is crucial for medical image analysis. However, current registration approaches are not practical enough in terms of registration accuracy and clinical applicability. In…

Computer Vision and Pattern Recognition · Computer Science 2025-06-26 Tao Guo , Yinuo Wang , Shihao Shu , Weimin Yuan , Diansheng Chen , Zhouping Tang , Cai Meng , Xiangzhi Bai

Physiological motion, such as cardiac and respiratory motion, during Magnetic Resonance (MR) image acquisition can cause image artifacts. Motion correction techniques have been proposed to compensate for these types of motion during…

Image and Video Processing · Electrical Eng. & Systems 2021-07-21 Thomas Küstner , Jiazhen Pan , Haikun Qi , Gastao Cruz , Christopher Gilliam , Thierry Blu , Bin Yang , Sergios Gatidis , René Botnar , Claudia Prieto

The increasing prevalence of compact UAVs has introduced significant risks to public safety, while traditional drone detection systems are often bulky and costly. To address these challenges, we present TAME, the Temporal Audio-based Mamba…

Sound · Computer Science 2025-03-04 Zhenyuan Xiao , Huanran Hu , Guili Xu , Junwei He

Head movement during scanning impedes activation detection in fMRI studies. Head motion in fMRI acquired using slice-based Echo Planar Imaging (EPI) can be estimated and compensated by aligning the images onto a reference volume through…

Computer Vision and Pattern Recognition · Computer Science 2015-11-12 Yu-Hui Chen , Roni Mittelman , Boklye Kim , Charles Meyer , Alfred Hero

Advances in Deep Learning have made possible reliable landmark tracking of human bodies and faces that can be used for a variety of tasks. We test a recent Computer Vision solution, MediaPipe Holistic (MPH), to find out if its tracking of…

Computer Vision and Pattern Recognition · Computer Science 2024-03-27 Anna Kuznetsova , Vadim Kimmelman

Predictive shift-reduce (PSR) parsing for hyperedge replacement (HR) grammars is very efficient, but restricted to a subclass of unambiguous HR grammars. To overcome this restriction, we have recently extended PSR parsing to generalized PSR…

Formal Languages and Automata Theory · Computer Science 2019-12-23 Mark Minas

This paper addresses the issue of matching rigid and articulated shapes through probabilistic point registration. The problem is recast into a missing data framework where unknown correspondences are handled via mixture models. Adopting a…

Computer Vision and Pattern Recognition · Computer Science 2020-12-10 Radu Horaud , Florence Forbes , Manuel Yguel , Guillaume Dewaele , Jian Zhang

Sequential decision-making and motion planning for robotic manipulation induce combinatorial complexity. For long-horizon tasks, especially when the environment comprises many objects that can be interacted with, planning efficiency becomes…

Robotics · Computer Science 2022-03-08 Cornelius V. Braun , Joaquim Ortiz-Haro , Marc Toussaint , Ozgur S. Oguz

Tracking Any Point (TAP) plays a crucial role in motion analysis. Video-based approaches rely on iterative local matching for tracking, but they assume linear motion during the blind time between frames, which leads to point loss under…

Computer Vision and Pattern Recognition · Computer Science 2025-07-29 Han Han , Wei Zhai , Yang Cao , Bin Li , Zheng-jun Zha

Recent studies highlight the potential of textual modalities in conditioning the speech separation model's inference process. However, regularization-based methods remain underexplored despite their advantages of not requiring auxiliary…

Audio and Speech Processing · Electrical Eng. & Systems 2025-08-06 Tsun-An Hsieh , Heeyoul Choi , Minje Kim

We introduce an adaptive scheduling for adaptive sampling as a novel way of machine learning in the construction of part-of-speech taggers. The goal is to speed up the training on large data sets, without significant loss of performance…

Computation and Language · Computer Science 2024-02-06 Manuel Vilares Ferro , Victor M. Darriba Bilbao , Jesús Vilares Ferro

An effective way to increase the noise robustness of automatic speech recognition is to label noisy speech features as either reliable or unreliable (missing) prior to decoding, and to replace the missing ones by clean speech estimates. We…

Sound · Computer Science 2009-01-19 J. F. Gemmeke , B. Cranen

We propose a novel approach for aerial video action recognition. Our method is designed for videos captured using UAVs and can run on edge or mobile devices. We present a learning-based approach that uses customized auto zoom to…

Computer Vision and Pattern Recognition · Computer Science 2023-07-20 Xijun Wang , Ruiqi Xian , Tianrui Guan , Celso M. de Melo , Stephen M. Nogar , Aniket Bera , Dinesh Manocha

Previous real-time MRI (rtMRI)-based speech synthesis models depend heavily on noisy ground-truth speech. Applying loss directly over ground truth mel-spectrograms entangles speech content with MRI noise, resulting in poor intelligibility.…

Sound · Computer Science 2025-01-20 Neil Shah , Ayan Kashyap , Shirish Karande , Vineet Gandhi

In robotics, motion capture systems have been widely used to measure the accuracy of localization algorithms. Moreover, this infrastructure can also be used for other computer vision tasks, such as the evaluation of Visual (-Inertial) SLAM…

Robotics · Computer Science 2024-03-05 Junlin Song , Antoine Richard , Miguel Olivares-Mendez

Electromagnetic articulography (EMA) captures the position and orientation of a number of markers, attached to the articulators, during speech. As such, it performs the same function for speech that conventional motion capture does for…

Human-Computer Interaction · Computer Science 2013-11-01 Ingmar Steiner , Korin Richmond , Slim Ouni

Light detection and ranging (LiDAR) point clouds and building information modeling (BIM) represent two distinct data modalities in the fields of robot perception and construction. These modalities originate from different sources and are…

Robotics · Computer Science 2025-03-11 Zhijian Qiao , Haoming Huang , Chuhao Liu , Zehuan Yu , Shaojie Shen , Fumin Zhang , Huan Yin

We introduce a new cross-modal fusion technique designed for generative error correction in automatic speech recognition (ASR). Our methodology leverages both acoustic information and external linguistic representations to generate accurate…

Humans use multiple communication channels to interact with each other. For instance, body gestures or facial expressions are commonly used to convey an intent. The use of such non-verbal cues has motivated the development of prediction…

Robotics · Computer Science 2024-10-02 Christian Arzate Cruz , Yotam Sechayk , Takeo Igarashi , Randy Gomez
‹ Prev 1 4 5 6 7 8 10 Next ›