中文
相关论文

相关论文: Contour-based 3d tongue motion visualization using…

200 篇论文

We present an approach to robustly track the geometry of an object that deforms over time from a set of input point clouds captured from a single viewpoint. The deformations we consider are caused by applying forces to known locations on…

计算机视觉与模式识别 · 计算机科学 2015-03-31 Stefanie Wuhrer , Jochen Lang , Motahareh Tekieh , Chang Shu

Visual speech recognition aims to identify the sequence of phonemes from continuous speech. Unlike the traditional approach of using 2D image feature extraction methods to derive features of each video frame separately, this paper proposes…

计算机视觉与模式识别 · 计算机科学 2016-09-08 Toni Heidenreich , Michael W. Spratling

Ultrasound (US) imaging is a critical tool in medical diagnostics, offering real-time visualization of physiological processes. One of its major advantages is its ability to capture temporal dynamics, which is essential for assessing motion…

图像与视频处理 · 电气工程与系统科学 2025-09-03 Yves Stebler , Thomas M. Sutter , Ece Ozkan , Julia E. Vogt

Existing text-driven 3D human motion editing methods have demonstrated significant progress, but are still difficult to precisely control over detailed, part-specific motions due to their global modeling nature. In this paper, we propose…

图形学 · 计算机科学 2026-01-01 Yujie Yang , Zhichao Zhang , Jiazhou Chen , Zichao Wu

To ease the difficulty of acquiring annotation labels in 3D data, a common method is using unsupervised and open-vocabulary semantic segmentation, which leverage 2D CLIP semantic knowledge. In this paper, unlike previous research that…

计算机视觉与模式识别 · 计算机科学 2024-10-14 Fuyang Yu , Runze Tian , Zhen Wang , Xiaochuan Wang , Xiaohui Liang

Speech-driven facial animation involves using a speech signal to generate realistic videos of talking faces. Recent deep learning approaches to facial synthesis rely on extracting low-dimensional representations and concatenating them,…

This paper presents a novel approach for generating 3D talking heads from raw audio inputs. Our method grounds on the idea that speech related movements can be comprehensively and efficiently described by the motion of a few control points…

计算机视觉与模式识别 · 计算机科学 2023-07-27 Federico Nocentini , Claudio Ferrari , Stefano Berretti

Silent speech interfaces (SSI) aim to reconstruct the speech signal from a recording of the articulatory movement, such as an ultrasound video of the tongue. Currently, deep neural networks are the most successful technology for this task.…

声音 · 计算机科学 2021-04-26 László Tóth , Amin Honarmandi Shandiz

Investigating the relationship between internal tissue point motion of the tongue and oropharyngeal muscle deformation measured from tagged MRI and intelligible speech can aid in advancing speech motor control theories and developing novel…

图像与视频处理 · 电气工程与系统科学 2023-02-15 Xiaofeng Liu , Fangxu Xing , Jerry L. Prince , Maureen Stone , Georges El Fakhri , Jonghye Woo

Ultrasound offers a radiation-free, cost-effective solution for real-time visualization of spinal landmarks, paraspinal soft tissues and neurovascular structures, making it valuable for intraoperative guidance during spinal procedures.…

计算机视觉与模式识别 · 计算机科学 2025-11-20 Miruna-Alexandra Gafencu , Yordanka Velikova , Nassir Navab , Mohammad Farid Azampour

Understanding speech production both visually and kinematically can inform second language learning system designs, as well as the creation of speaking characters in video games and animations. In this work, we introduce a data-driven…

图像与视频处理 · 电气工程与系统科学 2024-09-25 Hong Nguyen , Sean Foley , Kevin Huang , Xuan Shi , Tiantian Feng , Shrikanth Narayanan

Speech-driven 3D facial animation is challenging due to the complex geometry of human faces and the limited availability of 3D audio-visual data. Prior works typically focus on learning phoneme-level features of short audio windows with…

计算机视觉与模式识别 · 计算机科学 2022-03-18 Yingruo Fan , Zhaojiang Lin , Jun Saito , Wenping Wang , Taku Komura

We present a novel audio-driven facial animation approach that can generate realistic lip-synchronized 3D facial animations from the input audio. Our approach learns viseme dynamics from speech videos, produces animator-friendly viseme…

图形学 · 计算机科学 2023-01-18 Linchao Bao , Haoxian Zhang , Yue Qian , Tangli Xue , Changhai Chen , Xuefei Zhe , Di Kang

Simulating facial appearance change following bony movement is a critical step in orthognathic surgical planning for patients with jaw deformities. Conventional biomechanics-based methods such as the finite-element method (FEM) are labor…

Temporal volume images with 3D+t (4D) information are often used in medical imaging to statistically analyze temporal dynamics or capture disease progression. Although deep-learning-based generative models for natural images have been…

图像与视频处理 · 电气工程与系统科学 2022-06-28 Boah Kim , Jong Chul Ye

In ultrasound tomography, the speed of sound inside an object is estimated based on acoustic measurements carried out by sensors surrounding the object. An accurate forward model is a prominent factor for high-quality image reconstruction,…

图像与视频处理 · 电气工程与系统科学 2021-11-24 Janne Koponen , Timo Lähivaara , Jari Kaipio , Marko Vauhkonen

Ultrasound imaging is a cost-effective and radiation-free modality for visualizing anatomical structures in real-time, making it ideal for guiding surgical interventions. However, its limited field-of-view, speckle noise, and imaging…

图像与视频处理 · 电气工程与系统科学 2024-04-26 Remi Delaunay , Ruisi Zhang , Filipe C. Pedrosa , Navid Feizi , Dianne Sacco , Rajni Patel , Jayender Jagadeesan

In this paper, we describe a system for generating three-dimensional visual simulations of natural language motion expressions. We use a rich formal model of events and their participants to generate simulations that satisfy the minimal…

计算与语言 · 计算机科学 2016-10-04 Nikhil Krishnaswamy , James Pustejovsky

Recent advances in 3D facial expression reconstruction have demonstrated remarkable performance in capturing macro-expressions, yet the reconstruction of micro-expressions remains unexplored. This novel task is particularly challenging due…

计算机视觉与模式识别 · 计算机科学 2026-03-10 Che Sun , Xinjie Zhang , Rui Gao , Xu Chen , Yuwei Wu , Yunde Jia

In ultrasound (US) imaging, various types of adaptive beamforming techniques have been investigated to improve the resolution and contrast-to-noise ratio of the delay and sum (DAS) beamformers. Unfortunately, the performance of these…

图像与视频处理 · 电气工程与系统科学 2020-02-25 Shujaat Khan , Jaeyoung Huh , Jong Chul Ye