中文
相关论文

相关论文: A Human Ear Reconstruction Autoencoder

200 篇论文

Autoencoder (AE) is a neural network (NN) architecture that is trained to reconstruct an input at its output. By measuring the reconstruction errors of new input samples, AE can detect anomalous samples deviated from the trained data…

机器学习 · 计算机科学 2023-02-16 Jinho Choi , Jihong Park , Abhinav Japesh , Adarsh

Learning a good 3D human pose representation is important for human pose related tasks, e.g. human 3D pose estimation and action recognition. Within all these problems, preserving the intrinsic pose information and adapting to view…

计算机视觉与模式识别 · 计算机科学 2021-11-03 Qiang Nie , Ziwei Liu , Yunhui Liu

Self-supervised models allow (pre-)training on unlabeled data and therefore have the potential to overcome the need for large annotated cohorts. One leading self-supervised model is the masked autoencoder (MAE) which was developed on…

图像与视频处理 · 电气工程与系统科学 2023-03-13 Daniel M. Lang , Eli Schwartz , Cosmin I. Bercea , Raja Giryes , Julia A. Schnabel

We present a novel face reconstruction method capable of reconstructing detailed face geometry, spatially varying face reflectance from a single monocular image. We build our work upon the recent advances of DNN-based auto-encoders with…

计算机视觉与模式识别 · 计算机科学 2022-04-06 Abdallah Dib , Junghyun Ahn , Cedric Thebault , Philippe-Henri Gosselin , Louis Chevallier

Sound processing in the human auditory system is complex and highly non-linear, whereas hearing aids (HAs) still rely on simplified descriptions of auditory processing or hearing loss to restore hearing. Even though standard HA…

音频与语音处理 · 电气工程与系统科学 2023-06-21 Fotios Drakopoulos , Sarah Verhulst

Reconstructing the 3D pose of a person in metric scale from a single view image is a geometrically ill-posed problem. For example, we can not measure the exact distance of a person to the camera from a single view image without additional…

计算机视觉与模式识别 · 计算机科学 2021-12-06 Zhijian Yang , Xiaoran Fan , Volkan Isler , Hyun Soo Park

Hearable devices, equipped with one or more microphones, are commonly used for speech communication. Here, we consider the scenario where a hearable is used to capture the user's own voice in a noisy environment. In this scenario, own voice…

音频与语音处理 · 电气工程与系统科学 2025-08-20 Mattes Ohlenbusch , Christian Rollwage , Simon Doclo

Recent progress in human shape learning, shows that neural implicit models are effective in generating 3D human surfaces from limited number of views, and even from a single RGB image. However, existing monocular approaches still struggle…

计算机视觉与模式识别 · 计算机科学 2024-03-20 Marco Pesavento , Yuanlu Xu , Nikolaos Sarafianos , Robert Maier , Ziyan Wang , Chun-Han Yao , Marco Volino , Edmond Boyer , Adrian Hilton , Tony Tung

The attention-based encoder-decoder (AED) speech recognition model has been widely successful in recent years. However, the joint optimization of acoustic model and language model in end-to-end manner has created challenges for text…

音频与语音处理 · 电气工程与系统科学 2024-09-17 Shaoshi Ling , Guoli Ye , Rui Zhao , Yifan Gong

Automated segmentation of intracranial arteries on magnetic resonance angiography (MRA) allows for quantification of cerebrovascular features, which provides tools for understanding aging and pathophysiological adaptations of the…

图像与视频处理 · 电气工程与系统科学 2017-12-21 Li Chen , Yanjun Xie , Jie Sun , Niranjan Balu , Mahmud Mossa-Basha , Kristi Pimentel , Thomas S. Hatsukami , Jenq-Neng Hwang , Chun Yuan

Self-supervised learning has shown great success in Speech Recognition. However, it has been observed that finetuning all layers of the learned model leads to lower performance compared to resetting top layers. This phenomenon is attributed…

计算与语言 · 计算机科学 2024-05-15 Valentin Vielzeuf

Compared to humans, machine learning models generally require significantly more training examples and fail to extrapolate from experience to solve previously unseen challenges. To help close this performance gap, we augment single-task…

机器学习 · 计算机科学 2018-07-27 Tailin Wu , John Peurifoy , Isaac L. Chuang , Max Tegmark

The reconstruction of 3D objects from brain signals has gained significant attention in brain-computer interface (BCI) research. Current research predominantly utilizes functional magnetic resonance imaging (fMRI) for 3D reconstruction…

图形学 · 计算机科学 2025-05-06 Xia Deng , Shen Chen , Jiale Zhou , Lei Li

This study addresses the problem of 3D human mesh reconstruction from multi-view images. Recently, approaches that directly estimate the skinned multi-person linear model (SMPL)-based human mesh vertices based on volumetric heatmap…

计算机视觉与模式识别 · 计算机科学 2023-06-30 Sungho Chun , Sungbum Park , Ju Yong Chang

This paper proposes a novel model fitting algorithm for 3D facial expression reconstruction from a single image. Face expression reconstruction from a single image is a challenging task in computer vision. Most state-of-the-art methods fit…

计算机视觉与模式识别 · 计算机科学 2018-08-20 Fanzi Wu , Songnan Li , Tianhao Zhao , King Ngi Ngan , Lv Sheng

Merging multi-exposure images is a common approach for obtaining high dynamic range (HDR) images, with the primary challenge being the avoidance of ghosting artifacts in dynamic scenes. Recent methods have proposed using deep neural…

计算机视觉与模式识别 · 计算机科学 2024-02-29 Zhilu Zhang , Haoyu Wang , Shuai Liu , Xiaotao Wang , Lei Lei , Wangmeng Zuo

Multi-view human mesh recovery (HMR) is broadly deployed in diverse domains where high accuracy and strong generalization are essential. Existing approaches can be broadly grouped into geometry-based and learning-based methods. However,…

计算机视觉与模式识别 · 计算机科学 2026-04-02 Haoyu Xie , Shengkai Xu , Cheng Guo , Muhammad Usama Saleem , Wenhan Wu , Chen Chen , Ahmed Helmy , Pu Wang , Hongfei Xue

We present a self-supervised human mesh recovery framework to infer human pose and shape from monocular images in the absence of any paired supervision. Recent advances have shifted the interest towards directly regressing parameters of a…

计算机视觉与模式识别 · 计算机科学 2020-08-05 Jogendra Nath Kundu , Mugalodi Rakesh , Varun Jampani , Rahul Mysore Venkatesh , R. Venkatesh Babu

In this work we present a novel method for reconstructing 3D surfaces using a multi-beam imaging sonar. We integrate the intensities measured by the sonar from different viewpoints for fixed cell positions in a 3D grid. For each cell we…

机器人学 · 计算机科学 2022-06-08 Sascha Arnold , Bilal Wehbe

Unsupervised learning is becoming more and more important recently. As one of its key components, the autoencoder (AE) aims to learn a latent feature representation of data which is more robust and discriminative. However, most AE based…

机器学习 · 计算机科学 2019-04-02 Jingcai Guo , Song Guo
‹ 上一页 1 8 9 10 下一页 ›