中文
相关论文

相关论文: Frame Rate Up-Conversion Using Key Point Agnostic …

200 篇论文

As supercomputers advance towards exascale capabilities, computational intensity increases significantly, and the volume of data requiring storage and transmission experiences exponential growth. Adaptive Mesh Refinement (AMR) has emerged…

分布式、并行与集群计算 · 计算机科学 2023-07-20 Daoce Wang , Jesus Pulido , Pascal Grosset , Jiannan Tian , Sian Jin , Houjun Tang , Jean Sexton , Sheng Di , Zarija Lukić , Kai Zhao , Bo Fang , Franck Cappello , James Ahrens , Dingwen Tao

With the increasing demand of capturing our environment in three-dimensions for AR/ VR applications and autonomous driving among others, the importance of high-resolution point clouds rises. As the capturing process is a complex task, point…

计算机视觉与模式识别 · 计算机科学 2023-01-30 Viktoria Heimann , Andreas Spruck , André Kaup

This paper presents improvements and novel additions to our recent work on end-to-end optimized hierarchical bi-directional video compression to further advance the state-of-the-art in learned video compression. As an improvement, we…

图像与视频处理 · 电气工程与系统科学 2022-06-29 Eren Cetin , M. Akin Yilmaz , A. Murat Tekalp

High-frequency components are crucial for maintaining video clarity and realism, but they also significantly impact coding bitrate, resulting in increased bandwidth and storage costs. This paper presents an end-to-end learning-based…

计算机视觉与模式识别 · 计算机科学 2025-08-13 Yingxue Pang , Shijie Zhao , Junlin Li , Li Zhang

Frame aggregation is a mechanism by which multiple frames are combined into a single transmission unit over the air. Frames aggregated at the AMSDU level use a common CRC check to enforce integrity. For longer aggregated AMSDU frames, the…

网络与互联网体系结构 · 计算机科学 2017-07-11 Gautam Bhanage

We develop an effective point cloud rendering pipeline for novel view synthesis, which enables high fidelity local detail reconstruction, real-time rendering and user-friendly editing. In the heart of our pipeline is an adaptive frequency…

计算机视觉与模式识别 · 计算机科学 2023-03-21 Yi Zhang , Xiaoyang Huang , Bingbing Ni , Teng Li , Wenjun Zhang

Time-resolved femtosecond stimulated Raman spectroscopy (FSRS) is a powerful tool for investigating ultrafast structural and vibrational dynamics in light absorbing systems. However, the technique generally requires exposing a sample to…

仪器与探测器 · 物理学 2018-07-10 Matthew N. Ashner , William A. Tisdale

Conventional speech enhancement technique such as beamforming has known benefits for far-field speech recognition. Our own work in frequency-domain multi-channel acoustic modeling has shown additional improvements by training a spatial…

声音 · 计算机科学 2020-02-10 Taejin Park , Kenichi Kumatani , Minhua Wu , Shiva Sundaram

State-of-the-art end-to-end automatic speech recognition (ASR) extracts acoustic features from input speech signal every 10 ms which corresponds to a frame rate of 100 frames/second. In this report, we investigate the use of high-frame-rate…

音频与语音处理 · 电气工程与系统科学 2019-07-15 Cong-Thanh Do

State-of-the-art video deblurring methods use deep network architectures to recover sharpened video frames. Blurring especially degrades high-frequency (HF) information, yet this aspect is often overlooked by recent models that focus more…

计算机视觉与模式识别 · 计算机科学 2024-12-03 Bo Ji , Angela Yao

In this paper, we introduce a novel approach that combines multiresolution (MR) techniques with the flux reconstruction (FR) method to accurately and effciently simulate compressible flows. We achieve further enhancements in effciency…

流体动力学 · 物理学 2023-06-21 Yixuan Lian , Jinsheng Cai , Shucheng Pan

We propose an adaptive form of frameless rendering with the potential to dramatically increase rendering speed over conventional interactive rendering approaches. Without the rigid sampling patterns of framed renderers, sampling and…

图形学 · 计算机科学 2025-10-21 Abhinav Dayal , Cliff Woolley , Benjamin Watson , David Luebke

Most neural speech codecs achieve bitrate adjustment through intra-frame mechanisms, such as codebook dropout, at a Constant Frame Rate (CFR). However, speech segments inherently have time-varying information density (e.g., silent intervals…

音频与语音处理 · 电气工程与系统科学 2025-09-09 Hanglei Zhang , Yiwei Guo , Zhihan Li , Xiang Hao , Xie Chen , Kai Yu

Caching and reusing intermediate features across consecutive frames is a common technique to reduce redundant computation and transmission for edge-cloud video analytics in mobile edge computation. Existing methods manage the cache in a…

网络与互联网体系结构 · 计算机科学 2026-05-08 Xiuxian Guan , Zongyuan Zhang , Zheng Lin , Zekai Sun , Tianyang Duan , Zihan Fang , Rui Wang , Heming Cui , Wei Ni , Jun Luo , Yuanwei Liu

High-resolution fMRI provides a window into the brain's mesoscale organization. Yet, higher spatial resolution increases scan times, to compensate for the low signal and contrast-to-noise ratio. This work introduces a deep learning-based 3D…

图像与视频处理 · 电气工程与系统科学 2024-03-20 Hongwei Bran Li , Matthew S. Rosen , Shahin Nasr , Juan Eugenio Iglesias

Neural video compression (NVC) has demonstrated superior compression efficiency, yet effective rate control remains a significant challenge due to complex temporal dependencies. Existing rate control schemes typically leverage frame content…

图像与视频处理 · 电气工程与系统科学 2026-02-03 Wuyang Cong , Junqi Shi , Lizhong Wang , Weijing Shi , Ming Lu , Hao Chen , Zhan Ma

There exist many scenarios where pixel information is available only on a non-regular subset of pixel positions. For further processing, however, it is required to reconstruct such images on a regular grid. Besides many other algorithms,…

图像与视频处理 · 电气工程与系统科学 2022-04-08 Markus Jonscher , Jürgen Seiler , André Kaup

Face recognition has made great progress with the development of deep learning. However, video face recognition (VFR) is still an ongoing task due to various illumination, low-resolution, pose variations and motion blur. Most existing…

计算机视觉与模式识别 · 计算机科学 2017-08-29 Yibo Hu , Xiang Wu , Ran He

Recent approaches for fast semantic video segmentation have reduced redundancy by warping feature maps across adjacent frames, greatly speeding up the inference phase. However, the accuracy drops seriously owing to the errors incurred by…

计算机视觉与模式识别 · 计算机科学 2023-07-12 Songyuan Li , Junyi Feng , Xi Li

High spatial frequency information, including fine details like textures, significantly contributes to the accuracy of semantic segmentation. However, according to the Nyquist-Shannon Sampling Theorem, high-frequency components are…

计算机视觉与模式识别 · 计算机科学 2025-07-24 Linwei Chen , Ying Fu , Lin Gu , Dezhi Zheng , Jifeng Dai