English
Related papers

Related papers: ESC-MVQ: End-to-End Semantic Communication With Mu…

200 papers

Conventional approaches for video captioning leverage a variety of offline-extracted features to generate captions. Despite the availability of various offline-feature-extractors that offer diverse information from different perspectives,…

Computer Vision and Pattern Recognition · Computer Science 2024-10-23 Tian-Zi Niu , Zhen-Duo Chen , Xin Luo , Xin-Shun Xu

Semantic communication with joint semantic-channel coding robustly transmits diverse data modalities but faces challenges in mitigating semantic information loss due to packet drops in packet-based systems. Under current protocols, packets…

Emerging Technologies · Computer Science 2025-08-05 Lei Teng , Senran Fan , Chen Dong , Haotai Liang , Zhicheng Bao , Xiaodong Xu , Rui Meng , Ping Zhang

Semantic communications (SemCom) have emerged as a new paradigm for supporting sixth-generation applications, where semantic features of data are transmitted using artificial intelligence algorithms to attain high communication…

Information Theory · Computer Science 2024-03-15 Jianhao Huang , Kai Yuan , Chuan Huang , Kaibin Huang

Generative models with discrete latent representations have recently demonstrated an impressive ability to learn complex high-dimensional data distributions. However, their performance relies on a long sequence of tokens per instance and a…

Machine Learning · Computer Science 2024-03-26 David D. Nguyen , David Leibowitz , Surya Nepal , Salil S. Kanhere

With the recent advancements in edge artificial intelligence (AI), future sixth-generation (6G) networks need to support new AI tasks such as classification and clustering apart from data recovery. Motivated by the success of deep learning,…

Networking and Internet Architecture · Computer Science 2023-04-06 Zhonghao Lyu , Guangxu Zhu , Jie Xu , Bo Ai , Shuguang Cui

Satellite-terrestrial communications are severely constrained by high path loss, limited spectrum resources, and time-varying channel conditions, rendering conventional bit-level transmission schemes inefficient and fragile, particularly in…

Information Theory · Computer Science 2026-03-04 Jinghong Huang , Mengying Sun , Xiaodong Xu , Jianchi Zhu , Zechuan Fang , Jingxuan Zhang , Ruichen Zhang , Chen Dong , Ping Zhang , Dusit Niyato

A novel learnable dictionary encoding layer is proposed in this paper for end-to-end language identification. It is inline with the conventional GMM i-vector approach both theoretically and practically. We imitate the mechanism of…

Audio and Speech Processing · Electrical Eng. & Systems 2018-04-03 Weicheng Cai , Zexin Cai , Xiang Zhang , Xiaoqi Wang , Ming Li

The proliferation of deep learning-based machine vision applications has given rise to a new type of compression, so called video coding for machine (VCM). VCM differs from traditional video coding in that it is optimized for machine vision…

Computer Vision and Pattern Recognition · Computer Science 2023-08-09 Yeongwoong Kim , Hyewon Jeong , Janghyun Yu , Younhee Kim , Jooyoung Lee , Se Yoon Jeong , Hui Yong Kim

Semantic communication (SemCom) emerges as a transformative paradigm for traffic-intensive visual data transmission, shifting focus from raw data to meaningful content transmission and relieving the increasing pressure on communication…

Image and Video Processing · Electrical Eng. & Systems 2026-02-02 Runze Cheng , Yao Sun , Ahmad Taha , Xuesong Liu , David Flynn , Muhammad Ali Imran

Quantum Error Correction (QEC) decoding faces a fundamental accuracy-efficiency tradeoff. Classical methods like Minimum Weight Perfect Matching (MWPM) exhibit variable performance across noise models and suffer from polynomial complexity,…

Quantum Physics · Physics 2026-04-16 David Zenati , Eliya Nachmani

Image quantization is a crucial technique in image generation, aimed at learning a codebook that encodes an image into a discrete token sequence. Recent advancements have seen researchers exploring learning multi-modal codebook (i.e.,…

Computer Vision and Pattern Recognition · Computer Science 2025-03-12 Guotao Liang , Baoquan Zhang , Zhiyuan Wen , Junteng Zhao , Yunming Ye , Kola Ye , Yao He

Semantic communication is a new paradigm that exploits deep learning models to enable end-to-end communications processes, and recent studies have shown that it can achieve better noise resiliency compared with traditional communication…

Signal Processing · Electrical Eng. & Systems 2023-06-28 Wenyu Zhang , Kaiyuan Bai , Sherali Zeadally , Haijun Zhang , Hua Shao , Hui Ma , Victor C. M. Leung

Video-based dialog task is a challenging multimodal learning task that has received increasing attention over the past few years with state-of-the-art obtaining new performance records. This progress is largely powered by the adaptation of…

Computer Vision and Pattern Recognition · Computer Science 2022-10-27 Huda Alamri , Anthony Bilic , Michael Hu , Apoorva Beedu , Irfan Essa

Quantization is a key method for deploying deep neural networks on edge devices with limited memory and computation resources. Recent improvements in Post-Training Quantization (PTQ) methods were achieved by an additional local optimization…

Computer Vision and Pattern Recognition · Computer Science 2024-09-27 Ofir Gordon , Elad Cohen , Hai Victor Habi , Arnon Netzer

Vector quantization (VQ) is a method for deterministically learning features through discrete codebook representations. Recent works have utilized visual tokenizers to discretize visual regions for self-supervised representation learning.…

Computer Vision and Pattern Recognition · Computer Science 2024-09-11 Chenjing Ding , Chiyu Wang , Boshi Liu , Xi Guo , Weixuan Tang , Wei Wu

Vector quantization is a fundamental tool for compressing high-dimensional embeddings, yet existing multi-codebook methods rely on static codebooks that limit expressiveness under heterogeneous data geometry. While recent dynamic quantizers…

Machine Learning · Computer Science 2026-05-15 Zhengjia Zhong , Shuyan Ke , Zaizhou Lin , Jiaqi Song , Hongyi Lan , Hui Li

With the rapid development of multimodal learning, the image-text matching task, as a bridge connecting vision and language, has become increasingly important. Based on existing research, this study proposes an innovative visual semantic…

Computer Vision and Pattern Recognition · Computer Science 2024-12-30 Wenjing Chen

The rapid progress of artificial intelligence (AI) and computer vision (CV) has facilitated the development of computation-intensive applications like Visual Question Answering (VQA), which integrates visual perception and natural language…

Computer Vision and Pattern Recognition · Computer Science 2024-11-28 Sige Liu , Nan Li , Yansha Deng , Tony Q. S. Quek

Quantization has become essential for the efficient deployment of speech processing systems. Although widely studied, most existing quantization methods were developed for vision and NLP architectures, while the specific challenges of audio…

Sound · Computer Science 2026-03-10 Lucas Rakotoarivony

Variational quantum algorithm (VQA), which is comprised of a classical optimizer and a parameterized quantum circuit, emerges as one of the most promising approaches for harvesting the power of quantum computers in the noisy intermediate…

Quantum Physics · Physics 2021-12-01 Samuel Stein , Yufei Ding , Nathan Wiebe , Bo Peng , Karol Kowalski , Nathan Baker , James Ang , Ang Li
‹ Prev 1 4 5 6 7 8 10 Next ›