中文
相关论文

相关论文: SA-EMO: Structure-Aligned Encoder Mixture of Opera…

200 篇论文

End-to-end design of communication systems using deep autoencoders (AEs) is gaining attention due to its flexibility and excellent performance. Besides single-user transmission, AE-based design is recently explored in multi-user setup,…

信号处理 · 电气工程与系统科学 2023-06-26 Vukan Ninkovic , Dejan Vukobratovic , Adriano Pastore , Carles Anton-Haro

Mixture-of-Experts (MoE) architectures leverage sparse activation to enhance the scalability of large language models (LLMs), making them suitable for deployment in resource-constrained edge networks. However, the sheer number of experts…

信息论 · 计算机科学 2026-03-26 Qian Chen , Xianhao Chen , Kaibin Huang

Advancements in 6G wireless technology have elevated the importance of beamforming, especially for attaining ultra-high data rates via millimeter-wave (mmWave) frequency deployment. Although promising, mmWave bands require substantial beam…

网络与互联网体系结构 · 计算机科学 2024-07-31 Avi Deb Raha , Kitae Kim , Apurba Adhikary , Mrityunjoy Gain , Zhu Han , Choong Seon Hong

Full waveform inversion (FWI) is a process in which seismic numerical simulations are fit to observed data by changing the wave velocity model of the medium under investigation. The problem is non-linear, and therefore optimization…

计算工程、金融与科学 · 计算机科学 2017-06-06 Eran Treister , Eldad Haber

Sparse mixture-of-experts (MoE) layers have been shown to substantially increase model capacity without a proportional increase in computational cost and are widely used in transformer architectures, where they typically replace…

计算机视觉与模式识别 · 计算机科学 2026-04-16 Svetlana Pavlitska , Haixi Fan , Konstantin Ditschuneit , J. Marius Zöllner

We propose a Weak-form Physics-Informed Neural Operator (WINO), a data-free framework that combines the efficiency of neural operators with the geometric flexibility of the $\varphi$-finite element method ($\varphi$-FEM). $\varphi$-FEM is…

数值分析 · 数学 2026-05-26 Bokai Zhu , Qinghui Zhang , Timon Rabczuk

Medical image segmentation remains challenging in low-data regimes, where scarce annotations often yield poor generalization and ambiguous boundaries with missing fine structures. Recent self-supervised pretraining has improved…

计算机视觉与模式识别 · 计算机科学 2026-05-15 Zhiquan Chen , Haitao Wang , Guowei Zou , Hejun Wu

Recently, Convolutional Neural Network (CNN) or Transformer architecture based models have been proposed to overcome the limitations of Recurrent Neural Network (RNN) based models in spatiotemporal prediction. These models prevent the…

计算机视觉与模式识别 · 计算机科学 2026-05-05 Hyeonseok Jin

Mixture of Experts (MoE) models have emerged as the de facto architecture for scaling up language models without significantly increasing the computational cost. Recent MoE models demonstrate a clear trend towards high expert granularity…

机器学习 · 计算机科学 2026-03-30 Wentao Guo , Mayank Mishra , Xinle Cheng , Ion Stoica , Tri Dao

Reliable channel estimation (CE) is fundamental for robust communication in dynamic wireless environments, where models must generalize across varying conditions such as signal-to-noise ratios (SNRs), the number of resource blocks (RBs),…

信号处理 · 电气工程与系统科学 2025-09-22 Tianyu Li , Yan Xin , Jianzhong , Zhang

Recent studies have highlighted the potential of adapting the Segment Anything Model (SAM) for various downstream tasks. However, constructing a more powerful and generalizable encoder to further enhance performance remains an open…

计算机视觉与模式识别 · 计算机科学 2025-08-06 Xinyu Xiong , Zihuang Wu , Lei Zhang , Lei Lu , Ming Li , Guanbin Li

Stack autoencoder (SAE), as a representative deep network, has unique and excellent performance in feature learning, and has received extensive attention from researchers. However, existing deep SAEs focus on original samples without…

机器学习 · 计算机科学 2022-10-28 Chuanyan Zhou , Jie Ma , Fan Li , Yongming Li , Pin Wang , Xiaoheng Zhang

End-to-end autoencoder (AE) learning has the potential of exceeding the performance of human-engineered transceivers and encoding schemes, without a priori knowledge of communication-theoretic principles. In this work, we aim to understand…

Full-waveform inversion (FWI) is known as a seismic data processing method that achieves high-resolution imaging. In the inversion part of the method that brings high resolution in finding a convergence point in the model space, a local…

地球物理 · 物理学 2023-07-11 Jiahang Li , Hitoshi Mikada , Junichi Takekawa

Hybrid beamforming (HBF) and antenna selection are promising techniques for improving the energy efficiency~(EE) of massive multiple-input multiple-output~(mMIMO) systems. However, the transmitter architecture may contain several parameters…

信号处理 · 电气工程与系统科学 2024-07-01 Hamed Hojatian , Zoubeir Mlika , Jérémy Nadal , Jean-François Frigon , François Leduc-Primeau

As revealed by the scaling law of fine-grained MoE, model performance ceases to be improved once the granularity of the intermediate dimension exceeds the optimal threshold, limiting further gains from single-dimension fine-grained design.…

计算机视觉与模式识别 · 计算机科学 2026-03-17 Ning Liao , Xiaoxing Wang , Xiaohan Qin , Junchi Yan

Full waveform inversion (FWI) is an iterative identification process that serves to minimize the misfit of model-based simulated and experimentally measured wave field data, with the goal of identifying a field of parameters for a given…

计算工程、金融与科学 · 计算机科学 2023-12-05 Tim Bürchner , Philipp Kopp , Stefan Kollmannsberger , Ernst Rank

The application of hybrid precoding in millimeter wave (mmWave) multiple-input multiple-output (MIMO) systems has been proved effective for reducing the number of radio frequency (RF) chains. However, the maximum number of independent data…

信息论 · 计算机科学 2017-09-11 Longzhuang He , Jintao Wang , Jian Song

Full-waveform inversion (FWI) plays a vital role in geoscience to explore the subsurface. It utilizes the seismic wave to image the subsurface velocity map. As the machine learning (ML) technique evolves, the data-driven approaches using ML…

机器学习 · 计算机科学 2024-01-09 Junhuan Yang , Hanchen Wang , Yi Sheng , Youzuo Lin , Lei Yang

The combination of Mixture-of-Experts (MoE) and Low-Rank Adaptation (LoRA) has shown significant potential for enhancing the multi-task learning capabilities of Large Language Models. However, existing methods face two primary challenges:…

计算与语言 · 计算机科学 2026-04-22 Boyan Shi , Wei Chen , Shuyuan Zhao , Junfeng Shen , Shengnan Guo , Shaojiang Wang , Huaiyu Wan