English
Related papers

Related papers: Self-Augmented Mixture-of-Experts for QoS Predicti…

200 papers

In cross-border e-commerce, search relevance modeling faces the dual challenge of extreme linguistic diversity and fine-grained semantic nuances. Existing approaches typically rely on scaling up a single monolithic Large Language Model…

Information Retrieval · Computer Science 2026-02-04 Ye Liu , Xu Chen , Wuji Chen , Mang Li

The coexistence of parallel applications in shared computing nodes, each one featuring different Quality of Service (QoS) requirements, carries out new challenges to improve resource occupation while keeping acceptable rates in terms of…

Distributed, Parallel, and Cluster Computing · Computer Science 2024-02-08 Luis Costero , Francisco D. Igual , Katzalin Olcoz , Francisco Tirado

The Mixture of Experts (MoE) for language models has been proven effective in augmenting the capacity of models by dynamically routing each input token to a specific subset of experts for processing. Despite the success, most existing…

Machine Learning · Computer Science 2024-07-26 Hao Zhao , Zihan Qiu , Huijia Wu , Zili Wang , Zhaofeng He , Jie Fu

In this paper, we present the Gaussian process regression as the predictive model for Quality-of-Service (QoS) attributes in Web service systems. The goal is to predict performance of the execution system expressed as QoS attributes given…

Networking and Internet Architecture · Computer Science 2013-05-09 Jakub M. Tomczak , Jerzy Swiatek , Krzysztof Latawiec

Ranking model plays an essential role in e-commerce search and recommendation. An effective ranking model should give a personalized ranking list for each user according to the user preference. Existing algorithms usually extract a user…

Information Retrieval · Computer Science 2023-06-09 Juan Gong , Zhenlin Chen , Chaoyi Ma , Zhuojian Xiao , Haonan Wang , Guoyu Tang , Lin Liu , Sulong Xu , Bo Long , Yunjiang Jiang

With the prevalence of big-data-driven applications, such as face recognition on smartphones and tailored recommendations from Google Ads, we are on the road to a lifestyle with significantly more intelligence than ever before. Various…

Distributed, Parallel, and Cluster Computing · Computer Science 2022-09-26 Ying Mao , Weifeng Yan , Yun Song , Yue Zeng , Ming Chen , Long Cheng , Qingzhi Liu

The Sparsely-Activated Mixture-of-Experts (MoE) has gained increasing popularity for scaling up large language models (LLMs) without exploding computational costs. Despite its success, the current design faces a challenge where all experts…

Machine Learning · Computer Science 2024-09-20 Manxi Sun , Wei Liu , Jian Luan , Pengzhi Gao , Bin Wang

The increasing demand for video streaming services with high Quality of Experience (QoE) has prompted a lot of research on client-side adaptation logic approaches. However, most algorithms use the client's previous download experience and…

Multimedia · Computer Science 2017-08-15 Ran Dubin , Amit Dvir , Ofir Pele , Ofer Hadar , Itay Katz , Ori Mashiach

The Mixture of Experts (MoE) is a widely known neural architecture where an ensemble of specialized sub-models optimizes overall performance with a constant computational cost. However, conventional MoEs pose challenges at scale due to the…

Computation and Language · Computer Science 2023-09-12 Ted Zadouri , Ahmet Üstün , Arash Ahmadian , Beyza Ermiş , Acyr Locatelli , Sara Hooker

For distributed systems to properly react to peaks of requests, their adaptation activities would benefit from the estimation of the amount of requests. This paper proposes a solution to produce a short-term forecast based on data…

Neural and Evolutionary Computing · Computer Science 2014-06-13 Christian Napoli , Giuseppe Pappalardo , Emiliano Tramontana

With the introduction of QUIC, a modern transport-layer network protocol, HTTP/3 leverages its benefits to enhance web content delivery. This paper proposes a mechanism based on the recently standardized Extensible Prioritization Scheme…

Networking and Internet Architecture · Computer Science 2024-04-29 Abhinav Gupta , Radim Bartos

To meet the growing demand for smarter, faster, and more efficient embodied AI solutions, we introduce a novel Mixture-of-Expert (MoE) method that significantly boosts reasoning and learning efficiency for embodied autonomous systems.…

Artificial Intelligence · Computer Science 2025-08-14 Lu Xu , Jiaqian Yu , Xiongfeng Peng , Yiwei Chen , Weiming Li , Jaewook Yoo , Sunghyun Chunag , Dongwook Lee , Daehyun Ji , Chao Zhang

Quality of Service (QoS) is now regarded as a requirement for all networks in managing resources like bandwidth and avoidance of network impairments like packet loss, jitter, and delay. Media transfer or streaming would be virtually…

Networking and Internet Architecture · Computer Science 2021-06-15 Thulani Phakathi , Bukohwo Michael Esiefarienrhe , Francis Lugayizi

The sparse Mixture-of-Experts (MoE) model is powerful for large-scale pre-training and has achieved promising results due to its model capacity. However, with trillions of parameters, MoE is hard to be deployed on cloud or mobile…

Machine Learning · Computer Science 2022-06-03 Tianyu Chen , Shaohan Huang , Yuan Xie , Binxing Jiao , Daxin Jiang , Haoyi Zhou , Jianxin Li , Furu Wei

With the rapid advancement of internet technologies, network services have become critical for delivering diverse and reliable applications to users. However, the exponential growth in the number of available services has resulted in many…

Machine Learning · Computer Science 2025-10-28 Guanchen Du , Jianlong Xu , Mingtong Li , Ruiqi Wang , Qianqing Guo , Caiyi Chen , Qingcao Dai , Yuxiang Zeng

Specifying QoS properties can limit the selection of some good web services that the user will have considered; this is because the algorithm used strictly ensures that there is a match between QoS properties of the consumer with that of…

Other Computer Science · Computer Science 2010-04-13 Agushaka J. O. , Lawal M. M. , Bagiwa , A. M. , Abdullahi B. F

Sparse Mixture of Experts (SMoEs) models scale the capacity of models while maintaining constant computational overhead. Early designs typically relied on a fixed value of $k$, where $k$ represents either the number of experts selected per…

Computation and Language · Computer Science 2025-10-28 Giang Do , Hung Le , Truyen Tran

With rapid advancements in large language models (LLMs), AI-generated content (AIGC) has emerged as a key driver of technological innovation and economic transformation. Personalizing AIGC services to meet individual user demands is…

Computer Science and Game Theory · Computer Science 2025-11-04 Hongjia Wu , Minrui Xu , Zehui Xiong , Lin Gao , Haoyuan Pan , Dusit Niyato , Tse-Tin Chan

Mixture-of-Experts (MoE) models enable scalable performance by activating large parameter sets sparsely, minimizing computational overhead. To mitigate the prohibitive cost of training MoEs from scratch, recent work employs upcycling,…

Machine Learning · Computer Science 2025-11-13 Qi Wang , Hanyang Peng , Yue Yu

Sparsely Mixture of Experts (MoE) has received great interest due to its promising scaling capability with affordable computational overhead. MoE converts dense layers into sparse experts, and utilizes a gated routing network to make…

Computation and Language · Computer Science 2022-07-20 Yuan Xie , Shaohan Huang , Tianyu Chen , Furu Wei