English
Related papers

Related papers: SwiftCache: Model-Based Learning for Dynamic Conte…

200 papers

Next-generation mobile networks are expected to facilitate fast AI model downloading to end users. By caching models on edge servers, mobile networks can deliver models to end users with low latency, resulting in a paradigm called edge…

Networking and Internet Architecture · Computer Science 2024-05-21 Guanqiao Qu , Zheng Lin , Fangming Liu , Xianhao Chen , Kaibin Huang

Online social networks (OSNs) are emerging as the most popular mainstream platform for content cascade diffusion. In order to provide satisfactory quality of experience (QoE) for users in OSNs, much research dedicates to proactive content…

Social and Information Networks · Computer Science 2020-03-26 Qiong Wu , Muhong Wu , Xu Chen , Zhi Zhou , Kaiwen He , Liang Chen

Small-cell networks have been proposed to meet the demand of ever growing mobile data traffic. One of the prominent challenges faced by small-cell networks is the lack of sufficient backhaul capacity to connect small-cell base stations…

Networking and Internet Architecture · Computer Science 2014-07-07 Yang Guan , Yao Xiao , Hao Feng , Chien-Chung Shen , Leonard J. Cimini

Low latency communication is one of the fundamental requirements for 5G wireless networks and beyond. In this paper, a novel approach for joint caching, user scheduling and resource allocation is proposed for minimizing the queuing latency…

Networking and Internet Architecture · Computer Science 2023-09-22 Tamoor-ul-Hassan Syed , Samarakoon Sumudu , Bennis Mehdi , Matti Latva-aho

Stream media content caching is a key enabling technology to promote the value chain of future urban vehicular networks. Nevertheless, the high mobility of vehicles, intermittency of information transmissions, high dynamics of user…

Information Theory · Computer Science 2023-05-15 Biqian Feng , Chenyuan Feng , Daquan Feng , Yongpeng Wu , Xiang-Gen Xia

While Deep Learning (DL) technologies are a promising tool to solve networking problems that map to classification tasks, their computational complexity is still too high with respect to real-time traffic measurements requirements. To…

Networking and Internet Architecture · Computer Science 2022-10-04 Alessandro Finamore , James Roberts , Massimo Gallo , Dario Rossi

Mobile Edge Caching is a promising technique to enhance the content delivery quality and reduce the backhaul link congestion, by storing popular content at the network edge or mobile devices (e.g. base stations and smartphones) that are…

Computer Science and Game Theory · Computer Science 2021-09-15 Mingyu Li , Changkun Jiang , Lin Gao , Tong Wang , Yufei Jiang

As mobile services are shifting from "connection-centric" communications to "content-centric" communications, content-centric wireless networking emerges as a promising paradigm to evolve the current network architecture. Caching popular…

Information Theory · Computer Science 2016-11-15 Rui Wang , Xi Peng , Jun Zhang , K. B. Letaief

The deployment of small cells is expected to gain huge momentum in the near future, as a solution for managing the skyrocketing mobile data demand growth. Local caching of popular files at the small cell base stations has been recently…

Networking and Internet Architecture · Computer Science 2014-03-03 Konstantinos Poularakis , George Iosifidis , Vasilis Sourlas , Leandros Tassiulas

This paper focuses on similarity caching systems, in which a user request for an {object~$o$} that is not in the cache can be (partially) satisfied by a similar stored {object~$o'$}, at the cost of a loss of user utility. Similarity caching…

Networking and Internet Architecture · Computer Science 2021-05-28 Michele Garetto , Emilio Leonardi , Giovanni Neglia

Edge caching will play a critical role in facilitating the emerging content-rich applications. However, it faces many new challenges, in particular, the highly dynamic content popularity and the heterogeneous caching configurations. In this…

Networking and Internet Architecture · Computer Science 2021-01-18 Tongyu Zong , Chen Li , Yuanyuan Lei , Guangyu Li , Houwei Cao , Yong Liu

Predictive Autoscaling is used to forecast the workloads of servers and prepare the resources in advance to ensure service level objectives (SLOs) in dynamic cloud environments. However, in practice, its prediction task often suffers from…

Machine Learning · Computer Science 2023-09-06 Hongyan Hao , Zhixuan Chu , Shiyi Zhu , Gangwei Jiang , Yan Wang , Caigao Jiang , James Zhang , Wei Jiang , Siqiao Xue , Jun Zhou

Large language models (LLMs) have excelled in various applications, yet serving them at scale is challenging due to their substantial resource demands and high latency. Our real-world studies reveal that over 70% of user requests to LLMs…

Machine Learning · Computer Science 2025-09-05 Yifan Yu , Yu Gan , Nikhil Sarda , Lillian Tsai , Jiaming Shen , Yanqi Zhou , Arvind Krishnamurthy , Fan Lai , Henry M. Levy , David Culler

With data durability, high access speed, low power efficiency and byte addressability, NVMe and SSD, which are acknowledged representatives of emerging storage technologies, have been applied broadly in many areas. However, one key issue…

Performance · Computer Science 2020-11-17 Nan Wu , Pengcheng Li

This letter proposes a novel three-tier content caching architecture for Vehicular Fog Caching (VFC)-assisted platoon, where the VFC is formed by the vehicles driving near the platoon. The system strategically coordinates storage across…

Networking and Internet Architecture · Computer Science 2026-02-05 Bowen Tan , Qiong Wu , Pingyi Fan , Kezhi Wang , Nan Cheng , Wen Chen

Pushing popular content to cheap "helper" nodes (e.g., small cells) during off-peak hours has recently been proposed to cope with the increase in mobile data traffic. User requests can be served locally from these helper nodes, if the…

Networking and Internet Architecture · Computer Science 2017-02-17 Pavlos Sermpezis , Thrasyvoulos Spyropoulos , Luigi Vigneri , Theodoros Giannakas

Replicating or caching popular content in memories distributed across the network is a technique to reduce peak network loads. Conventionally, the main performance gain of this caching was thought to result from making part of the requested…

Information Theory · Computer Science 2015-09-08 Mohammad Ali Maddah-Ali , Urs Niesen

Federated Learning (FL) trains a shared model using data and computation power on distributed agents coordinated by a central server. Decentralized FL (DFL) utilizes local model exchange and aggregation between agents to reduce the…

Machine Learning · Computer Science 2025-02-05 Xiaoyu Wang , Guojun Xiong , Houwei Cao , Jian Li , Yong Liu

Deep Neural Networks (DNNs) have become an essential component in many application domains including web-based services. A variety of these services require high throughput and (close to) real-time features, for instance, to respond or…

Machine Learning · Computer Science 2022-09-20 Mohammadamin Abedi , Yanni Iouannou , Pooyan Jamshidi , Hadi Hemmati

In this work, we propose a content caching and delivery strategy to maximize throughput capacity in cache-enabled wireless networks. To this end, efficient betweenness (EB), which indicates the ratio of content delivery paths passing…

Information Theory · Computer Science 2020-05-11 Chenxi Zhao , Junyu Liu , Min Sheng , Yanpeng Dai