中文
相关论文

相关论文: Incentivizing Permissionless Distributed Learning …

200 篇论文

With the rising emergence of decentralized and opportunistic approaches to machine learning, end devices are increasingly tasked with training deep learning models on-devices using crowd-sourced data that they collect themselves. These…

机器学习 · 计算机科学 2023-04-12 Haoxiang Yu , Hsiao-Yuan Chen , Sangsu Lee , Sriram Vishwanath , Xi Zheng , Christine Julien

DeepSeek-R1 has successfully enhanced Large Language Model (LLM) reasoning capabilities through its rule-based reward system. While it's a ''perfect'' reward system that effectively mitigates reward hacking, such reward functions are often…

机器学习 · 计算机科学 2025-10-27 Chenxing Wei , Jiarui Yu , Ying Tiffany He , Hande Dong , Yao Shu , Fei Yu

Recently, there has been increased interest in globally distributed training, which has the promise to both reduce training costs and democratize participation in building large-scale foundation models. However, existing models trained in a…

分布式、并行与集群计算 · 计算机科学 2026-03-11 Joel Lidin , Amir Sarfi , Erfan Miahi , Quentin Anthony , Shivam Chauhan , Evangelos Pappas , Benjamin Thérien , Eugene Belilovsky , Samuel Dare

Incentives that compensate for the involved costs in the decentralized training of a Federated Learning (FL) model act as a key stimulus for clients' long-term participation. However, it is challenging to convince clients for quality…

机器学习 · 计算机科学 2022-11-04 Shashi Raj Pandey , Lam Duc Nguyen , Petar Popovski

Peer-to-peer (p2p) networks are not independent of their peers, and the network efficiency depends on peers contributing resources. Because shared resources are not free, this contribution must be rewarded. Peers across the network may…

网络与互联网体系结构 · 计算机科学 2022-08-16 Vahid Heidaripour Lakhani , Leander Jehl , Rinke Hendriksen , Vero Estrada-Galiñanes

Blockchain-based federated learning (BCFL) has recently gained tremendous attention because of its advantages such as decentralization and privacy protection of raw data. However, there has been few research focusing on the allocation of…

密码学与安全 · 计算机科学 2022-02-23 Zhilin Wang , Qin Hu , Ruinian Li , Minghui Xu , Zehui Xiong

Peer-to-peer deep learning algorithms are enabling distributed edge devices to collaboratively train deep neural networks without exchanging raw training data or relying on a central server. Peer-to-Peer Learning (P2PL) and other algorithms…

机器学习 · 计算机科学 2023-12-22 Srinivasa Pranav , José M. F. Moura

Modern learning algorithms use gradient descent updates to train inferential models that best explain data. Scaling these approaches to massive data sizes requires proper distributed gradient descent schemes where distributed worker nodes…

We present a unified framework for Large Language Model (LLM) fine-tuning that integrates Imitation Learning and Reinforcement Learning. By analyzing the gradient of a composite objective combining trajectory-level KL divergence with task…

机器学习 · 计算机科学 2025-12-30 Yingru Li , Ziniu Li , Jiacai Liu

This paper investigates whether Bittensor can be considered the Bitcoin of decentralized Artificial Intelligence by directly comparing its tokenomics, decentralization properties, consensus mechanism, and incentive structure against those…

密码学与安全 · 计算机科学 2025-07-08 Elizabeth Lui , Jiahao Sun

In open Federated Learning (FL) environments where no central authority exists, ensuring collaboration fairness relies on decentralized reward settlement, yet the prohibitive cost of permissionless blockchains directly clashes with the…

密码学与安全 · 计算机科学 2026-02-27 Shuang Liang , Yang Hua , Linshan Jiang , Peishen Yan , Tao Song , Bin Yao , Haibing Guan

LLM agents are increasingly relevant to research domains such as vulnerability discovery. Yet, the strongest systems remain closed and cloud-only, making them resource-intensive, difficult to reproduce, and unsuitable for work involving…

密码学与安全 · 计算机科学 2026-03-19 Philipp Normann , Andreas Happe , Jürgen Cito , Daniel Arp

This paper proposes a model that enables permissionless and decentralized networks for complex computations. We explore the integration and optimize load balancing in an open, decentralized computational network. Our model leverages…

计算金融 · 定量金融 2025-01-03 German Rodikov

Distributed learning has gained significant attention due to its advantages in scalability, privacy, and fault tolerance.In this paradigm, multiple agents collaboratively train a global model by exchanging parameters only with their…

机器学习 · 计算机科学 2026-03-31 Ziqin Chen , Yongqiang Wang

This paper presents an empirical analysis of Steemit, a key representative of the emerging incentivized social media platforms over Blockchains, to understand and evaluate the actual level of decentralization and the practical effects of…

社会与信息网络 · 计算机科学 2021-02-03 Chao Li , Balaji Palanisamy

We consider deep deterministic policy gradient (DDPG) in the context of reinforcement learning with sparse rewards. To enhance exploration, we introduce a search procedure, \emph{${\epsilon}{t}$-greedy}, which generates exploratory options…

机器学习 · 计算机科学 2026-02-18 Ehsan Futuhi , Shayan Karimi , Chao Gao , Martin Müller

In this paper, we question the rationale behind propagating large numbers of parameters through a distributed system during federated learning. We start by examining the rank characteristics of the subspace spanned by gradients across…

机器学习 · 计算机科学 2022-02-02 Sheikh Shams Azam , Seyyedali Hosseinalipour , Qiang Qiu , Christopher Brinton

Federated Learning is an emerging distributed collaborative learning paradigm used by many of applications nowadays. The effectiveness of federated learning relies on clients' collective efforts and their willingness to contribute local…

计算机科学与博弈论 · 计算机科学 2022-05-24 Shuyu Kong , You Li , Hai Zhou

We consider distributed optimization under communication constraints for training deep learning models. We propose a new algorithm, whose parameter updates rely on two forces: a regular gradient step, and a corrective direction dictated by…

机器学习 · 计算机科学 2022-04-29 Yunfei Teng , Wenbo Gao , Francois Chalus , Anna Choromanska , Donald Goldfarb , Adrian Weller

There exist a number of reinforcement learning algorithms which learnby climbing the gradient of expected reward. Their long-runconvergence has been proved, even in partially observableenvironments with non-deterministic actions, and…

机器学习 · 计算机科学 2013-01-14 Lex Weaver , Nigel Tao
‹ 上一页 1 2 3 10 下一页 ›