中文
相关论文

相关论文: Crowdfunding Dynamics Tracking: A Reinforcement Le…

200 篇论文

In this paper we consider the problem of how a reinforcement learning agent tasked with solving a set of related Markov decision processes can use knowledge acquired early in its lifetime to improve its ability to more rapidly solve novel,…

人工智能 · 计算机科学 2019-02-26 Francisco M. Garcia , Bruno C. da Silva , Philip S. Thomas

Trajectory following is one of the complicated control problems when its dynamics are nonlinear, stochastic and include a large number of parameters. The problem has significant difficulties including a large number of trials required for…

机器人学 · 计算机科学 2019-02-14 Ali Lenjani

Diffusion models have become popular for policy learning in robotics due to their ability to capture high-dimensional and multimodal distributions. However, diffusion policies are stochastic and typically trained offline, limiting their…

机器人学 · 计算机科学 2025-05-28 Ralf Römer , Alexander von Rohr , Angela P. Schoellig

We study a problem of allocating divisible jobs, arriving online, to workers in a crowdsourcing setting which involves learning two parameters of strategically behaving workers. Each job is split into a certain number of tasks that are then…

人工智能 · 计算机科学 2016-02-15 Satyanath Bhat , Divya Padmanabhan , Shweta Jain , Y Narahari

Recent successes combine reinforcement learning algorithms and deep neural networks, despite reinforcement learning not being widely applied to robotics and real world scenarios. This can be attributed to the fact that current…

机器学习 · 计算机科学 2020-09-01 Vinicius G. Goecks

Recently online advertisers utilize Recommender systems (RSs) for display advertising to improve users' engagement. The contextual bandit model is a widely used RS to exploit and explore users' engagement and maximize the long-term rewards…

信息检索 · 计算机科学 2022-10-27 Shion Ishikawa , Young-joo Chung , Yu Hirate

Tracking controllers enable robotic systems to accurately follow planned reference trajectories. In particular, reinforcement learning (RL) has shown promise in the synthesis of controllers for systems with complex dynamics and modest…

机器人学 · 计算机科学 2025-05-05 Jake Welde , Nishanth Rao , Pratik Kunapuli , Dinesh Jayaraman , Vijay Kumar

Estimating uncertainty for AI agents in real-world multi-turn tool-using interaction with humans is difficult because failures are often triggered by sparse critical episodes (e.g., looping, incoherent tool use, or user-agent…

人工智能 · 计算机科学 2026-02-13 Sina Tayebati , Divake Kumar , Nastaran Darabi , Davide Ettori , Ranganath Krishnan , Amit Ranjan Trivedi

Current state-of-the-art crowd navigation approaches are mainly deep reinforcement learning (DRL)-based. However, DRL-based methods suffer from the issues of generalization and scalability. To overcome these challenges, we propose a method…

机器人学 · 计算机科学 2023-09-26 Hafiq Anas , Ong Wee Hong , Owais Ahmed Malik

Despite achieving state-of-the-art generation quality, diffusion models are hindered by the substantial computational burden of their iterative sampling process. While feature caching techniques achieve effective acceleration at higher step…

计算机视觉与模式识别 · 计算机科学 2026-03-04 Benlei Cui , Shaoxuan He , Bukun Huang , Zhizeng Ye , Yunyun Sun , Longtao Huang , Hui Xue , Yang Yang , Jingqun Tang , Zhou Zhao , Haiwen Hong

Crowdfunding has been used as one of the effective ways for entrepreneurs to raise funding especially in creative industries. Individuals as well as organizations are paying more attentions to the emergence of new crowdfunding platforms. In…

计算机与社会 · 计算机科学 2015-03-03 Yang Song , Robert van Boeschoten

Imitation learning is a promising approach for enabling generalist capabilities in humanoid robots, but its scaling is fundamentally constrained by the scarcity of high-quality expert demonstrations. This limitation can be mitigated by…

机器人学 · 计算机科学 2025-08-21 Quentin Rouxel , Clemente Donoso , Fei Chen , Serena Ivaldi , Jean-Baptiste Mouret

Crowdfunding, which is the act of raising funds from a large number of people's contributions, is among the most popular research topics in economic theory. Due to the fact that crowdfunding platforms (CFPs) have facilitated the process of…

数理金融 · 定量金融 2022-07-18 Fatemeh Nosrat

Autonomous racing is becoming popular for academic and industry researchers as a test for general autonomous driving by pushing perception, planning, and control algorithms to their limits. While traditional control methods such as MPC are…

机器人学 · 计算机科学 2023-07-07 Edoardo Ghignone , Nicolas Baumann , Mike Boss , Michele Magno

We develop a methodology for index tracking and risk exposure control using financial derivatives. Under a continuous-time diffusion framework for price evolution, we present a pathwise approach to construct dynamic portfolios of…

数理金融 · 定量金融 2017-05-31 Tim Leung , Brian Ward

Online narratives spread unevenly across platforms, with content emerging on one site often appearing on others, hours, days or weeks later. Existing cross-platform information diffusion models often treat platforms as isolated systems,…

社会与信息网络 · 计算机科学 2025-10-22 Patrick Gerard , Luca Luceri , Leonardo Blas , Emilio Ferrara

Mining the underlying patterns in gigantic and complex data is of great importance to data analysts. In this paper, we propose a motion pattern approach to mine frequent behaviors in trajectory data. Motion patterns, defined by a set of…

计算机视觉与模式识别 · 计算机科学 2015-01-06 Mahdi M. Kalayeh , Stephen Mussmann , Alla Petrakova , Niels da Vitoria Lobo , Mubarak Shah

The inference of outcomes in dynamic processes from structural features of systems is a crucial endeavor in network science. Recent research has suggested a machine learning-based approach for the interpretation of dynamic patterns emerging…

物理与社会 · 物理学 2022-11-15 Aruane M. Pineda , Caroline L. Alves , Colm Connaughton , Francisco A. Rodrigues

Cyclic pursuit frameworks, which are built upon pursuit interactions between neighboring agents in a cycle graph, provide an efficient way to create useful global behaviors in a collective of autonomous robots. Previous work had considered…

系统与控制 · 计算机科学 2017-11-16 Kevin S. Galloway , Biswadip Dey

Flow-based policies have recently emerged as a powerful tool in offline and offline-to-online reinforcement learning, capable of modeling the complex, multimodal behaviors found in pre-collected datasets. However, the full potential of…

机器学习 · 计算机科学 2025-09-30 Deshu Chen , Yuchen Liu , Zhijian Zhou , Chao Qu , Yuan Qi