中文
相关论文

相关论文: Learning while Competing -- 3D Modeling & Design

200 篇论文

Reinforcement learning has been widely applied in automated bidding. Traditional approaches model bidding as a Markov Decision Process (MDP). Recently, some studies have explored using generative reinforcement learning methods to address…

机器学习 · 计算机科学 2025-07-23 Kaiyuan Li , Pengyu Wang , Yunshan Peng , Pengjia Yuan , Yanxiang Zeng , Rui Xiang , Yanhua Cheng , Xialong Liu , Peng Jiang

Engineering problems that apply machine learning often involve computationally intensive methods but rely on limited datasets. As engineering data evolves with new designs and constraints, models must incorporate new knowledge over time.…

机器学习 · 计算机科学 2025-04-18 Kaira M. Samuel , Faez Ahmed

Recent advancements in Large Language Models have yielded significant improvements in complex reasoning tasks such as mathematics and programming. However, these models remain heavily dependent on annotated data and exhibit limited…

机器学习 · 计算机科学 2025-09-01 Jia Liu , ChangYi He , YingQiao Lin , MingMin Yang , FeiYang Shen , ShaoGuo Liu

Robotic skills can be learned via imitation learning (IL) using user-provided demonstrations, or via reinforcement learning (RL) using large amountsof autonomously collected experience.Both methods have complementarystrengths and…

This article introduces a novel sample-efficient curriculum learning (CL) approach for training an end-to-end reinforcement learning (RL) policy for robust stabilization of a Quadrotor. The learning objective is to simultaneously stabilize…

We present Catalyst.RL, an open-source PyTorch framework for reproducible and sample efficient reinforcement learning (RL) research. Main features of Catalyst.RL include large-scale asynchronous distributed training, efficient…

机器学习 · 计算机科学 2020-04-09 Sergey Kolesnikov , Valentin Khrulkov

Multi-Agent Deep Reinforcement Learning (MDRL) is a promising research area in which agents learn complex behaviors in cooperative or competitive environments. However, MDRL comes with several challenges that hinder its usability, including…

机器学习 · 计算机科学 2025-01-22 Ahmed Alagha , Jamal Bentahar , Hadi Otrok , Shakti Singh , Rabeb Mizouni

In imitation and reinforcement learning, the cost of human supervision limits the amount of data that robots can be trained on. An aspirational goal is to construct self-improving robots: robots that can learn and improve on their own, from…

机器人学 · 计算机科学 2023-03-03 Archit Sharma , Ahmed M. Ahmed , Rehaan Ahmad , Chelsea Finn

High-quality and representative data is essential for both Imitation Learning (IL)- and Reinforcement Learning (RL)-based motion planning tasks. For real robots, it is challenging to collect enough qualified data either as demonstrations…

机器人学 · 计算机科学 2023-06-13 Sha Luo , Lambert Schomaker

The open world is inherently dynamic, characterized by ever-evolving concepts and distributions. Continual learning (CL) in this dynamic open-world environment presents a significant challenge in effectively generalizing to unseen test-time…

计算机视觉与模式识别 · 计算机科学 2024-10-01 Youngeun Kim , Jun Fang , Qin Zhang , Zhaowei Cai , Yantao Shen , Rahul Duggal , Dripta S. Raychaudhuri , Zhuowen Tu , Yifan Xing , Onkar Dabeer

The combination of deep neural network models and reinforcement learning algorithms can make it possible to learn policies for robotic behaviors that directly read in raw sensory inputs, such as camera images, effectively subsuming both…

机器学习 · 计算机科学 2019-05-17 Avi Singh , Larry Yang , Kristian Hartikainen , Chelsea Finn , Sergey Levine

Autonomous car racing is a challenging task in the robotic control area. Traditional modular methods require accurate mapping, localization and planning, which makes them computationally inefficient and sensitive to environmental changes.…

机器人学 · 计算机科学 2021-07-20 Peide Cai , Hengli Wang , Huaiyang Huang , Yuxuan Liu , Ming Liu

Dynamics problem solving is highly specific to the problem at hand and to develop the general mind framework to become an effective problem solver requires ingenuity and creativity on top of a solid grounding on theoretical and conceptual…

物理教育 · 物理学 2014-09-23 Luciano Fleischfresser

In this paper, we describe our winning approach to solving the Lane Following Challenge at the AI Driving Olympics Competition through imitation learning on a mixed set of simulation and real-world data. AI Driving Olympics is a two-stage…

机器学习 · 计算机科学 2020-07-08 Mikita Sazanovich , Konstantin Chaika , Kirill Krinkin , Aleksei Shpilman

Robots are used in more and more complex environments, and are expected to be able to adapt to changes and unknown situations. The easiest and quickest way to adapt is to change the control system of the robot, but for increasingly complex…

机器人学 · 计算机科学 2019-05-15 Tønnes F. Nygaard , Jørgen Nordmoen , Charles P. Martin , Kyrre Glette

Who gets to decide how generative AI tools enter students' classrooms? We report on a five-week participatory design program in which three 11th-grade Latinx students and three high school teachers in California negotiated how generative AI…

人机交互 · 计算机科学 2026-05-01 Santiago Ojeda-Ramirez , Eva Durall Gazulla , Kylie Peppler

Existing reinforcement learning environment libraries use monolithic environment classes, provide shallow methods for altering agent observation and action spaces, and/or are tied to a specific simulation environment. The Core Reinforcement…

Model-based reinforcement learning (MBRL) is believed to have much higher sample efficiency compared to model-free algorithms by learning a predictive model of the environment. However, the performance of MBRL highly relies on the quality…

机器学习 · 计算机科学 2022-11-16 Xin-Yang Liu , Jian-Xun Wang

The demand of finite raw materials will keep increasing as they fuel modern society. Simultaneously, solutions for stopping carbon emissions in the short term are not available, thus making the net zero target extremely challenging to…

计算机与社会 · 计算机科学 2025-12-17 Federico Zocco , Andrea Corti , Monica Malvezzi

Electric endurance racing is characterized by severe energy constraints and strong aerodynamic interactions. Determining race-winning policies therefore becomes a fundamentally multi-agent, game-theoretic problem. These policies must…

系统与控制 · 电气工程与系统科学 2026-05-13 Wytze de Vries , Erik van den Eshof , Jorn van Kampen , Mauro Salazar