中文
相关论文

相关论文: DOA: A Degeneracy Optimization Agent with Adaptive…

200 篇论文

Object manipulation, which focuses on learning to perform tasks on similar parts across different types of objects, can be divided into an approaching stage and a manipulation stage. However, previous works often ignore this characteristic…

机器人学 · 计算机科学 2025-12-17 Bin Fan , Jian-Jian Jiang , Zhuohao Li , Xiao-Ming Wu , Yi-Xiang He , YiHan Yang , Shengbang Liu , Wei-Shi Zheng

Policy optimization methods with function approximation are widely used in multi-agent reinforcement learning. However, it remains elusive how to design such algorithms with statistical guarantees. Leveraging a multi-agent performance…

机器学习 · 计算机科学 2023-05-09 Yulai Zhao , Zhuoran Yang , Zhaoran Wang , Jason D. Lee

Direct Preference Optimization (DPO) guides large language models (LLMs) to generate recommendations aligned with user historical behavior distributions by minimizing preference alignment loss. However, our systematic empirical research and…

信息检索 · 计算机科学 2026-05-28 Chu Zhao , Enneng Yang , Jianzhe Zhao , Guibing Guo

High-speed, low-latency obstacle avoidance that is insensitive to sensor noise is essential for enabling multiple decentralized robots to function reliably in cluttered and dynamic environments. While other distributed multi-agent collision…

人工智能 · 计算机科学 2017-07-07 Pinxin Long , Wenxi Liu , Jia Pan

Dynamic multi-objective optimization problems (DMOPs) remain a challenge to be settled, because of conflicting objective functions change over time. In recent years, transfer learning has been proven to be a kind of effective approach in…

神经与进化计算 · 计算机科学 2019-10-23 Zhenzhong Wang , Min Jiang , Xing Gao , Liang Feng , Weizhen Hu , Kay Chen Tan

In order to improve the accuracy and resolution for transmit beamspace multiple-input multiple-output (MIMO) radar, a search-free direction-of-arrival (DOA) estimation method based on tensor decomposition and polynomial rooting is proposed.…

信息论 · 计算机科学 2020-10-15 Feng Xu , Xiaopeng Yang , Tian Lan

Diffusion Transformers have become a dominant paradigm in visual generation, yet their low inference efficiency remains a key bottleneck hindering further advancement. Among common training-free techniques, caching offers high acceleration…

计算机视觉与模式识别 · 计算机科学 2026-03-25 Tong Shao , Yusen Fu , Guoying Sun , Jingde Kong , Zhuotao Tian , Jingyong Su

In this paper, we present a novel hybrid approach that combines Reinforcement Learning (RL) with Dynamic Window Approach (DWA) for adaptive 3D local navigation of high-degree-of-freedom robotic systems. Our method leverages sparse point…

机器人学 · 计算机科学 2026-05-14 Chiara Castellani , Enrico Turco , Domenico Prattichizzo

In this paper, we explore using deep reinforcement learning for problems with multiple agents. Most existing methods for deep multi-agent reinforcement learning consider only a small number of agents. When the number of agents increases,…

机器学习 · 计算机科学 2018-05-24 Arbaaz Khan , Clark Zhang , Daniel D. Lee , Vijay Kumar , Alejandro Ribeiro

Complex design problems are common in the scientific and industrial fields. In practice, objective functions or constraints of these problems often do not have explicit formulas, and can be estimated only at a set of sampling points through…

最优化与控制 · 数学 2022-10-12 Lulu Zhang , Zhi-Qin John Xu , Yaoyu Zhang

Direction-of-Arrival (DOA) estimation in sensor arrays faces limitations under demanding conditions, including low signal-to-noise ratio, single-snapshot scenarios, coherent sources, and unknown source counts. Conventional beamforming…

信号处理 · 电气工程与系统科学 2025-10-14 Xuyao Deng , Yong Dou , Kele Xu

It is important for deep reinforcement learning (DRL) algorithms to transfer their learned policies to new environments that have different visual inputs. In this paper, we introduce Prompt based Proximal Policy Optimization ($P^{3}O$), a…

计算机视觉与模式识别 · 计算机科学 2023-03-28 Guoliang You , Xiaomeng Chu , Yifan Duan , Jie Peng , Jianmin Ji , Yu Zhang , Yanyong Zhang

We propose an approach based on machine learning to solve two-stage linear adaptive robust optimization (ARO) problems with binary here-and-now variables and polyhedral uncertainty sets. We encode the optimal here-and-now decisions, the…

机器学习 · 计算机科学 2026-04-21 Dimitris Bertsimas , Cheol Woo Kim

Remote state estimation, where many sensors send their measurements of distributed dynamic plants to a remote estimator over shared wireless resources, is essential for mission-critical applications of Industry 4.0. Most of the existing…

信息论 · 计算机科学 2024-10-28 Gaoyang Pang , Wanchun Liu , Yonghui Li , Branka Vucetic

Parameter-efficient fine-tuning (PEFT) methods have become the standard paradigm for adapting large-scale models. Among these techniques, Weight-Decomposed Low-Rank Adaptation (DoRA) has been shown to improve both the learning capacity and…

机器学习 · 计算机科学 2026-02-09 Nghiem T. Diep , Hien Dang , Tuan Truong , Tan Dinh , Huy Nguyen , Nhat Ho

Efficient data replication in decentralized storage systems must account for diverse policies, especially in multi-organizational, data-intensive environments. This work proposes PSMOA, a novel Policy Support Multi-objective Optimization…

网络与互联网体系结构 · 计算机科学 2025-05-22 Xi Wang , Susmit Shannigrahi

This paper summarizes in depth the state of the art of aerial swarms, covering both classical and new reinforcement-learning-based approaches for their management. Then, it proposes a hybrid AI system, integrating deep reinforcement…

人工智能 · 计算机科学 2025-01-16 Raúl Arranz , David Carramiñana , Gonzalo de Miguel , Juan A. Besada , Ana M. Bernardos

Learning robot control policies from physics simulations is of great interest to the robotics community as it may render the learning process faster, cheaper, and safer by alleviating the need for expensive real-world experiments. However,…

机器人学 · 计算机科学 2021-06-22 Fabio Muratore , Michael Gienger , Jan Peters

Motivated by recent advance of machine learning using Deep Reinforcement Learning this paper proposes a modified architecture that produces more robust agents and speeds up the training process. Our architecture is based on Asynchronous…

机器学习 · 计算机科学 2018-04-18 Ibrahim M. Sobh , Nevin M. Darwish

The integration of artificial intelligence across multiple domains has emphasized the importance of replicating human-like cognitive processes in AI. By incorporating emotional intelligence into AI agents, their emotional stability can be…

人工智能 · 计算机科学 2024-07-31 Hari Prasad , Chinnu Jacob , Imthias Ahamed T. P