English
Related papers

Related papers: ACCoRD: Actor-Critic Conflict Resolution with Deep…

200 papers

In this work, a novel and model-based artificial neural network (ANN) training method is developed supported by optimal control theory. The method augments training labels in order to robustly guarantee training loss convergence and improve…

Optimization and Control · Mathematics 2023-03-17 Viktor Andersson , Balázs Varga , Vincent Szolnoky , Andreas Syrén , Rebecka Jörnsten , Balázs Kulcsár

Large language models (LLMs) equipped with retrieval--the Retrieval-Augmented Generation (RAG) paradigm--should combine their parametric knowledge with external evidence, yet in practice they often hallucinate, over-trust noisy snippets, or…

Artificial Intelligence · Computer Science 2026-01-13 Hua Ye , Siyuan Chen , Ziqi Zhong , Canran Xiao , Haoliang Zhang , Yuhan Wu , Fei Shen

This paper proposes a set of new error criteria and learning approaches, Adaptive Normalized Risk-Averting Training (ANRAT), to attack the non-convex optimization problem in training deep neural networks (DNNs). Theoretically, we…

Machine Learning · Computer Science 2016-06-10 Zhiguang Wang , Tim Oates , James Lo

Reinforcement Learning (RL) is currently one of the most commonly used techniques for traffic signal control (TSC), which can adaptively adjusted traffic signal phase and duration according to real-time traffic data. However, a fully…

Computer Science and Game Theory · Computer Science 2023-01-03 Yuli. Zhang , Shangbo. Wang , Ruiyuan. Jiang

With the rise of artificial intelligence, neural network simulations of biological neuron models are being explored to reduce the footprint of learning and inference in resource-constrained task scenarios. A mainstream type of such networks…

Hardware Architecture · Computer Science 2024-06-04 Alejandro Linares-Barranco , Luciano Prono , Robert Lengenstein , Giacomo Indiveri , Charlotte Frenkel

We consider an improper reinforcement learning setting where a learner is given $M$ base controllers for an unknown Markov decision process, and wishes to combine them optimally to produce a potentially new controller that can outperform…

Machine Learning · Computer Science 2022-07-20 Mohammadi Zaki , Avinash Mohan , Aditya Gopalan , Shie Mannor

As an alternative to both classical PID-type and modern model-based approaches to solving control problems, active disturbance rejection control (ADRC) has gained significant traction in recent years. With its simple tuning method and…

Systems and Control · Electrical Eng. & Systems 2024-08-01 Gernot Herbst

Neural network (NN) controllers achieve strong empirical performance on nonlinear dynamical systems, yet deploying them in safety-critical settings requires robustness to disturbances and uncertainty. We present a method for jointly…

Systems and Control · Electrical Eng. & Systems 2026-04-02 Neelay Junnarkar , Yasin Sonmez , Murat Arcak

The performance of automatic speech recognition systems under noisy environments still leaves room for improvement. Speech enhancement or feature enhancement techniques for increasing noise robustness of these systems usually add components…

Computation and Language · Computer Science 2016-09-19 Stefan Braun , Daniel Neil , Shih-Chii Liu

Both generative adversarial networks (GAN) in unsupervised learning and actor-critic methods in reinforcement learning (RL) have gained a reputation for being difficult to optimize. Practitioners in both fields have amassed a large number…

Machine Learning · Computer Science 2017-01-19 David Pfau , Oriol Vinyals

Deriving fast and effectively coordinated control actions remains a grand challenge affecting the secure and economic operation of today's large-scale power grid. This paper presents a novel artificial intelligence (AI) based methodology to…

Optimization and Control · Mathematics 2020-12-14 Ruisheng Diao , Di Shi , Bei Zhang , Siqi Wang , Haifeng Li , Chunlei Xu , Tu Lan , Desong Bian , Jiajun Duan

We propose an adversarial deep reinforcement learning (ADRL) algorithm for high-dimensional stochastic control problems. Inspired by the information relaxation duality, ADRL reformulates the control problem as a min-max optimization between…

Optimization and Control · Mathematics 2025-07-03 Nan Chen , Mengzhou Liu , Xiaoyan Wang , Nanyi Zhang

Driver assistance systems as well as autonomous cars have to rely on sensors to perceive their environment. A heterogeneous set of sensors is used to perform this task robustly. Among them, radar sensors are indispensable because of their…

Signal Processing · Electrical Eng. & Systems 2019-06-26 Johanna Rock , Mate Toth , Elmar Messner , Paul Meissner , Franz Pernkopf

A deep neural network (DNN) based power control method is proposed, which aims at solving the non-convex optimization problem of maximizing the sum rate of a multi-user interference channel. Towards this end, we first present PCNet, which…

Signal Processing · Electrical Eng. & Systems 2019-03-12 Fei Liang , Cong Shen , Wei Yu , Feng Wu

Artificial neural networks (NNs) can be implemented using chemical reaction networks (CRNs), where the concentrations of species act as inputs and outputs. In such biochemical computing, noise-robust computing is crucial due to the…

Molecular Networks · Quantitative Biology 2024-10-17 Sunghwa Kang , Jinsu Kim

Deep reinforcement learning (RL) methods have significant potential for dialogue policy optimisation. However, they suffer from a poor performance in the early stages of learning. This is especially problematic for on-line learning with…

Computation and Language · Computer Science 2017-07-06 Pei-Hao Su , Pawel Budzianowski , Stefan Ultes , Milica Gasic , Steve Young

Large language models (LLMs) and retrieval-augmented generation (RAG) techniques have revolutionized traditional information access, enabling AI agent to search and summarize information on behalf of users during dynamic dialogues. Despite…

Information Retrieval · Computer Science 2024-09-04 Yunxiao Shi , Min Xu , Haimin Zhang , Xing Zi , Qiang Wu

Air traffic control is a real-time safety-critical decision making process in highly dynamic and stochastic environments. In today's aviation practice, a human air traffic controller monitors and directs many aircraft flying through its…

Machine Learning · Computer Science 2019-05-07 Marc Brittain , Peng Wei

This work presents a novel reinforcement learning (RL) algorithm based on Y-wise Affine Neural Networks (YANNs). YANNs provide an interpretable neural network which can exactly represent known piecewise affine functions of arbitrary input…

Systems and Control · Electrical Eng. & Systems 2026-05-21 Austin Braniff , Yuhe Tian

Open Radio Access Network (O-RAN) enables network control through multi-vendor xApps operating both within and across layers, subnets, and domains, whose concurrent execution can trigger conflicts that are latent during the development…

Networking and Internet Architecture · Computer Science 2026-04-22 Zeyu Fang , Shu Hong , Huu Trung Thieu , Nakjung Choi , Tian Lan
‹ Prev 1 3 4 5 6 7 10 Next ›