Artificial Intelligence · Computer Science
Rethinking Agentic Reinforcement Learning In Large Language Models
Fangming Cui, Ruixiao Zhu, Cheng Fang, Sunan Li +1
2026-05-18
Artificial Intelligence · Computer Science
RetroAgent: From Solving to Evolving via Retrospective Dual Intrinsic Feedback
Xiaoying Zhang, Zichen Liu, Yipeng Zhang, Xia Hu +1
2026-03-31
Computation and Language · Computer Science
Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization
Weiran Yao, Shelby Heinecke, Juan Carlos Niebles, Zhiwei Liu +11
2024-05-07
Computation and Language · Computer Science
Training Agents with Weakly Supervised Feedback from Large Language Models
Dihong Gong, Pu Lu, Zelong Wang, Meng Zhou +1
2024-12-02
Computation and Language · Computer Science
MetaReflection: Learning Instructions for Language Agents using Past Reflections
Priyanshu Gupta, Shashank Kirtania, Ananya Singha, Sumit Gulwani +3
2024-10-11
Artificial Intelligence · Computer Science
Large Language Model as a Policy Teacher for Training Reinforcement Learning Agents
Zihao Zhou, Bin Hu, Chenyang Zhao, Pu Zhang +1
2024-05-28
Machine Learning · Computer Science
SAC-GLAM: Improving Online RL for LLM agents with Soft Actor-Critic and Hindsight Relabeling
Loris Gaven, Clement Romac, Thomas Carta, Sylvain Lamprier +2
2026-01-30
Computation and Language · Computer Science
Improving Retrospective Language Agents via Joint Policy Gradient Optimization
Xueyang Feng, Bo Lan, Quanyu Dai, Lei Wang +4
2025-03-04
Artificial Intelligence · Computer Science
Reflexion: Language Agents with Verbal Reinforcement Learning
Noah Shinn, Federico Cassano, Edward Berman, Ashwin Gopinath +2
2023-10-11
Artificial Intelligence · Computer Science
RAG-Modulo: Solving Sequential Tasks using Experience, Critics, and Language Models
Abhinav Jain, Chris Jermaine, Vaibhav Unhelkar
2024-09-20
Machine Learning · Computer Science
SkillRL: Evolving Agents via Recursive Skill-Augmented Reinforcement Learning
Peng Xia, Jianwen Chen, Hanyang Wang, Jiaqi Liu +9
2026-02-10
Machine Learning · Computer Science
Reinforcement Learning for Aligning Large Language Models Agents with Interactive Environments: Quantifying and Mitigating Prompt Overfitting
Mohamed Salim Aissi, Clement Romac, Thomas Carta, Sylvain Lamprier +4
2025-09-08
Computation and Language · Computer Science
Agentic Conversational Search with Contextualized Reasoning via Reinforcement Learning
Fengran Mo, Yifan Gao, Sha Li, Hansi Zeng +6
2026-04-16
Computation and Language · Computer Science
In Prospect and Retrospect: Reflective Memory Management for Long-term Personalized Dialogue Agents
Zhen Tan, Jun Yan, I-Hung Hsu, Rujun Han +11
2025-07-29
Computation and Language · Computer Science
Retrospective Learning from Interactions
Zizhao Chen, Mustafa Omer Gul, Yiwei Chen, Gloria Geng +2
2025-05-22
Artificial Intelligence · Computer Science
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents
Renxi Wang, Rifo Ahmad Genadi, Bilal El Bouardi, Yongxin Wang +4
2025-07-22
Machine Learning · Computer Science
Interactive Dialogue Agents via Reinforcement Learning on Hindsight Regenerations
Joey Hong, Jessica Lin, Anca Dragan, Sergey Levine
2024-11-11
Computation and Language · Computer Science
AgentLongBench: A Controllable Long Benchmark For Long-Contexts Agents via Environment Rollouts
Shicheng Fang, Yuxin Wang, Xiaoran Liu, Jiahao Lu +5
2026-02-02
Artificial Intelligence · Computer Science
A Comprehensive Survey on Reinforcement Learning-based Agentic Search: Foundations, Roles, Optimizations, Evaluations, and Applications
Minhua Lin, Zongyu Wu, Zhichao Xu, Hui Liu +6
2025-10-29
Machine Learning · Computer Science
ReLook: Vision-Grounded RL with a Multimodal LLM Critic for Agentic Web Coding
Yuhang Li, Chenchen Zhang, Ruilin Lv, Ao Liu +5
2025-10-14
Artificial Intelligence · Computer Science
Empowering Large Language Model Agents through Action Learning
Haiteng Zhao, Chang Ma, Guoyin Wang, Jing Su +4
2024-08-09
Computation and Language · Computer Science
ReSpAct: Harmonizing Reasoning, Speaking, and Acting Towards Building Large Language Model-Based Conversational AI Agents
Vardhan Dongre, Xiaocheng Yang, Emre Can Acikgoz, Suvodip Dey +2
2025-04-22
Information Retrieval · Computer Science
Beyond Static LLM Policies: Imitation-Enhanced Reinforcement Learning for Recommendation
Yi Zhang, Lili Xie, Ruihong Qiu, Jiajun Liu +1
2025-10-16
Cryptography and Security · Computer Science
Large Language Model Integration with Reinforcement Learning to Augment Decision-Making in Autonomous Cyber Operations
Konur Tholl, François Rivest, Mariam El Mezouar, Adrian Taylor +1
2026-02-17