Trading and Market Microstructure · Quantitative Finance
Behavioral Consistency Validation for LLM Agents: An Analysis of Trading-Style Switching through Stock-Market Simulation
Zeping Li, Guancheng Wan, Keyang Chen, Yu Chen +5
2026-03-25
Software Engineering · Computer Science
Same Signal, Different Semantics: A Cross-Framework Behavioral Analysis of Software Engineering Agents
Wei Ma, Zhi Chen, Jingxu Gu, Tianling Li +2
2026-05-19
Information Retrieval · Computer Science
Same Outcomes, Different Journeys: A Trace-Level Framework for Comparing Human and GUI-Agent Behavior in Production Search Systems
Maria Movin, Claudia Hauff, Aron Henriksson, Panagiotis Papapetrou
2026-04-10
Computation and Language · Computer Science
GroundAct: Can LLM Agents Ground Actions in Environmental States?
Zixuan Wang, Dingming Li, Hongxing Li, Yanrui Miao +7
2026-05-29
Software Engineering · Computer Science
Evaluating Plan Compliance in Autonomous Programming Agents
Shuyang Liu, Saman Dehghan, Jatin Ganhotra, Martin Hirzel +1
2026-04-29
Artificial Intelligence · Computer Science
Are LLM Agents Behaviorally Coherent? Latent Profiles for Social Simulation
James Mooney, Josef Woldense, Zheng Robert Jia, Shirley Anugrah Hayati +3
2026-05-11
Human-Computer Interaction · Computer Science
InconLens: Interactive Visual Diagnosis of Behavioral Inconsistencies in LLM-based Agentic Systems
Shuo Yan, Xiaolin Wen, Shaolun Ruan, Yanjie Zhang +4
2026-03-31
Artificial Intelligence · Computer Science
AgentMisalignment: Measuring the Propensity for Misaligned Behaviour in LLM-Based Agents
Akshat Naik, Patrick Quinn, Guillermo Bosch, Emma Gouné +3
2025-10-02
Artificial Intelligence · Computer Science
Echoing: Identity Failures when LLM Agents Talk to Each Other
Sarath Shekkizhar, Romain Cosentino, Adam Earle, Silvio Savarese
2026-03-04
Computation and Language · Computer Science
When Agents Look the Same: Quantifying Distillation-Induced Similarity in Tool-Use Behaviors
Chenghao Yang, Yuning Zhang, Zhoufutu Wen, Tao Gong +3
2026-04-24
Artificial Intelligence · Computer Science
On the Resilience of LLM-Based Multi-Agent Collaboration with Faulty Agents
Jen-tse Huang, Jiaxu Zhou, Tailin Jin, Xuhui Zhou +5
2025-05-30
Artificial Intelligence · Computer Science
On the Reliability of Computer Use Agents
Gonzalo Gonzalez-Pumariega, Saaket Agashe, Jiachen Yang, Ang Li +1
2026-04-21
Multiagent Systems · Computer Science
The Subtle Art of Defection: Understanding Uncooperative Behaviors in LLM based Multi-Agent Systems
Devang Kulshreshtha, Wanyu Du, Raghav Jain, Srikanth Doss +3
2026-01-13
Multiagent Systems · Computer Science
CONSCIENTIA: Can LLM Agents Learn to Strategize? Emergent Deception and Trust in a Multi-Agent NYC Simulation
Aarush Sinha, Arion Das, Soumyadeep Nag, Charan Karnati +6
2026-04-14
Artificial Intelligence · Computer Science
Agent psychometrics: Task-level performance prediction in agentic coding benchmarks
Chris Ge, Daria Kryvosheieva, Daniel Fried, Uzay Girit +1
2026-04-02
Multiagent Systems · Computer Science
Optimizing Sequential Multi-Step Tasks with Parallel LLM Agents
Enhao Zhang, Erkang Zhu, Gagan Bansal, Adam Fourney +2
2025-07-15
Computation and Language · Computer Science
ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent
Renat Aksitov, Sobhan Miryoosefi, Zonglin Li, Daliang Li +9
2023-12-18
Software Engineering · Computer Science
Coding Agents Don't Know When to Act
Thibaud Gloaguen, Niels Mündler, Mark Müller, Veselin Raychev +1
2026-05-11
Distributed, Parallel, and Cluster Computing · Computer Science
LogAct: Enabling Agentic Reliability via Shared Logs
Mahesh Balakrishnan, Ashwin Bharambe, Davide Testuggine, David Geraghty +6
2026-04-10