4-page short papers
Submissions on new agent evaluation methods, RL environment design, agentic benchmarks, and real-world case studies; work-in-progress submissions are encouraged.
RLEval: Methods and Reinforcement Learning Environments for Evaluating AI Agents
The RLEval workshop at ACM CAIS 2026 is the first research venue dedicated to evaluating AI agents, focusing on reinforcement learning environments, evaluation methods, benchmarks, and real-world deployment case studies. It aims to address critical open questions in measuring and improving agentic capabilities through interventional, causal, and counterfactual techniques.
Paper fit
A strong submission should clearly identify its contribution and evaluate it appropriately.
Submissions on new agent evaluation methods, RL environment design, agentic benchmarks, and real-world case studies; work-in-progress submissions are encouraged.
Compiled from the official call for papers. The organizers’ pages remain authoritative.
Last verified September 9, 2026