Talk Title: Toward Reliable Deployment of LLM Agents: From Uncertainty Quantification to Progress Advantage
LLM agents are increasingly deployed on long-horizon tasks with tool use, irreversible actions, and unpredictable feedback. Yet we have few principled ways to tell, mid-episode, whether an agent is on track or quietly failing. Most uncertainty quantification (UQ) research still centers on single-turn QA, a poor match for interactive agents. In this talk, I'll present a general formulation of agent UQ and the challenges unique to agentic settings, from choosing uncertainty estimators to modeling how uncertainty evolves over an interaction. I'll then show that a powerful answer has been hiding in plain sight: RL post-training already yields an implicit step-level signal, the progress advantage, which recovers the optimal advantage function with no annotation or reward-model training. Across test-time scaling, UQ, and failure attribution, this free byproduct beats confidence baselines and even dedicated trained reward models.
Bio: Sharon Li is an Associate Professor in the Department of Computer Sciences at the University of Wisconsin-Madison. Her research focuses on algorithmic and theoretical foundations of reliable machine learning, addressing challenges in both model development and deployment in the open world. Previously, she was a postdoc researcher in the Computer Science department at Stanford University. She completed her Ph.D. from Cornell University, advised by John E. Hopcroft. She has served as the Program Chair for ICML 2026. She was the recipient of Alfred P. Sloan Fellowship (2025), NSF CAREER Award (2023), MIT Innovators Under 35 Award (2023), AFOSR Young Investigator Award (2022), Forbes 30under30 in Science (2020), and multiple faculty research awards from Google, Meta, and Amazon. She was named the “Innovator of the Year 2023” by MIT Technology Review. Her work has won the Outstanding Paper Award at NeurIPS 2022 and ICLR 2022.