9:30 AM
9:50 AM
Welcome Address
By Danilo Bzdok & Siva Reddy
10:00 AM
Evaluating System-Level Reasoning in LLM Agents
Jacob Andreas
10:45 AM
Don’t Forget the User: Balancing the Scales in Agentic Training and Evaluation
Seraphina Goldfarb-Tarrant
11:30 AM
Recap. Discussion Audience/ G. Speakers
12:00 PM
Lunch on your own
1:45 PM
Memorization: Myth or Mystery?
Verna Dankers
2:30 PM
Towards Scalable and Actionable Interpretability
Yonatan Belinkov