Get to know one another
Share backgrounds, current work, and the evaluation questions each person is thinking about.
Better evaluations for more capable agents
A small recurring research group for people actively working on AI agent evaluation.
About
This is the first meeting of a recurring Boston-area group for researchers actively building or studying AI agent evaluations.
The immediate goal is to help participants get to know one another and build a shared agenda for future meetings.
First meeting
Share backgrounds, current work, and the evaluation questions each person is thinking about.
Identify the problems, research directions, and collaborations worth carrying into future meetings.
Who should apply
Applicants should have hands-on AI agent evaluation experience, including evaluations, agent systems, or related empirical research.
The group is limited to 10 participants. If this meeting does not work for you, or you want to hear about later sessions, send us a general expression of interest.