Oct 28-30: Exploring AI in Evaluation and Evaluation of AI at the 2026 European Evaluation Society (EES) Conference
Join MTI and MERL Tech Community of Practice members at the European Evaluation Society Conference in Lille, France, from October 28-30. Linda Raftree (MTI) is co-leading S13: How is emerging artificial intelligence intersecting with evaluation and democracy? Towards new practice and adoption with Bianca Montrosse-Moorhead (University of Connecticut), Steffen Bohni Nielsen (National Research Centre for the Working Environment) and Alix de Saint-Albin (Pluricité), with the goal of exploring Generative AI in Evaluation from various angles together with EES participants.
The strand has 6 sessions:
- Session 13a: “Level Up” Emerging AI and Evaluation Practice on Wed Oct 28, 10:15-11.45 CET (Bianca Montrosse-Moorhead, Linda Raftree, Steffen Bohni Nielsen and Alix de Saint-Albin). The session aims to bring conference participants up to speed on the state of the field of AI in Evaluation as a grounding and starting point for subsequent Strand 13 sessions.
- Session 13b: Living the tensions: Responsible AI evaluation cases on Wed 28 Oct 13-14.30 CET with Steffen Bohni Nielsen, Julian King, Youngjin Kim (Tetra Tech), Samantha Abbato (Visual Insights) Alana Kinarsky (UCLA), Leslie Fierro (McGill University) and Elyse McCall-Thomas (University of Ottawa). The session showcases several example of AI in Evaluation, with a focus on what makes them “responsible.”
- S13c: Designing MEL for GenAI programming: tensions and bright spots on Wed 28 15-16.30 CET with Tetyana Zelenska (Digital Green), Valentine Gandhi (Dev Cafe), Emeka Nwankwo (MTI), and Linda Raftree. The session offers examples of different emerging responsible MEL approaches and practices for evaluating global development and social programs that include AI as a key enabler.
- S13d: Responsible GenAI in Evaluation: Competencies, FRAME, and Emerging Practice on Thu 29 Oct 13-14.30 CET with Bianca Montrosse-Moorhead, Kerry Bruce (Clear Up Consulting), Anastasia (Tessie) Catsambas (Encompass LLC) and Valentine Gandhi (Dev Cafe). The session covers conceptual, practical, and case-based insights into how GenAI is being responsibly integrated into evaluation and what this means for the field.
- S13e: Critical approaches to AI in evaluation standard setting from African, Indigenous and European perspectives on Thu 29 Oct 15-16.30 CET with Varaidzo Matimba (MTI), Andrealisa Belzer (Canadian Evaluation Society), Awuor Ponge (EvalIndigenout Global Network), Larry Bremner (Proactive Information Services, Inc) and Alix de Saint-Albin. This session will discuss Made in Africa AI for Made in Africa Evaluation, the Wolastoq Declaration on Indigenous Evaluation, and EU language and cultural diversity hierarchies reinforced by commercial AI, and then open for a fishbowl on how this does or should influence how we use AI for Evaluation.
- S13f: Building a Shared Agenda for AI, Evaluation and Democracy on Fri 30 Oct 10.30-12.00 CET with Alix de Saint-Albin, Steffen Bohni Nielsen, Bianca Montrosse-Moorhead and Linda Raftree. This session sums up conversations from the strand and aims to generate ideas for shared global standards for the Responsible Use of AI in Evaluation as well as the Evaluation of the Use of AI. It proposes the formation of a global working group to create as set of Standards for the Field of Evaluation.
