This event is in the past.It took place on April 29, 2026 at West Hall.

Principled Evaluation of Large Language Models: A Statistical Perspective 1
Educational Talk
Principled Evaluation of Large Language Models: A Statistical Perspective 2
Principled Evaluation of Large Language Models: A Statistical Perspective 3
Principled Evaluation of Large Language Models: A Statistical Perspective 4
Principled Evaluation of Large Language Models: A Statistical Perspective 5

Principled Evaluation of Large Language Models: A Statistical Perspective

April 29 at 9am

At West Hall

Burns Park, Ann Arbor

🎓Open to the public, no registration needed
🤖Covers cutting-edge LLM evaluation methods

Explore how statistics can make AI evaluation more rigorous. This dissertation defense presents three novel frameworks for testing large language models with scientific precision, from prompt sensitivity analysis to bridging human and automated judgments.

Academic talk
Dissertation defense
AI research
Statistics focus