Third Workshop on Human-Centered Evaluation and Auditing of Language Models: AI Agents-in-the-Loop
- ,
- Wesley Hanwen Deng,
- Yu Lu Liu,
- Han Jiang,
- Valerie Chen,
- Haotian Li
- ,
- ,
- Carnegie Mellon University,
- Johns Hopkins University,
- Microsoft Research Asia,
- Korea Advanced Institute of Science and Technology
Research Output:
Conference Article in Proceeding or Book/Report chapter
Article in proceedings
Peer-reviewOpen access
Publication Information
Output type
Research Output:
Conference Article in Proceeding or Book/Report chapter
Article in proceedings
Peer-reviewOriginal language
EnglishArticle number
975Pages from-to (Number of pages)
Pages 1-7 (8 pages)Publication milestones
- Published - 13/04/2026
Publication status
Published - 13/04/2026
Publisher
Association for Computing Machinery, United StatesISBN (Print)
9798400722813Publication IDs
- ORCID: /0000-0003-0245-1633/work/211486309
- Scopus: 105038103359
Host publication title
CHI EA '26: Proceedings of the Extended Abstracts of the 2026 CHI Conference on Human Factors in Computing SystemsAbstract
Large Language Models (LLMs) are increasingly deployed in real-world applications but also pose significant risks. Our previous workshop iterations at CHI'24 and CHI'25 brought together HCI and AI researchers to address the “evaluation crisis” through human-centered evaluation approaches. As the demand for human-centered evaluation grows, a new frontier has emerged: practitioners are increasingly turning to AI agents themselves as tools to support evaluation processes. This third iteration introduces the theme of AI agents-in-the-loop, exploring the emerging frontier where human judgment meets agent automation in LLM evaluation workflows. The workshop will examine critical questions about task allocation between humans and agents, meta-evaluation of evaluator agents, and the design of safeguards that preserve human agency while benefiting from automation. Through position papers, keynote discussion, and collaborative activities, participants will identify key challenges, share emerging practices, and outline research directions for hybrid human-AI auditing and evaluation approaches.
Publication metrics
PlumX, opens in new tab
Captures
10
Access to documents
Related Event
Title
Conference on Human Factors in Computing Systems
Event type
ConferenceDegree of recognition
International eventDate
13/04/2026 - 17/04/2026Location
Centre de Convencions Internacional de Barcelona.BarcelonaSpain
