PhD or equivalent experience in Human–Computer Interaction, Computer Science, Cognitive Science, or a related field, with a strong emphasis on empirical evaluation of interactive AI/LLM systems.
3+ years of academic or industry research experience post-PhD, including leadership on complex research initiatives and analyzing data from a real AI product.
Strong publication record, with demonstrated impact in top-tier AI (NeurIPS, ICML, ICLR, ACL) and HCI (CHI) venues
Deep expertise in experimental design and measurement, particularly for:
Task performance and human activity
Comparative evaluation frameworks
Mixed-methods research grounded in real-world behavior
Strong technical and coding skills, including:
Python and data analysis / ML tooling
Experience building experimental systems and benchmark infrastructure
Familiarity working with LLM APIs, agent frameworks, or AI-assisted tooling
Proven ability to define and lead research agendas that connect human work, AI capability, and business or economic impact.
Strong collaboration skills, especially working across research, engineering, product, and UXR teams.