AI evaluation
For teams shipping systems that make claims about people.
- Dataset and benchmark design and collection
- Model evaluation
- Generative psychometrics
Consulting through Girard Consulting LLC
68publications, 2011–2026 10lab-led 30open preprints 23with data or code 19open-source tools 6datasets and benchmarks
About
I’m an Associate Professor of Psychology and Data Science at the University of Kansas, where I direct the Affective Communication and Computing Lab. I trained as a clinical psychologist — PhD, University of Pittsburgh, 2018 — and my research is about measuring things that resist measurement: emotion, clinical state, the texture of a conversation.
That work has appeared in Annual Review of Clinical Psychology, Journal of Consulting and Clinical Psychology and Affective Science, and it has produced open-source tools that other researchers use.
Through Girard Consulting I bring the same standards to applied work — dataset and benchmark design, model evaluation and generative psychometrics for teams building systems that reason about people. I’m also a co-founder of SMaRT Workshops and Principal Scientist at Fluid Concepts Research.
For teams shipping systems that make claims about people.
For work that has to survive review.
For teams who want the capability in-house.
Every tool carries its calibration: 9 in calibration, 8 under calibration, 2 withdrawn. I would rather you know which is which.
Analyzing and visualizing circular data.
R
Hierarchical structure (bass-ackwards) analysis.
R
Intraclass correlations (ICCs).
R
Working with Hierarchical Taxonomy of Psychopathology data.
R
Automating common affective computing and social signal processing tasks.
R
Multiarchitecture image combining RStudio and r2u.
Docker
Affective Science, 7, 166–180
Annual Review of Clinical Psychology, 22, 49–75
Journal of Consulting and Clinical Psychology, 93, 749–760
Proceedings of the 11th International Conference on Affective Computing and Intelligent Interaction (ACII), 1–8