0 papers in the Human Centered Ai Evaluation research area.
Discover 1 peer-reviewed study in Human Centered Ai Evaluation (2026). Explore research findings powered by Prolific's diverse participant panel.
This page lists 1 peer-reviewed paper in the research area of Human Centered Ai Evaluation in the Prolific Citations Library, a curated collection of research powered by high-quality human data from Prolific.
Authors: N Petrova, A Gordon, E Blindow
Year: 2026
Published in: Open review
Institution: Prolific
Research Area: Human-centered AI evaluation, Bayesian statistics, Responsible AI, AI alignment, LLM Evaluation
Discipline: Machine Learning, Artificial Intelligence
The study introduces HUMAINE, a multidimensional evaluation framework for LLMs, revealing demographic-specific preference variations and ranking google/gemini-2.5-pro as the top-performing model with a posterior probability of 95.6%.
Methods: Multi-turn naturalistic conversations analyzed using a hierarchical Bayesian Bradley-Terry-Davidson model with post-stratification to census data, stratified across 22 demographic groups.
Key Findings: Performance of 28 LLMs across five human-centric dimensions, accounting for demographic-specific preferences.
Sample Size: 23404
Related Disciplines: Machine Learning, Artificial Intelligence
Related Institutions: Prolific
Researchers: N Petrova, A Gordon, E Blindow
Publication Years: 2026
Browse all papers in the Prolific Citations Library