#Alejandro Vidal
2 articles with this tag

AI Research
AI Benchmarking: Beyond 50s Metrics
Alejandro Vidal argues for adopting psychometric principles like IRT to move beyond outdated LLM benchmarking, enabling more nuanced and reliable model evaluation.
2 months ago

AI Research
Alejandro Vidal on Rethinking LLM Evaluation with Psychometrics
Alejandro Vidal of Mindmakers advocates for integrating psychometrics into LLM evaluation to move beyond simplistic accuracy scores and gain deeper insights into model intelligence and benchmark quality.
2 months ago