All presentationsPoster Mar 30, 2026Analyzing AI Evaluation Benchmarks Through Information Retrieval and Network ScienceGaia Simeoni, Michael Soprano, Riccardo Lunardi, Kevin Roitero, Stefano MizzaroPDFHITSbenchmarkslarge language modelsRelatedAnalyzing AI Evaluation Benchmarks Through Information Retrieval and Network ScienceLarge Language Models as Assessors: On the Impact of Relevance ScalesPILs of Knowledge: A Synthetic Benchmark for Evaluating Question Answering Systems in HealthcareLarge Language Models as Assessors: On the Impact of Relevance ScalesLarge Language Models for Combinatorial Optimization: A Systematic Review