Analyzing AI Evaluation Benchmarks Through Information Retrieval and Network Science
Poster - The 48th European Conference on Information Retrieval (ECIR 2026). Delft, The Netherlands.
Poster - The 48th European Conference on Information Retrieval (ECIR 2026). Delft, The Netherlands.
Poster - The 48th European Conference on Information Retrieval (ECIR 2026). Delft, The Netherlands.
Traditionally, relevance judgments have relied on human annotators, but recent advances in Large Language Models (LLMs) have prompted growing interest in their use as a proxy for …
Many analyses have been performed on Information Retrieval (IR) evaluation benchmarks. Benchmarking also plays a central role in evaluating the capabilities of Large Language …
Patient Information Leaflets (PILs) provide essential information about medication usage, side effects, precautions, and interactions, making them a valuable resource for Question …
The scholarly publishing process relies on peer review to uphold the quality of scientific knowledge. However, challenges such as increasing submission volumes and potential …
While the concept of responsible AI is becoming more and more popular, practitioners and researchers may often struggle to characterize responsible practices in their own work. …