TY - UNPB
T1 - Deriving Consensus Rankings from Benchmarking Experiments
AU - Hornik, Kurt
AU - Meyer, David
PY - 2006
Y1 - 2006
N2 - Whereas benchmarking experiments are very frequently used to investigate the performance of statistical or machine learning algorithms for supervised and unsupervised learning tasks, overall analyses of such experiments are typically only carried out on a heuristic basis, if at all. We suggest to determine winners, and more generally, to derive a consensus ranking of the algorithms, as the linear order on the algorithms which minimizes average symmetric distance (Kemeny-Snell distance) to the performance relations on the individual benchmark data sets. This leads to binary programming problems which can typically be solved reasonably efficiently. We apply the approach to a medium-scale benchmarking experiment to assess the performance of Support Vector Machines in regression and classification problems, and compare the obtained consensus ranking with rankings obtained by simple scoring and Bradley-Terry modeling.
AB - Whereas benchmarking experiments are very frequently used to investigate the performance of statistical or machine learning algorithms for supervised and unsupervised learning tasks, overall analyses of such experiments are typically only carried out on a heuristic basis, if at all. We suggest to determine winners, and more generally, to derive a consensus ranking of the algorithms, as the linear order on the algorithms which minimizes average symmetric distance (Kemeny-Snell distance) to the performance relations on the individual benchmark data sets. This leads to binary programming problems which can typically be solved reasonably efficiently. We apply the approach to a medium-scale benchmarking experiment to assess the performance of Support Vector Machines in regression and classification problems, and compare the obtained consensus ranking with rankings obtained by simple scoring and Bradley-Terry modeling.
U2 - 10.57938/bc3712be-1fa9-4ff4-9b9e-dfebfb2c244f
DO - 10.57938/bc3712be-1fa9-4ff4-9b9e-dfebfb2c244f
M3 - WU Working Paper and Case
T3 - Research Report Series / Department of Statistics and Mathematics
BT - Deriving Consensus Rankings from Benchmarking Experiments
PB - Department of Statistics and Mathematics, WU Vienna University of Economics and Business
CY - Vienna
ER -