Skip to content
EvalSuite
Documentation menu

Statistical tests

Friedman test

Implementedstatistics.friedman

Definition

Rank-based test for three or more related samples, such as several models evaluated on the same datasets.

Formula

χ²_F = (12 / nk(k + 1)) Σ Rⱼ² − 3n(k + 1)

Inputs and outputs

  • n × k matrix of scores

Returns: StatisticalTestResult

Assumptions

No assumptions beyond valid, aligned inputs of the documented types.

Limitations

No metric-specific limitations are documented yet. Interpret the value alongside the task, data, and other metrics.

Python API

PythonSince v0.2.0
import evalsuite as es

es.friedman_test(scores_a, scores_b, scores_c)

References

  1. Friedman, M. (1937). The use of ranks to avoid the assumption of normality implicit in the analysis of variance. Journal of the American Statistical Association, 32(200), 675–701.

Implementation status