scikit_posthocs.posthoc_duncan

scikit_posthocs.posthoc_duncan(a: list | ndarray | DataFrame, val_col: str | None = None, group_col: str | None = None, sort: bool = False) DataFrame

Duncan’s multiple range test for normally distributed data with equal group variances, following a parametric ANOVA [1].

Parameters:
  • a (Union[list, np.ndarray, DataFrame]) – An array, any object exposing the array interface or a pandas DataFrame.

  • val_col (str, optional) – Name of a DataFrame column that contains dependent variable values (test or response variable). Values should have a non-nominal scale. Must be specified if a is a pandas DataFrame object.

  • group_col (str, optional) – Name of a DataFrame column that contains independent variable values (grouping or predictor variable). Values should have a nominal scale (categorical). Must be specified if a is a pandas DataFrame object.

  • sort (bool, optional) – If True, sort data by group columns.

Returns:

result – P values.

Return type:

pandas.DataFrame

Notes

Like posthoc_snk, this is a stepwise range test using as nmeans the number of ordered means a pair spans. The studentized-range p value for each pair is further Bonferroni-adjusted as 1 - (1 - p) ** (1 / (nmeans - 1)), following Duncan (1955). There is no separate p value adjustment argument, since the step-down procedure is itself the adjustment.

References

Examples

>>> import scikit_posthocs as sp
>>> import pandas as pd
>>> x = pd.DataFrame({"a": [1,2,3,5,1], "b": [12,31,54,62,12], "c": [10,12,6,74,11]})
>>> x = x.melt(var_name='groups', value_name='values')
>>> sp.posthoc_duncan(x, val_col='values', group_col='groups')