Skip to contents

The difference between two evaluations of the same rows, metric by metric: the mean score before and after, their difference with its 95% interval (paired: each row's difference), and how many rows got better, worse, or stayed the same. A difference whose interval holds 0 may be noise.

Usage

compare(before, after)

Arguments

before, after

Evaluations (evaluate()) of the same rows, in the same order.

Value

A tibble: metric, before, after, diff, conf.low, conf.high, better, worse, same, n.