{# Paired statistical evidence for "B vs A". Expects `p` (PairedComparison). #} {% if p %}
{{ p.b }}: {{ p.verdict }} vs {{ p.a }} over {{ p.n_tasks }} paired task{{ '' if p.n_tasks == 1 else 's' }} (wins {{ p.wins }}, losses {{ p.losses }}, ties {{ p.ties }}{% if p.sign_test_p is not none %}, sign test p = {{ '%.3f'|format(p.sign_test_p) }}{% endif %})
| per-task difference (B − A) | mean | {{ ((p.pass_rate_diff.level if p.pass_rate_diff else 0.95) * 100)|round|int }}% interval | P(> 0) | |||||
|---|---|---|---|---|---|---|---|---|
| {{ label }} | {% if iv %} {% if unit == "pts" %}{{ (iv.estimate * 100)|delta(1) }} pts | [{{ (iv.low * 100)|delta(1) }}, {{ (iv.high * 100)|delta(1) }}] | {% else %}{{ iv.estimate|delta(4) }} | [{{ iv.low|delta(4) }}, {{ iv.high|delta(4) }}] | {% endif %}{{ iv.p_positive|pct }} | {% else %}— | — | — | {% endif %}
Fewer than {{ p.min_tasks }} paired tasks: the interval is shown, but no verdict is given. Add tasks (repetitions do not count as tasks).
{% endif %}