Board 3 of 3

Efficiency

What that capability costs. Em per dollar and Em per thousand context tokens, over the same rows as the other two boards. This is where equal accuracy at very different token bills stops being a footnote and becomes the ranking.

Efficiency

0 ranked

#ModelMemoryEffortEm per $EmCostCtx tokRunsCorpus
No rows can be ranked on this view with the current filters.

How this is scored

Unit
Effective Minutes (Em) — quality-adjusted minutes of expert work replaced.
Lift
Em(rig) minus Em(same model, same effort, vendor-native) within the same surface and corpus. Null when no matched baseline exists — never zero. A positive lift on "None" means the bare model beat having the whole corpus in context, which is a real result at large corpus sizes, not an error.
Anchors
T_human anchors are PROVISIONAL expert estimates, not panel-measured.