Board 3 of 3
Efficiency
What that capability costs. Em per dollar and Em per thousand context tokens, over the same rows as the other two boards. This is where equal accuracy at very different token bills stops being a footnote and becomes the ranking.

0 ranked
| # | Model | Memory | Effort | Em per $ | Em | Cost | Ctx tok | Runs | Corpus |
|---|---|---|---|---|---|---|---|---|---|
| No rows can be ranked on this view with the current filters. | |||||||||
How this is scored
- Unit
- Effective Minutes (Em) — quality-adjusted minutes of expert work replaced.
- Lift
- Em(rig) minus Em(same model, same effort, vendor-native) within the same surface and corpus. Null when no matched baseline exists — never zero. A positive lift on "None" means the bare model beat having the whole corpus in context, which is a real result at large corpus sizes, not an error.
- Anchors
- T_human anchors are PROVISIONAL expert estimates, not panel-measured.