These get confused constantly in review meetings. They are separate quantities with separate owners.
The Rasch model ties all three together with one line: P(pass) = 1 / (1 + e−(θ − b)). When an LoB's maturity equals an item's difficulty, it has a coin-flip chance of passing. Above the difficulty, better than even. Below it, worse.
Flip any response below. Every number on this page recomputes from what you set here.
| Item | Assessment check | Difficulty scale −2 … +2 | b | Observed |
|---|
For a candidate θ, the model predicts a pass probability for every item. We then ask: how likely is the pattern we actually saw? Pass items contribute P. Fail items contribute 1 − P. Multiply the four together and you get the likelihood of that candidate.
The peak sits near the point where the passed items are comfortably below the LoB's maturity and the failed items are comfortably above it. That crossover is the estimate.
Weight belongs to the framework: how much Security matters in the Northstar rating, set once by architecture governance. θ belongs to the LoB: how mature this one team actually is. Drag the weights — the overall score moves, the θ bars do not.
Weight changes the question — what this portfolio cares about. θ changes only when the LoB's evidence changes. Raising the Security weight does not make Payments any more secure; it just makes their security gap count for more.
LoB Ledger passes the two easy items. LoB Vault passes the two hard ones. A checklist counts both as 2 of 4. Watch what the model does.