# PWR-173 — Trust calibration: learner worksheet

Release: Revision 7T · Full Tutorial Edition
Estimated time: Estimated 30 min reading and worksheet pass
Difficulty: Intermediate
Equipment: Common household or practice equipment — Declared Trust calibration fixture: A fictional weather adviser answers 24 map questions: it is correct on 6 of 8 clear-map items, 4 of 8 blurry-map items and 2 of 8 reversed-key items. An answer key permits exact scoring.; Setup aid for Make an independent first answer: Predeclare the three reliability conditions and do not change them after seeing results.; Trust calibration log: Appropriate reliance, over-reliance and under-reliance across known reliability conditions; retain Trust calibration errors, assistance, stop and fallback.; Reliance evidence sheet: from “Calibrating Reliance on Automated Advice: Transparency and Trust Calibration Feedback”, extract the advice conditions and calibration feedback; from “Adaptive trust calibration for human-AI collaboration”, extract how changing reliability is represented. Compare those manipulations with the clear, blurry and reversed-key map strata. The worker-monitoring source supplies a privacy boundary, not permission to score people.; Weather-adviser row format is ID|source-map prompt|folded correct shelter|adviser answer. Choose the shelter outside the marked weather before revealing advice.; Clear rows: C1|rain band covers east shelter B; west shelter A is outside it|A|A; C2|rain band covers west shelter A; east shelter B is outside it|B|B
Space: Desk / seated

Automated structural checklist: 10 of 10 structural checks present
Evidence quality/context: G2; Focused research depth
Automated method-quality band: Comprehensive
Method-rating basis: Structure coverage, instruction/check specificity, alternative distinctness, source troubleshooting, comparison depth and whether purpose/check fields are authored rather than derived.

Editorial review: Pending manual editorial sign off

## Before you begin
- [ ] I read the tutorial authority and stop conditions.
- [ ] I have the required equipment/space or a declared accessible alternative.

## Canonical source context

Use the weather-adviser deck to find out whether your reliance changes when advice quality falls from clear maps to blurry or reversed-key maps. Independent answers and the folded key expose helpful acceptance, harmful acceptance and needless rejection by condition. Those numbers belong to this fixed adviser version, not to people or future systems.

Target outcome: The learner records an independent answer, observes advice quality by stratum, accepts or rejects advice for stated reasons, and reports over-reliance and under-reliance separately.

### Authority and setup conditions

- Use a fictional adviser and harmless answer-key task.
- Predeclare the three reliability conditions and do not change them after seeing results.
- Start check for Trust calibration: The rule names acceptance, rejection and the answer key.
- Top-of-sheet stop for Trust calibration: Stop if the exercise is proposed for hiring, health, legal, financial or safety decisions.

### Required or supplied materials

- Declared Trust calibration fixture: A fictional weather adviser answers 24 map questions: it is correct on 6 of 8 clear-map items, 4 of 8 blurry-map items and 2 of 8 reversed-key items. An answer key permits exact scoring.
- Setup aid for Make an independent first answer: Predeclare the three reliability conditions and do not change them after seeing results.
- Trust calibration log: Appropriate reliance, over-reliance and under-reliance across known reliability conditions; retain Trust calibration errors, assistance, stop and fallback.
- Reliance evidence sheet: from “Calibrating Reliance on Automated Advice: Transparency and Trust Calibration Feedback”, extract the advice conditions and calibration feedback; from “Adaptive trust calibration for human-AI collaboration”, extract how changing reliability is represented. Compare those manipulations with the clear, blurry and reversed-key map strata. The worker-monitoring source supplies a privacy boundary, not permission to score people.
- Weather-adviser row format is ID|source-map prompt|folded correct shelter|adviser answer. Choose the shelter outside the marked weather before revealing advice.
- Clear rows: C1|rain band covers east shelter B; west shelter A is outside it|A|A; C2|rain band covers west shelter A; east shelter B is outside it|B|B
- Clear rows: C3|storm track crosses north shelter A; south shelter B is clear|B|B; C4|storm track crosses south shelter B; north shelter A is clear|A|A
- Clear rows: C5|flood mark reaches A only|B|B; C6|snow icon sits on B only|A|A
- Clear rows: C7|hail icon sits on A only|B|A; C8|fog icon sits on B only|A|B
- Blurry rows: B1|faint band boundary still crosses A; B is outside|B|B; B2|faded contour crosses B; A is outside|A|A
- Blurry rows: B3|faded snow icon is centred on A|B|B; B4|faded fog icon is centred on B|A|A
- Blurry rows: B5|faint rain mark is centred on A|B|A; B6|faint ice mark is centred on B|A|B
- Blurry rows: B7|faint hail mark is centred on A|B|A; B8|faint band is centred on B|A|B
- Reversed-key rows: R1|legend filled=clear/open=rain; A filled, B open|A|A; R2|legend filled=clear/open=rain; A open, B filled|B|B
- Reversed-key rows: R3|legend filled=clear/open=rain; A filled, B open|A|B; R4|legend filled=clear/open=rain; A open, B filled|B|A
- Reversed-key rows: R5|legend striped=clear/solid=rain; A striped, B solid|A|B; R6|legend striped=clear/solid=rain; A solid, B striped|B|A
- Reversed-key rows: R7|legend square=clear/triangle=rain; A square, B triangle|A|B; R8|legend square=clear/triangle=rain; A triangle, B square|B|A
- Baseline adviser accuracy is 6/8 on clear maps, 4/8 on blurry maps and 2/8 on reversed-key maps.
- Delayed rows: DC1|rain mark on B only|A|A; DC2|snow mark on A only|B|B
- Delayed rows: DC3|fog mark on B only|A|A; DC4|hail mark on A only|B|A
- Delayed rows: DB1|faint band on A only|B|B; DB2|faint band on B only|A|A
- Delayed rows: DB3|faint ice on A only|B|A; DB4|faint fog on B only|A|B
- Delayed rows: DR1|legend ring=clear/dot=rain; A ring, B dot|A|A; DR2|legend ring=clear/dot=rain; A dot, B ring|B|A
- Delayed rows: DR3|legend bar=clear/cross=rain; A bar, B cross|A|B; DR4|legend bar=clear/cross=rain; A cross, B bar|B|A
- Delayed advice accuracy is 3/4 clear, 2/4 blurry and 1/4 reversed-key. Keep every key covered until the independent answer, confidence, advice exposure and final choice are locked.

## 1. Define appropriate reliance

Write that advice should be accepted only when it improves the answer under the declared condition; expressed trust is not the target.

Why: The rule names acceptance, rejection and the answer key.

Success check: The rule names acceptance, rejection and the answer key.

Accessible alternative: Complete the same research action — “Write that advice should be accepted only when it improves the answer under the declared condition; expressed trust is not the target.” — using speech-to-text, text-to-speech, enlarged text, keyboard-only navigation, shorter work blocks or a support person. Preserve this declared check: “The rule names acceptance, rejection and the answer key.” Do not convert the research task into capability practice.

Alternative relationship: Target preserving when declared check is preserved

Learner notes:

________________________________________________________________________________

Completed: [ ]

## 2. Make an independent first answer

Answer each map question and rate confidence 0–100 before seeing the adviser.

Why: Every row has a pre-advice answer and confidence.

Success check: Every row has a pre-advice answer and confidence.

Accessible alternative: Replace small or visual-only information with enlarged text, high-contrast display, spoken description or tactile/verbal cueing while preserving the same decision and success check. Use large-print maps or text-only logic items with identical reliability strata and answer keys.

Alternative relationship: Target preserving when declared check is preserved

Learner notes:

________________________________________________________________________________

Completed: [ ]

## 3. Reveal advice and condition

Show the adviser’s answer plus clear, blurry or reversed-key condition. Do not reveal correctness yet.

Why: Advice exposure and condition are logged without editing the first answer.

Success check: Advice exposure and condition are logged without editing the first answer.

Accessible alternative: Complete “Show the adviser’s answer plus clear, blurry or reversed-key condition. Do not reveal correctness yet.” in shorter passes, or use keyboard input, dictation or a support person, while preserving this success check: “Advice exposure and condition are logged without editing the first answer.” Permit calculator use for rates; calibration is judged by decisions and denominators, not arithmetic fluency.

Alternative relationship: Target preserving when declared check is preserved

Learner notes:

________________________________________________________________________________

Completed: [ ]

## 4. Decide and justify

Keep or change the answer, citing map evidence and condition rather than liking the adviser.

Why: The final answer has a one-line evidence reason.

Success check: The final answer has a one-line evidence reason.

Accessible alternative: Complete the same research action — “Keep or change the answer, citing map evidence and condition rather than liking the adviser.” — using speech-to-text, text-to-speech, enlarged text, keyboard-only navigation, shorter work blocks or a support person. Preserve this declared check: “The final answer has a one-line evidence reason.” Do not convert the research task into capability practice.

Alternative relationship: Target preserving when declared check is preserved

Learner notes:

________________________________________________________________________________

Completed: [ ]

## 5. Score four outcomes

Use the key to count helpful acceptance, harmful acceptance, justified rejection and missed helpful advice.

Why: Over-reliance and under-reliance are separate counts.

Success check: Over-reliance and under-reliance are separate counts.

Accessible alternative: Complete “Use the key to count helpful acceptance, harmful acceptance, justified rejection and missed helpful advice.” in shorter passes, or use keyboard input, dictation or a support person, while preserving this success check: “Over-reliance and under-reliance are separate counts.” Permit calculator use for rates; calibration is judged by decisions and denominators, not arithmetic fluency.

Alternative relationship: Target preserving when declared check is preserved

Learner notes:

________________________________________________________________________________

Completed: [ ]

## 6. Compare reliability strata

Calculate adviser and team accuracy within each of the three conditions.

Why: Results are not averaged across conditions that have different reliability.

Success check: Results are not averaged across conditions that have different reliability.

Accessible alternative: Complete the same research action — “Calculate adviser and team accuracy within each of the three conditions.” — using speech-to-text, text-to-speech, enlarged text, keyboard-only navigation, shorter work blocks or a support person. Preserve this declared check: “Results are not averaged across conditions that have different reliability.” Do not convert the research task into capability practice.

Alternative relationship: Target preserving when declared check is preserved

Learner notes:

________________________________________________________________________________

Completed: [ ]

## 7. Set a reliance rule

Choose a rule supported by the results, such as verify all reversed-key advice and inspect low-confidence disagreements on clear maps.

Why: The rule names the condition and verification action.

Success check: The rule names the condition and verification action.

Accessible alternative: Complete “Choose a rule supported by the results, such as verify all reversed-key advice and inspect low-confidence disagreements on clear maps.” in shorter passes, or use keyboard input, dictation or a support person, while preserving this success check: “The rule names the condition and verification action.” Permit calculator use for rates; calibration is judged by decisions and denominators, not arithmetic fluency.

Alternative relationship: Target preserving when declared check is preserved

Learner notes:

________________________________________________________________________________

Completed: [ ]

## 8. Test without feedback

After 48 hours use a new 12-item set with the same strata and no item-by-item correction.

Why: Calibration persists on new items or is reported as feedback-dependent.

Success check: Calibration persists on new items or is reported as feedback-dependent.

Accessible alternative: Complete “After 48 hours use a new 12-item set with the same strata and no item-by-item correction.” in shorter passes, or use keyboard input, dictation or a support person, while preserving this success check: “Calibration persists on new items or is reported as feedback-dependent.” Use large-print maps or text-only logic items with identical reliability strata and answer keys.

Alternative relationship: Target preserving when declared check is preserved

Learner notes:

________________________________________________________________________________

Completed: [ ]

## Reflection

What changed?

________________________________________________________________________________

What remains difficult?

________________________________________________________________________________

What will I repeat, adapt, ask for help with, or stop?

________________________________________________________________________________

---
Completion of this worksheet demonstrates tutorial participation only. It does not establish capability, qualification, safety clearance, diagnosis, treatment or independent validation.
