Section 139 of 440

FAMILY 16: METACOGNITION & EPISTEMIC RESILIENCE

Superdomain: Cognition & Learning. This family contains eight stable Power records.

PWR-121 · Confidence calibration

Evidence route: Evidence dossier with a non-prescriptive guide frame. Use the linked canonical tutorial for current teaching, accessibility, safety and editorial-review status; this evidence dossier does not enlarge that tutorial’s authority.

Summary. Calibration means matching confidence to outcomes—not simply becoming less confident.

Definition. Capability to develop or express confidence calibration in a declared context without inheriting broader claims.

What the current evidence supports. Confidence can be scored against outcomes, but scalable calibration training has conflicting task-dependent evidence.

Measurement boundary. Declare question class, resolution rules, probability scale, calibration, resolution, base rate, scoring, sample size and missing outcomes; average confidence is not calibration.

Myth. Always know exactly how sure to be

Metric. Calibration and resolution over a predeclared set of resolved judgments

Boundary. Good calibration in one question class does not prove wisdom or transfer to another.

Negative and limiting findings.

  • Practical-scoring feedback failed in two large experiments.

  • Automated feedback improved calibration in modified blackjack but not overall calibration in a more realistic baseball task.

Claims this evidence cannot support.

  • always know what you do not know

  • eliminate overconfidence

  • calibration score proves wisdom

  • AI gives objective confidence

Bounded non-prescriptive guide frame

Purpose. Explain confidence calibration against resolved outcomes without turning uncertainty into a personal score or universal wisdom claim.

This frame supports conceptual observation and reflection only. It collects no answers, stores nothing, produces no score, and does not become a training protocol.

What this frame may discuss.

  • Evidence and measurement boundaries for confidence calibration

  • Task specificity, access configurations and negative findings

  • Limits on transfer, efficacy and authority

Conceptual observations.

  • Conceptually distinguish average confidence, calibration and resolution.

  • Treat the question class, base rate, outcome rules and missing outcomes as part of the result.

Reflection prompts.

  • What would count as a resolved outcome?

  • Could calibration in one question class transfer poorly to another?

Accessibility alternatives.

  • Fictional benign judgments are sufficient; no submission or computation is needed.

  • Words, percentages, frequencies and visual scales are alternative representations.

  • No confidence diary, tracking schedule or score is requested.

Stop the frame if.

  • Stop if the frame becomes repeated self-scoring, a certainty target or an optimisation loop.

  • Do not use confidence labels as competence, credibility or clinical judgments about a person.

Escalation boundary.

  • Consequential judgments require domain evidence, independent checks and accountable human review.

Prohibited uses.

  • Calibration exercise, score, rank, target, diary or progression.

  • Clinical, employment, credibility or competence assessment.

  • Claims of universal wisdom or reliable transfer.

What it cannot establish.

  • Good calibration in one question class does not prove wisdom or transfer to another.

  • It cannot establish: always know what you do not know.

  • It cannot establish: eliminate overconfidence.

  • It cannot establish: calibration score proves wisdom.

  • It cannot establish: AI gives objective confidence.

Evidence references. [1] primary empirical support; limiting or contrary evidence; [3] primary empirical support; limiting or contrary evidence; [10] primary empirical support; limiting or contrary evidence; [16] limiting or contrary evidence; official boundary context.

Next gate. No external confirmation is required for the current bounded explainer and non-prescriptive guide-frame permissions; any stronger efficacy, generalisation, protocol, promotion or independent-validation claim requires a new adjudication.

PWR-122 · Error detection

Evidence route: Externally gated evidence dossier. Use the linked canonical tutorial for current teaching, accessibility, safety and editorial-review status; this evidence dossier does not enlarge that tutorial’s authority.

Summary. Error detection improves when people and interfaces are designed together—but rare mistakes remain easy to miss.

Definition. Capability to develop or express error detection in a declared context without inheriting broader claims.

What the current evidence supports. Selected interfaces can reduce specific human–AI and rare-target errors, but this is configured system performance rather than a universal human faculty.

Measurement boundary. Declare error taxonomy, prevalence, base rate, false alarms, misses, correction path, model/version, workload and consequences; high accuracy can hide rare catastrophic misses.

Myth. Spot every hidden error

Metric. Miss rate, false-alarm rate, time to correction and harm under declared prevalence

Boundary. One detection task cannot establish universal vigilance, honesty detection or safety.

Negative and limiting findings.

  • Very rare targets produced high miss rates despite extensive task exposure.

  • Cognitive forcing reduced overreliance but lowered user ratings and varied across people.

Claims this evidence cannot support.

  • spot every mistake

  • human lie detector

  • AI-proof your mind

  • error score identifies careless people

Evidence references. [9] primary empirical support; limiting or contrary evidence; [10] primary empirical support; limiting or contrary evidence; [12] primary empirical support; limiting or contrary evidence; [16] limiting or contrary evidence; official boundary context.

Next gate. External confirmation is required before higher-authority use because: a clinical, high-risk or regulated configuration.

PWR-123 · Bias resistance

Evidence route: Externally gated evidence dossier. Use the linked canonical tutorial for current teaching, accessibility, safety and editorial-review status; this evidence dossier does not enlarge that tutorial’s authority.

Summary. People can become better at spotting selected manipulation techniques—but nobody becomes immune to misinformation.

Definition. Capability to develop or express bias resistance in a declared context without inheriting broader claims.

What the current evidence supports. Brief prompts and transparent games can reduce selected misinformation-related errors, but they do not create general immunity to bias.

Measurement boundary. Predeclare bias/error type, content, ground truth, item novelty, intention versus behavior, delay, false positives and political/cultural context.

Myth. Vaccinate the mind against misinformation

Metric. Accuracy, false positives and observed behavior on novel benign items after delay

Boundary. A prompt or game cannot create general immunity, political correctness or permanent resistance.

Negative and limiting findings.

  • Most accuracy-prompt outcomes were sharing intentions rather than observed long-term behavior.

  • Inoculation estimates were sensitive to the examples used.

Claims this evidence cannot support.

  • become immune to misinformation

  • remove cognitive bias

  • mental vaccine

  • program people against wrong beliefs

Evidence references. [4] primary empirical support; limiting or contrary evidence; [11] primary empirical support; limiting or contrary evidence; [15] primary empirical support; limiting or contrary evidence.

Next gate. External confirmation is required before higher-authority use because: a clinical, high-risk or regulated configuration; cultural, community-rights or affected-person authority; a restricted safety or legitimacy boundary.

PWR-124 · Source evaluation and lateral reading

Evidence route: Evidence dossier with a non-prescriptive guide frame. Use the linked canonical tutorial for current teaching, accessibility, safety and editorial-review status; this evidence dossier does not enlarge that tutorial’s authority.

Summary. Source checking can be learned, but search access, language, platform design and time shape what is possible.

Definition. Capability to develop or express source evaluation and lateral reading in a declared context without inheriting broader claims.

What the current evidence supports. Lateral reading and deliberate source checking can improve selected credibility judgments, with small and not uniformly retained effects.

Measurement boundary. Declare claim, source chain, search environment, language, access barriers, independent corroboration, time, accuracy and delayed novel-item test.

Myth. Know instantly whether anything online is true

Metric. Accurate source/provenance judgments on novel claims with search path declared

Boundary. One checklist, search result or AI citation cannot certify truth or make someone scam-proof.

Negative and limiting findings.

  • Many intervention estimates were null at two weeks.

  • No between-intervention advantage was detected at follow-up.

Claims this evidence cannot support.

  • instantly know what is true

  • Google it and verify anything

  • become scam-proof

  • source score proves credibility

Bounded non-prescriptive guide frame

Purpose. Explain lateral source checking as one bounded aid to credibility judgment, not an instant truth detector.

This frame supports conceptual observation and reflection only. It collects no answers, stores nothing, produces no score, and does not become a training protocol.

What this frame may discuss.

  • Evidence and measurement boundaries for source evaluation and lateral reading

  • Task specificity, access configurations and negative findings

  • Limits on transfer, efficacy and authority

Conceptual observations.

  • Conceptually trace a benign claim from an apparent source to independent corroboration and original provenance.

  • Distinguish source credibility, claim accuracy and completeness.

Reflection prompts.

  • Is the page the original source, a report of it or a copy of a copy?

  • What independent source could contradict or limit the claim?

Accessibility alternatives.

  • A prewritten fictional example avoids live searches and exposure to harmful content.

  • Screen readers, plain language, translated material and assisted navigation are valid configurations.

  • No person, account or community needs to be investigated.

Stop the frame if.

  • Stop if the frame becomes surveillance, doxxing, scam engagement, public accusation or exposure to disturbing material.

  • Stop if repeated checking becomes compulsive or reassurance-seeking.

Escalation boundary.

  • Fraud, safeguarding, legal or security concerns belong with appropriate authorities or qualified support; do not investigate personally.

Prohibited uses.

  • Live investigation, scam testing, surveillance, profiling, doxxing or public accusation.

  • Checklist score, truth certification or claim of being scam-proof.

  • Realistic misinformation generation or persuasion experiments.

What it cannot establish.

  • One checklist, search result or AI citation cannot certify truth or make someone scam-proof.

  • It cannot establish: instantly know what is true.

  • It cannot establish: Google it and verify anything.

  • It cannot establish: become scam-proof.

  • It cannot establish: source score proves credibility.

Evidence references. [4] primary empirical support; limiting or contrary evidence; [13] primary empirical support; limiting or contrary evidence; [16] limiting or contrary evidence; official boundary context.

Next gate. No external confirmation is required for the current bounded explainer and non-prescriptive guide-frame permissions; any stronger efficacy, generalisation, protocol, promotion or independent-validation claim requires a new adjudication.

PWR-125 · Anomaly detection

Evidence route: Externally gated evidence dossier. Use the linked canonical tutorial for current teaching, accessibility, safety and editorial-review status; this evidence dossier does not enlarge that tutorial’s authority.

Summary. The rarer the target, the easier it can be to miss—even for experienced searchers.

Definition. Capability to develop or express anomaly detection in a declared context without inheriting broader claims.

What the current evidence supports. Rare anomalies are systematically easy to miss, and selected interface designs can reduce misses in bounded search tasks.

Measurement boundary. Declare target definition/prevalence, misses, false alarms, search time, fatigue, correction interface, ground truth and downstream escalation.

Myth. Never miss a hidden threat

Metric. Miss and false-alarm rates at declared target prevalence under realistic workload

Boundary. A game or lab search cannot certify professional detection, intuition or public safety.

Negative and limiting findings.

  • Millions of game trials still showed very high miss rates for ultra-rare targets.

  • Correction benefits were interface-dependent and not validated in professional field settings.

Claims this evidence cannot support.

  • spot the invisible pattern

  • never miss a threat

  • superhuman vigilance

  • anomaly score proves expert inspection

Evidence references. [9] primary empirical support; limiting or contrary evidence; [12] primary empirical support; limiting or contrary evidence.

Next gate. External confirmation is required before higher-authority use because: a clinical, high-risk or regulated configuration; a restricted safety or legitimacy boundary.

PWR-126 · Uncertainty communication

Evidence route: Evidence dossier with a non-prescriptive guide frame. Use the linked canonical tutorial for current teaching, accessibility, safety and editorial-review status; this evidence dossier does not enlarge that tutorial’s authority.

Summary. Uncertainty can be communicated honestly without automatically destroying trust—but format and audience matter.

Definition. Capability to develop or express uncertainty communication in a declared context without inheriting broader claims.

What the current evidence supports. Numeric uncertainty can often be communicated without large trust loss, while wording and cultural context materially affect reception.

Measurement boundary. Declare uncertainty type, format, source, audience, language, comprehension, trust, decision and equity; trust is not the sole success criterion.

Myth. Use the perfect phrase and everyone will understand risk

Metric. Comprehension, calibrated trust and decision quality for a declared uncertainty message

Boundary. A trusted message is not necessarily understood, accurate, fair or actionable.

Negative and limiting findings.

  • Vague verbal uncertainty could reduce trust.

  • Effects differed by country, source and topic, and comprehension/decision quality were not always established.

Claims this evidence cannot support.

  • one perfect uncertainty phrase

  • numbers make communication objective

  • trust proves understanding

  • AI can calibrate any message automatically

Bounded non-prescriptive guide frame

Purpose. Explain how uncertainty format and context affect comprehension and trust without prescribing the perfect message.

This frame supports conceptual observation and reflection only. It collects no answers, stores nothing, produces no score, and does not become a training protocol.

What this frame may discuss.

  • Evidence and measurement boundaries for uncertainty communication

  • Task specificity, access configurations and negative findings

  • Limits on transfer, efficacy and authority

Conceptual observations.

  • Conceptually compare a numeric range with vague verbal uncertainty in a benign fictional message.

  • Keep comprehension, calibrated trust, decision quality and equity as separate outcomes.

Reflection prompts.

  • What kind of uncertainty is being communicated?

  • Could a trusted message still be misunderstood, inaccurate or unfair?

Accessibility alternatives.

  • Plain language, numbers, frequencies, diagrams, audio and translation are alternative formats.

  • Cultural context and numeracy access are part of the audience configuration.

  • No live persuasion target or personal-risk disclosure is needed.

Stop the frame if.

  • Stop if the frame becomes persuasion, crisis communication, medical consent, political messaging or manipulation testing.

  • Do not test messages on people without consent and appropriate governance.

Escalation boundary.

  • Clinical, emergency, legal or public-risk communication requires qualified communicators, affected-person input and accountable review.

Prohibited uses.

  • Live persuasion experiment, audience manipulation, A/B test or optimisation target.

  • Medical, emergency, legal or political communication protocol.

  • Claims that one wording guarantees understanding or trust.

What it cannot establish.

  • A trusted message is not necessarily understood, accurate, fair or actionable.

  • It cannot establish: one perfect uncertainty phrase.

  • It cannot establish: numbers make communication objective.

  • It cannot establish: trust proves understanding.

  • It cannot establish: AI can calibrate any message automatically.

Evidence references. [5] primary empirical support; limiting or contrary evidence; [8] primary empirical support; limiting or contrary evidence; [16] limiting or contrary evidence; official boundary context.

Next gate. No external confirmation is required for the current bounded explainer and non-prescriptive guide-frame permissions; any stronger efficacy, generalisation, protocol, promotion or independent-validation claim requires a new adjudication.

PWR-127 · Belief updating

Evidence route: Externally gated evidence dossier. Use the linked canonical tutorial for current teaching, accessibility, safety and editorial-review status; this evidence dossier does not enlarge that tutorial’s authority.

Summary. Facts can shift specific beliefs, but changing a belief does not automatically change an attitude, intention or action.

Definition. Capability to develop or express belief updating in a declared context without inheriting broader claims.

What the current evidence supports. Corrections can improve accuracy for targeted false claims across countries, but effects on attitudes, intentions and lasting behavior are limited and inconsistent.

Measurement boundary. Declare claim, prior belief, evidence/source, correction, confidence, delay, attitudes, intentions, behavior and autonomy; movement toward authority is not automatically accuracy.

Myth. Change anyone’s mind with the right facts

Metric. Appropriate confidence change after verified evidence, plus delayed behavior where relevant

Boundary. Belief movement does not prove coercive success, durable action or universal agreement.

Negative and limiting findings.

  • A large vaccine-misinformation correction changed accuracy only slightly and did not change attitudes or intentions.

  • Two-week persistence in that study was uncertain.

Claims this evidence cannot support.

  • change anyone’s mind

  • correct beliefs permanently

  • facts automatically change behavior

  • belief update score identifies irrational people

Evidence references. [6] primary empirical support; limiting or contrary evidence; [7] primary empirical support; limiting or contrary evidence; [16] limiting or contrary evidence; official boundary context.

Next gate. External confirmation is required before higher-authority use because: a clinical, high-risk or regulated configuration; cultural, community-rights or affected-person authority; a restricted safety or legitimacy boundary.

PWR-128 · Epistemic humility

Evidence route: Evidence dossier with a non-prescriptive guide frame. Use the linked canonical tutorial for current teaching, accessibility, safety and editorial-review status; this evidence dossier does not enlarge that tutorial’s authority.

Summary. Intellectual humility can grow in specific contexts—but it is not self-erasure, indecision or obedience.

Definition. Capability to develop or express epistemic humility in a declared context without inheriting broader claims.

What the current evidence supports. Brief interventions and affiliated conversations can increase selected topic-specific humility measures, but broad behavior change is not established.

Measurement boundary. Declare topic, self-report versus behavior, accuracy, power relation, affiliation, dissent, delay and willingness to revise; lower confidence is not automatically humility.

Myth. Become ego-free and always open-minded

Metric. Accuracy-sensitive revision, expressed uncertainty and behavioral engagement on declared topics

Boundary. Lower confidence or polite language cannot prove wisdom, fairness or genuine openness.

Negative and limiting findings.

  • The intervention did not increase engagement with diverse viewpoints.

  • General humility and self-rated knowledge effects were inconsistent, and conversation results relied on self-report and affiliation.

Claims this evidence cannot support.

  • become ego-free

  • humility score proves wisdom

  • make opponents open-minded

  • doubt yourself more

Bounded non-prescriptive guide frame

Purpose. Explain topic-specific willingness to revise without treating lower confidence or politeness as proof of humility.

This frame supports conceptual observation and reflection only. It collects no answers, stores nothing, produces no score, and does not become a training protocol.

What this frame may discuss.

  • Evidence and measurement boundaries for epistemic humility

  • Task specificity, access configurations and negative findings

  • Limits on transfer, efficacy and authority

Conceptual observations.

  • Conceptually distinguish accurate belief revision, expressed uncertainty and behavioural engagement.

  • Notice how affiliation, dissent and power can change what a self-report means.

Reflection prompts.

  • What evidence would justify revision rather than mere deference?

  • Could polite language conceal unchanged belief or unequal power?

Accessibility alternatives.

  • A fictional benign disagreement is sufficient; personal belief disclosure is excluded.

  • Written, spoken, assisted and no-response participation are equally valid.

  • No ideology, diagnosis or personality inference is invited.

Stop the frame if.

  • Stop if the frame becomes compelled disclosure, debate pressure, belief scoring or an attempt to make someone compliant.

  • Stop if power imbalance, identity threat or distress makes reflection unsafe.

Escalation boundary.

  • High-conflict, workplace, clinical or safeguarding situations require appropriate facilitation and rights-respecting support.

Prohibited uses.

  • Humility score, personality rank, debate exercise, compliance target or progression.

  • Compelled belief disclosure, persuasion or ideological profiling.

  • Claims of ego removal, wisdom or fairness.

What it cannot establish.

  • Lower confidence or polite language cannot prove wisdom, fairness or genuine openness.

  • It cannot establish: become ego-free.

  • It cannot establish: humility score proves wisdom.

  • It cannot establish: make opponents open-minded.

  • It cannot establish: doubt yourself more.

Evidence references. [2] primary empirical support; limiting or contrary evidence; [14] primary empirical support; limiting or contrary evidence.

Next gate. No external confirmation is required for the current bounded explainer and non-prescriptive guide-frame permissions; any stronger efficacy, generalisation, protocol, promotion or independent-validation claim requires a new adjudication.

Family source register

Numbers in each dossier refer to this family register. A source can support one bounded proposition while simultaneously limiting transfer, certainty, generalisation or safety.

[1] Eric R. Stone; Jason Luu; Cory K. Costello; Annie H. Somerville (2023). Automated calibration training for forecasters. Primary research. https://doi.org/10.1002/bdm.2334

[2] Larissa Knöchelmann; Adina Janßen; Julia Dörbaum; Daniel W. Heck; J. Christopher Cohrs (2025). Enhancing Intellectual Humility About Political Topics: An Intervention Tournament Including Five Conceptual Replications. Primary research. https://doi.org/10.1002/ejsp.3177

[3] Matthew Martin; David R. Mandel (2024). Calibration Feedback With the Practical Scoring Rule Does Not Improve Calibration of Confidence. Primary research. https://doi.org/10.1002/ffo2.199

[4] Gordon Pennycook; David G. Rand (2022). Accuracy prompts are a replicable and generalizable approach for reducing the spread of misinformation. Primary research. https://doi.org/10.1038/s41467-022-30073-5

[5] Anne Marthe van der Bles; Sander van der Linden; Alexandra L. J. Freeman; David J. Spiegelhalter (2020). The effects of communicating uncertainty on public trust in facts and numbers. Primary research. https://doi.org/10.1073/pnas.1913678117

[6] Ethan Porter; Thomas J. Wood (2021). The global effectiveness of fact-checking: Evidence from simultaneous experiments in Argentina, Nigeria, South Africa, and the United Kingdom. Primary research. https://doi.org/10.1073/pnas.2104235118

[7] Ethan Porter; Yamil Velez; Thomas J. Wood (2023). Correcting COVID-19 vaccine misinformation in 10 countries. Primary research. https://doi.org/10.1098/rsos.221097

[8] John Kerr; Anne-Marthe van der Bles; Sarah Dryhurst; Claudia R. Schneider; Vivien Chopurian; Alexandra L. J. Freeman; Sander van der Linden (2023). The effects of communicating uncertainty around statistics, on public trust. Primary research. https://doi.org/10.1098/rsos.230604

[9] Mathias S. Fleck; Stephen R. Mitroff (2007). Rare Targets Are Rarely Missed in Correctable Search. Primary research. https://doi.org/10.1111/j.1467-9280.2007.02006.x

[10] Zana Buçinca; Maja Barbara Malaya; Krzysztof Z. Gajos (2021). To Trust or to Think: Cognitive Forcing Functions Can Reduce Overreliance on AI in AI-assisted Decision-making. Primary research. https://doi.org/10.1145/3449287

[11] Jon Roozenbeek; Rakoen Maertens; William McClanahan; Sander van der Linden (2020). Disentangling Item and Testing Effects in Inoculation Research on Online Misinformation: Solomon Revisited. Primary research. https://doi.org/10.1177/0013164420940378

[12] Stephen R. Mitroff; Adam T. Biggs (2013). The Ultra-Rare-Item Effect. Primary research. https://doi.org/10.1177/0956797613504221

[13] Lisa Oswald; Anastasia Kozyreva; Stefan M. Herzog; Pietro Leonardo Nickl; Ralph Hertwig (2026). Boosting Media Literacy Using Lateral Reading and Online Search Interventions. Primary research. https://doi.org/10.1177/09567976261453813

[14] Katherine R. Thorson; Lindsey A. Beck; Sarah Ketay; Keith M. Welker (2023). Increases in Intellectual Humility From Guided Conversations Are Greater When People Perceive Affiliation With Conversation Partners. Primary research. https://doi.org/10.1177/19485506231213775

[15] Melisa Basol; Jon Roozenbeek; Sander van der Linden (2020). Good News about Bad News: Gamified Inoculation Boosts Confidence and Cognitive Immunity Against Fake News. Primary research. https://doi.org/10.5334/joc.91

[16] National Institute of Standards and Technology (2023). Artificial Intelligence Risk Management Framework (AI RMF 1.0). Official authority. https://www.nist.gov/itl/ai-risk-management-framework