Skip to content

PAN Lab levers

Lever

Escalate checks

When monitoring flags trouble, the organization raises how closely its people check the system's output. It does not wait for the next scheduled review. The extra checking is temporary, and it comes when the monitoring says it is needed.

What it is

Monitoring that nobody acts on is only a record of what went wrong. Escalation gives each alert a consequence: more reviewers on the flagged work, a second read of every affected case, or a lower bar for sending a case to a supervisor. The documented cases cited below show why the extra checking helps. Reviewers caught what the model missed because they knew things about the case that the model did not.

What it pushes on in the Lab

In the Lab, this lever raises the checking and correcting that the people using the system can do. It also weakens the pathway by which people adopt the automated system's failed output.

increasedPeople or agents using it
dampenedFailures adopted by people or agents

You can also pull this lever at a strong tier, which costs more. At the strong tier, each effect below the Lab's strongest setting pushes harder.

In the modes that offer aiming, you can aim this lever at particular parts and pathways of a network. Otherwise it applies to the whole network.

Its pattern in the Practice Library

The Practice Library describes the pattern behind this lever:State-feedback vigilance

The pressures it answers

These pressures list this lever among the levers that answer them:

A lever answers a pressure when it pushes the other way on something the pressure pushes on.

Where you can pull it

Networks in the Lab that offer this lever:144

Every network that offers it

The evidence behind its effects

The Lab cites these claims from the evidence registry for this lever's effects.

In the documented AFST evaluation, screener overrides of the tool — roughly a third of its recommendations — cut screen-in disparity from about 20% to 9% relative to the tool acting alone.[4]

goldhaberfiebertprince2019GroundingGovernment evaluationSave

Goldhaber-Fiebert & Prince (Stanford), Impact evaluation summary: Allegheny Family Screening Tool (Allegheny County DHS, April 2019) https://analytics.alleghenycounty.us/wp-content/uploads/2019/05/Impact-Evaluation-Summary-from-16-ACDHS-26_PredictiveRisk_Package_050119_FINAL-5.pdf

https://analytics.alleghenycounty.us/wp-content/uploads/2019/05/Impact-Evaluation-Summary-from-16-ACDHS-26_PredictiveRisk_Package_050119_FINAL-5.pdf

Appears in: PAN framework development

Grounds: capability governance: at-node control; model org: allegheny_afst

Topics: child-welfare

stapletonetal2022AcademicPeer-reviewedSave

Stapleton, L., Lee, M. H., Qing, D., Wright, M., Chouldechova, A., Holstein, K., Wu, Z. S., & Zhu, H. (2022). Imagining new futures beyond predictive systems in child welfare: A qualitative study with impacted stakeholders. 2022 ACM Conference on Fairness Accountability and Transparency, 1162–1177. https://doi.org/10.1145/3531146.3533177

doi.org/10.1145/3531146.3533177

Appears in: PAN framework development; Paramerge authored research

Grounds: deployment audit: Allegheny AFST

Topics: algorithmic-fairness, child-welfare

In the documented MiDAS case, error among no-review auto-adjudications ran roughly 93%, and determinations erred at about 85% without human review versus 44% with it.[4]

aiincidentdatabaseGroundingInvestigativeSave

AI Incident Database, Incident 373 (MiDAS false fraud claims) https://incidentdatabase.ai/cite/373/

https://incidentdatabase.ai/cite/373/

Grounds: model org: michigan_midas

In contextual inquiries with Allegheny AFST call screeners, workers calibrated reliance using contextual case knowledge unavailable to the model and reliably detected and overrode erroneous risk scores — complementary human information, not generic distrust, was the safeguard's mechanism.[2]

kawakami2022AcademicSave

Kawakami, A., Sivaraman, V., Cheng, H.-F., Stapleton, L., Cheng, Y., Qing, D., Perer, A., Wu, Z. S., Zhu, H., & Holstein, K. (2022). Improving Human-AI Partnerships in Child Welfare: Understanding Worker Practices, Challenges, and Desires for Algorithmic Decision Support. In CHI Conference on Human Factors in Computing Systems (CHI '22). ACM. https://doi.org/10.1145/3491102.3517439

doi.org/10.1145/3491102.3517439

Appears in: PAN framework development

Topics: algorithmic-fairness, child-welfare, human-ai-interaction

dearteaga2020AcademicSave

De-Arteaga, M., Fogliato, R., & Chouldechova, A. (2020). A Case for Humans-in-the-Loop: Decisions in the Presence of Erroneous Algorithmic Scores. In CHI Conference on Human Factors in Computing Systems (CHI 2020). ACM. https://doi.org/10.1145/3313831.3376638

doi.org/10.1145/3313831.3376638

Appears in: Evidence reverification (2026)

Topics: algorithmic-fairness, child-welfare, human-ai-interaction

Pull this lever in the PAN Lab and watch which way it pushes the network.

Open the PAN Lab