other
The Fairness Verdict Needed Its Own Error Bar
VFR-Audit measures how often a clinical model's pass-or-fail fairness decision reverses across resamples, sizes and hospitals.
Summary
VFR-Audit measures how often a clinical model's pass-or-fail fairness decision reverses across resamples, sizes and hospitals.
Clinical governance often turns a continuous fairness metric into a binary verdict, but an interval around the metric does not directly say whether that verdict will flip. VFR-Audit introduces a Verdict Flip Rate bounded between zero and 0.5, then reports resampling stability, audit-size sensitivity and cross-hospital agreement. It also tracks whether mitigation buys a stable pass at the expense of discrimination. The framework addresses reliability of audit decisions for length-of-stay prediction; it does not certify a model as clinically safe or fair.
Why it matters
VFR-Audit measures how often a clinical model's pass-or-fail fairness decision reverses across resamples, sizes and hospitals.
Limits and context
- Clinical governance often turns a continuous fairness metric into a binary verdict, but an interval around the metric does not directly say whether that verdict will flip.
- The framework addresses reliability of audit decisions for length-of-stay prediction; it does not certify a model as clinically safe or fair.
Key claims
VFR-Audit measures how often a clinical model's pass-or-fail fairness decision reverses across resamples, sizes and hospitals.
Qualification: Clinical governance often turns a continuous fairness metric into a binary verdict, but an interval around the metric does not directly say whether that verdict will flip.
Evidence: source-2026-09-01-012
Sources
- arXiv preprint 2608.30846arXiv · primary research
Corrections
No corrections have been recorded for this story.