Causal Evidentiary Governance for High-Risk Machine Learning Systems
Samah Kareem, Barış Çeliktaş
Abstract
Machine learning systems deployed for credit, hiring, and resource distribution are increasingly subject to regulatory oversight from policies such as the EU AI Act and GDPR. Current fairness governance practices rely on observational fairness metrics, post-hoc explainability, and immutable audit logs, but provide limited support for causal attribution and efficient evidentiary verification. We introduce Causal Evidentiary Governance (CEG), a framework in which regulated institutions commit to a versioned directed acyclic graph (DAG) that partitions causal pathways into allowable and disallowed groups. The Causal Harm Rate measures prediction variation attributable to disallowed causal pathways. Each decision is accompanied by a signed Decision-Evidence Packet (DEP), cryptographically binding the prediction to a digest of the published DAG and path-specific attributions. DEP digests can be appended to a Merkle tree to enable logarithmic-cost inclusion proofs. We validate CEG through a two-layer empirical methodology using demographic summaries from four years of PMA credit supervisory data to construct 10,000 synthetic credit applicants across four strategic DAG counterfactuals. Causal Harm Rate isolates injected causal effects more clearly than demographic parity or equalized odds. Cross-model validation and ablation studies assess robustness. Evaluation on the German Credit dataset shows that harm associated with specific causal pathways can be substantially understated by associational fairness metrics. Finally, a proof-of-concept implementation demonstrates operationally plausible throughput and highlights relevant performance tradeoffs.
Create a lesson
Related papers
Rights by Architecture: A Human-Compatible Sociotechnical Layer for Digital Protection Across Regulatory Regimes
Soheil Human
Addressing Trust in AI Systems through Education: A Didactic Perspective
Pierre Haritz, Hendrik Krone, Thomas Liebig
Meeting the Coming Wave: The Emerging Politics of AI and Work across 33 Parliaments
Juliana Chueri, Petter Törnberg
Fairness-Aware Multimodal Transformer Modeling for Real-Time Student Attention Estimation
Christoforos Fragkiadakis, Seyed Sahand Mohammadi Ziabari, Ali Mohammed Mansoor Alsahag
Privacy Washing: Detecting Internal Contradictions in Privacy Policies
Thomas Brackin
Accurate in space, unreliable in time: how LLMs represent national cultural change
Yalda Daryani, Miranda Bogen, Madeleine I. G. Daepp