arXiv Machine Learning

Not All Synthetic Data Are Equal: Expert-Committee Audit Screening for Imbalanced Crash-Injury-Severity Prediction in Automated Driving Systems

The paper introduces Expert-Committee Audit Screening (ECAS), a framework that evaluates the credibility of synthetic minority samples for predicting crash injury severity in automated driving systems. Using real incident data from the NHTSA, ECAS filters generated samples based on label support, boundary separation, committee agreement, and local plausibility, then selects accepted samples via within‑class percentile normalization and Pareto non‑dominated sorting. The best ECAS configuration, combined with normalizing flow augmentation and a TabPFN classifier, outperformed other evidence settings in balanced accuracy, macro‑F1, and minor‑injury recall, and analysis showed ECAS‑accepted samples were better supported by nearby real crashes.

arXiv AI
Sep 23

Toward Auditable and Calibrated AI for Dementia-Related Crash Severity Prediction: A Selective Deferral Framework to Support Human Review

The paper presents a decision‑aware framework for predicting dementia‑related crash severity that emphasizes auditability and selective deferral. Using 4,781 Texas crash records, the authors evaluate several models—including structured, narrative, fusion, calibrated fusion, BERT‑family, and local large‑language‑model baselines—under a stratified 70/15/15 split. The leakage‑controlled Gemma model achieves the highest macro‑F1 of 0.545, while a calibrated fusion model reaches 0.522 macro‑F1 with an expected calibration error of 0.033; selective deferral further improves performance, raising macro‑F1 to 0.573 at 70% coverage and reducing severity cost to 0.577.

By Gaurab Chhetri, Anika Baitullah, Subasish Das
arXiv AI
Sep 2

Causal Evidentiary Governance for High-Risk Machine Learning Systems

The paper proposes Causal Evidentiary Governance (CEG), a framework that requires regulated institutions to maintain a versioned directed acyclic graph (DAG) separating allowable from disallowed causal pathways in high‑risk machine learning systems. CEG introduces the Causal Harm Rate to quantify prediction variation due to disallowed pathways and pairs each decision with a signed Decision‑Evidence Packet (DEP) that cryptographically links the prediction to the DAG and path‑specific attributions, enabling efficient inclusion proofs via a Merkle tree. Empirical validation on synthetic credit data and the German Credit dataset demonstrates that CEG more clearly isolates causal effects than traditional fairness metrics and that a proof‑of‑concept implementation shows operational feasibility with manageable performance tradeoffs.

By Samah Kareem, Bar{\i}\c{s} \c{C}elikta\c{s}
arXiv Machine Learning
Jul 30

Cost-Sensitive Conformal Prediction and Human-in-the-Loop Abstention for Imbalanced High-Stakes Decision Support: A Multi-Domain Benchmark

arXiv:2607. 27143v1 Announce Type: new Abstract: High-stakes decision systems in credit scoring, fraud detection, healthcare, and industrial safety require reliable uncertainty quantification under severe class imbalance and asymmetric error costs.

By Manpreet Singh, Akshatha Srikantha, Shyamal Lakhanpal
arXiv Machine Learning
Aug 19

Proactive Road Safety Intervention in Australia: Predicting Risky Driving Hotspots from Connected Vehicle Data

The paper proposes a proactive approach to road safety in Greater Sydney by using connected vehicle telemetry to predict risky driving events before crashes occur. It quantifies risky driving with g‑force thresholds and builds spatio‑temporal heatmaps to locate high‑risk zones. Eight predictive models were compared, with ARIMA achieving the lowest error and showing that simple time‑series methods can rival deep learning when data are limited, highlighting the value of IoT data for targeted safety interventions.

By Adriana-Simona Mih\u{a}i\c{t}\u{a}, Clarence Cheung, Artur Grigorev, Tuo Mao, David Lillo-Trynes
arXiv Computer Vision
Aug 21

CAViAR: A Causal Video Dataset for Fine-Grained Accident Reasoning in Real-World Scenarios

arXiv:2608. 19380v1 Announce Type: new Abstract: While modern autonomous driving systems excel at perception tasks such as object detection and trajectory prediction, they lack the high-level causal reasoning required to interpret traffic accidents.

By Sparsh Garg, Yi-Wen Chen, Vijay Kumar B G, Abhishek Aich
arXiv AI
Aug 25

EG-ARSA: An Expert-Grounded Open Model for Visual Road Safety Auditing in Low-Resource Settings

EG-ARSA introduces an Expert‑Grounded Distillation (EGD) framework that transfers institutional road safety expertise into a compact vision‑language model for visual road safety auditing. The method calibrates a teacher model against authoritative field audits, achieving a Cohen’s kappa of 0.74 before generating structured supervision for an 8‑billion‑parameter student model via Low‑Rank Adaptation. The authors also release Bangladesh Road Safety Audit (BD‑ARSA), an open dataset of 21,947 image‑audit records, and demonstrate that the student model outperforms both its larger teacher and Gemini‑2.5‑Flash in ordinal risk assessment and expert evaluation.

By Md Thamed Bin Zaman Chowdhury, Moazzem Hossain