PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
May 21, 20260 citationsOpen Access

Why Aggregate Accuracy is Inadequate for Evaluating Fairness in Law Enforcement Facial Recognition Systems

View Full Paper
KAKhalid Adnan Alsayed

Key Points

  • The aim is to show that aggregate accuracy does not effectively assess fairness in facial recognition systems used by law enforcement.
  • Analyzed subgroup-level error distributions, focusing on false positive and false negative rates.
  • Reviewed literature on demographic performance disparities in facial recognition.
  • Discussed model-agnostic fairness auditing methods for assessing deployed systems.
  • Demonstrated that overall performance metrics obscure significant disparities in error rates across demographic groups.
  • Highlighted the operational risks related to an accuracy-centric evaluation in law enforcement settings.
  • Emphasized the need for comprehensive fairness-aware evaluation strategies in AI systems.

Abstract

Facial recognition systems are increasingly deployed in law enforcement and security contexts, where algorithmic decisions can carry significant societal consequences. Despite high reported accuracy, growing evidence demonstrates that such systems often exhibit uneven performance across demographic groups, leading to disproportionate error rates and potential harm. This paper argues that aggregate accuracy is an insufficient metric for evaluating the fairness and reliability of facial recognition systems for high-stakes environments. Through analysis of subgroup-level error distribution, including false positive and false negative rates, we demonstrate how overall performance metrics can obscure critical disparities across demographic groups. Drawing on existing literature and empirical observations from classification-based systems, the paper highlights the operational risks associated with accuracy-centric evaluation practices, particularly in law enforcement applications where misclassification may result in wrongful suspicion or missed identification. We further discuss the importance of model-agnostic fairness auditing approaches that enable post-deployment evaluation without access to proprietary systems. Finally, the paper outlines the inherent trade-offs between fairness and accuracy and emphasizes the need for more comprehensive fairness-aware evaluation strategies in high-stakes AI systems.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Khalid Adnan Alsayed (2026) studied this question.

synapsesocial.com/papers/6a0ea127be05d6e3efb5f885https://doi.org/10.5281/zenodo.20298396
Ask AI
Helpful
Bookmark
Share
View Full Paper