A public dashboard observing signals, trends and issues.
SubscribeLogin한국어
Latest observation
2026-10-08
Public objects
4434
Build time
2026-10-08 19:44 KST
The Futures

Signal Bias Probes: A Framework for Active Multi-Group Fairness Auditing Without Model Reconstruction

Summary

This paper addresses the gap between fairness-aware machine learning training, which in practice offers only limited improvement over standard empirical risk minimization, and the resulting need for reliable post-hoc auditing of deployed models. Existing black-box auditing approaches either require reconstructing the model, exposing it to extraction attacks, or directly estimate fairness metrics while providing little insight into which regions of the data distribution drive bias. The authors note that property-specific auditing, which extracts only targeted fairness information without full model reconstruction, has remained poorly understood until now. They introduce a 'bias probe' framework enabling targeted, adaptive queries that reveal bias structure while preserving model confidentiality, built on by an active auditor called ALeBi that efficiently estimates multi-group fairness metrics. The work establishes new sample complexity guarantees governed by a property-specific complexity measure and extends the analysis to adversarial settings where a model owner might strategically obscure bias, demonstrating a fundamental trade-off between confidentiality and reliable auditing.

Classification

Main topicAI & Computing
Region menusGlobal
Impactscope:global
Time horizon0-3 years (2026-10-02)
Last updated2026-10-02 11:32 KST

Evidence 1

Part of trends 0

No objects.

Directly linked issues 0

No objects.

Public id: fm-6059579203ff