Researchers have developed a new protocol for conducting audits of machine learning models that aims to prevent providers from manipulating the evaluation process. This protocol utilizes a Private Information Retrieval (PIR) mechanism, allowing auditors to query models without the provider knowing which specific data points will be audited. The method is designed to be efficient, require minimal overhead, and not necessitate changes to the model or its inference pipeline. Theoretical guarantees and experimental results suggest that this approach significantly increases the detectability of manipulation by forcing providers to falsify a larger number of responses to hide unfairness. AI
IMPACT Enhances the trustworthiness of ML model evaluations by making manipulation more detectable.
RANK_REASON The cluster contains a research paper detailing a novel protocol for auditing machine learning models. [lever_c_demoted from research: ic=1 ai=1.0]
- alphaXiv
- arXiv
- arXivLabs
- CatalyzeX Code Finder for Papers
- CORE Recommender
- DagsHub
- Gotit.pub
- Hugging Face
- IArxiv Recommender
- Influence Flower
- private information retrieval
- ScienceCast
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →