NSF PAR Search | NSF Public Access Repository

Note: When clicking on a Digital Object Identifier (DOI) number, you will be taken to an external site maintained by the publisher. Some full text articles may not yet be available without a charge during the embargo (administrative interval).
What is a DOI Number?

Some links on this page may take you to non-federal websites. Their policies may differ from this site.

Fairness Issues and Mitigations in (Differentially Private) Socio-Demographic Data Processes

https://doi.org/10.1609/aaai.v39i27.35035

Ko, Joonhyuk; Ziani, Juba; Das, Saswat; Williams, Matt; Fioretto, Ferdinando (April 2025, Proceedings of the AAAI Conference on Artificial Intelligence)

Statistical agencies rely on sampling techniques to collect socio-demographic data crucial for policy-making and resource allocation. This paper shows that surveys of important societal relevance introduce sampling errors that unevenly impact group-level estimates, thereby compromising fairness in downstream decisions. To address these issues, this paper introduces an optimization approach modeled on real-world survey design processes, ensuring sampling costs are optimized while maintaining error margins within prescribed tolerances. Additionally, privacy-preserving methods used to determine sampling rates can further impact these fairness issues. This paper explores the impact of differential privacy on the statistics informing the sampling process, revealing a surprising effect: not only is the expected negative effect from the addition of noise for differential privacy negligible, but also this privacy noise can in fact reduce unfairness as it positively biases smaller counts. These findings are validated over an extensive analysis using datasets commonly applied in census statistics.
more » « less
Free, publicly-accessible full text available April 11, 2026
Fairness Issues and Mitigations in (Differentially Private) Socio-demographic Data Processes

Ko, Joonhyuk; Ziani, Juba; Das, Saswat; Williams, Matt; Fioretto, Ferdinando (February 2025, Proceedings of the AAAI Conference on Artificial Intelligence)

Free, publicly-accessible full text available February 28, 2026
Low-rank finetuning for LLMs: A fairness perspective

Das, Saswat; Romanelli, Marco; Tran, Cuong; Kailkhura, Bhavya; Fioretto, Ferdinando (February 2025, AAAI CoLoRA Workshop, 2025)

Free, publicly-accessible full text available February 27, 2026
Low-rank finetuning for LLMs: A fairness perspective

Das, Saswat; Romanelli, Marco; Tran, Cuong; Reza, Zarreen; Kailkhura, Bhavya; Fioretto, Ferdinando (May 2024, arXivorg)

Low-rank approximation techniques have become the de facto standard for fine-tuning Large Language Models (LLMs) due to their reduced computational and memory requirements. This paper investigates the effectiveness of these methods in capturing the shift of fine-tuning datasets from the initial pre-trained data distribution. Our findings reveal that there are cases in which low-rank fine-tuning falls short in learning such shifts. This, in turn, produces non-negligible side effects, especially when fine-tuning is adopted for toxicity mitigation in pre-trained models, or in scenarios where it is important to provide fair models. Through comprehensive empirical evidence on several models, datasets, and tasks, we show that low-rank fine-tuning inadvertently preserves undesirable biases and toxic behaviors. We also show that this extends to sequential decision-making tasks, emphasizing the need for careful evaluation to promote responsible LLMs development.
more » « less
Full Text Available
Disparate Impact on Group Accuracy of Linearization for Private Inference

Das, Saswat; Romanelli, Marco; Fioretto, Ferdinando (February 2024, Proceedings of Machine Learning Research)

Full Text Available
Finding ε and δ of Traditional Disclosure Control Systems

https://doi.org/10.1609/aaai.v38i20.30204

Das, Saswat; Zhu, Keyu; Task, Christine; Van_Hentenryck, Pascal; Fioretto, Ferdinando (March 2024, Proceedings of the AAAI Conference on Artificial Intelligence)

This paper analyzes the privacy of traditional Statistical Disclosure Control (SDC) systems under a differential privacy interpretation. SDCs, such as cell suppression and swapping, promise to safeguard the confidentiality of data and are routinely adopted in data analyses with profound societal and economic impacts. Through a formal analysis and empirical evaluation of demographic data from real households in the U.S., the paper shows that widely adopted SDC systems not only induce vastly larger privacy losses than classical differential privacy mechanisms, but, they may also come at a cost of larger accuracy and fairness.
more » « less
Full Text Available
Privacy and Bias Analysis of Disclosure Avoidance Systems

Zhu, Keyu; Fioretto, Ferdinando; Van Hentenryck, Pascal; Das, Saswat; Task, Christine (January 2023, arXivorg)

Full Text Available

Search for: All records