NSF PAR Search | NSF Public Access Repository

Note: When clicking on a Digital Object Identifier (DOI) number, you will be taken to an external site maintained by the publisher. Some full text articles may not yet be available without a charge during the embargo (administrative interval).
What is a DOI Number?

Some links on this page may take you to non-federal websites. Their policies may differ from this site.

The statistical fairness field guide: perspectives from social and formal sciences

https://doi.org/10.1007/s43681-022-00183-3

Carey, Alycia N.; Wu, Xintao (June 2022, AI and Ethics)

Abstract Over the past several years, a multitude of methods to measure the fairness of a machine learning model have been proposed. However, despite the growing number of publications and implementations, there is still a critical lack of literature that explains the interplay of fair machine learning with the social sciences of philosophy, sociology, and law. We hope to remedy this issue by accumulating and expounding upon the thoughts and discussions of fair machine learning produced by both social and formal (i.e., machine learning and statistics) sciences in this field guide. Specifically, in addition to giving the mathematical and algorithmic backgrounds of several popular statistics-based fair machine learning metrics used in fair machine learning, we explain the underlying philosophical and legal thoughts that support them. Furthermore, we explore several criticisms of the current approaches to fair machine learning from sociological, philosophical, and legal viewpoints. It is our hope that this field guide helps machine learning practitioners identify and remediate cases where algorithms violate human rights and values.
more » « less
Root Cause Analysis of Anomalies in Multivariate Time Series Through Granger Causal Discovery

Han, Xiao; Absar, Saima; Zhang, Lu; Yuan, Shuhan (April 2025, The Thirteenth International Conference on Learning Representations (ICLR))

Free, publicly-accessible full text available April 24, 2026
A Causal Lens for Learning Long-term Fair Policies

Lear, Jacob; Zhang, Lu (April 2025, The Thirteenth International Conference on Learning Representations (ICLR))

Free, publicly-accessible full text available April 24, 2026
A Causal Lens for Learning Long-term Fair Policies

Lear, Jacob; Zhang, Lu (April 2025, The Thirteenth International Conference on Learning Representations)

Free, publicly-accessible full text available April 24, 2026
Root Cause Analysis of Anomalies in Multivariate Time Series through Granger Causal Discovery

Han, Xiao; Absar, Saima; Zhang, Lu; Yuan, Shuhan (April 2025, The Thirteenth International Conference on Learning Representations)

Free, publicly-accessible full text available April 24, 2026
Causal Diffusion Autoencoders: Toward Counterfactual Generation via Diffusion Probabilistic Models

https://doi.org/10.3233/FAIA240780

Komanduri, Aneesh; Zhao, Chen; Chen, Feng; Wu, Xintao (October 2024, IOS Press)

Diffusion probabilistic models (DPMs) have become the state-of-the-art in high-quality image generation. However, DPMs have an arbitrary noisy latent space with no interpretable or controllable semantics. Although there has been significant research effort to improve image sample quality, there is little work on representation-controlled generation using diffusion models. Specifically, causal modeling and controllable counterfactual generation using DPMs is an underexplored area. In this work, we propose CausalDiffAE, a diffusion-based causal representation learning framework to enable counterfactual generation according to a specified causal model. Our key idea is to use an encoder to extract high-level semantically meaningful causal variables from high-dimensional data and model stochastic variation using reverse diffusion. We propose a causal encoding mechanism that maps high-dimensional data to causally related latent factors and parameterize the causal mechanisms among latent factors using neural networks. To enforce the disentanglement of causal variables, we formulate a variational objective and leverage auxiliary label information in a prior to regularize the latent space. We propose a DDIM-based counterfactual generation procedure subject to do-interventions. Finally, to address the limited label supervision scenario, we also study the application of CausalDiffAE when a part of the training data is unlabeled, which also enables granular control over the strength of interventions in generating counterfactuals during inference. We empirically show that CausalDiffAE learns a disentangled latent space and is capable of generating high-quality counterfactual images.
more » « less
Full Text Available
Achieving Counterfactual Explanation for Sequence Anomaly Detection

Cheng, He; Xu, Depeng; Yuan, Shuhan; Wu, Xintao (September 2024, Springer)

Machine Learning and Knowledge Discovery in Databases. Research Track and Demo Track - European Conference, ECML PKDD 2024, Vilnius, Lithuania, September 9-13, 2024.
more » « less
Full Text Available
Learning Causally Disentangled Representations via the Principle of Independent Causal Mechanisms

https://doi.org/10.24963/ijcai.2024/476

Komanduri, Aneesh; Wu, Yongkai; Chen, Feng; Wu, Xintao (August 2024, International Joint Conferences on Artificial Intelligence Organization)

Learning disentangled causal representations is a challenging problem that has gained significant attention recently due to its implications for extracting meaningful information for downstream tasks. In this work, we define a new notion of causal disentanglement from the perspective of independent causal mechanisms. We propose ICM-VAE, a framework for learning causally disentangled representations supervised by causally related observed labels. We model causal mechanisms using nonlinear learnable flow-based diffeomorphic functions to map noise variables to latent causal variables. Further, to promote the disentanglement of causal factors, we propose a causal disentanglement prior learned from auxiliary labels and the latent causal structure. We theoretically show the identifiability of causal factors and mechanisms up to permutation and elementwise reparameterization. We empirically demonstrate that our framework induces highly disentangled causal factors, improves interventional robustness, and is compatible with counterfactual generation.
more » « less
Full Text Available
Fair Weak-Supervised Learning: A Multiple-Instance Learning Approach

https://doi.org/10.1109/IJCNN60899.2024.10651225

Dai, Yucong; Jiang, Xiangyu; Hu, Yaowei; Zhang, Lu; Wu, Yongkai (June 2024, IEEE)

With the prevalence of machine learning in many high-stakes decision-making processes, e.g., hiring and admission, it is important to take fairness into account when practitioners design and deploy machine learning models, especially in scenarios with imperfectly labeled data. Multiple-Instance Learning (MIL) is a weakly supervised approach where instances are grouped in labeled bags, each containing several instances sharing the same label. However, current fairness-centric methods in machine learning often fall short when applied to MIL due to their reliance on instance-level labels. In this work, we introduce a Fair Multiple-Instance Learning (FMIL) framework to ensure fairness in weakly supervised learning. In particular, our method bridges the gap between bag-level and instance-level labeling by leveraging the bag labels, inferring high-confidence instance labels to improve both accuracy and fairness in MIL classifiers. Comprehensive experiments underscore that our FMIL framework substantially reduces biases in MIL without compromising accuracy.
more » « less
Full Text Available
From Identifiable Causal Representations to Controllable Counterfactual Generation: A Survey on Causal Generative Modeling

Komanduri, A; Wu, X; Wu, Y; Chen, F (May 2024, Transactions on machine learning research)

Full Text Available

« Prev Next »

Search for: All records