Using Text-Based Causal Inference to Disentangle Factors Influencing Online Review Ratings

Li, Linsen; Culotta, Aron; Mattei, Nicholas

doi:10.18653/v1/2025.naacl-long.562

Citation Details

Using Text-Based Causal Inference to Disentangle Factors Influencing Online Review Ratings

Online reviews provide valuable insights into the perceived quality of facets of a product or service. While aspect-based sentiment analysis has focused on extracting these facets from reviews, there is less work understanding the impact of each aspect on overall perception. This is particularly challenging given correlations among aspects, making it difficult to isolate the effects of each. This paper introduces a methodology based on recent advances in text-based causal analysis, specifically CausalBERT, to disentangle the effect of each factor on overall review ratings. We enhance CausalBERT with three key improvements: temperature scaling for better calibrated treatment assignment estimates; hyperparameter optimization to reduce confound overadjustment; and interpretability methods to characterize discovered confounds. In this work, we treat the textual mentions in reviews as proxies for real-world attributes. We validate our approach on real and semi-synthetic data from over 600K reviews of U.S. K-12 schools. We find that the proposed enhancements result in more reliable estimates, and that perception of school administration and performance on benchmarks are significant drivers of overall school ratings. more »

Award ID(s):: 2333537 2107505

PAR ID:: 10629571

Author(s) / Creator(s):: Li, Linsen; Culotta, Aron; Mattei, Nicholas

Publisher / Repository:: Association for Computational Linguistics

Date Published:: 2025-01-01

Page Range / eLocation ID:: 11259 to 11277

Format(s):: Medium: X

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
Accepted Manuscript1.0
Conference Paper:
https://doi.org/10.18653/v1/2025.naacl-long.562

More Like this