Incorporating background knowledge in symbolic regression using a computer algebra system

Fox, Charles; Tran, Neil_D (ORCID:0000000190331759); Nacion, F_Nikki; Sharlin, Samiha (ORCID:0000000263799206); Josephson, Tyler_R (ORCID:0000000201000227)

doi:10.1088/2632-2153/ad4a1e

Citation Details

Incorporating background knowledge in symbolic regression using a computer algebra system

Abstract Symbolic regression (SR) can generate interpretable, concise expressions that fit a given dataset, allowing for more human understanding of the structure than black-box approaches. The addition of background knowledge (in the form of symbolic mathematical constraints) allows for the generation of expressions that are meaningful with respect to theory while also being consistent with data. We specifically examine the addition of constraints to traditional genetic algorithm (GA) based SR (PySR) as well as a Markov-chain Monte Carlo (MCMC) based Bayesian SR architecture (Bayesian Machine Scientist), and apply these to rediscovering adsorption equations from experimental, historical datasets. We find that, while hard constraints prevent GA and MCMC SR from searching, soft constraints can lead to improved performance both in terms of search effectiveness and model meaningfulness, with computational costs increasing by about an order of magnitude. If the constraints do not correlate well with the dataset or expected models, they can hinder the search of expressions. We find incorporating these constraints in Bayesian SR (as the Bayesian prior) is better than by modifying the fitness function in the GA. more »

Award ID(s):: 2138938

PAR ID:: 10511871

Author(s) / Creator(s):: Fox, Charles; Tran, Neil_D; Nacion, F_Nikki; Sharlin, Samiha; Josephson, Tyler_R

Publisher / Repository:: IOP Publishing

Date Published:: 2024-06-03

Journal Name:: Machine Learning: Science and Technology

Volume:: 5

Issue:: 2

ISSN:: 2632-2153

Format(s):: Medium: X Size: Article No. 025057

Size(s):: Article No. 025057

Sponsoring Org:: National Science Foundation

Journal Article:
https://doi.org/10.1088/2632-2153/ad4a1e

More Like this