Incorporating Background Knowledge in Symbolic Regression using a Computer Algebra System

01/27/2023
by   Charles Fox, et al.
0

Symbolic Regression (SR) can generate interpretable, concise expressions that fit a given dataset, allowing for more human understanding of the structure than black-box approaches. The addition of background knowledge (in the form of symbolic mathematical constraints) allows for the generation of expressions that are meaningful with respect to theory while also being consistent with data. We specifically examine the addition of constraints to traditional genetic algorithm (GA) based SR (PySR) as well as a Markov-chain Monte Carlo (MCMC) based Bayesian SR architecture (Bayesian Machine Scientist), and apply these to rediscovering adsorption equations from experimental, historical datasets. We find that, while hard constraints prevent GA and MCMC SR from searching, soft constraints can lead to improved performance both in terms of search effectiveness and model meaningfulness, with computational costs increasing by about an order-of-magnitude. If the constraints do not correlate well with the dataset or expected models, they can hinder the search of expressions. We find Bayesian SR is better these constraints (as the Bayesian prior) than by modifying the fitness function in the GA

READ FULL TEXT

page 1

page 2

page 3

page 4

research
10/20/2019

Bayesian Symbolic Regression

Interpretability is crucial for machine learning in many scenarios such ...
research
10/21/2020

Logic Guided Genetic Algorithms

We present a novel Auxiliary Truth enhanced Genetic Algorithm (GA) that ...
research
07/03/2022

Symbolic Regression is NP-hard

Symbolic regression (SR) is the task of learning a model of data in the ...
research
04/13/2023

Priors for symbolic regression

When choosing between competing symbolic models for a data set, a human ...
research
06/14/2023

Probabilistic Regular Tree Priors for Scientific Symbolic Reasoning

Symbolic Regression (SR) allows for the discovery of scientific equation...
research
03/24/2022

Explainable Artificial Intelligence for Exhaust Gas Temperature of Turbofan Engines

Data-driven modeling is an imperative tool in various industrial applica...
research
03/13/2023

Transformer-based Planning for Symbolic Regression

Symbolic regression (SR) is a challenging task in machine learning that ...

Please sign up or login with your details

Forgot password? Click here to reset