Library Header Image
LSE Research Online LSE Library Services

Simplification of genetic programs: a literature survey

Javed, Noman ORCID: 0000-0001-7770-6616, Gobet, Fernand ORCID: 0000-0002-9317-6886 and Lane, Peter (2022) Simplification of genetic programs: a literature survey. Data Mining and Knowledge Discovery, 36 (4). 1279 - 1300. ISSN 1384-5810

[img] Text (Javed2022_Article_SimplificationOfGeneticProgram) - Published Version
Available under License Creative Commons Attribution.

Download (412kB)

Identification Number: 10.1007/s10618-022-00830-7


Genetic programming (GP), a widely used evolutionary computing technique, suffers from bloat—the problem of excessive growth in individuals’ sizes. As a result, its ability to efficiently explore complex search spaces reduces. The resulting solutions are less robust and generalisable. Moreover, it is difficult to understand and explain models which contain bloat. This phenomenon is well researched, primarily from the angle of controlling bloat: instead, our focus in this paper is to review the literature from an explainability point of view, by looking at how simplification can make GP models more explainable by reducing their sizes. Simplification is a code editing technique whose primary purpose is to make GP models more explainable. However, it can offer bloat control as an additional benefit when implemented and applied with caution. Researchers have proposed several simplification techniques and adopted various strategies to implement them. We organise the literature along multiple axes to identify the relative strengths and weaknesses of simplification techniques and to identify emerging trends and areas for future exploration. We highlight design and integration challenges and propose several avenues for research. One of them is to consider simplification as a standalone operator, rather than an extension of the standard crossover or mutation operators. Its role is then more clearly complementary to other GP operators, and it can be integrated as an optional feature into an existing GP setup. Another proposed avenue is to explore the lack of utilisation of complexity measures in simplification. So far, size is the most discussed measure, with only two pieces of prior work pointing out the benefits of using time as a measure when controlling bloat.

Item Type: Article
Official URL:
Additional Information: © 2022 The Authors
Divisions: CPNSS
Subjects: B Philosophy. Psychology. Religion > B Philosophy (General)
Q Science > Q Science (General)
Date Deposited: 11 Apr 2022 13:36
Last Modified: 12 Jul 2024 22:12

Actions (login required)

View Item View Item


Downloads per month over past year

View more statistics