Title: Constraint based induction of multi-objective regression trees
Authors: Struyf, Jan
D┼żeroski, Sašo
Issue Date: 2005
Host Document: Proceedings of the Fourth International Workshop on Knowledge Discovery in Inductive Databases pages:110-121
Conference: International Workshop on Knowledge Discovery in Inductive Databases edition:4 location:Porto, Portugal date:October 3-7, 2005
Abstract: Constrained based inductive systems are a key component of inductive databases and responsible for building the models that satisfy the constraints in the inductive queries. In this paper, we propose a constraint based system for building multi-objective regression trees. A multi-objective regression tree is a decision tree capable of predicting several numeric variables at once. We focus on size and accuracy constraints. By either specifying maximum size or minimum accuracy, the user can trade-off size (and thus interpretability) for accuracy. Our approach is to first build a large tree based on the training data and to prune it in a second step to satisfy the user constraints. This has the advantage that the tree can be stored in the inductive database and used for answering inductive queries with different constraints. Besides size and accuracy constraints, we also briefly discuss syntactic constraints. We evaluate our system on a number of real world data sets and measure the size versus accuracy trade-off.
Publication status: published
KU Leuven publication type: IC
Appears in Collections:Informatics Section

Files in This Item:
File Status SizeFormat
41827.pdf Published 200KbAdobe PDFView/Open


All items in Lirias are protected by copyright, with all rights reserved.