Download Data Mining in Bioinformatics Day 1: Classification

Survey
yes no Was this document useful for you?
   Thank you for your participation!

* Your assessment is very important for improving the work of artificial intelligence, which forms the content of this project

Document related concepts

Nonlinear dimensionality reduction wikipedia , lookup

K-nearest neighbors algorithm wikipedia , lookup

Transcript
Decision Tree
Information gain
The information gain is the loss of entropy (increase in
information) that is caused by splitting with respect to
attribute A
Gain(A) = Info(D) − InfoA(D)
(11)
We pick A such that this gain is maximised.
Karsten Borgwardt: Data Mining in Bioinformatics, Page 21