Ranking the Uniformity of Interval Pairs
author:Tapio Elomaa, Tampere University of Technology
published: Oct. 10, 2008, recorded: September 2008, views: 17
Related content
18:35
77 views - Frank Eichinger, Klemens Böhm, Matthias Huber, 2008
16:37
140 views - Alexander J. Smola, Alexandros Karatzoglou, Markus Weimer, 2008
33:36
35 views - Ingo Mierswa, 2008
01:36:27
10204 views - Jure Leskovec, 2008
21:44
167 views - Ajit Singh, Geoffrey J. Gordon, 2008
31:36
117 views - Serge Waterschool, 2008
12:04
111 views - Petteri Nurmi, 2005
18:48
91 views - Pauli Miettinen, 2008
16:54
247 views - Kathryn Hempstalk, Eibe Frank, Ian H. Witten, 2008
12:51
67 views - Ulf Brefeld, Tobias Scheffer, Thoralf Klein, 2008
Report a problem or upload files
If you have found a problem with this lecture or would like to send us extra material, articles, exercises, etc., please use our ticket system to describe your request and upload the data.Enter your e-mail into the 'Cc' field, and we will keep you updated with your request's status.
Description
We study the problem of finding the most uniform partition of a label distribution on an interval. This problem occurs, e.g., in discretization of continuous features, where evaluation heuristics need to find the location of the best place to split the current feature. The weighted average of empirical entropies of the interval label distributions is often used for this task. We observe that this rule is sub-optimal, because it prefers short intervals too much. Therefore, we proceed to study alternative approaches. A solution that is based on compression turns out to be the best in our empirical experiments. We also study how these alternative methods affect the performance of classification algorithms.
Link this page
Would you like to put a link to this lecture on your homepage?Go ahead! Copy the HTML snippet !




Write your own review or comment: