Empirical Comparisons of Learning Methods & Case Studies

author:Rich Caruana, Cornell University
published: Feb. 25, 2007,   recorded: May 2005,   views: 120
Categories
You might be experiencing some problems with Your Video player.

Related content

Visitors who watched this lecture also watched...
01:16:36
Which Supervised Learning Method Works Best for What? An Empirical Comparison of Learning Methods and Metrics

857 views - Rich Caruana, 2006
02:07:12
Boosting

3938 views - Robert Schapire, 2005
01:35:26
KDD - CUP 2004

50 views - Rich Caruana, 1970
01:33:19
Model Compression

100 views - Rich Caruana, 1970
03:54:31
Support Vector Machines

12685 views - Chih-Jen Lin, 2006
01:04:59
Evidence Integration in Bioinformatics

98 views - Phil Long, 2005
01:42:36
Generalization bounds

153 views - John Langford, 2005
01:02:16
Spooky Stuff in Metric Space

168 views - Rich Caruana, 1970
32:08
Model Compression: Bagging your Cake and Eating it too (part 2)

55 views - Rich Caruana, 2007
04:58:57
Reinforcement Learning

1393 views - Satinder Singh, 2006

Report a problem or upload files

If you have found a problem with this lecture or would like to send us extra material, articles, exercises, etc., please use our ticket system to describe your request and upload the data.
Enter your e-mail into the 'Cc' field, and we will keep you updated with your request's status.
Lecture popularity: You need to login to cast your vote.

 Watch videos:   (click on thumbnail to launch)

Watch Part 1
Part 1 1:25:08 Flash video Windows Media video
!NOW PLAYING
Watch Part 2
Part 2 0:36:16 Flash video Windows Media video

Description

Decision trees may be intelligible, but can they cut the mustard? Have SVMs replaced neural nets, or are neural nets still best for regression, and SVMs best for classification? Boosting maximizes a margin much like SVMs, but can boosting compete with SVMs? And is it better to boost weak models, as theory suggests, or to boost stronger models? Bagging is much easier than boosting, so how well does bagging stack up against boosting? Bagging is supposed to be best with low bias high variance methods like decision trees, so if we bag lower variance models like neural nets are they as good as bagged trees? What happens if we do bagging with steroids, i.e. switch to random forests? And what about old friends like k-nearest neighbor — should they just be put out to pasture? In this lecture I'll compare the performance of a variety of popular machine learning methods on nine performance criteria: Accuracy, F-score, Lift, Precision/Recall Break-Even? Point, Area under the ROC, Average Precision, Squared Error, Cross-Entropy?, and Probabilistic Calibration. I'll show that while no one learning method does it all, it is possible to "repair" some of them so that they do well on all metrics. I'll then describe NACHOS, a new ensemble method that does even better by by building on top of these other learning methods. Finally, I'll discuss how the nine performance metrics relate to each other, and look at a few case-studies to show why it is important to use the right metric for each problem.

Link this page  

Would you like to put a link to this lecture on your homepage?
Go ahead! Copy the HTML snippet !

Write your own review or comment:

make sure you have javascript enabled or clear this field: