Hacker Timesnew | past | comments | ask | show | jobs | submitlogin

Crudely: combine the speed of classification trees with the generally low error rates (flexibility) of the SVM.


There is a good implementation in Mahout and it scales well on Hadoop, so its a pretty good out of the box whole data set learning algorithm.

No advantage over taking a decent statistically valid subset and running C4.5 (or a variant with floating point splits) apart from it's less work.




Consider applying for YC's Fall 2026 batch! Applications are open till July 27.

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: