This file is part of IDEAS, which uses RePEc data


[ Papers | Articles | Software | Books | Chapters | Authors | Institutions | JEL Classification | NEP reports | Search | New papers by email | Author registration | Rankings | Volunteers | FAQ | Blog | Help! ]

Boosted regression (boosting): An introductory tutorial and a Stata plugin

Author info | Abstract | Publisher info | Download info | Related research | Statistics
Author Info
Matthias Schonlau (RAND)
Abstract

Boosting, or boosted regression, is a recent data-mining technique that has shown considerable success in predictive accuracy. This article gives an overview of boosting and introduces a new Stata command, boost, that im- plements the boosting algorithm described in Hastie, Tibshirani, and Friedman (2001, 322). The plugin is illustrated with a Gaussian and a logistic regression example. In the Gaussian regression example, the R2 value computed on a test dataset is R2 = 21.3% for linear regression and R2 = 93.8% for boosting. In the logistic regression example, stepwise logistic regression correctly classifies 54.1% of the observations in a test dataset versus 76.0% for boosted logistic regression. Currently, boost accommodates Gaussian (normal), logistic, and Poisson boosted regression. boost is implemented as a Windows C++ plugin. Copyright 2005 by StataCorp LP.

Download Info
To download:

If you experience problems downloading a file, check if you have the proper application to view it first. Information about this may be contained in the File-Format links below. In case of further problems read the IDEAS help page. Note that these files are not on the IDEAS site. Please be patient as the files may be large.

File URL: http://www.stata-journal.com/article.html?article=st0087
File Format:
File Function:
Download Restriction: no
File URL: http://www.stata-journal.com/software/sj5-3/st0087/
File Format: text/html
File Function:
Download Restriction: no

Publisher Info
Article provided by StataCorp LP in its journal Stata Journal.

Volume (Year): 5 (2005)
Issue (Month): 3 (September)
Pages: 330-354
Download reference. The following formats are available: HTML (with abstract), plain text (with abstract), BibTeX, RIS (EndNote, RefMan, ProCite), ReDIF
Handle: RePEc:tsj:stataj:v:5:y:2005:i:3:p:330-354

Contact details of provider:
Web page: http://www.stata-journal.com/

Order Information:
Web: http://www.stata-journal.com/subscription.html

For technical questions regarding this item, or to correct its listing, contact: (Christopher F. Baum).

Related research
Keywords: boost; boosted regression; boosting; data mining;

Statistics
Access and download statistics

Did you know? You too can volunteer with RePEc.

This page was last updated on 2009-12-20.


This information is provided to you by IDEAS at the Department of Economics, College of Liberal Arts and Sciences, University of Connecticut using RePEc data on a server sponsored by the Society for Economic Dynamics.