Many thanks go to Knewton for providing the space for this event!
Special thanks go to Yibo Chen for giving such a great workshop!
--------------------------------
This is a step by step intensive 2 hours demonstration session of model building which focus on a ongoing kaggle competition.
Yibo Chen is a data analyst with experience in model building such as response model in CRM and credit score card. Recently he is interested in Kaggle's competitions. After participating in some of these competitions, he has learned some knowledge about the data mining, and also get a score not very bad (currently 231st of 104993 data scientists).
Introduced our solution to the Amazon Employee Access Challenge.
Use stacking based on 5-fold cv for combining predictions of the base learners. The software we use is R (2.15.1) and some add-on packages including gbm, randomForest, glmnet, kernlab and Matrix.
--------------------------------
Other Useful Info Link:
The Kaggle competition The Hewlett Foundation: Short Answer Scoring and the winners' solutions. http://www.kaggle.com/c/asap-sas/details/winners
The first step in becoming a data scientist is to complete your Data Science Bootcamp Application. Just click the button to apply. It's free and will only take you about 5 minutes.