October 2019

New Introduction to the data_algebra

October 31, 2019 | 0 Comments

We’ve had really good progress in bringing the Python data_algebra to feature parity with R rquery. In fact we are able to reproduced the New Introduction to rquery article as a “New Introduction to the data_algebra” here. The idea is: you may have good reasons to want to work in R or to want to … Continue reading New Introduction to the data_algebra [...Read more...]

Free R/datascience Extract: Evaluating a Classification Model with a Spam Filter

October 15, 2019 | 0 Comments

We are excited to share a free extract of Zumel, Mount, Practical Data Science with R, 2nd Edition, Manning 2019: Evaluating a Classification Model with a Spam Filter. This section reflects an important design decision in the book: teach model evaluation first, and as a step separate from model construction. It is funny, but it … Continue reading Free R/datascience Extract: Evaluating a Classification Model with a Spam Filter [...Read more...]

AI for Engineers

October 9, 2019 | 0 Comments

For the last year we (Nina Zumel, and myself: John Mount) have had the honor of teaching the AI200 portion of LinkedIn’s AI Academy. John Mount at the LinkedIn campus Nina Zumel designed most of the material, and John Mount has been delivering it and bringing her feedback. We’ve just started our 9th cohort. We … Continue reading AI for Engineers [...Read more...]

vtreat Cross Validation

October 6, 2019 | 0 Comments

Nina Zumel finished new documentation on how vtreat‘s cross validation works, which I want to share here. vtreat is a system that makes data preparation for machine learning a “one-liner” (available in R or available in Python). We have a set of starting off points here. These documents describe what vtreat does for you, you … Continue reading vtreat Cross Validation [...Read more...]

New vtreat Documentation (Starting with Multinomial Classification)

October 1, 2019 | 0 Comments

Nina Zumel finished some great new documentation showing how to use Python vtreat to prepare data for multinomial classification mode. And I have finally finished porting the documentation to R vtreat. So we now have good introductions on how to use vtreat to prepare data for the common tasks of: Regression: R regression example, Python … Continue reading New vtreat Documentation (Starting with Multinomial Classification) [...Read more...]