[This article was first published on R – Win Vector LLC, and kindly contributed to R-bloggers]. (You can report issue about the content on this page here)
Want to share your content on R-bloggers? click here if you have a blog, or here if you don't.
Want to share your content on R-bloggers? click here if you have a blog, or here if you don't.
Chapter 8 “Advanced Data Preparation” of Practical Data Science with R is a study in:
- Using the
R vtreat
package for advanced data preparation. - Cross-validated data preparation.
It is the professionally edited, ready to cite version of an important data preparation methodology. An advantage being: a number of well documented result improving transforms are added to your predictive analytic work in one documented step.
We also have a number of free data preparation resources (for both the R vtreat
package and the Python vtreat
package; notice we believe in cross-language data science tools and practice):
- Free
R
video lecture on advanced variable preparation. - Free
Python
video lecture on advanced variable preparation. - Task oriented (and cross-linked) examples in
R
andPython
.
To leave a comment for the author, please follow the link and comment on their blog: R – Win Vector LLC.
R-bloggers.com offers daily e-mail updates about R news and tutorials about learning R and many other topics. Click here if you're looking to post or find an R/data-science job.
Want to share your content on R-bloggers? click here if you have a blog, or here if you don't.