This course is designed for data scientists, machine learning practitioners, and researchers who want to understand how resampling techniques must be adapted to the structure of the problem at hand. You will learn how standard validation methods such as cross-validation can fail when applied blindly, and how to design problem-dependent resampling strategies for spatial data, pair-input data, and other dependent observation structures. The course also covers spatial cross-validation, dependency-aware evaluation design, and statistical testing methods to assess whether performance estimates are reliable. By the end of the course, you will be able to choose and construct appropriate resampling strategies that reflect the true structure of your data and provide trustworthy performance estimates.
What you'll learn
understand how standard validation methods like cross-validation can fail
design problem-dependent resampling strategies
apply spatial cross-validation and dependency-aware evaluation methods
conduct statistical tests to assess performance estimates