CoRE Lab Berkeley School of Education

How to be 'Choosy': Wrangling Big Datasets

“How to be Choosy” provides pedagogical strategies, technical methods, and interactive Jupyter Notebooks and CODAP templates to help educators and curriculum designers make large datasets manageable for classroom instruction. The collection accompanies the Teaching Statistics publication by Wilkerson et al. (2025).

Intended Audience & Use Case: Aimed at middle school, high school, and undergraduate data science and statistics instructors, as well as educational content developers. It serves as a guide for selecting, filtering, and preparing authentic large-scale datasets for teaching.

Available Resources

Previous post
WDS: Exploration Units Collection
Next post
MoDa: Models and Real-World Data