CoRE Lab Berkeley School of Education

Resources for Teaching and Design

Explore our collection of discussion guides, datasets, lesson slides, and interactive notebook modules designed to integrate computational thinking and data investigation into classrooms. Search or filter below by project, target level, and resource type.

Writing Data Stories: Integrating Data into Middle School Science
Writing Data Stories: Integrating Data into Middle School Science

“Writing Data Stories” is an NSF-funded research project that integrates computational data analysis into middle school science classrooms. The project teaches students to construct “syncretic data stories”—multimodal projects that blend academic statistical analysis of scientific datasets with personal narrative and social reflection.

Show Your Work Computational Notebooks for Educators
Show Your Work! Computational Notebooks for Educators

“Show Your Work!” is a suite of free, web-based introductory Jupyter Notebooks designed for K-12 educators and curriculum designers with little to no prior programming experience. Built on learning sciences principles, the project introduces computational notebooks as epistemic tools, letting teachers experience firsthand what it feels like to conduct notebook-based computational data investigations in specific subject domains. Each notebook introduces a “bundle” of data and statistics ideas, code, and content that focus on specific kinds of subject area analyses such as time series analysis, descriptive statistics, spatial analysis, and more.

Rivulet: Python Notebook Tools for Educational Data Retrieval
Rivulet: Python Notebook Tools for Educational Data Retrieval

The Rivulet project provides teachers, curriculum designers, and educational researchers with Python notebooks and CODAP plugins that streamline querying and fetching environmental datasets from public API endpoints. These notebooks act as automated wrangling pipelines, translating complex database APIs into clean, structured data that highlight “signature” scientific patterns.

MoDa: Models and Real-World Data
MoDa: Models and Real-World Data

The MoDa (Modeling + Data) project combines agent-based simulation, domain-specific block-based programming, and data visualization tools to support students’ reasoning about complex environmental phenomena. By building and running simulations and then comparing them to real-world datasets, students can test, validate, and refine their scientific theories.

How to be Choosy : Wrangling Big Datasets
How to be 'Choosy': Wrangling Big Datasets

“How to be Choosy” provides a framework for reducing the size of very large datasets (with too many rows of columns) according to pedagogical goals. It describes technical methods for reducing dataset size using CODAP or computational notebooks (Jupyter or CoLab), provides interactive templates and practice activites, and offers a DIY guide to help educators and curriculum designers make large datasets manageable for classroom instruction.

WDS: Exploration Units Collection
WDS: Exploration Units Collection

Writing Data Stories Exploration Units are comprehensive, 2-3 week long curriculum units designed to build deep data literacy and statistical inquiry skills in middle school classrooms. Centered around authentic socioscientific datasets, these units provide bilingual (English/Spanish) student handouts, slides, and teacher lesson guides.

WDS: Data Launchpads Collection
WDS: Data Launchpads Collection

Data Launchpads are highly interactive documents built inside the Common Online Data Analysis Platform (CODAP) using the Story Builder plugin. Launchpads act as scaffolded “on-ramps” for students exploring complex public datasets. Each launchpad features built-in background information, multimedia context-setters, data activators, and guided tutorials on graphs, maps, and filtering.

WDS: DataBytes Collection
WDS: DataBytes Collection

DataBytes are quick, bite-sized classroom activities (designed to take 30 minutes or less) that encourage students to interpret and analyze data visualizations related to everyday scientific issues. The resources include several specific lessons focused on visualizations sourced from news media and scientific agency reports, as well as a general framework and DIY guide for teachers to create their own DataBytes activities.

WDS: Yellowstone Cascade Launchpad
WDS: Yellowstone Cascade Launchpad

This dedicated data launchpad provides a curated ecological dataset focusing on Yellowstone National Park’s trophic levels. It allows students to investigate how the reintroduction of gray wolves triggered a trophic cascade, impacting elk populations, willow and aspen growth, and beaver dams.

WDS: Spotify Billboard Hot 100 Launchpad
WDS: Spotify/Billboard Hot 100 Launchpad

This launchpad integrates Billboard Hot 100 chart histories (1958–2021) with Spotify’s audio analysis parameters (such as energy, tempo, and danceability). It provides a high-interest dataset for students to examine how popular music has changed over time.

WDS: Emerald Lake Aquatic Ecosystem Launchpad
WDS: Emerald Lake Aquatic Ecosystem Launchpad

This launchpad explores high-altitude lake water chemistry and ecosystem indicators using long-term monitoring data from Emerald Lake in the Sierra Nevada. It allows students to examine temperature, pH, and nutrient variables over several decades.

WDS: COVID-19 Dataset Launchpad
WDS: COVID-19 Dataset Launchpad

This launchpad provides students with structured epidemiological data tracking cases, vaccination rates, and mortality statistics from the COVID-19 pandemic. It introduces key concepts in public health data representation and disease transmission modeling.

WDS: CalEnviroScreen Data Launchpad
WDS: CalEnviroScreen Data Launchpad

This launchpad introduces students to California’s CalEnviroScreen dataset to explore pollution. It allows students to map and analyze various health, environmental, and demographic metrics across different California census tracts.

SyW: Spatial Analyses Mapping Module
SyW: Spatial Analyses Mapping Module

This module introduces educators to GIS and spatial mapping. It guides teachers through plotting geographic coordinates, layering data points, and analyzing location-based datasets.

SyW: Center and Spread Statistics Module
SyW: Center and Spread Statistics Module

This module walks educators through basic statistical analysis of datasets, illustrating concepts like mean, median, and data spread. It provides interactive visualizations to help teachers explore mathematical distribution shapes and trends.

SyW: Time Series Analysis Module
SyW: Time Series Analysis Module

This module focuses on longitudinal data and time series analysis, showing educators how to plot chronological data and measure change over time. It provides code examples for analyzing historical sensor or environmental data.

SyW: Intro to Jupyter Notebooks Module
SyW: Intro to Jupyter Notebooks Module

This introductory module guides educators through the basic mechanics of Jupyter Notebooks. It covers how code cells are executed, how comments are read, and how variables are modified, serving as a gentle entry point into coding in Python and R.

Rivulet: USGS EPA WQX Water Quality Notebook
Rivulet: USGS/EPA WQX Water Quality Notebook

This Jupyter notebook connects to the Water Quality Portal (WQP) to query the EPA and USGS Water Quality Exchange (WQX) database. It automates the extraction of water chemistry and ecological indicators—such as dissolved oxygen, salinity, pH, nitrates, heavy metals, and bacterial levels—for local monitoring sites.

Rivulet Next: Interactive CODAP Data Plugins
Rivulet Next: Interactive CODAP Data Plugins

Rivulet Next is an experimental collection of interactive browser plugins designed for the Common Online Data Analysis Platform (CODAP). These plugins bypass the need for python programming or server hosting by allowing users to directly query APIs from the CODAP interface to fetch, map, and analyze real-time environmental data.

Rivulet: NOAA CoastWatch Ocean Data Notebook
Rivulet: NOAA CoastWatch Ocean Data Notebook

This Jupyter notebook queries oceanographic datasets hosted on the National Oceanic and Atmospheric Administration’s (NOAA) CoastWatch ERDDAP servers. It streamlines the retrieval of sea surface temperatures (SST), sea level height deviations, and thermal expansion telemetry across custom coordinates and historical time spans.

Rivulet: EPA AQS Data Retrieval Notebook
Rivulet: EPA AQS Data Retrieval Notebook

This Jupyter notebook connects directly to the US Environmental Protection Agency’s Air Quality System (AQS) API. It allows users to programmatically extract air quality telemetry—including particulate matter (PM2.5 and PM10), ozone, nitrogen dioxide, and overall Air Quality Index (AQI) values—for any US county or monitoring station across custom time ranges.

Rivulet: NASA POWER Albedo Notebook
Rivulet: NASA POWER Albedo Notebook

This Jupyter notebook connects to the NASA POWER API to fetch surface albedo levels anywhere on the globe. It automates the extraction of these data and allows for easy comparison of different sites.

Choosy: EPA Toxic Release Inventory Wrangling Demo
Choosy: EPA Toxic Release Inventory Wrangling Demo

This interactive data wrangling resource features the EPA Toxic Release Inventory (TRI) database for the state of California to illustrate data reduction strategies. It provides structured examples to help educators find the right sub-selections of records and variables from the full dataset, while still addressing key learning objectives.

Choosy: How to be Choosy Quick Reference Guide
Choosy: How to be 'Choosy' Quick Reference Guide

This single-page reference guide describes six pedagogical strategies for preparing large datasets for classroom use: three row or case-reduction strategies (Random, Purposeful, and Build-Your-Own) and three column or attribute-reduction strategies (Thematic, Mathematical, and Question-Driven). For each strategy, the guide defines the method, evaluates its pedagogical benefits and drawbacks (especially in terms of statistical and data science learning goals), and maps it to GAISE II and IDSSP educational standards.

Choosy: Billboard Hot 100 Data Wrangling Demo
Choosy: Billboard Hot 100 Data Wrangling Demo

This interactive data wrangling resource features a curated Billboard Hot 100 dataset along with guided workflows demonstrating how to filter, sort, and slice long, case-heavy datasets (with too many records) to fit specific lesson objectives. It helps educators find the right sub-selections of music charts from the full database for the kinds of investigations they want to do.