Attention:The NSF Public Access Repository (PAR) system and access will be unavailable from 11:00 PM ET on Thursday, August 13 until 12:00 AM ET on Friday, August 14 due to maintenance. We apologize for the inconvenience.


Title: Getting Messy with Authentic Data: Exploring the Potential of Using Data from Scientific Research to Support Student Data Literacy
Data are becoming increasingly important in science and society, and thus data literacy is a vital asset to students as they prepare for careers in and outside science, technology, engineering, and mathematics and go on to lead productive lives. In this paper, we discuss why the strongest learning experiences surrounding data literacy may arise when students are given opportunities to work with authentic data from scientific research. First, we explore the overlap between the fields of quantitative reasoning, data science, and data literacy, specifically focusing on how data literacy results from practicing quantitative reasoning and data science in the context of authentic data. Next, we identify and describe features that influence the complexity of authentic data sets (selection, curation, scope, size, and messiness) and implications for data-literacy instruction. Finally, we discuss areas for future research with the aim of identifying the impact that authentic data may have on student learning. These include defining desired learning outcomes surrounding data use in the classroom and identification of teaching best practices when using data in the classroom to develop students’ data-literacy abilities.  more » « less
Award ID(s):
1832042 1637653 1027253
PAR ID:
10112628
Author(s) / Creator(s):
; ;
Date Published:
Journal Name:
CBE—Life Sciences Education
Volume:
18
Issue:
2
ISSN:
1931-7913
Page Range / eLocation ID:
es2
Format(s):
Medium: X
Sponsoring Org:
National Science Foundation
More Like this
  1. Data literacy is a critical skill for developing the next generation of scientists and informed citizens. Authentic engagement with real-world Earth science datasets in K–12 classrooms is often limited by the complexity of research data and the need for effective pedagogical strategies to support sensemaking. Data Puzzles, a collaborative effort between scientists, curriculum developers, and educators housed at the Cooperative Institute for Research in Environmental Sciences (CIRES), addresses this challenge by transforming authentic Earth and climate science datasets into structured, instructional modules designed for secondary classrooms. Each Data Puzzle is anchored in a relevant Earth system phenomenon and pairs authentic datasets (e.g., from satellite observations, field sensors, and physical samples) with research-based instructional practices from Ambitious Science Teaching (AST, Windschitl et al., 2020 ) to support data-driven reasoning, uncertainty exploration, and evidence-based explanation. Data Puzzle lessons have been adopted by hundreds of educators, empowering thousands of students to analyze and interpret authentic datasets while building critical data literacy skills. In parallel, the emerging Puzzle Piece resources allow students and teachers to extend structured learning into open-ended investigations using large, secondary datasets within a free, web-based analysis platform (CODAP). This combination of structured and open-ended resources creates an accessible pathway from curated datasets toward exploratory science practices, amplifying the’ broader impacts of science by translating complex datasets into classroom-ready learning experiences. This presentation will demonstrate how Data Puzzles serve as a model for scientists to broaden the impact of their research, enabling authentic data use and discovery in classrooms nationwide. We will share examples of how scientists’ data have been integrated into Data Puzzle lessons, highlight outcomes from teacher professional learning experiences, and discuss how this framework helps foster equitable data access and supports the next generation of Earth scientists and informed citizens. 
    more » « less
  2. Over the last decade, reform in science education has placed an emphasis on the science practices as a way to engage students in the process of science and improve scientific literacy. A critical component of developing scientific literacy is learning to apply quantitative reasoning to authentic scientific phenomena and problems. Students need practice moving fluidly (or fluently) between math and science to develop a habit of mind that encourages the application of quantitative reasoning to real-world scenarios. Here we present a student-facing model that challenges students to think across these two fields. The model brings together math and science with a goal to increase scientific literacy by engaging students in quantitative reasoning within the context of scientific questions and phenomena. In the classroom, the model serves to help students visualize the logical and necessary moves they make as they use quantitative reasoning to connect science practices with mathematical thinking. 
    more » « less
  3. Abstract Incorporating authentic research skills and practices into K‐12 science, technology, engineering, and mathematics (STEM) instruction is a challenging yet crucial approach for introducing students to authentic science inquiry. While recommendations for emphasizing data literacy and quantitative reasoning in science classroom contexts are well‐established, implementation remains challenging. Over the span of 4 years (2019–2023), a multi‐institution team of teachers, education researchers, and forest scientists established a partnership with the overarching goal of integrating authentic forest research and data into middle and high school classrooms. The education researchers played a critical role in facilitating effective scientist and teacher interactions while addressing classroom implementation challenges. Importantly, the effectiveness and mutual benefits of the research partnership were greatly influenced by specific practices implemented by the education research team, and the assumption of different collaborative roles by all stakeholders involved. In this study, we examine these roles, relationships, and interactions of all stakeholders in the partnership, with “stakeholder” referring to participating teachers, education researchers, and collaborating forest scientists. 
    more » « less
  4. null (Ed.)
    Authentic, “messy data” contain variability that comes from many sources, such as natural variation in nature, chance occurrences during research, and human error. It is this messiness that both deters potential users of authentic data and gives data the power to create unique learning opportunities that reveal the nature of science itself. While the value of bringing contemporary research and messy data into the classroom is recognized, implementation can seem overwhelming. We discuss the importance of frequent interactions with messy data throughout K–16 science education as a mechanism for students to engage in the practices of science, such as visualizing, analyzing, and interpreting data. Next, we describe strategies to help facilitate the use of messy data in the classroom while building complexity over time. Finally, we outline one potential sequence of activities, with specific examples, to highlight how various activity types can be used to scaffold students' interactions with messy data. 
    more » « less
  5. Data science education can help broaden participation in computer science (CS) because it provides rich, authentic contexts for students to apply their computing knowledge. Data literacy, particularly among underrepresented students, is critical to everyone in this increasingly digital world. However, the integration of data science into K-12 schools is nascent, and the pedagogical training of CS teachers in data science remains limited. Our research-practice partnership modified an existing data science unit to include two pedagogical techniques known to support minoritized students: rich classroom discourse and personally-relevant problem-solving. This paper describes the iterative design process we used to revise and pilot this new data science unit. 
    more » « less