What happens when a diverse group of youth ages 11 through 14 are introduced to data science using authentic, public, multivariate data in an out-of-school context assuming no special prerequisite knowledge? We designed three 10-hour Data Club modules in which real-world data and the questions students asked of such data drove the learning process. Each module was grounded in a topic that youth connected with at a personal level. Youth learned how to use a free online data platform that made it easy to rearrange, group, filter, and graph data. Within the progression of the module, we used youths’ own questions, data moves, and data visualizations to engage them in critical inquiry and foster productive habits of mind for working with data. Our goal was for youth to emerge from the Data Clubs experience feeling empowered to interact with, ask questions of, and reason about and from data.
more »
« less
Absolute Paleointensity Estimates from Precambrian India and the Long-Term Thermal Evolution of the Earth (Dataset)
Paleomagnetic, rock magnetic, or geomagnetic data found in the MagIC data repository from a paper titled: Absolute Paleointensity Estimates from Precambrian India and the Long-Term Thermal Evolution of the Earth
more »
« less
- Award ID(s):
- 2016763
- PAR ID:
- 10673677
- Publisher / Repository:
- Magnetics Information Consortium (MagIC)
- Date Published:
- Edition / Version:
- 1
- Subject(s) / Keyword(s):
- Igneous Volcanic Dike Basalt 740000000 2363000000 Years BP
- Format(s):
- Medium: X
- Sponsoring Org:
- National Science Foundation
More Like this
-
-
Abstract. The Global Ocean Data Analysis Project (GLODAP) is a synthesis effort providing regular compilations of surface to bottom ocean biogeochemical bottle data, with an emphasis on seawater inorganic carbon chemistry and related variables determined through chemical analysis of seawater samples. GLODAPv2.2023 is an update of the previous version, GLODAPv2.2022 (Lauvset et al., 2022). The major changes are as follows: data from 23 new cruises were added. In addition, a number of changes were made to the data included in GLODAPv2.2022. GLODAPv2.2023 includes measurements from more than 1.4 million water samples from the global oceans collected on 1108 cruises. The data for the now 13 GLODAP core variables (salinity, oxygen, nitrate, silicate, phosphate, dissolved inorganic carbon, total alkalinity, pH, chlorofluorocarbon-11 (CFC-11), CFC-12, CFC-113, CCl4, and SF6) have undergone extensive quality control with a focus on the systematic evaluation of bias. The data are available in two formats: (i) as submitted by the data originator but converted to World Ocean Circulation Experiment (WOCE) exchange format and (ii) as a merged data product with adjustments applied to minimize bias. For the present annual update, adjustments for the 23 new cruises were derived by comparing those data with the data from the 1085 quality-controlled cruises in the GLODAPv2.2022 data product using crossover analysis. SF6 data from all cruises were evaluated by comparison with CFC-12 data measured on the same cruises. For nutrients and ocean carbon dioxide (CO2), chemistry comparisons to estimates based on empirical algorithms provided additional context for adjustment decisions. The adjustments that we applied are intended to remove potential biases from errors related to measurement, calibration, and data-handling practices without removing known or likely time trends or variations in the variables evaluated. The compiled and adjusted data product is believed to be consistent to better than 0.005 in salinity, 1 % in oxygen, 2 % in nitrate, 2 % in silicate, 2 % in phosphate, 4 µmol kg−1 in dissolved inorganic carbon, 4 µmol kg−1 in total alkalinity, 0.01–0.02 in pH (depending on region), and 5 % in the halogenated transient tracers. The other variables included in the compilation, such as isotopic tracers and discrete CO2 fugacity (fCO2), were not subjected to bias comparison or adjustments. The original data, their documentation, and DOI codes are available at the Ocean Carbon and Acidification Data System of NOAA National Centers for Environmental Information (NCEI), which also provides access to the merged data product. This is provided as a single global file and as four regional ones – the Arctic, Atlantic, Indian, and Pacific oceans – under https://doi.org/10.25921/zyrq-ht66 (Lauvset et al., 2023). These bias-adjusted product files also include significant ancillary and approximated data, which were obtained by interpolation of, or calculation from, measured data. This living data update documents the GLODAPv2.2023 methods and provides a broad overview of the secondary quality control procedures and results.more » « less
-
Abstract LIGO, Virgo, and KAGRA form a network of gravitational-wave observatories. Data and analysis results from this network are made publicly available through the Gravitational Wave Open Science Center. This paper describes open data from this network, including the addition of data from the first part of the fourth observing run and selected periods from the preceding engineering run, collected from 2023 May to 2024 January. The public dataset includes calibrated strain time series for each instrument, data from additional channels used for noise subtraction and detector characterization, and analysis data products from version 4.0 of the Gravitational-Wave Transient Catalog.more » « less
-
The 21 cm signal of neutral hydrogen contains a wealth of information about the poorly constrained era of cosmological history, the Epoch of Reionization (EoR). Recently, AI models trained on EoR simulations have gained significant attention as a powerful and flexible option for inferring parameters from 21 cm observations. However, previous works show that AI models trained on data from one simulator fail to generalize to data from another, raising doubts about AI models’ ability to accurately infer parameters from observation. We develop a new strategy for training AI models on cosmological simulations based on the principle that increasing the diversity of the training data set improves model robustness by averaging out spurious and contradictory information. We train AI models on data from different combinations of four simulators, then compare the models’ performance when predicting on data from held-out simulators acting as proxies for our Universe. We find that models trained on data from multiple simulators perform better on data from a held-out simulator than models trained on data from a single simulator, indicating that increasing the diversity of the training data set improves a model’s ability to generalize. This result suggests that future EoR parameter inference methods can mitigate simulator-specific bias by incorporating multiple simulation approaches into their analyses.more » « less
An official website of the United States government
