Geochemical data from sedimentary rocks are the primary source of information regarding Earth's surface evolution through time, including its air and water envelopes and interactions with life and deep Earth processes. The Sedimentary Geochemistry and Paleoenvironments Project (SGP) is a scientific consortium centered around open data and community-driven development of cyberinfrastructure tools and resources for sedimentary geochemistry and Earth history. Here we describe the SGP Phase 2 data release, which focused on incorporating Paleoproterozoic and Mesoproterozoic (2500–1000 million years ago) data and better accommodating carbonate data. This data release was built through the involvement of >200 researchers worldwide in academia, government, and industry, and provides the largest available public data resource for our user community in the academic fields of geochemistry, sedimentology, tectonics, paleontology, Earth history, and paleoclimate, as well as the petroleum and minerals industries. The dataset now encompasses 126,006 samples and 4,132,371 geochemical analyses. In addition to direct entry by SGP Team Members, we have ingested and incorporated datasets from the Geoscience Australia OZCHEM database, the Alberta Geological Survey, and the Deep-Time Marine Sedimentary Element Database (DM-SED) compilation. This paper details sampling in the Phase 2 dataset with respect to age, geography, lithology, and other geological characteristics, documents access via our search website and API, discusses possible issues and/or biases in the dataset that could impact analyses, describes plans for governance and stewardship of data from Indigenous lands, and serves as the citable reference paper for the data release.
Neoproterozoic strata of the Lomfjorden and lower Hinlopenstretet Supergroups (early-middle Tonian Veteranen Group, middle-late Tonian Akademikerbreen Group, and Cryogenian-Ediacaran Polarisbreen Group) in northeastern Svalbard, Norway, represent one of the most complete and wellpreserved terminal Proterozoic sedimentary successions worldwide. These strata have been used in numerous geochemical compilations to probe Neoproterozoic paleoenvironments, including patterns of global chemical weathering prior to the snowball Earth events. Here, we present a new high-resolution geochemical data set incorporating delta 13C, delta 18O, and 87Sr/86Sr data from carbonates with mineralogy, major-and trace-element data, and epsilon Nd data from mudrocks to critically examine previously published data sets and better ally the succession with global Neoproterozoic records through a refined age model. We found that the geochemistry of Lomfjorden and lower Hinlopenstretet Supergroup mudrocks is greatly affected by carbonate contamination and that rigorous filtering is imperative for extracting reliable provenance and weathering information. Previous studies ascribed relatively juvenile detrital Nd isotope signatures in these units to a mafic detrital source region in support of the hypothesis that basalt weathering contributed to global cooling during the Tonian prior to the Sturtian snowball Earth event. In contrast, we present a new temporally extensive epsilon Nd record that, when combined with elemental geochemical data, is best explained by the weathering of a combination of young and old granitic to granodioritic crustal sources, most likely sourced from the Grenville Province of Laurentia or Sveconorwegian Province of Scandinavia. Additionally, this new high-resolution mudrock data set allows us to assess paleoenvironmental and weathering trends within the Lomfjorden and lower Hinlopenstretet Supergroups, including the factors leading to exceptional fossil preservation in the ca. 790 Ma Algal Dolomite member of the Svanbergfjellet Formation, and post-snowball Earth weathering dynamics.
Exceptionally preserved organic-walled microfossils of the Svanbergfjellet Formation, Svalbard document the diversity of eukaryotes in a Tonian (1000–720 myr ago) shallow sea. We review this fossil Lagerstätte and re-sample it at the highest stratigraphic resolution to date. We place the Lagerstätte in an updated age model and evaluate fossil taphonomy. The Svanbergfjellet Formation is one of the most biodiverse units in the Tonian, c . 50% more diverse than the average fossiliferous Tonian unit. It notably preserves the green alga Proterocladus and the possible green alga Palaeastrum . These multicellular fossils, when combined with others elsewhere, suggest that green algae were well established in Tonian marine ecosystems, but substantially predate organic biomarkers that record a dominant green algal contribution to marine primary productivity only by the Ediacaran. Svanbergfjellet also includes several problematic taxa with complex multicellular morphologies (e.g. Jacutinema and Valkyria ) that require further investigation into their phylogenetic placement, but that could substantially add to the known diversity of crown eukaryotes in the Proterozoic.
Oolitic ironstones are iron-rich and chert-poor sedimentary rocks containing concentrically coated grains composed of iron (oxyhydr)oxides and iron phyllosilicates that offer a unique window into iron cycling in ancient coastal environments. These enigmatic deposits are common in the Phanerozoic stratigraphic record yet lack clear modern analogues, and curiously are thought to be absent from Precambrian strata, suggesting a secular control on their deposition. Here we describe a previously unreported ironstone from the middle Tonian (ca. 850 Ma) Katherine Group in the Wernecke Inlier (Yukon, Canada), and show that similar deposits can be found-albeit rarely-throughout the Proterozoic. We investigate the origin of this unit and evaluate its palaeoenvironmental significance, and in light of an extensive literature review, present a holistic model for Precambrian ironstone deposition. The Katherine ironstone occurs in multiple horizons in the McClure and Abraham Plains formations and contains iron ooids and oncoids composed dominantly of authigenic hematite and berthierine, with detrital quartz grains. Textural relationships demonstrate that these coated grains formed on the seafloor with synsedimentary reworking, and the fine interlamination of these phases in grain coatings suggests redox and pH fluctuation during ironstone genesis. Facies associations indicate that the ironstones accumulated in a range of low-energy, shallow marine environments (tidal mudflats and coastal embayments). Geochemical analyses offer insights into genetic processes, and the radiogenic Nd isotope composition and negative Eu anomalies of the Katherine ironstone suggest a continental iron source. We present a model whereby abundant iron, cations, and silica-requisite for the authigenesis of iron phyllosilicates-were supplied from chemical weathering and preferentially enriched in coastal environments due to gradients in pH, Eh and salinity. This continental input would have led to intense iron cycling coupled to organic matter respiration, iron phyllosilicate authigenesis (i.e., reverse weathering). The enrichment of authigenic Fe(III) (oxyhydr)oxides and Fe(II) phyllosilicates took place on a broad coastal plain influenced by both autogenic and allogenic fluctuations in relative sea level, likely in a humid, tropical climate. The lenticular and episodic nature of ironstones in the Proterozoic stratigraphic record suggests that a unique combination of environmental conditions fostered ironstone accumulation. By reviewing the literature on oolitic ironstones, we re-evaluate the temporal distribution of these deposits compared to Archaean-Palaeoproterozoic iron formations, and show that the Great Oxidation Event may have been a prerequisite for ironstone deposition, which may implicate oxidative chemical weathering or suboxic, marine iron cycling. In general, we suggest that the Precambrian record of oolitic ironstones represents an important deep-time archive of iron and nutrient cycling in coastal settings.
Tonian (ca. 1000-720 Ma) marine environments are hypothesised to have experienced major redox changes coinciding with the evolution and diversification of multicellular eukaryotes. In particular, the earliest Tonian stratigraphic record features the colonisation of benthic habitats by multicellular macroscopic algae, which would have been powerful ecosystem engineers that contributed to the oxygenation of the oceans and the reorganisation of biogeochemical cycles. However, the paleoredox context of this expansion of macroalgal habitats in Tonian nearshore marine environments remains uncertain due to limited well-preserved fossils and stratigraphy. As such, the interdependent relationship between early complex life and ocean redox state is unclear. An assemblage of macrofossils including the chlorophyte macroalga Archaeochaeta guncho was recently discovered in the lower Mackenzie Mountains Supergroup in Yukon (Canada), which archives marine sedimentation from ca. 950-775 Ma, permitting investigation into environmental evolution coincident with eukaryotic ecosystem evolution and expansion. Here we present multi-proxy geochemical data from the lower Mackenzie Mountains Supergroup to constrain the paleoredox environment within which these large benthic macroalgae thrived. Two transects show evidence for basin-wide anoxic (ferruginous) oceanic conditions (i.e., high FeHR/FeT, low Fepy/FeHR), with muted redox-sensitive trace metal enrichments and possible seasonal variability. However, the weathering of sulfide minerals in the studied samples may obscure geochemical signatures of euxinic conditions. These results suggest that macroalgae colonized shallow environments in an ocean that remained dominantly anoxic with limited evidence for oxygenation until ca. 850 Ma. Collectively, these geochemical results provide novel insights into the environmental conditions surrounding the evolution and expansion of benthic macroalgae and the eventual dominance of oxygenated oceanic conditions required for the later emergence of animals.
A geologically rapid Neoproterozoic oxygenation event is commonly linked to the appearance of marine animal groups in the fossil record. However, there is still debate about what evidence from the sedimentary geochemical record—if any—provides strong support for a persistent shift in surface oxygen immediately preceding the rise of animals. We present statistical learning analyses of a large dataset of geochemical data and associated geological context from the Neoproterozoic and Palaeozoic sedimentary record and then use Earth system modelling to link trends in redox-sensitive trace metal and organic carbon concentrations to the oxygenation of Earth’s oceans and atmosphere. We do not find evidence for the wholesale oxygenation of Earth’s oceans in the late Neoproterozoic era. We do, however, reconstruct a moderate long-term increase in atmospheric oxygen and marine productivity. These changes to the Earth system would have increased dissolved oxygen and food supply in shallow-water habitats during the broad interval of geologic time in which the major animal groups first radiated. This approach provides some of the most direct evidence for potential physiological drivers of the Cambrian radiation, while highlighting the importance of later Palaeozoic oxygenation in the evolution of the modern Earth system. Oxygen in shallow shelf waters rose linearly with atmospheric oxygen in the Neoproterozoic era, potentially driving the first radiation of marine animals, but widespread ocean oxygenation came later, according to reconstructions of oxygen levels and marine productivity.
The Bylot basins of northeastern Canada and northwestern Greenland comprise the Borden, Aston-Hunting, Fury and Hecla, and Thule basins. This system of late Mesoproterozoic ( c. 1.27–1.0 Ga) sedimentary basins preserves an important record of present day northeastern Laurentia coincident with the emplacement of the Mackenzie large igneous province, the Shawinigan and Ottawan phases of the Grenville Orogeny, and the development of the Midcontinent Rift. However, establishing correlations between the sedimentary successions of the Bylot basins has been hindered by the absence of robust chronostratigraphic constraints. As a result, the degree to which these basins were interconnected, whether they share a common tectonostratigraphic history, and how their sedimentary patterns relate to regional tectonic events remain open questions. Recent Re–Os geochronology from organic-rich strata has yielded depositional ages from the Borden (1048 and 1046 Ma) and Fury and Hecla (1087 Ma) basins, which we integrate with existing models for the depositional history of these basins to derive three tectonostratigraphic assemblages from the Bylot basins. We project our refined tectonostratigraphic framework for the Borden and Fury and Hecla successions to Greenland to establish a testable hypothesis for how the Thule Supergroup fits into this tectonostratigraphic picture.
Europa is a premier target for advancing both planetary science and astrobiology, as well as for opening a new window into the burgeoning field of comparative oceanography. The potentially habitable subsurface ocean of Europa may harbor life, and the globally young and comparatively thin ice shell of Europa may contain biosignatures that are readily accessible to a surface lander. Europa's icy shell also offers the opportunity to study tectonics and geologic cycles across a range of mechanisms and compositions. Here we detail the goals and mission architecture of the Europa Lander mission concept, as developed from 2015 through 2020. The science was developed by the 2016 Europa Lander Science Definition Team (SDT), and the mission architecture was developed by the preproject engineering team, in close collaboration with the SDT. In 2017 and 2018, the mission concept passed its mission concept review and delta-mission concept review, respectively. Since that time, the preproject has been advancing the technologies, and developing the hardware and software, needed to retire risks associated with technology, science, cost, and schedule.
The volcanic emplacement and subsequent weathering of the Deccan Traps of India is believed to have had a significant influence in driving global climatic shifts from the Late Cretaceous and through the Cenozoic. The magnitude of the Deccan Traps' impact on Earth's surface environment is largely dependent on the speculated original footprint of the large igneous province. To test established estimates for the pre-erosive northern extent of the Deccan Traps, we applied low-temperature apatite (U-Th)/He thermochronology (AHe) on rocks from the Bundelkhand craton and overlying Proterozoic Vindhyan successions of central India similar to 150-200 km northeast of the northernmost preservation of Deccan basalts. New AHe data reveal young similar to 5-85 Ma AHe dates with low effective uranium concentrations (eU) between 5-22 ppm, with a steep positive date-eU correlation that plateaus at similar to 350 Ma in grains with eU values >50 ppm. Inverse thermal history modeling-utilizing AHe diffusion parameters of the Radiation Damage Accumulation and Annealing Model (RDAAM)-indicate that observed AHe date-eU correlations are most consistent with thermal histories that require the craton and Vindhyan strata to be at or near surface temperatures by similar to 66 Ma, followed by a discrete reheating event associated with Deccan volcanism. These results establish new minimal areal constraints for the northern extent of Deccan volcanism which thermally perturbed much of the Vindhyan succession. Thermal alteration of organic rich Vindhyan sediment may have provided an additional source of volatile emissions that facilitated late Maastrichtian warming at the onset of Deccan volcanism. New minimal northern constraints on Deccan volcanism additionally confirm that large volumes of Deccan basalts have been stripped away since the time of their emplacement, which poses considerable implications for unraveling their role in Cenozoic cooling. (C) 2021 Elsevier B.V. All rights reserved.
Detrital zircon U-Pb geochronology is one of the most common methods used to constrain the provenance of ancient sedimentary systems. Yet, its efficacy for precisely constraining paleogeographic reconstructions is often complicated by geological, analytical, and statistical uncertainties. To test the utility of this technique for reconstructing complex, margin-parallel terrane displacements, we compiled new and previously published U-Pb detrital zircon data (n = 7924; 70 samples) from Neoproterozoic-Cambrian marine sandstone-bearing units across the Porcupine shear zone of northern Yukon and Alaska, which separates the North Slope subterrane of Arctic Alaska from northwestern Laurentia (Yukon block). Contrasting tectonic models for the North Slope subterrane indicate it originated either near its current position as an autochthonous continuation of the Yukon block or from a position adjacent to the northeastern Laurentian margin prior to >1000 km of Paleozoic?Mesozoic translation. Our statistical results demonstrate that zircon U-Pb age distributions from the North Slope subterrane are consistently distinct from the Yukon block, thereby supporting a model of continent-scale strike-slip displacement along the Arctic margin of North America. Further examination of this dataset highlights important pitfalls associated with common methodological approaches using small sample sizes and reveals challenges in relying solely on detrital zircon age spectra for testing models of terranes displaced along the same continental margin from which they originated. Nevertheless, large -n detrital zircon datasets interpreted within a robust geologic framework can be effective for evaluating translation across complex tectonic boundaries.
The Hecla Hoek succession of northeastern Svalbard, Norway, is an similar to 7 km thick Tonian-Ordovician sedimentary succession that overlies Stenian-Tonian felsic igneous and metasedimentary rocks. The carbonate-dominated upper Tonian-Ediacaran (ca. 820-600 Ma) Akademikerbreen and Polarisbreen groups have yielded important insights into Earth's Neoproterozoic climate, environment, and biological evolution. However, the underlying siliciclastic-dominated lower Tonian (ca. 950-820 Ma) Veteranen Group has garnered little attention despite the fact that it is remarkably well-preserved and hosts diverse microfossil assemblages. Here, we present the first detailed sedimentological analysis of the Veteranen Group from a continuous similar to 4.4km thick stratigraphic section at Faksevagen, Ny Friesland, Spitsbergen. Integrated facies analysis, sequence stratigraphy, and carbonate delta C-13(carb) and delta O-18(carb) chemostratigraphy elucidate the early depositional history of the Hecla Hoek basin and provide fundamental paleoenvironmental constraints for future investigations of this succession as an archive of Tonian Earth History. The Veteranen Group records a long-lived deltaic and storm-influenced marine sedimentary system that reveals dynamics of Precambrian clastic sedimentation prior to the evolution of land plants. Five asymmetric transgressive-regressive (T-R) sequences within the Veteranen Group thin upwards, providing support for the hypothesis that the contact with the Akademikerbreen Group represents a rift-to-drift transition. This complex record of Tonian deltaic and storm-influenced marine sedimentation along the Laurentian margin strengthens correlation between the Veteranen Group and coeval strata from East Greenland and sets the stage to better understand the Proterozoic tectonic evolution of the North Atlantic-circum-Arctic region following the Grenville orogeny. (C) 2021 Elsevier B.V. All rights reserved.
Geobiology explores how Earth's system has changed over the course of geologic history and how living organisms on this planet are impacted by or are indeed causing these changes. For decades, geologists, paleontologists, and geochemists have generated data to investigate these topics. Foundational efforts in sedimentary geochemistry utilized spreadsheets for data storage and analysis, suitable for several thousand samples, but not practical or scalable for larger, more complex datasets. As results have accumulated, researchers have increasingly gravitated toward larger compilations and statistical tools. New data frameworks have become necessary to handle larger sample sets and encourage more sophisticated or even standardized statistical analyses. In this paper, we describe the Sedimentary Geochemistry and Paleoenvironments Project (SGP; Figure 1), which is an open, community-oriented, database-driven research consortium. The goals of SGP are to (1) create a relational database tailored to the needs of the deep-time (millions to billions of years) sedimentary geochemical research community, including assembling and curating published and associated unpublished data; (2) create a website where data can be retrieved in a flexible way; and (3) build a collaborative consortium where researchers are incentivized to contribute data by giving them priority access and the opportunity to work on exciting questions in group papers. Finally, and more idealistically, the goal was to establish a culture of modern data management and data analysis in sedimentary geochemistry. Relative to many other fields, the main emphasis in our field has been on instrument measurement of sedimentary geochemical data rather than data analysis (compared with fields like ecology, for instance, where the post-experiment ANOVA (analysis of variance) is customary). Thus, the longer-term goal was to build a collaborative environment where geobiologists and geologists can work and learn together to assess changes in geochemical signatures through Earth history. With respect to the data product, SGP is focused on assembling a well-vetted and comprehensive dataset that is tractable to multivariate statistical analyses accounting for multiple geological and methodological biases. Phase 1 of the project, which focused on the Neoproterozoic and Paleozoic, has been completed. Future phases will capture a broader range of geologic time, data types, and geography. The database contains tens of thousands of unpublished data points provided by consortium members, as well as detailed metadata that go beyond what is contained in papers. In many cases, these represent measurements that are tangential to a given published study but still of high utility to database studies; these allow the community to address questions that would be impossible to answer solely with the published data. For instance, in order to use a proxy such as Mo/TOC (total organic carbon) ratios in mudrocks deposited under a euxinic water column, the full suite of trace metal, iron speciation, and total organic carbon data is needed. Likewise, geospatial information is required to account for sampling biases, and many statistical learning approaches cannot accept, or have difficulty with, incomplete geological predictor variables. Ultimately, it is this complete data matrix that will allow for SGP's most insightful analyses. This paper serves as an introduction to SGP, the process by which our data products are created, a description of the Phase 1 data product and a citable reference for that product, a description of the SGP website and API (Application Programming Interface) for open access, and a statement of our future goals. In recent years, there has been a welcome trend in the broader geochemical community toward increased data accessibility, documentation of sample context, and sample curation, albeit with challenges still ahead (Brantley et al., 2020; Cutcher-Gershenfeld et al., 2016; Planavsky et al., 2020). First, progress has been made through journals and organizations adopting stringent data archiving rules and promoting adherence to FAIR principles—findability, accessibility, interoperability, and reusability ("FAIR Play in Geoscience Data," 2019; Wilkinson et al., 2016). Second, several databases now house geochemical data at different scales and with different focuses (Brantley et al., 2020; Gard et al., 2019; He et al., 2019; Lehnert et al., 2000). Among the largest and most active are projects such as EarthChem (earthchem.org), the Geobiodiversity Database (geobiodiversity.com), Pangaea (https://www.pangaea.de), and the StabisoDB (https://cnidaria.nat.uni-erlangen.de/stabisodb/). The SGP database was built with the data structures and standards of these other projects in mind, in keeping with FAIR principles and with the hope that data can be easily shared in the future. Consistent with the stance taken by other organizations in the community (Hanson, 2016), we also strongly encourage all members to register their samples for an International Geo Sample Number (IGSN; i.e., globally unique alphanumeric sample identifiers), which can be obtained from the System for Earth Sample Registration (www.geosamples.org). However, SGP is a domain-specific project that differs from other databases in the way the data are collected, the nature of the data collected, and the tailored way in which they are presented to our research community. Although some other databases contain sedimentary geochemical data, the vast majority of deep-time data is not available from any single source, and samples are not readily associated with critical contextual data—such as age constraints and environmental data—necessary for the types of proxy-through-time and/or environmental studies typically conducted in historical geobiology. When the SGP was founded in 2015, we believed that a "team science" philosophy would be the most effective way to move beyond spreadsheets to the type and abundance of data required. The research consortium framework we have implemented is modeled after mature consortia in human statistical genetics, such as the Psychiatric Genomics Consortium (PGC). In the PGC, researchers have aggregated data to make statistically robust observations and landmark findings not possible with the data generated by any single research group alone (Duncan et al., 2017; Schizophrenia Working group of the Psychiatric Genomics Consortium, 2014; Wray et al., 2018). Similar to biomedical research consortia, we hope that the intellectual and collaborative environment fostered by SGP will ultimately be as important as our data products or specific insights in research papers. The first priority for Phase 1 of SGP was to assemble or generate multi-proxy sedimentary geochemical data (carbon and sulfur abundances and isotopes, iron speciation, major and trace metal abundances, and trace metal isotopes, primarily from fine-grained siliciclastic rocks) from multiple regions worldwide for every Paleozoic Epoch and equivalent ~25 Myr Neoproterozoic time slice. In addition to data compilation, this has involved an effort by SGP members to generate new geochemical data from "background" intervals in the Paleozoic (i.e., not associated with events such as mass extinctions or significant climatic shifts). The first phase of data collection came to an end in 2019. At that point, a copy of the database was vetted by SGP team members and then archived—the first data "freeze" (following the best-practices approach used in medical consortia). Working groups were formed (with working group leadership established through an open call to SGP team members), and data were made available to Working group analysts via the website and through tailored queries. The first working group papers have recently been published (LeRoy et al., 2021; Lipp et al., 2021; Mehra et al., 2021), and more are in progress. Meanwhile, data collection continues, and the Phase 2 goal is to include more Mesozoic–Cenozoic and pre-Neoproterozoic time intervals and to expand the geochemical record to more diverse lithologies and grain-specific phases. The Phase 2 data freeze is currently anticipated for 2023, followed by data vetting and analyses toward group papers. SGP utilizes a relational database implemented with the PostgreSQL database management system. A full database diagram and documentation are available at https://github.com/ufarrell/sgp_phase1, and a simplified diagram is shown in Figure 2. The design was inspired by several existing data models in the geological and natural history museum communities. Tables for analytical geochemistry are from the British Geological Survey (BGS) geochemistry data model (Watson et al., 2014), with minor modifications. Tables for geological, geographical, and sample details are based on established museum collection management databases (Specify 6 https://www.specifysoftware.org/ and Arctos https://arctosdb.org/) in addition to the Observations Data Model 2 (ODM2, Horsburgh et al., 2016; Hsu et al., 2017), an information model for Earth observations. The SGP database is centered on the sample table (Figure 2). Samples are generally characterized by an individual rock sample and all resulting analyzed powders. The three key sections of the database linked to samples are (1) analytical results and associated methods, (2) geographical context, and (3) geological context. Dictionary tables (standardized lists of terms, also known as "controlled vocabularies") are based on existing community vocabularies where possible (e.g., from EarthChem, ODM2, Macrostrat, U.S. Geological Survey (USGS), and BGS). However, in many cases, these vocabularies required additions, such as the inclusion of specific sedimentary geochemical experimental methods (e.g., sequential iron extraction techniques; Poulton & Canfield, 2005). The BGS data model for analytical methods and geochemical results has been adopted almost without modification. We store analytical data in their submitted or published format and do not standardize the results to any given unit. An analytical result may be empty (NULL) only if it is below or above detection limits, and those values are also stored if they are available. If the results are published, they are linked directly to a reference work on an individual basis so that a fine-level distinction can be made between published and related unpublished data from the same samples. Any geostandards that are analyzed alongside samples in a study are also recorded. In the SGP, we make every effort not to include the same result twice. However, replicates may legitimately be added if the same sample has undergone analysis for the same analyte more than once (this could include anything from true replicate analyses using the same methods in the same laboratory to analyses of the same sample by different research groups using different methods). We do not currently assign new sample identifiers to sub-samples. A parent–child relationship may be added in Phase 2 when the focus will expand to include carbonate data. The SGP welcomes contributions from any interested researchers. Specifically, contributing data automatically makes a researcher part of the SGP Collaborative Team, rather than one needing to "join" SGP to contribute data. In the first consortium-building stage, potential collaborators were targeted if their work was particularly relevant to the Phase 1 goals, and additional researchers were recruited via SGP representation at multiple conferences. SGP collaborators are involved in providing details about their samples and providing published data tables and unpublished data from their own archives. In addition, some data have been collected from relevant published studies where the authors are not directly involved. In such cases, contextual information was coded by SGP team members using information provided in the paper. SGP collaborators are asked to fill in a template with contextual information as completely as possible, but with an emphasis on key fields such as modern latitude and longitude, stratigraphic unit name, depositional environment, and lithology. A particularly important field is interpreted age, which is a numerical estimate for the age of each sample in millions of years (Ma). Whenever possible, the original authors, who are most familiar with the samples and stratigraphic sections, are asked to provide the interpreted age. They can use whatever method with which they feel most comfortable; for example, ages may be estimated based on assumed sedimentation rates and/or linear interpolation, or groups of samples can be assigned one age based on proximity to any available time markers. A brief justification is required for each age provided, which may be used in the future to refine ages further. Maximum and minimum age estimates can also be stored, and indeed, are critical for the type of re-weighted bootstrap analyses employed by many SGP working groups (Mehra et al., 2021). A subset of samples from two USGS databases has been integrated into the SGP database. The first of the databases used is the National Geochemical Database: Rock (USGS NGDB, U.S. Geological Survey, 2008), comprising data from USGS projects from the 1960s to1990s, largely from North America. The second is the Global Geochemical Database for Critical Metals in Black Shales project (USGS CMIBS, Granitto et al., 2017), which includes predominantly Phanerozoic shale data from all continents. Data from both USGS databases lack much of the contextual information available for samples directly coded by the SGP team members (most specifically basin type, metamorphic/maturity grade, depositional environment, and detailed age justification) and there are a higher proportion of analytes with less detailed geochemical methodology. Nevertheless, they represent large numbers of samples (74% of samples in Phase 1 are from USGS sources) with age, lithology, and geographic information that can be utilized for many types of analysis. In the case of USGS NGDB, only sedimentary samples were incorporated into SGP, and in the case of USGS CMIBS, we did not include samples with lithologies indicative of ore or studies where the authors were primarily concerned with mineral deposits or studying the effects of metamorphism on shales. An attempt was made to match USGS fields to SGP fields, with some data cleaning needed in order to extract important information such as up-to-date stratigraphic names. Samples can easily be traced back to the original USGS databases using their original identifiers. The USGS NGDB data were enhanced by adding interpreted ages. Samples were matched, using a combination of stratigraphy and location, to the continuous-time age model in Macrostrat (Peters et al., 2018). Specifically, the minimum and maximum age estimates from the Macrostrat model were entered, and the interpreted age was entered as the average of these values. Only samples with matched interpreted ages were included from USGS NGDB. The USGS CMIBS samples were associated with Macrostrat continuous-time age models where possible and given age information by SGP team members where not. However, a proportion (36%) remain without ages, and filling those in is a key goal for Phase 2. These three sources of data (direct entry by SGP team members (26% of samples), the CMIBS compilation (16% of samples), and the USGS NGDB (58% of samples)) provide a robust base platform for statistical analyses of aggregated sedimentary geochemical data through Earth history. Moving forward, we will continue direct entry from SGP team members, and work toward incorporating geochemical data compiled by additional geological surveys (for instance, incorporation of the OZCHEM whole-rock database from Geoscience Australia is currently in progress). Phase 1 of data collection ended in August 2019. A static version of the database was archived and made available to collaborators through the website (sgp-search.io) and via tailored queries. Time was allowed for vetting, and any errors discovered were corrected before the final freeze in February 2020. The Phase 1 data freeze includes 82,578 samples, with 2,701,236 analytical results, and was made public through our search website in December 2020. This paper should be cited in the future use of Phase 1 data downloads. More complete information on the Phase 1 data product can be found on the SGP wiki (https://github.com/ufarrell/sgp_phase1/wiki), including summaries by age, lithology, and geochemical methodology, as well as the specifics of how USGS databases were incorporated into the SGP structure. The SGP-contributed dataset includes 20,811 samples with 518,291 results. Approximately two thirds of the data (64%) come from 160 published sources (https://github.com/ufarrell/sgp_phase1/wiki/SGP-data-references). The remaining 36% are from unpublished sources, including new and legacy data. The samples come from 942 individual sites from 46 countries (Figure 3). Consistent with the Phase 1 goals, 84% of samples were from the Neoproterozoic–Paleozoic (Figure 4). Sixty-four percent of samples are fine-grained siliciclastic rocks (shale, mudstone, or siltstone), as are the majority of uncoded lithologies (Figure 5). The data from USGS NGDB that are incorporated into the SGP database include 48,234 samples with 1,769,696 results. Nearly all (99%) of the samples are from the United States. Nineteen percent are sandstone, 13% are shale, and 29% do not have a specific lithology (although lithological details may be available in verbatim fields; Figure 5). Contextual details, including depositional environment and low-grade metamorphic bin, are mostly not available for these samples, and methodological information is sparse. In general, the USGS NGDB samples skew younger than the SGP samples: 39% are from the Paleozoic, 25% from the Mesozoic, and 33% from the Cenozoic (~3% of samples are from the Proterozoic/Archean). The USGS database provides excellent coverage of the United States, but given the remit of the organization, with strong focus on economic deposits (petroleum-producing units, phosphatic units, and sedimentary mineral deposits), the sampling may not be representative of the entire country. This is distinct from the bias present in geochemical data produced by academic researchers, which are often focused on mass extinction intervals, Earth system perturbations, and other stratigraphic boundaries. The data incorporated from USGS CMIBS into the SGP database include 12,797 samples with 409,188 results. The samples are from 45 countries, with 40% from Canada, 27% from the United States, and 13% from Australia. The majority of samples are fine-grained siliciclastic sediments (69% shale, mudstone, siltstone, or argillite; Figure 5). Sixty percent of samples with interpreted ages are Paleozoic, 24% are Mesozoic, 2% are Cenozoic, and 15% are Proterozoic/Archean. As was the case for USGS NGDB, contextual details, including depositional environment and low-grade metamorphic bin, are often missing for these samples. However, more detailed geochemical methodological information is available. Each sample in CMIBS has a "best value" result per analyte, selected from multiple values that were originally available (Granitto et al., 2017). The choice of "best value" was made using a rubric which included consideration of the sample weight, the sample "decomposition" (e.g., full vs. partial acid digestion), the instruments used in the analysis, and the detection limits (Granitto et al., 2013). The SGP search website (sgp-search.io) utilizes an intuitive user interface to query the Phase 1 database via an API. The two main search types are "samples" and "analyses," with "nhhxrf" simply being a "samples" search that excludes any handheld XRF (X-ray fluorescence) data. This methodological distinction is made because while handheld XRF data can be accurate for some elements (e.g., Ca and Fe), it is highly inaccurate for many others (e.g., S, Ni) (Rowe et al., 2012). Handheld XRF data represent 1% of the total results and 4% of SGP-contributed data; although this is a small percentage now, we anticipate continued growth given the popularity and utility of handheld XRFs. A "samples" search will list an individual sample on each row, with geological context information and geochemical analytes taking up the columns. Data are converted to one standard unit, and oxides are converted to elements (e.g., Al2O3 to Al), and values are averaged if more than one analysis was made per sample. Note, this search may average values produced using different analytical methods, although the number of samples in the database with multiple analytical values for a specific analyte is relatively small. Further, any analyses below or above detection limit are removed, as these cannot be averaged. This has implications for queries involving very low abundance elements (e.g., Ag in sedimentary rocks), as only results above detection limits, and thus higher values, will be included. We anticipate that this search will produce the optimal data output for most end-users interested in Earth history: a file with age, geological context, and geochemical data for each sample. If users are looking to delve deeper into the data and understand the analyses and procedures that were executed to obtain each sample's geochemical data, then the "analyses" search is useful because it lists every analysis recorded in the database in a separate row. The "analyses" search also allows users to show data relating to the laboratory where the sample was analyzed, the person who made the measurement, geochemical methodology, etc. At the current time, aside from the ability to exclude handheld XRF data, the "samples" and "nhhxrf" search types will not report information about, or have the ability to filter by, geochemical methodology. Users who are interested in methodological details or who would like to export a data file beyond the size limit (10 Mb) should contact the SGP Leadership Team regarding a custom SQL query. Once the user has selected a search type, samples can be filtered based on both geological context and geochemical attributes. Note that for many samples some aspects of geological contextual information are incomplete. Thus, for example, a search filtering for samples deposited in a rift basin will only return samples positively described as such and not necessarily all samples in the database deposited in rift basins. Given that samples will have non-overlapping missing data, too many filters may result in a smaller-than-expected dataset. Search results will appear in a "preview" window that can be used to check the output. Each sample also has an information icon associated with it; clicking this icon will bring up a lightbox with detailed sample information. Finally, the user may request to show reference information for their search. For "analyses" searches (where every analysis is shown as an individual row), this will return the specific literature citation for that individual analytic result. For other search types, this will return, for every sample, a concatenated list of all references whose geochemical data contributed to that specific search. When the user is satisfied with their search, they can then download a.csv file of the data and export a map showing the location and age of samples in their search. Thus, an example API call would be {"type":"samples","filters":{"country":["Argentina","Brazil","Chile","Bolivia","Colombia","Venezuela"],"toc":[2,100]},"show":["toc","fe","height_meters","section_name","country","interpreted_age"]}. This API call is making a "samples" type search for samples that originate from Argentina, Brazil, Chile, Bolivia, Colombia, or Venezuela and have 2%–100% total organic carbon (TOC) content. In other words, searching for organic-rich samples from South America. In addition, the API call is asking for a results output table with columns that show TOC (wt%), Fe (wt%), section or core name, collection height in meters, each sample's country, and the age in millions of years. Full documentation and a tutorial video are available on the website. The overarching goal of SGP was to provide intellectual and geoinformatic resources for the Earth Science community to advance our understanding of environmental changes on Earth through time. A better understanding of Earth's history requires sufficient data density, but equally importantly it means training a new generation of researchers with the data science and statistical skills to make meaningful conclusions from large sedimentary geochemical datasets. Much of the focus in SGP Phase 1 was in initiating the consortium and increasing the data product to the point where it was useful for analyses by the community. We now aim to increasingly move toward developing a community-initiated set of best practices for data management, a culture of publishing metadata, and a shared intellectual framework for analyzing such datasets. Over the course of Phase 2, we plan to continue holding annual meetings at Goldschmidt while also beginning regular video calls to share progress and ideas for data analysis. We will also develop accessible "Proxy Primer" videos to help the geobiological community understand the strengths and weaknesses of different proxies. Echoing this final point, we reiterate that the SGP is a community-oriented research consortium, and we welcome suggestions on how to best move toward our shared goals. We thank Sufian Lattouf for developing the initial version of the SGP website, and Kai Lenz, Kassie Sharp, Aaron Cole, Clare Swan, Lyna Kim, and John Freshwaters for computational assistance. We thank Erin Saupe, Itay Halevy, Jordon Hemingway, Minming Cui, Maya Gomes, Matthew Granitto, Alf Lenz, Charles Henderson, Chengsheng Jin, Clint Scott, David Champion, Jinghai Yang, Joe Shaffer, Kathy Doyle, Lei Xiang, Liam Bhajan, Patrick Sack, Paul Hoffman, Paulo Linarde Dantas Mascena, Will Thompson-Butler, and Yu Liu for their contributions to SGP. We thank Patrick Sullivan and Laramie Duncan for discussions regarding the PGC and research consortium organization. We thank the donors of The American Chemical Society Petroleum Research Fund for partial support of SGP website development (61017-ND2). EAS is funded by National Science Foundation grant (NSF) EAR-1922966. BGS authors (JE, PW) publish with permission of the Executive Director of the British Geological Survey, UKRI. Any use of trade, firm, or product names is for descriptive purposes only and does not imply endorsement by the U.S. Government. The authors declare no conflicts of interest.
Molecular phylogenetic data suggest that photosynthetic eukaryotes first evolved in freshwater environments in the early Proterozoic and diversified into marine environments by the Tonian Period, but early algal evolution is poorly reflected in the fossil record. Here, we report newly discovered, millimeter- to centimeter-scale macrofossils from outershelf marine facies of the ca. 950–900 Ma (Re-Os minimum age constraint = 898 ± 68 Ma) Dolores Creek Formation in the Wernecke Mountains, northwestern Canada. These fossils, variably preserved by iron oxides and clay minerals, represent two size classes. The larger forms feature unbranching thalli with uniform cells, differentiated cell walls, longitudinal striations, and probable holdfasts, whereas the smaller specimens display branching but no other diagnostic features. While the smaller population remains unresolved phylogenetically and may represent cyanobacteria, we interpret the larger fossils as multicellular eukaryotic macroalgae with a plausible green algal affinity based on their large size and presence of rib-like wall ornamentation. Considered as such, the latter are among the few green algae and some of the largest macroscopic eukaryotes yet recognized in the early Neoproterozoic. Together with other Tonian fossils, the Dolores Creek fossils indicate that eukaryotic algae, including green algae, colonized marine environments by the early Neoproterozoic Era.
The geologic processes involved in attaining strengthened cratonic lithosphere remain debated despite their importance for stabilization and long-term preservation. In central India, stabilization of the Bundelkhand craton has conventionally been attributed to the youngest magmatic event impacting the craton at similar to 2.5 Ga, though the post-amalgamation evolution of the craton prior to Proterozoic basin development is poorly understood. This study presents new basement zircon and apatite U-Pb age data along with new detrital zircon U-Pb age data from Proterozoic marginal sedimentary basin deposits to explore the post-magmatic and burial evolution of the Bundelkhand craton. Apatite from similar to 3.4-2.5 Ga granitoids and gneisses collected across the similar to 250 km wide craton yielded near uniform U-Pb ages between similar to 2.4-2.3 Ga, indicating broad-scale exhumation of the Bundelkhand craton through mid-crustal depths following amalgamation and felsic magmatism. Unroofing of the Bundelkhand craton at this time is corroborated by similar to 2.7-2.5 Ga detrital zircon U-Pb age peaks from basal sandstones of the Bijawar and Gwalior groups, which lie in direct nonconformable contact with the craton along both its southeastern and northwestern margins, respectively. These age populations reveal an abundance of zircon sourced directly from the Bundelkhand craton, and a sub-population of similar to 2.2-2.3 Ga grains provide a maximum depositional age for the oldest strata deposited on the craton. We speculate that the redistribution of heat producing elements associated with shallow emplacement of Bundelkhand granitoids and subsequent erosion may have enhanced lithospheric strengthening and facilitated a long-term thermal regime that promoted craton stability by similar to 2.2 Ga. Following stabilization, far-field marginal tectonism likely influenced the vertical motions within the Bundelkhand craton, inducing stages of broad subsidence and erosion recorded within the strata of the Lower and Upper Vindhyan successions.
The latest Mesoproterozoic Arctic Bay Formation (Borden Basin, Nunavut, Canada) is up to similar to 1130 m-thick and contains a significant proportion of unusually organic-rich black shale (up to 12.3 wt% total organic carbon). Insofar as increased biological productivity is related to organic matter burial, this organic-rich succession is seemingly incongruent with the low biological productivity world hypothesised for much of the Proterozoic. To better understand the conditions leading to development of this organic-rich unit, we explore the redox geochemistry of the Arctic Bay Formation using a multi-proxy approach (nitrogen isotopes, iron speciation, total organic carbon, total sulphur, and trace metal abundances). Redox proxy data support a stratified water column, with oxic surface waters underlain by intermittently euxinic waters, which are in turn underlain by persistently ferruginous deeper waters. The highly alkaline, restricted marine basin in which the Arctic Bay Formation was deposited may have allowed for rapid sequestration of highly reactive iron in carbonate minerals, resulting in an 'excess' of sulphur that resulted in sulphurisation of organic matter. Estimates for organic matter burial rates during deposition of the Arctic Bay Formation suggest that they were perhaps similar to 5-6 times mid-Proterozoic average values (although there are permissible scenarios in which it was extremely productive), underscoring that such organic-rich sedimentary rocks could be produced in a low productivity world. (C) 2020 Elsevier B.V. All rights reserved.