Geology ReportsSearch

SEARCH · Geology Reports

Results for “Data”

Search indexed USGS publications on groundwater, aquifers, geologic maps, mineral resources and earthquakes. Explore source records by subject and place.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4Linked to original sources

Evaluation of the Global Land Data Assimilation System (GLDAS) air temperature data products

There is a high demand for agrohydrologic models to use gridded near-surface air temperature data as the model input for estimating regional and global water budgets and cycles. The Global Land Data Assimilation System (GLDAS) developed by combining simulation models with observations provides a long-term gridded meteorological dataset at the global scale. However, the GLDAS air temperature products have not been comprehensively evaluated, although the accuracy of the products was assessed in limited areas. In this study, the daily 0.25° resolution GLDAS air temperature data are compared with two reference datasets: 1) 1-km-resolution gridded Daymet data (2002 and 2010) for the conterminous United States and 2) global meteorological observations (2000–11) archived from the Global Historical Climatology Network (GHCN). The comparison of the GLDAS datasets with the GHCN datasets, including 13 511 weather stations, indicates a fairly high accuracy of the GLDAS data for daily temperature. The quality of the GLDAS air temperature data, however, is not always consistent in different regions of the world; for example, some areas in Africa and South America show relatively low accuracy. Spatial and temporal analyses reveal a high agreement between GLDAS and Daymet daily air temperature datasets, although spatial details in high mountainous areas are not sufficiently estimated by the GLDAS data. The evaluation of the GLDAS data demonstrates that the air temperature estimates are generally accurate, but caution should be taken when the data are used in mountainous areas or places with sparse weather stations.

Journal of Hydrometeorology

Data model and relational database design for the New England Water-Use Data System (NEWUDS)

The New England Water-Use Data System (NEWUDS) is a database for the storage and retrieval of water-use data. NEWUDS can handle data covering many facets of water use, including (1) tracking various types of water-use activities (withdrawals, returns, transfers, distributions, consumptive-use, wastewater collection, and treatment); (2) the description, classification and location of places and organizations involved in water-use activities; (3) details about measured or estimated volumes of water associated with water-use activities; and (4) information about data sources and water resources associated with water use. In NEWUDS, each water transaction occurs unidirectionally between two site objects, and the sites and conveyances form a water network. The core entities in the NEWUDS model are site, conveyance, transaction/rate, location, and owner. Other important entities include water resources (used for withdrawals and returns), data sources, and aliases. Multiple water-exchange estimates can be stored for individual transactions based on different methods or data sources. Storage of user-defined details is accommodated for several of the main entities. Numerous tables containing classification terms facilitate detailed descriptions of data items and can be used for routine or custom data summarization. NEWUDS handles single-user and aggregate-user water-use data, can be used for large or small water-network projects, and is available as a stand-alone Microsoft? Access database structure. Users can customize and extend the database, link it to other databases, or implement the design in other relational database applications.

Open-File Report

Definitions of components of the Water Data Sources Directory maintained by the National Water Data Exchange

This report contains a definition and description of each component of the Water Data Souces Directory data base maintained by the National Water Data Exchange (NAWDEX). It is intended, primarily, to assist those persons using the Water Data Sources Directory in understanding information obtained from the data base. The Water Data Sources Directory is a computerized data base maintained and operated using the System 2000 data base management system. It contains information about organizations that collect, store, and disseminate water data. (Woodard-USGS)

Open-File Report

Definitions of components of the Water Data Sources Directory maintained by the National Water Data Exchange

This report contains a definition and description of each component of the Water Data Sources Directory data base maintained by the National Water Data Exchange (NAWDEX). It is intended, primarily, to assist those persons using the Water Data Sources Directory in understanding information obtained from the data base. The Water Data Sources Directory is a computerized data base maintained and operated using the System 2000 data base management system. It contains information about organizations that collect, store, and disseminate water data. (USGS)

Open-File Report

National water information system user's manual; Volume 2, Chapter 5, Water-Use Data System; Part 1, Site-Specific Water-Use Data System (SSWUDS)

The Water-Use Data System (WUDS) is a water-use data storage and retrieval system. The WUDS is part of the National Water Information System (NWIS), which was developed by the U.S. Geological Survey (USGS). The National Water Information System is a distributed national data base in which data can be processed over a network of minicomputers at USGS offices throughout the United States. This system comprises the Automated Data Processing System, the Ground-Water Site Inventory System, the Water-Quality System, and the WUDS. This manual reflects the status of WUDS for the 90.2 version of NWIS, and contains user information and discussion on the general operating procedures for the programs found within the WUDS menus. The WUDS comprises two subsystems, the Site- Specific Water-Use Data System (SSWUDS) and the Aggregate Water- Use Data System (AWUDS). This manual documents SSWUDS and covers the following major topics: Overall SSWUDS Concepts SSWUDS Menu Structures SSWUDS Data Entry Program SSWUDS Data Dictionary

Open-File Report

Computer Programs to Display and Modify Data in Geographic Coordinates and Methods to Transfer Positions to and from Maps, with Applications to Gravity Data Processing, Global Positioning Systems, and 30-Meter Digital Elevation Models

Computer programs were written in the Fortran language to process and display gravity data with locations expressed in geographic coordinates. The programs and associated processes have been tested for gravity data in an area of about 125,000 square kilometers in northwest Nevada, southeast Oregon, and northeast California. This report discusses the geographic aspects of data processing. Utilization of the programs begins with application of a template (printed in PostScript format) to transfer locations obtained with Global Positioning Systems to and from field maps and includes a 5-digit geographic-based map naming convention for field maps. Computer programs, with source codes that can be copied, are used to display data values (printed in PostScript format) and data coverage, insert data into files, extract data from files, shift locations, test for redundancy, and organize data by map quadrangles. It is suggested that 30-meter Digital Elevation Models needed for gravity terrain corrections and other applications should be accessed in a file search by using the USGS 7.5-minute map name as a file name, for example, file '40117_B8.DEM' contains elevation data for the map with a southeast corner at lat 40? 07' 30' N. and lon 117? 52' 30' W.

Open-File Report

Evaluation of volatile organic compound (VOC) blank data and application of study reporting levels to groundwater data collected for the California GAMA Priority Basin Project, May 2004 through September 2010

Volatile organic compounds (VOCs) were analyzed in quality-control samples collected for the California Groundwater Ambient Monitoring and Assessment (GAMA) Program Priority Basin Project. From May 2004 through September 2010, a total of 2,026 groundwater samples, 211 field blanks, and 109 source-solution blanks were collected and analyzed for concentrations of 85 VOCs. Results from analyses of these field and source-solution blanks and of 2,411 laboratory instrument blanks during the same time period were used to assess the quality of data for the 2,026 groundwater samples. Eighteen VOCs were detected in field blanks or source-solution blanks: acetone, benzene, bromodichloromethane, 2-butanone, carbon disulfide, chloroform, 1,1-dichloroethene, dichloromethane, ethylbenzene, tetrachloroethene, styrene, tetrahydrofuran, toluene, trichloroethene, trichlorofluoromethane, 1,2,4-trimethylbenzene, m - and p -xylenes, and o -xylene. The objective of the evaluation of the VOC-blank data was to determine if study reporting levels (SRLs) were needed for any of the VOCs detected in blanks to ensure the quality of the data from groundwater samples. An SRL is equivalent to a raised reporting level that is used in place of the reporting level used by the analyzing laboratory [long‑term method detection level (LT-MDL) or laboratory reporting level (LRL)] to reduce the probability of reporting false-positive detections. Evaluation of VOC-blank data was done in three stages: (1) identification of a set of representative quality‑control field blanks (QCFBs) to be used for calculation of SRLs and identification of VOCs amenable to the SRL approach, (2) evaluation of potential sources of contamination to blanks and groundwater samples by VOCs detected in field blanks, and (3) selection of appropriate SRLs from among four potential SRLs for VOCs detected in field blanks and application of those SRLs to the groundwater data. An important conclusion from this study is that to ensure the quality of the data from groundwater samples, it was necessary to apply different methods of determining SRLs from field blank data to different VOCs, rather than use the same method for all VOCs. Four potential SRL values were defined by using three approaches: two values were defined by using a binomial probability method based on one-sided, nonparametric upper confidence limits, one was defined as equal to the maximum concentration detected in the field blanks, and one was defined as equal to the maximum laboratory method detection level used during the period when samples were collected for the project. The differences in detection frequencies and concentrations among different types of blanks (laboratory instrument blanks, source-solution blanks, and field blanks collected with three different sampling equipment configurations) and groundwater samples were used to infer the sources and mechanisms of contamination for each VOC detection in field blanks. Other chemical data for the groundwater samples (oxidation-reduction state, co-occurrence of VOCs, groundwater age) and ancillary information about the well sites (land use, presence of known sources of contamination) were used to evaluate whether the patterns of detections of VOCs in groundwater samples before and after application of potential SRLs were plausible. On this basis, the appropriate SRL was selected for each VOC that was determined to require an SRL. The SRLs for ethylbenzene [0.06 microgram per liter (μg/L)], m - and p -xylenes (0.33 μg/L), o -xylene (0.12 μg/L), toluene (0.69 μg/L), and 1,2,4-trimethylbenzene (0.56 μg/L) corresponded to the highest concentrations detected in the QCFBs and were selected because they resulted in the most censoring of groundwater data. Comparisons of hydrocarbon ratios in groundwater samples and blanks and comparisons between detection frequencies of the five hydrocarbons in groundwater samples and different types of blanks suggested three dominant sources of contamination that affected groundwater samples and blanks: (1) ethylbenzene, m - and p -xylenes, o -xylene, and toluene from fuel or exhaust components sorbed onto sampling lines, (2) toluene from vials and the source blank water, and (3) 1,2,4-trimethylbenzene from materials used for collection of samples for radon-222 analysis.

California

Accessibility of geotechnical earthquake Engineering data and the need for data storage and dissemination standards

Ease of data access and data standards are two issues critical to the success of GIS technology when applied to earthquake hazards research problems that require geotechnical engineering and related data. Efforts to reduce data accession costs and to streamline the data exchange process will result in short-term cost and time saving and will add long-term value to the data sets themselves. Such efforts might include centralized data centers, standardized data base designs and formats, cooperative efforts to fill data gaps, and standardized distribution methods and media.

Conference Paper

Logical data model for hydrographic data based on HY_Features concepts

This report describes background and design of the “hydrofabric data model” which defines logic for implementation of data schemas and software that deals with hydrologic geospatial data. As a “logical” data model, the hydrofabric data model specifies details necessary to support compatibility of data and software that satisfy diverse needs without unnecessarily restricting implementation details. The logic presented in this report is based on concepts defined in WaterML2 Part 3 Surface Hydrology Features Concepts and is designed to serve the needs of a range of hydroscience use cases. Development of international community standards applicable to hydrofabrics began, prompted by the World Meteorological Organization Commission for Hydrology, in 2012 [5] . More than 10 years later, this report documents one aspect of a long-term research and development activity that traces its roots back that far. This report describes terminology, use cases, and background as context preceding presentation of the logical model and discussion of its design. Three appendices document related data models, an example encoding of the hydrofabric data model, and an artificial schematic and tabular data example. The sections of the report can be accessed in the Clause 5 section.

OCG Public Engineering Report

Groundwater-quality and select quality-control data from the National Water-Quality Assessment Project, January through December 2015, and previously unpublished data from 2013 to 2014

Groundwater-quality data were collected from 502 wells as part of the National Water-Quality Assessment Project of the U.S. Geological Survey National Water-Quality Program and are included in this report. Most of the wells (500) were sampled from January through December 2015, and 2 of them were sampled in 2013. The data were collected from five types of well networks: principal aquifer study networks, which are used to assess the quality of groundwater used for public water supply; land-use study networks, which are used to assess land-use effects on shallow groundwater quality; major aquifer study networks, which are used to assess the quality of groundwater used for domestic supply; enhanced trends networks, which are used to evaluate the time scales during which groundwater quality changes; and vertical flow-path study networks, which are used to evaluate changes in groundwater quality from shallow to deeper depths. Groundwater samples were analyzed for a large number of water-quality indicators and constituents, including major ions, nutrients, trace elements, volatile organic compounds, pesticides, radionuclides, and some constituents of special interest (arsenic speciation, chromium [VI], and perchlorate). These groundwater-quality data, along with data from quality-control samples, are tabulated in this report and in an associated data release. Some data from environmental samples collected in 2013 and quality-control samples collected in 2014 also are included in the associated data release; these data are associated with networks described in this report and have not been published previously.

Data Series

State of the data: Assessing the FAIRness of USGS data

In response to recent shifts towards open science that emphasize transparency, reproducibility, and access to research data, the US Geological Survey (USGS) conducted a study to assess the degree to which USGS data assets meet the FAIR data principles (Findable, Accessible, Interoperable, and Reusable). The USGS designed and applied a methodology for quantitative analysis of FAIR characteristics. A new rubric was derived from a crosswalk of existing FAIR evaluation frameworks and customized for the USGS. The rubric, consisting of 62 yes/no questions, was applied to 392 metadata records of USGS data products published between 1987 and 2022. Results were analyzed to show which FAIR characteristics were most and least present in the metadata and how these scores changed after the implementation of data policy requirements in 2016. Aggregated scores showed specific areas of strength and needed improvements. The greatest increases in FAIR scores over time were for elements that were required by new data policies, especially in the ‘Findable’ category. Based on the results, this paper presents strategies to further improve USGS alignment with FAIR. The suggested strategies are organized in four key areas: USGS data repository characteristics, training and communities of practice, data management policy considerations, and metadata standards, tools, and best practices.

Data Science Journal

Modern pollen-assemblages data from small lakes paired with local forest-composition data in northeastern United States

For the past century, pollen analysis has served as a primary tool for inferring past changes in vegetation composition and structure (Birks et al. 2016, Edwards et al. 2017). Pollen-based inferences are supported by empirical studies comparing modern pollen assemblages with modern vegetation composition. In one approach, pollen abundances (usually percentages) for individual taxa are compared directly with quantitative estimates of abundance in surrounding vegetation (Jackson 1994, Davis 2000). This approach has been applied most frequently using spatially extensive but coarse-scale forest inventory data (Webb et al. 1981, Bradshaw and Webb 1985, Prentice & Webb 1986, Prentice et al. 1987, Paciorek & McLachlan 2009, Dawson et al. 2016, Kujawa et al. 2016). In these studies, forest composition cannot usually be estimated accurately within a 1- to 10 km radius of the individual sites owing to limited spatial density of forest inventory data. A few studies have compared vegetation composition within 50-100 m of pollen-sampling sites, but in these cases the pollen is from forest-floor assemblages (Bradshaw 1981, Jackson & Wong 1994, Jackson & Kearsley 1998) or from small forest hollows (Calcote 1995, 1998, Parshall & Calcote 2001). Largely lacking are pollen assemblage data from lake sediments paired with local forest composition, measured within 100 to 1000 m of the lake margins (Jackson 1990). This absence represents a substantial gap in ability to understand and model pollen-vegetation relationships, because lakes are the primary source of fossil-pollen sequences worldwide, and because the leptokurtic nature of pollen dispersal ensures that local vegetation has an important effect on pollen composition in sediments (Jackson 1994, Sugita 1994, 2007a, 2007b, Jackson & Lyford 1999). Here, I present a data set pairing modern pollen assemblages from 33 small lakes in the forested northeastern United States (Fig. 1) with forest composition data measured within 20, 50, 100, 500, and 1000 metres of the lake margins. This data set incorporates most of the sites used in Jackson (1990), adding 16 new sites and delivering the vegetation data by species in absolute units (i.e., total basal area), which allows various weightings and transformations to be applied. The data set should be of value to paleoecologists and forest ecologists in understanding, modeling, and validating the pollen-vegetation relationships that are at the heart of paleoecological inference.

Northeastern United States

Topographic mapping data semantics through data conversion and enhancement

This paper presents research on the semantics of topographic data for triples and ontologies to blend the capabilities of the Semantic Web and The National Map of the U.S. Geological Survey. Automated conversion of relational topographic data of several geographic sample areas to the triple data model standard resulted in relatively poor semantic associations. Further research employed vocabularies of feature type and spatial relation terms. A user interface was designed to model the capture of non-standard terms relevant to public users and to map those terms to existing data models of The National Map through the use of ontology. Server access for the study area triple stores was made publicly available, illustrating how the development of linked data may transform institutional policies to open government data resources to the public. This paper presents these data conversion and research techniques that were tested as open linked data concepts leveraged through a user-centered interface and open USGS server access to the public.

Book chapter

MTH5: An archive and exchangeable data format for magnetotelluric time series data

Magnetotellurics (MT) is a passive electromagnetic geophysical method that measures variations in subsurface electrical resistivity. MT data are collected in the time domain and processed in the frequency domain to produce estimates of a transfer function representing the Earth’s electrical structure. Unfortunately, the MT community lacks metadata and data standards for time series data. As the community grows and findability, accessibility, interoperability, and reuse of digital assets (FAIR) data principles are enforced by government and funding agencies, a standard is needed for time series data. Presented here is a hierarchical data format (MTH5) that is logically formatted to how MT data are collected. Open-source Python packages are also described to read, write, and manipulate MTH5 files. These include a package to deal with metadata ( mt_metadata ) based on standards developed by the Working Group for Magnetotelluric Data Handling and Software assembled by the Incorporated Research Institutions for Seismology (IRIS), and mth5 : a package to interact with MTH5 files that uses mt_metadata . Example code and workflows are presented.

Computers & Geosciences

An analysis of water data systems to inform the Open Water Data Initiative

Improving access to data and fostering open exchange of water information is foundational to solving water resources issues. In this vein, the Department of the Interior's Assistant Secretary for Water and Science put forward the charge to undertake an Open Water Data Initiative (OWDI) that would prioritize and accelerate work toward better water data infrastructure. The goal of the OWDI is to build out the Open Water Web (OWW). We therefore considered the OWW in terms of four conceptual functions: water data cataloging, water data as a service, enriching water data, and community for water data. To describe the current state of the OWW and identify areas needing improvement, we conducted an analysis of existing systems using a standard model for describing distributed systems and their business requirements. Our analysis considered three OWDI-focused use cases—flooding, drought, and contaminant transport—and then examined the landscape of other existing applications that support the Open Water Web. The analysis, which includes a discussion of observed successful practices of cataloging, serving, enriching, and building community around water resources data, demonstrates that we have made significant progress toward the needed infrastructure, although challenges remain. The further development of the OWW can be greatly informed by the interpretation and findings of our analysis.

Journal of the American Water Resources Associatio

Integration of paleoseismic data from multiple sites to develop an objective earthquake chronology: Application to the Weber segment of the Wasatch fault zone, Utah

We present a method to evaluate and integrate paleoseismic data from multiple sites into a single, objective measure of earthquake timing and recurrence on discrete segments of active faults. We apply this method to the Weber segment (WS) of the Wasatch fault zone using data from four fault-trench studies completed between 1981 and 2009. After systematically reevaluating the stratigraphic and chronologic data from each trench site, we constructed time-stratigraphic OxCal models that yield site probability density functions (PDFs) of the times of individual earthquakes. We next qualitatively correlated the site PDFs into a segment-wide earthquake chronology, which is supported by overlapping site PDFs, large per-event displacements, and prominent segment boundaries. For each segment-wide earthquake, we computed the product of the site PDF probabilities in common time bins, which emphasizes the overlap in the site earthquake times, and gives more weight to the narrowest, best-defined PDFs. The product method yields smaller earthquake-timing uncertainties compared to taking the mean of the site PDFs, but is best suited to earthquakes constrained by broad, overlapping site PDFs. We calculated segment-wide earthquake recurrence intervals and uncertainties using a Monte Carlo model. Five surface-faulting earthquakes occurred on the WS at about 5.9, 4.5, 3.1, 1.1, and 0.6 ka. With the exception of the 1.1-ka event, we used the product method to define the earthquake times. The revised WS chronology yields a mean recurrence interval of 1.3 kyr (0.7–1.9-kyr estimated two-sigma [2δ] range based on interevent recurrence). These data help clarify the paleoearthquake history of the WS, including the important question of the timing and rupture extent of the most recent earthquake, and are essential to the improvement of earthquake-probability assessments for the Wasatch Front region.

Utah

Progress on water data integration and distribution: a summary of select U.S. Geological Survey data systems

Critical water-resources issues ranging from flood response to water scarcity make access to integrated water information, services, tools, and models essential. Since 1995 when the first water data web pages went online, the U.S. Geological Survey has been at the forefront of water data distribution and integration. Today, real-time and historical streamflow observations are available via web pages and a variety of web service interfaces. The Survey has built partnerships with Federal and State agencies to integrate hydrologic data providing continuous observations of surface and groundwater, temporally discrete water quality data, groundwater well logs, aquatic biology data, water availability and use information, and tools to help characterize the landscape for modeling. In this paper, we summarize the status and design patterns implemented for selected data systems. We describe how these systems contribute to a U.S. Federal Open Water Data Initiative and present some gaps and lessons learned that apply to global hydroinformatics data infrastructure.

Journal of Hydroinformatics

Data Delivery and Mapping Over the Web: National Water-Quality Assessment Data Warehouse

The U.S. Geological Survey began its National Water-Quality Assessment (NAWQA) Program in 1991, systematically collecting chemical, biological, and physical water-quality data from study units (basins) across the Nation. In 1999, the NAWQA Program developed a data warehouse to better facilitate national and regional analysis of data from 36 study units started in 1991 and 1994. Data from 15 study units started in 1997 were added to the warehouse in 2001. The warehouse currently contains and links the following data: -- Chemical concentrations in water, sediment, and aquatic-organism tissues and related quality-control data from the USGS National Water Information System (NWIS), -- Biological data for stream-habitat and ecological-community data on fish, algae, and benthic invertebrates, -- Site, well, and basin information associated with thousands of descriptive variables derived from spatial analysis, like land use, soil, and population density, and -- Daily streamflow and temperature information from NWIS for selected sampling sites.

Fact Sheet