USGS ScienceSearch

Geology topics

David I. Donato

Publications and source records attributed to David I. Donato.

14 recordsLinked to original sources

Semi-automated methods to develop a unified geographic information system dataset

Geospatial data describing the topography, natural features, human-built features, and land uses of a particular area or region can come from independent data providers and, therefore, vary in format, data encoding, and geographic coverage. Because of the complexity of the processes and procedures required for unifying these heterogeneous data into a dataset with consistent format, encoding, and coverage, fully automated procedures for data unification do not exist. However, a combination of manual and automated procedures—semi-automated methods—can substantially reduce the time required for data unification while improving accuracy. This report presents three semi-automated data-unification methods in detail. Although these methods are not new in principle, their details are the result of original development work, and they serve as examples that can be reused, adapted, or generalized to provide head starts to future data-unification projects. The format of this report can be used and refined to encourage the publication of future reports and more widespread sharing of semi-automated methods.

Techniques and Methods

Efficient processing of two-dimensional arrays with C or C++

Because fast and efficient serial processing of raster-graphic images and other two-dimensional arrays is a requirement in land-change modeling and other applications, the effects of 10 factors on the runtimes for processing two-dimensional arrays with C and C++ are evaluated in a comparative factorial study. This study’s factors include the choice among three C or C++ source-code techniques for array processing; the choice of Microsoft Windows 7 or a Linux operating system; the choice of 4-byte or 8-byte array elements and indexes; and the choice of 32-bit or 64-bit memory addressing. This study demonstrates how programmer choices can reduce runtimes by 75 percent or more, even after compiler optimizations. Ten points of practical advice for faster processing of two-dimensional arrays are offered to C and C++ programmers. Further study and the development of a C and C++ software test suite are recommended. Key words : array processing, C, C++, compiler, computational speed, land-change modeling, raster-graphic image, two-dimensional array, software efficiency

Techniques and Methods

Coding conventions and principles for a National Land-Change Modeling Framework

This report establishes specific rules for writing computer source code for use with the National Land-Change Modeling Framework (NLCMF). These specific rules consist of conventions and principles for writing code primarily in the C and C++ programming languages. Collectively, these coding conventions and coding principles create an NLCMF programming style. In addition to detailed naming conventions, this report provides general coding conventions and principles intended to facilitate the development of high-performance software implemented with code that is extensible, flexible, and interoperable. Conventions for developing modular code are explained in general terms and also enabled and demonstrated through the appended templates for C++ base source-code and header files. The NLCMF limited-extern approach to module structure, code inclusion, and cross-module access to data is both explained in the text and then illustrated through the module templates. Advice on the use of global variables is provided.

Techniques and Methods

Simple, efficient allocation of modelling runs on heterogeneous clusters with MPI

In scientific modelling and computation, the choice of an appropriate method for allocating tasks for parallel processing depends on the computational setting and on the nature of the computation. The allocation of independent but similar computational tasks, such as modelling runs or Monte Carlo trials, among the nodes of a heterogeneous computational cluster is a special case that has not been specifically evaluated previously. A simulation study shows that a method of on-demand (that is, worker-initiated) pulling from a bag of tasks in this case leads to reliably short makespans for computational jobs despite heterogeneity both within and between cluster nodes. A simple reference implementation in the C programming language with the Message Passing Interface (MPI) is provided.

Environmental Modelling and Software

Building unified geospatial data for land-change modeling—A case study in the area of Richmond, Virginia

An effort to build a unified collection of geospatial data for use in land-change modeling (LCM) led to new insights into the requirements and challenges of building an LCM data infrastructure. A case study of data compilation and unification for the Richmond, Va., Metropolitan Statistical Area (MSA) delineated the problems of combining and unifying heterogeneous data from many independent localities such as counties and cities. The study also produced conclusions and recommendations for use by the national LCM community, emphasizing the critical need for simple, practical data standards and conventions for use by localities. This report contributes an uncopyrighted core glossary and a much needed operational definition of data unification.

Virginia

Disparity between state fish consumption advisory systems for methylmercury and US Environmental Protection Agency recommendations: A case study of the South Central United States

Fish consumption advisories are used to inform citizens in the United States about noncommercial game fish with hazardous levels of methylmercury (MeHg). The US Environmental Protection Agency (USEPA) suggests issuing a fish consumption advisory when concentrations of MeHg in fish exceed a human health screening value of 300 ng/g. However, states have authority to develop their own systems for issuing fish consumption advisories for MeHg. Five states in the south central United States (Arkansas, Louisiana, Mississippi, Oklahoma, and Texas) issue advisories for the general human population when concentrations of MeHg exceed 700 ng/g to 1000 ng/g. The objective of the present study was to estimate the increase in fish consumption advisories that would occur if these states followed USEPA recommendations. The authors used the National Descriptive Model of Mercury in Fish to estimate the mercury concentrations in 5 size categories of largemouth bass–equivalent fish at 766 lentic and lotic sites within the 5 states. The authors found that states in this region have not issued site‐specific fish consumption advisories for most of the water bodies that would have such advisories if USEPA recommendations were followed. One outcome of the present study may be to stimulate discussion between scientists and policy makers at the federal and state levels about appropriate screening values to protect the public from the health hazards of consuming MeHg‐contaminated game fish.

Arkansas, Louisiana, Mississippi, Oklahoma, Texas

Historic and forecasted population and land-cover change in eastern North Carolina, 1992-2030

The Southeast Regional Partnership for Planning and Sustainability (SERPPAS) was formed in 2005 as a partnership between the Department of Defense (DOD) and State and Federal agencies to promote better collaboration in making resource-use decisions. In support of this goal, the U.S. Geological Survey (USGS) conducted a study to evaluate historic population growth and land-cover change, and to model future change, for the 13-county SERPPAS study area in southeastern North Carolina (fig. 1). Improved understanding of trends in land-cover change and the ability to forecast land-cover change that is consistent with these trends will be a key component of efforts to accommodate local military-mission imperatives while also promoting sustainable economic growth throughout the 13-county study area. The study had three principal objectives: 1. Evaluate historic changes in population and land cover for the period 1992–2006 using both previously existing as well as newly generated land-cover data. 2. Develop models to forecast future change in land cover using the data gathered in objective 1 in conjunction with ancillary data on the suitability of the various sub-areas within the study area for low- and high-intensity urban development. 3. Deliver these results—including an executive-level briefing and a USGS technical report—to DOD, other project cooperators, and local counties in hard-copy and digital formats and via the Web through a map-based data viewer. This report provides a general overview of the study and is intended for general distribution to non-technical audiences.

North Carolina

Effects of mercury deposition and coniferous forests on the mercury contamination of fish in the south central United States

Mercury (Hg) is a toxic metal that is found in aquatic food webs and is hazardous to human and wildlife health. We examined the relationship between Hg deposition, land coverage by coniferous and deciduous forests, and average Hg concentrations in largemouth bass (Micropterus salmoides)-equivalent fish (LMBE) in 14 ecoregions located within all or part of six states in the South Central U.S. In 11 ecoregions, the average Hg concentrations in 35.6-cm total length LMBE were above 300 ng/g, the threshold concentration of Hg recommended by the U.S. Environmental Protection Agency for the issuance of fish consumption advisories. Percent land coverage by coniferous forests within ecoregions had a significant linear relationship with average Hg concentrations in LMBE while percent land coverage by deciduous forests did not. Eighty percent of the variance in average Hg concentrations in LMBE between ecoregions could be accounted for by estimated Hg deposition after adjusting for the effects of coniferous forests. Here we show for the first time that fish from ecoregions with high atmospheric Hg pollution and coniferous forest coverage pose a significant hazard to human health. Our study suggests that models that use Hg deposition to predict Hg concentrations in fish could be improved by including the effects of coniferous forests on Hg deposition.

Arkansas;Louisiana;Mississippi;Oklahoma;Tennessee;

Computing ordinary least-squares parameter estimates for the National Descriptive Model of Mercury in Fish

A specialized technique is used to compute weighted ordinary least-squares (OLS) estimates of the parameters of the National Descriptive Model of Mercury in Fish (NDMMF) in less time using less computer memory than general methods. The characteristics of the NDMMF allow the two products X'X and X'y in the normal equations to be filled out in a second or two of computer time during a single pass through the N data observations. As a result, the matrix X does not have to be stored in computer memory and the computationally expensive matrix multiplications generally required to produce X'X and X'y do not have to be carried out. The normal equations may then be solved to determine the best-fit parameters in the OLS sense. The computational solution based on this specialized technique requires O(8 p 2 +16 p ) bytes of computer memory for p parameters on a machine with 8-byte double-precision numbers. This publication includes a reference implementation of this technique and a Gaussian-elimination solver in preliminary custom software.

Techniques and Methods

Computing maximum-likelihood estimates for parameters of the National Descriptive Model of Mercury in Fish

This report presents the mathematical expressions and the computational techniques required to compute maximum-likelihood estimates for the parameters of the National Descriptive Model of Mercury in Fish (NDMMF), a statistical model used to predict the concentration of methylmercury in fish tissue. The expressions and techniques reported here were prepared to support the development of custom software capable of computing NDMMF parameter estimates more quickly and using less computer memory than is currently possible with available general-purpose statistical software. Computation of maximum-likelihood estimates for the NDMMF by numerical solution of a system of simultaneous equations through repeated Newton-Raphson iterations is described. This report explains the derivation of the mathematical expressions required for computational parameter estimation in sufficient detail to facilitate future derivations for any revised versions of the NDMMF that may be developed.

Open-File Report

Designing and implementing a regional urban modeling system using the SLEUTH cellular urban model

This paper presents a fine-scale (30 meter resolution) regional land cover modeling system, based on the SLEUTH cellular automata model, that was developed for a 257000 km 2 area comprising the Chesapeake Bay drainage basin in the eastern United States. As part of this effort, we developed a new version of the SLEUTH model (SLEUTH-3r), which introduces new functionality and fit metrics that substantially increase the performance and applicability of the model. In addition, we developed methods that expand the capability of SLEUTH to incorporate economic, cultural and policy information, opening up new avenues for the integration of SLEUTH with other land-change models. SLEUTH-3r is also more computationally efficient (by a factor of 5) and uses less memory (reduced 65%) than the original software. With the new version of SLEUTH, we were able to achieve high accuracies at both the aggregate level of 15 sub-regional modeling units and at finer scales. We present forecasts to 2030 of urban development under a current trends scenario across the entire Chesapeake Bay drainage basin, and three alternative scenarios for a sub-region within the Chesapeake Bay watershed to illustrate the new ability of SLEUTH-3r to generate forecasts across a broad range of conditions.

Computers, Environment and Urban Systems

Fast, Inclusive Searches for Geographic Names Using Digraphs

An algorithm specifies how to quickly identify names that approximately match any specified name when searching a list or database of geographic names. Based on comparisons of the digraphs (ordered letter pairs) contained in geographic names, this algorithmic technique identifies approximately matching names by applying an artificial but useful measure of name similarity. A digraph index enables computer name searches that are carried out using this technique to be fast enough for deployment in a Web application. This technique, which is a member of the class of n-gram algorithms, is related to, but distinct from, the soundex, PHONIX, and metaphone phonetic algorithms. Despite this technique's tendency to return some counterintuitive approximate matches, it is an effective aid for fast, inclusive searches for geographic names when the exact name sought, or its correct spelling, is unknown.

Techniques and Methods

Secure Web-Site Access with Tickets and Message-Dependent Digests

Although there are various methods for restricting access to documents stored on a World Wide Web (WWW) site (a Web site), none of the widely used methods is completely suitable for restricting access to Web applications hosted on an otherwise publicly accessible Web site. A new technique, however, provides a mix of features well suited for restricting Web-site or Web-application access to authorized users, including the following: secure user authentication, tamper-resistant sessions, simple access to user state variables by server-side applications, and clean session terminations. This technique, called message-dependent digests with tickets, or MDDT, maintains secure user sessions by passing single-use nonces (tickets) and message-dependent digests of user credentials back and forth between client and server. Appendix 2 provides a working implementation of MDDT with PHP server-side code and JavaScript client-side code.

Techniques and Methods

EMMMA: A web-based system for environmental mercury mapping, modeling, and analysis

Mercury in our environment - in our air, water, soil, and especially our food - poses significant hazards to human health, particularly for developing fetuses and young children. Because of the importance of this issue and the length of time it has been studied, large and complex data sets of mercury concentrations in various media and associated ancillary data have been generated by many Federal, State, Tribal, and local agencies. To facilitate efficient and effective use of these data in managing and mitigating human and wildlife exposure to mercury, the U.S. Geological Survey (USGS) and the National Institute of Environmental Health Sciences have developed a website for visualizing and studying the distribution of mercury in our environment. The Environmental Mercury Mapping, Modeling, and Analysis (EMMMA) website (http://emmma.usgs.gov) provides health and environmental researchers, managers, and other decision-makers the ability to: 1) Interactively view and access a nationwide collection of environmental mercury data (fish tissue, atmospheric emissions and deposition, stream sediments, soils, and coal) and mercuryrelated data (mine locations); 2) Interactively view and access predictions of the National Descriptive Model of Mercury in Fish (NDMMF) at 4,976 sites and 6,829 sampling events (events are unique combinations of site and sampling date) across the United States; and 3) Use interactive mapping and graphing capabilities to visualize spatial and temporal trends and study relationships between mercury and other variables.

Open-File Report