USGS ScienceSearch

USGS · 70231381

Classifying Worldwide Standardized Seismograph Network records using a simple convolution neural network

Abstract

The U.S. Geological Survey (USGS) maintains an archive of 189,180 digitized scans of analog seismic records from the World‐Wide Standardized Seismograph Network (WWSSN). Although these scans have been made public, the archive is too large to manually review, and few researchers have utilized large numbers of these records. To facilitate further research using this historical dataset, we develop a simple convolutional neural network (CNN) that rapidly (∼4.75 s/film chip) classifies scanned film chip images (called “chips,” because they are individually cut segments of 70 mm film) into four categories of “interestingness” to earthquake seismologists based on the presence of earthquakes and other seismic signals in the record: “no interest,” “little interest,” “interest,” and “high interest.” The CNN, dubbed “Seismic Analog Record Network” (SARNet), can identify four types of seismic traces (“no events,” “minor events,” “major events,” and “errors”) in 200 × 200 pixel subcrops with an accuracy of 92% using a confidence threshold of 85%. SARNet then converts 100 random subcrops from each film chip into the overall classification of interestingness. In this task, SARNet performed as well as expert human classifiers in determining the film chip’s overall interest grade. Applying SARNet to 34,000 film chips in the WWSSN archive found that 21% of the images were of “high interest” and had an “indeterminate” rate of only 4%. Thus, the need for the manual review of images was reduced by 79%. Sorting of film chips derived from SARNet will expedite further exploration of the archive of digitized analog seismic records stored at the USGS.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Nagle Nagle-McNaughton, Adam T. Ringler, Robert E. Anthony, Alexis Casondra Bianca Alejandro, David C. Wilson, Justin Thomas Wilgus. 2022-05-09. Classifying Worldwide Standardized Seismograph Network records using a simple convolution neural network. https://doi.org/10.1785/0220220017

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related USGS reports

Digitizer Suite: The Albuquerque Seismological Laboratory Digitizer Testing Suite

Laboratory testing of digitizers and seismometers helps ensure that prior to deployment the instrumentation can produce high quality data and is operating within specifications. In this work we detail the software package called: the Albuquerque Seismological Laboratory (ASL) Digitizer Test Suite. This Java software package provides several algorithms to verify various performance parameters of digitizers commonly used for recording analog seismic instruments. The goal of these tests is not to be exhaustive, but to identify common failures that could compromise the integrity of seismic data being recorded on the digitizer. For example, Sandia National Laboratories (e.g., Slad and Merchant, 2018) routinely do comprehensive testing of digitizers for various monitoring missions. While these tests reports are valuable for comprehensively characterizing a recording system, it would be resource intensive to conduct such tests on every seismic recorder used in a network. We focus on tests that include ways to estimate the sensitivity, timing, self-noise, and clip-level of the digitizer, as well as the fidelity of the signal being recorded. The software is publicly available and provides a way for the community to verify the integrity of a digitizer using a minimum amount of outside equipment.

Seismological Research Letters

The digital archivist: Automating legacy macroseismic data processing using large language models

Macroseismic data are a key resource to investigate shaking and damage from preinstrumental and early instrumental eras. However, data are often stored as inconsistently formatted reports describing observed shaking and damage, making manually parsing and interpreting accounts labor‐intensive. We introduce a novel workflow using Google’s Gemini 2.5 Pro large language model (LLM) to automate the extraction and structuring of macroseismic observations from summary reports. We apply this workflow to the 22 March 1957 M 5.3 Daly City, California, earthquake as a case study. We used Gemini to extract addresses, originally assigned modified Mercalli intensity values, and descriptions from each report. To address coordinate precision limits, addresses were geocoded via Google’s Geocoding application programming interface. This workflow yielded over 2300 geocoded intensity reports for the Daly City earthquake. We use the geocoded accounts, with the original report intensity assignments, to develop a shaking intensity map that in some respects rivals modern Did You Feel It? Maps. We also extract and present data for the 9 February 1971 M L 6.7 Sylmar, California, earthquake. Our results demonstrate the potential of LLMs for reliably extracting and analyzing large, unstructured macroseismic datasets. LLMs offer a scalable solution for rapidly digitizing macroseismic archives, enabling their broader use to constrain ground‐motion models in modern seismic hazard analysis and to improve our understanding of site effects in urban areas. The concepts explored here may also be applied to the handling of other legacy seismological and earth science data.

Seismological Research Letters