Search USGSSearch

SEARCH · Search USGS

Results for “Computational Statistics and Data Analysis”

Search indexed USGS publications on groundwater, aquifers, geologic maps, mineral resources and earthquakes. Explore source records by subject and place.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 181 records · Page 10Linked to original sources

Volcanogenic Massive Sulfide Deposits of the World - Database and Grade and Tonnage Models

Grade and tonnage models are useful in quantitative mineral-resource assessments. The models and database presented in this report are an update of earlier publications about volcanogenic massive sulfide (VMS) deposits. These VMS deposits include what were formerly classified as kuroko, Cyprus, and Besshi deposits. The update was necessary because of new information about some deposits, changes in information in some deposits, such as grades, tonnages, or ages, revised locations of some deposits, and reclassification of subtypes. In this report we have added new VMS deposits and removed a few incorrectly classified deposits. This global compilation of VMS deposits contains 1,090 deposits; however, it was not our intent to include every known deposit in the world. The data was recently used for mineral-deposit density models (Mosier and others, 2007; Singer, 2008). In this paper, 867 deposits were used to construct revised grade and tonnage models. Our new models are based on a reclassification of deposits based on host lithologies: Felsic, Bimodal-Mafic, and Mafic volcanogenic massive sulfide deposits. Mineral-deposit models are important in exploration planning and quantitative resource assessments for two reasons: (1) grades and tonnages among deposit types vary significantly, and (2) deposits of different types occur in distinct geologic settings that can be identified from geologic maps. Mineral-deposit models combine the diverse geoscience information on geology, mineral occurrences, geophysics, and geochemistry used in resource assessments and mineral exploration. Globally based deposit models allow recognition of important features and demonstrate how common different features are. Well-designed deposit models allow geologists to deduce possible mineral-deposit types in a given geologic environment and economists to determine the possible economic viability of these resources. Thus, mineral-deposit models play a central role in presenting geoscience information in a useful form to policy makers. The foundation of mineral-deposit models is information about known deposits. The purpose of this publication is to present the latest geologic information and newly developed grade and tonnage models for VMS deposits in digital form. This publication contains computer files with information on VMS deposits from around the world. It also presents new grade and tonnage models for three subtypes of VMS deposits and a text file allowing locations of all deposits to be plotted in geographic information system (GIS) programs. The data are presented in FileMaker Pro and text files to make the information available to a wider audience. The value of this information and any derived analyses depends critically on the consistent manner of data gathering. For this reason, we first discuss the rules used in this compilation. Next, we provide new grade and tonnage models and analysis of the information in the file. Finally, the fields of the data file are explained. Appendix A gives the summary statistics for the new grade-tonnage models and Appendix B displays the country codes used in the database.

Open-File Report

A generalized Grubbs-Beck test statistic for detecting multiple potentially influential low outliers in flood series

he Grubbs-Beck test is recommended by the federal guidelines for detection of low outliers in flood flow frequency computation in the United States. This paper presents a generalization of the Grubbs-Beck test for normal data (similar to the Rosner (1983) test; see also Spencer and McCuen (1996)) that can provide a consistent standard for identifying multiple potentially influential low flows. In cases where low outliers have been identified, they can be represented as “less-than” values, and a frequency distribution can be developed using censored-data statistical techniques, such as the Expected Moments Algorithm. This approach can improve the fit of the right-hand tail of a frequency distribution and provide protection from lack-of-fit due to unimportant but potentially influential low flows (PILFs) in a flood series, thus making the flood frequency analysis procedure more robust.

Water Resources Research

Preliminary synthesis and assessment of environmental flows in the middle Verde River watershed, Arizona

A 3-year study was undertaken to evaluate the suitability of the available modeling tools for characterizing environmental flows in the middle Verde River watershed of central Arizona, describe riparian vegetation throughout the watershed, and estimate sediment mobilization in the river. Existing data on fish and macroinvertebrates were analyzed in relation to basin characteristics, flow regimes, and microhabitat, and a pilot study was conducted that sampled fish and macroinvertebrates and the microhabitats in which they were found. The sampling for the pilot study took place at five different locations in the middle Verde River watershed. This report presents the results of this 3-year study. The Northern Arizona Groundwater Flow Model (NARGFM) was found to be capable of predicting long-term changes caused by alteration of regional recharge (such as may result from climate variability) and groundwater pumping in gaining, losing, and dry reaches of the major streams in the middle Verde River watershed. Over the period 1910 to 2006, the model simulated an increase in dry reaches, a small increase in reaches losing discharge to the groundwater aquifer, and a concurrent decrease in reaches gaining discharge from groundwater. Although evaluations of the suitability of using the NARGFM and Basin Characteristic Model to characterize various streamflow intervals showed that smallerscale basin monthly runoff could be estimated adequately at locations of interest, monthly stream-flow estimates were found unsatisfactory for determining environmental flows. Orthoimagery and Moderate Resolution Imaging Spectroradiometer data were used to quantify stream and riparian vegetation properties related to biotic habitat. The relative abundance of riparian vegetation varied along the main channel of the Verde River. As would be expected, more upland plant species and fewer lowland species were found in the upper-middle section compared to the lower-middle section, and vice-versa. Vegetation changes within the upper-middle and lower-middle reaches are related to differences in climate and hydrology. In general, the riparian vegetation of the middle Verde River watershed is that of a healthy ecosystem’s mixed age, mixed patch structure, likely a result of the mostly unaltered disturbance regime. The frequency of in-river hydrogeomorphic features (pool, riffle, run) varied along the middle Verde River channel. There was a greater abundance of riffle habitat in the upper-middle reach; the lower-middle reach included more pool habitat. The Oak Creek tributary was more homogenous in geomorphic stream habitat composition than West Clear Creek, where runs dominated the upper reaches and pools dominated many of the lower reaches. On the basis of the period of record and discharges recorded at 15-minute intervals, five flows were found to reach the gravel-transport threshold. Sediment mobilization computed with flows averaged over daily time steps yielded just three flows that reached the gravel-transport threshold, and monthly averaged flows yielded none. In the middle Verde River watershed, 15-minute data should be used when possible to evaluate sediment transport in the river system. Data from more than 300 fish surveys conducted from 1992 to 2011 were analyzed using two schemes, one that divided the river into five reaches based on basin characteristics, and a second that divided the river into five reaches based on degree of flow alteration (specifically, diversions). Fish community metrics and assemblage data were used to analyze patterns of species composition and abundance in the two approaches. Overall, native and non-native species were regularly interacting and probably competing for similar resources. Fish abundances were also analyzed in response to floods and other flow metrics. Although the data are limited, native fish abundances increased more rapidly than non-native fish abundances in response to large floods. The basin-characteristic reach analysis showed native fish in greater abundance in the upper-middle reaches of the Verde River watershed and generally decreasing with downstream distance. The median relative abundance of native fish decreased by 50 percent from reach 1 to reach 5. Using the reach scheme based on degree of flow alteration, nondiverted reaches were found to have a greater abundance of native fish than diverted reaches. In heavily diverted reaches, non-native species outnumbered native species. Fish metrics and stream-flow metrics for the 30, 90, and 365-day periods before collection were computed and the results analyzed statistically. Only abundance of all fish species was associated with the 30-day flow metrics. The 90-day flow metrics were generally positively associated with fish metrics, whereas the 365-day flow metrics had more negative correlations. In particular, significant relations were found between fish metrics and the magnitude and frequency of high flows, including maximum monthly flow, median annual number of high-flow events, and median annual maximum streamflow. Native sucker (Catostomidae) populations tended to decrease in periods of extended base flow, and fish in the non-native sunfish family (Centrarchidae) decreased in periods of flashy, high magnitude flows. A pilot study surveyed fish at five locations in the upper part of the middle Verde River watershed as a means to measure microhabitat availability and quantify native and non-native fish use of that available microhabitat. Results indicated that native and non-native species exhibit some clear differences in microhabitat use. Although at least some native and non-native fish were found in each velocity, depth, and substrate category, preferential microhabitat use was common. On a percentage basis, non-native species had a strong preference for slow-moving and deeper water with silt and sand substrate, with a secondary preference for faster moving and very shallow water and a coarse gravel substrate. Native species showed a general preference for somewhat faster, moderate depth water over coarse gravel and had no clear secondary preference. Macroinvertebrate-variables index period, high-flow year, and collection location (upper-middle Verde River, lowermiddle Verde River, or Verde River tributaries) were found to be important explanatory variables in differentiating among community metrics. Overall richness (number of unique taxa), Shannon’s diversity index, and the percent of the most dominant taxa were all highly correlated, but their response to each macroinvertebrate variable was different. The percentage of mayfly (order Ephemeroptera) taxa was significantly higher in Oak Creek and the upper-middle and lower-middle Verde River reaches, locations which have higher flows and more urbanization than other reaches. When community metrics were related to hydrologic metrics, caddisfly (order Trichoptera) populations appeared to increase and mayfly populations to decrease in response to less flashy and more stable streamflows. Conversely, caddisfly populations appeared to decrease and mayfly populations to increase in response to greater flow variability. Six locations along the Verde River were sampled for macroinvertebrates as part of a pilot study associated with this report—(1) below Granite Creek, (2) near Campbell Ranch, (3) at the U.S. Geological Survey Paulden gage, (4) at the Perkinsville Bridge, (5) at the USGS Clarkdale gage, and (6) near the Reitz Ranch property. A nonmetric multidimensional scaling ordination of macroinvertebrate assemblages showed that the Verde River below Granite Creek site was different from the five other sites and that the Perkinsville Bridge and near Reitz Ranch samples had similar community structure. The near Campbell Ranch and Paulden gage locations had similar microhabitat characteristics, with the exception of riparian cover, yet the assemblage structure was very different. The different community composition at Verde River below Granite Creek was likely due to it having the smallest substrate sizes, lowest velocities, shallowest depths, and most riparian cover of the six sites.

Arizona

General water-quality conditions, long-term trends, and network analysis at selected sites within the Ambient Water-Quality Monitoring Network in Missouri, water years 1993–2017

The U.S. Geological Survey, in cooperation with the Missouri Department of Natural Resources, collects data pertaining to the surface-water resources of Missouri. Established in 1964, the Ambient Water-Quality Monitoring Network (AWQMN) consisted of 69 sites in 2017. Two additional sites from the National Water-Quality Program are included with the AWQMN sites for the analyses in this report. The sites are sampled typically from 2 to 12 times per year for physical properties, total suspended solids, nutrients, fecal indicator bacteria, and trace elements. The period of analysis for this study was from 1993 through 2017 and data analysis included 71 sites and 15 water-quality constituents plus discharge. Data analysis involved retrieving the data, conditioning the data for analysis, analyzing the data for trends, and analyzing the monitoring network to determine if potential data gaps or data redundancies exist in the network. Results from these analyses can be used to help manage the monitoring network into the future. Water-quality data were analyzed using several software packages to provide graphical and statistical information for interpretation of trends in the data at selected sites. Discharge data at selected sites were analyzed to determine the general trends during the analysis period and how the water-quality samples represented the range of daily mean discharges at each site. Water-quality data also were analyzed at selected sites to determine the relative sensitivity of selected sites and constituents to changes in data collection frequency. Trend analysis at selected sites using a simulated reduction in sampling frequency was completed to compare to trends obtained using monthly data to determine the potential degradation in the ability of determining trends from a reduced sampling frequency. The viability of using estimated discharge to evaluate long-term trends for sites with no continuous discharge was investigated. Data from sites were statistically compared in groups to determine the relative similarity (or difference) between sites for each water-quality constituent to identify potentially redundant sites in the monitoring network. Discharge-weighted long-term trends during 1993 through 2017 were analyzed for 15 water-quality constituents at 58 sites and results indicated there were significant single- or two-period trends in about 17 percent of the analyses. Some trends indicated improvement and some trends indicated deterioration of the general water quality at some sites in the AWQMN. No trend was indicated in about 31 percent of the analyses. The constituents pH, specific conductance, and total phosphorus showed the most frequent significant trends, and each of the 15 constituents examined had a significant trend at one or more sites. A total of 42 sites indicated at least 1 constituent with a significant single- or two-period trend, and 10 sites indicated 6 or more significant trends. Potential data gaps identified for computing discharge-weighted long-term trends in the monitoring network included the lack of collection of continuous discharge at 23 sites, insufficient sampling frequency for some constituents (dissolved chloride and total and dissolved lead and zinc) at some sites, insufficient temporal sample distribution (lack of at least one sample in each season per year) at some sites, and insufficient sampling frequency for some highly censored constituents (nutrients and total and dissolved lead and zinc) at some sites. Potential data gaps based on site spatial distribution were identified in 7 basins greater than 800 square miles. Potential site redundancies were identified in 4 basins that had an area greater than 500 square miles with a site density greater than 2 sites per 1,000 square miles. Potential site redundancies also were identified for nine site pairs by observing statistical similarities in the constituent data distributions. Sampling frequency was investigated to determine if reducing the sampling frequency of select constituents could provide a statistically similar data distribution. At 28 of 71 sites, 11 constituents had sufficient data collection frequency (approximately monthly) to allow for the creation of simulated datasets of various reduced data collection frequency. For the selected monitoring network sites analyzed, the data distribution of a simulated sampling frequency of four times per year or greater, roughly evenly distributed over the year, was not significantly different than the data distribution of the original monthly sampling frequency. Sites analyzed using varying simulated sampling frequencies tended to be more sensitive to sampling frequency changes if they were in basins classified as large or very large size and tended to be least sensitive in basins classified as small and medium size in the Ozark Plateaus Province. Simulated reduced frequency sampling analysis indicated that the constituents and measurements most sensitive to changes in sampling frequencies were water temperature, dissolved oxygen, discharge, and dissolved nitrate, and least sensitive were pH, total suspended solids, dissolved phosphorus, and total phosphorus. Discharge-weighted long-term trend analysis was repeated at 22 sites for 11 constituents using a simulated quarterly sampling frequency, and matched about 46 percent of the significant single-period trends identified using monthly data and about 65 percent of the analyses that indicated no trend using the monthly data.

Missouri

Regional regression equations for estimation of four hydraulic properties of streams at approximate bankfull conditions for different ecoregions in Texas

The U.S. Geological Survey, in cooperation with the U.S. Army Corps of Engineers, assessed statistical relations between hydraulic properties of streams at approximate bankfull conditions for different ecological regions (ecoregions) in Texas. Data from more than 103,000 records of measured discharge and ancillary hydraulic properties were assembled from summaries of discharge measurements for 424 U.S. Geological Survey streamgages in Texas. The data were subsequently subsetted at each streamgage for a streamgage-specific discharge interval centered on the estimated median annual peak discharge (0.5 annual exceedance probability) obtained from previously published regional regression equations in Texas in conjunction with the streamgage-specific sample median annual peak discharge for the period of record for each streamgage. Discharge measurements at gaged locations representing bankfull conditions (approximated from a discharge interval centered on the estimated median annual peak discharge at a given site) and associated watershed properties were subjected to rigorous statistical analysis. For most discharge measurements (where discharge is symbolically represented as Q ), the following hydraulic properties are available: cross-section area ( A ), water-surface top width ( B ), and reported mean velocity ( V ). Statewide summary statistics were computed by using these four hydraulic properties ( Q , A , B , and V ) and the following five watershed properties: (1) watershed area (contributing drainage area), (2) a multiple of main-channel slope (1,000 times main-channel slope), (3) mean annual precipitation, (4) drainage density, and (5) sinuosity ratio. From the initial set of 424 streamgages, summary statistics were computed for 372 selected streamgages in Texas and constitute the subsetted measurements dataset described in this report. Eight of the 10 ecoregions in Texas are represented in the statewide summary statistics. The resulting statistical relations, expressed as regression equations, can be used to estimate cross-section area, water-surface top width, discharge, and mean velocity of streams in different Texas ecoregions, at approximate bankfull conditions. In the regression equations, watershed properties were the independent variables for applicable watersheds, and predictions from the equations might be useful for estimating the four hydraulic properties at ungaged or unmonitored locations from selected characteristics measured at both the ungaged locations and gaged locations. Four regression equations to estimate the four hydraulic properties were identified as the preferred equations from this study. The four preferred equations use watershed area, mean annual precipitation, and aggregated ecoregion (treated as a categorical variable) to estimate the hydraulic properties, and justification is provided for this preference. For the four equations, the proportions of variance explained by the regression equations as measured by Nash-Sutcliffe efficiency are about 71 percent for cross-section area, 36 percent for top width, 76 percent for discharge, and 25 percent for mean velocity. Residual standard error (RSEs) of the regression equations are 0.252 log10 square feet for cross-section area, 0.319 log10 feet for top width, 0.247 log10 cubic feet per second for discharge, and 0.190 log10 feet per second for mean velocity, and the corresponding standard deviations of response are 0.465 log10 square feet, 0.397 log10 feet, 0.507 log10 cubic feet per second, and 0.220 log10 feet per second, respectively. The residual standard errors are less than the standard deviations as anticipated but show that the uncertainty reduction (percent change) for cross-section area is about −46 percent, about −20 percent for top width, about −51 percent for discharge, and about −14 percent for mean velocity.

Texas

Simulation of a proposed emergency outlet from Devils Lake, North Dakota

From 1993 to 2001, Devils Lake rose more than 25 feet, flooding farmland, roads, and structures around the lake and causing more than $400 million in damages in the Devils Lake Basin. In July 2001, the level of Devils Lake was at 1,448.0 feet above sea level 1 , which was the highest lake level in more than 160 years. The lake could continue to rise to several feet above its natural spill elevation to the Sheyenne River (1,459 feet above sea level) in future years, causing extensive additional flooding in the basin and, in the event of an uncontrolled natural spill, downstream in the Red River of the North Basin as well. The outlet simulation model described in this report was developed to determine the potential effects of various outlet alternatives on the future lake levels and water quality of Devils Lake. Lake levels of Devils Lake are controlled largely by precipitation on the lake surface, evaporation from the lake surface, and surface inflow. For this study, a monthly water-balance model was developed to compute the change in total volume of Devils Lake, and a regression model was used to estimate monthly water-balance data on the basis of limited recorded data. Estimated coefficients for the regression model indicated fitted precipitation on the lake surface was greater than measured precipitation in most months, fitted evaporation from the lake surface was less than estimated evaporation in most months, and ungaged inflow was about 2 percent of gaged inflow in most months. Dissolved sulfate was considered to be the key water-quality constituent for evaluating the effects of a proposed outlet on downstream water quality. Because large differences in sulfate concentrations existed among the various bays of Devils Lake, monthly water-balance data were used to develop detailed water and sulfate mass-balance models to compute changes in sulfate load for each of six major storage compartments in response to precipitation, evaporation, inflow, and outflow from each compartment. The storage compartments--five for Devils Lake and one for Stump Lake--were connected by bridge openings, culverts, or natural channels that restricted mixing between compartments. A numerical algorithm was developed to calculate inflow and outflow from each compartment. Sulfate loads for the storage compartments first were calculated using the assumptions that no interaction occurred between the bottom sediments and the water column and no wind- or buoyancy-induced mixing occurred between compartments. However, because the fitted sulfate loads did not agree with the estimated sulfate loads, which were obtained from recorded sulfate concentrations, components were added to the sulfate mass-balance model to account for the flux of sulfate between bottom sediments and the lake and for mixing between storage compartments. Mixing between compartments can occur during periods of open water because of wind and during periods of ice cover because of water-density differences between compartments. Sulfate loads calculated using the sulfate mass-balance model with sediment interaction and mixing between compartments closely matched sulfate loads computed from historical concentrations. The water and sulfate mass-balance models were used to calculate potential future lake levels and sulfate concentrations for Devils Lake and Stump Lake given potential future values of monthly precipitation, evaporation, and inflow. Potential future inputs were generated using a scenario approach and a stochastic approach. In the scenario approach, historical values of precipitation, evaporation, and inflow were repeated in the future for a particular sequence of historical years. In the stochastic approach, a statistical time-series model was developed to randomly generate potential future inputs. The scenario approach was used to evaluate the effectiveness of various outlet alternatives, and the stochastic approach was used to evaluate the hydrologic and water-quality effects of the potential outlet alternatives that were selected on the basis of the scenario analysis. Given potential future lake levels and sulfate concentrations generated using either the scenario or stochastic approach and potential future ambient flows and sulfate concentrations for the Sheyenne River receiving waters, daily outlet discharges could be calculated for virtually any outlet alternative. For the scenario approach, future ambient flows and sulfate concentrations for the Sheyenne River were generated using the same sequence of years used for generating water-balance data for Devils Lake. For the stochastic approach, a procedure was developed for generating daily Sheyenne River flows and sulfate concentrations that were "in-phase" with the generated water-balance data for Devils Lake. Simulation results for the scenario approach indicated that neither of the West Bay outlet alternatives provided effective flood-damage reduction without exceeding downstream water-quality constraints. However, both Pelican Lake outlet alternatives provided significant flood-damage reduction with only minor downstream water-quality changes. The most effective alternative for controlling rising lake levels was a Pelican Lake outlet with a 480-cubic-foot-per-second pump capacity and a 250-milligram-per-liter downstream sulfate constraint. However, this plan is costly because of the high pump capacity and the requirement of a control structure on Highway 19 to control the level of Pelican Lake. A less costly, though less effective for flood-damage reduction, plan is a Pelican Lake outlet with a 300-cubic-foot-per-second pump capacity and a 250-milligram-per-liter downstream sulfate constraint. The plan is less costly because the pump capacity is smaller and because the control structure on Highway 19 is not required. The less costly Pelican Lake alternative with a 450-milligramper- liter downstream sulfate constraint rather than a 250-milligram-per-liter downstream sulfate constraint was identified by the U.S. Army Corps of Engineers as the preferred alternative for detailed design and engineering analysis. Simulation results for the stochastic approach indicated that the geologic history of lake-level fluctuations of Devils Lake for the past 2,500 years was consistent with a climatic history that consisted of two climate states--a wet state, similar to conditions during 1980-99, and a normal state, similar to conditions during 1950-78. The transition times between the wet and normal climatic periods occurred randomly. The average duration of the wet climatic periods was 20 years, and the average duration of the normal climatic periods was 120 years. The stochastic approach was used to generate 10,000 independent sequences of lake levels and sulfate concentrations for Devils Lake for water years 2001-50. Each trace began with the same starting conditions, and the duration of the current wet cycle was generated randomly for each trace. Each trace was generated for the baseline (natural) condition and for the Pelican Lake outlet with a 300-cubic-foot-per-second pump capacity and a 450-milligram-per-liter downstream sulfate constraint. The outlet significantly lowered the probabilities of future lake-level increases within the next 50 years and did not substantially increase the probabilities of reaching low lake levels or poor water-quality conditions during the same period.

Water-Resources Investigations Report

Identification of spectral units on Phoebe

We apply a multivariate statistical method to the Phoebe spectra collected by the VIMS experiment onboard the Cassini spacecraft during the flyby of June 2004. The G-mode clustering method, which permits identification of the most important features in a spectrum, is used on a small subset of data, characterized by medium and high spatial resolution, to perform a raw spectral classification of the surface of Phoebe. The combination of statistics and comparative analysis of the different areas using both the VIMS and ISS data is explored in order to highlight possible correlations with the surface geology. In general, the results by Clark et al. [Clark, R.N., Brown, R.H., Jaumann, R., Cruikshank, D.P., Nelson, R.M., Buratti, B.J., McCord, T.B., Lunine, J., Hoefen, T., Curchin, J.M., Hansen, G., Hibbitts, K., Matz, K.-D., Baines, K.H., Bellucci, G., Bibring, J.-P., Capaccioni, F., Cerroni, P., Coradini, A., Formisano, V., Langevin, Y., Matson, D.L., Mennella, V., Nicholson, P.D., Sicardy, B., Sotin, C., 2005. Nature 435, 66-69] are confirmed; but we also identify new signatures not reported before, such as the aliphatic CH stretch at 3.53 ??m and the ???4.4 ??m feature possibly related to cyanide compounds. On the basis of the band strengths computed for several absorption features and for the homogeneous spectral types isolated by the G-mode, a strong correlation of CO2 and aromatic hydrocarbons with exposed water ice, where the uniform layer covering Phoebe has been removed, is established. On the other hand, an anti-correlation of cyanide compounds with CO2 is suggested at a medium resolution scale. ?? 2007 Elsevier Inc. All rights reserved.

Icarus

A functional model for characterizing long-distance movement behaviour

Advancements in wildlife telemetry techniques have made it possible to collect large data sets of highly accurate animal locations at a fine temporal resolution. These data sets have prompted the development of a number of statistical methodologies for modelling animal movement. Telemetry data sets are often collected for purposes other than fine-scale movement analysis. These data sets may differ substantially from those that are collected with technologies suitable for fine-scale movement modelling and may consist of locations that are irregular in time, are temporally coarse or have large measurement error. These data sets are time-consuming and costly to collect but may still provide valuable information about movement behaviour. We developed a Bayesian movement model that accounts for error from multiple data sources as well as movement behaviour at different temporal scales. The Bayesian framework allows us to calculate derived quantities that describe temporally varying movement behaviour, such as residence time, speed and persistence in direction. The model is flexible, easy to implement and computationally efficient. We apply this model to data from Colorado Canada lynx ( Lynx canadensis ) and use derived quantities to identify changes in movement behaviour.

Methods in Ecology and Evolution

Methods for estimating annual exceedance probability discharges for streams in Arkansas, based on data through water year 2013

In 2013, the U.S. Geological Survey initiated a study to update regional skew, annual exceedance probability discharges, and regional regression equations used to estimate annual exceedance probability discharges for ungaged locations on streams in the study area with the use of recent geospatial data, new analytical methods, and available annual peak-discharge data through the 2013 water year. An analysis of regional skew using Bayesian weighted least-squares/Bayesian generalized-least squares regression was performed for Arkansas, Louisiana, and parts of Missouri and Oklahoma. The newly developed constant regional skew of -0.17 was used in the computation of annual exceedance probability discharges for 281 streamgages used in the regional regression analysis. Based on analysis of covariance, four flood regions were identified for use in the generation of regional regression models. Thirty-nine basin characteristics were considered as potential explanatory variables, and ordinary least-squares regression techniques were used to determine the optimum combinations of basin characteristics for each of the four regions. Basin characteristics in candidate models were evaluated based on multicollinearity with other basin characteristics (variance inflation factor < 2.5) and statistical significance at the 95-percent confidence level ( p ≤ 0.05). Generalized least-squares regression was used to develop the final regression models for each flood region. Average standard errors of prediction of the generalized least-squares models ranged from 32.76 to 59.53 percent, with the largest range in flood region D. Pseudo coefficients of determination of the generalized least-squares models ranged from 90.29 to 97.28 percent, with the largest range also in flood region D. The regional regression equations apply only to locations on streams in Arkansas where annual peak discharges are not substantially affected by regulation, diversion, channelization, backwater, or urbanization. The applicability and accuracy of the regional regression equations depend on the basin characteristics measured for an ungaged location on a stream being within range of those used to develop the equations.

Arkansas

Digital computer methods for water‐quality data

The digital computer is used on a routine basis in the ground-water program in Kansas for tasks ranging from the listing of water-quality data in tabular and publishable form to statistically and graphically analyzing a mass of data. In the past year a number of computer programs in FORTRAN IV have been developed by Charles O. Morgan and Jesse M. McNellis using an IBM-7040 computer to store, retrieve, and manipulate water-quality data. These programs: (1) Tabulate data at the rate of 40 chemical analyses of water per minute in a format similar to that found in the Kansas ground-water publications. (2) Perform necessary calculations and print Stiff diagrams at the rate of 30 per minute. (3) Perform necessary calculations and print Piper diagrams, including a square modification of the normally diamond-shaped cation-anion diagram, and trilinear diagrams of the cations and anions. The symbol representing the analyses located on the diagrams can be designated by either an analysis number or a geologic unit number. A cation-anion diagram showing the average chemical composition of water for an aquifer can also be printed. These diagrams for 50 analyses can be produced in 1.5 minutes. (4) Plot maps of 42 individual, combined, or calculated parameters obtained from the data cards. These maps can be plotted to any specified scale and for as many as 10 designated geologic units. Computer time involved for one map with 50 plotted points is 15 seconds. It is estimated that the use of these programs will save several man-months during a ground-water study, and the error inherent in the manual manipulation of data is greatly reduced. The present cost for running 50 analyses through the four water-quality programs on the computer is approximately $20.

Groundwater

Wave data processing toolbox manual

Researchers routinely deploy oceanographic equipment in estuaries, coastal nearshore environments, and shelf settings. These deployments usually include tripod-mounted instruments to measure a suite of physical parameters such as currents, waves, and pressure. Instruments such as the RD Instruments Acoustic Doppler Current Profiler (ADCP(tm)), the Sontek Argonaut, and the Nortek Aquadopp(tm) Profiler (AP) can measure these parameters. The data from these instruments must be processed using proprietary software unique to each instrument to convert measurements to real physical values. These processed files are then available for dissemination and scientific evaluation. For example, the proprietary processing program used to process data from the RD Instruments ADCP for wave information is called WavesMon. Depending on the length of the deployment, WavesMon will typically produce thousands of processed data files. These files are difficult to archive and further analysis of the data becomes cumbersome. More imperative is that these files alone do not include sufficient information pertinent to that deployment (metadata), which could hinder future scientific interpretation. This open-file report describes a toolbox developed to compile, archive, and disseminate the processed wave measurement data from an RD Instruments ADCP, a Sontek Argonaut, or a Nortek AP. This toolbox will be referred to as the Wave Data Processing Toolbox. The Wave Data Processing Toolbox congregates the processed files output from the proprietary software into two NetCDF files: one file contains the statistics of the burst data and the other file contains the raw burst data (additional details described below). One important advantage of this toolbox is that it converts the data into NetCDF format. Data in NetCDF format is easy to disseminate, is portable to any computer platform, and is viewable with public-domain freely-available software. Another important advantage is that a metadata structure is embedded with the data to document pertinent information regarding the deployment and the parameters used to process the data. Using this format ensures that the relevant information about how the data was collected and converted to physical units is maintained with the actual data. EPIC-standard variable names have been utilized where appropriate. These standards, developed by the NOAA Pacific Marine Environmental Laboratory (PMEL) (http://www.pmel.noaa.gov/epic/), provide a universal vernacular allowing researchers to share data without translation.

Open-File Report

Predicted uranium and radon concentrations in New Hampshire (USA) groundwater—Using Multi Order Hydrologic Position as predictors

Two radioactive elements, uranium (U) and radon (Rn), which are of potential concern in New Hampshire (NH) groundwater, are investigated. Exceedance probability maps are tools to highlight locations where the concentrations of undesirable substances in the groundwater may be elevated. Two forms of statistical analysis are used to create exceedance probability maps for U and Rn in NH groundwater. The first, Boosted Regression Tree (BRT), was selected for estimating U exceedance values. It computes exceedance values directly using the Bernoulli distribution function. The second method of statistical analysis used for Rn to determine exceedance probabilities is ordinary least squares (OLS) regression. In the process of determining exceedance probabilities for U and Rn, the utility of a new dataset is investigated. That new predictor dataset is the Multi-Order Hydrologic Position (MOHP) dataset. MOHP raster datasets have been produced nationally for the conterminous United States at a 30-m resolution. The concept behind MOHP is that, for any given point on the earth's surface, there is the potential for a longer groundwater flow path as one goes deeper beneath the land surface. MOHP predictors were tested in both models. Three MOHP predictors were found useful in the BRT model and two in the OLS model. MOHP data were found useful as predictors along with other site characteristics in predicting U and Rn exceedance probabilities in New Hampshire groundwater.

New Hampshire

Statistical analysis of water-quality data containing multiple detection limits: S-language software for regression on order statistics

Trace contaminants in water, including metals and organics, often are measured at sufficiently low concentrations to be reported only as values below the instrument detection limit. Interpretation of these "less thans" is complicated when multiple detection limits occur. Statistical methods for multiply censored, or multiple-detection limit, datasets have been developed for medical and industrial statistics, and can be employed to estimate summary statistics or model the distributions of trace-level environmental data. We describe S-language-based software tools that perform robust linear regression on order statistics (ROS). The ROS method has been evaluated as one of the most reliable procedures for developing summary statistics of multiply censored data. It is applicable to any dataset that has 0 to 80% of its values censored. These tools are a part of a software library, or add-on package, for the R environment for statistical computing. This library can be used to generate ROS models and associated summary statistics, plot modeled distributions, and predict exceedance probabilities of water-quality standards. ?? 2005 Elsevier Ltd. All rights reserved.

Computers & Geosciences

Percentage entrainment of constituent loads in urban runoff, south Florida

Runoff quantity and quality data from four urban basins in south Florida were analyzed to determine the entrainment of total nitrogen, total phosphorus, total carbon, chemical oxygen demand, suspended solids, and total lead within the stormwater runoff. Land use of the homogeneously developed basins are residential (single family), highway, commercial, and apartment (multifamily). A computational procedure was used to calculate, for all storms that had water-quality data, the percentage of constituent load entrainment in specified depths of runoff. The plot of percentage of constituent load entrained as a function of runoff is termed the percentage-entrainment curve. Percentage-entrainment curves were developed for three different source areas of basin runoff: (1) the hydraulically effective impervious area, (2) the contributing area, and (3) the drainage area. With basin runoff expressed in inches over the contributing area, the depth of runoff required to remove 90 percent of the constituent load ranged from about 0.4 inch to about 1.4 inches; and to remove 80 percent, from about 0.3 to 0.9 inch. Analysis of variance, using depth of runoff from the contributing area as the response variable, showed that the factor 'basin' is statistically significant, but that the factor 'constituent' is not statistically significant in the forming of the percentage-entrainment curve. Evidently the sewerage design, whether elongated or concise in plan dictates the shape of the percentage-entrainment curve. The percentage-entrainment curves for all constituents were averaged for each basin and plotted against basin runoff for three source areas of runoff-the hydraulically effective impervious area, the contributing area, and the drainage area. The relative positions of the three curves are directly related to the relative sizes of the three source areas considered. One general percentage-entrainment curve based on runoff from the contributing area was formed by averaging across both constituents and basins. Its coordinates are: 0.25 inch of runoff for 50-percent entrainment, 0.65 inch of runoff for 80-percent entrainment, and 0.95 inch of runoff for 90-percent entrainment. The general percentage-entrainment curve based on runoff from the hydraulically effective impervious area has runoff values of 0.35, 0.95, 1.6 inches, respectively.

Florida

The Digital Shoreline Analysis System (DSAS) Version 4.0 - An ArcGIS extension for calculating shoreline change

The Digital Shoreline Analysis System (DSAS) version 4.0 is a software extension to ESRI ArcGIS v.9.2 and above that enables a user to calculate shoreline rate-of-change statistics from multiple historic shoreline positions. A user-friendly interface of simple buttons and menus guides the user through the major steps of shoreline change analysis. Components of the extension and user guide include (1) instruction on the proper way to define a reference baseline for measurements, (2) automated and manual generation of measurement transects and metadata based on user-specified parameters, and (3) output of calculated rates of shoreline change and other statistical information. DSAS computes shoreline rates of change using four different methods: (1) endpoint rate, (2) simple linear regression, (3) weighted linear regression, and (4) least median of squares. The standard error, correlation coefficient, and confidence interval are also computed for the simple and weighted linear-regression methods. The results of all rate calculations are output to a table that can be linked to the transect file by a common attribute field. DSAS is intended to facilitate the shoreline change-calculation process and to provide rate-of-change information and the statistical data necessary to establish the reliability of the calculated results. The software is also suitable for any generic application that calculates positional change over time, such as assessing rates of change of glacier limits in sequential aerial photos, river edge boundaries, land-cover changes, and so on.

Open-File Report

Structure of high latitude currents in global magnetospheric-ionospheric models

Using three resolutions of the Lyon-Fedder-Mobarry global magnetosphere-ionosphere model (LFM) and the Weimer 2005 empirical model we examine the structure of the high latitude field-aligned current patterns. Each resolution was run for the entire Whole Heliosphere Interval which contained two high speed solar wind streams and modest interplanetary magnetic field strengths. Average states of the field-aligned current (FAC) patterns for 8 interplanetary magnetic field clock angle directions are computed using data from these runs. Generally speaking the patterns obtained agree well with results obtained from the Weimer 2005 computing using the solar wind and IMF conditions that correspond to each bin. As the simulation resolution increases the currents become more intense and narrow. A machine learning analysis of the FAC patterns shows that the ratio of Region 1 (R1) to Region 2 (R2) currents decreases as the simulation resolution increases. This brings the simulation results into better agreement with observational predictions and the Weimer 2005 model results. The increase in R2 current strengths also results in the cross polar cap potential (CPCP) pattern being concentrated in higher latitudes. Current-voltage relationships between the R1 and CPCP are quite similar at the higher resolution indicating the simulation is converging on a common solution. We conclude that LFM simulations are capable of reproducing the statistical features of FAC patterns.

Space Science Reviews

Peak-discharge frequency and potential extreme peak discharge for natural streams in the Brazos River basin, Texas

The 2-, 5-, 10-, 25-, 50-, and 100-year peak discharges were estimated for 186 streamflow-gaging stations with at least 8 years of data for natural streams in and near the Brazos River Basin, Texas. Multiple regression equations were developed to estimate peak-discharge frequency for the 2-, 5-, 10-, 25-, 50-, and 100-year recurrence intervals for each of three hydrologic regions that compose the Brazos River Basin. The equations for each region are a function of significant basin characteristics (explanatory variables). The significant explanatory variables among six that were tested are the contributing drainage area and stream slope for regions 1 and 2 and the contributing drainage area for region 3. For the three sets of equations, the coefficient of determination ranges from 0.59 to 0.93, and the standard error ranges from 0.184 to 0.391 log units. A larger coefficient of determination and a lower standard error generally are associated with the equations for hydrologic regions 2 and 3. Statistics from the regression analysis allow computation of the prediction interval associated with a given significance level for a peak-discharge frequency estimate. The regression equations can be used to estimate peak discharges for sites at, near, or away from sites with streamflow-gaging stations. The potential extreme peak-discharge curves as related to contributing drainage area were estimated for each of the three hydrologic regions from measured extreme peaks of record at 186 sites with streamflow-gaging stations and from measured extreme peaks at 37 sites without streamflow-gaging stations in and near the Brazos River Basin. The potential extreme peak-discharge curves generally are similar for hydrologic regions 1 and 2, and the curve for region 3 consistently is below the curves for regions 1 and 2, which indicates smaller peak discharges.

Texas

Increasing seasonal variation in the extent of rivers and lakes from 1984 to 2022

Knowledge of the spatial and temporal distribution of surface water is important for water resource management, flood risk assessment, monitoring ecosystem health, constraining estimates of biogeochemical cycles and understanding our climate. While global-scale spatiotemporal change detection of surface water has significantly improved in recent years due to planetary-scale remote sensing and computing, it has remained challenging to distinguish the changing characteristics of rivers and lakes. Here we analyze the spatial extent of permanent and seasonal rivers and lakes globally over the past 38 years based on new data of river system extents and surface water trends. Results show that while the total permanent surface area of both rivers and lakes has remained relatively constant, the areas with intermittent seasonal coverage have increased by 12 % and 27 % for rivers and lakes, respectively. The increase is statistically significant in over 84 % of global water catchments based on Spearman's rank correlations (rho) above 0.05 and p values less than 0.05. The seasonal river extent is nearly 32 % larger than the previously observed annual mean river extent, suggesting large seasonal variations that impact not only ecosystem health but also estimations of terrestrial biogeochemical cycles of carbon. The outcomes of our analysis are shared as the Surface Area of Rivers and Lakes (SARL) database, serving as a valuable resource for monitoring and research of hydrological cycles, ecosystem accounting, and water management.

Hydrology and Earth System Sciences