Learn R Programming

⚠️There's a newer version (4.3.0) of this package.Take me there.

FedData

FedData is an R package implementing functions to automate downloading geospatial data available from several federated data sources (mainly sources maintained by the US Federal government). Currently, the package allows for retrieval of five datasets:

Additional data sources are in the works.

This package is designed with the large-scale geographic information system (GIS) use-case in mind: cases where the use of dynamic web-services is impractical due to the scale (spatial and/or temporal) of analysis. It functions primarily as a means of downloading tiled or otherwise spatially-defined datasets; additionally, it can preprocess those datasets by extracting data within an area of interest (AoI), defined spatially. It relies heavily on the sp, raster, and rgdal packages.

This package has been built and tested on a source (Homebrew) install of R on macOS 10.12 (Sierra), and has been successfully run on Ubuntu 14.04.5 LTS (Trusty), Ubuntu 16.04.1 LTS (Xenial) and binary installs of R on Mac OS 10.12 and Windows 10.

Development

Contributors

Install FedData

  • From CRAN:

    install.packages('FedData')
  • Development version from GitHub:

    install.packages("devtools")
    devtools::install_github("bocinsky/FedData")
    library(FedData)
  • Linux (Ubuntu 14.04.5 or 16.04.1):

    First, in terminal:

    sudo add-apt-repository ppa:ubuntugis/ppa -y
    sudo apt-get update -q
    sudo apt-get install libssl-dev libcurl4-openssl-dev netcdf-bin libnetcdf-dev gdal-bin libgdal-dev

    Then, in R:

    update.packages("survival")
    install.packages("devtools")
    devtools::install_github("bocinsky/FedData")
    library(FedData)

Demonstration

This demo script is available in the /inst folder at the location of the installed package.

Load FedData and define a study area

# FedData Tester
library(FedData)
library(magrittr)

# Set a directory for testing
testDir <- "~/FedData Test"
# and create it if necessary
dir.create(testDir, showWarnings=F, recursive=T)
setwd(testDir)

# Extract data for the Village Ecodynamics Project "VEPIIN" study area:
# http://village.anth.wsu.edu
vepPolygon <- polygon_from_extent(raster::extent(672800,740000,4102000,4170000),
                                  proj4string="+proj=utm +datum=NAD83 +zone=12")

Get and plot the National Elevation Dataset for the study area

# Get the NED (USA ONLY)
# Returns a raster
NED <- get_ned(template=vepPolygon,
               label="VEPIIN")
# Plot with raster::plot
raster::plot(NED)

Get and plot the Daymet dataset for the study area

# Get the DAYMET (North America only)
# Returns a raster
DAYMET <- get_daymet(template=vepPolygon,
               label="VEPIIN",
               elements = c("prcp","tmax"),
               years = 1980:1985)
# Plot with raster::plot
raster::plot(DAYMET$tmax$X1985.10.23)

Get and plot the daily GHCN precipitation data for the study area

# Get the daily GHCN data (GLOBAL)
# Returns a list: the first element is the spatial locations of stations,
# and the second is a list of the stations and their daily data
GHCN.prcp <- get_ghcn_daily(template=vepPolygon, 
                            label="VEPIIN", 
                            elements=c('prcp'))
# Plot the NED again
raster::plot(NED)
# Plot the spatial locations
sp::plot(GHCN.prcp$spatial, pch=1, add=T)
legend('bottomleft', pch=1, legend="GHCN Precipitation Records")

Get and plot the daily GHCN temperature data for the study area

# Elements for which you require the same data
# (i.e., minimum and maximum temperature for the same days)
# can be standardized using standardize==T
GHCN.temp <- get_ghcn_daily(template = vepPolygon, 
                            label = "VEPIIN", 
                            elements = c('tmin','tmax'), 
                            years = 1980:1985,
                            standardize = T)
# Plot the NED again
raster::plot(NED)
# Plot the spatial locations
sp::plot(GHCN.temp$spatial, add=T, pch=1)
legend('bottomleft', pch=1, legend="GHCN Temperature Records")

Get and plot the National Hydrography Dataset for the study area

# Get the NHD (USA ONLY)
NHD <- get_nhd(template=vepPolygon, 
               label="VEPIIN")
# Plot the NED again
raster::plot(NED)
# Plot the NHD data
NHD %>%
  lapply(sp::plot, col='black', add=T)

Get and plot the NRCS SSURGO data for the study area

# Get the NRCS SSURGO data (USA ONLY)
SSURGO.VEPIIN <- get_ssurgo(template=vepPolygon, 
                     label="VEPIIN")
# Plot the NED again
raster::plot(NED)
# Plot the SSURGO mapunit polygons
plot(SSURGO.VEPIIN$spatial,
     lwd=0.1,
     add=T)

Get and plot the NRCS SSURGO data for particular soil survey areas

# Or, download by Soil Survey Area names
SSURGO.areas <- get_ssurgo(template=c("CO670","CO075"), 
                           label="CO_TEST")

# Let's just look at spatial data for CO675
SSURGO.areas.CO675 <- SSURGO.areas$spatial[SSURGO.areas$spatial$AREASYMBOL=="CO075",]

# And get the NED data under them for pretty plotting
NED.CO675 <- get_ned(template=SSURGO.areas.CO675,
                            label="SSURGO_CO675")
               
# Plot the SSURGO mapunit polygons, but only for CO675
plot(NED.CO675)
plot(SSURGO.areas.CO675,
     lwd=0.1,
     add=T)

Get and plot the ITRDB chronology locations in the study area

# Get the ITRDB records
ITRDB <- get_itrdb(template=vepPolygon,
                        label="VEPIIN",
                        makeSpatial=T)
# Plot the NED again
raster::plot(NED)
# Map the locations of the tree ring chronologies
plot(ITRDB$metadata, pch=1, add=T)
legend('bottomleft', pch=1, legend="ITRDB chronologies")

Acknowledgements

This package is a product of SKOPE (Synthesizing Knowledge of Past Environments) and the Village Ecodynamics Project. This software is licensed under the MIT license.

Copy Link

Version

Install

install.packages('FedData')

Monthly Downloads

911

Version

2.4.0

License

MIT + file LICENSE

Maintainer

R. Bocinsky

Last Published

January 21st, 2017

Functions in FedData (2.4.0)

get_itrdb

Download the latest version of the ITRDB, and extract given parameters.
get_ghcn_daily

Download and crop the Global Historical Climate Network-Daily data.
get_ghcn_inventory

Download and crop the inventory of GHCN stations.
get_daymet_tile

Download and crop a netcdf tile from the 1-km DAYMET daily weather dataset.
FedData-package

Scripts to automate downloading geospatial data available from the several federated data sources
download_ghcn_daily_station

Download the daily data for a GHCN weather station.
get_ned

Download and crop the 1 (~30 meter) or 1/3 (~10 meter) arc-second National Elevation Dataset.
get_ned_tile

Download and crop tile from the 1 (~30 meter) or 1/3 (~10 meter) arc-second National Elevation Dataset.
get_huc4

Download and crop a shapefile of the HUC4 regions of the National Hydrography Dataset.
get_daymet

Download and crop the 1-km DAYMET daily weather dataset.
get_ssurgo

Download and crop data from the NRCS SSURGO soils database.
pkg_test

Install and load a package.
get_ghcn_daily_station

Download and extract the daily data for a GHCN weather station.
polygon_from_extent

Turn an extent object into a polygon
get_ssurgo_inventory

Download and crop a shapefile of the SSURGO study areas.
read_crn_data

Read chronology data from a Tucson-format chronology file.
tiles

The DAYMET tiles SpatialPolygonsDataFrame.
get_ssurgo_study_area

Download and crop the spatial and tabular data for a SSURGO study area.
unwrap_rows

Unwraps a matrix and only keep the first n elements.
get_nhd_subregion

Download and crop data from a zipped HUC4 subregion of the National Hydrography Dataset.
get_nhd

Download and crop the National Hydrography Dataset.
read_crn_metadata

Read metadata from a Tucson-format chronology file.
read_crn

Read a Tucson-format chronology file.
sequential_duplicated

Get a logical vector of which elements in a vector are sequentially duplicated.
spdf_from_polygon

Turn a SpatialPolygons object into a SpatialPolygonsDataFrame.
substr_right

Get the rightmost 'n' characters of a character string.
station_to_data_frame

Convert a list of station data to a single data frame.
extract_ssurgo_data

Extract data from a SSURGO databse pertaining to a set of mapunits.
download_itrdb

Download the latest version of the ITRDB.
download_nhd_subregion

Download a zipped NHD HUC4 subregion.
download_ssurgo_study_area

Download a zipped directory containing the spatial and tabular data for a SSURGO study area.
download_huc4

Download a zipped directory containing a shapefile of the HUC4 subregions of the NHD.
download_ssurgo_inventory

Download a zipped directory containing a shapefile of the SSURGO study areas.
download_daymet_tile

Download a netcdf tile from the 1-km DAYMET daily weather dataset.
download_data

Use curl to download a file.
download_ned_tile

Download a zipped tile from the 1 (~30 meter) or 1/3 (~10 meter) arc-second National Elevation Dataset.