gbifdb: High Performance Interface to 'GBIF'

A high performance interface to the Global Biodiversity Information Facility, 'GBIF'. In contrast to 'rgbif', which can access small subsets of 'GBIF' data through web-based queries to a central server, 'gbifdb' provides enhanced performance for R users performing large-scale analyses on servers and cloud computing providers, providing full support for arbitrary 'SQL' or 'dplyr' operations on the complete 'GBIF' data tables (now over 1 billion records, and over a terabyte in size). 'gbifdb' accesses a copy of the 'GBIF' data in 'parquet' format, which is already readily available in commercial computing clouds such as the Amazon Open Data portal and the Microsoft Planetary Computer, or can be accessed directly without downloading, or downloaded to any server with suitable bandwidth and storage space. The high-performance techniques for local and remote access are described in <> and <> respectively.

Version: 0.1.2
Depends: R (≥ 4.0)
Imports: arrow (≥ 6.0.1), duckdb (≥ 0.2.9), DBI, dplyr
Suggests: spelling, dbplyr, testthat (≥ 3.0.0), covr, knitr, rmarkdown, aws.s3
Published: 2022-05-21
Author: Carl Boettiger ORCID iD [aut, cre]
Maintainer: Carl Boettiger <cboettig at>
License: Apache License (≥ 2)
NeedsCompilation: no
Language: en-US
Citation: gbifdb citation info
Materials: README NEWS
CRAN checks: gbifdb results


Reference manual: gbifdb.pdf
Vignettes: Intro to gbifdb


Package source: gbifdb_0.1.2.tar.gz
Windows binaries: r-devel:, r-release:, r-oldrel:
macOS binaries: r-release (arm64): gbifdb_0.1.2.tgz, r-oldrel (arm64): gbifdb_0.1.2.tgz, r-release (x86_64): gbifdb_0.1.2.tgz, r-oldrel (x86_64): gbifdb_0.1.2.tgz


Please use the canonical form to link to this page.