Server definition
- Hash
- sha256:088191def31d39344769be87ea036f6e9280da4f790cfcae1d52fbcf84a82ef6
- What it is
- What a remote MCP server returned when asked what it offers: 13 tools
The blob, as servednamed by its sha256
{
"instructions": "Use the gbif_* tools to query species taxonomy, occurrences, datasets, and publishers via the GBIF API. Keyless — these endpoints need no credentials. Resolve any name with gbif_match_species first — it returns the backbone taxonKey the occurrence tools expect, resolving a synonym to its accepted taxon and reporting the matched synonym key alongside as matchedTaxonKey, unlike the raw scientificName filter. The kingdom, family, and genus filters on gbif_search_species take names and are resolved to backbone keys the same way before the search runs, since /species/search scopes by taxon key alone: the narrowest of the three supplied is what scopes, a name is matched exactly and capitalized as GBIF writes it, one that reaches no backbone taxon at that rank fails rather than widening the search, and pairing any of them with datasetKey matches nothing unless that checklist is the GBIF backbone. Countries use ISO 3166-1 alpha-2; datasets and publishers are keyed by UUID, occurrences by integer key. Omit a filter you do not want rather than sending it blank: GBIF ignores a parameter it is given with no value and answers with the whole unfiltered scope, so every optional filter on these tools rejects a blank or whitespace-only value instead of quietly widening the query. That rejection is a typed invalid_filter carrying the recovery hint the tool declares, except on country and publishingCountry, where the uppercase-only pattern rejects it at the schema and the failure arrives as JSON-RPC -32602 with empty structuredContent. On the occurrence tools country and publishingCountry, and on gbif_search_datasets publishingCountry, the uppercase two-letter form is the only one accepted — lowercase and alpha-3 forms (\"gb\", \"USA\") match nothing on those routes, so those fields reject anything else rather than answering zero, while gbif_search_publishers resolves either form case-insensitively against the registry and needs no such constraint. An organization key from gbif_search_publishers chains into gbif_search_datasets two ways: publishingOrg for the datasets that organization published, hostingOrg for the ones its own installation serves. publishingOrg is almost always the one meant — most organizations publish through an installation someone else runs, so hostingOrg matches nothing for them — and supplying both intersects the two rather than combining them. On the occurrence tools country is where the record was observed and publishingCountry is the country of the organization that published it — different questions that disagree on most records, so pick deliberately; stateProvince is matched verbatim, exactly and case-sensitively, and an unrecognized value returns zero records rather than an error, so take one from a STATE_PROVINCE facet instead of guessing a spelling. GBIF indexes absence records — a documented survey that looked for a taxon and did not find it — alongside sightings, so gbif_search_occurrences, gbif_count_occurrences, and gbif_occurrence_facets all default occurrenceStatus to PRESENT and report the applied filter in their enrichment; pass ANY to include absences, or ABSENT for absences alone. For some taxa absences dominate: an unfiltered count can run five orders of magnitude above the sightings it appears to report. The dataset recordCount on gbif_search_datasets, gbif_get_dataset, and the gbif://dataset/{datasetKey} resource is not filtered that way — it spans every occurrenceStatus, so it will exceed a gbif_count_occurrences total for the same datasetKey by design. Occurrence paging caps at offset+limit=100,001 and GBIF exposes no cursor or scroll, so a larger result set is reached only by partitioning it: gbif_occurrence_facets with facet DATASET_KEY splits a scope into buckets that sum to its full total, since every occurrence carries exactly one datasetKey, and each bucket is then searchable on its own; a facet on a dimension a record can lack — YEAR, MONTH, STATE_PROVINCE — leaves those records in no bucket, stateProvince included even though the occurrence tools can filter on it, while BASIS_OF_RECORD and PUBLISHING_COUNTRY are gap-free and both have a matching search filter, making either a sound second cut on a bucket still over the cap though too coarse for the first. Bucket sums reconcile only against the same occurrenceStatus, and only when the facet call repeats the search filters — gbif_occurrence_facets accepts a narrower set than gbif_search_occurrences and gbif_count_occurrences, so re-apply whatever it cannot take on each per-datasetKey search. This server cannot retrieve a result set in bulk — the GBIF Download API needs a GBIF.org account and returns an archive asynchronously, and the monthly GBIF Parquet snapshot on AWS Open Data is a bulk dataset; both are routes to take outside this server.",
"tools": [
{
"description": "Resolve up to 50 scientific names to GBIF backbone taxon keys in one call — the batch counterpart to gbif_match_species for checklist, inventory, and species-list workflows that would otherwise need one round trip per name. Each name is matched independently and results are returned in input order, one entry per name. A name with no backbone match yields matchType NONE (no taxonKey) instead of failing the batch; a per-name lookup failure yields matchType ERROR carrying that name's error message and, when the failure was classified, a machine-readable reason — the rest of the batch is unaffected, and the call as a whole still succeeds. When a queried name is a synonym, taxonKey is the accepted taxon it resolves to and matchedTaxonKey carries the synonym's own key. Common names are not supported — use gbif_search_species for vernacular searches. Below confidence 80, review the match.",
"inputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"properties": {
"names": {
"description": "Scientific names to match against the GBIF backbone. 1–50 per call, matched in parallel.",
"items": {
"description": "A scientific name to match, e.g. \"Panthera leo\".",
"minLength": 1,
"type": "string"
},
"maxItems": 50,
"minItems": 1,
"type": "array"
},
"strict": {
"default": false,
"description": "When true, require an exact match for every name (no fuzzy matching). When false (default), GBIF applies fuzzy matching to tolerate minor misspellings.",
"type": "boolean"
}
},
"required": [
"names"
],
"type": "object"
},
"name": "gbif_bulk_match_species",
"outputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"anyOf": [
{
"not": {
"required": [
"error"
]
},
"required": [
"results"
]
},
{
"required": [
"error"
]
}
],
"properties": {
"error": {
"additionalProperties": {},
"description": "Present when the call failed. Absent on success.",
"properties": {
"code": {
"description": "JSON-RPC error code for this failure.",
"maximum": 9007199254740991,
"minimum": -9007199254740991,
"type": "integer"
},
"data": {
"additionalProperties": {},
"properties": {
"reason": {
"description": "Machine-readable failure mode.",
"type": "string"
},
"recovery": {
"additionalProperties": {},
"description": "Actionable next step for the caller.",
"properties": {
"hint": {
"type": "string"
}
},
"required": [
"hint"
],
"type": "object"
},
"retryable": {
"description": "Whether retrying may succeed.",
"type": "boolean"
}
},
"type": "object"
},
"message": {
"description": "Human-readable description of what went wrong.",
"type": "string"
}
},
"required": [
"code",
"message"
],
"type": "object"
},
"results": {
"description": "One result per input name, in input order.",
"items": {
"additionalProperties": false,
"description": "Match outcome for one input name.",
"properties": {
"canonicalName": {
"description": "Matched scientific name without authorship. Absent when unmatched.",
"type": "string"
},
"confidence": {
"description": "Match confidence 0–100. Below 80 warrants review. Absent on ERROR.",
"type": "number"
},
"error": {
"description": "Failure message when matchType is ERROR. Absent otherwise.",
"type": "string"
},
"matchType": {
"description": "EXACT, FUZZY, or HIGHERRANK for a match; NONE when GBIF found no usable match; ERROR when the lookup itself failed for this name (see error).",
"type": "string"
},
"matchedTaxonKey": {
"description": "Backbone key of the name that actually matched. Present only when it differs from taxonKey — that is, when a synonym was resolved to its accepted taxon.",
"type": "number"
},
"name": {
"description": "The input name this entry corresponds to.",
"type": "string"
},
"rank": {
"description": "Taxonomic rank of the matched taxon.",
"type": "string"
},
"reason": {
"description": "Machine-readable failure identifier when matchType is ERROR — e.g. invalid_filter when GBIF rejected a supplied value. Absent when the failure carried no classification, and absent on every non-ERROR entry.",
"type": "string"
},
"scientificName": {
"description": "Full matched scientific name with authorship. Absent when unmatched.",
"type": "string"
},
"status": {
"description": "Taxonomic status: ACCEPTED, SYNONYM, or DOUBTFUL.",
"type": "string"
},
"taxonKey": {
"description": "GBIF backbone taxon key to pass to gbif_search_occurrences, gbif_count_occurrences, and gbif_occurrence_facets — the accepted taxon's key when this name is a synonym, otherwise the matched taxon's own key. Absent when matchType is NONE or ERROR.",
"type": "number"
}
},
"required": [
"name",
"matchType"
],
"type": "object"
},
"type": "array"
}
},
"type": "object"
}
},
{
"description": "Count occurrences matching a taxon + location filter without fetching records. Use for quick totals (\"how many Aves records in Sweden?\") or before deciding whether to paginate a full search. Accepts taxonKey, country (uppercase ISO 3166-1 alpha-2), publishingCountry, stateProvince, isGeoreferenced, datasetKey, year, occurrenceStatus, and iucnRedListCategory. Counts sightings only by default, matching gbif_search_occurrences — GBIF also indexes absence records, and for some taxa they are the overwhelming majority. A count above 100,001 is the signal to partition rather than page: gbif_search_occurrences cannot reach past that offset, so split the query by DATASET_KEY via gbif_occurrence_facets and search each dataset separately.",
"inputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"properties": {
"country": {
"description": "ISO 3166-1 alpha-2 code, uppercase, of where the occurrence was recorded (e.g., \"GB\", \"US\"). Not the publisher's country — that is publishingCountry, and the two disagree on most records. Lowercase and alpha-3 forms (\"gb\", \"USA\") match nothing upstream, which is why only the uppercase two-letter form is accepted here. Take a value from a COUNTRY facet on gbif_occurrence_facets; an uppercase pair GBIF does not know (\"XX\") is rejected upstream by name.",
"pattern": "^[A-Z]{2}$",
"type": "string"
},
"datasetKey": {
"description": "Filter to a specific dataset UUID (8-4-4-4-12 hex) from gbif_search_datasets. Omit the field to count across every dataset — an empty string is rejected rather than read as no filter, because GBIF answers a blank datasetKey with the unfiltered total. The result is not the recordCount the dataset tools and the gbif://dataset/{datasetKey} resource report for the same key: that figure spans every occurrenceStatus, while this count applies occurrenceStatus below, PRESENT by default.",
"type": "string"
},
"isGeoreferenced": {
"description": "When true, count only georeferenced records. When false, count only non-georeferenced records.",
"type": "boolean"
},
"iucnRedListCategory": {
"description": "Count only records whose taxon carries this IUCN Red List category: CR Critically Endangered, EN Endangered, VU Vulnerable, NT Near Threatened, LC Least Concern, DD Data Deficient, EX Extinct, EW Extinct in the Wild, CD Conservation Dependent. Records with no category are excluded when this is set.",
"enum": [
"CR",
"EN",
"VU",
"NT",
"LC",
"DD",
"EX",
"EW",
"CD"
],
"type": "string"
},
"occurrenceStatus": {
"default": "PRESENT",
"description": "Presence/absence filter. Defaults to PRESENT: an ABSENT record documents a survey that looked for the taxon and did not find it, so counting one inflates the total with the opposite of a sighting. Use ANY for both (GBIF's own default), or ABSENT for non-observations alone. Matches the gbif_search_occurrences default, so the two tools agree.",
"enum": [
"PRESENT",
"ABSENT",
"ANY"
],
"type": "string"
},
"publishingCountry": {
"description": "ISO 3166-1 alpha-2 code, uppercase, of the organization that published the record — not where the occurrence was observed, which is country. The two differ constantly: of 60,290,950 records observed in GB, 1,548,928 were published by US organizations. Take a value from a PUBLISHING_COUNTRY facet on gbif_occurrence_facets. Lowercase and alpha-3 forms (\"us\", \"USA\") match nothing upstream, which is why only the uppercase two-letter form is accepted here.",
"pattern": "^[A-Z]{2}$",
"type": "string"
},
"stateProvince": {
"description": "State, province, or first-level administrative division, matched as a verbatim string — exact and case-sensitive. GBIF stores what each dataset recorded without normalizing it, so there is no vocabulary to guess from: \"England\", \"England - Greater London\", and \"Greater London\" are three distinct values, and \"england\" is none of them. Take one from a STATE_PROVINCE facet on gbif_occurrence_facets scoped the same way and pass it back unchanged — an unmatched value counts zero rather than erroring. Omit the field to count across every state or province — a blank or whitespace-only value is rejected rather than dropped, because GBIF answers one with the unfiltered total.",
"type": "string"
},
"taxonKey": {
"description": "GBIF backbone taxon key from gbif_match_species. Matches the given taxon and all descendant taxa (subspecies, varieties, etc.).",
"type": "number"
},
"year": {
"description": "Year or year range (e.g., \"2024\" or \"2020,2024\"). Both endpoints inclusive. Omit the field to count across every year — a blank or whitespace-only value is rejected rather than dropped, because GBIF answers one with the unfiltered total.",
"type": "string"
}
},
"type": "object"
},
"name": "gbif_count_occurrences",
"outputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"anyOf": [
{
"not": {
"required": [
"error"
]
},
"required": [
"count",
"occurrenceStatus"
]
},
{
"required": [
"error"
]
}
],
"properties": {
"count": {
"description": "Total occurrences matching the supplied filters.",
"type": "number"
},
"error": {
"additionalProperties": {},
"description": "Present when the call failed. Absent on success.",
"properties": {
"code": {
"description": "JSON-RPC error code for this failure.",
"maximum": 9007199254740991,
"minimum": -9007199254740991,
"type": "integer"
},
"data": {
"additionalProperties": {},
"properties": {
"reason": {
"description": "Machine-readable failure mode. Declared by this tool: `invalid_filter`: A filter was supplied blank or whitespace-only, datasetKey is not an 8-4-4-4-12 hex UUID, a two-letter country or publishingCountry code is one GBIF does not know, or GBIF rejected another filter value as malformed. Other values are possible when a failure originates below the handler.",
"examples": [
"invalid_filter"
],
"type": "string"
},
"recovery": {
"additionalProperties": {},
"description": "Actionable next step for the caller.",
"properties": {
"hint": {
"type": "string"
}
},
"required": [
"hint"
],
"type": "object"
},
"retryable": {
"description": "Whether retrying may succeed.",
"type": "boolean"
}
},
"type": "object"
},
"message": {
"description": "Human-readable description of what went wrong.",
"type": "string"
}
},
"required": [
"code",
"message"
],
"type": "object"
},
"notice": {
"description": "Guidance when the count is zero under a verbatim stateProvince filter, larger than gbif_search_occurrences can page to, or narrowed by a presence/absence filter. Absent when none applies.",
"type": "string"
},
"occurrenceStatus": {
"description": "The presence/absence filter applied upstream — PRESENT, ABSENT, or ANY when no filter was sent. Says what the count covers.",
"type": "string"
}
},
"type": "object"
}
},
{
"description": "Fetch full dataset metadata by UUID key — title, description, citation text, contacts, license, DOI, record count, numConstituents (sub-datasets), and temporal/geographic coverage. Use after gbif_search_datasets or when an occurrence record's datasetKey needs provenance detail. Contacts are capped by contactLimit (default 10); contactsTotal and contactsReturned report the full count.",
"inputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"properties": {
"contactLimit": {
"default": 10,
"description": "Maximum number of contacts to include (default 10, max 100). Set to 0 to omit contact detail while still reporting contactsTotal — useful when citation, license, and record count are all you need from a high-contact dataset like eBird.",
"maximum": 100,
"minimum": 0,
"type": "integer"
},
"datasetKey": {
"description": "Dataset UUID (8-4-4-4-12 hex) from gbif_search_datasets or an occurrence record.",
"type": "string"
}
},
"required": [
"datasetKey"
],
"type": "object"
},
"name": "gbif_get_dataset",
"outputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"anyOf": [
{
"not": {
"required": [
"error"
]
}
},
{
"required": [
"error"
]
}
],
"properties": {
"cap": {
"description": "contactLimit applied when the list was capped. Raise it (max 100) to see more. Absent otherwise.",
"type": "number"
},
"citationText": {
"description": "Full citation text for academic reference. May be absent.",
"type": "string"
},
"contacts": {
"description": "Dataset contacts, capped at contactLimit. Absent when the dataset has no contacts or contactLimit is 0.",
"items": {
"additionalProperties": false,
"description": "A dataset contact with role, name, organization, and email.",
"properties": {
"email": {
"description": "Contact email addresses. May be absent.",
"items": {
"type": "string"
},
"type": "array"
},
"firstName": {
"description": "First name. May be absent.",
"type": "string"
},
"lastName": {
"description": "Last name. May be absent.",
"type": "string"
},
"organization": {
"description": "Organization name. May be absent.",
"type": "string"
},
"type": {
"description": "Contact type (e.g., ADMINISTRATIVE_POINT_OF_CONTACT).",
"type": "string"
}
},
"type": "object"
},
"type": "array"
},
"contactsReturned": {
"description": "Number of contacts included in this response (≤ contactLimit). Present when the dataset has any contacts.",
"type": "number"
},
"contactsTotal": {
"description": "Total contacts on the dataset before applying contactLimit. Present when the dataset has any contacts.",
"type": "number"
},
"description": {
"description": "Full dataset description. May be absent.",
"type": "string"
},
"doi": {
"description": "DOI for citation. May be absent.",
"type": "string"
},
"error": {
"additionalProperties": {},
"description": "Present when the call failed. Absent on success.",
"properties": {
"code": {
"description": "JSON-RPC error code for this failure.",
"maximum": 9007199254740991,
"minimum": -9007199254740991,
"type": "integer"
},
"data": {
"additionalProperties": {},
"properties": {
"reason": {
"description": "Machine-readable failure mode. Declared by this tool: `not_found`: The datasetKey UUID does not match any dataset in GBIF. `invalid_filter`: datasetKey is not a UUID, or GBIF rejected the request as malformed. Other values are possible when a failure originates below the handler.",
"examples": [
"not_found",
"invalid_filter"
],
"type": "string"
},
"recovery": {
"additionalProperties": {},
"description": "Actionable next step for the caller.",
"properties": {
"hint": {
"type": "string"
}
},
"required": [
"hint"
],
"type": "object"
},
"retryable": {
"description": "Whether retrying may succeed.",
"type": "boolean"
}
},
"type": "object"
},
"message": {
"description": "Human-readable description of what went wrong.",
"type": "string"
}
},
"required": [
"code",
"message"
],
"type": "object"
},
"geographicCoverages": {
"description": "Geographic coverage descriptions declared by the dataset. May be absent.",
"items": {
"additionalProperties": false,
"description": "A geographic coverage entry.",
"properties": {
"description": {
"description": "Geographic coverage description (e.g. \"Worldwide\"). May be absent.",
"type": "string"
}
},
"type": "object"
},
"type": "array"
},
"key": {
"description": "Dataset UUID.",
"type": "string"
},
"license": {
"description": "License identifier. May be absent.",
"type": "string"
},
"notice": {
"description": "How to reach the contacts contactLimit held back. Absent when every contact was returned.",
"type": "string"
},
"numConstituents": {
"description": "Number of constituent sub-datasets. May be absent.",
"type": "number"
},
"publishingCountry": {
"description": "Country code of the publishing organization.",
"type": "string"
},
"recordCount": {
"description": "Occurrence records GBIF has indexed for this dataset, matching the figure gbif_search_datasets reports. Spans every occurrenceStatus: absence records — surveys that looked for a taxon and did not find it — are counted alongside sightings, and on some datasets they are the overwhelming majority. gbif_count_occurrences with this datasetKey answers the other question, defaulting to occurrenceStatus PRESENT, so the two figures are expected to differ rather than one being wrong. Fetched separately because the detail endpoint omits it; absent when that lookup does not return in time.",
"type": "number"
},
"shown": {
"description": "Contacts included in this response when the list was capped. Absent otherwise.",
"type": "number"
},
"temporalCoverages": {
"description": "Temporal coverage ranges declared by the dataset. May be absent.",
"items": {
"additionalProperties": false,
"description": "A temporal coverage range.",
"properties": {
"end": {
"description": "Coverage end as an ISO 8601 date-time. May be absent.",
"type": "string"
},
"start": {
"description": "Coverage start as an ISO 8601 date-time. May be absent.",
"type": "string"
}
},
"type": "object"
},
"type": "array"
},
"title": {
"description": "Dataset title.",
"type": "string"
},
"truncated": {
"description": "True when the dataset carries more contacts than contactLimit allowed through. Absent when every contact was returned.",
"type": "boolean"
},
"type": {
"description": "Dataset type (OCCURRENCE, CHECKLIST, etc.).",
"type": "string"
}
},
"type": "object"
}
},
{
"description": "Fetch a single occurrence record by its GBIF occurrence key. Returns the complete Darwin Core record — all coordinates, administrative geography (GADM levels 0–3), dates, collections metadata, collector identifiers, conservation status, media links, and quality issue flags. Check occurrenceStatus before reading the record as a sighting: ABSENT means a survey looked for the taxon and did not find it. Use the occurrence key from gbif_search_occurrences results.",
"inputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"properties": {
"occurrenceKey": {
"description": "GBIF occurrence key from gbif_search_occurrences results.",
"type": "number"
}
},
"required": [
"occurrenceKey"
],
"type": "object"
},
"name": "gbif_get_occurrence",
"outputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"anyOf": [
{
"not": {
"required": [
"error"
]
}
},
{
"required": [
"error"
]
}
],
"properties": {
"basisOfRecord": {
"description": "How the occurrence was recorded.",
"type": "string"
},
"canonicalName": {
"description": "Canonical name without authorship.",
"type": "string"
},
"catalogNumber": {
"description": "Catalog number within the collection. May be absent.",
"type": "string"
},
"class": {
"description": "Class classification. May be absent.",
"type": "string"
},
"classKey": {
"description": "Backbone taxon key for the class. May be absent.",
"type": "number"
},
"collectionCode": {
"description": "Collection code within the institution. May be absent.",
"type": "string"
},
"continent": {
"description": "Continent name. May be absent.",
"type": "string"
},
"coordinateUncertaintyInMeters": {
"description": "Coordinate uncertainty radius in meters. May be absent.",
"type": "number"
},
"country": {
"description": "Country name. May be absent.",
"type": "string"
},
"countryCode": {
"description": "ISO 3166-1 alpha-2 country code. May be absent.",
"type": "string"
},
"datasetKey": {
"description": "UUID of the source dataset.",
"type": "string"
},
"day": {
"description": "Observation day. May be absent.",
"type": "number"
},
"decimalLatitude": {
"description": "Latitude in decimal degrees (WGS84). May be absent.",
"type": "number"
},
"decimalLongitude": {
"description": "Longitude in decimal degrees (WGS84). May be absent.",
"type": "number"
},
"error": {
"additionalProperties": {},
"description": "Present when the call failed. Absent on success.",
"properties": {
"code": {
"description": "JSON-RPC error code for this failure.",
"maximum": 9007199254740991,
"minimum": -9007199254740991,
"type": "integer"
},
"data": {
"additionalProperties": {},
"properties": {
"reason": {
"description": "Machine-readable failure mode. Declared by this tool: `not_found`: The occurrenceKey does not exist in GBIF. `invalid_filter`: GBIF rejected the occurrenceKey as unparseable — a fraction, or a value past the largest integer the endpoint accepts. Other values are possible when a failure originates below the handler.",
"examples": [
"not_found",
"invalid_filter"
],
"type": "string"
},
"recovery": {
"additionalProperties": {},
"description": "Actionable next step for the caller.",
"properties": {
"hint": {
"type": "string"
}
},
"required": [
"hint"
],
"type": "object"
},
"retryable": {
"description": "Whether retrying may succeed.",
"type": "boolean"
}
},
"type": "object"
},
"message": {
"description": "Human-readable description of what went wrong.",
"type": "string"
}
},
"required": [
"code",
"message"
],
"type": "object"
},
"eventDate": {
"description": "Observation date as ISO 8601 string. May be absent.",
"type": "string"
},
"eventTime": {
"description": "Time of day of the observation, with seconds and UTC offset (e.g. 20:15:00+01:00) — the offset eventDate omits when it carries a local time. May be absent.",
"type": "string"
},
"family": {
"description": "Family classification.",
"type": "string"
},
"gadm": {
"additionalProperties": false,
"description": "GADM administrative geography — stable GIDs and names at levels 0–3. May be absent.",
"properties": {
"level0": {
"additionalProperties": false,
"description": "GADM level 0 — country. May be absent.",
"properties": {
"gid": {
"description": "GADM GID — stable administrative-unit identifier (e.g. SWE, SWE.2_1). May be absent.",
"type": "string"
},
"name": {
"description": "Administrative-unit name. May be absent.",
"type": "string"
}
},
"type": "object"
},
"level1": {
"additionalProperties": false,
"description": "GADM level 1 — state/province. May be absent.",
"properties": {
"gid": {
"description": "GADM GID — stable administrative-unit identifier (e.g. SWE, SWE.2_1). May be absent.",
"type": "string"
},
"name": {
"description": "Administrative-unit name. May be absent.",
"type": "string"
}
},
"type": "object"
},
"level2": {
"additionalProperties": false,
"description": "GADM level 2 — county/district. May be absent.",
"properties": {
"gid": {
"description": "GADM GID — stable administrative-unit identifier (e.g. SWE, SWE.2_1). May be absent.",
"type": "string"
},
"name": {
"description": "Administrative-unit name. May be absent.",
"type": "string"
}
},
"type": "object"
},
"level3": {
"additionalProperties": false,
"description": "GADM level 3 — municipality/ward, the finest level GBIF indexes. Absent where the country does not subdivide that far.",
"properties": {
"gid": {
"description": "GADM GID — stable administrative-unit identifier (e.g. SWE, SWE.2_1). May be absent.",
"type": "string"
},
"name": {
"description": "Administrative-unit name. May be absent.",
"type": "string"
}
},
"type": "object"
}
},
"type": "object"
},
"genus": {
"description": "Genus classification.",
"type": "string"
},
"identifiedBy": {
"description": "Identifier name(s). May be absent.",
"type": "string"
},
"identifiers": {
"description": "Alternative record identifiers from the source. May be absent.",
"items": {
"additionalProperties": false,
"description": "An alternative identifier for the occurrence record.",
"properties": {
"identifier": {
"description": "The identifier value. May be absent.",
"type": "string"
},
"type": {
"description": "Identifier type (e.g. URL, DOI, GBIF_PORTAL). May be absent.",
"type": "string"
}
},
"type": "object"
},
"type": "array"
},
"individualCount": {
"description": "Number of individuals. May be absent.",
"type": "number"
},
"institutionCode": {
"description": "Code of the contributing institution. May be absent.",
"type": "string"
},
"issues": {
"description": "GBIF data quality issue flags.",
"items": {
"type": "string"
},
"type": "array"
},
"iucnRedListCategory": {
"description": "IUCN Red List category of the taxon — CR Critically Endangered, EN Endangered, VU Vulnerable, NT Near Threatened, LC Least Concern, DD Data Deficient, EX Extinct, EW Extinct in the Wild, CD Conservation Dependent. May be absent.",
"type": "string"
},
"key": {
"description": "GBIF occurrence key.",
"type": "number"
},
"kingdom": {
"description": "Kingdom classification.",
"type": "string"
},
"lifeStage": {
"description": "Life stage of the individual(s). May be absent.",
"type": "string"
},
"locality": {
"description": "Locality description. May be absent.",
"type": "string"
},
"media": {
"description": "Associated media (images, audio, video). May be absent.",
"items": {
"additionalProperties": false,
"description": "A media item (image, audio, video) associated with the occurrence.",
"properties": {
"format": {
"description": "MIME format of the media.",
"type": "string"
},
"identifier": {
"description": "URL to the media file.",
"type": "string"
},
"license": {
"description": "License for the media.",
"type": "string"
},
"title": {
"description": "Media title.",
"type": "string"
},
"type": {
"description": "Media type (StillImage, Sound, etc.).",
"type": "string"
}
},
"type": "object"
},
"type": "array"
},
"month": {
"description": "Observation month (1–12). May be absent.",
"type": "number"
},
"occurrenceID": {
"description": "Darwin Core occurrenceID — the source record identifier, often a URL back to the origin record. May be absent.",
"type": "string"
},
"occurrenceStatus": {
"description": "PRESENT when the record asserts the taxon was there, ABSENT when it documents a survey that looked and did not find it. An ABSENT record is not a sighting — it carries coordinates, a date, and a recorder all the same. May be absent.",
"type": "string"
},
"order": {
"description": "Order classification.",
"type": "string"
},
"phylum": {
"description": "Phylum classification.",
"type": "string"
},
"publishingCountry": {
"description": "Country code of the publishing organization.",
"type": "string"
},
"recordedBy": {
"description": "Collector name(s). May be absent.",
"type": "string"
},
"scientificName": {
"description": "Scientific name from occurrence record.",
"type": "string"
},
"sex": {
"description": "Sex of the individual(s). May be absent.",
"type": "string"
},
"species": {
"description": "Species canonical name.",
"type": "string"
},
"stateProvince": {
"description": "State or province. May be absent.",
"type": "string"
},
"taxonKey": {
"description": "Backbone taxon key.",
"type": "number"
},
"taxonRank": {
"description": "Taxonomic rank of the identified taxon.",
"type": "string"
},
"taxonomicStatus": {
"description": "Status of the identification carried on this record — ACCEPTED, PROVISIONALLY_ACCEPTED, SYNONYM, DOUBTFUL, and so on. Says whether the occurrence was filed under an accepted name or a synonym. May be absent.",
"type": "string"
},
"year": {
"description": "Observation year. May be absent.",
"type": "number"
}
},
"type": "object"
}
},
{
"description": "Fetch a single backbone taxon by its GBIF taxon key. Returns full classification, authorship, taxonomic status, vernacular name, descendant count, and publication reference. Use after gbif_match_species when you need the complete record rather than the match summary. When taxonomicStatus is SYNONYM, acceptedKey and accepted fields identify the accepted taxon. The extinct field is absent (not false) on most records — only present on explicitly flagged taxa.",
"inputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"properties": {
"taxonKey": {
"description": "GBIF backbone taxon key from gbif_match_species or another taxonomy tool.",
"type": "number"
}
},
"required": [
"taxonKey"
],
"type": "object"
},
"name": "gbif_get_species",
"outputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"anyOf": [
{
"not": {
"required": [
"error"
]
}
},
{
"required": [
"error"
]
}
],
"properties": {
"accepted": {
"description": "Scientific name of the accepted taxon when this record is a synonym.",
"type": "string"
},
"acceptedKey": {
"description": "Backbone key of the accepted taxon when this record is a synonym.",
"type": "number"
},
"authorship": {
"description": "Taxonomic authorship of the name.",
"type": "string"
},
"canonicalName": {
"description": "Scientific name without authorship.",
"type": "string"
},
"class": {
"description": "Class classification.",
"type": "string"
},
"classKey": {
"description": "Taxon key for the class.",
"type": "number"
},
"error": {
"additionalProperties": {},
"description": "Present when the call failed. Absent on success.",
"properties": {
"code": {
"description": "JSON-RPC error code for this failure.",
"maximum": 9007199254740991,
"minimum": -9007199254740991,
"type": "integer"
},
"data": {
"additionalProperties": {},
"properties": {
"reason": {
"description": "Machine-readable failure mode. Declared by this tool: `not_found`: The taxonKey does not exist in the GBIF backbone. `invalid_filter`: GBIF rejected the taxonKey as unparseable — a fraction, or a value outside the 32-bit signed integer range. Other values are possible when a failure originates below the handler.",
"examples": [
"not_found",
"invalid_filter"
],
"type": "string"
},
"recovery": {
"additionalProperties": {},
"description": "Actionable next step for the caller.",
"properties": {
"hint": {
"type": "string"
}
},
"required": [
"hint"
],
"type": "object"
},
"retryable": {
"description": "Whether retrying may succeed.",
"type": "boolean"
}
},
"type": "object"
},
"message": {
"description": "Human-readable description of what went wrong.",
"type": "string"
}
},
"required": [
"code",
"message"
],
"type": "object"
},
"extinct": {
"description": "True when the taxon is explicitly flagged as extinct. Absent on most records.",
"type": "boolean"
},
"family": {
"description": "Family classification.",
"type": "string"
},
"familyKey": {
"description": "Taxon key for the family.",
"type": "number"
},
"genus": {
"description": "Genus classification.",
"type": "string"
},
"genusKey": {
"description": "Taxon key for the genus.",
"type": "number"
},
"key": {
"description": "GBIF backbone taxon key.",
"type": "number"
},
"kingdom": {
"description": "Kingdom classification.",
"type": "string"
},
"kingdomKey": {
"description": "Taxon key for the kingdom.",
"type": "number"
},
"numDescendants": {
"description": "Count of child taxa in the backbone under this taxon.",
"type": "number"
},
"numOccurrences": {
"description": "Occurrence record count in GBIF.",
"type": "number"
},
"order": {
"description": "Order classification.",
"type": "string"
},
"orderKey": {
"description": "Taxon key for the order.",
"type": "number"
},
"parent": {
"description": "Name of the immediate parent taxon.",
"type": "string"
},
"parentKey": {
"description": "Taxon key of the immediate parent.",
"type": "number"
},
"phylum": {
"description": "Phylum classification.",
"type": "string"
},
"phylumKey": {
"description": "Taxon key for the phylum.",
"type": "number"
},
"publishedIn": {
"description": "Original description citation when available.",
"type": "string"
},
"rank": {
"description": "Taxonomic rank (SPECIES, GENUS, FAMILY, etc.).",
"type": "string"
},
"scientificName": {
"description": "Full scientific name with authorship.",
"type": "string"
},
"species": {
"description": "Species canonical name.",
"type": "string"
},
"speciesKey": {
"description": "Taxon key for the species.",
"type": "number"
},
"taxonomicStatus": {
"description": "ACCEPTED, SYNONYM, DOUBTFUL, etc. SYNONYM means acceptedKey/accepted are populated.",
"type": "string"
},
"vernacularName": {
"description": "English common name when available.",
"type": "string"
}
},
"type": "object"
}
},
{
"description": "List direct children of a backbone taxon — genera within a family, species within a genus, subspecies within a species. Paginated. Use gbif_match_species to get the taxonKey first, then iterate with offset for large groups.",
"inputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"properties": {
"limit": {
"default": 20,
"description": "Number of children to return (default 20, max 1000).",
"maximum": 1000,
"minimum": 1,
"type": "number"
},
"offset": {
"default": 0,
"description": "Pagination offset.",
"minimum": 0,
"type": "number"
},
"taxonKey": {
"description": "GBIF backbone taxon key from gbif_match_species or another taxonomy tool.",
"type": "number"
}
},
"required": [
"taxonKey"
],
"type": "object"
},
"name": "gbif_get_species_children",
"outputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"anyOf": [
{
"not": {
"required": [
"error"
]
},
"required": [
"children",
"offset",
"limit",
"endOfRecords"
]
},
{
"required": [
"error"
]
}
],
"properties": {
"cap": {
"description": "Limit applied when the result was truncated. Re-call with offset to page on.",
"type": "number"
},
"children": {
"description": "Direct child taxa.",
"items": {
"additionalProperties": false,
"description": "A direct child taxon with key, name, rank, and status.",
"properties": {
"canonicalName": {
"description": "Scientific name without authorship.",
"type": "string"
},
"key": {
"description": "GBIF backbone taxon key.",
"type": "number"
},
"numDescendants": {
"description": "Count of child taxa under this node.",
"type": "number"
},
"numOccurrences": {
"description": "Occurrence record count.",
"type": "number"
},
"rank": {
"description": "Taxonomic rank.",
"type": "string"
},
"scientificName": {
"description": "Full scientific name with authorship.",
"type": "string"
},
"taxonomicStatus": {
"description": "ACCEPTED, SYNONYM, DOUBTFUL, etc.",
"type": "string"
},
"vernacularName": {
"description": "Common name when available.",
"type": "string"
}
},
"type": "object"
},
"type": "array"
},
"endOfRecords": {
"description": "True when there are no more results after this page.",
"type": "boolean"
},
"error": {
"additionalProperties": {},
"description": "Present when the call failed. Absent on success.",
"properties": {
"code": {
"description": "JSON-RPC error code for this failure.",
"maximum": 9007199254740991,
"minimum": -9007199254740991,
"type": "integer"
},
"data": {
"additionalProperties": {},
"properties": {
"reason": {
"description": "Machine-readable failure mode. Declared by this tool: `not_found`: The taxonKey does not exist in the GBIF backbone. `invalid_filter`: GBIF rejected the taxonKey as unparseable — a fraction, or a value outside the 32-bit signed integer range. Other values are possible when a failure originates below the handler.",
"examples": [
"not_found",
"invalid_filter"
],
"type": "string"
},
"recovery": {
"additionalProperties": {},
"description": "Actionable next step for the caller.",
"properties": {
"hint": {
"type": "string"
}
},
"required": [
"hint"
],
"type": "object"
},
"retryable": {
"description": "Whether retrying may succeed.",
"type": "boolean"
}
},
"type": "object"
},
"message": {
"description": "Human-readable description of what went wrong.",
"type": "string"
}
},
"required": [
"code",
"message"
],
"type": "object"
},
"limit": {
"description": "Records returned in this page.",
"type": "number"
},
"notice": {
"description": "Agent guidance — a no-children note for a valid taxon, or a pagination note when the page was capped. Absent on a complete single page.",
"type": "string"
},
"offset": {
"description": "Current pagination offset.",
"type": "number"
},
"shown": {
"description": "Children returned in this page when the result was truncated.",
"type": "number"
},
"truncated": {
"description": "True when more children exist beyond this page. Absent on the final page.",
"type": "boolean"
}
},
"type": "object"
}
},
{
"description": "Return the parent chain for a taxon — from kingdom (or domain) down to the immediate parent of the queried taxon — as an ordered array. Each entry has its rank, canonical name, and taxon key. The array is returned root-first (kingdom → phylum → class → … → immediate parent of the queried taxon); the queried taxon itself is not included — call gbif_get_species for its own record. Useful for building taxonomic trees or understanding placement without navigating the backbone level-by-level.",
"inputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"properties": {
"taxonKey": {
"description": "GBIF backbone taxon key from gbif_match_species or another taxonomy tool.",
"type": "number"
}
},
"required": [
"taxonKey"
],
"type": "object"
},
"name": "gbif_get_species_classification",
"outputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"anyOf": [
{
"not": {
"required": [
"error"
]
},
"required": [
"classification"
]
},
{
"required": [
"error"
]
}
],
"properties": {
"classification": {
"description": "Classification chain ordered from root (kingdom) to the immediate parent of the queried taxon. The queried taxon itself is not included — call gbif_get_species for its own record.",
"items": {
"additionalProperties": false,
"description": "A single rank entry in the classification chain.",
"properties": {
"key": {
"description": "Backbone taxon key for this rank.",
"type": "number"
},
"name": {
"description": "Canonical name at this rank.",
"type": "string"
},
"rank": {
"description": "Taxonomic rank (KINGDOM, PHYLUM, CLASS, etc.).",
"type": "string"
},
"scientificName": {
"description": "Full scientific name with authorship.",
"type": "string"
}
},
"type": "object"
},
"type": "array"
},
"error": {
"additionalProperties": {},
"description": "Present when the call failed. Absent on success.",
"properties": {
"code": {
"description": "JSON-RPC error code for this failure.",
"maximum": 9007199254740991,
"minimum": -9007199254740991,
"type": "integer"
},
"data": {
"additionalProperties": {},
"properties": {
"reason": {
"description": "Machine-readable failure mode. Declared by this tool: `not_found`: The taxonKey does not exist in the GBIF backbone. `invalid_filter`: GBIF rejected the taxonKey as unparseable — a fraction, or a value outside the 32-bit signed integer range. Other values are possible when a failure originates below the handler.",
"examples": [
"not_found",
"invalid_filter"
],
"type": "string"
},
"recovery": {
"additionalProperties": {},
"description": "Actionable next step for the caller.",
"properties": {
"hint": {
"type": "string"
}
},
"required": [
"hint"
],
"type": "object"
},
"retryable": {
"description": "Whether retrying may succeed.",
"type": "boolean"
}
},
"type": "object"
},
"message": {
"description": "Human-readable description of what went wrong.",
"type": "string"
}
},
"required": [
"code",
"message"
],
"type": "object"
},
"notice": {
"description": "Guidance when the chain is empty because the taxon sits at the root of the backbone. Absent when the chain has entries.",
"type": "string"
}
},
"type": "object"
}
},
{
"description": "Match a scientific name against the GBIF backbone taxonomy. Returns the best-matching taxon with full classification and a confidence score (0–100). This is the mandatory first step for any GBIF workflow — it returns the backbone taxonKey required by gbif_search_occurrences, gbif_count_occurrences, and gbif_occurrence_facets. When the queried name is a synonym, taxonKey is the accepted taxon it resolves to and matchedTaxonKey carries the synonym's own key; occurrence counts differ sharply between the two, so pass taxonKey. Below confidence 80, the match should be reviewed. matchType NONE means no usable match was found — try removing the strict flag or broadening the name.",
"inputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"properties": {
"kingdom": {
"description": "Narrow the match to a specific kingdom (e.g., \"Animalia\", \"Plantae\", \"Fungi\") to disambiguate names that appear in multiple kingdoms. Omit the field to match against the whole backbone — a blank or whitespace-only value is rejected rather than dropped, because GBIF answers one with the undisambiguated match, which is indistinguishable from a match that honored the kingdom.",
"type": "string"
},
"name": {
"description": "Scientific name to match. Examples: \"Parus major\", \"Agaricus bisporus\", \"Homo sapiens\". Fuzzy matching handles minor spelling variations. Common names are not supported — use gbif_search_species for vernacular name searches.",
"type": "string"
},
"rank": {
"description": "Expected taxonomic rank. Use to avoid matching a genus when you expect a species.",
"enum": [
"KINGDOM",
"PHYLUM",
"CLASS",
"ORDER",
"FAMILY",
"GENUS",
"SPECIES",
"SUBSPECIES"
],
"type": "string"
},
"strict": {
"default": false,
"description": "When true, only return an exact match. When false (default), GBIF applies fuzzy matching — useful for minor spelling variations and abbreviated names.",
"type": "boolean"
}
},
"required": [
"name"
],
"type": "object"
},
"name": "gbif_match_species",
"outputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"anyOf": [
{
"not": {
"required": [
"error"
]
}
},
{
"required": [
"error"
]
}
],
"properties": {
"canonicalName": {
"description": "Scientific name without authorship.",
"type": "string"
},
"class": {
"description": "Class of the matched taxon.",
"type": "string"
},
"classKey": {
"description": "Backbone taxon key for the class.",
"type": "number"
},
"confidence": {
"description": "Match confidence score 0–100. Below 80 warrants review.",
"type": "number"
},
"error": {
"additionalProperties": {},
"description": "Present when the call failed. Absent on success.",
"properties": {
"code": {
"description": "JSON-RPC error code for this failure.",
"maximum": 9007199254740991,
"minimum": -9007199254740991,
"type": "integer"
},
"data": {
"additionalProperties": {},
"properties": {
"reason": {
"description": "Machine-readable failure mode. Declared by this tool: `no_match`: matchType is NONE — no candidate met the match threshold. `invalid_filter`: kingdom was supplied blank or whitespace-only, which disambiguates nothing. Other values are possible when a failure originates below the handler.",
"examples": [
"no_match",
"invalid_filter"
],
"type": "string"
},
"recovery": {
"additionalProperties": {},
"description": "Actionable next step for the caller.",
"properties": {
"hint": {
"type": "string"
}
},
"required": [
"hint"
],
"type": "object"
},
"retryable": {
"description": "Whether retrying may succeed.",
"type": "boolean"
}
},
"type": "object"
},
"message": {
"description": "Human-readable description of what went wrong.",
"type": "string"
}
},
"required": [
"code",
"message"
],
"type": "object"
},
"family": {
"description": "Family of the matched taxon.",
"type": "string"
},
"familyKey": {
"description": "Backbone taxon key for the family.",
"type": "number"
},
"genus": {
"description": "Genus of the matched taxon.",
"type": "string"
},
"genusKey": {
"description": "Backbone taxon key for the genus.",
"type": "number"
},
"kingdom": {
"description": "Kingdom of the matched taxon.",
"type": "string"
},
"kingdomKey": {
"description": "Backbone taxon key for the kingdom.",
"type": "number"
},
"matchType": {
"description": "EXACT, FUZZY, HIGHERRANK, or NONE. NONE means no usable match.",
"type": "string"
},
"matchedTaxonKey": {
"description": "Backbone key of the name that actually matched. Present only when it differs from taxonKey — that is, when a synonym was resolved to its accepted taxon.",
"type": "number"
},
"notice": {
"description": "Guidance when the queried name was a synonym and taxonKey was resolved to the accepted taxon. Absent when the matched name is already the accepted one.",
"type": "string"
},
"order": {
"description": "Order of the matched taxon.",
"type": "string"
},
"orderKey": {
"description": "Backbone taxon key for the order.",
"type": "number"
},
"phylum": {
"description": "Phylum of the matched taxon.",
"type": "string"
},
"phylumKey": {
"description": "Backbone taxon key for the phylum.",
"type": "number"
},
"rank": {
"description": "Taxonomic rank of the matched taxon.",
"type": "string"
},
"scientificName": {
"description": "Full scientific name with authorship.",
"type": "string"
},
"species": {
"description": "Species canonical name of the matched taxon.",
"type": "string"
},
"speciesKey": {
"description": "Backbone taxon key for the species.",
"type": "number"
},
"status": {
"description": "Taxonomic status: ACCEPTED, SYNONYM, or DOUBTFUL.",
"type": "string"
},
"taxonKey": {
"description": "GBIF backbone taxon key to pass to downstream tools. The accepted taxon's key when the queried name is a synonym, otherwise the matched taxon's own key.",
"type": "number"
}
},
"type": "object"
}
},
{
"description": "Aggregate occurrence counts across a dimension (COUNTRY, STATE_PROVINCE, YEAR, BASIS_OF_RECORD, DATASET_KEY, KINGDOM_KEY, etc.). Returns one page of facet values ranked by count descending — the top facetLimit at facetOffset 0, a later slice of the same ranking past that. No record payloads returned. Core tool for distribution analysis and trend queries: \"which countries have the most records for this species?\", \"how has observation volume changed since 2010?\". Scope the aggregation with taxonKey, country (uppercase ISO 3166-1 alpha-2), publishingCountry, stateProvince, year, geometry, basisOfRecord, datasetKey, occurrenceStatus, or iucnRedListCategory filters. Also the way to split a result set too large for gbif_search_occurrences to page (offset+limit caps at 100,001): facet by DATASET_KEY, then search each datasetKey on its own. Aggregates sightings only by default, matching gbif_search_occurrences and gbif_count_occurrences; to measure the presence/absence split itself, pass facet OCCURRENCE_STATUS with occurrenceStatus ANY.",
"inputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"properties": {
"basisOfRecord": {
"description": "Scope to a specific basis of record.",
"enum": [
"HUMAN_OBSERVATION",
"MACHINE_OBSERVATION",
"PRESERVED_SPECIMEN",
"LIVING_SPECIMEN",
"MATERIAL_SAMPLE",
"MATERIAL_CITATION",
"OCCURRENCE",
"LITERATURE"
],
"type": "string"
},
"country": {
"description": "ISO 3166-1 alpha-2 code, uppercase, of where the occurrence was recorded, to scope to one country. Not the publisher's country — that is publishingCountry, and the two disagree on most records. Scope to one country, or pass back a value this tool returned under facet COUNTRY to drill into that bucket. Lowercase and alpha-3 forms (\"gb\", \"USA\") match nothing upstream, which is why only the uppercase two-letter form is accepted here.",
"pattern": "^[A-Z]{2}$",
"type": "string"
},
"datasetKey": {
"description": "Scope the aggregation to a single dataset by its GBIF dataset UUID (8-4-4-4-12 hex). Obtain one from gbif_search_datasets, gbif_get_dataset, a DATASET_KEY facet, or the datasetKey field on an occurrence record. Omit the field to aggregate across every dataset — an empty string is rejected rather than read as no scope, because GBIF answers a blank datasetKey with the unfiltered aggregation.",
"type": "string"
},
"facet": {
"description": "Dimension to aggregate by (e.g., COUNTRY, YEAR, BASIS_OF_RECORD, SPECIES_KEY, OCCURRENCE_STATUS, IUCN_RED_LIST_CATEGORY). DATASET_KEY is the dimension to split on when a result set is too large to page: every occurrence carries exactly one datasetKey, so its buckets sum to totalOccurrences with no gap and no overlap, and it has the cardinality to cut a large scope into pageable pieces. BASIS_OF_RECORD and PUBLISHING_COUNTRY are gap-free too and both have a matching filter on the occurrence tools, so either can drive a further split of a bucket still too large — but on that same scope they return 9 and 41 buckets against DATASET_KEY's 550, so neither replaces it as the first cut. A dimension a record can lack silently drops that record: faceting one 60,290,950-record scope by YEAR returned 224 buckets summing to 59,407,400, leaving 883,550 undated records in no bucket at all, and MONTH, STATE_PROVINCE, and SPECIES_KEY lose records the same way — stateProvince included, even though the occurrence tools can now filter on it. Sums are comparable only across the same occurrenceStatus scope.",
"enum": [
"BASIS_OF_RECORD",
"COUNTRY",
"STATE_PROVINCE",
"YEAR",
"DATASET_KEY",
"KINGDOM_KEY",
"PHYLUM_KEY",
"CLASS_KEY",
"ORDER_KEY",
"FAMILY_KEY",
"GENUS_KEY",
"SPECIES_KEY",
"PUBLISHING_COUNTRY",
"MONTH",
"OCCURRENCE_STATUS",
"IUCN_RED_LIST_CATEGORY"
],
"type": "string"
},
"facetLimit": {
"default": 10,
"description": "Maximum number of facet values to return (default 10, max 100).",
"maximum": 100,
"minimum": 1,
"type": "number"
},
"facetOffset": {
"default": 0,
"description": "Zero-based offset into the ranked facet values, for paging past the first facetLimit values on high-cardinality dimensions like DATASET_KEY. Advance by facetLimit to fetch the next page (0, then facetLimit, then 2×facetLimit, …).",
"minimum": 0,
"type": "number"
},
"geometry": {
"description": "WKT polygon to scope the aggregation to a geographic area (e.g., POLYGON((8 47, 9 47, 9 48, 8 48, 8 47))). Coordinates are longitude latitude. Omit the field to aggregate everywhere — a blank or whitespace-only value is rejected rather than dropped, because GBIF answers one with the unfiltered aggregation.",
"type": "string"
},
"iucnRedListCategory": {
"description": "Scope to records whose taxon carries this IUCN Red List category: CR Critically Endangered, EN Endangered, VU Vulnerable, NT Near Threatened, LC Least Concern, DD Data Deficient, EX Extinct, EW Extinct in the Wild, CD Conservation Dependent. Leave unset and facet on IUCN_RED_LIST_CATEGORY to see the whole distribution instead.",
"enum": [
"CR",
"EN",
"VU",
"NT",
"LC",
"DD",
"EX",
"EW",
"CD"
],
"type": "string"
},
"occurrenceStatus": {
"default": "PRESENT",
"description": "Presence/absence scope. Defaults to PRESENT so the aggregation counts sightings, not the surveys that looked and found nothing, and agrees with gbif_count_occurrences on the same filters. Use ANY for both — required to see both buckets when facet is OCCURRENCE_STATUS — or ABSENT for non-observations alone.",
"enum": [
"PRESENT",
"ABSENT",
"ANY"
],
"type": "string"
},
"publishingCountry": {
"description": "ISO 3166-1 alpha-2 code, uppercase, of the organization that published the record — not where the occurrence was observed, which is country. Scope to one publisher country, or pass back a value this tool returned under facet PUBLISHING_COUNTRY to drill into that bucket. Lowercase and alpha-3 forms (\"us\", \"USA\") match nothing upstream, which is why only the uppercase two-letter form is accepted here.",
"pattern": "^[A-Z]{2}$",
"type": "string"
},
"stateProvince": {
"description": "State, province, or first-level administrative division, matched as a verbatim string — exact and case-sensitive. Pass back a value this tool returned under facet STATE_PROVINCE rather than a guessed one: GBIF stores what each dataset recorded without normalizing it, so \"England\", \"England - Greater London\", and \"Greater London\" are three distinct values, \"england\" is none of them, and an unmatched value aggregates zero records rather than erroring. Omit the field to aggregate across every state or province — a blank or whitespace-only value is rejected rather than dropped, because GBIF answers one with the unfiltered aggregation.",
"type": "string"
},
"taxonKey": {
"description": "Backbone taxon key to scope the aggregation. Matches the given taxon and all descendant taxa (subspecies, varieties, etc.).",
"type": "number"
},
"year": {
"description": "Year or year range (e.g., \"2020,2024\") to scope the aggregation. Both endpoints inclusive. Omit the field to aggregate across every year — a blank or whitespace-only value is rejected rather than dropped, because GBIF answers one with the unfiltered aggregation.",
"type": "string"
}
},
"required": [
"facet"
],
"type": "object"
},
"name": "gbif_occurrence_facets",
"outputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"anyOf": [
{
"not": {
"required": [
"error"
]
},
"required": [
"facet",
"totalOccurrences",
"counts",
"facetLimit",
"facetOffset",
"occurrenceStatus"
]
},
{
"required": [
"error"
]
}
],
"properties": {
"cap": {
"description": "facetLimit applied when the page was capped. Re-call with facetOffset advanced by this value to page on. Absent otherwise.",
"type": "number"
},
"counts": {
"description": "Facet values ranked by count descending — one page of up to facetLimit entries starting at facetOffset, not necessarily the top ones.",
"items": {
"additionalProperties": false,
"description": "A facet value with its occurrence count.",
"properties": {
"count": {
"description": "Occurrence count for this facet value.",
"type": "number"
},
"name": {
"description": "Facet value (country code, year, basisOfRecord, etc.).",
"type": "string"
}
},
"required": [
"name",
"count"
],
"type": "object"
},
"type": "array"
},
"error": {
"additionalProperties": {},
"description": "Present when the call failed. Absent on success.",
"properties": {
"code": {
"description": "JSON-RPC error code for this failure.",
"maximum": 9007199254740991,
"minimum": -9007199254740991,
"type": "integer"
},
"data": {
"additionalProperties": {},
"properties": {
"reason": {
"description": "Machine-readable failure mode. Declared by this tool: `invalid_filter`: A scope filter was supplied blank or whitespace-only, datasetKey is not an 8-4-4-4-12 hex UUID, a two-letter country or publishingCountry code is one GBIF does not know, or GBIF rejected the geometry or year scope as malformed. Other values are possible when a failure originates below the handler.",
"examples": [
"invalid_filter"
],
"type": "string"
},
"recovery": {
"additionalProperties": {},
"description": "Actionable next step for the caller.",
"properties": {
"hint": {
"type": "string"
}
},
"required": [
"hint"
],
"type": "object"
},
"retryable": {
"description": "Whether retrying may succeed.",
"type": "boolean"
}
},
"type": "object"
},
"message": {
"description": "Human-readable description of what went wrong.",
"type": "string"
}
},
"required": [
"code",
"message"
],
"type": "object"
},
"facet": {
"description": "The facet dimension aggregated.",
"type": "string"
},
"facetLimit": {
"description": "Maximum facet values requested.",
"type": "number"
},
"facetOffset": {
"description": "Zero-based offset applied to the ranked facet values.",
"type": "number"
},
"notice": {
"description": "Guidance when no facet values were returned, the page came back full and more values may remain, a verbatim stateProvince filter matched nothing, or a presence/absence filter narrowed the aggregation. Absent only when none applies.",
"type": "string"
},
"occurrenceStatus": {
"description": "The presence/absence filter applied upstream — PRESENT, ABSENT, or ANY when no filter was sent. Says what totalOccurrences and every bucket cover.",
"type": "string"
},
"shown": {
"description": "Facet values returned in this page when the page was capped. Absent otherwise.",
"type": "number"
},
"totalOccurrences": {
"description": "Total matching occurrences across all facet values.",
"type": "number"
},
"truncated": {
"description": "Heuristic continuation flag: present and true when this page returned a full facetLimit of values, so more distinct values may exist past facetOffset + facetLimit. GBIF exposes no total distinct-value count, so this is an estimate, not exact. Absent when the page came back short, which is the only proof the ranking is exhausted.",
"type": "boolean"
}
},
"type": "object"
}
},
{
"description": "Search GBIF datasets by keyword, type, publishing country (uppercase ISO 3166-1 alpha-2), publishing organization, or hosting organization. The two organization filters answer different questions — publishingOrg matches the organization whose data it is, hostingOrg the organization whose installation serves it — and an organization key from gbif_search_publishers usually wants publishingOrg. Returns dataset title, description, license, record count, and DOI. Use to find the source dataset behind a set of records, or to explore what data collections are available for a taxon, country, or organization.",
"inputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"properties": {
"hostingOrg": {
"description": "UUID (8-4-4-4-12 hex, lowercase — matched case-sensitively, as publishingOrg is) of the organization whose installation serves the dataset — not the organization that published it, which is publishingOrg. Most organizations publish through an installation someone else runs, so a key from gbif_search_publishers matches nothing here for them: of the first 25 GB organizations the registry lists, all 25 host no datasets while 13 publish one or two. Supplied together the two filters are intersected, not combined.",
"type": "string"
},
"limit": {
"default": 20,
"description": "Number of datasets to return (default 20, max 1000).",
"maximum": 1000,
"minimum": 1,
"type": "number"
},
"offset": {
"default": 0,
"description": "Pagination offset.",
"minimum": 0,
"type": "number"
},
"publishingCountry": {
"description": "ISO 3166-1 alpha-2 code, uppercase, of the organization that published the dataset (e.g., \"GB\", \"US\", \"DE\", \"SE\"). Lowercase and alpha-3 forms (\"gb\", \"GBR\") match nothing upstream, which is why only the uppercase two-letter form is accepted here — unlike the country filter on gbif_search_publishers, which resolves either form. Take a value from a PUBLISHING_COUNTRY facet on gbif_occurrence_facets; an uppercase pair GBIF does not assign (\"XX\") is rejected upstream by name.",
"pattern": "^[A-Z]{2}$",
"type": "string"
},
"publishingOrg": {
"description": "UUID (8-4-4-4-12 hex, lowercase — GBIF matches these two keys case-sensitively, so an upper-cased rendering of a real key matches nothing) of the organization that published the dataset — the organization whose data it is, and the question a key from gbif_search_publishers is usually asking. Not the organization that serves it, which is hostingOrg and matches a different set: Butterfly Conservation (0d72dd7f-6f05-46af-85c2-8b6e77ce5534) publishes 3 datasets and hosts none, while the National Biodiversity Network (07f617d0-c688-11d8-bf62-b8a03c50a862) hosts 984 — those 3 among them — and publishes 1. Supplied together the two filters are intersected, not combined, so the same key in both fields returns only what that organization both published and serves.",
"type": "string"
},
"q": {
"description": "Free-text search across dataset title and description. Omit the field to browse without a term — a blank or whitespace-only value is rejected rather than sent, because GBIF answers a blank one with all 123,527 indexed datasets and a whitespace-only one with none, and neither is the search a caller who filled the field was asking for.",
"type": "string"
},
"type": {
"description": "Filter by dataset type. OCCURRENCE for observation records, CHECKLIST for species lists.",
"enum": [
"OCCURRENCE",
"CHECKLIST",
"METADATA",
"SAMPLING_EVENT"
],
"type": "string"
}
},
"type": "object"
},
"name": "gbif_search_datasets",
"outputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"anyOf": [
{
"not": {
"required": [
"error"
]
},
"required": [
"datasets",
"totalCount",
"offset",
"limit",
"endOfRecords"
]
},
{
"required": [
"error"
]
}
],
"properties": {
"datasets": {
"description": "Matching datasets.",
"items": {
"additionalProperties": false,
"description": "A GBIF dataset with key, title, type, license, and record count.",
"properties": {
"description": {
"description": "Brief description, truncated to a 300-character preview. May be absent.",
"type": "string"
},
"descriptionTruncated": {
"description": "True when the description was shortened to the 300-char preview; call gbif_get_dataset with this key for the full text. Omitted when the dataset has no description.",
"type": "boolean"
},
"doi": {
"description": "DOI for citation. May be absent.",
"type": "string"
},
"key": {
"description": "Dataset UUID for gbif_get_dataset chaining.",
"type": "string"
},
"license": {
"description": "License identifier. May be absent.",
"type": "string"
},
"publishingCountry": {
"description": "Country code of the publisher.",
"type": "string"
},
"recordCount": {
"description": "Occurrence records GBIF has indexed for this dataset, spanning every occurrenceStatus: absence records — surveys that looked for a taxon and did not find it — are counted alongside sightings, and on some datasets they are the overwhelming majority. For the sightings-only figure, call gbif_count_occurrences with this key; it defaults to occurrenceStatus PRESENT, so the two figures are expected to differ rather than one being wrong.",
"type": "number"
},
"title": {
"description": "Dataset title.",
"type": "string"
},
"type": {
"description": "Dataset type (OCCURRENCE, CHECKLIST, etc.).",
"type": "string"
}
},
"type": "object"
},
"type": "array"
},
"endOfRecords": {
"description": "True when there are no more results after this page.",
"type": "boolean"
},
"error": {
"additionalProperties": {},
"description": "Present when the call failed. Absent on success.",
"properties": {
"code": {
"description": "JSON-RPC error code for this failure.",
"maximum": 9007199254740991,
"minimum": -9007199254740991,
"type": "integer"
},
"data": {
"additionalProperties": {},
"properties": {
"reason": {
"description": "Machine-readable failure mode. Declared by this tool: `invalid_filter`: A filter was supplied blank or whitespace-only, publishingOrg or hostingOrg is not a lowercase 8-4-4-4-12 hex UUID, publishingCountry is a two-letter code GBIF does not assign, or GBIF rejected another filter value as malformed. Other values are possible when a failure originates below the handler.",
"examples": [
"invalid_filter"
],
"type": "string"
},
"recovery": {
"additionalProperties": {},
"description": "Actionable next step for the caller.",
"properties": {
"hint": {
"type": "string"
}
},
"required": [
"hint"
],
"type": "object"
},
"retryable": {
"description": "Whether retrying may succeed.",
"type": "boolean"
}
},
"type": "object"
},
"message": {
"description": "Human-readable description of what went wrong.",
"type": "string"
}
},
"required": [
"code",
"message"
],
"type": "object"
},
"limit": {
"description": "Datasets returned in this page.",
"type": "number"
},
"notice": {
"description": "Guidance when results are empty or paging overshot. Absent on successful result pages.",
"type": "string"
},
"offset": {
"description": "Current pagination offset.",
"type": "number"
},
"totalCount": {
"description": "Total matching datasets before pagination.",
"type": "number"
}
},
"type": "object"
}
},
{
"description": "Search 3.9B+ GBIF occurrence records with Darwin Core filters. Use taxonKey from gbif_match_species for reliable results — it resolves synonyms automatically. Accepts country (uppercase ISO 3166-1 alpha-2, where the record was observed), publishingCountry (the publishing organization's country — a different question), stateProvince, bounding box (decimalLatitude/decimalLongitude ranges), WKT polygon geometry, year range, month, basis of record, coordinate filter, and dataset key. Returns sightings only by default — GBIF also indexes absence records (surveys that looked and found nothing), and occurrenceStatus controls whether they are included. Pagination is capped at offset+limit=100,001 and GBIF offers no cursor or scroll, so a larger result set is covered only by partitioning it — facet it by DATASET_KEY with gbif_occurrence_facets and search each datasetKey separately. This server cannot download a result set in bulk; that needs the GBIF Download API with a GBIF.org account, or the GBIF snapshot on AWS Open Data.",
"inputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"properties": {
"basisOfRecord": {
"description": "Filter by how the occurrence was recorded. HUMAN_OBSERVATION covers citizen science. PRESERVED_SPECIMEN covers natural history collections.",
"enum": [
"HUMAN_OBSERVATION",
"MACHINE_OBSERVATION",
"PRESERVED_SPECIMEN",
"LIVING_SPECIMEN",
"MATERIAL_SAMPLE",
"MATERIAL_CITATION",
"OCCURRENCE",
"LITERATURE"
],
"type": "string"
},
"coordinateUncertaintyInMeters": {
"description": "Filter by coordinate uncertainty radius in meters. Range format: \"min,max\" (e.g., \"0,1000\" for sub-kilometer precision). Both endpoints inclusive. Omit the field to accept any uncertainty — a blank or whitespace-only value is rejected rather than dropped, because GBIF answers one with the unfiltered result set.",
"type": "string"
},
"country": {
"description": "ISO 3166-1 alpha-2 code, uppercase, of where the occurrence was recorded (e.g., \"GB\", \"US\", \"DE\", \"SE\"). Not the publisher's country — that is publishingCountry, and the two disagree on most records. Lowercase and alpha-3 forms (\"gb\", \"USA\") match nothing upstream, which is why only the uppercase two-letter form is accepted here. Take a value from a COUNTRY facet on gbif_occurrence_facets; an uppercase pair GBIF does not know (\"XX\") is rejected upstream by name.",
"pattern": "^[A-Z]{2}$",
"type": "string"
},
"datasetKey": {
"description": "Restrict results to a single dataset by its GBIF dataset UUID (8-4-4-4-12 hex). Obtain one from gbif_search_datasets, gbif_get_dataset, a DATASET_KEY facet (gbif_occurrence_facets), or the datasetKey field on an occurrence record. Omit the field to search every dataset — an empty string is rejected rather than read as no filter, because GBIF answers a blank datasetKey with the unfiltered result set.",
"type": "string"
},
"decimalLatitude": {
"description": "Latitude range as \"min,max\" (e.g., \"47.0,48.5\"). Decimal degrees, WGS84. Combine with decimalLongitude for a bounding box. Omit the field to leave latitude unbounded — a blank or whitespace-only value is rejected rather than dropped, because GBIF answers one with the unfiltered result set.",
"type": "string"
},
"decimalLongitude": {
"description": "Longitude range as \"min,max\" (e.g., \"8.0,9.5\"). Decimal degrees, WGS84. Combine with decimalLatitude for a bounding box. Omit the field to leave longitude unbounded — a blank or whitespace-only value is rejected rather than dropped, because GBIF answers one with the unfiltered result set.",
"type": "string"
},
"geometry": {
"description": "WKT polygon for geographic filtering (e.g., POLYGON((8 47, 9 47, 9 48, 8 48, 8 47))). Coordinates are longitude latitude. Takes precedence over decimalLatitude/decimalLongitude. Omit the field to search everywhere — a blank or whitespace-only value is rejected rather than dropped, because GBIF answers one with the unfiltered result set.",
"type": "string"
},
"hasCoordinate": {
"description": "When true, return only georeferenced records (those with coordinates). When false, return ONLY records without coordinates. Omit the parameter entirely to include all records regardless of coordinate presence.",
"type": "boolean"
},
"isInCluster": {
"description": "Filter to records flagged as likely duplicates (true) or exclude them (false). Omit to include all. Note: GBIF does not expose a cluster identifier — only the membership flag. To de-duplicate, set isInCluster: false to exclude all clustered records.",
"type": "boolean"
},
"iucnRedListCategory": {
"description": "Restrict to records whose taxon carries this IUCN Red List category: CR Critically Endangered, EN Endangered, VU Vulnerable, NT Near Threatened, LC Least Concern, DD Data Deficient, EX Extinct, EW Extinct in the Wild, CD Conservation Dependent. Records with no category are excluded when this is set.",
"enum": [
"CR",
"EN",
"VU",
"NT",
"LC",
"DD",
"EX",
"EW",
"CD"
],
"type": "string"
},
"limit": {
"default": 20,
"description": "Number of records to return (default 20, max 300).",
"maximum": 300,
"minimum": 1,
"type": "number"
},
"month": {
"description": "Calendar month (1–12). Useful for seasonal distribution queries.",
"maximum": 12,
"minimum": 1,
"type": "number"
},
"occurrenceStatus": {
"default": "PRESENT",
"description": "Presence/absence filter. Defaults to PRESENT: an ABSENT record documents a survey that looked for the taxon and did not find it, so including one would read as a sighting of the opposite. Use ANY for both (GBIF's own default), or ABSENT for non-observations alone.",
"enum": [
"PRESENT",
"ABSENT",
"ANY"
],
"type": "string"
},
"offset": {
"default": 0,
"description": "Pagination offset. GBIF serves offset+limit up to 100,001 and rejects anything past it, with no cursor or scroll to continue from. To reach a result set larger than that, split it into per-datasetKey searches using a DATASET_KEY facet from gbif_occurrence_facets — gap-free and high-cardinality, unlike YEAR, which leaves undated records in no bucket — rather than paging deeper.",
"minimum": 0,
"type": "number"
},
"publishingCountry": {
"description": "ISO 3166-1 alpha-2 code, uppercase, of the organization that published the record — not where the occurrence was observed, which is country. The two differ constantly: of 60,290,950 records observed in GB, 1,548,928 were published by US organizations. Take a value from a PUBLISHING_COUNTRY facet on gbif_occurrence_facets. Lowercase and alpha-3 forms (\"us\", \"USA\") match nothing upstream, which is why only the uppercase two-letter form is accepted here.",
"pattern": "^[A-Z]{2}$",
"type": "string"
},
"scientificName": {
"description": "Scientific name filter. Less precise than taxonKey — does not match synonyms. Use taxonKey from gbif_match_species for reliable results. Supplying both does not narrow the search: GBIF combines the two taxon filters with OR, so the result is the union of the two, not their intersection. Omit the field to search every name — a blank or whitespace-only value is rejected rather than dropped, because GBIF answers one with the unfiltered result set.",
"type": "string"
},
"stateProvince": {
"description": "State, province, or first-level administrative division, matched as a verbatim string — exact and case-sensitive. GBIF stores what each dataset recorded without normalizing it, so there is no vocabulary to guess from: \"England\", \"England - Greater London\", and \"Greater London\" are three distinct values, and \"england\" is none of them. Take one from a STATE_PROVINCE facet on gbif_occurrence_facets scoped the same way and pass it back unchanged — an unmatched value returns zero records rather than an error. Omit the field to search every state or province — a blank or whitespace-only value is rejected rather than dropped, because GBIF answers one with the unfiltered result set. Records carrying no stateProvince match no value, so this cannot partition a scope.",
"type": "string"
},
"taxonKey": {
"description": "GBIF backbone taxon key from gbif_match_species. Preferred over scientificName — matches all synonyms automatically. Matches the given taxon and all descendant taxa (subspecies, varieties, etc.).",
"type": "number"
},
"year": {
"description": "Year or year range. Single year: \"2024\". Range: \"2020,2024\". Filters by observation year. Both endpoints inclusive. Omit the field to search every year — a blank or whitespace-only value is rejected rather than dropped, because GBIF answers one with the unfiltered result set.",
"type": "string"
}
},
"type": "object"
},
"name": "gbif_search_occurrences",
"outputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"anyOf": [
{
"not": {
"required": [
"error"
]
},
"required": [
"occurrences",
"totalCount",
"offset",
"limit",
"endOfRecords",
"occurrenceStatus"
]
},
{
"required": [
"error"
]
}
],
"properties": {
"endOfRecords": {
"description": "True when there are no more results after this page.",
"type": "boolean"
},
"error": {
"additionalProperties": {},
"description": "Present when the call failed. Absent on success.",
"properties": {
"code": {
"description": "JSON-RPC error code for this failure.",
"maximum": 9007199254740991,
"minimum": -9007199254740991,
"type": "integer"
},
"data": {
"additionalProperties": {},
"properties": {
"reason": {
"description": "Machine-readable failure mode. Declared by this tool: `pagination_cap_exceeded`: offset + limit exceeds 100,001, the deepest page GBIF serves. `invalid_filter`: A filter value is unusable — any filter supplied blank or whitespace-only, a datasetKey that is not an 8-4-4-4-12 hex UUID, a two-letter country or publishingCountry code GBIF does not know, or a WKT geometry or range GBIF rejects. Other values are possible when a failure originates below the handler.",
"examples": [
"pagination_cap_exceeded",
"invalid_filter"
],
"type": "string"
},
"recovery": {
"additionalProperties": {},
"description": "Actionable next step for the caller.",
"properties": {
"hint": {
"type": "string"
}
},
"required": [
"hint"
],
"type": "object"
},
"retryable": {
"description": "Whether retrying may succeed.",
"type": "boolean"
}
},
"type": "object"
},
"message": {
"description": "Human-readable description of what went wrong.",
"type": "string"
}
},
"required": [
"code",
"message"
],
"type": "object"
},
"limit": {
"description": "Records returned in this page.",
"type": "number"
},
"notice": {
"description": "Guidance when results are empty, paging overshot, the match is larger than the pagination cap can reach, or a presence/absence filter narrowed the result. Absent only when none applies.",
"type": "string"
},
"occurrenceStatus": {
"description": "The presence/absence filter applied upstream — PRESENT, ABSENT, or ANY when no filter was sent. Says what totalCount and the returned records cover.",
"type": "string"
},
"occurrences": {
"description": "Occurrence records matching the filters.",
"items": {
"additionalProperties": false,
"description": "A single occurrence record with location, taxon, date, and provenance fields.",
"properties": {
"basisOfRecord": {
"description": "How the occurrence was recorded.",
"type": "string"
},
"canonicalName": {
"description": "Canonical name without authorship.",
"type": "string"
},
"coordinateUncertaintyInMeters": {
"description": "Coordinate uncertainty radius in meters. May be absent.",
"type": "number"
},
"country": {
"description": "Country name. May be absent.",
"type": "string"
},
"countryCode": {
"description": "ISO 3166-1 alpha-2 country code. May be absent.",
"type": "string"
},
"datasetKey": {
"description": "UUID of the source dataset.",
"type": "string"
},
"datasetName": {
"description": "Name of the source dataset. May be absent.",
"type": "string"
},
"day": {
"description": "Observation day. May be absent.",
"type": "number"
},
"decimalLatitude": {
"description": "Latitude in decimal degrees (WGS84). May be absent.",
"type": "number"
},
"decimalLongitude": {
"description": "Longitude in decimal degrees (WGS84). May be absent.",
"type": "number"
},
"eventDate": {
"description": "Observation date as ISO 8601 string. May be absent.",
"type": "string"
},
"eventTime": {
"description": "Time of day of the observation, with seconds and UTC offset (e.g. 20:15:00+01:00) — the offset eventDate omits when it carries a local time. May be absent.",
"type": "string"
},
"individualCount": {
"description": "Number of individuals. May be absent.",
"type": "number"
},
"issues": {
"description": "GBIF data quality issue flags for this record.",
"items": {
"type": "string"
},
"type": "array"
},
"iucnRedListCategory": {
"description": "IUCN Red List category of the taxon — CR, EN, VU, NT, LC, DD, EX, EW, or CD. May be absent.",
"type": "string"
},
"key": {
"description": "GBIF occurrence key for gbif_get_occurrence chaining.",
"type": "number"
},
"locality": {
"description": "Locality description. May be absent.",
"type": "string"
},
"month": {
"description": "Observation month (1–12). May be absent.",
"type": "number"
},
"occurrenceStatus": {
"description": "PRESENT when the record asserts the taxon was there, ABSENT when it documents a survey that looked and did not find it. An ABSENT record is not a sighting. May be absent.",
"type": "string"
},
"publishingCountry": {
"description": "Country code of the publishing organization.",
"type": "string"
},
"rank": {
"description": "Taxonomic rank of the identified taxon.",
"type": "string"
},
"recordedBy": {
"description": "Collector name(s). May be absent.",
"type": "string"
},
"scientificName": {
"description": "Scientific name from occurrence record.",
"type": "string"
},
"stateProvince": {
"description": "State or province name. May be absent.",
"type": "string"
},
"taxonKey": {
"description": "Backbone taxon key.",
"type": "number"
},
"taxonomicStatus": {
"description": "Status of the identification carried on this record — ACCEPTED, PROVISIONALLY_ACCEPTED, SYNONYM, DOUBTFUL, and so on. Says whether the occurrence was filed under an accepted name or a synonym. May be absent.",
"type": "string"
},
"year": {
"description": "Observation year. May be absent.",
"type": "number"
}
},
"type": "object"
},
"type": "array"
},
"offset": {
"description": "Current pagination offset.",
"type": "number"
},
"totalCount": {
"description": "Total matching occurrences before pagination.",
"type": "number"
}
},
"type": "object"
}
},
{
"description": "Search organizations registered with GBIF by name fragment or country. Returns organization key, title, and country — sufficient to chain into gbif_search_datasets as publishingOrg for the datasets an organization published, or as hostingOrg for the ones its own installation serves, or to understand who publishes data for a region. publishingOrg is the usual chain: most organizations publish through an installation someone else runs, so hostingOrg matches nothing for them.",
"inputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"properties": {
"country": {
"description": "ISO 3166-1 country code to filter organizations by country. The alpha-2 form (\"GB\") is canonical; unlike the country codes on the occurrence tools and gbif_search_datasets, this one also resolves the alpha-3 form (\"GBR\") and is case-insensitive, because the registry endpoint matches the parsed country rather than the string. A value GBIF cannot parse as a country errors rather than returning an empty list. Omit the field to search every country — an empty string is rejected rather than read as no filter, because the registry answers a blank country with all 3,561 organizations.",
"type": "string"
},
"limit": {
"default": 20,
"description": "Number of organizations to return (default 20, max 1000).",
"maximum": 1000,
"minimum": 1,
"type": "number"
},
"offset": {
"default": 0,
"description": "Pagination offset.",
"minimum": 0,
"type": "number"
},
"q": {
"description": "Name fragment to search for. Matches organization names. Omit the field to browse without a term — a blank or whitespace-only value is rejected rather than sent, because the registry answers either with all 3,561 registered organizations.",
"type": "string"
}
},
"type": "object"
},
"name": "gbif_search_publishers",
"outputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"anyOf": [
{
"not": {
"required": [
"error"
]
},
"required": [
"publishers",
"totalCount",
"offset",
"limit",
"endOfRecords"
]
},
{
"required": [
"error"
]
}
],
"properties": {
"endOfRecords": {
"description": "True when there are no more results after this page.",
"type": "boolean"
},
"error": {
"additionalProperties": {},
"description": "Present when the call failed. Absent on success.",
"properties": {
"code": {
"description": "JSON-RPC error code for this failure.",
"maximum": 9007199254740991,
"minimum": -9007199254740991,
"type": "integer"
},
"data": {
"additionalProperties": {},
"properties": {
"reason": {
"description": "Machine-readable failure mode. Declared by this tool: `invalid_filter`: q was supplied blank or whitespace-only, country is the empty string, or GBIF could not parse the country value as a country. Other values are possible when a failure originates below the handler.",
"examples": [
"invalid_filter"
],
"type": "string"
},
"recovery": {
"additionalProperties": {},
"description": "Actionable next step for the caller.",
"properties": {
"hint": {
"type": "string"
}
},
"required": [
"hint"
],
"type": "object"
},
"retryable": {
"description": "Whether retrying may succeed.",
"type": "boolean"
}
},
"type": "object"
},
"message": {
"description": "Human-readable description of what went wrong.",
"type": "string"
}
},
"required": [
"code",
"message"
],
"type": "object"
},
"limit": {
"description": "Organizations returned in this page.",
"type": "number"
},
"notice": {
"description": "Guidance when results are empty or paging overshot. Absent on successful result pages.",
"type": "string"
},
"offset": {
"description": "Current pagination offset.",
"type": "number"
},
"publishers": {
"description": "Matching organizations.",
"items": {
"additionalProperties": false,
"description": "A GBIF-registered publishing organization.",
"properties": {
"city": {
"description": "City. May be absent.",
"type": "string"
},
"country": {
"description": "ISO 3166-1 alpha-2 country code.",
"type": "string"
},
"key": {
"description": "Organization UUID. Chains into gbif_search_datasets as publishingOrg for the datasets this organization published, or as hostingOrg for the ones its own installation serves — publishingOrg is the usual one, since most organizations host nothing.",
"type": "string"
},
"title": {
"description": "Organization name.",
"type": "string"
}
},
"type": "object"
},
"type": "array"
},
"totalCount": {
"description": "Total matching organizations before pagination.",
"type": "number"
}
},
"type": "object"
}
},
{
"description": "Search or browse the GBIF backbone taxonomy. Accepts scientific name fragments, rank filters, and higher-taxon constraints. Useful for exploring what species exist under a higher taxon (e.g., \"list all families of Coleoptera\"), for simple name-fragment searches, or when gbif_match_species returns too narrow a result. kingdom, family, and genus scope the browse to a higher taxon: each is resolved to its backbone key before the search runs, so the narrowest one supplied is what scopes, an alternative name resolves to the taxon it is a synonym of, and a name that matches no backbone taxon at that rank fails rather than returning the whole index. Names are capitalized as GBIF writes them (\"Paridae\", not \"paridae\") and are matched exactly, not fuzzily. Paginated — use limit and offset to walk through results.",
"inputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"properties": {
"datasetKey": {
"description": "Scope to a specific checklist dataset UUID (8-4-4-4-12 hex). Omit the field to search the GBIF backbone — an empty string is rejected rather than read as no scope, because GBIF answers a blank datasetKey with the unfiltered backbone result.",
"type": "string"
},
"family": {
"description": "Scope the search to a family, by name — \"Paridae\", \"Fagaceae\". Resolved to its backbone key before the search runs, so an alternative family name lands on the taxon it is a synonym of (\"Compositae\" scopes to Asteraceae). Matched exactly and capitalized as GBIF writes it; a name that is not a backbone family fails rather than being ignored. Supplied with genus, it must be that genus's own family. Omit the field to browse every family; a blank or whitespace-only value is rejected rather than dropped.",
"type": "string"
},
"genus": {
"description": "Scope the search to a genus, by name — \"Quercus\", \"Parus\". Resolved to its backbone key before the search runs, and it is the narrowest of the three, so it is what scopes when kingdom or family is supplied too. Matched exactly and capitalized as GBIF writes it; a name shared across kingdoms (\"Prunella\", \"Oenanthe\") resolves only when kingdom is supplied with it. Omit the field to browse every genus; a blank or whitespace-only value is rejected rather than dropped.",
"type": "string"
},
"isExtinct": {
"description": "Filter to extinct (true) or extant (false) taxa.",
"type": "boolean"
},
"kingdom": {
"description": "Scope the search to a kingdom, by name — \"Animalia\", \"Plantae\", \"Fungi\". Resolved to its backbone key before the search runs, and matched exactly: capitalize it as GBIF writes it, since \"animalia\" resolves to nothing. Supplied alongside family or genus it disambiguates that name rather than scoping on its own — \"Prunella\" alone names both a bird genus and a plant genus and resolves to neither. Omit the field to browse every kingdom; a blank or whitespace-only value is rejected rather than dropped.",
"type": "string"
},
"limit": {
"default": 20,
"description": "Number of records to return (default 20, max 1000).",
"maximum": 1000,
"minimum": 1,
"type": "number"
},
"offset": {
"default": 0,
"description": "Pagination offset.",
"minimum": 0,
"type": "number"
},
"q": {
"description": "Name fragment to search for. Matches scientific and vernacular names. Omit the field to browse without a name term — a blank or whitespace-only value is rejected rather than sent, because GBIF answers a blank one with the whole 46,623,754-name index and a whitespace-only one with nothing, and neither is the search a caller who filled the field was asking for.",
"type": "string"
},
"rank": {
"description": "Filter to a specific taxonomic rank.",
"enum": [
"KINGDOM",
"PHYLUM",
"CLASS",
"ORDER",
"FAMILY",
"GENUS",
"SPECIES",
"SUBSPECIES"
],
"type": "string"
}
},
"type": "object"
},
"name": "gbif_search_species",
"outputSchema": {
"$schema": "https://json-schema.org/draft/2020-12/schema",
"additionalProperties": false,
"anyOf": [
{
"not": {
"required": [
"error"
]
},
"required": [
"taxa",
"totalCount",
"offset",
"limit",
"endOfRecords"
]
},
{
"required": [
"error"
]
}
],
"properties": {
"endOfRecords": {
"description": "True when there are no more results after this page.",
"type": "boolean"
},
"error": {
"additionalProperties": {},
"description": "Present when the call failed. Absent on success.",
"properties": {
"code": {
"description": "JSON-RPC error code for this failure.",
"maximum": 9007199254740991,
"minimum": -9007199254740991,
"type": "integer"
},
"data": {
"additionalProperties": {},
"properties": {
"reason": {
"description": "Machine-readable failure mode. Declared by this tool: `invalid_filter`: q, kingdom, family, genus, or datasetKey was supplied blank or whitespace-only, datasetKey is not an 8-4-4-4-12 hex UUID, or GBIF rejected another filter value as malformed. `unresolved_taxon_scope`: kingdom, family, or genus named no GBIF backbone taxon at that rank — a misspelling, a lowercase name, a name entered under the wrong rank, or a name shared across kingdoms with no kingdom supplied to separate them. `conflicting_taxon_scope`: the supplied family and genus each resolved, but to taxa in different lineages — the genus does not sit in that family. Other values are possible when a failure originates below the handler.",
"examples": [
"invalid_filter",
"unresolved_taxon_scope",
"conflicting_taxon_scope"
],
"type": "string"
},
"recovery": {
"additionalProperties": {},
"description": "Actionable next step for the caller.",
"properties": {
"hint": {
"type": "string"
}
},
"required": [
"hint"
],
"type": "object"
},
"retryable": {
"description": "Whether retrying may succeed.",
"type": "boolean"
}
},
"type": "object"
},
"message": {
"description": "Human-readable description of what went wrong.",
"type": "string"
}
},
"required": [
"code",
"message"
],
"type": "object"
},
"limit": {
"description": "Records returned in this page.",
"type": "number"
},
"notice": {
"description": "Guidance when results are empty or paging overshot. Absent on successful result pages.",
"type": "string"
},
"offset": {
"description": "Current pagination offset.",
"type": "number"
},
"taxa": {
"description": "Matching taxa.",
"items": {
"additionalProperties": false,
"description": "A backbone taxon with classification, status, and occurrence counts.",
"properties": {
"canonicalName": {
"description": "Scientific name without authorship.",
"type": "string"
},
"class": {
"description": "Class classification.",
"type": "string"
},
"extinct": {
"description": "True when explicitly flagged as extinct.",
"type": "boolean"
},
"family": {
"description": "Family classification.",
"type": "string"
},
"genus": {
"description": "Genus classification.",
"type": "string"
},
"key": {
"description": "GBIF backbone taxon key.",
"type": "number"
},
"kingdom": {
"description": "Kingdom classification.",
"type": "string"
},
"numDescendants": {
"description": "Count of child taxa in the backbone.",
"type": "number"
},
"numOccurrences": {
"description": "Occurrence record count in GBIF.",
"type": "number"
},
"order": {
"description": "Order classification.",
"type": "string"
},
"phylum": {
"description": "Phylum classification.",
"type": "string"
},
"rank": {
"description": "Taxonomic rank.",
"type": "string"
},
"scientificName": {
"description": "Full scientific name with authorship.",
"type": "string"
},
"taxonomicStatus": {
"description": "ACCEPTED, SYNONYM, DOUBTFUL, etc.",
"type": "string"
},
"vernacularName": {
"description": "Common name when available.",
"type": "string"
}
},
"type": "object"
},
"type": "array"
},
"taxonScope": {
"description": "The higher-taxon scope actually applied — which of kingdom, family, or genus scoped the search, the backbone taxon its name resolved to, and that taxon key. Absent when none of the three was supplied.",
"type": "string"
},
"totalCount": {
"description": "Total matches before pagination.",
"type": "number"
}
},
"type": "object"
}
}
]
}Verify it yourself
curl -s https://api.teppi.xyz/v1/evidence/sha256:088191def31d39344769be87ea036f6e9280da4f790cfcae1d52fbcf84a82ef6 | sha256sum