Biodiversity informatics: the challenge of linking data and the role of shared identifiers

15 years 6 months ago

Download precedings.nature.com

A major challenge facing biodiversity informatics is integrating data stored in widely distributed databases. Initial efforts have relied on taxonomic names as the shared identifier linking records in different databases. However, taxonomic names have limitations as identifiers, being neither stable nor globally unique, and the pace of molecular taxonomic and phylogenetic research means that a lot of information in public sequence databases is not linked to formal taxonomic names. This review explores the use of other identifiers, such as specimen codes and GenBank accession numbers, to link otherwise disconnected facts in different databases. The structure of these links can also be exploited using the PageRank algorithm to rank the results of searches on biodiversity databases. The key to rich integration is a commitment to deploy and reuse globally unique, shared identifiers (such as DOIs and LSIDs), and the implementation services that link those identifiers.

Roderic D. M. Page

Real-time Traffic

BIB 2008 | Databases | Identifiers | Taxonomic Names |

claim paper

Post Info
More Details (n/a)

Added	08 Dec 2010
Updated	08 Dec 2010
Type	Journal
Year	2008
Where	BIB
Authors	Roderic D. M. Page

Comments (0)

Sciweavers

Biodiversity informatics: the challenge of linking data and the role of shared identifiers

BIB 2008 | Databases | Identifiers | Taxonomic Names |

Explore & Download

Productivity Tools

Sciweavers