LinkHub: a Semantic Web system that facilitates cross-database queries and information retrieval in proteomics

Cheung, Kei-Hoi; Gerstein, Mark B; Schultz, Martin; Smith, Andrew K; Yip, Kevin Y

LinkHub: a Semantic Web system that facilitates cross-database queries and information retrieval in proteomics

Authors: Kei-Hoi Cheung
Mark B Gerstein
Martin Schultz
Andrew K Smith
Kevin Y Yip
Publication date: 1 January 2007
Publisher: BioMed Central
Doi

Abstract

Abstract Background A key abstraction in representing proteomics knowledge is the notion of unique identifiers for individual entities (e.g. proteins) and the massive graph of relationships among them. These relationships are sometimes simple (e.g. synonyms) but are often more complex (e.g. one-to-many relationships in protein family membership). Results We have built a software system called LinkHub using Semantic Web RDF that manages the graph of identifier relationships and allows exploration with a variety of interfaces. For efficiency, we also provide relational-database access and translation between the relational and RDF versions. LinkHub is practically useful in creating small, local hubs on common topics and then connecting these to major portals in a federated architecture; we have used LinkHub to establish such a relationship between UniProt and the North East Structural Genomics Consortium. LinkHub also facilitates queries and access to information and documents related to identifiers spread across multiple databases, acting as "connecting glue" between different identifier spaces. We demonstrate this with example queries discovering "interologs" of yeast protein interactions in the worm and exploring the relationship between gene essentiality and pseudogene content. We also show how "protein family based" retrieval of documents can be achieved. LinkHub is available at hub.gersteinlab.org and hub.nesg.org with supplement, database models and full-source code. Conclusion LinkHub leverages Semantic Web standards-based integrated data to provide novel information retrieval to identifier-related documents through relational graph queries, simplifies and manages connections to major hubs such as UniProt, and provides useful interactive and query interfaces for exploring the integrated data.</p

Similar works

Full text

Open in the Core reader

Download PDF

Available Versions

Crossref

Last time updated on 01/04/2019

Directory of Open Access Journals

oai:doaj.org/article:ca4d083e0...

Last time updated on 17/12/2014