Search bioRxiv⌕ Search

Biology subjects

Imam, F. T.

Publications and source records attributed to Imam, F. T..

2 recordsLinked to original sources

The Data Distillery: A Graph Framework for Semantic Integration and Querying of Biomedical Data

The Data Distillery Knowledge Graph (DDKG) is a framework for semantic integration and querying of biomedical data across domains. Built for the NIH Common Fund Data Ecosystem, it supports translational research by linking clinical and experimental datasets in a unified graph model. Clinical standards such as ICD-10, SNOMED, and DrugBank are integrated through UMLS, while genomics and basic science data are structured using ontologies and standards such as HPO, GENCODE, Ensembl, STRING, and ClinVar. The DDKG uses a property graph architecture based on the UBKG infrastructure and supports ontology-based ingestion, identifier normalization, and graph-native querying. The system is modular and can be extended with new datasets or schema modules. We demonstrate its utility for informatics queries across eight use cases, including regulatory variant analysis, tissue-specific expression, biomarker discovery, and cross-species variant prioritization. The DDKG is accessible via a public interface, a programmatic API, and downloadable builds for local use.

bioinformatics↗

Developing a Multiscale Neural Connectivity Knowledgebase of the Autonomic Nervous System

The Stimulating Peripheral Activity to Relieve Conditions (SPARC) program is a U.S. National Institutes of Health (NIH) funded effort to enhance our understanding of the neural circuitry responsible for visceral control. SPARCs mission is to identify, extract, and compile our overall existing knowledge and understanding of the autonomic nervous system (ANS) connectivity between the central nervous system and end organs. A major goal of SPARC is to use this knowledge to promote the development of the next generation of neuromodulation devices and bioelectronic medicine for nervous system diseases. As part of the SPARC program, we have been developing SCKAN, a dynamic knowledge base of ANS connectivity that contains information about the origins, terminations, and routing of ANS projections. The distillation of SPARCs connectivity knowledge into this knowledge base involves a rigorous curation process to capture connectivity information provided by experts, published literature, textbooks, and SPARC scientific data. SCKAN is used to automatically generate anatomical and functional connectivity maps on the SPARC portal. In this article, we present the design and functionality of SCKAN, including the detailed knowledge engineering process developed to populate the resource with high quality and accurate data. We discuss the process from both the perspective of SCKANs ontological representation as well as its practical applications in developing information systems. We share our techniques, strategies, tools and insights for developing a practical knowledgebase of ANS connectivity that supports continual enhancement.

bioinformatics↗