Search bioRxiv⌕ Search

bioRxiv · 10.1101/2024.03.15.585133

Geographical distribution, disease association and diversity of Klebsiella pneumoniae KL and O antigens in India: roadmap for vaccine development

Abstract

Klebsiella pneumoniae poses a significant healthcare challenge due to its multidrug resistance and diverse serotype landscape. This study aimed to explore the serotype diversity of 1072 K. pneumoniae and its association with geographical distribution, disease severity and antimicrobial/virulence patterns in India. Whole-genome sequencing was performed on the Illumina platform, and genomic analysis was carried out using the Kleborate tool. KL64 (n=264/1072, 26%), KL51 (249/1072, 24%), KL2 (n=88/1072, 8%), O1/O2v1 (n=471/1072, 44%), O1/O2v2 (n=353/1072, 33%), and OL101 (n=66/1072, 6%) were the most prevalent serotypes. The study identified 119 different sequence types (STs) with varying serotypes, with KL64 being the most predominant in ST231 (26%). O serotypes were strongly linked with STs, with O1/O2v1 predominantly associated with ST231 (44%). Simpsons diversity index and Fishers exact test revealed higher serotype diversity in the north and east regions, along with intriguing associations between specific serotypes and resistance profiles. No significant association between KL or O types and disease severity was observed. Furthermore, we found no specific association of virulence factors with KL types or O antigen types (P>0.05). Conventionally described hypervirulent clones (i.e., KL1 and KL2) in India lacked typical virulent markers (i.e., aerobactin), contrasting with other regional serotypes. The cumulative distribution of KL and O serotypes suggests that future vaccines may have to include either [~]20 KL types or 4 O types to cover >85% of the carbapenemase-producing Indian K. pneumoniae population. The findings underscore the need for a vaccine with broad coverage to address the diverse landscape of K. pneumoniae strains in different regions of India. Understanding regional serotype dynamics is pivotal for targeted surveillance, interventions, and tailored vaccine strategies to tackle the diverse landscape of K. pneumoniae infections across India. Data SummaryO_LIAll the sequenced data has been submitted to the European Nucleotide Archive (ENA) under the Bioproject numbers PRJEB29740 and PRJEB50614. Run Accessions and Biosample numbers are provided in Supplementary Table 1 with corresponding metadata for each sample used in the study. C_LIO_LIThe Microreact link for the genomic analysis is provided (https://microreact.org/project/oqKM84GBszEPW9Emt2FKnP-klebsiella-pneumoniae-indian-serotypes). C_LIO_LIThe pipelines used in the study are published in gitlab (https://gitlab.com/cgps/ghru/pipelines). C_LIO_LIThe tools details and the implementation of the pipelines are described in protocols.io (https://www.protocols.io/view/ghru-genomic-surveillance-of-antimicrobial-resista-bp2l6b11kgqe/v4). C_LIO_LIThe R scripts used with all the input files used for each script have been published in Fishare (https://doi.org/10.6084/m9.figshare.25414807.v1) C_LI Impact StatementKlebsiella pneumoniae produces polysaccharide capsules, which serve as both epidemiological markers and significant virulence factors. The increasing accessibility of whole genome sequencing has made it easier than ever to investigate this capsule diversity. This study is the first of its kind in India to comprehensively investigate the serotype diversity of K. pneumoniae strains and their association with disease severity, antimicrobial resistance/virulence patterns, and geographical distribution across various regions of the subcontinent. This multi-dimensional analysis not only provides valuable insights into the molecular epidemiology of K. pneumoniae in India but also offers crucial data for the development of targeted interventions, including vaccine formulations tailored to address the prevailing serotypes. These findings serve as a foundation for informed decision-making in the management and prevention of K. pneumoniae infections, ultimately contributing to improved public health outcomes in the region.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Shamanna, V., Srinivas, S., Couto, N., Nagaraj, G., Sajankila, S. P., Krishnappa, H. G., Kumar, K. A., Aanensen, D., Lingegowda, R. K., GHRU India Consortium,. 2024-03-19. Geographical distribution, disease association and diversity of Klebsiella pneumoniae KL and O antigens in India: roadmap for vaccine development. https://doi.org/10.1101/2024.03.15.585133

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

Reconstruction of the Prox gene family evolution in vertebrates reveals multiple lineage-specific gene losses

Prospero-related homeobox (Prox) genes encode a family of transcription factors that play essential roles in the development of several organs and systems, including the central nervous system, lymphatic endothelium, musculature, and liver. Despite their developmental importance, the evolutionary history of the vertebrate Prox gene family remains poorly understood. In this study we combined phylogenetic and synteny analysis to characterise the evolution of the Prox family in vertebrates. Our results reveal that two to three Prox subfamilies were already present in the last common ancestor of jawed vertebrates. We clarify the identity and evolutionary relationships of well-studied members of this family and identify multiple independent losses of Prox2 and Prox3 genes in specific vertebrate lineages. Furthermore, we uncover evidence for the existence of a fourth Prox gene in the ancestral vertebrate genome, which was subsequently lost. Overall, this study provides the first comprehensive analysis of the evolutionary history of the vertebrate Prox gene family and establishes a foundations for future studies on the functional roles of these genes.

genomics↗

Gene flux shapes diversity and evolution of the ancient 17q21.31 inversion polymorphism

A hallmark of chromosomal inversions is that they suppress recombination between haplotypes, allowing inversion haplotypes to persist as single co-inherited units. To determine the extent to which inversions nevertheless permit genetic exchange, we investigated a common 979-kb inversion polymorphism at the human 17q21.31 locus. This locus exhibits deep divergence between the reference (H1) and inverted (H2) haplotypes, extensive segmental duplications (SDs) flanking the inversion, and association with neurodegenerative diseases, developmental disorders, and fertility-related phenotypes. Using single-cell sperm genome sequencing data, we directly measured recombination rates between H1 and H2 haplotypes and found near-complete suppression of single-crossover events between the haplotypes. The rare single crossovers that did occur were mediated by non-allelic homologous recombination between shared H1 and H2 SDs, generating novel duplication architectures. In contrast, two-switch events consistent with gene conversion or double crossovers, spanning 17-150 kb, occurred throughout the inversion at rates exceeding genome-wide estimates for events of comparable size. Consistent with recurring genetic exchange, we identified 99 distinct H1-H2 recombinant haplotypes segregating in All of Us genomes, including 26 with combinations of H1 and H2 SDs. These recombinant haplotypes facilitated dissection of the inversion's effects on fertility-related phenotypes; using a large parent-embryo dataset, we found that H2 additively increases female crossover rates across chromosomes and that KANSL1 duplications do not explain this effect. Finally, ancestral recombination graphs dated H1-H2 gene flux (the exchange of genetic material between alternative arrangements) to approximately 100-500 thousand years ago, revealing that H1 and H2 haplotypes have co-segregated for at least half a million years. Together, these results demonstrate that inversions can be permeable barriers to recombination, with ongoing gene flux influencing the diversity and evolution of inversion polymorphisms.

genomics↗

Genomic correlates of metastatic competence and progression in human melanoma

Genomic events and their timing that grant a primary tumour the competence to disseminate remain poorly defined. We performed sequencing of 247 stage I/II primary cutaneous melanomas (CMs) and 60 matched metastases without intervening therapy from a prospectively followed registry cohort with a median followup of 92 months, integrating copy-number, mutational, protein and spatial-transcriptomic analyses. Relapse was not distinguished by oncogenic point mutations, which were largely shared between primaries and metastases, but by somatic copy-number alterations (SCNAs) and global chromosomal instability. We defined OncoCycle, a six-gene copy-number signature (amplification of CDK4, MCL1 and CD276; biallelic loss of CDKN2A, CDKN2B and TP53BP1) that predicted relapse independently of established clinicopathological features in melanoma, and a pan-cancer analysis. In matched pairs, metastatic progression was driven by continued copy-number evolution and reduction in intra-tumoural heterogeneity, rather than by acquired point mutations, and OncoCycle alterations from primary tumours were preserved in metastasis seeding clones. Clonal reconstruction revealed both monoclonal and polyclonal metastasis seeding, and spatial transcriptomics resolved copy-number-defined metastatic subclones occupying and programming distinct immune and stromal niches. Thus, metastatic competence was primed early by focal SCNAs on a background of chromosomal instability, elaborated by continued copy-number evolution during dissemination and spatio-temporal interactions with the tumour-microenvironment.

genomics↗