Search bioRxivSearch

Biology subjects

Bensmail, H.

Publications and source records attributed to Bensmail, H..

5 recordsLinked to original sources

Integrative Statistical Inferences for Drug Sensitivity Biomarkers in Cancer

Personal medicine has been associated with different patient responses to different anti-cancer therapies. Recently, scientists are looking not only for new biomarkers associated with a disease such as cancer but also identifying biomarkers that predict patients who are most likely to respond to a particular cancer treatment. Orderly endeavors to relate cancer mutational information with biological conditions may encourage the interpretation of somatic mutation indexes into significant biomarkers for patient stratification.\n\nWe have screened and incorporated a board of cancer cell lines from Genomics of Drug Sensitivity in Cancer (GDSC) database to recognize genomic highlights related with drug sensitivity. We used mutation, DNA copy number variation, and gene expression information from Catalogue of Somatic Mutations in Cancer (COSMIC) and The Cancer Genome ATLAS (TCGA) for cell lines with their reactions to associate focused and cytotoxic treatments with approved drugs and drugs under clinical and preclinical examination.\n\nWe discovered mutated cancer genes were related with cell reaction to, mostly accessible, cancer medications and some mutated genes were related with sensitivity to an expansive scope of therapeutic agents. By connecting drug activity to the useful many-sided quality of cancer genomes, efficient pharmacogenomic profiling in tumor cell lines gives an intense biomarker revelation stage to guide balanced malignancy remedial systems.\n\nOur study highlights that gene ANK2 amplification, and gene CELSER1 amplification and deletion are highly associated with anti-leukemic drug candidate LFM-A13. It also highlights that gene NUP214 and ROS1 copy number and gene NSD1 amplification are as a group highly associated with the parkinson drug Nilotinib. Finally, our study confirms that gene BRAF mutation is interacting with the BRAF-selective inhibitors drugs PLX4720 and SB590885. On the other hand, our study provides two open source analysis packages: bastah for the multitask-association analysis, and UNGeneAnno for automatic annotation of the variants.

genomics

Application of High-Dimensional Statistics and Network based Visualization techniques on Arab Diabetes and Obesity data

BackgroundObesity and its co-morbidities are characterized by a chronic low-grade inflammatory state, uncontrolled expression of metabolic measurements and dis-regulation of various forms of stress response. However, the contribution and correlation of inflammation, metabolism and stress responses to the disease are not fully elucidated. In this paper a cross-sectional case study was conducted on clinical data comprising 117 human male and female subjects with and without type 2 diabetes (T2D). Characteristics such as anthropometric, clinical and bio-chemical measurements were collected.\n\nMethodsAssociation of these variables with T2D and BMI were assessed using penalized hierarchical linear and logistic regression. In particular, elastic net, hdi and glinternet were used as regularization models to distinguish between cases and controls. Differential network analysis using closed-form approach was performed to identify pairwise-interaction of variables that influence prediction of the phenotype.\n\nResultsFor the 117 participants, physical variables such as PBF, HDL and TBW had absolute coefficients 0.75, 0.65 and 0.34 using the glinternet approach, biochemical variables such as MIP, ROS and RANTES were identified as determinants of obesity with some interaction between inflammatory markers such as IL4, IL-6, MIP, CSF, Eotaxin and ROS. Diabetes was associated with a significant increase in thiobarbituric acid reactive substances (TBARS) which are considered as an index of endogenous lipid peroxidation and an increase in two inflammatory markers, MIP-1 and RANTES. Furthermore, we obtained 13 pairwise effects. The pairwise effects include pairs from and within physical, clinical and biochemical features, in particular metabolic, inflammatory, and oxidative stress markers.\n\nConclusionsWe showcase that markers of oxidative stress (derived from lipid peroxidation) such as MIP-1 and RANTES participate in the pathogenesis of diseases such as diabetes and obesity in the Arab population.

bioinformatics

Differential Community Detection in Paired Biological Networks

MotivationBiological networks unravel the inherent structure of molecular interactions which can lead to discovery of driver genes and meaningful pathways especially in cancer context. Often due to gene mutations, the gene expression undergoes changes and the corresponding gene regulatory network sustains some amount of localized re-wiring. The ability to identify significant changes in the interaction patterns caused by the progression of the disease can lead to the revelation of novel relevant signatures.\n\nMethodsThe task of identifying differential sub-networks in paired biological networks (A:control,B:case) can be re-phrased as one of finding dense communities in a single noisy differential topological (DT) graph constructed by taking absolute difference between the topological graphs of A and B. In this paper, we propose a fast two-stage approach, namely Differential Community Detection (DCD), to identify differential sub-networks as differential communities in a de-noised version of the DT graph. In the first stage, we iteratively re-order the nodes of the DT graph to determine approximate block diagonals present in the DT adjacency matrix using neighbourhood information of the nodes and Jaccard similarity. In the second stage, the ordered DT adjacency matrix is traversed along the diagonal to remove all the edges associated with a node, if that node has no immediate edges within a window. We then apply community detection methods on this de-noised DT graph to discover differential sub-networks as communities.\n\nResultsOur proposed DCD approach can effectively locate differential sub-networks in several simulated paired random-geometric networks and various paired scale-free graphs with different power-law exponents. The DCD approach easily outperforms community detection methods applied on the original noisy DT graph and recent statistical techniques in simulation studies. We applied DCD method on two real datasets: a) Ovarian cancer dataset to discover differential DNA co-methylation sub-networks in patients and controls; b) Glioma cancer dataset to discover the difference between the regulatory networks of IDH-mutant and IDH-wild-type. We demonstrate the potential benefits of DCD for finding network-inferred bio-markers/pathways associated with a trait of interest.\n\nConclusionThe proposed DCD approach overcomes the limitations of previous statistical techniques and the issues associated with identifying differential sub-networks by use of community detection methods on the noisy DT graph. This is reflected in the superior performance of the DCD method with respect to various metrics like Precision, Accuracy, Kappa and Specificity. The code implementing proposed DCD method is available at https://sites.google.com/site/ raghvendramallmlresearcher/codes.

systems biology

RGBM: Regularized Gradient Boosting Machines For The Identification of Transcriptional Regulators Of Discrete Glioma Subtypes

The transcription factors (TF) which regulate gene expressions are key determinants of cellular phenotypes. Reconstructing large-scale genome-wide networks which capture the influence of TFs on target genes are essential for understanding and accurate modelling of living cells. We propose RGBM: a gene regulatory network (GRN) inference algorithm, which can handle data from heterogeneous information sources including dynamic time-series, gene knockout, gene knockdown, DNA microarrays and RNA-Seq expression profiles. RGBM allows to use an a priori mechanistic of active biding network consisting of TFs and corresponding target genes. RGBM is evaluated on the DREAM challenge datasets where it surpasses the winners of the competitions and other established methods for two evaluation metrics by about 10-15%.\n\nWe use RGBM to identify the main regulators of the molecular subtypes of brain tumors. Our analysis reveals the identity and corresponding biological activities of the master regulators driving transformation of the G-CIMP-high into the G-CIMP-low subtype of glioma and PA-like into LGm6-GBM, thus, providing a clue to the yet undetermined nature of the transcriptional events driving the evolution among these novel glioma subtypes.\n\nRGBM is available for download on CRAN at https://cran.rproject.org/web/packages/RGBM/index.html

bioinformatics

Unified Framework for Representing and Ranking

In the database retrieval and nearest neighbor classification tasks, the two basic problems are to represent the query and database objects, and to learn the ranking scores of the database objects to the query. Many studies have been conducted for the representation learning and the ranking score learning problems, however, they are always learned independently from each other. In this paper, we argue that there are some inner relationships between the representation and ranking of database objects, and try to investigate their relationships by learning them in a unified way. To this end, we proposed the Unified framework for Representation and Ranking (UR2) of objects for the database retrieval and nearest neighbor classification tasks. The learning of representation parameter and the ranking scores are modeled within one single unified objective function. The objective function is optimized alternately with regarding to representation parameter and the ranking scores. Based on the optimization results; iterative algorithms are developed to learn the representation parameter and the ranking scores on a unified way. Moreover, with two different formulas of representation (feature selection and subspace learning), we give two versions of UR2. The proposed algorithms are tested on two challenging tasks - MRI image based brain tumor retrieval and nearest neighbor classification based protein identification. The experiments show the advantage of the proposed unified framework over the state-of-the-art independent representation and ranking methods.

bioinformatics