Search bioRxivSearch

Biology subjects

Vazquez, A.

Publications and source records attributed to Vazquez, A..

3 recordsLinked to original sources

A physical model of cell metabolism

Cell metabolism is characterized by three fundamental energy demands to sustain cell maintenance, to trigger aerobic fermentation and to achieve maximum metabolic rate. Here we report a physical model of cell metabolism that explains the origin of these three energy scales. Our key hypothesis is that the maintenance energy demand is rooted on the energy expended by molecular motors to fluidize the cytoplasm and counteract molecular crowding. Using this model and independent parameter estimates we make predictions for the three energy scales that are in quantitative agreement with experimental values. The model also recapitulates the dependencies of cell growth with extracellular osmolarity and temperature. This theory brings together biophysics and cell biology in a tractable model that can be applied to understand key principles of cell metabolism.

biophysics

Accurate Genomic Prediction Of Human Height

We construct genomic predictors for heritable and extremely complex human quantitative traits (height, heel bone density, and educational attainment) using modern methods in high dimensional statistics (i.e., machine learning). Replication tests show that these predictors capture, respectively, ~40, 20, and 9 percent of total variance for the three traits. For example, predicted heights correlate ~0.65 with actual height; actual heights of most individuals in validation samples are within a few cm of the prediction. The variance captured for height is comparable to the estimated SNP heritability from GCTA (GREML) analysis, and seems to be close to its asymptotic value (i.e., as sample size goes to infinity), suggesting that we have captured most of the heritability for the SNPs used. Thus, our results resolve the common SNP portion of the \"missing heritability\" problem - i.e., the gap between prediction R-squared and SNP heritability. The ~20k activated SNPs in our height predictor reveal the genetic architecture of human height, at least for common SNPs. Our primary dataset is the UK Biobank cohort, comprised of almost 500k individual genotypes with multiple phenotypes. We also use other datasets and SNPs found in earlier GWAS for out-of-sample validation of our results.

genomics

Bacterial genome reduction as a result of short read sequence assembly

High-throughput comparative genomics has changed our view of bacterial evolution and relatedness. Many genomic comparisons, especially those regarding the accessory genome that is variably conserved across strains in a species, are performed using assembled genomes. For completed genomes, an assumption is made that the entire genome was incorporated into the genome assembly, while for draft assemblies, often constructed from short sequence reads, an assumption is made that genome assembly is an approximation of the entire genome. To understand the potential effects of short read assemblies on the estimation of the complete genome, we downloaded all completed bacterial genomes from GenBank, simulated short reads, assembled the simulated short reads and compared the resulting assembly to the completed assembly. Although most simulated assemblies demonstrated little reduction, others were reduced by as much as 25%, which was correlated with the repeat structure of the genome. A comparative analysis of lost coding region sequences demonstrated that up to 48 CDSs or up to ~112,000 bases of coding region sequence, were missing from some draft assemblies compared to their finished counterparts. Although this effect was observed to some extent in 32% of genomes, only minimal effects were observed on pan-genome statistics when using simulated draft genome assemblies. The benefits and limitations of using draft genome assemblies should be fully realized before interpreting data from assembly-based comparative analyses.

genomics