Search bioRxiv⌕ Search

bioRxiv · 10.64898/2026.06.21.733609

qPCR Guru, a free browser-based platform, strengthens microRNA analyses using full-curve Cq estimation

Abstract

Quantitative PCR (qPCR) depends on reliable quantification cycle (Cq) estimation from amplification curves, which are not always well-behaved. We developed qPCR Guru to provide a complete analysis pipeline including data quality assessment, relative quantification, standard-curve diagnostics, and dual-method Cq evaluation. The latter compares the conventional instrument-derived threshold ("Reported") Cq versus the full-curve five-parameter logistic (5PL) second-derivative-maximum ("Fit") Cq and automatically flags curve-shape abnormalities and disagreement between the two estimates. On high-expressing targets (mRNA and microRNA), the two methods showed strong convergence, confirming general-purpose performance. On low-expressing targets, such as serum microRNA, baseline artifacts and biphasic amplification result in threshold miscalls that standard instrument analysis does not flag. Fit Cq restored replicate-concordant values where Reported Cq split the technical replicates by 17-20 cycles, recovered MIQE-compliant amplification efficiencies lost to biphasic miscalls (from 74% to 102% and 387% to 98%), and lowered within-group variability by 48% and 68% in feline and bovine samples, respectively. Together, these results demonstrate that full-curve estimation, with integrated curve-level diagnostics, strengthens qPCR analyses against threshold miscalls. ARTICLE HIGHLIGHTSO_LIqPCR Guru is a free, browser-based platform that provides a complete analysis pipeline and facilitates side-by-side comparisons of an instruments threshold (Reported) Cq and a full-curve (Fit) Cq, from the five-parameter logistic fitting with second-derivative-maximum (SDM/cpD2). C_LIO_LIFor every well the application automatically flags curve-shape abnormalities and disagreement between the two Cq estimates. C_LIO_LIOn clean, high-expressing mRNA and microRNA targets, the two estimators (Reported Cq and Fit Cq) were strongly concordant and produced equivalent relative quantification with comparable precision. C_LIO_LIIn low-expressing serum microRNA, baseline artifacts and biphasic amplification produced threshold Cq miscalls of up to [~]20 cycles and were detected by curve-shape flags and/or method disagreement. C_LIO_LIThe full-curve Cq estimate recovered replicate-concordant values, restored MIQE-compliant amplification efficiencies, and reduced within-group variability in serum microRNA. C_LI

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Singh, A., Singh, O., Sarkar, M., Coultous, R., Stice, S.. 2026-06-23. qPCR Guru, a free browser-based platform, strengthens microRNA analyses using full-curve Cq estimation. https://doi.org/10.64898/2026.06.21.733609

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

Functional primary human 3D skeletal muscle organoids enable exercise and metabolic research

Human skeletal muscle is the principal site of insulin-stimulated glucose disposal and a major mediator of exercise-induced metabolic benefits, yet human models that preserve metabolic and exercise responsiveness remain limited. We generated primary human skeletal muscle organoids from donor-derived CD56+ myoblasts using a collagen-based extracellular matrix and serum-free IGF1-guided differentiation. The organoids formed aligned contractile tissues containing oxidative and glycolytic fiber type-like myotubes, displayed enhanced mitochondrial respiration, insulin-stimulated glucose uptake, and reproducible force generation. Electrical pulse stimulation induced AMPK activation, increased glucose utilization and lactate production, and upregulated canonical exercise-responsive genes including NR4A3 and PPARGC1A. Notably, transcriptional responses to in vitro exercise overlapped with acute exercise responses observed in skeletal muscle biopsies from the same donors. The organoids further detected functional impairments of skeletal muscle performance induced by TGF-{beta}1 and metformin and increased speed generation by testosterone treatment. These findings establish a donor-specific human skeletal muscle platform that recapitulates key features of insulin action and exercise adaptation and may enable mechanistic studies of skeletal muscle metabolism, exercise responsiveness, and therapeutic interventions relevant to diabetes.

Molecular Biology↗

TRIDENT (Taxonomic Resolution and IDentification using Environmental dNa Traces): An Optimized Algorithm for Vertebrate Taxonomic Assignments in eDNA Metabarcoding, Integrating Molecular, Taxonomic, and Ecological Criteria

Environmental DNA (eDNA) metabarcoding has become a powerful approach for large-scale biodiversity assessment, yet taxonomic assignment remains one of its most critical error-prone steps. Current bioinformatic pipelines rely on molecular similarity searches against reference databases, but assignment accuracy is constrained not only by short marker length and database incompleteness, but also by fundamental limitations, including recent species radiations, incomplete lineage sorting, introgression, NUMTs, and the imperfect correspondence between genetic variation and species boundaries. Here, we present TRIDENT (Taxonomic Resolution and IDentification using Environmental dNa Traces), an automated and simple protocol designed to improve taxonomic assignments in eDNA metabarcoding. Initially developed for marine vertebrates, TRIDENT may be used with any barcode and integrates three complementary sources of evidence: molecular similarity (NCBI/GenBank and BOLD), curated taxonomic information (WoRMS), and ecological plausibility derived from biogeographic occurrence data (GBIF). The workflow sequentially constructs candidate taxon lists based on sequence similarity, expands them through taxonomic hierarchies, and filters them using spatial occurrence constraints. It further identifies possible taxa lacking reference barcodes and evaluates their plausibility through CO1-based similarity if data exist in BOLD. TRIDENT has been implemented as a source-available Python tool and tested using empirical eDNA datasets from marine vertebrates as well as simulated communities. Results demonstrate that the tool produces taxonomic assignments consistent with expert manual curation while substantially reducing processing time and attention errors caused by manual processing of large datasets. By combining molecular, taxonomic, and ecological criteria within a single framework, TRIDENT improves transparency and reproducibility and provides a robust and flexible solution strengthening confidence in taxonomic identifications in eDNA-based biodiversity assessments.

Molecular Biology↗

Identifying and Addressing Systematic Data Leakage in Protein-Ligand Affinity Benchmarks

Accurate prediction of protein-ligand binding affinity is a crucial goal in structure-based drug discovery, with the potential to significantly shorten development timelines. Recently, a new wave of machine learning models based on co-folding, such as Boltz-2 and IsoDDE, has demonstrated performance that matches or exceeds that of gold-standard physics-based methods like Free Energy Perturbation (FEP). This paper provides a critical assessment of these claims, revealing that current benchmarks are heavily influenced by data leakage, and proposes a new benchmark that explicitly controls for data leakage. We demonstrate that splitting by protein-sequence identity is inherently insufficient to prevent data leakage due to "target mirroring," in which homologous proteins with low overall sequence identity still exhibit highly correlated binding profiles. Our meta-analysis of documents in the ChEMBL 36 database identifies more than 6,000 such assay pairs and finds that leakage persists for sequence-identity thresholds as low as 0.2, well below the values commonly used in benchmarks today. Additionally, we show that a ligand-only baseline model, which lacks protein structural information, achieves surprisingly high performance on the FEP+ 4 and OpenFE benchmarks (r = 0.66 and r = 0.36, respectively). Our results indicate that current benchmarks tend to reward models for memorizing training data and exploiting localized leakage rather than truly learning biophysical principles. To address this issue, we propose the Novelty-Tiered Affinity Benchmark, in which the test data is partitioned into ligand novelty tiers. In the most challenging tier (Tanimoto similarity < 0.35), ligand-only models perform notably worse (r = 0.14), offering a clear baseline for evaluating genuine generalization. We argue that the field must move beyond sequence-based splits to ensure that AI-driven discovery translates into successful prospective laboratory research.

Molecular Biology↗