Search bioRxiv⌕ Search

bioRxiv · 10.64898/2026.06.25.734422

BIRD-Seq: B2 Protein Integrated End-to-End Pipeline for dsRNA Detection and Nanopore Sequencing for Virus Monitoring

Abstract

Double-stranded RNA (dsRNA) is a near-universal hallmark of active viral infection. Despite its role as a pan-viral replication intermediate, dsRNA-centred technologies for virus monitoring remain scarce and largely rely on monoclonal antibody-based approaches that, while highly sensitive, are costly and difficult to engineer, scale, or integrate with downstream assays. Here, we present a modular and antibody-free pipeline for quantitative and qualitative dsRNA analysis built around an engineered B2 protein from Flock House virus with nanomolar-range binding affinity. The pipeline is compatible with absorbance- or luminescence-based measurement formats. In a sandwich assay configuration (Sand-BIRD), sub-ng mL-1 quantification of dsRNA is achieved, comparable to the gold standard J2 monoclonal antibody, directly from crude biological samples without RNA extraction. Sand-BIRD reliably detects viral infection in both plant samples (Tomato bushy stunt virus and Grapevine fanleaf virus) and mosquitoes (West Nile virus and Dengue virus) with commercial-grade reliability. dsRNA eluted from positive samples were further processed directly by Oxford Nanopore direct sequencing, enabling identification of virus species without prior sequence knowledge or total RNA extraction. Together, this work establishes an end-to-end, sequence-agnostic workflow for direct RNA-duplex quantification and sequencing (BIRD-Seq), which has compelling potential for emerging infectious disease surveillance and next-generation point-of-care (PoC) diagnostics. Technology ReadinessBIRD-Seq is an integrated pipeline for agnostic dsRNA detection and sequencing designed for broad-spectrum virus monitoring. A Technology Readiness Level (TRL) 5 under NASAs classification framework has been reached as BIRD-Seq has been validated in laboratory-relevant environments using real-world samples, including virus-infected plants and mosquitoes. The ELISA-based sensing platform employs engineered variants of the B2 protein (from Flock House virus) in a sandwich assay format for dsRNA capture and detection, achieving sensitivity comparable to the gold-standard J2 monoclonal antibody. Unlike traditional antibody-based methods, the B2 protein offers key practical advantages: straightforward production in bacterial expression systems, high versatility, and reduced manufacturing costs, as well as direct compatibility with crude extract monitoring, eliminating the need for RNA extraction. Captured B2/RNA duplexes can then be directly eluted from the ELISA microplates and subjected to downstream nanopore direct RNA sequencing, providing both quantitative and qualitative information on the underlying virus infection, a capacity enabled by the near-universal nature of dsRNA as a pathogen-associated molecular pattern. That said, further validation on large-scale field-collected and clinical samples will be essential before widespread deployment can be envisioned. While the B2 sandwich assay offers favorable cost-efficiency over antibody-based alternatives, the relatively high cost of Oxford Nanopore direct RNA sequencing remains an important economic constraint. Nevertheless, the growing importance of dsRNA detection across virus sensing, mRNA vaccine development, innate immunity research, and human disease diagnostics, combined with the increasing role of portable long-read sequencing in emerging infectious disease (EIDs) surveillance, positions BIRD-Seq as an innovative and competitive diagnostic platform. Highlights- Double-stranded RNA (dsRNA) is one of the critical pathogen-associated molecular patterns for viral invasion in the host. A protein-based sandwich assay for dsRNA detection in crude biological samples with a sub-ng mL-1 order detection limit was developed, achieving similar sensing efficiency in comparison to expensive and proprietary monoclonal antibody-based ELISA methods. - The quantitative detection of RNA duplex is coupled with an Oxford nanopore direct dsRNA sequencing method for virus species identification and qualitative analysis. - This is one of the very first dsRNA-centered end-to-end workflows for virus monitoring and sequencing, validated for both infected plant and animal samples.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Xhurxhi, A. N., Sahu, S., Scheer, H., Miott, E. F., Alioua, A., Clesse, D., Monsion, B., Pompon, J., Szunerits, S., Blevins, T., Ritzenthaler, C.. 2026-06-26. BIRD-Seq: B2 Protein Integrated End-to-End Pipeline for dsRNA Detection and Nanopore Sequencing for Virus Monitoring. https://doi.org/10.64898/2026.06.25.734422

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

RNASeek: A Cross-Phyla Generative Foundation Model for Multipurpose RNA Modeling and Reinforcement Learning-Based Design

RNA plays central roles in regulating information flow and provides a versatile substrate for engineering biological functions. While large language models (LLMs) have transformed natural language processing and protein design, a general framework connecting RNA foundation models to functional sequence design remains limited. Here, we present RNASeek, a 1.6-billion-parameter generative foundation model built on a DeepSeek architecture and trained on a cross-phyla transcriptomic corpus for RNA sequence representation and generation. Natural-language tokens enable flexible conditional prediction and sequence design using a unified backbone. RNASeek captures species-specific transcript features and intron-exon boundaries in a zero-shot setting. We then fine-tune RNASeek to predict ribozyme self-cleavage activity and viral mRNA stability, revealing interpretable sequence features associated with function, including ribozyme loop flexibility and stem stability, as well as AU-rich motifs associated with mRNA stability. We use these functional predictors as reward models and apply Group Relative Policy Optimization (GRPO) to update the generation policy of RNASeek toward sequences with desired properties. GRPO-guided generation produces faster-cleaving ribozymes and stability-enhancing 3' UTRs while satisfying user-specified IUPAC constraints. Experimentally validated RNASeek-generated ribozymes achieve wild-type levels of activity, while RNASeek-generated 3' UTR sequences exceed the performance of the training data and benchmarked AI-generated 3' UTRs. Together, RNASeek establishes a unified pretrain-predict-optimize framework that connects learned RNA function to controllable de novo sequence design and provides a general strategy for engineering regulatory RNAs with desired properties.

bioengineering↗

Joint Vector Flow Mapping and Segmentation: Ill-Posedness,Differentiable Bayesian Inference, and Synthetic Vortex-FlowBenchmarks

Vector flow mapping (VFM) reconstructs left-ventricular (LV) blood velocity from color-Doppler echocardiography by combining the measured beamwise component with physical and regularizing constraints. Analysis of the discrete VFM formulation shows that the inverse problem is intrinsically ill posed: the occurrence of singular modes can be predicted from the geometry of the segmented blood-pool domain, the imposed boundary conditions, and the degree of smoothing. These modes can propagate uncertainty along entire transverse bands of the reconstructed velocity field, yet conventional VFM neither quantifies this uncertainty nor allows for correcting the blood-pool segmentation. We introduce Bayesian VFM (B--VFM), a hierarchical framework that jointly infers radial and transverse velocities, a probabilistic blood-pool mask, their spatially resolved uncertainties, and hyperparameters weighting Doppler and segmentation fidelity, mass conservation, boundary conditions, and smoothness. The discretized posterior admits a closed-form gradient and exact Hessian, enabling computationally efficient, gradient-based MAP estimation, sampling, and direct analysis of ill-posed modes. Posterior inference combines Gibbs sampling of conjugate Gamma-distributed hyperparameters with conditional maximum-a-posteriori estimation and a Laplace approximation for the high-dimensional velocity and mask fields. To accommodate systematic departures from planar mass conservation, B-VFM can learn the covariance of the planar divergence residual from an ensemble of flows and incorporate it as a structured model-discrepancy prior. Independent chains converged reproducibly, while covariance priors learned from flow ensembles illustrated how model discrepancies can be incorporated into the inference. B--VFM was evaluated using Lamb-Chaplygin dipoles under ideal conditions and with Doppler corruption, Doppler voids, and segmentation defects, and using the Hicks-Moffatt family of spherical vortices to assess violations of planar mass conservation. The method produced smooth reconstructions, localized uncertainty near unreliable measurements and regions of model inconsistency, and used flow information to correct segmentation errors. Within the tested vortex family, the data-informed planar divergence prior reduced velocity bias and mask distortion. B--VFM thus provides an uncertainty-aware reconstruction method and a flexible foundation for future VFM formulations incorporating additional priors, observations, and physical models. Future work will evaluate the method using clinical data and more complex three-dimensional benchmark flows.

bioengineering↗

Computational design of a versatile, zero-radius proximity labeling enzyme

The ability to map protein interactomes and organelle proteomes is foundational for achieving a molecular understanding of living cells. Proximity labeling (PL) provides a powerful strategy for this, but existing enzymes and photocatalysts are limited by their spatial resolution, reliance on biotin, and/or in vivo compatibility. Here we report FlexID, an engineered promiscuous ligase that catalyzes the rapid attachment of diverse small-molecule probes to proximal endogenous proteins. Critically, FlexID operates through a zero-radius, direct-contact mechanism, offering superior spatial precision compared to existing PL tools. We engineered FlexID by combining the strengths of sequence- and structure-trained computational models to enhance its catalytic activity and structural stability. Biophysical analysis revealed that specific conformational changes in FlexID improve its ability to recognize diverse target proteins while simultaneously preventing the premature release of the reactive intermediate. We demonstrate FlexID's versatility through in vivo proximity labeling, comprehensive organelle proteome mapping, and a high-throughput, fluorescence-based screen for molecular glues. Our work shows that computational methods can be harnessed to create mechanistically distinct PL enzymes and establishes FlexID as a flexible, high-resolution tool for mapping protein interactions and proteomes in living cells.

bioengineering↗