Search bioRxiv⌕ Search

bioRxiv · 10.64898/2026.09.17.751682

Inferring the latent network of pairwise mutualistic preferences from observed plant-pollinator interactions

Abstract

Plant-pollinator communities are typically represented as bipartite networks, whose edges are taken directly from field records of visits. These visits, however, are only a proxy for the object of ecological interest: the latent mutualistic preference between two species. While counts are shaped by preference, they also carry confounding factors such as species abundances, sampling effort, and site- or time-specific conditions. We introduce a hierarchical Bayesian framework that treats visit counts as a realisation of a Poisson process and, on the log scale, decomposes the corresponding pairwise rate into a baseline (community-wide activity together with sampling effort), individual species effects representing abundance, and pairwise mutualistic preferences. The model extends to data replicated across sites and time points, and to the inclusion of environmental or experimental covariates. Because the whole system is fitted jointly, we obtain posterior not only for the preferences but for every latent quantity, each carrying ecological signal of its own, with uncertainty propagated through every level of the model, down to any derived network metric. On synthetic data, we show that common practices, such as reading preferences off raw counts or aggregating replicated observations into a single network, confound abundance with preference. In contrast, our framework recovers the underlying preference structure. On empirical datasets, including a seasonal multi-site pollination study where urbanisation level enters as a covariate, the inferred preference network departs markedly from the observed visits, revealing structure hidden in the raw counts: how species vary across sites and time, and which parts of the community respond most to the covariate. When communities are compared along the urbanisation gradient, standard network metrics on the preference layer revise the conclusions drawn from visits alone. The framework offers a principled way to move from networks of observed visits to networks of underlying mutualistic preferences, carrying uncertainty from the data through to the ecological conclusions and accommodating the spatial, temporal, and covariate structure of modern plant-pollinator datasets. Because it acts on the foundational step of network construction, its implications are broad, placing network-based approaches on firmer ground.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Federici, L., Matechou, E., Iacopini, I.. 2026-09-18. Inferring the latent network of pairwise mutualistic preferences from observed plant-pollinator interactions. https://doi.org/10.64898/2026.09.17.751682

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

PlanktonLake-CEREEP- A Freshwater Plankton Image Dataset with Semi-Automated Label Cleaning

Plankton plays a fundamental role in aquatic ecosystems, influencing biogeochemical cycles and serving as a key food source for many organisms. Recent high-throughput imaging technologies enable the rapid acquisition of large volumes of microscopic images, creating new opportunities for monitoring planktonic ecosystems. However, the manual processing and annotation of the vast amounts of data generated by these devices remain time-consuming tasks. In this context, machine learning-based classification models offer a promising solution. In this data paper, we introduce a new labeled freshwater plankton dataset comprising approximately 88,000 images distributed across 43 taxa. We also present the labeling assistance method we used to facilitate dataset annotation. Finally, we present a baseline based on a convolutional neural network (CNN), which achieves a classification accuracy of 93% on our dataset.

ecology↗

A training protocol for human classification of Asian elephant images from trail cameras

Trail cameras have become ubiquitous tools for ecological data collection over recent decades. Despite progress in the development of automated algorithms and artificial intelligence for image classification, our ability to process large volumes of data remain limited by the need for trained human observers to make refined judgements. We provide guidance on placement of trail cameras for observing Asian elephants (Elephas maximus) and outline a protocol for training and testing naive human observers in performing image classifications (age/sex class and group composition) that cannot yet be automated. This process can be used to develop a high-throughput workflow capable of extracting useful data from large volumes of images. Our training material consisted of 14,007 images collected from 6 trail cameras around Udawalawe National Park in Sri Lanka from 2017-2019. In the first stage, expert observers (n=3) trained a group of inexperienced participants (n=4), who engaged in an iterative process to develop a protocol document. The document was then tested on a second set of subjects (n=6) each of whom classified 350 test images in four separate sequential batches using quantitative measures of precision and accuracy. The test set was sampled from 54,435 images from an additional 25 cameras. When compared to expert observers, they achieved a fair level of precision (Fleiss' kappa = 0.247) and 82.6% accuracy. Our approach can usefully be extended to other species and contexts.

ecology↗

Forest belowground productivity and carbon allocation predominantly driven by soil properties rather than climate

Forests are threatened by a multitude of stressors, including anthropogenic disturbances and climate change. Assessing how forests will respond to these stressors requires a comprehensive understanding of net primary productivity (Npp), environmental constraints on growth, and adaptive capacity. A parameter of significant uncertainty is belowground Npp (bNpp), which can account for up to 80% of total Npp but is poorly estimated and rarely measured directly. We used a cross-biome dataset of direct, field-based measurements of aboveground and belowground primary productivity and 21 climatic and soil variables to identify potential constraints on bNpp and belowground carbon allocation in boreal and cold temperate forests. Soil variables, rather than climate variables, were the main drivers of bNpp and belowground allocation across biomes. The importance of soil variables suggests that soil nutrient dynamics, especially soil nutrient pool and flux variables, must be explicitly modeled to more accurately predict feedbacks between climate, productivity, and within-tree carbon allocation. Within biomes, environmental drivers of belowground allocation varied between low versus high allocation forests, indicating that environmental drivers are site-specific and the development of within-biome, site-scale classifications for forest ecosystems could be useful. Changes in soil variables, such as increasing soil nitrogen pools, caused abrupt and large decreases in bNpp for boreal, but not cold temperate forests. Threshold-like shifts indicate that boreal forests might have lower adaptive capacity and higher sensitivity to disturbances than cold temperate forests. With 70% of boreal forests characterized by low bNpp, disturbances such as anthropogenic nitrogen deposition could cause large-scale decreases in bNpp that could push these forests beyond their adaptive capacity.

ecology↗