Search bioRxiv⌕ Search

bioRxiv · 10.64898/2026.02.23.707022

Geospatial foundation models enable data-efficient tree species mapping in temperate montane forests

Abstract

Accurate mapping of tree species from satellite data remains challenging in heterogeneous mountain forests due to environmental gradients, mixed stands, limited availability of high-purity training labels, and strong illumination-angle effects. Recent geospatial foundation models offer a new approach by learning generic, cloud-agnostic, information-rich representations from large multi-sensor archives suitable for a range of downstream tasks, but their ecological utility for species-level mapping remains incompletely understood. Here, we evaluate two geospatial foundation-model embeddings, AlphaEarth and Tessera, for tree species classification in the Trentino region of northern Italy, using parcel-level forest inventories as reference data (18 species and species groups). We compare their performance against conventional Sentinel-1+2 satellite composites across a series of controlled experiments examining classification accuracy, label efficiency, classifier complexity, robustness to label impurity, and temporal transferability. Foundation-model embeddings consistently outperform composite-based multispectral satellite baselines (weighted F1 = 0.83 vs. 0.80; macro F1 = 0.55 vs. 0.50), reaching near-asymptotic accuracy with as few as 5% of available training parcels and preserving ecologically meaningful structure aligned with functional and taxonomic groupings. However, realising this advantage requires a nonlinear classifier: a compact neural network provides better results than classic machine learning (i.e. Random Forest) and performs as well as deeper neural networks, while a linear classifier on foundation-model embeddings underperforms a neural network on conventional composites. Ancillary environmental covariates offer no additional classification benefit when added to embedding-based models. Classification accuracy remains robust to moderate levels of label impurity, allowing mixed parcels to be retained in the training dataset without substantial penalties, while training with parcel-level species proportions as soft labels achieves higher peak performance (macro F1 = 0.586 for Tessera, 0.589 for AlphaEarth) and lower Proportion L1 error than hard labels without requiring purity filtering, maximising the value of the full range of input data. However, temporal transfer across years reveals performance degradation, with weighted F1 declining by 9% for Tessera and 15% for AlphaEarth, and disproportionate losses for rare species. Overall, our results show that geospatial foundation models shift a primary bottleneck in species mapping from feature engineering toward the availability, quality, and temporal alignment of ecological reference data, while opening new opportunities for scalable biodiversity monitoring and the analysis of ecological change.

Source connections

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Ball, J. G. C., Wicklein, J. A., Feng, Z., Knezevic, J., Jaffer, S., Atzberger, C., Dalponte, M., Coomes, D.. 2026-02-24. Geospatial foundation models enable data-efficient tree species mapping in temperate montane forests. https://doi.org/10.64898/2026.02.23.707022

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

Beyond Single-Metric Assessments: Uncovering Masked Butterfly Declines via Multi-Scalar Analysis in Central Alberta

1. This study analyzed 21 years (2000-2025) of butterfly count data from Central Alberta, integrated with intensive 5-year (2021-2025) high-resolution intra-seasonal sampling. 2. Long-term macro-scale analysis revealed a significant decline in Shannon Diversity, a change that remained obscured when relying solely on traditional metrics of species richness and evenness. 3. This diversity decline was primarily driven by the severe, long-term collapse of the native Common Ringlet (Coenonympha tullia). 4. Four other dominant species--Cabbage White (Pieris rapae), Clouded Sulphur (Colias eriphyle), European Skipper (Thymelicus lineola), and Common Wood Nymph (Cercyonis pegala)--maintained long-term population stability, though their abundances were significantly constrained by extreme winter minimum temperatures and rapid spring warming. 5. High-resolution intra-seasonal analysis (2021-2025) demonstrated that community indices and species-specific abundances were strongly limited by daily weather, particularly wind velocity and temperature. 6. These findings illustrate that while traditional metrics like richness and evenness are fundamental to community ecology, they provide incomplete insights when applied in isolation; they are most effective when utilized as part of a complementary, multi-scalar framework. 7. This study highlights the necessity of coupling multi-decadal historical datasets with high-frequency, fine-scale sampling to accurately identify the mechanisms of community turnover that simpler metrics may overlook. 8. The results underscore the critical importance of standardized citizen science monitoring in quantifying environmental impacts and establishing conservation priorities for terrestrial insect groups.

ecology↗

From concentration to export: resource contrasts and bee traits shape pollinator spillover to crops

Floral plantings can either concentrate bees or export them to adjacent crops, yet the ecological conditions influencing these outcomes remain unclear. Here, we develop a mathematical model as proof of concept for our previous integrative hypothesis: concentrator and exporter outcomes can arise as alternative, context-dependent outcomes of the same underlying resource-selection process. Using bees as a model and focusing specifically on spillover from floral plantings to crops, we identified resource-specific thresholds separating concentration- and export-favoring conditions. Our model translates differences in relative patch attractiveness into context-dependent concentration and export outcomes and generates resource-specific, testable predictions about the conditions favoring pollinator movement into crops. In our simulations, the concentrator-exporter transition occurred at a lower flowering-intensity contrast than at pollen or nectar contrasts, which suggests that flowering intensity may provide an initial cue for bee movement, whereas nectar and pollen rewards refine or sustain bee responses once crops are perceived as attractive. Spillover thresholds differed among resource contrasts, whereas response steepness varied across bee-trait and community scenarios. Under the model's trait-sensitivity formulation, predicted spillover probability responded more strongly to flowering contrast for specialists than for generalists; colony size amplified this response, whereas bee richness dampened it. Together, these patterns show how flowering and resource contrasts interact with bee traits and community context to shape predicted spillover. Our results confirm that the concentrator and exporter hypotheses can be understood as context-dependent outcomes of the same ecological process rather than as mutually exclusive alternatives. Experimental tests of the predicted thresholds conducted in the field could reveal when and where floral plantings are most likely to promote bee spillover to crops, potentially supporting crop pollination.

ecology↗

A Computational Re-evaluation of Spatial Trials for Zoonotic Tuberculosis Control: Model Misspecification, Diagnostic Miss-classification, and the Illusion of Wildlife Culling Efficacy

1. Wildlife reservoir management frequently relies on the Randomised Badger Culling Trial's (RBCT) trade-off hypothesis, which posits that reductions in cattle herd infections are offset by a perturbation effect driven by disrupted host dispersal. This paper evaluates the computational and epidemiological robustness of this historical trial, which serves as the foundational empirical experiment guiding zoonotic tuberculosis (Mycobacterium bovis) control policies. 2. Using generalized linear mixed models with a generalized Poisson error distribution to explicitly address historical data overdispersion, this study contrasts traditional parametric inference against exact cluster-constrained permutation tests across distinct operational definitions of disease incidence. 3. Non-parametric diagnostics reveal that previously reported treatment and perturbation effects render as statistical artifacts under exact non-parametric permutation. Inside culling zones, parametric significance fails to withstand exact permutation verification due to extreme data leverage in localized cluster blocks. 4. Crucially, when diagnostic misclassification biases are eliminated by analysing total reactor datasets, all apparent culling effects disappear, and information criteria overwhelmingly favour nested null architectures. Unconfirmed reactors likely represent true biological infections missed by low-sensitivity post-mortem macro-necropsy, proving that host removal tracks observation noise rather than genuine zoonotic transmission pathways. 5. Finally, empirical scaling conducted in this study identifies a novel mathematical saturation effect, demonstrating that this sub-linear scaling is an operational artifact of unmodelled herd-level disease recurrence over time. 6. Policy implications. Because current zoonotic tuberculosis intervention frameworks are built upon a structurally misspecified statistical model, they have driven large-scale veterinary policies resulting in substantial, unevidenced ecological and economic interventions while failing to provide genuine public health, animal health, or disease control benefits.

ecology↗