Trait diversity metrics can perform well with highly incomplete datasets
Stewart, K.; Carmona, C. P.; Clements, C.; Venditti, C.; Tobias, J. A.; Gonzalez-Suarez, M.
Show abstract
O_LICharacterizing changes in trait diversity at large spatial scales provides insight into the impact of human activity on ecosystem structure and function. However, the approach is often based on trait datasets that are incomplete and unrepresentative, with uncertain impacts on trait diversity estimates. C_LIO_LITo address this knowledge gap, we simulated random and biased removal of data from a near complete avian trait dataset (9579 species) and assessed whether trait diversity metrics were robust to data incompleteness with and without using imputation to fill data gaps. Specifically, we compared two commonly used metrics each calculated with two methods: trait richness (calculated with convex hulls and trait probabilities densities) and trait divergence (calculated with distance-based Rao and trait probability densities). C_LIO_LIWithout imputation, estimates of global avian trait diversity (richness and divergence) were robust when 30-70% of species had missing data for four out of 11 continuous traits, depending on severity of bias and the method used. However, when missing traits were imputed based on present morphological trait data and phylogeny, trait diversity metrics consistently remained representative of the true value, even when 70% of species were missing data for four out of 11 traits and data were not missing at random (biased with respect to body mass). Trait probability densities and distance-based Rao were particularly robust to missingness and bias when combined with imputation, with convex hull-based trait richness being less reliable. C_LIO_LIExpanding global morphometric datasets to represent more taxa and traits, and to quantify intraspecific variation, remains a priority. In the meantime, our results show that widely used methods can successfully quantify large-scale trait diversity even when data are missing for two-thirds of species, so long as missing traits are estimated using imputation. C_LI
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- The best of two worlds: using stacked generalization for integrating expert range maps in species distribution models 94%
- Biological traits of seabirds predict extinction risk and vulnerability to anthropogenic threats 93%
- Genetic diversity varies with species traits and latitude in predatory soil arthropods (Myriapoda: Chilopoda) 92%
Similar papers in this journal
- The Hidden Side Of Diversity: Effects Of Imperfect Detection On Multiple Dimensions Of Biodiversity 95%
- How citizen science could improve Species Distribution Models and their independent assessment 93%
- Unlocking the potential of historical abundance datasets to study biomass change in flying insects 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.