Back

Global observation of plankton communities from space

Kaneko, H.; Endo, H.; Henry, N.; Berney, C.; Mahe, F.; Poulain, J.; Labadie, K.; Beluche, O.; El Hourany, R.; Tara Oceans Coordinators, ; Chaffron, S.; Wincker, P.; Nakamura, R.; Karp-Boss, L.; Boss, E.; Bowler, C.; de Vargas, C.; Tomii, K.; Ogata, H.; Acinas, S. G.; Babin, M.; Bork, P.; Cochrane, G.; Gorsky, G.; Guidi, L.; Grimsley, N.; Hingamp, P.; Iudicone, D.; Jaillon, O.; Kandels, S.; Karsenti, E.; Not, F.; Poulton, N.; Pesant, S.; Sardet, C.; Speich, S.; Stemmann, L.; Sullivan, M. B.; Sunagawa, S.

2022-09-26 ecology
10.1101/2022.09.23.508961 bioRxiv
Show abstract

Satellite remote sensing from space is a powerful way to monitor the global dynamics of marine plankton. Previous research has focused on developing models to predict the size or taxonomic groups of phytoplankton. Here we present an approach to identify representative communities from a global plankton network that included both zooplankton and phytoplankton and using global satellite observations to predict their biogeography. Six representative plankton communities were identified from a global co-occurrence network inferred using a novel rDNA 18S V4 planetary-scale eukaryotic metabarcoding dataset. Machine learning techniques were then applied to train a model that predicted these representative communities from satellite data. The model showed an overall 67% accuracy in the prediction of the representative communities. The prediction based on 17 satellite-derived parameters showed better performance than based only on temperature and/or the concentration of chlorophyll a. The trained model allowed to predict the global spatiotemporal distribution of communities over 19-years. Our model exhibited strong seasonal changes in the community compositions in the subarctic-subtropical boundary regions, which were consistent with previous field observations. This network-oriented approach can easily be extended to more comprehensive models including prokaryotes as well as viruses.

Matching journals

The top 8 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.