Can a history of crop rotations improve the prediction of soil organic carbon in the Andes? integrating machine learning multi-annual crop classification as a proxy of soil management
Bueno, M.; Loayza, H.; Ninanya, J.; Rinza, J.; Briceno, P.; Silva, L.; Mestanza, C.; Otiniano, R.; Kreuze, J. F.; Ramirez, D. A.
Show abstract
Soil organic carbon (SOC) is a crucial component related to various processes that ensure soil health and function. Its modeling is vital for assessing and monitoring soil degradation caused by the potential impact of agricultural activities. This study aimed to model SOC in the Northern highlands of Peru, characterized by a high amount of SOC, which is being affected by crop expansion. Crop rotation (CR) was incorporated into a modeling exercise using remote sensing data, fieldwork, and farmer surveys. A multi-year classification model with seven cropland classes was developed using data collected from 534 fields across 2022-2024, including 189 soil samples. Each cropland field was represented as a polygon delineating its boundaries and indicating its dominant crop cover. Time series of multispectral Sentinel-2 Level-2A Top of Canopy imagery were used to derive phenological features--such as the timing of maximum canopy cover and the length of the growing period--based on Normalized Difference Vegetation Index (NDVI) time series. A Random Forest classifier was used as the baseline model. The cropland classification model demonstrated strong overall performance, with F1 scores ranging from 0.81 to 0.98 across the different classes. The model performed well for lupin and pasture but scored lower for beans and potatoes. Predictions of cropland classes from 2019 to 2022 were created, resulting in frequency layers that represent crop rotations. Four feature configurations were evaluated: (i) including all features as a benchmark, (ii) excluding climatology, (iii) excluding crop rotation history, and (iv) excluding soil properties. Configurations including all features and excluding crop rotation history showed the highest performance (R2 = 0.63), while those excluding climatology or soil properties performed worse (R2 {approx} 0.52-0.53). Although soil features were the most important, fallow frequency emerged as the most critical predictor of SOC in crop rotations. When soil data were excluded, fallow frequency, combined with climatic features, explained over half of the SOC variability. The findings emphasize the importance of incorporating CR into SOC mapping efforts.
Matching journals
The top 1 journal accounts for 50% of the predicted probability mass.
Similar papers in this journal
- Open Soil Spectral Library (OSSL): Building reproducible soil calibration models through open development and community engagement 96%
- Long-term assessment of ecosystem services at ecological restoration sites using Landsat time series 96%
- Downscaling Satellite Soil Moisture using Geomorphometry and Machine Learning 96%
Similar papers in this journal
- YieldNet: A Convolutional Neural Network for Simultaneous Corn and Soybean Yield Prediction Based on Remote Sensing Data 93%
- Organic manure managements increases soil microbial community structure and diversity in double-cropping paddy field of Southern China 91%
- Coupling Day Length Data and Genomic Prediction tools for Predicting Time-Related Traits under Complex Scenarios 91%
Similar papers in this journal
- Monitoring of drought stress and transpiration rate using proximal thermal and hyperspectral imaging in an indoor automated plant phenotyping platform 92%
- PI-Plat: A high-resolution image-based 3D reconstruction method to estimate growth dynamics of rice inflorescence traits 91%
- Wheat grain width: A clue for re-exploring visual indicators of grain weight 91%
Similar papers in this journal
- Prediction of harvest-related traits in barley using high-throughput phenotyping data and machine learning 91%
- An integrative process-based model for biomass and yield estimation of hardneck garlic (Allium sativum) 91%
- A High-Throughput Physiological Functional Phenotyping System for Time- and Cost-Effective Screening of Potential Biostimulants 91%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.