Skip to Main content Skip to Navigation
Journal articles

Global gridded crop model evaluation: benchmarking, skills, deficiencies and implications

Christoph Müller Joshua Elliott 1 James Chryssanthacopoulos Almut Arneth 2 Juraj Balkovic 3 Philippe Ciais 4, 5 Delphine Deryng 6 Christian Folberth Michael Glotter Steven Hoek 7 Toshichika Iizumi Roberto C. Izaurralde 8 Curtis Jones Nikolay Khabarov Peter Lawrence 9 Wenfeng Liu 10, 11, 12 Stefan Olin Thomas A. M. Pugh Deepak K. Ray 13 Ashwan Reddy 14 Cynthia Rosenzweig 15 Alex C. Ruane 15 Gen Sakurai Erwin Schmid 16 Rastislav Skalsky Carol X. Song 17 Xuhui Wang Allard de Wit 18 Hong Yang
Abstract : Crop models are increasingly used to simulate crop yields at the global scale, but so far there is no general framework on how to assess model performance. Here we evaluate the simulation results of 14 global gridded crop modeling groups that have contributed historic crop yield simulations for maize, wheat, rice and soybean to the Global Gridded Crop Model Intercomparison (GGCMI) of the Agricultural Model Intercomparison and Improvement Project (AgMIP). Simulation results are compared to reference data at global, national and grid cell scales and we evaluate model performance with respect to time series correlation, spatial correlation and mean bias. We find that global gridded crop models (GGCMs) show mixed skill in reproducing time series correlations or spatial patterns at the different spatial scales. Generally, maize, wheat and soybean simulations of many GGCMs are capable of reproducing larger parts of observed temporal variability (time series correlation coefficients (r) of up to 0.888 for maize, 0.673 for wheat and 0.643 for soybean at the global scale) but rice yield variability cannot be well reproduced by most models. Yield variability can be well reproduced for most major producing countries by many GGCMs and for all countries by at least some. A comparison with gridded yield data and a statistical analysis of the effects of weather variability on yield variability shows that the ensemble of GGCMs can explain more of the yield variability than an ensemble of regression models for maize and soybean, but not for wheat and rice. We identify future research needs in global gridded crop modeling and for all individual crop modeling groups. In the absence of a purely observation-based benchmark for model evaluation, we propose that the best performing crop model per crop and region establishes the benchmark for all others, and modelers are encouraged to investigate how crop model performance can be increased. We make our evaluation system accessible to all crop modelers so that other modeling groups can also test their model performance against the reference data and the GGCMI benchmark.
Complete list of metadatas

Cited literature [78 references]  Display  Hide  Download
Contributor : Florence Aptel <>
Submitted on : Tuesday, October 27, 2020 - 1:45:46 PM
Last modification on : Thursday, October 29, 2020 - 3:16:09 AM


Publisher files allowed on an open archive



Christoph Müller, Joshua Elliott, James Chryssanthacopoulos, Almut Arneth, Juraj Balkovic, et al.. Global gridded crop model evaluation: benchmarking, skills, deficiencies and implications. Geoscientific Model Development, European Geosciences Union, 2017, 10 (4), pp.1403 - 1422. ⟨10.5194/gmd-10-1403-2017⟩. ⟨hal-01584200⟩



Record views


Files downloads