When can similar fields improve yield predictions?

A traceable correction for frozen yield models, scored forward in time.

2,062 field-seasons3 crops4 countries2016–2024

Scroll Imagery © Esri, Maxar · illustrative

Across all fields

Without labels, the best correction depends on the crop.

–methods clearly help all three crops

Where the error lives

A field-level shift can reach most of the Base error.

–of in‑season Base MSE is a year or field offset

When it helps

Corn and soybean gain early, wheat late.

Year by year

Some years hurt.

–field-level harm across compared methods

With same-season labels

A few labels go furthest in wheat.

–of wheat MSE removed with labels on 10% of fields (shift only), in‑season

Limitations

What this study does—and does not—establish.

A short record, not a future guarantee.

Only 3–4 test years per crop. Intervals describe these observed years; they do not establish future-year reliability.

One benchmark, a bounded scope.

YieldSAT2,062 field-seasons · 2016–2024
AR ArgentinaBR BrazilUY UruguayDE Germany

Corn, soybean and wheat. Other crops and regions are untested. The three worked fields were selected after the fact, for illustration.

Traceable outputs; untested explanations.

Donors→Candidates→Action→Output
Selector fidelityUser trustUsabilityNot evaluated

Replay checks the sources and prediction composition. It does not validate the selector’s internal reasoning.

Same-season labels are an extra assumption.

A uses same-stage support labels, which can arrive later in calendar time than a query. B uses no same-season yield labels.

Evidence notes and unresolved choices

Exploratory, retrospective evidence; no prospective validation or multiple-comparison adjustment. An interval crossing zero does not establish equivalence. Year effects are entangled with annual Base checkpoints.

SO/SA ensemble versus seed-mean reporting remains an author decision. Existing website values retain their stated convention; neither estimates nor intervals are mixed across conventions.

Comparisons change inputs, model capacity and selection procedures together. B excludes parameter-based adaptation by scope. GT-based diagnostics cannot serve as deployment rules; Shapley applies only to its analysed retrieval object.

Explorer

Now correct one field yourself.

Pick a correction, set λ, lock. Then see the harvest.

Open the Explorer →