Determine which foundation-model design choices control wildfire-disturbance sensitivity
Determine which design choices—including spatial context, temporal sampling, input modalities, training targets, and loss functions—make annual Earth-observation embeddings more sensitive to abrupt wildfire disturbance, thereby explaining the differing disturbance sensitivity of Tessera and AlphaEarth.
References
The source of this difference remains unclear, as the models differ in training data, input sensors, spatial context, architecture, loss functions, temporal sampling, and training targets. One notable pattern is that adding spatial context through U-Net models improved performance much more for AlphaEarth than for Tessera, despite AlphaEarth embeddings already containing patch-level spatial context. The pattern suggests that when a pixel's annual trajectory already contains a clear disturbance signal, strong temporal encoding reduces dependence on spatial decoding. AlphaEarth's latent projections also seem more strongly organized by land cover in the fire year, which may partly reflect differences in training objectives, including its use of land-cover information as a target. Without systematic ablations, however, we cannot attribute the observed differences to any single design choice. Future work should isolate the effects of spatial context, temporal sampling, input modalities, training targets, and loss functions to determine what makes annual embeddings more sensitive to abrupt disturbance.