TerraMind

IBM, ESA, Forschungszentrum Jülich · 2025-04-15 · image · geospatial · generative · anchor · artifact

Any-to-any generative multimodal foundation model for Earth observation, able to translate between optical, SAR and auxiliary geospatial modalities. Pre-trained on TerraMesh with a dual-scale encoder-decoder over both pixel patches and modality tokens.

Note
the model is multimodal in the Earth-observation sense of several sensor modalities, all of which are raster, so it carries a single modality value on this axis
Variants
TerraMind v1 tiny, TerraMind v1 small, TerraMind v1 base, TerraMind v1 large

Findings

Shared mechanisms