REVIEW 3 cited by
GEO-Bench: Toward Foundation Models for Earth Monitoring
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Recent progress in self-supervision has shown that pre-training large neural networks on vast amounts of unsupervised data can lead to substantial increases in generalization to downstream tasks. Such models, recently coined foundation models, have been transformational to the field of natural language processing. Variants have also been proposed for image data, but their applicability to remote sensing tasks is limited. To stimulate the development of foundation models for Earth monitoring, we propose a benchmark comprised of six classification and six segmentation tasks, which were carefully curated and adapted to be both relevant to the field and well-suited for model evaluation. We accompany this benchmark with a robust methodology for evaluating models and reporting aggregated results to enable a reliable assessment of progress. Finally, we report results for 20 baselines to gain information about the performance of existing models. We believe that this benchmark will be a driver of progress across a variety of Earth monitoring tasks.
Forward citations
Cited by 3 Pith papers
-
SIMPLER: Efficient Foundation Model Adaptation via Similarity-Guided Layer Pruning for Earth Observation
CKA-based pre-fine-tuning layer pruning selects redundant ViT depth on unlabeled EO task data, cutting up to ~79% parameters while retaining most task performance and speeding both train and inference.
-
The View From Space: Navigating Instrumentation Differences with EOFMs
EOFM embeddings are strongly partitioned by sensor architecture, so matching spectral bands is not enough to make cross-sensor embedding search reliable.
-
GeoChain: Multimodal Chain-of-Thought for Geographic Reasoning
A new 21-step geographic reasoning benchmark built from 1.46 million street-view images shows current multimodal LLMs handle simple visual questions well but rarely localize precisely.
Discussion (0). Continue with ORCID to comment.