SYNTOLOGY HomeExplorerAtlasCodeMethodologyAboutDevelopersFeedPricing
Paper · 2306.03831 · 2023

GEO-Bench: Toward Foundation Models for Earth Monitoring

arXiv · PDF · Open in the Atlas

Code that ran

We lifted 6 functions out of this paper's own repositories and ran 0 of them in a sandbox. "Ran" means the function executed on a synthesized input and returned a value. It is not a reproduction of the paper's results.

RepositoryRoleRan
servicenow/geo-bench canonical 0 of 6
FunctionStatusWhere it lives
biqm Not yet run servicenow/geo-bench/geobench/plot_tools.py
code served (permissive licence) · get_code("a152a592c2d89882")
bootstrap_iqm Not yet run servicenow/geo-bench/geobench/plot_tools.py
code served (permissive licence) · get_code("d1b48a3d91d5cb10")
iqm Not yet run servicenow/geo-bench/geobench/plot_tools.py
code served (permissive licence) · get_code("0e5e544e9fd463f6")
overlay_mask Not yet run servicenow/geo-bench/make_benchmark/dataset_converters/forestnet.py
code served (permissive licence) · get_code("2918c0bc1d90bf98")
safe_load Not yet run servicenow/geo-bench/geobench/_safe_pickle.py
code served (permissive licence) · get_code("4cf91d58782c1644")
safe_loads Not yet run servicenow/geo-bench/geobench/_safe_pickle.py
code served (permissive licence) · get_code("c7568e4f73ee2a27")

Repositories linked to this paper

Some links come from the archived Papers with Code dataset (CC BY-SA 4.0): attribution and licence.

Abstract

Recent progress in self-supervision has shown that pre-training large neural networks on vast amounts of unsupervised data can lead to substantial increases in generalization to downstream tasks. Such models, recently coined foundation models, have been transformational to the field of natural language processing. Variants have also been proposed for image data, but their applicability to remote sensing tasks is limited. To stimulate the development of foundation models for Earth monitoring, we propose a benchmark comprised of six classification and six segmentation tasks, which were carefully curated and adapted to be both relevant to the field and well-suited for model evaluation. We accompany this benchmark with a robust methodology for evaluating models and reporting aggregated results to enable a reliable assessment of progress. Finally, we report results for 20 baselines to gain information about the performance of existing models. We believe that this benchmark will be a driver of progress across a variety of Earth monitoring tasks.

For agents

The same record, over MCP at https://syntology.ai/mcp:

get_harvested_code_for_paper("2306.03831")
get_code_for_paper("2306.03831")
have("2306.03831")

Connect an agent — have() is free.