Source code
github.com/the-ai-alliance/geo-bench-vlm
the official repository with harness, tasks and verifiers
Geospatial agent benchmark · not-downloaded
Also known as: GEOBench-VLM, geo-bench-vlm
Evaluates: model answers only
Geospatial VLM multiple-choice question-answering accuracy on remote-sensing imagery.
github.com/the-ai-alliance/geo-bench-vlm
the official repository with harness, tasks and verifiers
In shortAgents do not run code inside this benchmark, and answers are checked by recomputation, not by another model's opinion. Part of the evidence base (dataset or code, not both) is published. Not a primary spatial verifier. MCQ accuracy does not prove code execution, artifact production, or workflow correctness.
Geospatial VLM multiple-choice question-answering accuracy on remote-sensing imagery.
Not a primary spatial verifier. MCQ accuracy does not prove code execution, artifact production, or workflow correctness.
not-downloaded — Not downloaded. VLM sidecar benchmark and possible small pinned image-understanding seed.
This entry is available as JSON at /api/registry#geobench-vlm. See the registry endpoint.