Pith. sign in

The table shows zero-shot performance with accuracy (%), failure rate (%), and selective accuracy (%)

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.AI 1

years

2025 1

verdicts

CONDITIONAL 1

representative citing papers

Towards LLM Agents for Earth Observation

cs.AI · 2025-04-16 · conditional · novelty 7.0

On a new 140-question Earth observation benchmark, the best LLM agent scores 33% accuracy with Google Earth Engine access because generated code fails to run over 58% of the time.

citing papers explorer

Showing 1 of 1 citing paper.

  • Towards LLM Agents for Earth Observation cs.AI · 2025-04-16 · conditional · none · ref 6

    On a new 140-question Earth observation benchmark, the best LLM agent scores 33% accuracy with Google Earth Engine access because generated code fails to run over 58% of the time.