Pith. sign in

Paper Citation Record · LEDGER

RS-GPT4V: A Unified Multimodal Instruction-Following Dataset for Remote Sensing Image Understanding

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2406.12479.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.12479 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T18:17:17.851335Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T05:27:39.992089Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e1d15310-df64-4800-afd8-219bfce622f3 · inbound

Multimodal Mathematical Reasoning Embedded in Aerial Vehicle Imagery: Benchmarking, Analysis, and Exploration cites this paper.

Multimodal Mathematical Reasoning Embedded in Aerial Vehicle Imagery: Benchmarking, Analysis, and Exploration RS-GPT4V: A Unified Multimodal Instruction-Following Dataset for Remote Sensing Image Understanding

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-04T18:17:17.851335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:17:17.851335Z digest=sha256:5202c644c5f1264bab66a091f944fd2d34a843557f016541f57807a6aff66192

Observation 9b2b62d2-1cd9-4ea9-9d7e-55e2b0cd535e · inbound

MMLANDMARKS: a Cross-View Instance-Level Benchmark for Geo-Spatial Understanding cites this paper.

MMLANDMARKS: a Cross-View Instance-Level Benchmark for Geo-Spatial Understanding RS-GPT4V: A Unified Multimodal Instruction-Following Dataset for Remote Sensing Image Understanding

Reference 96

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:51:15.286707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T20:49:11.961079Z digest=sha256:ded41b0be6a4a60e981343fb0bd280491a08e07894a8e299541d0cc24bd496f0

Observation b9240cff-dbb0-47e1-ab6e-3b89350f566b · inbound

Vision-and-Language Navigation for UAVs: Progress, Challenges, and a Research Roadmap cites this paper.

Vision-and-Language Navigation for UAVs: Progress, Challenges, and a Research Roadmap RS-GPT4V: A Unified Multimodal Instruction-Following Dataset for Remote Sensing Image Understanding

Reference 139

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:10:29.793391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T13:48:08.135538Z digest=sha256:b84dd8e66ebe4b9fe572f3822a845e426844a6f6d8806dd8b778560eb15ee165

Observation 8dd208b8-a452-4e44-8c97-cd91f28824fe · inbound

Beyond GSD-as-Token: Continuous Scale Conditioning for Remote Sensing VLMs cites this paper.

Beyond GSD-as-Token: Continuous Scale Conditioning for Remote Sensing VLMs RS-GPT4V: A Unified Multimodal Instruction-Following Dataset for Remote Sensing Image Understanding

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:15:56.904525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-11T01:52:51.777228Z digest=sha256:51d46d09ee8a154cd104451d30feef172b359cc51385b8d1cffe8c8fd2efd524

Observation c3b7e2df-fb52-4088-8464-e9f1477270ee · inbound

Earth-OneVision: Extending Remote Sensing Multimodal Large Language Models to More Sensor Modalities and Tasks cites this paper.

Earth-OneVision: Extending Remote Sensing Multimodal Large Language Models to More Sensor Modalities and Tasks RS-GPT4V: A Unified Multimodal Instruction-Following Dataset for Remote Sensing Image Understanding

Reference 89

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:27:39.993530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T13:16:12.768793Z digest=sha256:2086ed6e593dadef0f713ad429c609a6bcbd49476aaab86cea417c3f31ebedb6

Observation 47114dda-ffff-4cdb-9ab6-8f15fcc4680d · inbound

Multimodal Large Language Models for Remote Sensing Image Understanding: Domain-Specific or General-Purpose? cites this paper.

Multimodal Large Language Models for Remote Sensing Image Understanding: Domain-Specific or General-Purpose? RS-GPT4V: A Unified Multimodal Instruction-Following Dataset for Remote Sensing Image Understanding

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-01T10:20:58.536227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T10:20:58.536227Z digest=sha256:2b63f8103605b6c263fc3c6f9cc854d78d541e3c78443a16f75e95d8cdb1c4f3