Pith. sign in

Paper Citation Record · LEDGER

RSGPT: A Remote Sensing Vision Language Model and Benchmark

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2307.15266.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.15266 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T10:14:30.133950Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T17:05:51.453799Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 52a68107-ddfe-483d-9ca3-3aee66aa6a90 · inbound

SARChat-Bench-2M: A Multi-Task Vision-Language Benchmark for SAR Image Interpretation cites this paper.

SARChat-Bench-2M: A Multi-Task Vision-Language Benchmark for SAR Image Interpretation RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T10:14:30.133950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:14:30.133950Z digest=sha256:a36a56d2198954709931bf83a21dd108ae3610b9b01718b66be6e1cb1967d938

Observation 4ea8d0c7-8626-4ed6-97ff-c88b562bda53 · inbound

Chain-of-Talkers (CoTalk): Fast Human Annotation of Dense Image Captions cites this paper.

Chain-of-Talkers (CoTalk): Fast Human Annotation of Dense Image Captions RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:54.615881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:54.615881Z digest=sha256:ecb50a845d277a45ef80c9f0b99976f3507630de84c691227dc38baea2eba7bf

Observation 89386da8-4e18-4d7c-801e-a4d5714a0262 · inbound

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding cites this paper.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:21.928081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:21.928081Z digest=sha256:5fec762d6fb686b07441d30458bc4e3b6855daa5528109ff3a8028f0dae1649c

Observation 3edb0cda-0e2e-45c5-b5c1-bc48b5709486 · inbound

GeoProg3D: Compositional Visual Reasoning for City-Scale 3D Language Fields cites this paper.

GeoProg3D: Compositional Visual Reasoning for City-Scale 3D Language Fields RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T21:49:43.574175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:49:43.574175Z digest=sha256:e67a061bab1b979ba9e6fd05363b232bc8b84df2b7d48f31cc49dfbc7325bbad

Observation d91cd1ef-41d9-4f02-9fbf-215499192c50 · inbound

A Satellite-Ground Synergistic Large Vision-Language Model System for Earth Observation cites this paper.

A Satellite-Ground Synergistic Large Vision-Language Model System for Earth Observation RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-06T19:24:45.014869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:24:45.014869Z digest=sha256:885ffb811d32243510ae0b38ed1256705aa24e5679a8f8396aa3d153109c0d2c

Observation 4e9ac6fd-9126-4258-8354-b240f88382ba · inbound

GeoMag: A Vision-Language Model for Pixel-level Fine-Grained Remote Sensing Image Parsing cites this paper.

GeoMag: A Vision-Language Model for Pixel-level Fine-Grained Remote Sensing Image Parsing RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:13.189322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:13.189322Z digest=sha256:1e9d01a63206a899dd2f3534cbe00fe1b4404396712ba62a3a4a1c2219ea4a25

Observation 88f2835b-5272-4196-83ea-6c24caab3b65 · inbound

Enhancing Remote Sensing Vision-Language Models Through MLLM and LLM-Based High-Quality Image-Text Dataset Generation cites this paper.

Enhancing Remote Sensing Vision-Language Models Through MLLM and LLM-Based High-Quality Image-Text Dataset Generation RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T15:07:32.895752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:07:32.895752Z digest=sha256:90259be351016b87b896227dd652040651dbd592e750059402af31e3cb72b887

Observation 4fc331a4-3056-4123-b880-b41c7de012e1 · inbound

Few-Shot Vision-Language Reasoning for Satellite Imagery via Verifiable Rewards cites this paper.

Few-Shot Vision-Language Reasoning for Satellite Imagery via Verifiable Rewards RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T12:29:37.161294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:29:37.161294Z digest=sha256:34ed5405d2ae81936af935358a49f56ad9090b463acf3cf0efdeb8a284810f9a

Observation 84425ee0-fc77-48b2-ba75-c0e776e709ec · inbound

WildfireVLM: AI-powered Analysis for Early Wildfire Detection and Risk Assessment Using Satellite Imagery cites this paper.

WildfireVLM: AI-powered Analysis for Early Wildfire Detection and Risk Assessment Using Satellite Imagery RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:17:22.437778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T05:14:25.778239Z digest=sha256:d723c6f8cebd2a2679bd1603af3cb63bdd3da5933c6fa5b49981d94e4666c74d

Observation fdae9e8b-69a3-4b7a-b2d3-c0fbc8a65ea0 · inbound

Geo2Sound: A Scalable Geo-Aligned Framework for Soundscape Generation from Satellite Imagery cites this paper.

Geo2Sound: A Scalable Geo-Aligned Framework for Soundscape Generation from Satellite Imagery RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:59:03.522206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T09:52:35.741400Z digest=sha256:e070e0ae7f6c304271b52bc69f6fb7af00950564f38fce30589cda9389dba19c

Observation 149d635b-41a2-4a53-be7b-bcbc09121d74 · inbound

ChangeQuery: Advancing Remote Sensing Change Analysis for Natural and Human-Induced Disasters from Visual Detection to Semantic Understanding cites this paper.

ChangeQuery: Advancing Remote Sensing Change Analysis for Natural and Human-Induced Disasters from Visual Detection to Semantic Understanding RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:06:09.807361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T12:41:38.571997Z digest=sha256:97cfb774d4934a275a148263b7f3be137abf2d4169e40bd7850a12348422c8b7

Observation fcf196cd-2b09-4b70-bd31-e1c570e8d72e · inbound

Beyond GSD-as-Token: Continuous Scale Conditioning for Remote Sensing VLMs cites this paper.

Beyond GSD-as-Token: Continuous Scale Conditioning for Remote Sensing VLMs RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:15:56.985551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T01:52:51.777228Z digest=sha256:dfc6e6651d4633a7ab81f1609436554c6219db377ed44c0a41270738fdbe737d

Observation 8dcd6151-2a25-462d-a736-66e18c306fc0 · inbound

GeoVista: Visually Grounded Active Perception for Ultra-High-Resolution Remote Sensing Understanding cites this paper.

GeoVista: Visually Grounded Active Perception for Ultra-High-Resolution Remote Sensing Understanding RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:53:33.974344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T02:39:56.666424Z digest=sha256:4b01b675f3599d60eee9b5277c0c075e6dada993b1fd2bef804605a9338fc465

Observation 4419f07d-7ece-4962-831b-b018f9877d56 · inbound

OmniCD: A Foundational Framework for Remote Sensing Image Change Detection Guided by Multimodal Semantics cites this paper.

OmniCD: A Foundational Framework for Remote Sensing Image Change Detection Guided by Multimodal Semantics RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T08:13:16.011552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T08:03:27.451288Z digest=sha256:2f90f028651de0c2941ebac8ada2c5c5073768d67910c4a83e89aeb5e76a39f3

Observation 905bb316-54d7-40f5-a2d2-cd856d52b5c4 · inbound

RSICCLLM: A Multimodal Large Language Model for Remote Sensing Image Change Captioning cites this paper.

RSICCLLM: A Multimodal Large Language Model for Remote Sensing Image Change Captioning RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T17:05:51.455236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T04:09:32.397341Z digest=sha256:743ab1d1b707b309b88cadac3da55054a6583840afe8eb709e01432f36aef0a9

Observation 702be6a7-ebeb-4edb-9787-cc5e61a184f9 · inbound

WeaveEarth: Structured Evidence Construction and Reasoning for Training-Free UHR Remote Sensing Understanding cites this paper.

WeaveEarth: Structured Evidence Construction and Reasoning for Training-Free UHR Remote Sensing Understanding RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-14T14:09:30.395518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T14:09:30.395518Z digest=sha256:7d0f0a6cc2787e42ef4f0964effe86d135c2ec918662663e44e9d557ebcdda56