Pith. sign in

Paper Citation Record · LEDGER

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing

As of 13 August 2026, this Paper Citation Record lists 78 of 78 outbound references and 0 inbound Pith citation observations for arXiv:2607.25993.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.25993 v1

Coverage vector

measured 78 of 78 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T00:59:26.218778Z

measured 78 of 78 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

78 of 78 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved78
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c11bfa9e-2aa5-4ea6-975d-4cb1758b7ac5 · outbound

This paper cites Geollava-8k: Scaling remote-sensing multimodal large language models to 8k resolution.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Geollava-8k: Scaling remote-sensing multimodal large language models to 8k resolution

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:16.872554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:16.872554Z digest=sha256:ddce8f1ae84af9c89c0525398c63c13c1cb7b87c0ed17516797332f54509139a

Observation 52b5e044-50a9-4d0c-b289-e25af79c9551 · outbound

This paper cites XLRS-Bench: Could Your Multimodal LLMs Understand Extremely Large Ultra-High-Resolution Remote Sensing Imagery?.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing XLRS-Bench: Could Your Multimodal LLMs Understand Extremely Large Ultra-High-Resolution Remote Sensing Imagery?

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:16.958440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:16.958440Z digest=sha256:dfec5ef060014101d76e46314b1e4c772c4bdf4b2bd7ee25f0836f14783e8838

Observation 814304eb-51a7-4736-82ce-a7c1f09ffb4a · outbound

This paper cites A benchmark for ultra-high-resolution remote sensing mllms.arXiv preprint arXiv:2512.17319, 2025.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing A benchmark for ultra-high-resolution remote sensing mllms.arXiv preprint arXiv:2512.17319, 2025

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:17.014362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:17.014362Z digest=sha256:2e79ab941f1194d331bd5a261f1e86136844437e2d1aae1b3d1da3182e12704f

Observation 1f60195a-2db3-4a4c-bf89-603f0b4f4429 · outbound

This paper cites Geochat: Grounded large vision-language model for remote sensing.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Geochat: Grounded large vision-language model for remote sensing

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:17.090328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:17.090328Z digest=sha256:b1f0362ab0addf418609fa3dd803e380ee5feb3023beebfb480d9dea9e98d66e

Observation 2f6ca9cc-3a94-4c7b-8b69-ad244cf260e6 · outbound

This paper cites Earthgpt: A universal multi-modal large language model for multi-sensor image comprehension in remote sensing domain.IEEE Transactions on Geoscience and Remote Sensing, 2024.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Earthgpt: A universal multi-modal large language model for multi-sensor image comprehension in remote sensing domain.IEEE Transactions on Geoscience and Remote Sensing, 2024

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:17.215092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:17.215092Z digest=sha256:7aab1739b9e2f139af0d3c6f777d7509e4c1064a5ced8603e856ec6436adc98d

Observation 12285465-c255-4adc-b81b-fb0327377206 · outbound

This paper cites Rsgpt: A remote sensing vision language model and benchmark.ISPRS Journal of Photogrammetry and Remote Sensing, 224:272–286, 2025.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Rsgpt: A remote sensing vision language model and benchmark.ISPRS Journal of Photogrammetry and Remote Sensing, 224:272–286, 2025

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:17.318819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:17.318819Z digest=sha256:85e349345962aec0212d563a9b98fef0341dd2b2ac5aa72319bc8a4d917c011f

Observation 09b689b6-fee9-451c-8730-1b2f1e95c2aa · outbound

This paper cites Skyeyegpt: Unifying remote sensing vision- language tasks via instruction tuning with large language model.ISPRS Journal of Photogram- metry and Remote Sensing, 221:64–77, 2025.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Skyeyegpt: Unifying remote sensing vision- language tasks via instruction tuning with large language model.ISPRS Journal of Photogram- metry and Remote Sensing, 221:64–77, 2025

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:17.439421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:17.439421Z digest=sha256:e07e171d1555ddcb733a52c132299a5ac6f3fd7b355a1a2ffb632a776c90bc40

Observation e8b1fcf9-a2d4-49f6-847d-d4841751cd23 · outbound

This paper cites When Large Vision-Language Model Meets Large Remote Sensing Imagery: Coarse-to-Fine Text-Guided Token Pruning.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing When Large Vision-Language Model Meets Large Remote Sensing Imagery: Coarse-to-Fine Text-Guided Token Pruning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:17.520344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:17.520344Z digest=sha256:0c10381fb3e73da1a17b25bf3c7bdddf95f11735287e87d81391c89e02e93e3c

Observation ece5afc0-163a-4312-8ea0-0ee64030683d · outbound

This paper cites DeepEyes: Incentivizing "Thinking with Images" via Reinforcement Learning.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing DeepEyes: Incentivizing "Thinking with Images" via Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:17.566992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:17.566992Z digest=sha256:3e632fc4e31e9095fddf1c855b8c57ad892f5d1ecc5b1898c134392377962322

Observation 18bd7123-58d7-4a9a-a111-c29594e4941d · outbound

This paper cites DeepEyesV2: Toward Agentic Multimodal Model.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing DeepEyesV2: Toward Agentic Multimodal Model

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:17.736485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:17.736485Z digest=sha256:199ec8c8a42e34d6393db4b764146e2f618ee8dd9aa73c4c91159b560f499519

Observation 266e68f1-a109-4f4e-ba5c-636564131924 · outbound

This paper cites an unresolved cited work.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:17.889133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:17.889133Z digest=sha256:192c050c8016d2e1e853dcc25e12b6f9d5313f8db9c8a2c1259b4b095b604ebd

Observation f8d5eb53-742d-41f6-aa39-8408a6330817 · outbound

This paper cites an unresolved cited work.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:18.019185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:18.019185Z digest=sha256:ef30aa5779cb2fe525fda54a6abe7f4d6ed2bd1aa524ac84a235d723e5ce8036

Observation 08aba2c1-45c0-46a2-a252-526c7ffe7386 · outbound

This paper cites Zoomearth: Active perception for ultra-high-resolution geospatial vision-language tasks.arXiv preprint arXiv:2511.12267, 2025.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Zoomearth: Active perception for ultra-high-resolution geospatial vision-language tasks.arXiv preprint arXiv:2511.12267, 2025

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:18.151653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:18.151653Z digest=sha256:0148d820e524d918bc93565fc889217f738e71de043c170f85bb9d25ac0c95c4

Observation 2d4f191d-b8e0-4086-bc56-bd4518801e57 · outbound

This paper cites Geoeyes: On-demand visual focusing for evidence-grounded understanding of ultra-high-resolution remote sensing imagery.arXiv preprint arXiv:2602.14201, 2026.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Geoeyes: On-demand visual focusing for evidence-grounded understanding of ultra-high-resolution remote sensing imagery.arXiv preprint arXiv:2602.14201, 2026

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:18.214762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:18.214762Z digest=sha256:893a5e15ccaec71e0545e2389275b00a3c247d372dfbd13c126beebadfb7624c

Observation 045eb384-c08e-49d9-af74-e8ac2de503a9 · outbound

This paper cites Codev: Code with images for faithful visual reasoning via tool-aware policy optimization.arXiv preprint arXiv:2511.19661, 2025.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Codev: Code with images for faithful visual reasoning via tool-aware policy optimization.arXiv preprint arXiv:2511.19661, 2025

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:18.286594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:18.286594Z digest=sha256:0145d08ce234cae2aa66263611644e1966a100237b40a5b23054f88d9f1dd988

Observation 7fbffac7-fd71-454a-857e-340f563b60cc · outbound

This paper cites Zooming without zooming: Region-to-image distillation for fine-grained multimodal perception.arXiv preprint arXiv:2602.11858, 2026.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Zooming without zooming: Region-to-image distillation for fine-grained multimodal perception.arXiv preprint arXiv:2602.11858, 2026

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:18.418541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:18.418541Z digest=sha256:3b444df461b8c3534c1f5ce9d336814f107616a5b3e1fe17d6dc339cab2cfe9d

Observation 8d8d84ba-c45b-4320-850e-cf1df276870b · outbound

This paper cites What Does Vision Tool-Use Reinforcement Learning Really Learn? Disentangling Tool-Induced and Intrinsic Effects for Crop-and-Zoom.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing What Does Vision Tool-Use Reinforcement Learning Really Learn? Disentangling Tool-Induced and Intrinsic Effects for Crop-and-Zoom

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:18.518291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:18.518291Z digest=sha256:8d71c16ec34b2e7b966b1bb733eca0590d805334665ad6a2c99bb555c0217ce9

Observation ab95e434-3b5a-4c3d-9e9b-bc7bfe185316 · outbound

This paper cites Reinforced attention learning.arXiv preprint arXiv:2602.04884, 2026.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Reinforced attention learning.arXiv preprint arXiv:2602.04884, 2026

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:18.640925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:18.640925Z digest=sha256:f1ed160edd22ba8125f944c20f5b3b96a7448e6026eb55e3e3ef68dfbd639bbe

Observation 68b5ea20-e143-40b3-a051-c50550194ebd · outbound

This paper cites Smith, and Ranjay Krishna.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Smith, and Ranjay Krishna

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:18.738278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:18.738278Z digest=sha256:959eaa92ef4237b3d87b6e534bbcd1d72eee15bef92855de6a3b38150c7c25c9

Observation df08d272-e693-4579-855c-cdac8a691be9 · outbound

This paper cites Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:18.845444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:18.845444Z digest=sha256:7d6875c477b851b4cb982a3f59cdadafecf2651a6c45a08c4d1893bb52672293

Observation 13047702-0a6e-46ae-ae43-c7c2fe135170 · outbound

This paper cites an unresolved cited work.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:18.999170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:18.999170Z digest=sha256:42b695e1dd2c660fb10d484d934981e82c10fc6b2192879f475b4ffda678b0d0

Observation aff2f9cf-eeee-4b84-8efd-093bea6c8f14 · outbound

This paper cites an unresolved cited work.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:19.149637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:19.149637Z digest=sha256:44e957a15b6979201c25d86cb4fac09efd0887c69826204898abb30da48030ef

Observation 4f68a163-986b-43dd-a9f1-828720e18bef · outbound

This paper cites Lhrs-bot: Em- powering remote sensing with vgi-enhanced large multimodal language model.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Lhrs-bot: Em- powering remote sensing with vgi-enhanced large multimodal language model

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:19.257501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:19.257501Z digest=sha256:75a2671b85a52905b9421e81be917918c51f6539acf2976c7348510570ed90d1

Observation 848e696c-d13d-4c81-8ba3-f004776e211c · outbound

This paper cites LHRS-Bot-Nova: Improved Multimodal Large Language Model for Remote Sensing Vision-Language Interpretation.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing LHRS-Bot-Nova: Improved Multimodal Large Language Model for Remote Sensing Vision-Language Interpretation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:19.359533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:19.359533Z digest=sha256:c52fb62f719055fa244d96c435f7651146fe59b6e0c0380ee1d8e72c378da9ce

Observation b29b282b-bab2-41eb-835f-e708125f6286 · outbound

This paper cites Earthmind: Leveraging cross-sensor data for advanced earth observation interpretation with a unified multimodal llm.arXiv preprint arXiv:2506.01667, 2025.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Earthmind: Leveraging cross-sensor data for advanced earth observation interpretation with a unified multimodal llm.arXiv preprint arXiv:2506.01667, 2025

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:19.435240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:19.435240Z digest=sha256:26e8a8cbc76400b7fa624bff250b26f3cac4fdc00e4a002b7774dc22f5356697

Observation b5817d38-dbd8-4c31-aca0-310942558b45 · outbound

This paper cites Klein, Salman Khan, and Fahad Khan.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Klein, Salman Khan, and Fahad Khan

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:19.486302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:19.486302Z digest=sha256:2742ebb2b2760b4a3cd6fcb485ed1bee17b82d66f280086e1e1939eda686a0bc

Observation a97d61d4-c369-45e6-a60b-b6fa5fc67fb6 · outbound

This paper cites an unresolved cited work.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:19.578996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:19.578996Z digest=sha256:2c56f20892b7163ad0d09f1d9b2ed3876050f9e8a0f41a0be2e6fb07b334b4d8

Observation d23355bd-07ba-4550-bdd9-d3925996e483 · outbound

This paper cites an unresolved cited work.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:19.658554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:19.658554Z digest=sha256:3bd159400e2a4c72f2f4cb91ad46269bd8e1de9278ae914e2200a021bd5311b7

Observation 2563d57c-1345-4621-baed-f8ab87791800 · outbound

This paper cites Earthmarker: A visual prompting multi-modal large language model for remote sensing.IEEE Transactions on Geoscience and Remote Sensing, 2024.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Earthmarker: A visual prompting multi-modal large language model for remote sensing.IEEE Transactions on Geoscience and Remote Sensing, 2024

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:19.789342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:19.789342Z digest=sha256:f6085283dfa622c5bda815a3591a127769e8daa87e357dd1f95b8dedb612c3d5

Observation 53f4792a-f73e-43c8-a7ad-4263692d2508 · outbound

This paper cites Rsunivlm: A unified vision-language model for remote sensing via granularity-oriented moe.Pattern Recognition, 179:113717, 2026.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Rsunivlm: A unified vision-language model for remote sensing via granularity-oriented moe.Pattern Recognition, 179:113717, 2026

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:19.925428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:19.925428Z digest=sha256:1372dca0a01cbe5a39dd19b51269e2ef80e13afff246e1e6a212a62635777a12

Observation 59ae6266-280f-4309-814a-f39ee181ff7e · outbound

This paper cites Bermano, and Ohad Fried.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Bermano, and Ohad Fried

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:20.039300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:20.039300Z digest=sha256:0e8fd2d2d7612a342b4df099cc83f2c81c698a768f879b245440b0382b9490db

Observation bc98720a-f71e-4bdf-8341-3d066909722d · outbound

This paper cites Chatearthnet: A global- scale image-text dataset empowering vision-language geo-foundation models.Earth System Science Data Discussions, pages 1–24, 2024.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Chatearthnet: A global- scale image-text dataset empowering vision-language geo-foundation models.Earth System Science Data Discussions, pages 1–24, 2024

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:20.113408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:20.113408Z digest=sha256:8e9730696f9933e0d64d9f234bba73e8ebba1790b4e658f6db2cccd98bbb4f43

Observation 26665e13-7435-416b-8642-190480e9833e · outbound

This paper cites Rsvqa: Visual question answering for remote sensing data.IEEE Transactions on Geoscience and Remote Sensing, 58(12):8555– 8566, 2020.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Rsvqa: Visual question answering for remote sensing data.IEEE Transactions on Geoscience and Remote Sensing, 58(12):8555– 8566, 2020

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:20.219206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:20.219206Z digest=sha256:70c341813d5d054bcc634abc95ae94f2839fb236008f91bfcfa6bd80d3415603

Observation a384ffea-9a3a-4e32-80c3-00cc45c5a7b0 · outbound

This paper cites Mutual attention inception network for remote sensing visual question answering.IEEE Transactions on Geoscience and Remote Sensing, 60:1–14, 2021.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Mutual attention inception network for remote sensing visual question answering.IEEE Transactions on Geoscience and Remote Sensing, 60:1–14, 2021

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:20.396216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:20.396216Z digest=sha256:2fc61708c786f42348ff90e898f6f387ef378441ec3f3dd8dc45615b6b818416

Observation 38bcf153-7b20-4316-b4a6-0ae3f5e5e6bd · outbound

This paper cites an unresolved cited work.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:20.604144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:20.604144Z digest=sha256:07f5debdf288fc72a8682d817daf31e8c7edb0c04867a4032b5b51ba43b651aa

Observation 6e171e44-4f04-46e4-97f5-1d0ae0d5def9 · outbound

This paper cites VHM: Versatile and Honest Vision Language Model for Remote Sensing Image Analysis.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing VHM: Versatile and Honest Vision Language Model for Remote Sensing Image Analysis

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:20.770802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:20.770802Z digest=sha256:ca5d100f9981374684e5beb5849a3fa10f094d3066cee6f4ba8e419a5b68549f

Observation efde5a10-6ec6-4f5a-a206-f4331c1d89bd · outbound

This paper cites SkySenseGPT: A Fine-Grained Instruction Tuning Dataset and Model for Remote Sensing Vision-Language Understanding.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing SkySenseGPT: A Fine-Grained Instruction Tuning Dataset and Model for Remote Sensing Vision-Language Understanding

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:20.911720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:20.911720Z digest=sha256:8c81f06bf753e603f71ce264fc13f033021be81250f0d6d27a18467390f16bf1

Observation e31ab13a-0a6f-49c5-83a1-e314d39e4076 · outbound

This paper cites Vhm: Versatile and honest vision language model for remote sensing image analysis.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Vhm: Versatile and honest vision language model for remote sensing image analysis

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:21.061018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:21.061018Z digest=sha256:76994f82e57b22729a3616c80ffed1c678ef18b3a3f901583758881b63522081

Observation 3dab8c60-d90d-4613-93e0-11591fb83241 · outbound

This paper cites VRSBench: A Versatile Vision-Language Benchmark Dataset for Remote Sensing Image Understanding.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing VRSBench: A Versatile Vision-Language Benchmark Dataset for Remote Sensing Image Understanding

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:21.171823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:21.171823Z digest=sha256:553b15277262e161e441c8d08aab8d636a55ba3734ccef42c34547f00da56c68

Observation 8aae2d44-7d1e-4172-ad1c-c475cfd70279 · outbound

This paper cites Thinking with Images for Multimodal Reasoning: Foundations, Methods, and Future Frontiers.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Thinking with Images for Multimodal Reasoning: Foundations, Methods, and Future Frontiers

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:21.369593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:21.369593Z digest=sha256:fc6ea4350c75787b57505b91e80319cfb932b6f385527b74565ef269ce3371e9

Observation 12c40016-1136-4111-a9a1-a24b220f201d · outbound

This paper cites VisualToolAgent (VisTA): A Reinforcement Learning Framework for Visual Tool Selection.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing VisualToolAgent (VisTA): A Reinforcement Learning Framework for Visual Tool Selection

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:21.527737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:21.527737Z digest=sha256:7b646ea177b845cf0771a52227f86519a0954fa78eaa86581a5fcfbc95e63e8d

Observation b05c17d7-b308-4078-b392-782864e90a7f · outbound

This paper cites Spacetools: Tool-augmented spatial reasoning via double interactive rl, 2025.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Spacetools: Tool-augmented spatial reasoning via double interactive rl, 2025

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:21.686719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:21.686719Z digest=sha256:abe7cd54852b1ab632afdb8fc67e151cb6dc92d89aec62d660557f08c12e5db4

Observation b16b97a9-0b9a-4018-ab06-12f1346796d1 · outbound

This paper cites Reinforcing VLMs to Use Tools for Detailed Visual Reasoning Under Resource Constraints.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Reinforcing VLMs to Use Tools for Detailed Visual Reasoning Under Resource Constraints

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:21.754125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:21.754125Z digest=sha256:3671f43750ddaf9d6001530a3087b6fc3b7cc725168118719af394ef15480890

Observation 81d1eaf5-cb09-45af-abb6-2528bd1a5cad · outbound

This paper cites CropVLM: Learning to Zoom for Fine-Grained Vision-Language Perception.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing CropVLM: Learning to Zoom for Fine-Grained Vision-Language Perception

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:21.920265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:21.920265Z digest=sha256:a8a150cb3ab6567f726bedec34f56dd0a4769779c3d20a707971d7043933322c

Observation b5c9ff74-6e74-48f0-b20a-0614f1eb827d · outbound

This paper cites Adaptive Chain-of-Focus Reasoning via Dynamic Visual Search and Zooming for Efficient VLMs.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Adaptive Chain-of-Focus Reasoning via Dynamic Visual Search and Zooming for Efficient VLMs

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:22.118346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:22.118346Z digest=sha256:ca410447f4baa009f198a58f43ff2fec1ad5c748d526858f156eed8cf4e8a1fa

Observation 5ea80939-15a8-41db-919c-9e1fd68a0d8b · outbound

This paper cites Sensenova- mars: Empowering multimodal agentic reasoning and search via reinforcement learning.arXiv preprint arXiv:2512.24330, 2025.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Sensenova- mars: Empowering multimodal agentic reasoning and search via reinforcement learning.arXiv preprint arXiv:2512.24330, 2025

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:22.279050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:22.279050Z digest=sha256:558b50eab3fec92e3c60464d647937bbc73ea690d48b1c481c29a1524144113f

Observation f3b598d4-7731-463b-9f71-738c39bc372d · outbound

This paper cites Pixel Reasoner: Incentivizing Pixel-Space Reasoning with Curiosity-Driven Reinforcement Learning.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Pixel Reasoner: Incentivizing Pixel-Space Reasoning with Curiosity-Driven Reinforcement Learning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:22.526800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:22.526800Z digest=sha256:526a5c442e2053086489b0b5e9d472392999857f455d21cae726186cdbcf060b

Observation 4da40ed0-c0d3-441c-9daa-7a8a0c18bd8f · outbound

This paper cites Reinforcing spatial reasoning in vision-language models with interwoven thinking and visual drawing, 2025.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Reinforcing spatial reasoning in vision-language models with interwoven thinking and visual drawing, 2025

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:22.676414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:22.676414Z digest=sha256:b605b5778228b3fa50ad6de82a92c746cb4532f5ceeb040bb23fc4ae9f2627aa

Observation 9813b036-09b0-4d65-ac6a-ba7a1bd105bd · outbound

This paper cites Qwen2.5-VL Technical Report.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Qwen2.5-VL Technical Report

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:22.748692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:22.748692Z digest=sha256:fb65e28c1675339ccb771ae79da53910c1013c3995a4cb76c50826e9f9a3f198

Observation 110b0c09-67c1-4d65-b886-2b4f74bfd8a8 · outbound

This paper cites Introducing gpt-5.4.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Introducing gpt-5.4

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:22.837480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:22.837480Z digest=sha256:8769f52ce5c61a92182930572763aa1447324de5ec01ee6064542033ba48be29

Observation 969d2ab0-ff69-4ca3-ac98-c6f804f4bb31 · outbound

This paper cites The claude 4.6 model family.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing The claude 4.6 model family

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:23.081848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:23.081848Z digest=sha256:24fc9a708f7f8da63c559e849f50b7680e4e30c7ef9d6ca320460d1ddc0123ae

Observation 2f458609-188c-4961-b143-3132fb560d04 · outbound

This paper cites Qwen3-VL Technical Report.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Qwen3-VL Technical Report

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:23.284240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:23.284240Z digest=sha256:f50217f0da13c8ba9960286aaf73c58d6b9f4c4056e8a2b82e49f641f12e8068

Observation 8fcdab0b-ec14-4b69-a1d5-7ab571857588 · outbound

This paper cites Gpt-5.2: Advancing science and math.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Gpt-5.2: Advancing science and math

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:23.420815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:23.420815Z digest=sha256:307e9d14a29eeaccfc4cb54a0a3705846f2bd3350f4cd1e53061fbae58ac3848

Observation ed036522-578f-41d6-b106-3daf4b5e86d6 · outbound

This paper cites CogCoM: A Visual Language Model with Chain-of-Manipulations Reasoning.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing CogCoM: A Visual Language Model with Chain-of-Manipulations Reasoning

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:23.594318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:23.594318Z digest=sha256:b739f20ecf18341b9d869e392402a024bb43d306399b1c20a8999351c7935e15

Observation 4623532c-32af-45c6-a840-567602617eb1 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:23.699554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:23.699554Z digest=sha256:938fd8c84d3a5b9e08ce880e579fb7b4ec66c9ff4e92704f5a5af2c43934af4f

Observation 173d56fb-755b-424c-afd9-e481bf4dd932 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:23.904221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:23.904221Z digest=sha256:173f4c8f4ec926b3e904fd470ef9dc89fb4ed679635f0185f1dd4687839fc46a

Observation c0ea6155-5476-498b-8c74-6e94153e5854 · outbound

This paper cites Claude 3.7 sonnet.https://www.anthropic.com, 2025.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Claude 3.7 sonnet.https://www.anthropic.com, 2025

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:24.119707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:24.119707Z digest=sha256:69019c7ba051065ebce4b7e9737473793c3b5bf1e4125e78452c5ca336311800

Observation ce446c75-0f10-49c5-b742-d489d7fe1c7f · outbound

This paper cites Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:24.259442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:24.259442Z digest=sha256:e1a1963eb7355141075759b086c063678af2163a831fc29e6eaea3c1b448b11d

Observation b27e1f69-2c10-42bc-8d0f-79df976f9b78 · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:24.347294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:24.347294Z digest=sha256:9d97ae89d02142bd6ef74585d35c8c17390baf3c44608fe41944a29cbf61a000

Observation e4edfa49-1e15-4af3-9132-b1e55c692976 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:24.455888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:24.455888Z digest=sha256:da4d7bd53a8f65ccd71ac73e6279cf0b04ee04246e3262716895e3b6f7c5019f

Observation c2acb3f3-628c-4a4d-aec6-4e9aa2b459a4 · outbound

This paper cites Internvl 2.5: Scaling up vision-language models with enhanced visual encoding.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Internvl 2.5: Scaling up vision-language models with enhanced visual encoding

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:24.529399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:24.529399Z digest=sha256:26da2cd8b3a05ebff8aefc18e395cdd2a965c453430d1e7c03ce76fde90998be

Observation ba836203-b390-4abe-9c4b-65f75ad38898 · outbound

This paper cites Intern-s1-mini.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Intern-s1-mini

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:24.602875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:24.602875Z digest=sha256:eea4b38d29c5c9d5c98cd87c0826e36819c0157d286cba4b70910527a37fdce7

Observation 04314910-d95d-46a0-ba5a-f4510459b4a7 · outbound

This paper cites Scaling vision pre-training to 4k resolution, 2025.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Scaling vision pre-training to 4k resolution, 2025

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:24.712561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:24.712561Z digest=sha256:933061314fe5b211480f1a6624d018d382209c38487169ac6e01eede7c35b6a1

Observation 89cc5eed-8ea0-411f-b726-447b9a64a00f · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:24.852448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:24.852448Z digest=sha256:3ab692c4c07a277014a07971d191244869b3bf86466152bc8852b40c0942f542

Observation 65d12c25-be27-4cbb-a7bf-f55fddb49ef5 · outbound

This paper cites InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:24.958431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:24.958431Z digest=sha256:9459c04d21a02b5be1a7d68263584fb58795f20618a44e2cbda3c3c21b3c5c8d

Observation 4cdacec1-aaf1-4bdd-ae93-d183aa36a112 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:25.064532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:25.064532Z digest=sha256:9a34f147c2a7354330cf510be44d5c6086a6e2e34769b9ed2279bc306197b4ea

Observation 131db01f-fad7-4211-b698-e3a62726865b · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:25.168541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:25.168541Z digest=sha256:dcab2bd0d135e212aa2b93a781b0173c46055d1c29452267b054ce645eb15b04

Observation fdd8507f-2f26-4c3f-84b7-6da8c8bacf4f · outbound

This paper cites Hello gpt-4o.https://openai.com/index/hello-gpt-4o, 2024.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Hello gpt-4o.https://openai.com/index/hello-gpt-4o, 2024

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:25.243527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:25.243527Z digest=sha256:7025d2e33d6fe7844e202d81d77761d58e9790c895242f7a2a799e76e48b2003

Observation d090fcbd-c8b2-458d-9669-43a09a9333b1 · outbound

This paper cites Gpt-4o mini: advancing cost-efficient intelligence.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Gpt-4o mini: advancing cost-efficient intelligence

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:25.345011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:25.345011Z digest=sha256:b977a87addac5fe8b3799a61b4101a53e59db3652481e2a02d847d44366e1b35

Observation 75d0d3e2-ae73-4ecf-b933-363859c22b00 · outbound

This paper cites LLaVA-OneVision-1.5: Fully Open Framework for Democratized Multimodal Training.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing LLaVA-OneVision-1.5: Fully Open Framework for Democratized Multimodal Training

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:25.422944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:25.422944Z digest=sha256:b07162f37cd5fcf91cdfbde1e383cf4927c9fac5955a0f5b0432dfe6a74ddd64

Observation fe6a7bc3-b2ba-48e8-89bd-d1f06d7c25d6 · outbound

This paper cites Qwen3 Technical Report.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Qwen3 Technical Report

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:25.492470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:25.492470Z digest=sha256:b32bf7d8cab3f958238b843eee386cae011cf5a080dd4769d996e7752221c586

Observation 3ade99b4-5603-491e-b1ab-50aab15b220c · outbound

This paper cites InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:25.571086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:25.571086Z digest=sha256:eaeb6ac679f85b965d601dfb7669cfa962fc3b71f6dcf349b5f195200e3cfbb3

Observation cffc18e7-6fd1-4e01-9dd5-0225277e5e2b · outbound

This paper cites Internlm2 technical report.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Internlm2 technical report

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:25.630835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:25.630835Z digest=sha256:e426ac35f3d9697521d2a76601e46d6eb972e59a9a874b0a6f5a5a451331702d

Observation f45abd77-c0e3-48b5-aa2a-f4fe01a12570 · outbound

This paper cites Internlm3.https://github.com/InternLM, 2024.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Internlm3.https://github.com/InternLM, 2024

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:25.734409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:25.734409Z digest=sha256:9f5b4edcc387e0c51438996c9451f593d4fc2c44dfc7e98a7c00b76dae75067d

Observation 5d6e9cef-d45f-4363-b862-c74f629bddba · outbound

This paper cites Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:25.930022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:25.930022Z digest=sha256:bf4453da6cd5162e2b26f5960a05109d2002f877bced4fd977eccd4284ac426f

Observation b360c749-1614-45a1-8676-f81404cd65e4 · outbound

This paper cites VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:25.996486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:25.996486Z digest=sha256:80f8468cbb58c50bd5440a6af54c50ef200b7fab804be0154c445768f1564347

Observation eaae9c14-519a-480c-a673-644432dd9cd2 · outbound

This paper cites Fair1m: A benchmark dataset for fine-grained object recognition in high-resolution remote sensing imagery.ISPRS Journal of Photogrammetry and Remote Sensing, 184:116–130, 2022.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Fair1m: A benchmark dataset for fine-grained object recognition in high-resolution remote sensing imagery.ISPRS Journal of Photogrammetry and Remote Sensing, 184:116–130, 2022

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:26.104529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:26.104529Z digest=sha256:4aed1badc843e6c7a36aae772938d5e5f1538c02545d9cd94e5c351b7e565f73

Observation 2567ef06-e15a-40c5-b4c0-98febc8c1b2c · outbound

This paper cites an unresolved cited work.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing Unresolved cited work

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:26.218778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:26.218778Z digest=sha256:8566723216f1939fc56b0d3d0d1550287df4c17750eaae37bafb7df51a76983f

Pith citing papers

No inbound Pith citation observations are available.