Pith. sign in

Paper Citation Record · LEDGER

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios

As of 11 August 2026, this Paper Citation Record lists 70 of 70 outbound references and 4 inbound Pith citation observations for arXiv:2512.24561.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2512.24561 v2

Coverage vector

measured 70 of 70 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T13:22:13.267737Z

measured 74 of 74 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T15:01:03.556212Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-13T23:18:25.494068Z

Reference resolution

70 of 70 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved69
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1407b649-e64b-40d9-8988-2a1232010af3 · outbound

This paper cites Qwen-vl: A versatile vision-language model for un- derstanding, localization, text reading, and beyond, 2023.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Qwen-vl: A versatile vision-language model for un- derstanding, localization, text reading, and beyond, 2023

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:07.475422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:07.475422Z digest=sha256:69e90a7e933e8090d0fed4917d2de9df3c1881064dd1fc7f3ece18f8514161e1

Observation bb100858-17d2-40bd-8059-6adb8e7479c1 · outbound

This paper cites Uniter: Universal image-text representation learning.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Uniter: Universal image-text representation learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:07.532784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:07.532784Z digest=sha256:98496f837703305de9bffe14a67ed570a1ecb4fbe40bc06ccadad8de1c38d392

Observation 1ee842d9-9938-44b3-8875-9f63b5e079d2 · outbound

This paper cites Cops-ref: A new dataset and task on composi- tional referring expression comprehension.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Cops-ref: A new dataset and task on composi- tional referring expression comprehension

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:07.598268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:07.598268Z digest=sha256:03555adf06eb0f56b618a6feba9b3eb52d8cc210053dcb405d93a1ba0c8c3417

Observation 4f01e7a2-af88-41f5-81fe-71e9cd1c16a8 · outbound

This paper cites Unit3d: A unified transformer for 3d dense captioning and visual grounding.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unit3d: A unified transformer for 3d dense captioning and visual grounding

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:07.673878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:07.673878Z digest=sha256:18bc1e6bd20f936e3f70fb8738b8e64b659c39df59a66dc19bee7c20837a4284

Observation 7e262af5-d149-42b1-92a0-23ca710517fe · outbound

This paper cites Advancing visual grounding with scene knowl- edge: Benchmark and method.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Advancing visual grounding with scene knowl- edge: Benchmark and method

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:07.741282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:07.741282Z digest=sha256:6d187fa9f9404001c303e025e1900c078f4cc41bfdaefe22e3a581ba7043f8d2

Observation c45022d3-1c20-4f82-9414-230b44a1a92e · outbound

This paper cites Transvg: End-to-end visual ground- ing with transformers.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Transvg: End-to-end visual ground- ing with transformers

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:07.814739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:07.814739Z digest=sha256:e8616a796018a974f5559dbeb6699b438422b55177a6840591257574f42c5f98

Observation 309e8538-795e-47ea-a566-b1d1f7993058 · outbound

This paper cites Bert: Pre-training of deep bidirectional trans- formers for language understanding.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Bert: Pre-training of deep bidirectional trans- formers for language understanding

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:07.875627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:07.875627Z digest=sha256:83768ac108f9952e6d0cd6e73537acd9f92bb50081c8f64e6a8435397e70e458

Observation 49a85298-a2b6-4d7a-99f7-a5e211ef41ab · outbound

This paper cites D3t: Distinctive dual-domain teacher zigzagging across rgb- thermal gap for domain-adaptive object detection.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios D3t: Distinctive dual-domain teacher zigzagging across rgb- thermal gap for domain-adaptive object detection

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:07.982432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:07.982432Z digest=sha256:0e9740226ff0e21328d5e23b074f53204670f0f122382acb5a68513dd5dec3cd

Observation 0af4a103-7809-46c1-9dc7-8b4af125131b · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:08.081744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:08.081744Z digest=sha256:f07ae236d67836d4d1404e14d6da5cc396789ec13d4da09f14195d7db742b64d

Observation bdd8dfea-ecb2-4268-8898-5ba49f84d6a4 · outbound

This paper cites Large-scale adversarial training for vision- and-language representation learning.NeurIPS, 33:6616– 6628, 2020.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Large-scale adversarial training for vision- and-language representation learning.NeurIPS, 33:6616– 6628, 2020

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:08.183519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:08.183519Z digest=sha256:d468ba14942fae63073ae1901d782d2df0145904099f7dc187d4e2842bce8884

Observation fec1444d-f117-40fc-b782-b787c38ec76d · outbound

This paper cites Room-and-object aware knowledge reasoning for remote embodied referring expression.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Room-and-object aware knowledge reasoning for remote embodied referring expression

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:08.253760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:08.253760Z digest=sha256:8c6015e18f0cbcc40ba192b93a04ad5e8d36a1227d1e006ee7a6ecee0e624290

Observation b0f35396-7226-4bdd-bae3-aeb7dbf0eff2 · outbound

This paper cites The iapr tc-12 benchmark: A new eval- uation resource for visual information systems.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios The iapr tc-12 benchmark: A new eval- uation resource for visual information systems

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:08.349314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:08.349314Z digest=sha256:f6e8a5567ad660701c3889d1ab036a952be172b31930b6b3ca38d2d86d8d4db6

Observation cc6a1b54-228b-4f93-a57e-b5970a8e6004 · outbound

This paper cites Deep residual learning for image recognition.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Deep residual learning for image recognition

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:08.447263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:08.447263Z digest=sha256:5c1be2b4d207713fdf1c23b78533f5811511cc072213726a76ca992c6b902e51

Observation a4bddb92-2038-49c6-b0a5-b5ccd64e6ac0 · outbound

This paper cites Lora: Low-rank adaptation of large language models.ICLR, 1(2):3, 2022.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Lora: Low-rank adaptation of large language models.ICLR, 1(2):3, 2022

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:08.559249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:08.559249Z digest=sha256:810b679d1123fa403e4f3c0ebb3f08cf98016fbef3b2b05563caa3fff9290f90

Observation 7db8483a-6139-4464-94e5-d1b3dd79514a · outbound

This paper cites Ei 2 det: Edge-guided illumination-aware interactive learning for visible-infrared object detection.TCSVT, 2025.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Ei 2 det: Edge-guided illumination-aware interactive learning for visible-infrared object detection.TCSVT, 2025

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:08.727448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:08.727448Z digest=sha256:2146ba606169dfa1f90c70fad917fb829a11d7b3d5e691a7144d28fd262b4855

Observation d41f7c38-41e1-478f-9bea-ac274ef82114 · outbound

This paper cites Beyond one-to-one: Re- thinking the referring image segmentation.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Beyond one-to-one: Re- thinking the referring image segmentation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:08.836138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:08.836138Z digest=sha256:bdb13816594eccb144d3bcd2bf1ef25f64987dbd5b867f2bf011cb028fe8a05f

Observation 2e61ec8c-4d4e-468c-8cd5-b5de161dd86f · outbound

This paper cites Referitgame: Referring to objects in pho- tographs of natural scenes.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Referitgame: Referring to objects in pho- tographs of natural scenes

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:08.924987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:08.924987Z digest=sha256:db25ee3e18def42e39ad7165e47d96ce693540e97077471aba5b439bc7c2d750

Observation 4f621396-d7f2-4d0c-857b-b445360ee236 · outbound

This paper cites mplug: Effective and efficient vision-language learning by cross-modal skip-connections.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios mplug: Effective and efficient vision-language learning by cross-modal skip-connections

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:08.997393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:08.997393Z digest=sha256:986d37c0df9c6fd560e8f1550d7fbffb080f9018d0d468d9718c564aefa14926

Observation 567a5f88-d237-42e1-a838-7af3bed1a193 · outbound

This paper cites Rgb-t semantic segmentation with location, activation, and sharpening.TCSVT, 33(3):1223–1235, 2022.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Rgb-t semantic segmentation with location, activation, and sharpening.TCSVT, 33(3):1223–1235, 2022

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:09.187025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:09.187025Z digest=sha256:62b2821dd9fcd389c63defd8b52471dd85289afde1f3221f066dfc48d00f4dcf

Observation da724d24-8ca5-4d5b-af96-fad6aa94cef9 · outbound

This paper cites Microsoft coco: Common objects in context.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Microsoft coco: Common objects in context

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:09.301713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:09.301713Z digest=sha256:44dc29386f58f52c30e47ce0d200817912f7b7b12626056a7c14fbf1e4f6431a

Observation 0f038089-9ff8-44a7-93b0-98721e65a4cd · outbound

This paper cites Gres: Gen- eralized referring expression segmentation.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Gres: Gen- eralized referring expression segmentation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:09.356679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:09.356679Z digest=sha256:6af46c9d96c403545f541a8e10fa4bf8e247364f841b5db79668d090fdf7ea89

Observation 91129e57-d05f-4298-9e37-598bc2d60a90 · outbound

This paper cites Refer-it-in-rgbd: A bottom-up ap- proach for 3d visual grounding in rgbd images.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Refer-it-in-rgbd: A bottom-up ap- proach for 3d visual grounding in rgbd images

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:09.469324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:09.469324Z digest=sha256:c3271cbc4de5d6d65fd9e6779d9e1cfe913743609ae1d6891ed5c1a12aca0de8

Observation 805d9a1c-b4cc-4c9d-a0bb-5c60bdf049ce · outbound

This paper cites Visual instruction tuning.NeurIPS, 36:34892–34916, 2023.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Visual instruction tuning.NeurIPS, 36:34892–34916, 2023

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:09.627654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:09.627654Z digest=sha256:61f0be577f7ecf79326d276c8cce50ba1611376d9d61c2ae2638204f327f8d98

Observation ef9c6c3c-4cfd-49bc-acf8-1c69cd8fed7d · outbound

This paper cites Target-aware dual adversarial learning and a multi-scenario multi-modality benchmark to fuse infrared and visible for object detection.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Target-aware dual adversarial learning and a multi-scenario multi-modality benchmark to fuse infrared and visible for object detection

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:09.685054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:09.685054Z digest=sha256:14f35af50f09e08e08ade6900ae7815159aaa514a3365fb2e86b6abd9184f674

Observation 2edb6309-ac5c-4d98-8c91-1603d2351ad4 · outbound

This paper cites Cross3dvg: Cross-dataset 3d visual grounding on different rgb-d scans.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Cross3dvg: Cross-dataset 3d visual grounding on different rgb-d scans

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:09.797495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:09.797495Z digest=sha256:72879bd1859925e73cc124bed1ad2003a55affc4cfcc3e071b277b1b53f64875

Observation 959bbfc9-1512-42d2-9499-e1385b44ed0f · outbound

This paper cites Mod- eling context between objects for referring expression under- standing.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Mod- eling context between objects for referring expression under- standing

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:09.957036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:09.957036Z digest=sha256:6aaf31178b88efbd2ce9fbfe137e5ae11712aa9524e31d1c010bfbdc05e669fa

Observation a25a0121-4167-462c-b42c-b046a3549897 · outbound

This paper cites Flickr30k entities: Collecting region-to-phrase corre- spondences for richer image-to-sentence models.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Flickr30k entities: Collecting region-to-phrase corre- spondences for richer image-to-sentence models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.034272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.034272Z digest=sha256:5e7b651e12d54f5b44f34f563a971ebad62c5f0b587a4ec14eecf0c6f56a09f2

Observation 031b6ad1-fc67-4964-bacc-bea7a1ee9cf7 · outbound

This paper cites Learn- ing transferable visual models from natural language super- vision.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Learn- ing transferable visual models from natural language super- vision

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.056944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.056944Z digest=sha256:da70fdc2c420d3c4bbf519dd14b24917bc3a69e29efcda102504b3466e1f5f4d

Observation eb5116eb-af16-4e42-937b-d9932eec4ee4 · outbound

This paper cites Dynamic mdetr: A dynamic multimodal transformer decoder for visual grounding.IEEE Transactions on Pattern Analysis and Machine Intelligence, 46(2):1181–1198, 2023.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Dynamic mdetr: A dynamic multimodal transformer decoder for visual grounding.IEEE Transactions on Pattern Analysis and Machine Intelligence, 46(2):1181–1198, 2023

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.132196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.132196Z digest=sha256:97882b4091cd0073d915ad2c0eae2b071b352b4d445e219b2e97dc198c510b52

Observation 869f0c3c-cdca-453f-9644-63d37f1169f5 · outbound

This paper cites Attention is all you need.NeurIPS, 30, 2017.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Attention is all you need.NeurIPS, 30, 2017

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.165003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.165003Z digest=sha256:ef0d78350d9d733d6d6ddfad33f3990684c23722b584668b76ebf6e8737f0d07

Observation e8659ac2-f1dd-4ebf-a7c2-4e7b6a7a8794 · outbound

This paper cites Ofa: Unifying architectures, tasks, and modalities through a simple sequence-to-sequence learning framework.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Ofa: Unifying architectures, tasks, and modalities through a simple sequence-to-sequence learning framework

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.222723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.222723Z digest=sha256:ee3ee14bc9be96ba24c8fa4680e1d696533fc48f5252cab325a3fbb0a7ea8852

Observation aa35b809-d83d-43ca-a718-0ed735563433 · outbound

This paper cites Image as a foreign language: Beit pretraining for vision and vision- language tasks.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Image as a foreign language: Beit pretraining for vision and vision- language tasks

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.291114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.291114Z digest=sha256:e9e100d4f546c2146ea098345f297488ec88c32e45c44e31bea3e2e3677acd23

Observation a43b6454-1678-48d8-b9d4-e4e4a696c5cb · outbound

This paper cites Clip-vg: Self-paced curriculum adapting of clip for visual grounding.TMM, 26:4334–4347,.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Clip-vg: Self-paced curriculum adapting of clip for visual grounding.TMM, 26:4334–4347,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.359561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.359561Z digest=sha256:06de645d9ccf44eb910dab2f197f98e487cec2b5f46b1a4e7d781500e6551055

Observation ceb9017f-6716-48c9-bdeb-10dcb348b875 · outbound

This paper cites Towards visual grounding: A survey.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Towards visual grounding: A survey

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.534984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.534984Z digest=sha256:071721a0f428c6972694dcb14709788565cf45e18f62865bf692b07898bda393

Observation 9cbcbc35-a951-4641-b7ab-2c7b948aa364 · outbound

This paper cites Hivg: Hierarchical multimodal fine- grained modulation for visual grounding.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Hivg: Hierarchical multimodal fine- grained modulation for visual grounding

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.608296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.608296Z digest=sha256:4ecf13ba5c843dc6c5ec0a1af8bef514905b992cc683c565ac72294e5e73c26d

Observation 1a5d5a3c-e6fa-4de4-81e3-4fa0ff4433bf · outbound

This paper cites Oneref: Unified one-tower expression grounding and segmentation with mask referring modeling.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Oneref: Unified one-tower expression grounding and segmentation with mask referring modeling

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.690297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.690297Z digest=sha256:5030955f74b0f51f611b959b5dc40e322eeef2b6576ef635728f76c8c8307488

Observation 894c0ae2-e442-40e5-bf4b-3d0ed8e4c595 · outbound

This paper cites Described object detection: Liberating ob- ject detection with flexible expressions.NeurIPS, 36:79095– 79107, 2023.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Described object detection: Liberating ob- ject detection with flexible expressions.NeurIPS, 36:79095– 79107, 2023

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.751759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.751759Z digest=sha256:d0216c7db488ea77162c8bd09c27ff7603bc42a1bae4b748bc01171d44f3fabe

Observation 12bd4718-e362-4862-b9b4-8c237dc051b9 · outbound

This paper cites Mc-bench: A bench- mark for multi-context visual grounding in the era of mllms.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Mc-bench: A bench- mark for multi-context visual grounding in the era of mllms

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.859018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.859018Z digest=sha256:71e5000e09729031710edd66865bac4122f502f364e61c8c04475632b04f1b38

Observation a5317b8b-90d2-4892-949b-1340eea2232c · outbound

This paper cites Improving one-stage visual grounding by recursive sub-query construction.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Improving one-stage visual grounding by recursive sub-query construction

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.963077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.963077Z digest=sha256:29bbd280a8e8874d7e8d4719d94440b344da10ef9c2028158b674e4dd1e033ba

Observation 0e63b9eb-7172-4b40-80d8-6b4f06347bfc · outbound

This paper cites Unitab: Unifying text and box outputs for grounded vision- language modeling.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unitab: Unifying text and box outputs for grounded vision- language modeling

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.074241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.074241Z digest=sha256:c2d29d50789a03db7aba398820ee736592367ecd34203212ddbd40208dd10b5e

Observation 1cb76a83-6046-4b61-a18b-eb439b5d1d27 · outbound

This paper cites Vi- sual grounding with multi-modal conditional adaptation.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Vi- sual grounding with multi-modal conditional adaptation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.136609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.136609Z digest=sha256:c7ca53d2cbcb9c57a37d73e64c8055f465583b19155bbde68b1d46e251d31d3f

Observation 2c7be307-2a62-4605-a87a-d3324ed593d0 · outbound

This paper cites Shifting more attention to visual backbone: Query-modulated refinement networks for end-to-end visual grounding.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Shifting more attention to visual backbone: Query-modulated refinement networks for end-to-end visual grounding

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.208551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.208551Z digest=sha256:3ff9f2df703a43134ba1f60e25ba2c16e7c8e5a90a038d083c8467958ba338d5

Observation 183bb79e-93ef-4873-9853-bf4a36d29ada · outbound

This paper cites Modeling context in referring expres- sions.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Modeling context in referring expres- sions

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.270069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.270069Z digest=sha256:c5873281288f8b5e2efece658f894c0c93c06795d1e298ec29ea5a366f36ac2d

Observation ac0055fa-fcb5-4ab3-915c-2aedee084a49 · outbound

This paper cites C 2former: Calibrated and complementary transformer for rgb-infrared object de- tection.TGRS, 62:1–12, 2024.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios C 2former: Calibrated and complementary transformer for rgb-infrared object de- tection.TGRS, 62:1–12, 2024

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.310614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.310614Z digest=sha256:e25e5fe7a323323ee487a4db07b7cca039a76220371265ebbdd436bde8fa9a70

Observation f90310fa-9edf-45f6-8e1c-d43a87e061ba · outbound

This paper cites Trans- lation, scale and rotation: Cross-modal alignment meets rgb-infrared vehicle detection.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Trans- lation, scale and rotation: Cross-modal alignment meets rgb-infrared vehicle detection

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.392019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.392019Z digest=sha256:94ad6782e5ca73398321d09aa569bc5964bdd2e3eb500ab146af58d4afd351a4

Observation 019f9ec1-96dc-4537-91fa-742ced008088 · outbound

This paper cites Improving rgb-infrared object detection with cascade alignment-guided transformer.Information Fusion, 105:102246, 2024.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Improving rgb-infrared object detection with cascade alignment-guided transformer.Information Fusion, 105:102246, 2024

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.462234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.462234Z digest=sha256:6b033e36638b0aa6bcd98f6ff65a773d11199bde61fa916157ae6719be56d50a

Observation 7933f4bf-8206-414e-9d2f-76a2fef01119 · outbound

This paper cites Unirgb-ir: A unified frame- work for visible-infrared semantic tasks via adapter tuning.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unirgb-ir: A unified frame- work for visible-infrared semantic tasks via adapter tuning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.542644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.542644Z digest=sha256:aaa2e974cfb453b1f6327727f317553993b0f0a48cb9b41913de37f021581af3

Observation 5dc45d55-c7c5-452b-95b5-0ee8441e4abf · outbound

This paper cites Multispectral fusion for object detection with cyclic fuse-and-refine blocks.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Multispectral fusion for object detection with cyclic fuse-and-refine blocks

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.610798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.610798Z digest=sha256:4baa19103b72b45afebbd6b730505ca1daca4166748d673424345eab3c263499

Observation ee7be17c-9ea4-40a8-88c4-0ed0c8c9468e · outbound

This paper cites Abmdrnet: Adaptive-weighted bi-directional modality difference reduc- tion network for rgb-t semantic segmentation.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Abmdrnet: Adaptive-weighted bi-directional modality difference reduc- tion network for rgb-t semantic segmentation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.644623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.644623Z digest=sha256:4d03098dd871a70206542d7181dda12d4b4e27cf5d0f7ce92e9506999c518e56

Observation 81ffd4b3-5722-4dc6-a3e6-57e85c43e6d4 · outbound

This paper cites Removal then selection: A coarse-to-fine fusion perspective for rgb-infrared object detection.arXiv e-prints, pages arXiv–2401, 2024.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Removal then selection: A coarse-to-fine fusion perspective for rgb-infrared object detection.arXiv e-prints, pages arXiv–2401, 2024

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.726883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.726883Z digest=sha256:685bd6bebc4d8cadfb247f9a2602d86b46eea4d58a5b1b6e42ba880f6ccc491a

Observation e354113d-a2f4-4527-8820-a55cb53aab71 · outbound

This paper cites be- hind the truck.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios be- hind the truck

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.831105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.831105Z digest=sha256:589c2625feb4ded452313d4c9b0a6d9dd44d93332b21c0c98dceae43ca0f144f

Observation c1623795-31fe-4d43-b53b-225543485a47 · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.914135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.914135Z digest=sha256:fc1dafa1b84c72e6452e2b8fcc999b85ba63cad5d2896918e4dd22ed46d82b02

Observation edfd31a8-d30e-41d2-b6fc-ff48332fbe89 · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:12.001270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:12.001270Z digest=sha256:6c096e83cedc75884d16c7d3ccf57ab308abb1a33c2c2ee960aa1ed1eb0aa036

Observation c2fb9113-c4ab-4fc6-b8f0-796a559df62c · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:12.187388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:12.187388Z digest=sha256:7c106ee50e68e62644c9977f281836e2e571d6126f9bd232dd67ec3470fe919b

Observation f96a5e63-4701-43df-b6b6-4fd64bd69e3e · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:12.218174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:12.218174Z digest=sha256:0083015b8e4e82ca68b2adb6b9b424f26709ea6668905ac00d339dd018373f30

Observation f32a7b6f-b590-4184-867f-954a7f47a86c · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:12.304908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:12.304908Z digest=sha256:529229e06633cdfd56853874f2b254b00143c90dd934f8a4c38966dbb882dac9

Observation 0a7c6681-4264-499d-bf3a-75c1b9f15d4d · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:12.488625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:12.488625Z digest=sha256:f46febe04c87f65769b93c5847705d11534fa1ac0ce151235f7ff9a006327a9a

Observation 38d3df62-5325-4b4b-a79a-586a1f9b87c8 · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:12.615969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:12.615969Z digest=sha256:6e73ff33e5f69ed371c6f744a6232b4085716b422a6c0dc8d716f755b058d83c

Observation 8540625b-e574-43fd-ac62-002f82383d36 · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:12.686063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:12.686063Z digest=sha256:7567ee85afb40068e70db5101f92baf7f69ad0c8a642a97d5d15a775686ad351

Observation 400a225d-e54d-46b3-838d-7cec33fbae08 · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:12.756090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:12.756090Z digest=sha256:866244fcd29621cd855cc5a3dbd53738dd3088d87702941506f1c1bdcf53690d

Observation 7beeb4e8-13a5-4879-bb26-19bbae99ffcc · outbound

This paper cites near the stairs.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios near the stairs

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:12.820399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:12.820399Z digest=sha256:94d34c4c20dbafe898db4ce0f8ecb35318007fe6a41679e7a5d4b50548ac3555

Observation 386a503d-75d2-4be4-93ec-f822478cd406 · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:12.879990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:12.879990Z digest=sha256:ae131b33d3f084e7be5600b28142828209596edbcb3f7418a179e729c1c542eb

Observation 4919d6e8-abb9-4622-94ab-c7e6a8db808f · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:12.957929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:12.957929Z digest=sha256:a0ed836606fc0ff1b9b57f7be88ef13f63e75bedd860140f58b5c4fa5f5667f0

Observation bf3a12b4-79e7-4ca7-9f1a-43aa71c047d2 · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:12.979078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:12.979078Z digest=sha256:a9242d705372de33eb5df7ec2044adca440fd38beb9be8cf8ebdaa701f6fa273

Observation 9c91cae0-6661-4ba5-a87e-7c2c9c824b6b · outbound

This paper cites Please return only one number corresponding to the lighting condition: 0 (very_weak_light), 1 (weak_light), 2 (normal_light), or 3 ( strong_light).

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Please return only one number corresponding to the lighting condition: 0 (very_weak_light), 1 (weak_light), 2 (normal_light), or 3 ( strong_light)

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:13.007925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:13.007925Z digest=sha256:c4ff3d03e17794f90d9bf0e29476c7b2b681b95f272394cfd7ec3fd52e53e9eb

Observation f67d87e2-450c-474c-ae53-981ab63e3435 · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:13.051235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:13.051235Z digest=sha256:a795c2fb46d015861191bfc103313b064a28ac1bfc8c067da838c1baf6f9a3cb

Observation 6b6846cd-bf47-4eb4-8db7-40c503b7c1e3 · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:13.131462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:13.131462Z digest=sha256:0d227b0593d84d3d08c9bb3331fd64c3ba0ea23af07de4f1e1082970e2cc3abc

Observation 540c878a-3b0f-4679-b930-fbd5c2801678 · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:13.221032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:13.221032Z digest=sha256:c545a3fa09f1b9d7ab06a894f64557703310b97da0b2c3ca5c35e3729cf867b9

Observation 806563eb-1840-4fbb-8b28-9cef8c06c8b0 · outbound

This paper cites small"ifsize_ratio < 0.01else.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios small"ifsize_ratio < 0.01else

Reference 75

Resolution
malformed identifier
no resolver link, observed 2026-08-03T13:22:13.267737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:13.267737Z digest=sha256:c2ac5f9b2764ff5210fa4c46f6be59fa6e72b3c816ea3affc9ca41c067365825

Observation ef0166ab-e496-4d55-88d5-ec756ac1c30e · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.429051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.429051Z digest=sha256:79882438caa7d933ae54281f7b2ebaedebf5246b24371440ff8aac484565b6b1

Pith citing papers

Observation 627a73fa-a568-4ebf-95f0-97a79ed31635 · inbound

OmniSch: A Multimodal PCB Schematic Benchmark For Structured Diagram Visual Reasoning cites this paper.

OmniSch: A Multimodal PCB Schematic Benchmark For Structured Diagram Visual Reasoning RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-07-01T02:17:18.600387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-13T23:17:45.005518Z digest=sha256:399da9ac00046aaf42cd12ec11ae188d59090b70c2d2de72d9512ca63e173dcf

Observation 88a6f25c-1961-4add-97de-c2a9524ed012 · inbound

OmniSch: A Multimodal PCB Schematic Benchmark For Structured Diagram Visual Reasoning cites this paper.

OmniSch: A Multimodal PCB Schematic Benchmark For Structured Diagram Visual Reasoning RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-13T15:15:25.977993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T15:15:25.977993Z digest=sha256:536f843860fd890f6d9326bb065e528c64746dbb5d123a62441c14c01c750b02

Observation bc4d7ac6-4525-4f87-9402-b4fb9e1bb77a · inbound

SpectraDINO: Modality-Conditioned Adaptation of RGB Vision Foundation Models Across Infrared Bands cites this paper.

SpectraDINO: Modality-Conditioned Adaptation of RGB Vision Foundation Models Across Infrared Bands RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-07-01T02:17:18.600387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-09T16:10:23.516635Z digest=sha256:35af4e84f0b9b93ba3495ed665b9eddc5b8bbc9fdffd3f36b94d14c890b1c94a

Observation 5199272c-3e9f-4c74-9573-f11b653460ee · inbound

SpectraDINO: Modality-Conditioned Adaptation of RGB Vision Foundation Models Across Infrared Bands cites this paper.

SpectraDINO: Modality-Conditioned Adaptation of RGB Vision Foundation Models Across Infrared Bands RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-02T15:01:03.556212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T15:01:03.556212Z digest=sha256:a009db4ab61425dc589a2bbe33b104f6b53dc645b9209c3b1f3a9c40a96c1789