Pith. sign in

Paper Citation Record · LEDGER

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs

As of 18 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 0 inbound Pith citation observations for arXiv:2607.27122.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.27122 v1

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-30T11:39:01.979348Z

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

25 of 25 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved22
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 013bc144-bd0a-4b2c-a645-2e6e5bd3f912 · outbound

This paper cites an unresolved cited work.

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs Unresolved cited work

Reference 1

Resolution
verified exact
doi, observed 2026-07-30T11:41:21.475467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-30T11:39:01.885452Z digest=sha256:6e84d4af80b4f2db484c2ad91960a972ffde295e00e4ece0f313b809e98ee235

Observation b2bc4ab2-af05-4ff2-a9cc-ad79a2684efb · outbound

This paper cites In: The Fourteenth International Conference on Learning Representations (ICLR) (2026).

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs In: The Fourteenth International Conference on Learning Representations (ICLR) (2026)

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-30T11:39:01.889906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T11:39:01.889906Z digest=sha256:89efa67a420f2e420e485021a97fe8609c69d44f347767b9320797becf48d70a

Observation 29cf6187-c670-4f9e-b93c-aa0290dcc1e0 · outbound

This paper cites In: MICCAI Workshop on Data Engineering in Medical Imaging, pp.

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs In: MICCAI Workshop on Data Engineering in Medical Imaging, pp

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-30T11:39:01.893597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T11:39:01.893597Z digest=sha256:475647288ddbed320957f744ca85cea340bc97de0b13abe772ef9cde6ce1c262

Observation f00068b0-3c79-45a7-8a1f-13214cad7a62 · outbound

This paper cites In: Proceedings of the IEEE/CVF Con- ference on Computer Vision and Pattern Recognition (CVPR), pp.

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs In: Proceedings of the IEEE/CVF Con- ference on Computer Vision and Pattern Recognition (CVPR), pp

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-30T11:39:01.897342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T11:39:01.897342Z digest=sha256:323f050603bc74ebe751a177edf5e64d8442ba8e19bb3ab9500db34de4775cba

Observation ffd8ed0f-949f-430d-8547-ad901a9b3f3d · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp.

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-30T11:39:01.901858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T11:39:01.901858Z digest=sha256:145522b0266f426dbffefadadda2053b56be0cb40a75f2f160015b80254e62d5

Observation 6c65abf5-c697-4713-831b-52fda506e540 · outbound

This paper cites In: Proceedings of the Computer Vision and Pattern Recognition Conference, pp.

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs In: Proceedings of the Computer Vision and Pattern Recognition Conference, pp

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-30T11:39:01.905290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T11:39:01.905290Z digest=sha256:d14cefca5a37f14419cb13b9066b70702091c70406d7413c0e42b1eecf179142

Observation f2e18f2b-9801-45c1-9e28-191e1689b652 · outbound

This paper cites arXiv:2511.04384 (2025).

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs arXiv:2511.04384 (2025)

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-30T11:39:01.909532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T11:39:01.909532Z digest=sha256:41df2d43b70ec9420cbc03b15ec0ddd79381b696d5adf05fa3f0ecd57219e4b5

Observation 5dadf50e-613c-42dd-b288-fab277d11804 · outbound

This paper cites Medico 2025: Visual Question Answering for Gastrointestinal Imaging.

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs Medico 2025: Visual Question Answering for Gastrointestinal Imaging

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-30T11:39:01.913748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T11:39:01.913748Z digest=sha256:25e1d0cbf7611d2044659b874afe3386ab685f1de283c81904fdf5851fc0f47d

Observation 24a72180-47cb-493f-8b87-d66eb0cd5fd6 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recog- nition (CVPR), pp.

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recog- nition (CVPR), pp

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-30T11:39:01.918507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T11:39:01.918507Z digest=sha256:98d12dc4ec119296e4b02fcf0a891cfe1e3fb959d579ea77bdfe07d43f3c61df

Observation a9b0048d-3cee-4f70-a6f8-ddaa1982ff0f · outbound

This paper cites Gastroen- terology170(1), 174–187 (2026).https://doi.org/10.1053/j.gastro.2025.07.

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs Gastroen- terology170(1), 174–187 (2026).https://doi.org/10.1053/j.gastro.2025.07

Reference 10

Resolution
verified exact
doi, observed 2026-07-30T11:41:21.377600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-30T11:39:01.922259Z digest=sha256:e27b7e36ed3ce5fd057f4941e4de0c605385c2d8ee9d84e7cd3eba16ae19d8e7

Observation 1a8f64eb-bbce-462a-92b2-e9d3f5cf03fe · outbound

This paper cites Gemma 3 Technical Report.

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs Gemma 3 Technical Report

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-30T11:39:01.926475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T11:39:01.926475Z digest=sha256:ee90ba443d3b32cef9cc01dc6ad7692f131d5d1b60df6e7a28d258c121996354

Observation 37611fd2-18f7-4457-8152-c71707bdb915 · outbound

This paper cites Advances in Neural Information Processing Systems 36, 10088–10115 (2023).

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs Advances in Neural Information Processing Systems 36, 10088–10115 (2023)

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-30T11:39:01.930329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T11:39:01.930329Z digest=sha256:9a8b831d6ead2204a9eb0d24ff33f97ca1b5321aed7aa78aff52391266bc0cb3

Observation b2661ba6-fb2e-4e89-8795-a8ac67162ba8 · outbound

This paper cites In: Proceedings of the AAAI Conference on Artificial Intelligence, vol.

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs In: Proceedings of the AAAI Conference on Artificial Intelligence, vol

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-30T11:39:01.933784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T11:39:01.933784Z digest=sha256:73987c1a9b517622cd64278b8c44a5e2502b6f53ad217ff0a58834b388d09a0a

Observation d3fffe8b-40ac-4541-bf95-05227866ead5 · outbound

This paper cites International Journal of Computer Vision128(2), 336–359 (2019).https://doi.

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs International Journal of Computer Vision128(2), 336–359 (2019).https://doi

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-30T11:39:01.937188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T11:39:01.937188Z digest=sha256:41d5560501fe6e8f0be0347a943870625045e0918571bbbc00004d66766f9f83

Observation 0aae1d75-68f5-406c-b278-eca81b723777 · outbound

This paper cites In: International Conference on Multimedia Modeling, pp.

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs In: International Conference on Multimedia Modeling, pp

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-30T11:39:01.940985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T11:39:01.940985Z digest=sha256:1c429a3894f5c126875342515b72d77b6f205256771ffcbe053b0fe515c60e30

Observation 86d17b14-e408-47ff-b413-0aca1c57f8e4 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp.

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-30T11:39:01.944328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T11:39:01.944328Z digest=sha256:aa9e0fb9dc9cf873b0646a58a7220414d274cbdc7611f9a51a73b23907bad3b4

Observation 509caa3c-e96c-44e5-878f-f8551dd95d34 · outbound

This paper cites an unresolved cited work.

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-30T11:39:01.947995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T11:39:01.947995Z digest=sha256:d30ec40c5e072f71d2443c4baf0d9b7d6e8d64e714c8d7bdaaa1e539bad81593

Observation e0c28141-0694-4c5d-a495-7e6785acc172 · outbound

This paper cites InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency.

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-30T11:39:01.951863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T11:39:01.951863Z digest=sha256:f6009a7e129d25c78a26d2452fa224d907ae007b8fa2e6b0314627522648a18b

Observation 7bef9fbf-507f-4d8c-8e0f-121feb599880 · outbound

This paper cites Pattern Recognition45(9), 3166–3182 (2012).https: //doi.org/10.1016/j.patcog.2012.03.002.

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs Pattern Recognition45(9), 3166–3182 (2012).https: //doi.org/10.1016/j.patcog.2012.03.002

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-30T11:39:01.956839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T11:39:01.956839Z digest=sha256:abd6075656144fc3d4717cf102dac70052b263f533a5d25c6cf06c009e3cfe6b

Observation c10b2203-fada-4b9f-874d-6a6672a1005f · outbound

This paper cites In: Proceedings of the 8th ACM on Multimedia Systems Conference (MMSys’17), pp.

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs In: Proceedings of the 8th ACM on Multimedia Systems Conference (MMSys’17), pp

Reference 20

Resolution
malformed identifier
no resolver link, observed 2026-07-30T11:39:01.961232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T11:39:01.961232Z digest=sha256:1ab5ea33498dabb83c23cad827e5e90e01e287d67b280824b9febf9a60537c19

Observation 4b1a8d4f-217e-4e03-adef-56ab96e873d8 · outbound

This paper cites In: MICCAI (2025, early accept, spotlight).

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs In: MICCAI (2025, early accept, spotlight)

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-30T11:39:01.964953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T11:39:01.964953Z digest=sha256:e99aabccce0a9fc9790b0f4d644c2de7d9d496508413e5b2e099a6245b67c663

Observation bdcc4662-a20d-4d0c-b2bf-782c00a6b8a3 · outbound

This paper cites an unresolved cited work.

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-30T11:39:01.968620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T11:39:01.968620Z digest=sha256:d665df115b29343f130a2e8f13a114297c88d2938e1fa61e228de8ebfdc6b397

Observation 933763f1-59c1-4332-a8cf-cd9b9d61e372 · outbound

This paper cites International Journal of Computer Vision 126(10), 1084–1102 (2018).

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs International Journal of Computer Vision 126(10), 1084–1102 (2018)

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-30T11:39:01.972222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T11:39:01.972222Z digest=sha256:730bdf59f0856b385b05a98c5da2d2053b239f96d8bd56281bfd024656efdebf

Observation f87d25f8-6aed-467a-8adb-308df7bf33c8 · outbound

This paper cites Medical Image Analysis, 103789 (2025).

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs Medical Image Analysis, 103789 (2025)

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-30T11:39:01.975877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T11:39:01.975877Z digest=sha256:736059db59f638f9a0f90489c768923d3ec78b8e15533cc0533ebaa6522d5018

Observation 4a238e02-9e01-4793-aaa8-43fe7f2c179d · outbound

This paper cites Surgical-VQLA: Transformer with Gated Vision-Language Embedding for Visual Question Localized-Answering in Robotic Surgery.

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs Surgical-VQLA: Transformer with Gated Vision-Language Embedding for Visual Question Localized-Answering in Robotic Surgery

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-30T11:39:01.979348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T11:39:01.979348Z digest=sha256:4f9c4aef89d054061b8d345824b09546b711f3160c3570369f4e987c0feb929d

Pith citing papers

No inbound Pith citation observations are available.