Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-10T05:50:40.056376Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2604.18134.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-10T05:50:40.056376Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
41 of 41 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation b7f0d21a-f965-4b5e-b9e3-83b9bf4d7c4c · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning.arXiv
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation eb9d6db4-31ee-4116-87ee-1b18518ada9f · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? EndoViT: pretraining vision transformers on a large collection of endoscopic images.International Journal of Computer Assisted Radiology and Surgery, 19(6): 1085–1091
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0eb4b824-f695-4f34-be9c-f24ead5cf0cc · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Unsupervised learning of visual features by contrasting cluster assignments
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation afc4ee06-ada0-4ec5-82f8-5d83b7684252 · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Emerg- ing Properties in Self-Supervised Vision Transformers
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation cf18a84f-5590-4f52-ad08-2f2e55c73f8f · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Garcia-Peraza-Herrera
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 98f648b4-9f88-438a-afb1-1d12c9059e19 · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Garcia-Peraza-Herrera
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1913f95a-34ef-472a-98ed-548408ff9818 · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? A simple framework for contrastive learning of visual representations
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation cdea6b76-9097-40cc-9655-f6789d4baf5c · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Improved Baselines with Momentum Contrastive Learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7b8ec03c-6238-40ad-9d15-d0d170a67c93 · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Unsupervised Hyper- spectral Image Super-Resolution via Self-Supervised Modal- ity Decoupling.International Journal of Computer Vision
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 10e846fa-174d-495c-9980-26a6a734eb98 · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Med-CMR: A Fine-Grained Benchmark Integrating Visual Evidence and Clinical Logic for Medical Complex Multi- modal Reasoning.arXiv
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation de9d3590-c5ef-445b-b9d0-d58eb1df448b · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Domain-Specific Language Model Pre- training for Biomedical Natural Language Processing.ACM Trans
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ca5d9f0e-510f-411a-91a9-ab1a8c83b80e · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Masked Autoencoders Are Scalable Vision Learners
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5f389e16-7219-4158-9638-4b73abb90d44 · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? LoRA: Low-Rank Adaptation of Large Language Models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 75433114-e759-4f05-8e9e-022c259bea0f · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Jaspers, Ronald L.P.D
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 541c6cd0-6380-45a4-a418-d620049f2a6c · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Survey of hallucination in natural language generation.ACM Comput
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1709b0b8-b947-4de1-a619-c03d21ed31df · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Scaling Up Visual and Vision-Language Representa- tion Learning With Noisy Text Supervision
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 877cf544-f528-4deb-969a-afb16d7894fe · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Align before Fuse: Vision and Language Representation Learning with Momentum Distillation
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e91a39ef-05a4-4d88-961c-4c261e2c9181 · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? BLIP: Bootstrapping Language-Image Pre-training for Uni- fied Vision-Language Understanding and Generation
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1ef1d810-821d-4cae-9a1e-cffa32104835 · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Meireles, Guy Rosman, Maria S
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7d5f0969-218d-4291-aee5-2e24498482fb · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? A systematic review of annotation for surgical pro- cess model analysis in minimally invasive surgery based on video.Surgical Endoscopy, 37(6):4298–4314
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b0186eec-df22-4b63-a242-42dfac215d3b · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Sur- gLaVi: Large-scale hierarchical dataset for surgical vision- language representation learning.Medical Image Analysis, 110:103982
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9db65e0d-a197-4d60-aaa2-32acba388688 · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Learning Transferable Visual Models From Natural Language Supervision
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7d43969e-6a10-416c-87aa-771ded0e530d · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? LAION-5B: an open large-scale dataset for training next generation image-text models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 603281e6-f751-49f3-a645-cea8b99297b8 · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Conceptual Captions: A Cleaned, Hypernymed, Im- age Alt-text Dataset For Automatic Image Captioning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 319c6820-bd52-47ca-83bd-a1f29179d3b3 · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Transnet v2: An effective deep network architecture for fast shot transition detection
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 36fa8543-bfbc-4e94-99d1-fa2f4b472c73 · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Regionaligner: Bridging ego-exo views for object correspondence via unified text- visual learning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 4068f84f-b9fc-4e44-88fd-4cf6e9ced7f3 · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? MedGRPO: Multi-Task Reinforcement Learning for Heterogeneous Medical Video Understanding
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 59998f53-caed-4f4c-b7be-3a9d1b319120 · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Parwani, and Muhammad Khalid Khan Niazi
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 325258b3-530d-4aa5-9970-d94ae1cc2704 · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Gemini: A Family of Highly Capable Multimodal Models
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 24c125e9-5c23-4de8-8e5e-81558f2b61e6 · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Twinanda, Sherif Shehata, Didier Mutter, Jacques Marescaux, Michel de Mathelin, and Nicolas Padoy
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 829e8f0e-8067-4477-a914-25a1a60b455b · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Representation Learning with Contrastive Predictive Coding
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 660ab2d5-d2f8-4f9e-ba7c-e6735183a1f4 · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? VideoMAE V2: Scaling Video Masked Autoencoders with Dual Masking
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation eea168c9-3f95-48e0-97d6-6d33d8b2000a · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? AutoLaparo: A New Dataset of Integrated Multi-tasks for Image-guided Sur- gical Automation in Laparoscopic Hysterectomy
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8939082c-6987-4233-9051-eb36b0592c0f · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Foun- dation Model for Endoscopy Video Analysis via Large-Scale Self-supervised Pre-train
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0825d2ff-0285-4f72-99d7-c4977d647622 · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Challenges in surgical video annotation.Computer Assisted Surgery, 26 (1):58–68
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1da81fec-5a34-4430-9f55-129142904ef9 · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? SpatiaLQA: A Benchmark for Evaluating Spatial Logical Reasoning in Vision-Language Models.arXiv
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 540dac1e-036a-4dff-8805-34781650609b · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Lavanchy, Jacques Marescaux, Pietro Mascagni, Nassir Navab, and Nicolas Padoy
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 4a04d59e-b710-4640-b6d5-c4f35d4e4ced · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Chain-of-Thought Compression Should Not Be Blind: V-Skip for Efficient Multimodal Reasoning via Dual-Path Anchoring.arXiv
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 56f9caa7-f108-49cd-929f-e279ab4fd2c3 · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Not All Errors Are Created Equal: ASCoT Addresses Late-Stage Fragility in Efficient LLM Reasoning.arXiv
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5f773138-3c4f-4f06-b07b-8a3d22f8aa1b · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? iBOT: Image BERT Pre- Training with Online Tokenizer.International Conference on Learning Representations (ICLR)
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 496ddb58-0d99-4c9a-8033-29bdae276a34 · outbound
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training? Can we trust AI doctors? a survey of medical hallucination in large language and large vision- language models
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
No inbound Pith citation observations are available.