Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T17:30:01.202763Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 68 of 68 outbound references and 0 inbound Pith citation observations for arXiv:2508.12108.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T17:30:01.202763Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
68 of 68 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f4146d61-2732-4516-b624-c2b7aeb8fb25 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Gloria: A multimodal global- local representation learning framework for label-efficient medical image recognition
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 8e0f9474-ea4f-4127-9d48-13630f15a833 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Medclip: Contrastive learning from unpaired medical images and text
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 1ca8d35d-1b76-42a3-a122-00b868588b1d · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine MedUnifier: Unifying Vision-and-Language Pre-training on Medical Data with Vision Generation Task using Discrete Visual Representations
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aeb99c0f-ea84-4b85-8a41-045a9e128482 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Towards unifying medical vision-and-language pre-training via soft prompts
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0da28f04-e7b2-43ac-8fe4-e3b7a886af18 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Learning transferable visual models from natural language supervision
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31fd3ddf-0019-47de-87cc-ccc666725f89 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Conditional prompt learning for vision- language models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a82cf4d4-0d5e-44a8-a9cc-09e4c9d47e98 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Learning to prompt for vision-language models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d24152b-dd7f-4b0f-b4a2-dc688ba3e347 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Groupvit: Semantic segmentation emerges from text supervision
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 31fd943f-85c2-41fd-80ff-adf890337742 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Learning to exploit temporal structure for biomedical vision-language processing
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 1b345788-b913-4dd1-9761-227c17a82168 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Cplip: zero-shot learning for histopathology with comprehensive vision-language alignment
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 86beeb90-f59c-497a-bb3a-c9b04b201fe4 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Lu, Bowen Chen, Andrew Zhang, Drew F
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 68253beb-c63f-4592-be58-8e38d3c53ec3 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Mimic-cxr, a de-identified publicly available database of chest radiographs with free-text reports
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 75d4139c-ac4d-494b-8a81-927e794d3f02 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Quilt-1M: One Million Image-Text Pairs for Histopathology
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e1c04ed-6500-4676-b46e-c47e48b92be3 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Towards generalist foundation model for radiology
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 1c8df07b-4282-4af2-aa22-24f49dff84f6 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49a9153b-bea6-4ccc-a6a9-ce080367e416 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Align before fuse: Vision and language representation learning with momentum distillation
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 005f1bc3-deb3-41eb-92ac-a0cec3d855d0 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Bert: Pre-training of deep bidi- rectional transformers for language understanding
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 496cac27-372c-429e-8797-9669ab826acc · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Momentum contrast for unsupervised visual representation learning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68afe0c2-a45f-4924-881b-173de1dd5779 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine A simple framework for contrastive learning of visual representations
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37c0d002-2bd2-46df-85b9-868fba55a5fc · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Bootstrap your own latent-a new approach to self-supervised learning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fca479a-bce1-4cd0-9d94-f14a8b21dab4 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Masked autoencoders are scalable vision learners
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a61a804-ebb8-4aeb-ab57-d0578605cec3 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Generative pretraining from pixels
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06722dc4-a9ea-472f-beba-ce93f12056e7 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Unsupervised learning of visual representations by solving jigsaw puzzles
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63e30dd3-80c0-4541-b335-0e6e2a759c11 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Colorization as a proxy task for visual understanding
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 1b4a9d23-d3ff-4d1a-ab2a-89b4ac9b4460 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Self-supervised representation learning by rotation feature decoupling
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0090d0f4-6a32-498e-bf27-7f1752ca3a0b · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine What makes instance discrimination good for transfer learning?
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d4d30880-b7a7-407a-a716-b7c9dd9545f4 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Scale-mae: A scale-aware masked autoencoder for multiscale geospatial representation learning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d8a30164-549d-4257-b6ae-c5d0719a8c23 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Attention is all you need
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bff68324-f635-4ec9-9917-50600d642e81 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Language models are few-shot learners
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 742dac01-52cc-4341-9119-f9f86ba32dc5 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine LLaMA: Open and Efficient Foundation Language Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5967b54d-c824-42c1-b5bc-2245925714b9 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Med3D: Transfer Learning for 3D Medical Image Analysis
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91e4ed54-2f39-4b4a-9c5e-f58f477b6416 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Models genesis
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 724bb59d-0b2c-45c6-99ab-a656e72b9bd5 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bac93bcb-f821-48f2-8ff5-f2c549c8a74f · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine LXMERT: Learning Cross-Modality Encoder Representations from Transformers
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0d91a95-a7b6-46a2-ac37-2c3d84f12e06 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Uniter: Universal image-text representation learning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f6b30fb9-6e03-4c3a-add7-a7fcd66b4145 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Oscar: Object-semantics aligned pre-training for vision-language tasks
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a5ce9f45-656d-4090-a193-445a6a8a5b1b · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Vinvl: Revisiting visual representations in vision-language models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edff5fdb-e67d-4bf3-a270-1ef58d6df557 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Coca: Contrastive captioners are image-text foundation models, 2022
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55eeb42c-144f-4743-adf4-d5521a0efe40 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65a2707a-9740-4ef6-b6b8-a553c277d1ab · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6df0aa4f-a65e-41d4-a663-6f98aff101bf · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Multi-modal understanding and generation for medical images and text via vision-language pre-training
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ff93f537-8c21-40d3-a176-520255c67df5 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa7cdee5-57b6-4bea-8382-678291a191c2 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Slip: Self-supervision meets language- image pre-training
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 30c66838-3b1e-4226-9be9-6e0896c0c7c8 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Supervision exists everywhere: A data efficient contrastive language-image pre-training paradigm, 2022
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e0504519-2856-4003-a9bc-3a893792e187 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Swin transformer: Hierarchical vision transformer using shifted windows
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35a0962c-c2a3-47dc-a28a-8495e21b17fd · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Self-supervised pre-training of swin transformers for 3d medical image analysis
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d36c25de-c9bd-462e-9e61-ddc1f30d4e7f · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Masked image modeling advances 3d medical image analysis, 2022
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a1337afe-146a-4800-bfa3-7a63abe36566 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine V oco: A simple-yet-effective volume contrastive learning framework for 3d medical image analysis
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 39c3030c-6260-4ec2-bd81-ceaee3ac0f57 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Context Encoders: Feature Learning by Inpainting
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d126cba3-332c-4507-9895-8e13a2929756 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Unsupervised Representation Learning by Predicting Image Rotations
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18bc6f70-c1eb-4cf3-8c96-c2732c5712ac · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Representation Learning with Contrastive Predictive Coding
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f456fc9-a844-4b98-8b47-c3c06dd978f3 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Fast WordPiece Tokenization
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9a74340-58c9-4891-88f7-ce9118f5469c · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Zero-shot text-to-image generation
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 35db7e1a-a920-464b-accd-d190d1e28d70 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Swin unetr: Swin transformers for semantic segmentation of brain tumors in mri images
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5231ee45-4b53-4ccf-bf1c-777711b32a05 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Abdomenct-1k: Is abdominal organ segmentation a solved problem? IEEE Transactions on Pattern Analysis and Machine Intelligence, 44(10):6695–6714, 2022
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ee2307a5-8a02-497e-bdf9-dbeedff77eff · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Ct-org, a new dataset for multiple organ segmentation in computed tomography
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 517d1b0f-fc32-4797-9814-9daa32543141 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Ledsam, and Olaf Ronneberger
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0419b32a-6f0d-4f36-92ed-505400bf3365 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Bleu: a method for automatic evaluation of machine translation
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6787a5d8-402d-4bd3-92a5-6ae847698a9b · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Rouge: A package for automatic evaluation of summaries
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 759c3292-5139-4c5c-b49f-4d11bdce6f35 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine BERTScore: Evaluating Text Generation with BERT
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b728e4d2-5e7e-44e3-bf75-bf7ca624c657 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Decoupled weight decay regularization, 2019
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77ac4f29-0498-4704-82b8-6e3cb1cf4238 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine https://huggingface.co/ContactDoctor/Bio- Medical-Llama-3-8B, 2024
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0e4abb36-f73f-4b20-98c6-4184023dd95c · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Lora: Low-rank adaptation of large language models
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff94ef9a-2544-4cf9-8091-f84bc85c13bd · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Segment anything
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 807c9353-43ba-4bfa-8ccb-a1f472aea0fc · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Segvol: Universal and interactive volumetric medical image segmentation
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 70085536-0c4e-4680-b283-50c93e0ab6bd · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Segment anything in medical images
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d7497ca5-54bc-45f8-9024-6fce54bf1fea · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Pmc- clip: Contrastive language-image pre-training using biomedical documents
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f449df5a-836c-4ed5-8d9d-055db0c5d757 · outbound
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine Accelerate: Training and inference at scale made simple, efficient and adaptable
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
No inbound Pith citation observations are available.