Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T18:38:19.066615Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 1 inbound Pith citation observation for arXiv:2411.11927.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T18:38:19.066615Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T04:20:50.578691Z
A source-named dated measurement, never combined with another source.
Source: cited_works
65 of 65 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 5fc2996f-d8e4-4fdc-8e1c-b26acbc91c16 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 771fe683-689f-441d-b6e9-682a5f800c57 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training GPT-4 Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 660777d5-cb9c-4b98-b151-3a197f179d4e · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Qwen Technical Report
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffdf892f-bd6e-4a73-abd7-c55e88b008fd · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Is a 3D-Tokenized LLM the Key to Reliable Autonomous Driving?
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c869d3f-597c-4e61-92cc-3786a74a9fb0 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Ar- trackv2: Prompting autoregressive tracker where to look and how to describe
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c4c56f93-3de8-484b-a1ad-86a340e98125 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training LLM2Vec: Large Language Models Are Secretly Powerful Text Encoders
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e25dad2-cdbc-4c67-a9af-08fa36885ddf · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Language models are few-shot learners
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0c0df006-4258-4b48-ba93-2aa5bcbf47b9 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Cross-lingual and multilingual clip
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f2176700-0419-4c48-b9f3-f506a023749c · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training ShareGPT4V: Improving Large Multi-Modal Models with Better Captions
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99a2709e-9a1a-4d99-890b-3cf78d420f12 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Pali: A jointly- scaled multilingual language-image model
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d1b03554-275a-471a-bd59-81be51e97053 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Altclip: Altering the language encoder in clip for extended language capabilities
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 818f101b-df01-47cb-9ef8-8d580dbea403 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Maskclip: Masked self-distillation advances contrastive language-image pretraining
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d7465e84-03ff-4c58-ba74-e7446b4560ab · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training An image is worth 16x16 words: Transform- ers for image recognition at scale
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 10bec2c1-eddb-469a-9391-df7cae57fefa · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Improving clip training with language rewrites
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 2dc147cf-df78-40ce-8f07-695493844f6b · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Pyramidclip: Hierarchical feature alignment for vision-language model pretraining
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f562151e-0523-4909-994a-422254de4be5 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Softclip: Softer cross-modal alignment makes clip stronger
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a44e35a2-0ca5-42d9-9767-00719897ee77 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Hiclip: Contrastive language-image pre- training with hierarchy-aware attention
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f8b962ac-6dc7-4488-a66d-56777c37f4a0 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Sugarcrepe: Fixing hack- able benchmarks for vision-language compositionality
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 8a850156-b50c-410f-a52f-a7e55f11e682 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Openclip, 2021
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation cd64749d-6852-4338-bceb-c524d5dfda4c · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Scaling up visual and vision-language representation learning with noisy text supervision
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 00cb0a3b-3899-409a-b0e7-2a1a893c0f8f · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Mistral 7B
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d033562-e4e6-40c2-97d8-eb4b9b788d4f · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Scaling Sentence Embeddings with Large Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9523231a-5fe3-433f-b323-a0f5e73eb07c · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Misalign, contrast then distill: Rethinking misalign- ments in language-image pre-training
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 03a6eaf3-1f33-47d8-b5fd-cbadd5602711 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Jina CLIP: Your CLIP Model Is Also Your Text Retriever
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df8046fa-aa08-4e92-9781-8bcb8859ac51 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training VeCLIP: Improving CLIP Training via Visual-enriched Captions
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e468e0e3-9f0b-4f11-9434-debbc4ba3d07 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Uni- clip: Unified framework for contrastive language-image pre- training
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e9c16874-4ff8-4c72-9d0b-036dbe818bdc · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Gecko: Versatile Text Embeddings Distilled from Large Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10ee5516-cff1-45b8-be52-193017adfe4b · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Meta-task prompting elicits embedding from large language models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 6eccc694-931a-4ab0-8990-fbe6fecc8037 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Scene graph generation: A comprehensive survey
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 16d083a4-7b19-4f7e-861f-e3aa071eb484 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Grounded language- image pre-training
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c291c074-5c93-4778-b7b3-73303cae886f · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training An inverse scaling law for clip training
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ad7acedf-ad08-4128-aa9b-576f699a806f · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Scene graph generation from objects, phrases and region captions
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a00fbd1f-04a6-4e30-bb3a-28cd27b1015b · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Supervi- sion exists everywhere: A data efficient contrastive language- image pre-training paradigm
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 27244eb8-92a5-4a14-9c51-e2409d99e880 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Scaling language-image pre-training via masking
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7358f3c8-a38c-4ecc-ac23-1f39083baa42 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Language quantized autoencoders: Towards unsupervised text-image alignment
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 07d857dd-9ef6-473b-aeff-8efc72d12d18 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training MLLMs-Augmented Visual-Language Representation Learning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb291ae4-5fcb-422f-bd0d-46559506bd91 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Decoupled Weight Decay Regularization
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0686173c-cdf0-4d4b-b1df-3615ec2ce85e · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Slip: Self-supervision meets language-image pre- training
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ecf32336-bbed-4254-bbdc-e17cb97b2bcf · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Generative Representational Instruction Tuning
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9fcf66c-1c73-49af-a6c7-910c486ad503 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Docci: De- scriptions of connected and contrasting images
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 37cbdc5f-e82c-4926-8245-1520109e40ca · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Learning transferable visual models from natural language supervision
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5d1b15dd-82f4-4705-b955-ae8943f16fcc · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training High-resolution image synthesis with latent diffusion models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00ea4c14-7c8a-47b0-b0eb-3be3273f8e55 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Conceptual captions: A cleaned, hypernymed, image alt-text dataset for automatic image captioning
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b94bf483-407b-4f26-b040-3338f133aeb3 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Repetition Improves Language Model Embeddings
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b1a5244-7d1c-4258-b856-f9a6ca575b3c · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training EVA-CLIP: Improved Training Techniques for CLIP at Scale
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64e4cd87-453e-4e07-aac4-042fc769be0f · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Crossmodal-3600: A massively multilingual multi- modal evaluation dataset
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a603a36f-318c-49e6-9fb0-e34cf50d15a4 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Winoground: Probing vision and language models for visio- linguistic compositionality
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 49222733-a491-46c9-accc-61f176f490a4 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Stablerep: Synthetic images from text-to- image models make strong visual representation learners
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 3d212d76-ad4e-4472-a9a6-86c71f9e7b44 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 181ea0ff-5cca-4aec-aaee-a07ffa20816e · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Multimodal few-shot learning with frozen language models
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ee888382-854d-4e60-8872-239e223ef6e8 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training NLLB-CLIP -- train performant multilingual image retrieval model on a budget
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b07fbd1e-0b69-4a54-bc5e-52af05dbb357 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Improving Text Embeddings with Large Language Models
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef407a48-25e3-4da7-8ac7-cbc02f5c9153 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Groupvit: Semantic segmentation emerges from text supervision
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5ba97f96-d30f-4d48-8a22-a1c68accd075 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d6bbb8b-f28e-4df6-b230-c09b5f4bffa4 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Alip: Adaptive language-image pre-training with synthetic caption
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 08a0c0fb-f71e-41cd-9394-4ac6a690eb78 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training FILIP: Fine-grained Interactive Language-Image Pre-Training
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9688f0f-3c42-4e1c-8b41-f20410f578c1 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training CoCa: Contrastive Captioners are Image-Text Foundation Models
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6700abaa-d14e-4096-b8d4-80892ca4f79f · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Spae: Semantic pyramid autoencoder for multimodal generation with frozen llms
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 46fbb959-e44a-465c-a66f-4232af836c81 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Lit: Zero-shot transfer with locked-image text tuning
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 24684c16-7b27-4655-9deb-003400c7547c · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Sigmoid loss for language image pre-training
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0a8f4552-a5fc-4809-bdf8-9683dbfab0d1 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Simple techniques for enhancing sentence embeddings in generative language models
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 688377fb-d2d8-4516-b32b-70dc64e34dc9 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Long-clip: Unlocking the long-text capability of clip
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d5678b42-79b4-4865-aee6-2f0a69950a5f · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Dreamlip: Language- image pre-training with long captions
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 31db733f-781e-4cc3-a199-1dcc8e19c9c8 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Beyond text: Frozen large language models in visual signal comprehension
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation bbb638db-a5df-4c67-a476-1f77e401f0d7 · outbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training yi”. After thinking step by step, the category of the main object in this image means in just one word:
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation dccbc202-eb71-431c-a90a-e929ada7b54e · inbound
Magic-MM-Embedding: Towards Visual-Token-Efficient Universal Multimodal Embedding with MLLMs FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.