Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 36 inbound Pith citation observations for arXiv:2211.01335.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T11:54:43.528354Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T17:18:43.901557Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 454367fc-f584-4108-83db-4f056ff99b3c · inbound
InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 163
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 3a5c388e-90f1-4a6b-8046-aee4026ec6ed · inbound
Transmission Line Defect Detection Based on UAV Patrol Images and Vision-language Pretraining Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ba97f96-d30f-4d48-8a22-a1c68accd075 · inbound
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e6af795-fd8d-4a68-98da-5fce212198a5 · inbound
Text-Video Multi-Grained Integration for Video Moment Montage Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b656fe77-a02f-4db8-9d53-252efc70b473 · inbound
From 2D CAD Drawings to 3D Parametric Models: A Vision-Language Approach Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e4a0350-2dad-43fd-aa68-10c8dfa1f7db · inbound
Controllable Satellite-to-Street-View Synthesis with Precise Pose Alignment and Zero-Shot Environmental Control Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af9a374b-3e77-49e9-8fbb-38a8765a354c · inbound
HarmonyCut: Supporting Creative Chinese Paper-cutting Design with Form and Connotation Harmony Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 100
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49c3e5ec-8fcb-444f-808a-65b31d513bdd · inbound
Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 279
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation ea8302fd-67df-4bee-8e33-90af132e764c · inbound
Multimodal Coreference Resolution for Chinese Social Media Dialogues: Dataset and Benchmark Approach Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9013c745-1044-4219-b1c2-e7133c872756 · inbound
VLM as Policy: Common-Law Content Moderation Framework for Short Video Platform Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24299011-d4e0-4667-84dc-87a2483404eb · inbound
SeriesBench: A Benchmark for Narrative-Driven Drama Series Understanding Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 810a1e1e-88ba-404a-8d59-e2005a2c0637 · inbound
Towards Cross-modal Retrieval in Chinese Cultural Heritage Documents: Dataset and Solution Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1967491-8899-4f2d-85a9-d16e43bfea0e · inbound
Beyond Cropped Regions: New Benchmark and Corresponding Baseline for Chinese Scene Text Retrieval in Diverse Layouts Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a473fc8a-db33-45b0-9760-cc662ae9c057 · inbound
FORGE: Forming Semantic Identifiers for Generative Retrieval in Industrial Datasets Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1be1cad4-722c-4c1d-acf1-a614c9c91e91 · inbound
FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2f3e6d3-f9ab-4560-bdf4-bc1cbad7a8e4 · inbound
Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 9313398e-7389-40e6-9162-02bc5618da1d · inbound
Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa561e56-aabf-4b11-825c-c256c14093d0 · inbound
Disentangling Fact from Sentiment: A Dynamic Conflict-Consensus Framework for Multimodal Fake News Detection Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 36fda9db-7c75-4e95-a259-dc1a30d28b6f · inbound
JARVIS: An Evidence-Grounded Retrieval System for Interpretable Deceptive Reviews Adjudication Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 14ec2737-114e-4706-9913-3eb083145973 · inbound
Maximizing Mutual Information Between Prompt and Response Improves LLM Performance With No Additional Data Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6002349a-9720-4901-b60c-02f1d9a9b57f · inbound
DRG-Font: Dynamic Reference-Guided Few-shot Font Generation via Contrastive Style-Content Disentanglement Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation aaf75cb8-8dac-4480-ae00-d935c580e1fb · inbound
Text-Guided Visual Representation Learning for Robust Multimodal E-Commerce Recommendation Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 92ed66d7-e17b-4c95-90da-fd86a351e428 · inbound
TIGER-FG: Text-Guided Implicit Fine-Grained Grounding for E-commerce Retrieval Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation f6f026c4-33d4-4b6c-984c-94d026951ae6 · inbound
MindAlign: Bridging EEG, Vision, and Language for Zero-Shot Visual Decoding Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 556cfd5d-0b4a-4bf6-92c8-72a3b1e2a459 · inbound
UniNote: A Unified Embedding Model for Multimodal Representation and Ranking Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation b05e1a17-c547-46a2-9f51-c82999cb66b6 · inbound
MyoSem: Aligning Electromyography to Natural-Language Action Semantics for Hand Action Understanding Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 2f3030fa-5072-48f1-a2d1-6265c12946d6 · inbound
Fine-grained Fragment Retrieval in Multi-modal Long-form Dialogues Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 558e0d6c-e320-4ea9-8f42-4e557a03d566 · inbound
Qwen-RobotWorld Technical Report: Unifying Embodied World Modeling through Language-Conditioned Video Generation Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 143
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 2888c23e-cc9c-4cee-a36c-e1716aece263 · inbound
JuZhou 1.0 Technical Report: The First Edge-Native Text-to-Image Foundation Model Trained Entirely on China-Developed AI Accelerators Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 8e187e19-ec37-44f5-819f-2758a2a1621a · inbound
JuZhou 1.0 Technical Report: The First Edge-Native Text-to-Image Foundation Model Trained Entirely on China-Developed AI Accelerators Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e24fe46-9617-4181-b2e5-3317bb8beb13 · inbound
BamiBERT: A New BERT-based Language Model for Vietnamese Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 101c3f87-480f-4fd5-bad1-0d7c46ca2163 · inbound
Qwen-Audio-VAE Technical Report Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 158
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98c1a34a-e216-4b65-a4c7-475005d12a3a · inbound
RecGPT-V3 Technical Report Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc146426-2b52-4402-b2f1-41c5f4ee38dc · inbound
CHaystack: Benchmarking Chinese Document Retrieval and VQA Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01b51685-1e22-4bc2-9f44-da41d03d2333 · inbound
GALA: Generative Aligned Learning for Adaptive Multimodal Representation in the Taobao Shangou Recommender System Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87e9efb2-e9f0-4622-8961-558d1fdb1567 · inbound
Illuminating Visual Identity in Universal Multimodal Embeddings Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.