Pith. sign in

Paper Citation Record · LEDGER

CLIP-GEN: Language-Free Training of a Text-to-Image Generator with CLIP

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2203.00386.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2203.00386 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T18:51:47.472317Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-10T16:55:57.755875Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1f4e4dd4-4b6a-492f-9cee-e4a6c034e945 · inbound

Hierarchical Text-Conditional Image Generation with CLIP Latents cites this paper.

Hierarchical Text-Conditional Image Generation with CLIP Latents CLIP-GEN: Language-Free Training of a Text-to-Image Generator with CLIP

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-10T16:55:57.758432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T16:55:57.612364Z digest=sha256:63fb49107e164b33e0150a8430950b8454a67bf571c830fde921820a76e069c3

Observation feb0ec46-cbde-4eec-955d-825dc8e0ba31 · inbound

ITACLIP: Boosting Training-Free Semantic Segmentation with Image, Text, and Architectural Enhancements cites this paper.

ITACLIP: Boosting Training-Free Semantic Segmentation with Image, Text, and Architectural Enhancements CLIP-GEN: Language-Free Training of a Text-to-Image Generator with CLIP

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-12T18:02:15.171099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:02:15.171099Z digest=sha256:d41e3976f9e7021f2281c90ff19be0a7f40d2bd6f99818b0b144904e904c86ff

Observation c312f0ed-0886-4efd-bc38-b84dafb08972 · inbound

CLIP Unreasonable Potential in Single-Shot Face Recognition cites this paper.

CLIP Unreasonable Potential in Single-Shot Face Recognition CLIP-GEN: Language-Free Training of a Text-to-Image Generator with CLIP

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T17:44:06.184255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:44:06.184255Z digest=sha256:a4cf22705a0d3815e1223b936037da714167bb0d686b145df8f4945bb2af748d

Observation 7c4614a6-cdc5-4e16-86ca-d30e5eb20a41 · inbound

Combining Genre Classification and Harmonic-Percussive Features with Diffusion Models for Music-Video Generation cites this paper.

Combining Genre Classification and Harmonic-Percussive Features with Diffusion Models for Music-Video Generation CLIP-GEN: Language-Free Training of a Text-to-Image Generator with CLIP

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T20:29:54.320641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:29:54.320641Z digest=sha256:69c9fa7a550b9999c8c18214bac78d28af1961f3c4c6eea0db2ee10e0cef55c8

Observation 2c4eff88-30ba-44b7-a738-1980efce8508 · inbound

BudgetFusion: Perceptually-Guided Adaptive Diffusion Models cites this paper.

BudgetFusion: Perceptually-Guided Adaptive Diffusion Models CLIP-GEN: Language-Free Training of a Text-to-Image Generator with CLIP

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-11T20:29:03.376858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:29:03.376858Z digest=sha256:a15095aa3d0ce564c4009415d86af3aeb3055896f012b80a129ad1f964a7fa53

Observation 7b895ac4-d2b2-4224-9f9e-fff8f3f4666c · inbound

How Vision-Language Tasks Benefit from Large Pre-trained Models: A Survey cites this paper.

How Vision-Language Tasks Benefit from Large Pre-trained Models: A Survey CLIP-GEN: Language-Free Training of a Text-to-Image Generator with CLIP

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-11T18:11:54.551267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T18:11:54.551267Z digest=sha256:580a6f956e604f328f0f179815b0ac172778806512c1dbecef7793cc86eaee58

Observation fc31449a-d54d-4159-8edc-c13736f48fc1 · inbound

Implicit Inversion turns CLIP into a Decoder cites this paper.

Implicit Inversion turns CLIP into a Decoder CLIP-GEN: Language-Free Training of a Text-to-Image Generator with CLIP

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T12:58:10.360905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:58:10.360905Z digest=sha256:5ba391b566711c1b1df47abf447d19dca964af1babdab68eb1adc6c63d205fc9

Observation 3bba8fd5-c2a1-4655-a9c9-e30ce919f66d · inbound

Normality Prior Guided Multi-Semantic Fusion Network for Unsupervised Image Anomaly Detection cites this paper.

Normality Prior Guided Multi-Semantic Fusion Network for Unsupervised Image Anomaly Detection CLIP-GEN: Language-Free Training of a Text-to-Image Generator with CLIP

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T18:51:47.472317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:51:47.472317Z digest=sha256:d8e5d2e5f5333058314f369a03a667d9222c2ab1d4e47fcf676346d28ad85787

Observation 58ee5fe9-b9a1-47fb-842b-0053e92d9dce · inbound

EdgeLoRA: An Efficient Multi-Tenant LLM Serving System on Edge Devices cites this paper.

EdgeLoRA: An Efficient Multi-Tenant LLM Serving System on Edge Devices CLIP-GEN: Language-Free Training of a Text-to-Image Generator with CLIP

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-06T20:59:02.389027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:59:02.389027Z digest=sha256:7f93ece1b4a2d6c3dcfc8c20923629a74e8f1e5c634d726dafa6ef73a69b7605

Observation 050fddba-2847-49b4-ae68-99af450117c7 · inbound

Language-based Color ISP Tuning cites this paper.

Language-based Color ISP Tuning CLIP-GEN: Language-Free Training of a Text-to-Image Generator with CLIP

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T17:39:24.326550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:39:24.326550Z digest=sha256:2946905e23d483a1200099e808cc55e2790df32c8a46e0e0ee3cd652869d05b5