Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T23:43:47.791790Z
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 0 inbound Pith citation observations for arXiv:2412.19997.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T23:43:47.791790Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
23 of 23 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 271d822e-5f7e-4d90-96c0-c01a9efb836d · outbound
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training VisualBERT: A Simple and Performant Baseline for Vision and Language
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84196178-abc1-4a18-af50-b40547dc5e9d · outbound
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training Vl-bert: Pre-training of generic visual-linguistic representations,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 50083edc-7a03-4a81-88c1-9d16d088e523 · outbound
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training Oscar: Object-semantics aligned pre-training for vision-language tasks,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a24a72c2-37e8-4397-ae85-02a35a4cc823 · outbound
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training Vilt: Vision-and-language transformer without convolution or region supervision,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation db1da44e-93f7-4d37-878d-b1f1f9a19176 · outbound
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2bf928b7-c73c-4027-8c32-048e63917beb · outbound
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training Align before fuse: Vision and language representation learning with momentum distillation,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7d98765a-5b95-4bca-9f8a-200cefcec455 · outbound
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training Self-distilled dynamic fusion network for language-based fashion retrieval,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 801820f7-f940-4f7d-8545-f4e0ffbf88a4 · outbound
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training Fashionvil: Fashion-focused vision-and-language representation learning,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e8665559-d64f-4865-859c-3dd2f6c8e402 · outbound
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training Fash- ionsap: Symbols and attributes prompt for fine-grained fashion vision- language pre-training,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 4337c9df-c258-418b-b96d-522d6a6510c0 · outbound
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training An image is worth 16x16 words: Transformers for image recognition at scale,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 580fb0c4-045d-406b-a96f-cece8fadede6 · outbound
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training Bert: Pre-training of deep bidirectional transformers for language understanding,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 00009df5-9eda-4b04-9ea9-94b08822592b · outbound
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training Fashion-Gen: The Generative Fashion Dataset and Challenge
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c084c952-1b71-4d6c-9b99-256d95f9be59 · outbound
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training Mmf: A multimodal framework for vision and language research,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 09303349-15de-459c-be10-03f04ed19acc · outbound
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training Pytorch: An imperative style, high-performance deep learning library,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1cb5497-9af9-45ba-9bde-2c75b574168d · outbound
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training Fashionbert: Text and image matching with adaptive loss for cross- modal retrieval,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 32ae1a74-f30d-435b-9a2e-7bcd1c3b16b4 · outbound
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training Kaleido-bert: Vision-language pre-training on fashion domain,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation dfcc346e-bbbb-451f-a195-94e2431bfce0 · outbound
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training Commercemm: Large-scale commerce multimodal represen- tation learning with omni retrieval,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 32b8c058-5e14-450c-a943-8c19cbdb8816 · outbound
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training Ei-clip: Entity-aware interventional contrastive learning for e-commerce cross-modal retrieval,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7dcb3607-e542-459c-8687-249d6d16986e · outbound
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training Masked vision-language transformer in fashion,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1b27ed18-7edd-4a84-bf01-f8dab45fa7a0 · outbound
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training Fashionklip: Enhancing e-commerce image-text retrieval with fashion multi-modal conceptual knowledge graph,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2ef245f4-9280-4682-b13d-f3a30b9924fb · outbound
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training Fad-vlp: Fashion vision-and-language pre-training towards unified retrieval and captioning,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation db0a5a0f-3224-4349-a82b-59db8bff87b7 · outbound
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training Fame-vil: Multi-tasking vision-language model for heterogeneous fashion tasks,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 26506c2f-03b7-4cb1-ade7-90472856ca40 · outbound
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training Syncmask: Synchronized attentional masking for fashion-centric vision-language pretraining,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
No inbound Pith citation observations are available.