Pith. sign in

Paper Citation Record · LEDGER

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation

As of 17 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 0 inbound Pith citation observations for arXiv:2412.08259.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.08259 v2

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T18:05:13.732515Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

57 of 57 outbound references displayed

  • verified exact1
  • verified fuzzy2
  • unresolved54
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 79c5c3e4-6b63-4d3b-b782-35608c87b3b4 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:15.315223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.306070Z digest=sha256:733a18c9ec7b018eaf42a7b0c461fed8e0a155fb53fce3490c03b5c12e6dacff

Observation 82539e37-b44a-4ec7-9709-0f2ddba53aea · outbound

This paper cites W.; Fidler, S.; and Kreis, K.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation W.; Fidler, S.; and Kreis, K

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.314063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.314063Z digest=sha256:70c9f75a0ebc2092a24b8bf902cba66fdf73f12e5cbf7f33561ce3c48cc6b759

Observation 4cdaff0b-aeac-435f-bc9c-263941199e24 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:15.272703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.320109Z digest=sha256:0facada63fd8812e7f793fdc69c9348967b2b80a46de2bf6d529a5d30d97b178

Observation 22e2223d-f340-4390-be54-77f182793c3f · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.326876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.326876Z digest=sha256:d0910edf17e9b895abbcd86c258c4b97896b5d7a6ad3c918b158f7b1ca4d2d86

Observation 2559d893-6ce4-46ea-98e9-f1cbb5a4a32b · outbound

This paper cites PP-OCR: A Practical Ultra Lightweight OCR System.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation PP-OCR: A Practical Ultra Lightweight OCR System

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.333706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.333706Z digest=sha256:c4f01eec2c0227bceac0ada13a28eed27455a7289fbbea9a6f50640796eca2e8

Observation 5dd59835-7482-4052-ae80-f7bcc701c165 · outbound

This paper cites Towards Expressive Communication with Internet Memes: A New Multimodal Conversation Dataset and Benchmark.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Towards Expressive Communication with Internet Memes: A New Multimodal Conversation Dataset and Benchmark

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.340758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.340758Z digest=sha256:b37755760c7f55d0ca2738dcc4f0414763a86c6b60713f59a161c92c53ed7b09

Observation 42417600-914c-425a-bb94-82cf03f996d1 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:15.221341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.348718Z digest=sha256:1e5909329aa0b2391231fe39d3eb9667adbb324d28fd173c3b4f00f1a92239ea

Observation 0afb9cf1-96ab-419d-a141-6fde8fe02aba · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.357393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.357393Z digest=sha256:7c1ed7df1a2523a8d8272918e46c421406b3289c3f3bf162b3ea4266f6dfdc8d

Observation 4b7cab7d-6831-4883-a652-5ff7f166481f · outbound

This paper cites Animate-A-Story: Storytelling with Retrieval-Augmented Video Generation.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Animate-A-Story: Storytelling with Retrieval-Augmented Video Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.364752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.364752Z digest=sha256:cdbc75525154e1944d8701e5d876a47838c58c87df6d148d157c34cd04ec62e8

Observation 6bce6305-c955-4e42-9763-d75c21be7f16 · outbound

This paper cites CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.372471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.372471Z digest=sha256:9a90ad785c645461753fadc648a0b7858234ff5184af10fc7b91e2cdd7c17f3a

Observation c3c3d260-91b2-4834-8bf8-fa544ea837cb · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:15.171497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.380643Z digest=sha256:5b8bca9ea5a467df19ab9d97484475288a5961938b0b0d12f4f823846933b1a6

Observation 720248a3-88cb-49d7-ba6a-4ff7bc009379 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:15.144517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.387955Z digest=sha256:f1aeb2cd32ddce250de570a00ea3292db477cc63e9871e883bcea5c6d655bed5

Observation c7aba3eb-7190-4b12-9e8a-d72e47eac37d · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:15.118023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.395650Z digest=sha256:1fc06d61a07a2fdd7b5fdb64083030cf38ac355ca88e0d050e87d4641cb1bcca

Observation 811f941a-0c4d-4976-8653-64601ac5f6af · outbound

This paper cites A.; et al.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation A.; et al

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.401749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.401749Z digest=sha256:8b541891d33aebed253bce18ae30e02da2138499b12bb03b2f82a53b78c0d993

Observation 8a3a9230-0a91-4e7e-9ab3-73e507fa98c5 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:15.061818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.408013Z digest=sha256:d6e960d0c50edd904a624ec661d1a1217d13cfd9b58a1359e1a4dcfed03d79df

Observation 9ebbc6a1-0d3b-4a96-9e2d-061410f520f9 · outbound

This paper cites BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.413807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.413807Z digest=sha256:44e80d3634f47122f184f6e24803e711813ffa6c6c03ca73524852f480393f6e

Observation 81f5138f-e152-4b17-be46-a1dab3a24d82 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:15.032060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.426081Z digest=sha256:2da68f152e31bd7e7a34c1889559fea5d4e83693f63c048998c9cd63ee8769ac

Observation 7d777018-c03a-44f7-811a-c166e8bea33d · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.433193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.433193Z digest=sha256:ea46a2de9a218cacd7884644aa141954eff87d3c458e10b444dc326b24c85a01

Observation 5c919640-4f31-43fa-ae73-c84224307328 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:14.973552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.440020Z digest=sha256:12ab0a73d9acd83462b0d3eec6167062e8758add1ad342aff02994169be73175

Observation 143cdedc-33f3-4e07-aa21-1483d0186d04 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:14.943518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.447383Z digest=sha256:4809cbf6146edd92dd657a274f8fa617ec79909c46ef882adab496dff504f8c5

Observation f802aa70-8ca5-4a8d-a7a4-f411a6aebbec · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:14.917915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.457131Z digest=sha256:78428088691f20d3d4d44108f779ca1efa70558db31fa2c82e30999749b6de78

Observation 8077bff0-de99-4f6c-9597-77b071d77a11 · outbound

This paper cites Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.463866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.463866Z digest=sha256:ae05bda82e1428f35ce9439b0e41200b5d7e63e29c68cfee046c0b469b70b7b3

Observation 9e678553-10db-4aca-b5cc-7b53bbea300e · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.470273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.470273Z digest=sha256:77357ab923e31c05d153355ab605775df9e8304d794e4ce37b83bab3886cd86d

Observation 3f27a1e9-3c54-4abe-a3dd-001c25a3552e · outbound

This paper cites X.; and Min, M.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation X.; and Min, M

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T18:05:14.881062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.477206Z digest=sha256:c06a2fb96f1e95606a7c745662d161d1958984ea089271f980aaa68733497340

Observation ff5b5a73-2589-41dc-b4ed-e2cf6094910d · outbound

This paper cites A.; Wang, L.; Cervantes, C.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation A.; Wang, L.; Cervantes, C

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.484233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.484233Z digest=sha256:b7fef7ab1fa832653f28e051160cbf6395e612bc41103e93f89e2a6ae6da8689

Observation 5005b19b-13ce-4266-8976-4333e6c7a9ac · outbound

This paper cites MELD: A Multimodal Multi-Party Dataset for Emotion Recognition in Conversations.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation MELD: A Multimodal Multi-Party Dataset for Emotion Recognition in Conversations

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.492090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.492090Z digest=sha256:87e1c014b5f1deb7b0879018014c6f05afceb561f4f69f92db1a12887792942c

Observation 21b185dd-9019-47e9-a615-9b52f33b50a6 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:14.838228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.499878Z digest=sha256:90358830da0a74b8f1a5d330f2f3e4ed18d7f7df04f0f4dc696440f1a8af8ec7

Observation 99a08d5e-c490-4f87-9776-c2e1fe23db39 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.506572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.506572Z digest=sha256:740dd545916126223cd9319969e907217e97e429b7a525fd7847a54819836240

Observation 29edc60e-de73-472f-b890-2a47a768ebf5 · outbound

This paper cites LAION-400M: Open Dataset of CLIP-Filtered 400 Million Image-Text Pairs.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation LAION-400M: Open Dataset of CLIP-Filtered 400 Million Image-Text Pairs

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.512709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.512709Z digest=sha256:4a36a13d24672ce8b63559802a41c03d75b76ccf181b051a8d2b0a23cdc61087

Observation fc265edc-2d6a-4218-80e8-d6ae6004265e · outbound

This paper cites KNN-Diffusion: Image Generation via Large-Scale Retrieval.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation KNN-Diffusion: Image Generation via Large-Scale Retrieval

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.520791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.520791Z digest=sha256:75e6095e96c5588e1dba9c42d4809fe88ee05bbbc82fb984b7841244f5a7da02

Observation 1e95f571-c327-4d3e-b717-026b5005a65a · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.530030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.530030Z digest=sha256:34d7ecd58ea303b4b68dd79806774264657e002412980e93a2f1397a2b0e82d8

Observation 9619dc6e-9f5c-4020-a62a-c26cd4617a74 · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.536746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.536746Z digest=sha256:8c8e4196e1b75a73fe558a3c768a02f8b52cc8d8427b7a694bd4b50501828b43

Observation 5adb4806-f174-4900-9b2d-11852ae3c975 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:14.774060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.543542Z digest=sha256:e70a60b2a2de191cb9f1e28d7dd66ffa1e58ddf4010dfae68c0c8317f5f6eb4a

Observation 96f79b53-b0d4-4d38-af4b-7bdb49b047b6 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.550734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.550734Z digest=sha256:ca9326322628538ee1aba427c0bb64dc9997fd8fa1a169e5d50efd5671c22ee9

Observation 4e9b16f8-348f-47b8-aa7d-ad083dc0d2d2 · outbound

This paper cites Denoising Diffusion Implicit Models.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Denoising Diffusion Implicit Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.558948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.558948Z digest=sha256:a6508731e672d19568b4bca85307af6c18ca8792d798a3167340961263a5b4cd

Observation 9ede53ab-144f-4ce2-b222-3f5684408ae4 · outbound

This paper cites Score-Based Generative Modeling through Stochastic Differential Equations.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Score-Based Generative Modeling through Stochastic Differential Equations

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.566395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.566395Z digest=sha256:56ee3f4ef2385121051401554c7896cf3ec0c5970d924a36c623a82383d1afad

Observation 258fa863-5324-437a-847c-3c648cc24479 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:14.736132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.572591Z digest=sha256:dbe161a1ab99b75ca20cc95a58ad7377bf43a7ace0956e43c5bd08fddf633b38

Observation 6a13ad1b-ce0e-448c-80ed-906e65058295 · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.579130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.579130Z digest=sha256:64e06f6e42db99269ead8d9a878b704badc4b3f63cbcfc60cc9fe81d390bc402

Observation 88f0f9c2-55f3-4bed-8a06-82a0d87e1180 · outbound

This paper cites N.; Kaiser, .; and Polosukhin, I.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation N.; Kaiser, .; and Polosukhin, I

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.585414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.585414Z digest=sha256:fe48c3b0d12c58d900218945bcb523bbf35f57b6f127065ff17cfeb61d9400d7

Observation 81e4186e-3df2-4fd5-bd42-aae55404c3c4 · outbound

This paper cites Swap Attention in Spatiotemporal Diffusions for Text-to-Video Generation.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Swap Attention in Spatiotemporal Diffusions for Text-to-Video Generation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.592846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.592846Z digest=sha256:8b22a5665823a2c3dde0e055056ff0f37fedca5477897d17e960b79e5fea07db

Observation b26df2c8-ad86-446b-a893-877c218e4cac · outbound

This paper cites VideoComposer: Compositional Video Synthesis with Motion Controllability.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation VideoComposer: Compositional Video Synthesis with Motion Controllability

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.599043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.599043Z digest=sha256:63127f11900fee60e9a03e847d3ade9d90c1c427cc32ba2da343725538036dc6

Observation dc70ab91-409b-40ca-a6bb-d4b8ba381730 · outbound

This paper cites GODIVA: Generating Open-DomaIn Videos from nAtural Descriptions.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation GODIVA: Generating Open-DomaIn Videos from nAtural Descriptions

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.607809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.607809Z digest=sha256:17a8c2f4f2e3bb3f8171e8709d39c048e77c9fa60b7168c8556aaa25d77e3c73

Observation fbc45108-7189-4ce1-890d-1c2519b3228d · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:14.694070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.620412Z digest=sha256:adcb955f44d20ec7519564372e9d533335685e2bf7e185425b3db856546a2f10

Observation 552703ce-3e89-41ab-b011-001acc2d78d3 · outbound

This paper cites Z.; Ge, Y.; Wang, X.; Lei, S.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Z.; Ge, Y.; Wang, X.; Lei, S

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.631121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.631121Z digest=sha256:eeeb03cbda7d64f3f57458fa9c88f842e4fe72815a0cb1a55994ca1896170615

Observation d1bce80a-b382-489e-9496-71d13b9dfcd9 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:14.648634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.639548Z digest=sha256:6803b715f974938f6d039caa831d7b63e77efe394f4f6424bb3d5bcb2c19706b

Observation a3a4b247-0834-419d-9299-516a767748c2 · outbound

This paper cites Patch-based Object-centric Transformers for Efficient Video Generation.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Patch-based Object-centric Transformers for Efficient Video Generation

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-08-11T18:05:13.988401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.645949Z digest=sha256:fd0cfbd2bc351ebc8442a9cf275884440f447178ccce4011b38079eb1768c090

Observation dfcc8d4d-978e-4f76-bb8e-5ed7a1f436f5 · outbound

This paper cites VideoGPT: Video Generation using VQ-VAE and Transformers.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation VideoGPT: Video Generation using VQ-VAE and Transformers

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.652243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.652243Z digest=sha256:85199b545f9976580d95a4e1d8a16e981a7cc1d674ec59b9ccd7ea6d1235e1f9

Observation 011c24ec-3fb9-4025-ae1b-cf00cb0df18e · outbound

This paper cites Baichuan 2: Open Large-scale Language Models.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Baichuan 2: Open Large-scale Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.660376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.660376Z digest=sha256:aa46371666674a5ffd63f8cae7068061eb01906aa5850f928f32421d4b70683b

Observation 8d1143f3-5244-47c6-bfd8-94827eef4d0c · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:14.626912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.666692Z digest=sha256:dfbd21c4b06d4f80e0ae720b042ffa6808105219ed39353b72b42a7cd81a8d0f

Observation 5f625d45-85ad-4c12-a3aa-6dad7654df97 · outbound

This paper cites Diffusion Probabilistic Modeling for Video Generation.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Diffusion Probabilistic Modeling for Video Generation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.681507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.681507Z digest=sha256:94fd7d32a3606832b550dfc0bac78882477412c3ecd2e7c5bae2aaf938bbb8a1

Observation 5044ea72-5286-4849-8a56-b200fd6285ba · outbound

This paper cites Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.690809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.690809Z digest=sha256:b1912a625abfb36c334c071777c2200d5ca551122124fc0e9e2aca42e7edd5df

Observation 62992de8-fed9-4a43-b863-d77b964768f5 · outbound

This paper cites I2VGen-XL: High-Quality Image-to-Video Synthesis via Cascaded Diffusion Models.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation I2VGen-XL: High-Quality Image-to-Video Synthesis via Cascaded Diffusion Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.697801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.697801Z digest=sha256:47d2821bc774418ac941e0553c65685a2c6aa529e2205e919ef72e33e83b7b2e

Observation 390699fb-c509-4905-a397-2feab8fa2f1e · outbound

This paper cites D.; and Langlotz, C.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation D.; and Langlotz, C

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T18:05:14.600132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.704382Z digest=sha256:45158847f2cfe01ffb7312c83f01c0e4fa3849372c5f786687cbf01d0b189b08

Observation fb5392fb-4538-4da6-a8c6-075ef8aa9073 · outbound

This paper cites Sticker820K: Empowering Interactive Retrieval with Stickers.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Sticker820K: Empowering Interactive Retrieval with Stickers

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.711261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.711261Z digest=sha256:e329973f52e92811966f5d75cc0a95c51241b062f003d5ddd97e2a5a7afca490

Observation ed3b9ef7-465b-48ae-bf81-89e3e5202b1d · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.719926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.719926Z digest=sha256:5d2c879e7b32245ab926f0c07ea6b81d17b12853110aaf4a9672b0a60a35986d

Observation 181a1b03-bc1d-44cf-8701-2699b379d9ca · outbound

This paper cites , " * write output.state after.block = add.period write newline.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation , " * write output.state after.block = add.period write newline

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.726189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.726189Z digest=sha256:0f36f03b720ca96d56d8bdf1b56024b1f644c72cead8fb7a9945564d2ff97a50

Observation 96af181a-ad7e-4cd6-9951-caa4eea8f455 · outbound

This paper cites write newline.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation write newline

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.732515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.732515Z digest=sha256:de1dac69946b3ed17d4dc72a738c76666a290626402a481dc63b4dad098d8fd7

Pith citing papers

No inbound Pith citation observations are available.