Pith. sign in

Paper Citation Record · LEDGER

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation

As of 23 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 0 inbound Pith citation observations for arXiv:2412.08259.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.08259 v2

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T18:05:13.732515Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

57 of 57 outbound references displayed

  • verified exact1
  • verified fuzzy2
  • unresolved54
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 79c5c3e4-6b63-4d3b-b782-35608c87b3b4 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:15.315223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.306070Z digest=sha256:3d0289b585499460ee9224ed9a079e71299ce01368befc3154305ea6571f798d

Observation 82539e37-b44a-4ec7-9709-0f2ddba53aea · outbound

This paper cites W.; Fidler, S.; and Kreis, K.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation W.; Fidler, S.; and Kreis, K

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.314063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.314063Z digest=sha256:4426c6e58ecac50fa7405e935398debaec16b1189455b0f2ebc5da652c5cfd6d

Observation 4cdaff0b-aeac-435f-bc9c-263941199e24 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:15.272703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.320109Z digest=sha256:bb77cf6de71346d758be7ffef6adf6f0a36f3a52e0d3a06ab1ef296a168d760d

Observation 22e2223d-f340-4390-be54-77f182793c3f · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.326876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.326876Z digest=sha256:95177b0d7e6b43bfa7ffaf9b66bde8962cc28a37eccfa3fa633cebf7c967af7d

Observation 2559d893-6ce4-46ea-98e9-f1cbb5a4a32b · outbound

This paper cites PP-OCR: A Practical Ultra Lightweight OCR System.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation PP-OCR: A Practical Ultra Lightweight OCR System

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.333706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.333706Z digest=sha256:46289029848839992ba5e0da4b38e26dd7ec058bc8aa7d37176b8b1500bcba6a

Observation 5dd59835-7482-4052-ae80-f7bcc701c165 · outbound

This paper cites Towards Expressive Communication with Internet Memes: A New Multimodal Conversation Dataset and Benchmark.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Towards Expressive Communication with Internet Memes: A New Multimodal Conversation Dataset and Benchmark

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.340758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.340758Z digest=sha256:18905e217d275e3e6e40b26f23226e2871fa2c590d523918ad5e04a890195345

Observation 42417600-914c-425a-bb94-82cf03f996d1 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:15.221341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.348718Z digest=sha256:f4a14ea5802a251881977f9bb739f9b0cef4f17e4574b8bafe2754c95e9c5b23

Observation 0afb9cf1-96ab-419d-a141-6fde8fe02aba · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.357393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.357393Z digest=sha256:11db95353a2c28e61404fa6dff6a4c29809447a592f2e789cd9a2370000cc097

Observation 4b7cab7d-6831-4883-a652-5ff7f166481f · outbound

This paper cites Animate-A-Story: Storytelling with Retrieval-Augmented Video Generation.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Animate-A-Story: Storytelling with Retrieval-Augmented Video Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.364752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.364752Z digest=sha256:22cd3d94a6aeab1fa3215aabf0dcc26eac18a52b8d644f197a5214e8f0e2295b

Observation 6bce6305-c955-4e42-9763-d75c21be7f16 · outbound

This paper cites CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.372471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.372471Z digest=sha256:f2a3a639018d72ee187c7b97e5a95993f6478c44f57e800b5a0098757fc8b247

Observation c3c3d260-91b2-4834-8bf8-fa544ea837cb · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:15.171497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.380643Z digest=sha256:ac5931633a3bbf76af723a97e4efdbf854ac14021b676a7c9f66bbe320994562

Observation 720248a3-88cb-49d7-ba6a-4ff7bc009379 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:15.144517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.387955Z digest=sha256:be7cbd8973b461be21ddaa92ff5bad2578645efd5098e8431f7d0f08f98312ea

Observation c7aba3eb-7190-4b12-9e8a-d72e47eac37d · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:15.118023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.395650Z digest=sha256:84f4b129cdc228b5f143a93df21569d63789fa8707ef183533392a26beb6374b

Observation 811f941a-0c4d-4976-8653-64601ac5f6af · outbound

This paper cites A.; et al.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation A.; et al

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.401749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.401749Z digest=sha256:8cc3cf09e4810ae90ca6ce27d1c91ba5dd8e9044319757f0c3fa8a98ae1f899f

Observation 8a3a9230-0a91-4e7e-9ab3-73e507fa98c5 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:15.061818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.408013Z digest=sha256:880a700360fa05b81842632cd55e68de71496ccea7c362276e31eebf16b6147a

Observation 9ebbc6a1-0d3b-4a96-9e2d-061410f520f9 · outbound

This paper cites BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.413807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.413807Z digest=sha256:37f7bed4658fa5771a8e774a6c5ca245b0ba1af09defec663bb903c95ad8cbe1

Observation 81f5138f-e152-4b17-be46-a1dab3a24d82 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:15.032060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.426081Z digest=sha256:dc505303c5377927aa4d9768b8b718e4e7990063499cc0e987ec3f163ec987a4

Observation 7d777018-c03a-44f7-811a-c166e8bea33d · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.433193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.433193Z digest=sha256:8e2a0839e7d2b2034bea3b1a7ccd557b658796515ec57e1f886692d8b823aa37

Observation 5c919640-4f31-43fa-ae73-c84224307328 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:14.973552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.440020Z digest=sha256:44e34e803800d7d043361e036ddbb2d9f7dc45b3df39e9169852127034f6ac62

Observation 143cdedc-33f3-4e07-aa21-1483d0186d04 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:14.943518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.447383Z digest=sha256:00f683790c8a82542e2852ce20c8ada1b003eab25246872f5627797c4e38a4c7

Observation f802aa70-8ca5-4a8d-a7a4-f411a6aebbec · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:14.917915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.457131Z digest=sha256:8c7fd13d0bcb84209015c28649318940063759b0602ad707a129f90afaec0b89

Observation 8077bff0-de99-4f6c-9597-77b071d77a11 · outbound

This paper cites Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.463866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.463866Z digest=sha256:fa2b85851756246d05166021facd60da57a2f1662647ef94a9bab5a363efc326

Observation 9e678553-10db-4aca-b5cc-7b53bbea300e · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.470273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.470273Z digest=sha256:bc21eb5e19c5782a9134da3f94e86f52590baf98556ce5a418b7bd0e13b09b7b

Observation 3f27a1e9-3c54-4abe-a3dd-001c25a3552e · outbound

This paper cites X.; and Min, M.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation X.; and Min, M

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T18:05:14.881062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.477206Z digest=sha256:4d00e1c983ca4e90d9bb8cd199926d806a444bbb0c3d49dfe75f54c5de5c2c7d

Observation ff5b5a73-2589-41dc-b4ed-e2cf6094910d · outbound

This paper cites A.; Wang, L.; Cervantes, C.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation A.; Wang, L.; Cervantes, C

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.484233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.484233Z digest=sha256:e8e3f6954f3e3d0d02bb7b857902a47ca3cc9b22bee65080ff76e1011cf5c49d

Observation 5005b19b-13ce-4266-8976-4333e6c7a9ac · outbound

This paper cites MELD: A Multimodal Multi-Party Dataset for Emotion Recognition in Conversations.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation MELD: A Multimodal Multi-Party Dataset for Emotion Recognition in Conversations

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.492090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.492090Z digest=sha256:36cfb2e6bdc2fd886ffcf245e12e518008529774b60ac887538e6a58ed74de70

Observation 21b185dd-9019-47e9-a615-9b52f33b50a6 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:14.838228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.499878Z digest=sha256:0a2207d4141a6e7890891670f24f74cb9969586afe523d93e2b0255385df74d7

Observation 99a08d5e-c490-4f87-9776-c2e1fe23db39 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.506572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.506572Z digest=sha256:b979bb41ac3cdcb36ac8009fd28705fcde4afbfab98cea5491546b4e72adf4d9

Observation 29edc60e-de73-472f-b890-2a47a768ebf5 · outbound

This paper cites LAION-400M: Open Dataset of CLIP-Filtered 400 Million Image-Text Pairs.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation LAION-400M: Open Dataset of CLIP-Filtered 400 Million Image-Text Pairs

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.512709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.512709Z digest=sha256:d2006c1b0cc9f175e92f5883f49e26428142e6a9de744934476368b382428ac7

Observation fc265edc-2d6a-4218-80e8-d6ae6004265e · outbound

This paper cites KNN-Diffusion: Image Generation via Large-Scale Retrieval.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation KNN-Diffusion: Image Generation via Large-Scale Retrieval

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.520791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.520791Z digest=sha256:23f0e6d755ed22b4b0691c1c84fe3bd5a7e42661a92f29b8d8b3a2069e44307c

Observation 1e95f571-c327-4d3e-b717-026b5005a65a · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.530030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.530030Z digest=sha256:1b3df57850bc3788506a2e86afd6714d92d29dc5ad8c8f241f732b6a4c509f2a

Observation 9619dc6e-9f5c-4020-a62a-c26cd4617a74 · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.536746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.536746Z digest=sha256:6035ba781b347edafa301d3769fab02dbbdc4a5de4afdf4ccd306aa20755a87f

Observation 5adb4806-f174-4900-9b2d-11852ae3c975 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:14.774060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.543542Z digest=sha256:fb4fa6a27144619a87f49bd9e2a4e232d08fd672fc7afddeeb4986f9dd9b23c3

Observation 96f79b53-b0d4-4d38-af4b-7bdb49b047b6 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.550734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.550734Z digest=sha256:d04acb4640044eeb6e47c03c8060bc980be6802ac44c26156f0eb5d964c54c2f

Observation 4e9b16f8-348f-47b8-aa7d-ad083dc0d2d2 · outbound

This paper cites Denoising Diffusion Implicit Models.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Denoising Diffusion Implicit Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.558948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.558948Z digest=sha256:1a0eea5e213c6f99e0cacfeb0d46a32c535e12351b5a81e9d07c07d47540ca18

Observation 9ede53ab-144f-4ce2-b222-3f5684408ae4 · outbound

This paper cites Score-Based Generative Modeling through Stochastic Differential Equations.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Score-Based Generative Modeling through Stochastic Differential Equations

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.566395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.566395Z digest=sha256:3a33c4caf739f837b6eef551fa54a7edeaf5a5b443dc82ba58e4deede33e70f5

Observation 258fa863-5324-437a-847c-3c648cc24479 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:14.736132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.572591Z digest=sha256:570ba7bfb0d633e4377dd2d7d5cde41fb43b462655dd8dc750ec34a7e0ef17e8

Observation 6a13ad1b-ce0e-448c-80ed-906e65058295 · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.579130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.579130Z digest=sha256:40f2e8400386115d365dd8712f93ee2a40cf2dec21394e01be49b470da5391a7

Observation 88f0f9c2-55f3-4bed-8a06-82a0d87e1180 · outbound

This paper cites N.; Kaiser, .; and Polosukhin, I.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation N.; Kaiser, .; and Polosukhin, I

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.585414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.585414Z digest=sha256:f054a2e45aebb3d06ffdc2b1bc35a015264f2b2fc4a1576144faf957dd5a8fb5

Observation 81e4186e-3df2-4fd5-bd42-aae55404c3c4 · outbound

This paper cites Swap Attention in Spatiotemporal Diffusions for Text-to-Video Generation.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Swap Attention in Spatiotemporal Diffusions for Text-to-Video Generation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.592846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.592846Z digest=sha256:3bcd7216354f6385ac1eb383aa119ac89651bebdb55eee3e229b7d44956041c4

Observation b26df2c8-ad86-446b-a893-877c218e4cac · outbound

This paper cites VideoComposer: Compositional Video Synthesis with Motion Controllability.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation VideoComposer: Compositional Video Synthesis with Motion Controllability

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.599043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.599043Z digest=sha256:663d120a09b167e7dfe42ca4cbbcecbc4d88221eaa8a1666910f0955ace0dcde

Observation dc70ab91-409b-40ca-a6bb-d4b8ba381730 · outbound

This paper cites GODIVA: Generating Open-DomaIn Videos from nAtural Descriptions.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation GODIVA: Generating Open-DomaIn Videos from nAtural Descriptions

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.607809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.607809Z digest=sha256:7d7eac0bf5955b2d4ed0bea05ce2b1146e8b367ee91d1705c4b0e92b231769a2

Observation fbc45108-7189-4ce1-890d-1c2519b3228d · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:14.694070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.620412Z digest=sha256:d66f328c8e831567e360e5d2a3391ab05863254486d61392a57b3f78bae85deb

Observation 552703ce-3e89-41ab-b011-001acc2d78d3 · outbound

This paper cites Z.; Ge, Y.; Wang, X.; Lei, S.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Z.; Ge, Y.; Wang, X.; Lei, S

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.631121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.631121Z digest=sha256:a29bb2ee8e7cfe73e0ffb8df76179d6c470166d68b91f216510a5794800b8f06

Observation d1bce80a-b382-489e-9496-71d13b9dfcd9 · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:14.648634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.639548Z digest=sha256:bda9d4f6c08de2659454cfaf3a975e8a74ee8cb10f36504223bcbf55272b00c5

Observation a3a4b247-0834-419d-9299-516a767748c2 · outbound

This paper cites Patch-based Object-centric Transformers for Efficient Video Generation.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Patch-based Object-centric Transformers for Efficient Video Generation

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-08-11T18:05:13.988401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.645949Z digest=sha256:acba972f3edcabcfe2d42840d7110f99ce3e74101f09e2b35290c834a5615339

Observation dfcc8d4d-978e-4f76-bb8e-5ed7a1f436f5 · outbound

This paper cites VideoGPT: Video Generation using VQ-VAE and Transformers.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation VideoGPT: Video Generation using VQ-VAE and Transformers

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.652243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.652243Z digest=sha256:993108246142246bc2390708895aafb4ee8a3274ace00d216463a771f6fdf09b

Observation 011c24ec-3fb9-4025-ae1b-cf00cb0df18e · outbound

This paper cites Baichuan 2: Open Large-scale Language Models.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Baichuan 2: Open Large-scale Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.660376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.660376Z digest=sha256:2e4a6ab735560e153c63b633cb828ed9c6890049c4442804a89055188628afda

Observation 8d1143f3-5244-47c6-bfd8-94827eef4d0c · outbound

This paper cites an unresolved cited work.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-11T18:05:14.626912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.666692Z digest=sha256:0fb4e85cc83ade4fc42d516241c758c877100b25dbaf3db51026c28b3519f096

Observation 5f625d45-85ad-4c12-a3aa-6dad7654df97 · outbound

This paper cites Diffusion Probabilistic Modeling for Video Generation.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Diffusion Probabilistic Modeling for Video Generation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.681507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.681507Z digest=sha256:6c4d88ab430b7cc1291184bb765c298524117f17b053979869507403cb5cd75a

Observation 5044ea72-5286-4849-8a56-b200fd6285ba · outbound

This paper cites Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.690809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.690809Z digest=sha256:006d11265af13b2dbf548289b97ba0dc8d58cb5b81d96c3e5d5e96fabc6860cd

Observation 62992de8-fed9-4a43-b863-d77b964768f5 · outbound

This paper cites I2VGen-XL: High-Quality Image-to-Video Synthesis via Cascaded Diffusion Models.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation I2VGen-XL: High-Quality Image-to-Video Synthesis via Cascaded Diffusion Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.697801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.697801Z digest=sha256:e4b9b779969b0b1b4ace78c8fd5ed7fbd8138fdc9cf689eac20d631493fe2c8d

Observation 390699fb-c509-4905-a397-2feab8fa2f1e · outbound

This paper cites D.; and Langlotz, C.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation D.; and Langlotz, C

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T18:05:14.600132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T18:05:13.704382Z digest=sha256:104c00204277f0f71e3d531c7ae31e37b8bfebe73c33cb339dca0dce588c8a65

Observation fb5392fb-4538-4da6-a8c6-075ef8aa9073 · outbound

This paper cites Sticker820K: Empowering Interactive Retrieval with Stickers.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation Sticker820K: Empowering Interactive Retrieval with Stickers

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.711261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.711261Z digest=sha256:a4fbbfb44af2a70893efc762b46be4ccaeb9adcc1e5c2bba97cdb0cee14d21b4

Observation ed3b9ef7-465b-48ae-bf81-89e3e5202b1d · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.719926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.719926Z digest=sha256:a8168f0f1f10b32de6c6e0f0f54eaa432f121d247459cadf14f0150eb8192f86

Observation 181a1b03-bc1d-44cf-8701-2699b379d9ca · outbound

This paper cites , " * write output.state after.block = add.period write newline.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation , " * write output.state after.block = add.period write newline

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.726189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.726189Z digest=sha256:c9ebcce1821b52c189a58adcd6bad493de2d240def991a94b3a5561bfa66a85c

Observation 96af181a-ad7e-4cd6-9951-caa4eea8f455 · outbound

This paper cites write newline.

VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation write newline

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-11T18:05:13.732515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:05:13.732515Z digest=sha256:a4e2e90a92e7a7bd3478da0fff11a042cddd11b749098fee8cde8a4042368a06

Pith citing papers

No inbound Pith citation observations are available.