Pith. sign in

Paper Citation Record · LEDGER

CAT: Content-Adaptive Image Tokenization

As of 11 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 10 inbound Pith citation observations for arXiv:2501.03120.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.03120 v1

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:56:44.078464Z

measured 76 of 76 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T13:40:31.721604Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T16:27:09.125374Z

Reference resolution

66 of 66 outbound references displayed

  • verified exact2
  • verified fuzzy22
  • unresolved38
  • parse uncertain1
  • malformed identifier3
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fba74c6f-75aa-4b04-b0dd-a0ca4af632f7 · outbound

This paper cites Taming transformers for high-resolution image synthesis, 2020.

CAT: Content-Adaptive Image Tokenization Taming transformers for high-resolution image synthesis, 2020

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:56:45.427876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:43.689710Z digest=sha256:7ba15bb0a61077c6b5362fcade6459087c57f57a88ba35820c7bdb2e9ee67365

Observation 4725b149-478e-4adc-9fec-e3b74a11a2ce · outbound

This paper cites Auto-encoding varia- tional bayes, 2014.

CAT: Content-Adaptive Image Tokenization Auto-encoding varia- tional bayes, 2014

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:56:45.412447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:43.695604Z digest=sha256:0bf5e338398d60843f5fac204e261f9d271cef4a9d38411c151f95d16cb4cb4d

Observation dd810094-b9a9-48ab-9b37-cdb4fd279304 · outbound

This paper cites Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation.

CAT: Content-Adaptive Image Tokenization Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.700330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.700330Z digest=sha256:170eb0767f1719602040efde2fe4ff28127cb8cd6c643b22cd65ff4561476239

Observation 9a6736b3-2c50-438f-bb63-3ebea2875677 · outbound

This paper cites An Image is Worth 32 Tokens for Reconstruction and Generation.

CAT: Content-Adaptive Image Tokenization An Image is Worth 32 Tokens for Reconstruction and Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.706541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.706541Z digest=sha256:b40f1c40a62746f27b653051e62e8210303400283d6e22fe15b73db76d8bdc31

Observation 0bd6685d-fa90-49b4-9572-c410e31c0eab · outbound

This paper cites Ef- ficient architecture search for diverse tasks.

CAT: Content-Adaptive Image Tokenization Ef- ficient architecture search for diverse tasks

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:56:45.390932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:43.712020Z digest=sha256:cbeccd123d343b4315bfd52d97fb444357048cb7010736e08079d4a34bd406d9

Observation 896fba71-a58a-4cc7-90c8-44ef6990e0f5 · outbound

This paper cites NAS-bench- 360: Benchmarking neural architecture search on diverse tasks.

CAT: Content-Adaptive Image Tokenization NAS-bench- 360: Benchmarking neural architecture search on diverse tasks

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:56:45.371391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:43.717112Z digest=sha256:daae59a6a974efa874a8b8690176dd7a02f45c86ce4651e759a5d7864c4ac2b5

Observation ddf4e365-e1bb-4967-8416-b5a05a6230ab · outbound

This paper cites Finite scalar quantization: Vq-vae made simple, 2023.

CAT: Content-Adaptive Image Tokenization Finite scalar quantization: Vq-vae made simple, 2023

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:56:45.353504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:43.723475Z digest=sha256:05e47ad9fe8e5e2b0daf7ad18591c5fa34bd405da210397d34530a254635b28e

Observation 32c73fe9-d893-4df8-976c-c96a3df6129b · outbound

This paper cites an unresolved cited work.

CAT: Content-Adaptive Image Tokenization Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.730493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.730493Z digest=sha256:2a7e9ccd2ad37dfa170362a4bfa0696cc3560cadc34a8bff3755d09d4c705cec

Observation 1d5351c5-a6aa-48f7-916a-d24f36b6382d · outbound

This paper cites Microsoft COCO: Common Objects in Context.

CAT: Content-Adaptive Image Tokenization Microsoft COCO: Common Objects in Context

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.736539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.736539Z digest=sha256:40ff35fb3710e2c33b05b3ddf2bf413ec842e67a3a819c9f7bdf33fe825591df

Observation 44d9325b-cc09-42d4-b80f-6c48ff8630e9 · outbound

This paper cites Elastictok: Adaptive tok- enization for image and video.

CAT: Content-Adaptive Image Tokenization Elastictok: Adaptive tok- enization for image and video

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:56:45.336847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:43.742114Z digest=sha256:f1de6491c12d55699332f1dbef0f865671476090b5a6360ea425ba1fdb7299b8

Observation cc3aacd2-6da8-4fbe-9d53-9e99d0f1738d · outbound

This paper cites High-resolution image syn- thesis with latent diffusion models, 2021.

CAT: Content-Adaptive Image Tokenization High-resolution image syn- thesis with latent diffusion models, 2021

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:56:45.320394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:43.747565Z digest=sha256:51f0cf0f0fe2c49fe88dc64827abf63942b0de3baf807440dc0cc1d70a3151df

Observation 71633840-d5d7-4b0b-8508-c822581fddcd · outbound

This paper cites an unresolved cited work.

CAT: Content-Adaptive Image Tokenization Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:56:45.299938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:43.753319Z digest=sha256:4c570798179d06baeed1b9c31579422947f68a1a4ba041b40a1bfd4fd1c6b268

Observation e95c0177-880d-4806-a223-da1d339fbcf0 · outbound

This paper cites Deep learning face attributes in the wild.

CAT: Content-Adaptive Image Tokenization Deep learning face attributes in the wild

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:56:45.280170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:43.759208Z digest=sha256:73371b69aae404713147bd4ac4d11eced0372f9088a2907bddaf26973f8b1708

Observation 52db2009-5ed6-44c9-b1cf-0a3e1c2dc416 · outbound

This paper cites ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning.

CAT: Content-Adaptive Image Tokenization ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.763638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.763638Z digest=sha256:c116cafa7b9145a40f813740847dfbb6c132d21191f7c00b710f2be31dd724bb

Observation dbb5b620-ce94-4cd0-9919-cfbb6f347642 · outbound

This paper cites Scalable Diffusion Models with Transformers.

CAT: Content-Adaptive Image Tokenization Scalable Diffusion Models with Transformers

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.769440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.769440Z digest=sha256:801304e0a13907f17e6f3bcdb4c971828b1740f8850fd4f71313004fa2cd552e

Observation 08467109-edb7-43a4-bbc7-77e24a972a6d · outbound

This paper cites Neural Discrete Representation Learning.

CAT: Content-Adaptive Image Tokenization Neural Discrete Representation Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.775234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.775234Z digest=sha256:7ce91c12ab509fd455773ffeb5e5f9fcc88b722a390409ad876d614a2cf4148d

Observation 9f10eda1-0f24-4e72-9db5-b94306916a66 · outbound

This paper cites Wiegand, G.J.

CAT: Content-Adaptive Image Tokenization Wiegand, G.J

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.780669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.780669Z digest=sha256:984678dabfa7030230ca84f08fb0d25827ff1905de88ff63358b9b4c027e8d13

Observation 4c416156-83d0-4f8b-8e6c-e8fc60c9e4d7 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

CAT: Content-Adaptive Image Tokenization An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.788229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.788229Z digest=sha256:03dd5db8d5aa9ed3e4e588f02d431e258c18e2c262669e1015bcedd98e702042

Observation b82361eb-cca7-459c-a7ad-1fa8badf29aa · outbound

This paper cites Dynamicvit: Efficient vision transformers with dynamic token sparsification.

CAT: Content-Adaptive Image Tokenization Dynamicvit: Efficient vision transformers with dynamic token sparsification

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:56:45.264194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:43.794553Z digest=sha256:838877a1f1df956540b124443b7dc6c3195fdf07cf0ce4ad1b926a9f0ae2b332

Observation a349ef33-4d8c-44df-948b-553ecbaa1836 · outbound

This paper cites A-ViT: Adaptive tokens for ef- ficient vision transformer.

CAT: Content-Adaptive Image Tokenization A-ViT: Adaptive tokens for ef- ficient vision transformer

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:56:45.247519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:43.799289Z digest=sha256:cd299267f282867815a548ea711c18368cfec63a791a80830f455065d68a4c32

Observation be924f65-5cc6-44a3-8bbe-d7414449e67e · outbound

This paper cites Token merging: Your ViT but faster.

CAT: Content-Adaptive Image Tokenization Token merging: Your ViT but faster

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:56:45.228133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:43.804638Z digest=sha256:857c708729d37a6a273722f024583deae12421dfd2ed5917103c1334744c3151

Observation ab99dc1b-fc35-4e26-ac69-b7c9a6d17209 · outbound

This paper cites Efficient Video Action Detection with Token Dropout and Context Refinement.

CAT: Content-Adaptive Image Tokenization Efficient Video Action Detection with Token Dropout and Context Refinement

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-10T21:56:44.576444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:43.810944Z digest=sha256:3f47a43602c7c35a4bd4573f0e21faf64718414c7e0d2fe1c9efbc7981b1ca38

Observation 5344084a-1c81-409f-84ca-60a6f79cc2ad · outbound

This paper cites Vision Transformers with Mixed-Resolution Tokenization.

CAT: Content-Adaptive Image Tokenization Vision Transformers with Mixed-Resolution Tokenization

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.816592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.816592Z digest=sha256:79d37987cc378ab9f8c4b9c50eeb79deb185d9a0fe19137b54df4ed4e61fcd55

Observation ef7c60dc-0343-4f08-8515-8c61c6bb7465 · outbound

This paper cites Adaptive Length Image Tokenization via Recurrent Allocation.

CAT: Content-Adaptive Image Tokenization Adaptive Length Image Tokenization via Recurrent Allocation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.822118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.822118Z digest=sha256:c99c2747dd19874aa42b3b3c29357c1ab969a96de842e113f9de4aa598ac96e3

Observation 689ad22b-5d7e-4741-af9e-31fe713bdab4 · outbound

This paper cites U-net: Convolutional networks for biomedical image segmentation,.

CAT: Content-Adaptive Image Tokenization U-net: Convolutional networks for biomedical image segmentation,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.827416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.827416Z digest=sha256:bf26d7d27102926fd9a2746ed23d85d5a05b467910f98f66d2d043221562dbbb

Observation 8ccfce6c-8771-465d-afe4-328015104a91 · outbound

This paper cites Matryoshka representation learning.

CAT: Content-Adaptive Image Tokenization Matryoshka representation learning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:56:45.196223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:43.837435Z digest=sha256:0c34d797838f81e9e1f0332f3c650e2f8baeed8a4fbded39cb95977fb78cbeee

Observation 581fada8-fa52-4455-843f-056e138f6073 · outbound

This paper cites Matryoshka Multimodal Models.

CAT: Content-Adaptive Image Tokenization Matryoshka Multimodal Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.843091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.843091Z digest=sha256:a7821e72ce81f0b750319d05ddcbca30ca6a17d3fe303e3bf5ef76cbd7c2e66e

Observation b09b575b-26b8-43a6-871f-1828ae6cedcb · outbound

This paper cites Matryoshka Diffusion Models.

CAT: Content-Adaptive Image Tokenization Matryoshka Diffusion Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.848801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.848801Z digest=sha256:3d9898758395380ca598033462ea9c0c20e445bc31baf0eacddc1fa23b23f54f

Observation dd3c8177-7d31-47c3-954a-81284f1f1d09 · outbound

This paper cites Transframer: Arbitrary Frame Prediction with Generative Models.

CAT: Content-Adaptive Image Tokenization Transframer: Arbitrary Frame Prediction with Generative Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.855083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.855083Z digest=sha256:40adaea9e4a69fe9d2f2c0a0bc275107c95f390ab11c1c1352a74305d2dfe082

Observation f608f7d0-3e62-41e8-a237-5013c93dca0a · outbound

This paper cites Matryoshka query trans- former for large vision-language models, 2024.

CAT: Content-Adaptive Image Tokenization Matryoshka query trans- former for large vision-language models, 2024

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:56:45.178890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:43.862153Z digest=sha256:fd9d1f7481f8c4debbc889695b68ba8f64bdf8fbd113eca2674c723f452ddc14

Observation 245738a3-97b5-44bd-8c97-15d191d4718a · outbound

This paper cites Dery, Corey Staten, Mikhail Khodak, Graham Neubig, and Ameet Talwalkar.

CAT: Content-Adaptive Image Tokenization Dery, Corey Staten, Mikhail Khodak, Graham Neubig, and Ameet Talwalkar

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:56:45.159011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:43.868549Z digest=sha256:a56561e65e721205e6afc3033a18ed87cba6d7a653503f133b6ef05c7739356a

Observation 1354c351-fba5-4c91-8e9c-04717aa0c30d · outbound

This paper cites UPS: Efficiently Building Foundation Models for PDE Solving via Cross-Modal Adaptation.

CAT: Content-Adaptive Image Tokenization UPS: Efficiently Building Foundation Models for PDE Solving via Cross-Modal Adaptation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.874099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.874099Z digest=sha256:4c36566fdf73f570d32d8c54071064903297cf00ac193f4b203cf1d11eb6b2cd

Observation 1fae709b-1b15-4bd2-8b23-0839c73809e9 · outbound

This paper cites an unresolved cited work.

CAT: Content-Adaptive Image Tokenization Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:56:45.139044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:43.879443Z digest=sha256:6ef179779cad43352b795b3c0d0ffd2826c8f2a04868d1f33dced3352e2ade67

Observation c13d75c8-dd44-437e-a40d-2ac453a53c5e · outbound

This paper cites Tag-llm: Repurposing general-purpose llms for specialized domains, 2024.

CAT: Content-Adaptive Image Tokenization Tag-llm: Repurposing general-purpose llms for specialized domains, 2024

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:56:45.105877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:43.894397Z digest=sha256:75d40877bf1b7b25e8aac3169fce7a45fb875c0dd094ada5893b3446ce243b12

Observation 18ea68a4-7a83-444a-965f-577352024f44 · outbound

This paper cites The unreasonable effectiveness of deep features as a perceptual metric.

CAT: Content-Adaptive Image Tokenization The unreasonable effectiveness of deep features as a perceptual metric

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:56:45.086950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:43.899667Z digest=sha256:b1a66b2e0a833f844299d18208bf889a36a96337e781080dd9df231fa1a73260

Observation b1abc578-9062-4ae9-844c-f1961c2b6ad1 · outbound

This paper cites Very Deep Convolutional Networks for Large-Scale Image Recognition.

CAT: Content-Adaptive Image Tokenization Very Deep Convolutional Networks for Large-Scale Image Recognition

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.905291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.905291Z digest=sha256:53f086db791a09c851f90fa2ca61b84d817972b5f0c731eba3f312619a2b1bd5

Observation a78100ea-4bc6-4df2-8ecd-067a982deb1e · outbound

This paper cites Instructblip: Towards general- purpose vision-language models with instruction tuning,.

CAT: Content-Adaptive Image Tokenization Instructblip: Towards general- purpose vision-language models with instruction tuning,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.911041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.911041Z digest=sha256:c970d970e60d175867e29557ee7e359afe6f6d283d38920b5b0bcdb44536ee66

Observation aa065f64-0b2f-4533-a846-f0149a060ec1 · outbound

This paper cites The Llama 3 Herd of Models.

CAT: Content-Adaptive Image Tokenization The Llama 3 Herd of Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.923167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.923167Z digest=sha256:c53e150ac88722b286130f833165acde34aae3c2270243e0163f72fa1c5e34aa

Observation 836c873a-a8c0-4e7d-9de5-f6b959ab23b9 · outbound

This paper cites Deep Residual Learning for Image Recognition.

CAT: Content-Adaptive Image Tokenization Deep Residual Learning for Image Recognition

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.928836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.928836Z digest=sha256:ef4dd87c3ae96f4ca73830f0965b80d7ddcf186e9d8a978b17a53c7fcc038e7c

Observation 6ab12f3d-4cd6-42aa-a93e-ee65128665ae · outbound

This paper cites Momentum Contrast for Unsupervised Visual Representation Learning.

CAT: Content-Adaptive Image Tokenization Momentum Contrast for Unsupervised Visual Representation Learning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.934729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.934729Z digest=sha256:b6e2374581554b8579cd8fa3a6be5e0353d7b4ccd0e71bddbf303c5bc8f20fe2

Observation 3452543c-ef77-4cf6-9618-acf26b7df7a8 · outbound

This paper cites Generative Adversarial Networks.

CAT: Content-Adaptive Image Tokenization Generative Adversarial Networks

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.943705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.943705Z digest=sha256:1d425b6e39ae5808dc9aa88357e980b065a1f8ae50590e97bd3b0d3a5e586640

Observation 7ff762b9-2c0b-4789-bf74-b3383bc6f2a4 · outbound

This paper cites Image quality metrics: Psnr vs.

CAT: Content-Adaptive Image Tokenization Image quality metrics: Psnr vs

Reference 42

Resolution
malformed identifier
no resolver link, observed 2026-08-10T21:56:43.950350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.950350Z digest=sha256:8dc2fe9c8deae1484147b7994dac79d51dcbb3e290d0f927ef9998ad1b53c4d7

Observation 7df01af4-5eaf-4840-bdfc-2976047d096d · outbound

This paper cites Emu: Enhancing Image Generation Models Using Photogenic Needles in a Haystack.

CAT: Content-Adaptive Image Tokenization Emu: Enhancing Image Generation Models Using Photogenic Needles in a Haystack

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.956292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.956292Z digest=sha256:89d784c081c6393a4e9e1af1252664d5b6a26222ed2a2214435bbd74e2f984f2

Observation 1bfc2511-0e92-48c1-98ec-5f20dc11f165 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilib- rium, 2018.

CAT: Content-Adaptive Image Tokenization Gans trained by a two time-scale update rule converge to a local nash equilib- rium, 2018

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:56:45.057932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:43.961715Z digest=sha256:c0bae801b0e5ff42581d10efc33ea1fa73bebdc6c2ad8b8522aebe7da8840e96

Observation b4ebcfb0-4e36-4f45-bd47-3146680c8f44 · outbound

This paper cites Continuous Conditional Generative Adversarial Networks: Novel Empirical Losses and Label Input Mechanisms.

CAT: Content-Adaptive Image Tokenization Continuous Conditional Generative Adversarial Networks: Novel Empirical Losses and Label Input Mechanisms

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-08-10T21:56:44.281547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:43.966283Z digest=sha256:d91ed7f045ff12a6bc0c166151c0790d1a567873e60b94e3a7e0b8fde52a73e2

Observation cf57f155-8c9d-4c1d-9ba0-59f850a792c5 · outbound

This paper cites Improved Techniques for Training GANs.

CAT: Content-Adaptive Image Tokenization Improved Techniques for Training GANs

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.972119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.972119Z digest=sha256:98ae70ebae4443d39af6ba3f96492013368c30cadd038a3ee47ac457bf4a9f7c

Observation 5ea93eba-8338-490f-8f0f-ae1a569fe50d · outbound

This paper cites Improved Precision and Recall Metric for Assessing Generative Models.

CAT: Content-Adaptive Image Tokenization Improved Precision and Recall Metric for Assessing Generative Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.977508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.977508Z digest=sha256:2dd6fc21c39dd7e8505f372bb3f9ad966825b00c6ed2be7bc591ac510f8d78eb

Observation f90a27c8-0694-462f-a513-6a4fb9e88efb · outbound

This paper cites Classifier-Free Diffusion Guidance.

CAT: Content-Adaptive Image Tokenization Classifier-Free Diffusion Guidance

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.983864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.983864Z digest=sha256:d5676a09474007be05f841de8f4095394253c11561505fdc75ad6558f1a6c962

Observation 0a478b81-018b-4d60-988a-a0871f770422 · outbound

This paper cites an unresolved cited work.

CAT: Content-Adaptive Image Tokenization Unresolved cited work

Reference 49

Resolution
malformed identifier
raw_fallback, observed 2026-08-10T21:56:45.040807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:43.991519Z digest=sha256:e6fef67e3e0f97056dcf860b7a414b66f3966b0e2f96eb94b0b95a5531baf293

Observation 61a9a0bc-f40a-4fbb-93ee-32571eda1467 · outbound

This paper cites ScribeAgent: Towards Specialized Web Agents Using Production-Scale Workflow Data.

CAT: Content-Adaptive Image Tokenization ScribeAgent: Towards Specialized Web Agents Using Production-Scale Workflow Data

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.996371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.996371Z digest=sha256:6c9be40584636d3c848cddfb702051650b8d4176c3182b2b7f60fb658330dc58

Observation ff4b7560-59d4-46f4-ac89-223f0fa434c4 · outbound

This paper cites Chameleon: Mixed-Modal Early-Fusion Foundation Models.

CAT: Content-Adaptive Image Tokenization Chameleon: Mixed-Modal Early-Fusion Foundation Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:44.001955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:44.001955Z digest=sha256:d02b94dd86828a308f0aa3bd26325dcbd5e68f6d9a7b0c7a61070e2075f8b27a

Observation fc7e872b-1719-4442-be81-b223937911a2 · outbound

This paper cites Transfusion: Pre- dict the next token and diffuse images with one multi-modal model, 2024.

CAT: Content-Adaptive Image Tokenization Transfusion: Pre- dict the next token and diffuse images with one multi-modal model, 2024

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:56:45.020162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:44.009461Z digest=sha256:6f9da27cb2aec82f692c72b79e3af9ee1608ca120d12f71c08d263ce2f350ddb

Observation 92e4e978-0a98-4045-9b01-3f55450d391e · outbound

This paper cites A style-based generator architecture for generative adversarial networks,.

CAT: Content-Adaptive Image Tokenization A style-based generator architecture for generative adversarial networks,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:44.016173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:44.016173Z digest=sha256:bcddac420ad901ba2fdf28861eff8b97c3f66de100d5167c0da01e3723855543

Observation 4a80a1eb-af9b-4497-84f8-6bd852528ad6 · outbound

This paper cites The open images dataset v4: Uni- fied image classification, object detection, and visual rela- tionship detection at scale.

CAT: Content-Adaptive Image Tokenization The open images dataset v4: Uni- fied image classification, object detection, and visual rela- tionship detection at scale

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:56:44.988101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:44.032020Z digest=sha256:4ed6c0d7743d3b394a3da131f14077071d18f3c85603fa82747d1bc9efeee8e8

Observation 83334c58-8ad6-4df3-bfd6-32c28caee51c · outbound

This paper cites • Are there faces in the image? → Yes/No.

CAT: Content-Adaptive Image Tokenization • Are there faces in the image? → Yes/No

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:56:44.971378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:44.045089Z digest=sha256:aef6c114d865a511ff659fe75c9bb95839faed24e7cdab6c9701db52509e9304

Observation e9454bba-2eb3-4726-9f8b-2228c1568863 · outbound

This paper cites an unresolved cited work.

CAT: Content-Adaptive Image Tokenization Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:56:44.954309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:44.050336Z digest=sha256:84495de23b2edf31a293b909f3ede53ce362abe7da197a44a0ac381703efe254

Observation e40a356b-b7c3-4ef5-b6ae-5f1655ce9c5d · outbound

This paper cites an unresolved cited work.

CAT: Content-Adaptive Image Tokenization Unresolved cited work

Reference 62

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:56:44.937162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:44.055922Z digest=sha256:322dcb93d5baf44823a8ce97033d970e3c55755d3ea172c97ed4d23032fdeca9

Observation 6dda0d0e-11d9-4d99-af8d-f7f860ec2c84 · outbound

This paper cites an unresolved cited work.

CAT: Content-Adaptive Image Tokenization Unresolved cited work

Reference 63

Resolution
parse uncertain
raw_fallback, observed 2026-08-10T21:56:44.919341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:44.062696Z digest=sha256:50830d8b9cd7af1128d67e27ca434b1164725807a215277a34e9dea1a196543d

Observation d6a455bc-8753-4bdb-9051-672e72da3003 · outbound

This paper cites Score: ? out of 9.

CAT: Content-Adaptive Image Tokenization Score: ? out of 9

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:56:44.902038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:44.067455Z digest=sha256:66ec4b4e798c213978e5794ebcc48f00bef6ef78611ed71c384af2daa4f6ed0d

Observation d9e41782-d961-41d8-8544-6dc7eb932db8 · outbound

This paper cites Architecture We implement the nested V AE similar to theAutoencoderKL implementation of the diffusers library.

CAT: Content-Adaptive Image Tokenization Architecture We implement the nested V AE similar to theAutoencoderKL implementation of the diffusers library

Reference 65

Resolution
malformed identifier
raw_fallback, observed 2026-08-10T21:56:44.884556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:44.072620Z digest=sha256:604e4567d7b6ac6d09e6c3ab9bcf99bc4b1ec323ddba5f6babf7c4e73905d8b3

Observation 026763e5-32e4-4baa-adeb-7a35b731b69f · outbound

This paper cites Architecture We use DiT-XL architecture with a patchify downsampler and patch size of 2.

CAT: Content-Adaptive Image Tokenization Architecture We use DiT-XL architecture with a patchify downsampler and patch size of 2

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:56:44.866856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:44.078464Z digest=sha256:aedddc8fc7ea1352ca72fbc5eacfcad050109b36b5a3dce5b32b5e0014276df2

Observation bf636b65-d38d-49c4-88de-4245d4e1b651 · outbound

This paper cites URL http: //dx.doi.org/10.1007/s11263-020-01316-z.

CAT: Content-Adaptive Image Tokenization URL http: //dx.doi.org/10.1007/s11263-020-01316-z

Reference 1405

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:44.039620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:44.039620Z digest=sha256:cbdc465146897eef165b2cb55b44b282b2fe3abe8e95bdb97263d5fdfa202924

Observation 3bfd9566-4c09-414d-b4c9-f0f8bd8c97b3 · outbound

This paper cites U-Net: Convolutional Networks for Biomedical Image Segmentation.

CAT: Content-Adaptive Image Tokenization U-Net: Convolutional Networks for Biomedical Image Segmentation

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.832491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.832491Z digest=sha256:9d0cae957fd2c40ec0d5c397c37a7f4ddb38398a415a50f87bc0a926582ebdc5

Observation b044cd92-913d-436d-8396-5def3a5c8520 · outbound

This paper cites A Style-Based Generator Architecture for Generative Adversarial Networks.

CAT: Content-Adaptive Image Tokenization A Style-Based Generator Architecture for Generative Adversarial Networks

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:44.026023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:44.026023Z digest=sha256:34a0668f676aa12a75c5dddcdfb63ef5463d63e71451715ec2c91e45f48a1afb

Observation a59f1b9c-d13f-4faa-b7ce-7acb0f72de92 · outbound

This paper cites an unresolved cited work.

CAT: Content-Adaptive Image Tokenization Unresolved cited work

Reference 2021

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:56:45.122987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:56:43.884352Z digest=sha256:e6e24484c6ba749071ef6f2372d83f441911c576304e0fc696ca99fe56f9ac20

Observation 3b6cb5fb-e623-40a4-859b-73f3eee76e44 · outbound

This paper cites InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning.

CAT: Content-Adaptive Image Tokenization InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:43.916875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:43.916875Z digest=sha256:2c7961f96e611f32ef060b9313ee83c267f6f2afdaf38444a29a0902b4b29a26

Pith citing papers

Observation 2576ec84-5f26-41ae-942b-d852836e6f68 · inbound

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity cites this paper.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity CAT: Content-Adaptive Image Tokenization

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.721604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.721604Z digest=sha256:3c08f91b54a17976e2d293f276c27ac22e631b6f41fe5c3fc5281d528cf23c40

Observation 210c2e1d-ef26-42b9-b227-58920dcdad78 · inbound

Thinking vs. Doing: Agents that Reason by Scaling Test-Time Interaction cites this paper.

Thinking vs. Doing: Agents that Reason by Scaling Test-Time Interaction CAT: Content-Adaptive Image Tokenization

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.047548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.047548Z digest=sha256:2ed551fe9a93c18242708577bb2929d3318ddbe3d4bf8ef05b5837651fa3e0c6

Observation 137f9ba6-5214-4a0e-bdc3-03c13c767b3d · inbound

Single-pass Adaptive Image Tokenization for Minimum Program Search cites this paper.

Single-pass Adaptive Image Tokenization for Minimum Program Search CAT: Content-Adaptive Image Tokenization

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T18:35:33.082701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:35:33.082701Z digest=sha256:34692798f76b02c7eac2fbc41b7a8f1a1695f39d35e125273e332705189e33f7

Observation 4b4406d8-d1a8-4cd8-8cba-9d22786eba8c · inbound

ELT: Elastic Looped Transformers for Visual Generation cites this paper.

ELT: Elastic Looped Transformers for Visual Generation CAT: Content-Adaptive Image Tokenization

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:05:59.720179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T17:19:22.543462Z digest=sha256:b7f1c516e697388a3a7fcf9ca2c3b5e24484527462712bc7b98f515c247fb951

Observation a736e1d1-5aff-4191-abeb-c193a8ce2685 · inbound

ELT: Elastic Looped Transformers for Visual Generation cites this paper.

ELT: Elastic Looped Transformers for Visual Generation CAT: Content-Adaptive Image Tokenization

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-02T16:35:03.359155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T16:35:03.359155Z digest=sha256:8349534dac7f360a1f7b98bea82666fb81151d8811fce42eefde676b085effd9

Observation 89064129-f04b-44c7-bcbb-460641706b8f · inbound

What Matters for Diffusion-Friendly Latent Manifold? Prior-Aligned Autoencoders for Latent Diffusion cites this paper.

What Matters for Diffusion-Friendly Latent Manifold? Prior-Aligned Autoencoders for Latent Diffusion CAT: Content-Adaptive Image Tokenization

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:05:56.961107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T01:57:24.033068Z digest=sha256:46c123e520fb45bb194db48d23632c31874c5458a3484fa5c75fd44dbf424a93

Observation f3dc957b-47c9-449a-a4fd-2019ef76f92d · inbound

Structure over Pixels: Learning Variable-Length Visual Programs cites this paper.

Structure over Pixels: Learning Variable-Length Visual Programs CAT: Content-Adaptive Image Tokenization

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:13:49.017325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T18:06:01.684713Z digest=sha256:3aa3f25d5a69f2af51a19e971d7c663d11dbe4236ea0252047c6e42789b948fc

Observation c54db9eb-1bd2-49ff-86e5-56f4e01bc11c · inbound

Diffusing in the Right Space: A Systematic Study of Latent Diffusability cites this paper.

Diffusing in the Right Space: A Systematic Study of Latent Diffusability CAT: Content-Adaptive Image Tokenization

Reference 105

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:36:27.682885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-28T10:44:24.318786Z digest=sha256:a83a7fb3fd3f02a0321ff7049c2a3b38eb7b53133648f8e83e8a0f88232722c1

Observation f175b7d2-8800-4d82-9c6d-55d93195bfcb · inbound

ChannelTok: Efficient Flexible-Length Vision Tokenization cites this paper.

ChannelTok: Efficient Flexible-Length Vision Tokenization CAT: Content-Adaptive Image Tokenization

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:06:44.300147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T07:09:25.049534Z digest=sha256:2e2c14fa5002168f6b75e153066a2b84b1fef6749639c0859647157e12e97a11

Observation 5120ffbd-936e-4178-a2fc-03f2f3d8bfdc · inbound

AdaTok: Self-Budgeting Image Tokenization with Quality-Preserving Dynamic Tokens cites this paper.

AdaTok: Self-Budgeting Image Tokenization with Quality-Preserving Dynamic Tokens CAT: Content-Adaptive Image Tokenization

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T16:27:09.126598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T22:43:09.489524Z digest=sha256:d0006821014c19e2f138c16d84707765a40cc50fccadd8de58329a99bd051aab