Pith. sign in

Paper Citation Record · LEDGER

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation

As of 12 August 2026, this Paper Citation Record lists 89 of 89 outbound references and 0 inbound Pith citation observations for arXiv:2412.01027.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.01027 v2

Coverage vector

measured 89 of 89 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T04:50:49.398206Z

measured 89 of 89 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

89 of 89 outbound references displayed

  • verified exact0
  • verified fuzzy47
  • unresolved42
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7e69f455-c992-483b-ab46-be8f0595c4d8 · outbound

This paper cites Qwen Technical Report.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Qwen Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:48.985145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:48.985145Z digest=sha256:76ea833e52b8287503b93c04662a90e642a612bf091f1a8b3b9b53a44509b3d0

Observation 9d2f5915-d586-40de-a0e7-a41625d08de3 · outbound

This paper cites Sequential modeling enables scalable learn- ing for large vision models.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Sequential modeling enables scalable learn- ing for large vision models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:48.990814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:48.990814Z digest=sha256:e67afd749b19a1fa9fb124a40d78b3c1c3159f8ad6ac37b5c863e44f9bfecea3

Observation ae4f521c-72e1-4ffc-b918-45150b077c5d · outbound

This paper cites eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:48.995942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:48.995942Z digest=sha256:b2f7a61a8488d57ba3cba12b49214007d5b4ebff06079b4872989b46147d46f4

Observation 95e91657-7daf-4258-b0a4-c75ccbd95d39 · outbound

This paper cites Towards in-context scene understanding.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Towards in-context scene understanding

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.001081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.001081Z digest=sha256:4a1850a9a46050fcdb6949d801a88af985c800cb6fa0d91db1e294bfb903ad9b

Observation 42e6cb1a-a23f-46b5-baca-f4b41ff54c40 · outbound

This paper cites Visual prompting via image inpaint- ing.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Visual prompting via image inpaint- ing

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.005923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.005923Z digest=sha256:3532287d93dffabb2e7fe1586541dbc1c6d80ac3a75faf59daf092e9bdffe34d

Observation d22faafb-1358-4135-8902-d8cf788066fb · outbound

This paper cites Ledits++: Limitless image editing using text-to-image models.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Ledits++: Limitless image editing using text-to-image models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.010710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.010710Z digest=sha256:16464f86a46964aab8833afb453c60b6811e2dc2947e01a77b4d17c966fda460

Observation 7a49f04f-0979-45c2-97aa-c8cc16113704 · outbound

This paper cites In- structpix2pix: Learning to follow image editing instructions.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation In- structpix2pix: Learning to follow image editing instructions

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.015800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.015800Z digest=sha256:bfa873f8d9167196a0fe6a03af6da27b701ebf5bdbe31b676ead05371ed910d0

Observation 8130c734-becc-4001-bd83-1af3670e7756 · outbound

This paper cites Lan- guage models are few-shot learners.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Lan- guage models are few-shot learners

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.020532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.020532Z digest=sha256:3f52635ad0d1e83d3aee07f59730b988a17c6b086825955b9a3c1174e73e085d

Observation 1e7c3b91-4bd3-47f7-8d94-840bdca509e8 · outbound

This paper cites Enhancing diffu- sion models with text-encoder reinforcement learning.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Enhancing diffu- sion models with text-encoder reinforcement learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.025163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.025163Z digest=sha256:43016b19fe84cf22fed759ce8433d8d7a6b99ab68ed8925e58622f54961bda8d

Observation 1e27c92e-222a-4859-8811-1fca3d2f5128 · outbound

This paper cites Gentron: Diffusion trans- formers for image and video generation.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Gentron: Diffusion trans- formers for image and video generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.029662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.029662Z digest=sha256:282e1839fceb4d174501add544ad5af5f512fed042058323d8fdf6786ae204da

Observation c69ae28e-a9a7-4e93-8545-cc077f0b27bc · outbound

This paper cites Scaling instruction- finetuned language models.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Scaling instruction- finetuned language models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.034521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.034521Z digest=sha256:c54ab4e37e62e35e0a6219bf868da7c444ec5dd315b36113e19202a50d3ae5b4

Observation efc84018-8926-49cf-8f07-46fa309446c8 · outbound

This paper cites Diffedit: Diffusion-based semantic image editing with mask guidance.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Diffedit: Diffusion-based semantic image editing with mask guidance

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.615935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.039271Z digest=sha256:2b1c333d8a18c65b12ffa6389ae6ed401c1345879280c8162203d22ca953b98e

Observation 550ceb1b-35c7-4783-9eaa-5108a325a2e8 · outbound

This paper cites Emu: Enhancing Image Generation Models Using Photogenic Needles in a Haystack.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Emu: Enhancing Image Generation Models Using Photogenic Needles in a Haystack

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.043738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.043738Z digest=sha256:9725c236622931d08bad5a517fab151cbb5f5aaa24daa7314fcdea99f766f96a

Observation 7c932f14-a5b3-4d0d-b046-20c3371e2a7a · outbound

This paper cites Dreamllm: Synergistic multimodal com- prehension and creation.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Dreamllm: Synergistic multimodal com- prehension and creation

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.598952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.048789Z digest=sha256:e27a5fec09c912e0c2605a529fe484516af912e28658a0b814c3c590837b3800

Observation da5df7ea-fe07-4d7d-8070-74cd828dd806 · outbound

This paper cites Diffusion self-guidance for control- lable image generation.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Diffusion self-guidance for control- lable image generation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.582913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.053140Z digest=sha256:ce0009ecaf78efca444dd126777cf9e937f15c2df876650616e8fd028d8446ab

Observation b22b22d0-667a-4dff-addf-8933ab823dab · outbound

This paper cites Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.057722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.057722Z digest=sha256:88e0151d0b293d70964247fd86980e3503d057e8b5690cb7df99f1f647c466dc

Observation e4aca15d-b16a-413f-9639-fba6578f3de1 · outbound

This paper cites PUMA: Empowering Unified MLLM with Multi-granular Visual Generation.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation PUMA: Empowering Unified MLLM with Multi-granular Visual Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.062861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.062861Z digest=sha256:dcaa20904e7add3f44a3705be6b63ba0f311b6d3bc65d30e352169e69f94f0e6

Observation 0978e7e5-8336-472e-ba08-de109ae41151 · outbound

This paper cites Explore in-context learning for 3d point cloud understanding.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Explore in-context learning for 3d point cloud understanding

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.567490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.067781Z digest=sha256:0624372ac3e26f831b56e9ce98e4a4c8c05a2e563deedc101260149c584d504b

Observation 02febbd3-042a-4d92-bd48-dc8c8dbdb870 · outbound

This paper cites Making llama see and draw with seed tokenizer.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Making llama see and draw with seed tokenizer

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.552297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.072661Z digest=sha256:8c87a87df22e23df5210f049d11b01c84f098dfbab79f05c99ddecb1ec05d764

Observation 16045736-77d1-47c0-ad8a-fc08cf5b5d08 · outbound

This paper cites SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.077160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.077160Z digest=sha256:6a4c502df452b102e3214e2d7cfa36e75841c59bf7a0f95fc468fc78250cb92c

Observation 22026619-52d2-4db6-b391-26cfe5f34d3d · outbound

This paper cites Analogist: Out-of-the-box visual in-context learning with image diffusion model.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Analogist: Out-of-the-box visual in-context learning with image diffusion model

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.536033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.081903Z digest=sha256:113bda56fed3f1f4fc7177100b10e1f95720009f08bd464b76dd8c68d54c4b71

Observation 2b57f07d-d7d7-415b-80c9-3d7dcee161b4 · outbound

This paper cites Generative Visual Instruction Tuning.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Generative Visual Instruction Tuning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.086350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.086350Z digest=sha256:b22fee633d17154c04ada4c311d024cada58953fa8ff009d7c9a4079d0e279ae

Observation e5f45ef6-bc05-4496-9aa7-e15d5c82bc5d · outbound

This paper cites Prompt-to-prompt image editing with cross-attention control.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Prompt-to-prompt image editing with cross-attention control

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.091059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.091059Z digest=sha256:91ddbc259637e713b2bbc84082b706cc414159cefe3f78a1474220884c19f63d

Observation be9aea8b-b3e5-4294-b858-41363877f6b8 · outbound

This paper cites Lora: Low- rank adaptation of large language models.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Lora: Low- rank adaptation of large language models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.095686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.095686Z digest=sha256:afc2741bca5a69c9b8124eb09191a7d95ebcb9e836b5f21b87b079b86a20a169

Observation e06625a7-57d5-468d-8939-a0636341a180 · outbound

This paper cites Multimodal Task Vectors Enable Many-Shot Multimodal In-Context Learning.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Multimodal Task Vectors Enable Many-Shot Multimodal In-Context Learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.099961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.099961Z digest=sha256:f2d12c971328be2192adc5ec28cbfed501de21211f55dabdbfc998cae4894d7f

Observation 782fb5de-b79b-4345-9bcb-0a8a3b6f57dd · outbound

This paper cites Customizing Text-to-Image Models with a Single Image Pair.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Customizing Text-to-Image Models with a Single Image Pair

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.104649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.104649Z digest=sha256:5c4f4d23e3295c2808f9b9c1c88706926e11df1a67df00d75e0d343395b779e5

Observation c42bbc05-7e5b-4236-9798-4f63002fa3af · outbound

This paper cites Chameleon: A data-efficient gener- alist for dense visual prediction in the wild.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Chameleon: A data-efficient gener- alist for dense visual prediction in the wild

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.501147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.109457Z digest=sha256:e4aafb384f3f4d200d5ef6a7237a63aca713a5fad2219c59c85c96c1e8280b01

Observation b479fb83-eae8-4815-88a2-d4ab8d328711 · outbound

This paper cites Gen- erating images with multimodal language models.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Gen- erating images with multimodal language models

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.485953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.113675Z digest=sha256:dcad408a495038a3a7bead64c4f44227132160c54e61cabb45cf70fda4c3970d

Observation 7fd27ea1-e316-4195-bbd3-ef813be05d1e · outbound

This paper cites Lego: Learning egocentric action frame generation via visual instruction tuning.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Lego: Learning egocentric action frame generation via visual instruction tuning

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.470763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.117972Z digest=sha256:5cf4b2f1f5e7bd654b1d8bd5d18a5a027b8b79c106dc3070beefa62fc524d786

Observation 8e1baba4-8a6c-4157-8c46-199fd945edae · outbound

This paper cites Blip-diffusion: pre-trained subject representation for controllable text-to- image generation and editing.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Blip-diffusion: pre-trained subject representation for controllable text-to- image generation and editing

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.455767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.122498Z digest=sha256:f44b5804c65a85512e99fc849a19e849e2ede7bf4f6b70a4f3d5d1b5831b2161

Observation 6680440f-6d9f-436b-b973-f0161e47a795 · outbound

This paper cites Visual in-context prompting.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Visual in-context prompting

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.440485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.126956Z digest=sha256:16bcc9c18f54dc92f23cf6be51ba87924988ef2bd70a54dd68d3190d6bf73b85

Observation 6de807b2-8fa7-47a7-ab82-3ec1cfa9b05a · outbound

This paper cites Autoregressive image generation without vector quantization.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Autoregressive image generation without vector quantization

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.424744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.131506Z digest=sha256:120e9a1f9269bf740dbd5fa2251a4edf1fe2b2accf080299d7c051df4ca5967f

Observation dbc1076f-ff8f-4905-a479-35d3bbbdacdf · outbound

This paper cites Visual atribute transfer through deep image analogy.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Visual atribute transfer through deep image analogy

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.409154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.136146Z digest=sha256:c53eb5d2157036e263a4c038cc2cf2d5c861ed0dcc4d85a80ff1de63b71d26ee

Observation 341e4589-69d9-4d11-87f9-214e4e0fdfec · outbound

This paper cites Text-driven image editing via learn- able regions.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Text-driven image editing via learn- able regions

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.393548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.140731Z digest=sha256:f32dca3d419d9d467f7cbba5e9ecf6af6ecd64c13e209bb98c9a45c126694fab

Observation 3ca134d0-1df1-47ed-b85b-00d74882ee49 · outbound

This paper cites Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.145289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.145289Z digest=sha256:bf5d8807743ed9fd68517ad68ff512502a9d75dad7acd6716a7f5066cbba14b1

Observation 1d50a909-574e-4428-bcc9-dbf2571601a7 · outbound

This paper cites Glid: Pre-training a generalist encoder-decoder vision model.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Glid: Pre-training a generalist encoder-decoder vision model

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.376543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.150110Z digest=sha256:2afc7f80d9f3fe451ab8b79c914a68736b9eca97cb3016f5a5e3c80cc81f9fea

Observation 990b57ef-10a4-4c2a-94c4-fcc155dc971c · outbound

This paper cites Decoupled weight de- cay regularization.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Decoupled weight de- cay regularization

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.360890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.154472Z digest=sha256:ea31a77e87696db930663eed5e37751e2c69dd36f9bb5d12b7e9f5422215eea7

Observation d04cfc0d-8b61-4f46-8337-7be81cf38622 · outbound

This paper cites Unified-io 2: Scaling autoregressive multimodal models with vision language audio and action.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Unified-io 2: Scaling autoregressive multimodal models with vision language audio and action

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.345267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.159126Z digest=sha256:347b3ab3bc99627dbdcc0256c208061fab3f76d4cb4d131edcec6afc8bde8276

Observation 5a2dbbc2-3bf4-42ed-a432-57ed16b64da5 · outbound

This paper cites STAR: Scale-wise Text-conditioned AutoRegressive image generation.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation STAR: Scale-wise Text-conditioned AutoRegressive image generation

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.164789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.164789Z digest=sha256:2710190f0e567ded8f71c4292bf4d6a064923af6d73877f62c46e5a65b64b070

Observation 94eba859-1dda-4caa-81f4-cecbf18da7ae · outbound

This paper cites Sdedit: Guided image synthesis and editing with stochastic differential equa- tions.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Sdedit: Guided image synthesis and editing with stochastic differential equa- tions

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.329938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.169597Z digest=sha256:af47d07f2d68b84ae407bf2854f8554e1824d16d46cd3e1a7d6767d0415c841a

Observation 4d7e444b-6c78-4263-bf1c-57d946bbedc4 · outbound

This paper cites Watch your steps: Local image and scene editing by text instructions.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Watch your steps: Local image and scene editing by text instructions

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.314971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.174342Z digest=sha256:0ad23ef60ac06e2f2ef4e8d06307c089c3741ffc40ec0d81642df9ad45917ef0

Observation 6c227a90-fd98-4521-8a7a-606aad5a5a62 · outbound

This paper cites Null-text inversion for editing real im- ages using guided diffusion models.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Null-text inversion for editing real im- ages using guided diffusion models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.178909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.178909Z digest=sha256:3ba97465ffd11c7f67cf07cd0e7be91e1ead32eda27074353f88522b869cddaa

Observation ccb5e8b5-db28-4486-919e-66f675d78a5a · outbound

This paper cites Visual instruction inversion: image editing via visual prompting.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Visual instruction inversion: image editing via visual prompting

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.288670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.183702Z digest=sha256:d89b97ceb88596a0f40dd1b6f84a17d0d5e9aad8c3251524757a46a09c76b8e8

Observation 77e8225d-7c1d-4ba5-b992-fbc5bbb713b4 · outbound

This paper cites Swiftbrush: One-step text-to-image diffusion model with variational score distilla- tion.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Swiftbrush: One-step text-to-image diffusion model with variational score distilla- tion

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.188120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.188120Z digest=sha256:34099b0a0d498e41d889de67c526cfc40f42d93e23b14af4e8a91dea59b4db62

Observation 5dc0784b-50d1-486c-acb5-ac55f5d3facf · outbound

This paper cites In-context Learning and Induction Heads.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation In-context Learning and Induction Heads

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.192908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.192908Z digest=sha256:0e380f1ecba9717417896618cde2f399cb9aa588850643093ea9307509fbf194

Observation 6f38064d-7310-4111-a58d-0d1f9968d45f · outbound

This paper cites Editing implicit assumptions in text-to-image diffusion models.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Editing implicit assumptions in text-to-image diffusion models

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.260748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.197556Z digest=sha256:b5235b8dd8e54b5079b23651d88b847faeb88e58dfc67b2219c1ef45afe6e90f

Observation c1c4d7f9-2ec4-40e2-b938-2aa0045c72af · outbound

This paper cites Effective real image editing with accelerated iter- ative diffusion inversion.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Effective real image editing with accelerated iter- ative diffusion inversion

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.244675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.202106Z digest=sha256:95b2cbbf24dc416016235d30430d914e6c1b73c44fb17583ab28f930a0e8176c

Observation 61b9e750-79d9-442d-ab5b-15f9306eeb80 · outbound

This paper cites Precisecontrol: En- hancing text-to-image diffusion models with fine-grained at- tribute control.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Precisecontrol: En- hancing text-to-image diffusion models with fine-grained at- tribute control

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.228279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.206707Z digest=sha256:fcef7b7225af2ab1b02aee46ae0e4837bc871d6605592df9badfd167932e440e

Observation 6f51dc45-e4e1-46f9-be90-46257cbfcb66 · outbound

This paper cites True few- shot learning with language models.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation True few- shot learning with language models

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.212993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.211022Z digest=sha256:1a52829fb85bd3ea51f522c1b90f7210fc95a6d359f30958f79dd1600009169f

Observation d09ed24f-437c-4b2a-bb06-4cbe5fa00ce2 · outbound

This paper cites Sdxl: Improving latent diffusion models for high-resolution image synthesis.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Sdxl: Improving latent diffusion models for high-resolution image synthesis

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.197379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.215637Z digest=sha256:f3b0dd710b6fc6cadeacb70416792c6cb7e4e034283178c2bb86cef81994833c

Observation 5f95d10b-3589-4797-817a-56ad676ca7e0 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Learning transferable visual models from natural language supervi- sion

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.220494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.220494Z digest=sha256:3c74a917aa39c756bf915f10538ff2d566637daa374f7176657382daab9e1bdf

Observation f2811d2f-f15d-4fdf-9682-c18b7ecaf352 · outbound

This paper cites Zero-shot text-to-image generation.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Zero-shot text-to-image generation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.225121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.225121Z digest=sha256:13969097324abccab607d28ff1664e175756cdf8584e5e0a3c403ef44f51677d

Observation dff739f4-a601-4e43-8cba-790e15407c08 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.229740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.229740Z digest=sha256:cf3d594799800081afe89444e747e6cac33e9eeb968fa981052f9200960bc994

Observation 1d16b7d2-e0b5-42c5-9c07-e62aec1fb9c7 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation High-resolution image synthesis with latent diffusion models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.234585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.234585Z digest=sha256:384e2a80875ef38bf335e2dacc4dfd25f56abbcbe6db1b19dc29b9b184e54ec6

Observation ed47628d-45e0-4c27-9934-2039ed92315f · outbound

This paper cites Photorealistic text-to-image diffusion models with deep language understanding.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Photorealistic text-to-image diffusion models with deep language understanding

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.239163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.239163Z digest=sha256:c02da02553a9db614fbb92a7b7136ab66e257acd4e72908ea9fca8448c4b669c

Observation b029c63b-4561-4006-9a52-1c8300393e3b · outbound

This paper cites Towards more unified in-context visual un- derstanding.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Towards more unified in-context visual un- derstanding

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.141887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.243734Z digest=sha256:ad982d8d3d3824e5b9f3f9d048ddba109fd816fb47609fa2e1a9bc0fee9e0338

Observation bed095a8-98d4-401c-a9ea-37c7e6087eeb · outbound

This paper cites Emu edit: Precise image editing via recognition and gen- eration tasks.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Emu edit: Precise image editing via recognition and gen- eration tasks

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.126573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.248312Z digest=sha256:cb66794efbba64c361ebb3c3451f5c3d3debe7416f02949d47775dcc956959ef

Observation c5eba021-39b3-46fa-a73a-2238c59c071e · outbound

This paper cites Diffusion image analo- gies.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Diffusion image analo- gies

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.111422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.252807Z digest=sha256:890d216c7d594b37f31ed5c0407b68b5bf448010348d6d587a516217b9aa4020

Observation cc3b2b50-102f-47c8-8c2d-aebe6a17deb0 · outbound

This paper cites Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.257448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.257448Z digest=sha256:e1681788341a8ea20f0f6ec72bc9e20959e3f8230c454b1240f06bda7f3628a2

Observation 5aef0648-b7cb-447d-b6b1-6cae5e0746f9 · outbound

This paper cites Emu: Generative pretraining in multimodality.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Emu: Generative pretraining in multimodality

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.096723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.262189Z digest=sha256:0ce4462aca524b2a06bafc7b686184974b4126e7a052aa210336b561ec175d95

Observation dc8e0ac2-4975-4b76-8973-798f14f1054c · outbound

This paper cites Generative multimodal mod- els are in-context learners.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Generative multimodal mod- els are in-context learners

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.080487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.266610Z digest=sha256:64822b719687d63be7e0753cab28cf7f4a52d5897d5502b2affc40bbc49acd5f

Observation f09ff6ce-4aba-4aa8-b7c7-a14a61329b2d · outbound

This paper cites Imagebrush: learning visual in-context instructions for exemplar-based image manipulation.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Imagebrush: learning visual in-context instructions for exemplar-based image manipulation

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.065599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.271112Z digest=sha256:4335ce47d2c51cdc2557eb041835b83e5418340d2b8b0f6358b6b943a8cabd68

Observation 2cd2823c-4bb5-486f-8f16-e4e0ba6f4c13 · outbound

This paper cites Rethinking and improving visual prompt selection for in-context learning segmentation.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Rethinking and improving visual prompt selection for in-context learning segmentation

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.050054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.275744Z digest=sha256:3dd29549fd3fd419c71b4169d8cbd57d40d26ba5c9822ed52f6fe0fb3a501472

Observation 9d92ecdb-2ea9-4f87-ae6d-f81c3459b945 · outbound

This paper cites Cognitive load during problem solving: Ef- fects on learning.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Cognitive load during problem solving: Ef- fects on learning

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.033456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.280195Z digest=sha256:3b2386ed8762c60fa8b860da2515f168a217e5ce965f62cad15b747f0d68055d

Observation 1f9602eb-e4d2-4d8a-b9d5-c1c61eaac87b · outbound

This paper cites Codi-2: In-context inter- leaved and interactive any-to-any generation.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Codi-2: In-context inter- leaved and interactive any-to-any generation

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.017976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.284699Z digest=sha256:fd7fd738c21f201ca924a9022aa045e784bbd3cf39b47e735da1ff09421d321d

Observation 1957869c-cf8d-4da1-b66c-63aa5d54a2cb · outbound

This paper cites Chameleon: Mixed-Modal Early-Fusion Foundation Models.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Chameleon: Mixed-Modal Early-Fusion Foundation Models

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.289474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.289474Z digest=sha256:c8c599726421a0c3fdb175fa8625b21ff5ca13371538e4f7299abf4617020b8d

Observation b9d81c44-c1b1-422c-ac07-1ce507fa7d5d · outbound

This paper cites How to grow a mind: Statistics, structure, and abstraction.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation How to grow a mind: Statistics, structure, and abstraction

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:50.001965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.294581Z digest=sha256:c9bb7d6af93e1c91f51ca07ffa07e71e8a66d866039b04563d4b57160a299d85

Observation 2e155071-68fc-4fc6-9c3c-0fb851ccc12d · outbound

This paper cites Visual autoregressive modeling: Scalable image gen- eration via next-scale prediction.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Visual autoregressive modeling: Scalable image gen- eration via next-scale prediction

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:49.985962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.299257Z digest=sha256:236492473a4fdb19cde1e2d6c6f74723e5953d67a777786d2869eb3e8725baf9

Observation 7c8a8970-2b8f-4d8b-8592-aa571bb076a8 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation LLaMA: Open and Efficient Foundation Language Models

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.303712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.303712Z digest=sha256:2298cd25f95ceb5d0c4c5310518af58f0b35c2efd2332e1ff5d2c7a43b4d7a9e

Observation b9840747-8159-45c9-b350-76927afe851f · outbound

This paper cites Edict: Exact diffusion inversion via coupled transformations.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Edict: Exact diffusion inversion via coupled transformations

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:49.970856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.308288Z digest=sha256:aebb5ecc976410a5237c9858738ae03788bec076a928727b1f5c1f08f3321207

Observation 7fb96bbf-4a92-49ee-a45e-24395db5a77a · outbound

This paper cites Explore In-Context Segmentation via Latent Diffusion Models.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Explore In-Context Segmentation via Latent Diffusion Models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.312831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.312831Z digest=sha256:cb5200ee4743d30ed8e0254a5d149e3b2d05e751167ca9bc4700fb3473db81ac

Observation c3e57cf5-eebc-4bb1-a4c6-e553d7fa561f · outbound

This paper cites Images speak in images: A generalist painter for in-context visual learning.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Images speak in images: A generalist painter for in-context visual learning

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.317769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.317769Z digest=sha256:c6b22461e80f53cbb00fcb3612962f1b189bb8e7d8ea96e37bd7ed0d2899e5ac

Observation 33a61fbc-7a86-473b-a7ac-a61bddb8f924 · outbound

This paper cites Seggpt: Segmenting ev- erything in context.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Seggpt: Segmenting ev- erything in context

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:49.945690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.322422Z digest=sha256:2b1b2e0c7a0cfe17c611de31ad118bcfa4968c2176216a2ac96ed6af308f94cd

Observation 2c0cfb8f-7b2a-4747-ad94-f0d3891f471d · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Emu3: Next-Token Prediction is All You Need

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.327105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.327105Z digest=sha256:4f99beaf68df374442af8b5499e502ebf36ba4ce3bcc8edadb00cbf199be06da

Observation 3999a648-31b7-4ea8-8165-4f370addb7fd · outbound

This paper cites In-context learning unlocked for diffu- sion models.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation In-context learning unlocked for diffu- sion models

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:49.929550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.332100Z digest=sha256:f7ae293e7510ff777a7e5a9be12445c88ee687db276f0e4845b07835f26a3fa1

Observation 8dd952f5-0926-4001-a046-19c665053bf9 · outbound

This paper cites The learn- ability of in-context learning.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation The learn- ability of in-context learning

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:49.912859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.336569Z digest=sha256:7d2453d69b2aef8ceb5d4d0080270e14db037972efbbbe20148be578f350d585

Observation 5566c6a6-e839-49b2-853b-3f927c3c0fa1 · outbound

This paper cites OmniGen: Unified Image Generation.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation OmniGen: Unified Image Generation

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.341022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.341022Z digest=sha256:e5a22852815410f6b6752aac8226ea7a32af74f1da819d34c899f978913e30b9

Observation d7961126-1edb-46f5-88ce-2ef177b4b95b · outbound

This paper cites Show-o: One Single Transformer to Unify Multimodal Understanding and Generation.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Show-o: One Single Transformer to Unify Multimodal Understanding and Generation

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.345818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.345818Z digest=sha256:bca43fe15ec2eac6a9fc2f35923c44c7efbbc0d8fb0f565dc24c3dc7866cbaf7

Observation 1e10c7fc-137b-467a-aea1-660224a9b7f0 · outbound

This paper cites To- wards global optimal visual in-context learning prompt se- lection.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation To- wards global optimal visual in-context learning prompt se- lection

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:49.897644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.350813Z digest=sha256:71ec804c6ed74cd60a96791fa48e571a377e342955d1647763b2e22fe270aba4

Observation a35469e6-92ec-4cb9-9e69-aa868b5fec94 · outbound

This paper cites Improv: Inpainting-based multimodal prompting for computer vision tasks.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Improv: Inpainting-based multimodal prompting for computer vision tasks

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:49.882072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.355291Z digest=sha256:64f84fa2475b82d5a58ca9f25c1243db88ce2c340371afb7c0f89425f0f99348

Observation 34e43b7a-a490-4f44-87bb-17f9d9b94885 · outbound

This paper cites Prompt-free diffusion: Taking” text” out of text-to-image diffusion models.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Prompt-free diffusion: Taking” text” out of text-to-image diffusion models

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:49.866525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.359799Z digest=sha256:ef17052ccca1d5b8624cf2d1f553ffb7d3b5f072b4a2c4153ef47e9f2a2ee567

Observation d883f07f-3acf-4ee2-83ae-83f1320fcc67 · outbound

This paper cites AnyGPT: Unified Multimodal LLM with Discrete Sequence Modeling.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation AnyGPT: Unified Multimodal LLM with Discrete Sequence Modeling

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.364318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.364318Z digest=sha256:012bd2f0f3b0800f350942fd03911d04d1d0c1d21323bf42075592b41ee4dd4b

Observation 409eb0b5-e82f-4c9b-91aa-d84bfe430e23 · outbound

This paper cites Magicbrush: a manually annotated dataset for instruction- guided image editing.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Magicbrush: a manually annotated dataset for instruction- guided image editing

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:49.851446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.369597Z digest=sha256:d90aacc39ca47c219bfbac57eb148aff57ff8983e743ba9b644d95e10116e8de

Observation 9b9caa51-1566-411b-a6cb-ca3cfa8ee9af · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Adding conditional control to text-to-image diffusion models

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.374025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.374025Z digest=sha256:057d3f7c6bb5e7bf26afc1b89ea8e975b059e5400fda25748a847cf6335e29cc

Observation 338b89d5-f98c-44bf-8b3f-3f69f603271a · outbound

This paper cites What makes good examples for visual in-context learning? In Proceed- ings of the 37th International Conference on Neural Infor- mation Processing Systems, pages 17773–17794, 2023.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation What makes good examples for visual in-context learning? In Proceed- ings of the 37th International Conference on Neural Infor- mation Processing Systems, pages 17773–17794, 2023

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:49.825502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.378517Z digest=sha256:8818c52b33631038298839cbbbb61b65d9bd6f133d536c80f4ab19660d42596a

Observation cc003e2c-a4eb-4b79-81de-68521e201cc2 · outbound

This paper cites InstructBrush: Learning Attention-based Instruction Optimization for Image Editing.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation InstructBrush: Learning Attention-based Instruction Optimization for Image Editing

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.383104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.383104Z digest=sha256:054ca6844fe40a1b92a52ddef627622e21c7cbf4fd993b38ae58ff9a7e748e5c

Observation 6e70062c-9bd1-4171-859c-82d4f202e702 · outbound

This paper cites Calibrate before use: Improving few-shot perfor- mance of language models.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Calibrate before use: Improving few-shot perfor- mance of language models

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:49.809961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.387913Z digest=sha256:da0822ee1ed66e1a01b0fbe2dcdcbec758212a785bd8f012c9956d35ebe356ab

Observation af8c770f-4c1f-4ab4-9229-7a16fef69820 · outbound

This paper cites Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.392554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.392554Z digest=sha256:c5687b9a21c0a6ce032660c9f3752a315778be16095eea991ec8109158366e6a

Observation 63076f3f-b888-4aae-8263-6352e5166f4c · outbound

This paper cites As a baseline of text-guided image editing model, InstructPix2Pix is trained only with textual instructions.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation As a baseline of text-guided image editing model, InstructPix2Pix is trained only with textual instructions

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:50:49.793577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-12T04:50:49.398206Z digest=sha256:0b56e4b49b8855b43eeba026a617b7334d513879796e58559c26ae525e9efc66

Pith citing papers

No inbound Pith citation observations are available.