Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:09:51.624919Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 1 inbound Pith citation observation for arXiv:2505.19874.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:09:51.624919Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-28T15:39:15.744954Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-01T22:16:15.736010Z
54 of 54 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation fca194cd-a8c3-4df0-bf7a-6449a1491836 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Flamingo: a visual language model for few-shot learning.Advances in neural information processing systems, 35:23716–23736, 2022
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 553fa905-4ac6-48b9-b1e9-47d505f6ce20 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Training Diffusion Models with Reinforcement Learning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3565bfa3-3c4d-4b1d-9819-e73f4624deba · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ac5ac4d-1524-4b64-9514-e256ac64affa · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation ANOLE: An Open, Autoregressive, Native Large Multimodal Models for Interleaved Image-Text Generation
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2aa0bf26-a6e9-4d30-973b-3ca8139ac5df · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Scaling rectified flow transformers for high-resolution image synthesis
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d85b30b-5d56-475e-99e7-0f8838bcd45d · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Taming transformers for high-resolution image synthesis
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52a0af3a-fa36-40b7-9763-9d6f62c03901 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60228cfa-672a-4918-8fd0-757974f8416a · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Dpok: Reinforcement learning for fine-tuning text-to-image diffusion models.Advances in Neural Information Processing Systems, 36:79858–79885, 2023
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f88e1f5a-04fb-407b-8b98-d43ede3e84eb · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation StyleShot: A Snapshot on Any Style
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eee7be35-5f74-40fc-ae14-402c4b3ccf6e · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Style aligned image generation via shared attention
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8d6ae7e1-01ae-43f3-aad5-da71b25ca712 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Lora: Low-rank adaptation of large language models.ICLR, 1(2):3, 2022
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9a81883-174d-4c7a-857b-a909adbb5bb3 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation ArtCrafter: Text-Image Aligning Style Transfer via Embedding Reframing
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0c6081a-9917-4462-95bb-3fa18a84a8df · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Perceiver: General perception with iterative attention
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7852db0d-bada-4321-8f55-cfaa4444fbcf · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Visual Style Prompting with Swapping Self-Attention
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fafef86f-84bd-4705-9e7a-981bb7cf327e · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Segment anything
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d1eaeae-a419-4d3f-aceb-f0793953549b · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Flux.https://github.com/black-forest-labs/flux, 2024
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62cec599-c947-4a54-b4b7-ad3b636aa6b3 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ab2158d-e721-4a63-b636-481678a58907 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation StyleCrafter: Enhancing Stylized Text-to-Video Generation with Style Adapter
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 223b7d21-e690-430f-b59f-43440290ec1b · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Subject- driven text-to-image generation via preference-based reinforcement learning.Advances in Neural Informa- tion Processing Systems, 37:123563–123591, 2024
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2b3f036d-3edd-4154-8dd7-c4506dfa1e04 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Null-text inversion for editing real images using guided diffusion models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b130d34c-c462-45db-b785-95a970aac38a · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation laion2b-en-aesthetic-square-cleaned
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 911f2656-fcfa-4452-8d16-551ae59a4052 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Training language models to follow instructions with human feedback.Advances in neural information processing systems, 35:27730–27744, 2022
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d50eb054-58ba-48c6-9ad2-6e456181a4b2 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5f0034e-e5c8-4564-985c-d8494719ec4f · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Learning transferable visual models from natural language supervision
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6c746c1-c854-42f7-ba74-323803890a41 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Direct preference optimization: Your language model is secretly a reward model.Advances in Neural Information Processing Systems, 36:53728–53741, 2023
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14e3aa5a-d94c-463c-9809-a11f8e8f719c · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation High-resolution image synthesis with latent diffusion models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3159c19-a960-4ef4-ab87-66015d4ed7c9 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Dream- booth: Fine tuning text-to-image diffusion models for subject-driven generation
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26cfabbe-36b6-43a5-aac6-36c58d806951 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Laion-5b: An open large-scale dataset for training next generation image-text models.Advances in neural information processing systems, 35:25278–25294, 2022
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70dba0c1-a7db-4ba6-8332-8b89f016be9b · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f323023-1822-4b25-992d-026a06635e34 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation StyleDrop: Text-to-Image Generation in Any Style
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b31ddba-8737-42d6-ba19-0449dca3b32a · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Personalized Text-to-Image Generation with Auto-Regressive Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d910e862-9f95-4dc7-bcee-fe054e31eabb · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07711783-a8f8-499c-9f8f-d8908afdaad4 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation HART: Efficient Visual Generation with Hybrid Autoregressive Transformer
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3bc4e6a-cc68-4997-a6ef-2975ad03ecab · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Chameleon: Mixed-Modal Early-Fusion Foundation Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32c66ba3-57c9-4e13-90f4-cf92abc8c600 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Neural discrete representation learning.Advances in neural information processing systems, 30, 2017
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e61c2d1-a85a-4246-8f3f-9c48c543c9d3 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Attention is all you need.Advances in neural information processing systems, 30, 2017
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43711ad9-0f2f-4974-a897-03930acc6342 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Diffusion model alignment using direct preference optimization
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d25ef643-909f-4c41-8b38-62c3b186eafd · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation InstantStyle: Free Lunch towards Style-Preserving in Text-to-Image Generation
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 167101e9-00ba-48f5-916f-a737441f8334 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation SimpleAR: Pushing the Frontier of Autoregressive Visual Generation through Pretraining, SFT, and RL
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 673a9799-3486-4b31-a59a-40c1f88c89bb · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Emu3: Next-Token Prediction is All You Need
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1837bbb7-6f50-4586-896c-b6370f689afc · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation StyleAdapter: A Unified Stylized Image Generation Model
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58afe69c-7688-4c7d-b860-5502e679028b · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation wikiart.https://huggingface.co/datasets/huggan/wikiart, 2022
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ddb66204-867c-4349-9b57-c15e7ea53cd0 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Infinite-id: Identity-preserved personaliza- tion via id-semantics decoupling paradigm
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cb0ad2dd-95d5-4dae-adb5-69ad8dbe9f34 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Proxy-tuning: Tailoring multimodal autoregressive models for subject-driven image generation.arXiv preprint arXiv:2503.10125, 2025
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e867ddc8-17ba-4b2c-a725-eb741b1236ca · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation OmniGen: Unified Image Generation
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a394c5d-423a-47e0-b993-212d765cccb3 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Show-o: One Single Transformer to Unify Multimodal Understanding and Generation
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3bcf9ed-255f-4ce8-ac85-953facd9d194 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Imagereward: Learning and evaluating human preferences for text-to-image generation.Advances in Neural Information Processing Systems, 36:15903–15935, 2023
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afb59c33-aab3-486e-b7ea-ddd39bb7179f · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation DanceGRPO: Unleashing GRPO on Visual Generation
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 423cda56-fd6a-4788-9cc6-8d764dc2b684 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Qwen2 Technical Report
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0149abf8-906e-4d68-a893-d6b5293125f1 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Qwen2.5 Technical Report
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74ff1e8c-efe5-446a-992d-dc056d8f6b86 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52a76abc-8b38-4013-b333-46543c7ce459 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Vector-quantized Image Modeling with Improved VQGAN
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08a76191-56d2-44a6-bd9d-86c777b74e32 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation DINO: DETR with Improved DeNoising Anchor Boxes for End-to-End Object Detection
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe1ac081-b2d4-4c35-bc7d-77d0e8557472 · outbound
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ef21c97-a801-42c5-9cf6-f5d435d29a09 · inbound
Residual Decoder Adapter: ID-Preserving Tokenizer Adaption for Autoregressive Text Rendering StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.