Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-31T02:16:08.275374Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 100 of 104 outbound references and 0 inbound Pith citation observations for arXiv:2607.28611.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-31T02:16:08.275374Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
100 of 104 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation b3f1285a-a8d2-4f8f-a4b9-b9dcc1c0ed75 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Gqa: Training generalizedmulti-querytransformermodelsfrommulti-headcheckpoints
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96695258-4f06-4a0b-a373-f9ea6da9bf1f · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Physics of language models: Part 4.1, architecture design and the magic of canon layers.Advances in Neural Information Processing Systems, 38:42349–42369, 2026
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d07a87d-a406-4eff-8c04-eb02c23dadc6 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Round and Round We Go! What makes Rotary Positional Encodings useful?
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a737b48e-98d1-426c-8c17-1cd3eae2f01d · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers FLUX.2: Frontier Visual Intelligence.https://bfl.ai/blog/flux-2, 2025
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68d82e32-8e07-4210-894f-267e0498a709 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29be113c-67ae-43d7-911f-a3b992baf929 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Align your latents: High-resolution video synthesis with latent diffusion models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d331e3b-93dd-4ff1-86b7-504d7c3dd1c5 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Video generation models as world simulators, 2024
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30e953e9-3eaf-4b06-8b0d-d25ac09c1202 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef575262-d288-4069-b27f-0b5ed773a300 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers The Effect of Mini-Batch Noise on the Implicit Bias of Adam
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fab3edf1-56a2-4e90-9f61-a2f66ca0cb59 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Aligning visual foundation encoders to tokenizers for diffusion models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36fa5535-064a-4f56-bae4-85e122f790c2 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Diffusion forcing: Next-token prediction meets full-sequence diffusion.Advances in Neural Information Processing Systems, 37:24081–24125, 2024
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3940baf-8bf9-44d4-a4d8-f8dc42210c58 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Pixart-alpha: Fast training of diffusion transformer for photorealistic text-to-image synthesis
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23fc2699-5ddc-4f6c-8ed8-89242d511d5b · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Sana-video: Efficient video generation with block linear diffusion transformer.arXiv preprint arXiv:2509.24695, 2025
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 953987e6-b508-4a14-8bd5-ddcd18846e6f · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Generating Long Sequences with Sparse Transformers
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efe1d071-f925-4e60-ae13-7cf14f5bb61f · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Conditional positional encodings for vision transformers
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8674616c-2509-4dbd-903f-49993d1da42a · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers UniMax: Fairer and more Effective Language Sampling for Large-Scale Multilingual Pretraining
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edcd3dbe-525e-4544-9742-04d20e705ff7 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Deepseekmoe: Towards ultimate expert specialization in mixture-of-experts language models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9116ccc-38c1-471e-9fd3-d8c5c2c9643f · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Don’t be lazy: Completep enables compute-efficient deep transformers.Advances in Neural Information Processing Systems, 38:137707–137739, 2026
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac60a4be-a081-44c8-b2a8-03848e0a3802 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Diffusion models beat gans on image synthesis.Advances in neural information processing systems, 34:8780–8794, 2021
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99b7920e-3e6b-4a0c-93f4-60cd195bd1c6 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Diffusion is spectral autoregression.https://sander.ai/2024/09/02/spectral-autoregression
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ed05fa7-6588-4b8f-a38b-f139a7e59178 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Efficient-vDiT: Efficient Video Diffusion Transformers With Attention Tile
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28838b84-62cb-4cdc-8185-7ac6cda2c1b3 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Cogview: Mastering text-to-image generation via transformers
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff79909b-1e87-4091-a68a-c00474cf92a1 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers A mathematical framework for transformer circuits.Transformer Circuits Thread, 1(1):12, 2021
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 671d716e-8ae5-4088-a419-c54af8afa06b · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Scaling rectified flow transformers for high-resolution image synthesis
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b81ca4ad-de40-42c4-b0ed-2f7a2faeed45 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Switch transformers: Scaling to trillion parameter models with simple and efficient sparsity.Journal of Machine Learning Research, 23(120):1–39, 2022
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0ee9087-c842-4324-ba7c-9f8d3087430b · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Scaling diffusion transformers to 16 billion parameters, 2024.URL https://arxiv
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 446d8ba5-35f7-4ccb-9656-986942f894cd · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Why Adam Works Better with $\beta_1 = \beta_2$: The Missing Gradient Scale Invariance Principle
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ce817d8-e67d-46df-9eb4-370966ed2a94 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Geneval: Anobject-focusedframeworkforevaluatingtext-to-image alignment.Advances in Neural Information Processing Systems, 36:52132–52152, 2023
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 671749fa-901a-4d51-ae7b-6987ce5f36a0 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Mamba: Linear-Time Sequence Modeling with Selective State Spaces
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b633e07-6b12-4c58-ac5e-566d2fb7de99 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Scaling Laws for Autoregressive Generative Modeling
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6986aedf-9bec-42d7-ad5b-7ffa62020ef6 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Deep Learning Scaling is Predictable, Empirically
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7e826fe-f2ee-40ab-a17b-119ab4738c20 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Classifier-Free Diffusion Guidance
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d368d901-0493-44a1-b018-5a816f0e4934 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Denoising diffusion probabilistic models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58a6f19e-8edf-4053-9290-88b2b7f56191 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Imagen Video: High Definition Video Generation with Diffusion Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ed24407-67a4-49d0-aaed-2698296a6679 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Videodiffusionmodels
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f224e44-c918-49f2-8d50-51c96bbc222f · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Training Compute-Optimal Large Language Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de8a3529-9465-4022-873d-0a7b8af82577 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Pushing the Boundaries of State Space Models for Image and Video Generation
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 907bfc44-2e6c-438d-b7cf-8da6a3be8b97 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Relic: Interactive video world model with long-horizon memory.arXiv preprint arXiv:2512.04040, 2025
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 854dee7c-e799-4fb1-9465-644bbdb961fd · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Zigma: A dit-style zigzag mamba diffusion model
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 347a662c-4aac-44ba-9bb8-68c2c9a11862 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0b02c73-1eb9-4353-861d-530ca9f69023 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Self forcing: Bridging the train-test gap in autoregressive video diffusion.Advances in Neural Information Processing Systems, 38:167283–167308, 2026
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afbafc47-653c-4a81-99c9-8a190606c0a9 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Muon: An optimizer for hidden layers in neural networks, 2024.URL https://kellerjordan
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a9554c0-a78e-4a22-ac5f-0dab9b9d65ca · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers EditVerse: Unifying Image and Video Editing and Generation with In-Context Learning
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43f4c58a-6810-4bfe-adda-4ece00621794 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Scaling Laws for Neural Language Models
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8164523-4a2b-4b2f-b976-2d7b763f56dc · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Transformers are RNNs: Fast autoregressive transformers with linear attention
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59fdf7c9-40b3-47c4-8afd-de47f0065aa7 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Theimpactofpositional encoding on length generalization in transformers.Advances in Neural Information Processing Systems, 36:24892–24928, 2023
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 687323ae-9ce3-4040-b70f-896365421064 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers GShard: Scaling Giant Models with Conditional Computation and Automatic Sharding
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9cff701e-1901-46ce-971c-f471979af551 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Back to basics: Let denoising generative models denoise
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1599da00-1c66-4942-8bfc-6092dc53883f · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Scalinglawsfordiffusiontransformers.arXivpreprintarXiv:2410.08184, 2024
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fe38c64-c4d4-47a9-92b6-d6260e640913 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Flow Matching for Generative Modeling
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c800bfc-44b0-41b5-ad7e-75edf89c5063 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab65fb43-f2e6-4737-b465-bf4a20dafbb7 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers DeepSeek-V3 Technical Report
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d9fb7bd-8261-4275-adfe-c6333a4ba733 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92036d02-b64c-4ea5-a004-471275f9a9ca · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bae04a0-1abc-43d3-8ca3-84f29ebf7d69 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Vmamba: Visual state space model.Advances in neural information processing systems, 37:103031–103063, 2024
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12de7f78-7b2e-4d10-a181-28fc38675a27 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Olmo Hybrid: From Theory to Practice and Back
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b914026-664f-43da-91dc-1397d6ac4913 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Completed hyperparameter transfer across modules, width, depth, batch and duration.arXiv preprint arXiv:2512.22382, 2025
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 212799eb-a387-45c3-a07e-2dd1616da5c8 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d656b73-33cd-494f-88f9-5966961b6d29 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Improveddenoisingdiffusionprobabilisticmodels
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8151033e-a9de-42d0-b2bd-8c0c052c7735 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers In-context Learning and Induction Heads
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a8e47bb-4377-48bb-9e15-e984c14596f8 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers In search of adam’s secret sauce.Advances in Neural Information Processing Systems, 38: 63404–63442, 2026
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c7a5ff3-66bd-4dc0-a055-d7724dd27987 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Scalable diffusion models with transformers
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42be2ed3-3502-4058-a0a4-6253d0ad9e89 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Smith, and Mike Lewis
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efdae082-a1d8-4568-982f-8cfc99517b03 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Exploring the limits of transfer learning with a unified text-to-text transformer.Journal of machine learning research, 21 (140):1–67, 2020
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f165b87e-08ef-49b9-aa6f-2f217c342229 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Hierarchical Text-Conditional Image Generation with CLIP Latents
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cd5d4c0-51d7-434e-8f5f-2b421631129c · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers High-resolution image synthesis with latent diffusion models
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation acc1cb7e-8e03-4849-8d53-eeb3e2d3cca8 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Photorealistic text-to-image diffusion models with deep language understanding.Advances in neural information processing systems, 35:36479–36494, 2022
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0493f57c-70eb-42ea-96f8-66c62baa9b60 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Fast Transformer Decoding: One Write-Head is All You Need
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4c7ed8b-d3a9-4ee9-989f-d43470775374 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers GLU Variants Improve Transformer
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ca1d191-b33b-4f77-b079-4ad0d40e6e08 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Make-A-Video: Text-to-Video Generation without Text-Video Data
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89d519d2-650e-482b-86ac-a7cfc9d78284 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers History-Guided Video Diffusion
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 448d3f28-3376-4196-bdc9-3f608c766d96 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Score-Based Generative Modeling through Stochastic Differential Equations
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94342501-a152-46bb-90b4-b7f4cc4a6ec6 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Roformer: Enhanced transformer with rotary position embedding.Neurocomputing, 568:127063, 2024
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ee61b72-b472-43af-b6ea-35b1efb7a054 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Retentive Network: A Successor to Transformer for Large Language Models
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e51f221d-5cc8-46df-9a7a-a6308546798d · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Dsv: Exploiting dynamic sparsity to accelerate large-scale video dit training.arXiv preprint arXiv:2502.07590, 2025
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db331e8e-1600-4fca-90d1-71d58df648a2 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Kimi k3: Open frontier intelligence, 2026.https://arxiv.org/abs/2607.24653
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cee43a81-79fd-4491-bbc7-17946f0ff2af · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Kimi Linear: An Expressive, Efficient Attention Architecture
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69380636-d526-4efb-baa6-b7f03c974c2a · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Qwen3.5-Omni Technical Report
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7183d94-e7ab-4d8d-bdba-4e77f9bd0d92 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Hunyuanvideo 1.5 technical report, 2025.https://arxiv.org/abs/2511
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18c04200-4605-452c-ba69-7a8fc7bf8e82 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Gomez, Lukasz Kaiser, and Illia Polosukhin
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94a2fd28-715b-4fce-a1cb-909bf7fdcc05 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Wan: Open and Advanced Large-Scale Video Generative Models
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c8f58a3-e9c7-45a3-bd6f-16f0cc273bbd · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Lingen: Towards high-resolution minute-length text-to-video generation with linear computational complexity
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b71a8927-3d4a-4512-9cbf-46e288aecc79 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Auxiliary-Loss-Free Load Balancing Strategy for Mixture-of-Experts
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 945ca130-8d03-433a-b5ba-12e39e89a6d8 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5afe614a-3d6a-4bdf-8a7a-c838d0837cf8 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Scaling laws, carefully.lilianweng.github.io, June 2026
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18181547-990b-45f6-9678-de8b78d31882 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Qwen-Image Technical Report
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f634608-880a-4d02-b295-6ed5f131e32f · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers VMoBA: Mixture-of-Block Attention for Video Diffusion Models
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f33261d-732d-49fe-878b-39018db1bfa9 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d45c22b-788f-44f7-9d0b-50113234b865 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers mHC: Manifold-Constrained Hyper-Connections
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63a96bc9-3167-463c-91a2-d6803eb34016 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Qwen3 Technical Report
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87e6a1ac-b78f-4ea7-964a-0246e98ecbed · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Tuning large neural networks via zero-shot hyperparameter transfer.Advances in Neural Information Processing Systems, 34:17084–17097, 2021
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9cc9133-94d9-4698-a52f-d1cec4bca3cb · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e73d0fd-b071-4be4-b8bc-1c395b8298f9 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Parallelizing linear transformers with the delta rule over sequence length.Advances in neural information processing systems, 37:115491–115522, 2024
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9828ea82-a305-4572-b5ab-f0574a808a06 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Cogvideox: Text-to-videodiffusionmodelswithanexperttransformer
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2002f6bb-7974-458f-918f-692d43cc7710 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Pixeldit: Pixel diffusion transformers for image generation
Reference 96
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55100718-4c93-403c-94ee-f06d3b4153f6 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Native sparse attention: Hardware-aligned and natively trainable sparse attention
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d12d853a-dc80-4086-a18f-c623c53a8907 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Bidirectional sparse attention for faster video diffusion training.arXiv preprint arXiv:2509.01085, 2025
Reference 98
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc1da28f-177a-49d9-83a5-abc89bf29834 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers KnapFormer: An Online Load Balancer for Efficient Diffusion Transformers Training
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a50d4019-bb90-4f5f-a6af-cf9dcc048071 · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Faster video diffusion with trainable sparse attention.Advances in Neural Information Processing Systems, 38:152509–152534, 2026
Reference 100
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1480be5-84c7-4953-8ea8-a5de0a860c4a · outbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers PyTorch FSDP: Experiences on Scaling Fully Sharded Data Parallel
Reference 101
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.