Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T22:56:05.289689Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 34 inbound Pith citation observations for arXiv:2509.06949.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T22:56:05.289689Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T15:55:08.275883Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T06:09:37.869529Z
39 of 39 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c2a8deb7-0d77-46e3-8cc2-b40afd9c582d · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models Block Diffusion: Interpolating Between Autoregressive and Diffusion Language Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0c69cd2-ed52-462a-a63f-4119fccde7ff · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models Bayesian Flow Networks
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4340bfb6-a48f-448a-b299-b24b837850f7 · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models Measuring Mathematical Problem Solving With the MATH Dataset
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ecaace8-23fc-4aec-b931-c867dd806537 · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models V-STaR: Training Verifiers for Self-Taught Reasoners
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4696d5f7-33fe-4330-bebf-3a806a2c5e95 · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 647c8c9a-07b9-4908-a2a5-e37755e586de · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models INTELLECT-1 Technical Report
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1af97e70-eab5-4919-ad3d-1ca9916da1f9 · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dee62877-ebc6-4646-bf44-2f95bb3a51bc · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models DeepRetrieval: Hacking Real Search Engines and Retrievers with Large Language Models via Reinforcement Learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2c29eb9-6039-497f-ad79-1f82578792ca · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1553b1ed-37af-4d21-b099-ebc5fba74812 · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models Train for the Worst, Plan for the Best: Understanding Token Ordering in Masked Diffusions
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58fa5b87-43e5-4959-93d9-2960109e2dcf · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models Mercury: Ultra-Fast Language Models Based on Diffusion
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eac7d5c2-c31f-4839-9d9c-7fa83ddffec4 · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models dLLM-Cache: Accelerating Diffusion Large Language Models with Adaptive Caching
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 713aadc2-9b0a-4e00-9321-87bea02c3a44 · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models dKV-Cache: The Cache for Diffusion Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 344c9699-b86a-4412-8cf5-2504ceb5806e · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models American invitational mathematics examination (aime) 2024: Aime i and aime ii.https://artofproblemsolving
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4c23a8d8-c948-4388-9fd0-1cde1550a3f6 · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models Large Language Diffusion Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e041e023-42bc-4b16-955d-fe8ba2f99933 · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models Your Absorbing Discrete Diffusion Secretly Models the Conditional Distributions of Clean Data
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7da51dd-ee38-40f6-b685-b570142d84fb · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models Proximal Policy Optimization Algorithms
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20ffdb78-e96d-4f0c-89a0-4880255a19d7 · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10d7909a-b5b7-4a44-a7c6-c5c7b603b704 · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 371879e1-9ea8-4d9b-9777-8da6103f9564 · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models Diffusion LLMs Can Do Faster-Than-AR Inference via Discrete Diffusion Forcing
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e7076c7-a6bf-4429-82d7-f5fd28c52b27 · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models LiveBench: A Challenging, Contamination-Limited LLM Benchmark
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64ae0e81-ac51-4fe7-a8eb-db8b1d41408f · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models Fast-dLLM: Training-free Acceleration of Diffusion LLM by Enabling KV Cache and Parallel Decoding
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7c5c330-352a-41e5-84c7-7db996ebe3fb · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models Qwen2.5 Technical Report
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a757482-b153-445f-95f8-21893f928c59 · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 092e5c71-68a6-4abd-9416-eda5d18239af · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebc6e2f6-74e0-4187-9043-3b8e52f5597b · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models Masked Diffusion Models are Secretly Time-Agnostic Masked Models and Exploit Inaccurate Categorical Sampling
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8398fa86-6311-4e9b-a738-e71432edd0fa · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models LLaDA 1.5: Variance-Reduced Preference Optimization for Large Language Diffusion Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 348caf80-0e6b-4a87-bdbc-457912b5e8d1 · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models Unresolved cited work
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6fde3fa0-fa9b-48e8-b17a-cf899a2460d2 · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models thinking
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 13e879c2-e038-4959-bf51-e42aa2909e15 · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models With dynamic sampling, we use a threshold ofT= 0.9 and 𝑡𝑜𝑝-𝑘= 0(i.e., all tokens are kept)
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b8909fce-c08c-4161-96bd-62086e9b73dc · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models By default, we use the𝑘= 3estimator for KL
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e205da69-4093-43b4-91d4-11fe07d7323d · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models We employ static decoding (one token per step) to enhance sampling quality (Gong et al., 2025), using the KV-cache
Reference 1024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0cee651a-fb40-4cc9-96d1-544b5c83be0c · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models Seed Diffusion: A Large-Scale Diffusion Language Model with High-Speed Inference
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72e53bf8-eeb4-4667-a74c-e66d14e01e9b · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models Unresolved cited work
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4632660b-9346-483b-ab34-4df096a97b8c · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models Continuous diffusion for categorical data
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d8dbddb-7f74-404b-85a6-026e071ed9c2 · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models URLhttps://www.science.org/ doi/10.1126/science.abq1158
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb933275-dd44-4bfc-9760-1cc1a8216f72 · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c30401bf-c138-45b5-9c2b-ee98055d6589 · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models DiffuCoder: Understanding and Improving Masked Diffusion Models for Code Generation
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19897308-9302-432a-b9e3-e32c4baf563e · outbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models Training Verifiers to Solve Math Word Problems
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37d9448c-a325-4dd9-863d-b4d6951ed9f1 · inbound
d2: Improving Reasoning in Diffusion Language Models via Trajectory Likelihood Estimation Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 800ad2a0-9efb-41ed-8a19-d5fddc32a4b2 · inbound
Enhancing Reasoning for Diffusion LLMs via Distribution Matching Policy Optimization Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf21409b-dd4a-47fa-8e82-8af6c646eade · inbound
T$^\star$: Progressive Block Scaling for Masked Diffusion Language Models Through Trajectory Aware Reinforcement Learning Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aac8d34a-7ed5-4678-9637-f856fbdece51 · inbound
The Flexibility Trap: Rethinking the Value of Arbitrary Order in Diffusion Language Models Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f38b6adf-8604-426a-9c2a-df30ef681b13 · inbound
FlashBlock: Attention Caching for Efficient Long-Context Block Diffusion Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f63daa97-6757-4967-94a6-cb94ef78eec6 · inbound
DICE: Diffusion Large Language Models Excel at Generating CUDA Kernels Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1841513-5677-477d-bdea-07091b1cae5d · inbound
Improving Sampling for Masked Diffusion Models via Information Gain Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8ab909dd-c28a-4395-b2ce-ff9ad4979467 · inbound
MemDLM: Memory-Enhanced DLM Training Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3562d992-73ec-4496-a18f-477cfc4c09da · inbound
TAD: Temporal-Aware Trajectory Self-Distillation for Fast and Accurate Diffusion LLM Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation d9e9c51b-a69c-4215-9b40-43c49577bfe8 · inbound
Relative Score Policy Optimization for Diffusion Language Models Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation dfda3f9e-43ee-47d2-9768-3826aefcf14a · inbound
Beyond Mode-Seeking RL: Trajectory-Balance Post-Training for Diffusion Language Models Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 9f028af6-4e13-4f28-b5e5-57169d77a00d · inbound
Sketch Then Paint: Hierarchical Reinforcement Learning for Diffusion Multi-Modal Large Language Models Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4931ca3e-ecc6-4b60-965e-d6d12050ac79 · inbound
Beyond Execution: Static-Analysis Rewards and Hint-Conditioned Diffusion RL for Code Generation Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1980fc2b-a08b-45dd-8dd2-795e2127eb22 · inbound
Fast-dDrive: Efficient Block-Diffusion VLM for Autonomous Driving Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation cddd6195-278a-4aed-a8f2-e4f8e212b7d3 · inbound
Fast-dDrive: Efficient Block-Diffusion VLM for Autonomous Driving Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 880e18e9-98e7-4abb-989f-c41f6c372eb3 · inbound
GDSD: Reinforcement Learning as Guided Denoiser Self-Distillation for Diffusion Language Models Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 811ec5d0-7205-45ae-890d-c281b6f09ed5 · inbound
Dynamic Infilling Anchors for Format-Constrained Generation in Diffusion Large Language Models Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3ad659e2-0aa5-4651-ac0b-b97de6064121 · inbound
Data-Efficient Autoregressive-to-Diffusion Language Models via On-Policy Distillation Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 971d982d-f78e-4362-af0e-ad8f44d0ae51 · inbound
Back on Track: Aligning Rewards and States for Reasoning in Diffusion Large Language Models Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation fdc14c71-1bd0-4457-993a-67d7a2af8218 · inbound
Unified Energy for Invariant and Independent Decoding in Diffusion Language Models Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3cd60031-8723-4cb0-affd-42cf8408f7dd · inbound
Beyond Fully Random Masking: Attention-Guided Denoising and Optimization for Diffusion Language Models Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b9a99040-ad11-47d1-892a-994b3be35b0b · inbound
VoidPadding: Let [VOID] Handle Padding in Masked Diffusion Language Models so that [EOS] Can Focus on Semantic Termination Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 62592c28-83d3-4cf7-9bac-dcbd43f62984 · inbound
Learning from the Self-future: On-policy Self-distillation for dLLMs Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 6de3e9a3-28d4-46cf-8bc3-e0ac677103c8 · inbound
HERALD: High-Throughput Block Diffusion LLM Serving via CPU-GPU Cooperative KV Cache Retrieval Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 2c5da919-983f-4db6-8b96-c090504ae5e5 · inbound
HERALD: High-Throughput Block Diffusion LLM Serving via CPU-GPU Cooperative KV Cache Retrieval Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bb9e1f9-9190-4dc1-8486-a5d7db9ce489 · inbound
Diffusion-GR2: Diffusion Generative Reasoning Re-ranker Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 75236aba-2ff5-4c7b-a001-513c1cb908a1 · inbound
Diffusion-GR2: Diffusion Generative Reasoning Re-ranker Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2de102df-3f46-41d4-b9f5-ed8e3a9f5c7b · inbound
Diffusion-GR2: Diffusion Generative Reasoning Re-ranker Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a5f90f7-0a02-4f5b-87fb-ee8d3ad77b09 · inbound
Spec-AUF: Accept-Until-Fail Training under Train-Inference Misalignment for Masked Block Drafters Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 32c74cee-776c-4044-b59d-3d4ef58b3e0f · inbound
Layer-Parallel Inference Reduces Encrypted Nonlinear Depth in Transformers Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0dbfbf5d-d8ae-4709-b6ff-7c4c11ad38b7 · inbound
Layer-Parallel Inference Reduces Encrypted Nonlinear Depth in Transformers Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7c9d9ec-48c9-4d3a-a49b-485c2a6f1130 · inbound
Trace-Based On-Policy Distillation for Masked Diffusion Language Models Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation add924f4-5a95-4b2d-a03c-b65f6aa1c66d · inbound
AdaFlash: Adaptive Speculative Decoding via On-Policy Distilled Diffusion Drafters Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2acfa3f7-fa86-4d7b-a71f-e7ace4c56226 · inbound
From Interface to Inference: Eliciting Any-Order Inference from Any-Order Models Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.