Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:40:17.599925Z
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 80 of 80 outbound references and 66 inbound Pith citation observations for arXiv:2504.12216.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:40:17.599925Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T15:55:08.290104Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-11T03:07:51.770001Z
80 of 80 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 6381cf53-f991-4ef0-baa5-a36fd6f69ce0 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41c95ad5-111b-4485-81bc-33061805b29a · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Back to Basics: Revisiting REINFORCE Style Optimization for Learning from Human Feedback in LLMs
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 297a88e9-9bd5-477d-98fe-c6c2d3802edc · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Arel’s sudoku generator
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 86840d8a-b479-4822-9e47-bb87b75357c5 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Block Diffusion: Interpolating Between Autoregressive and Diffusion Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d43724d7-8f52-4286-9629-296b750408cb · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Structured denoising diffusion models in discrete state-spaces.Advances in neural information processing systems, 34:17981–17993, 2021
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b7b51c4-446b-4aad-80fc-5994154c89ca · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Program Synthesis with Large Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09327b48-f87b-4b02-bf27-8157489bdb79 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4954ca3-bbd5-4b19-b915-7a7aa415bb3f · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Evaluating Large Language Models Trained on Code
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f95a0137-19cb-4991-950d-80541c9d01e1 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97f0e569-eb63-4dda-bd71-a0c032d4eeac · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Training Verifiers to Solve Math Word Problems
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2cb5f6a-76de-4c37-b3de-309b0729e9a2 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning FlashAttention-2: Faster attention with better parallelism and work partitioning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9684b0a4-13b4-4edd-a669-0681b0806c74 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning BERT: Pre-training of deep bidirectional transformers for language understanding
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 82c4f524-003b-486f-9924-86f0bb52b705 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning The Llama 3 Herd of Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aed3e921-9a1d-4c03-862b-3803fd15b4a2 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 338619f6-e7ac-42a5-81ad-0fbd1ae6d17f · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Scaling diffusion language models via adaptation from autoregressive models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47a4dae1-f08e-4f1b-9533-9045010858d8 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Likelihood-based diffusion language models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6723fde0-02a3-4090-b4b3-83ca4bfc4281 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f96cbd07-d1ee-432f-a643-bb3b6a750f4c · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Measuring Mathematical Problem Solving With the MATH Dataset
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a042447-9413-4a95-80e9-4bc6460c3dbb · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Denoising diffusion probabilistic models.Advances in neural information processing systems, 33:6840–6851, 2020
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e943232-8d48-4f25-8838-9ed52417f60c · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Mercury: Ultra-fast language models based on diffusion
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 8bc39d61-55e2-47c6-9d32-43e6ee4e52ba · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Numina- math
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e0aff48d-c51c-4efb-a865-de44583af816 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning ReMax: A Simple, Effective, and Efficient Reinforcement Learning Method for Aligning Large Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4662a7e3-c73c-4f43-ae7b-26fad1e489d1 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Let's Verify Step by Step
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 466fcac1-6e1e-4234-90bb-51e8681ccaf1 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Understanding R1-Zero-Like Training: A Critical Perspective
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20a14b4c-7804-4479-8283-411194561ef7 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Decoupled Weight Decay Regularization
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a23b65a1-729d-40bd-96a5-207668787007 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Discrete diffusion modeling by estimating the ratios of the data distribution
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 04303571-57fc-47ca-9143-f04c6c1651cd · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Dynamic Scaling of Unit Tests for Code Reward Modeling
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1445326c-c63e-4601-b6c7-63e2906dc7c6 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning s1: Simple test-time scaling
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3bd84b3-ae33-47d4-990d-bb405721cc1d · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Scaling up Masked Diffusion Models on Text
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a713c4a-70f5-4a54-b072-aad0428fabd5 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Large Language Diffusion Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfd2c49f-b980-4ede-97db-647bb01a1c44 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Learning to reason with llms, September 2024
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b981c54-9763-40bd-b3ed-f1940f08ed31 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Your Absorbing Discrete Diffusion Secretly Models the Conditional Distributions of Clean Data
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fafa3030-0956-46d6-a3a2-0a8aaf9dbab7 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Training language models to follow instructions with human feedback.Advances in neural information processing systems, 35:27730–27744, 2022
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4aa5b1c8-2d52-49b3-a03b-e8a2f8c4bf99 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Tinyzero
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 661f411b-b526-4cad-877a-8e35b2d83515 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Openwebmath: An open dataset of high-quality mathematical web text, 2023
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6df3e1f3-6268-4ba4-8396-cf36daeec210 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Simple and effective masked diffusion language models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5953b8ad-d60c-445f-9f4d-5a6865cacb27 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Proximal Policy Optimization Algorithms
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 035ddd47-a8d0-417b-92e9-b5b59e4ac685 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 760fa881-1ade-4537-9e13-5d4cf4cc6016 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Simplified and generalized masked diffusion for discrete data.Advances in neural information processing systems, 37:103131–103167, 2024
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05d97217-6205-4546-9c86-32c8a3731329 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Score-based generative modeling through stochastic differential equations
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf36f640-7466-4da4-874c-361fe70e8214 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2363d779-efe3-49f6-880d-434be0f2981b · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Open Thoughts
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ff689cb-05ca-4bf3-8696-3e672aa2b41a · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Trl: Transformer reinforce- ment learning.https://github.com/huggingface/trl, 2020
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e7838f69-76e7-4f63-a88c-d82646333a98 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Simple statistical gradient-following algorithms for connectionist reinforce- ment learning.Machine learning, 8:229–256, 1992
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 006348a4-84bb-4d76-a353-8d1b19d0fe37 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning KodCode: A Diverse, Challenging, and Verifiable Synthetic Dataset for Coding
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de7b6fb5-3b49-4fca-a6c4-e5159b94a4ec · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Beyond Autoregression: Discrete Diffusion for Complex Reasoning and Planning
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac86e258-44d0-4dd0-8dfd-b4ceb5f9ef82 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Diffusion of Thoughts: Chain-of-Thought Reasoning in Diffusion Language Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3fd64f8-25e1-4b64-8c27-bc7da95116f0 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Dream 7b, 2025
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 4fea2688-ab01-4b89-9261-10e8d52d5b40 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning LIMO: Less is More for Reasoning
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b9c8bac-0899-4a33-8d84-64664c99f524 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning MetaMath: Bootstrap Your Own Mathematical Questions for Large Language Models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1f7787c-cd72-4985-9064-bf3516b25a7c · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Fine-tuning discrete diffusion models with policy gradient methods.arXiv preprint arXiv:2502.01384, 2025
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddaa17c0-fd48-41e8-9179-b7e3d98c1fcc · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Lima: less is more for alignment
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 620bcf72-3eca-4eaa-a1b9-eca5506ead84 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Simply put, at any timestep, the probability that a token transitions to the masked state isαt
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e43718b8-34ed-4969-9c56-8696e10cb7ed · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Unresolved cited work
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 54dce162-5be9-4a55-9e50-9c7dd6442e6b · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Let’s go through each step in detail:
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation f6819624-3b62-44c0-8722-202622576a3f · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Therefore, the number of stars in the 5-star rows is: 76−36 = 40
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 60a07723-08c0-44e8-ab7a-a155ac847e62 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Two-thirds of the loaves are sold in the morning and half of what is left is sold equally in the afternoon and evening
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 3b232858-f673-440d-8c7a-6356dfc00fde · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Unresolved cited work
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c28e7d81-8807-4d40-9c5b-8dc9a7506c50 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Unresolved cited work
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d8b6e3d9-3d76-4eb6-a828-78d70510cfed · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Unresolved cited work
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 819e5a0b-b8c3-4554-8fac-d74203d951a7 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Unresolved cited work
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 3a68ef4c-529b-4f1b-bb3a-dac5288a4b99 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Unresolved cited work
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation f3288336-2e3a-403b-a3ba-1e79e57e628c · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Unresolved cited work
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation fd5652e6-e659-4de5-8fca-7949b9e7370e · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Unresolved cited work
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 05fcd3f9-9214-4025-b1a1-26c600c8fe76 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Unresolved cited work
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation ca99e707-a3a3-49fe-8d35-44e004ab2627 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Unresolved cited work
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c804e250-1bd9-4957-8d88-38c0ad831053 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Next, we need the total number of stars on the flag, which is 76
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 590f105d-e73d-409c-a87f-b872fa43e158 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Unresolved cited work
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 57d2df92-0106-42d4-8983-c2cc3487f889 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Unresolved cited work
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1ad2f604-4294-4761-9a30-1e3f4d331141 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Unresolved cited work
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e71d0a4b-3982-4dec-b272-88d78432cd21 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Unresolved cited work
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 5a9a5635-4bd2-4ff5-acd8-d4deb34c1af7 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Unresolved cited work
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 4b3e439f-69bc-46a6-a072-734a5d1f957e · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning < /reasoning><answer>8 < /answer> 25 Question:Jennifer’s dog has 8 puppies 3 of which have spots
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e65c4cc1-da41-42e3-8f42-7133761336cc · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Unresolved cited work
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 74015850-e3e6-49d2-82d1-ef9e212f4e22 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Unresolved cited work
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation afd7c678-33d7-4ed3-a64a-4d893462c67b · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning First, let’s find the total number of puppies from both dogs: - Jennifer’s dog has 8 puppies
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation fb592529-9f81-4b4f-ab33-307816770739 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Unresolved cited work
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 431cc05e-a8d9-434c-81b1-d78e4fc21042 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Unresolved cited work
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2bc6f677-a7b4-4993-9001-91417240d2e3 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Unresolved cited work
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 69f646d1-b3a3-4ebe-b22a-979957f1ad50 · outbound
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning Unresolved cited work
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c273ed96-ece8-4163-bc0c-c5b15126e428 · inbound
Decomposing Elements of Problem Solving: What "Math" Does RL Teach? d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b9288bf-20e8-44f4-9836-6d28f7abcf2e · inbound
SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2518048-a2e4-4371-aa52-f76b644542a7 · inbound
On a few pitfalls in KL divergence gradient estimation for RL d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3edbd26d-a14d-4fcf-a186-7929a40f4206 · inbound
DiffuCoder: Understanding and Improving Masked Diffusion Models for Code Generation d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19a8724e-48b8-41ce-90f1-fbb4acafa6b7 · inbound
A Survey on Latent Reasoning d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 138
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d18a6d1-3a90-43f4-9b80-b01d93858697 · inbound
Review, Remask, Refine (R3): Process-Guided Block Diffusion for Text Generation d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb227c3a-102e-453a-a558-62853cc6711d · inbound
A Survey on Diffusion Language Models d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 045f5a3e-1a77-486e-ad78-38fc1ad5c844 · inbound
Any-Order Flexible Length Masked Diffusion d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9df192f7-4d0f-4ebf-a9c7-98072d348ba8 · inbound
ParaThinker: Native Parallel Thinking as a New Paradigm to Scale LLM Test-time Compute d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 092e5c71-68a6-4abd-9416-eda5d18239af · inbound
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 649721c8-ba34-491e-b8c5-74c4f6654a65 · inbound
Inpainting-Guided Policy Optimization for Diffusion Large Language Models d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 746db5cb-acfe-4e8f-8f4b-6893ed5b4162 · inbound
GIFT: Guided Importance-Aware Fine-Tuning for Diffusion Language Models d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation aba1a0ea-dabe-4e8d-9a6c-1e472fb10393 · inbound
d2: Improving Reasoning in Diffusion Language Models via Trajectory Likelihood Estimation d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95c45515-7d88-4f4f-81ad-aeb64f48c3e1 · inbound
Error Analysis of Discrete Flow with Generator Matching d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5de9092-09f2-49dc-a7be-d917e547756a · inbound
Discrete Guidance Matching: Exact Guidance for Discrete Flow Matching d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation bb894fbc-d183-45f0-9476-45dd4fe5b6fe · inbound
Simple Policy Gradients for Reasoning with Diffusion Language Models d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 135b062e-3761-46b2-b702-32adbdd337b1 · inbound
ParallelBench: Understanding the Trade-offs of Parallel Decoding in Diffusion LLMs d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5be6dcc2-60a7-4261-bfd7-ba2f1b8900cf · inbound
Boundary-Guided Policy Optimization for Memory-efficient RL of Diffusion Large Language Models d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95802dca-1915-49bf-95e9-75e2fdfa9e6f · inbound
Efficient-DLM: From Autoregressive to Diffusion Language Models, and Beyond in Speed d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c8d58ff0-d20e-4ca6-8362-27c7ea08720f · inbound
The Flexibility Trap: Rethinking the Value of Arbitrary Order in Diffusion Language Models d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 008922e3-fc70-4ff2-8602-241dee5f854f · inbound
Streaming-dLLM: Accelerating Diffusion LLMs via Suffix Pruning and Dynamic Decoding d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 829c4aa9-9e56-4e5a-9a5a-9f4620b988f6 · inbound
Tuning the Implicit Regularizer of Masked Diffusion Language Models: Enhancing Generalization via Insights from $k$-Parity d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01373c77-85f0-4e11-9502-b8aea1c6ff96 · inbound
Can I Have Your Order? Monte-Carlo Tree Search for Slot Filling Ordering in Diffusion Language Models d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6cbe0675-721d-4b0f-9d84-b9fc1fab8da3 · inbound
LightningRL: Breaking the Accuracy-Parallelism Trade-off of Block-wise dLLMs via Reinforcement Learning d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecd4746f-30f2-48ea-94a7-a15f33f490b2 · inbound
LogicDiff: Logic-Guided Denoising Improves Zero-Shot Reasoning in Masked Diffusion Language Models d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation ef021ff4-084b-4518-9e04-33c8f5f32563 · inbound
Discrete Flow Matching Policy Optimization d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation bcb2d663-7f03-4a2c-8b0a-599043483cea · inbound
DMax: Aggressive Parallel Decoding for dLLMs d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 103
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation bf78d367-87dc-4715-9df8-702544f8ccea · inbound
DMax: Aggressive Parallel Decoding for dLLMs d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 103
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 91ab6570-e043-4ba3-a0cf-71800b2e7de7 · inbound
CAGenMol: Condition-Aware Diffusion Language Model for Goal-Directed Molecular Generation d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e8e19a53-fed5-49ae-8e60-3657a4156917 · inbound
Break the Block: Dynamic-size Reasoning Blocks for Diffusion Large Language Models via Monotonic Entropy Descent with Reinforcement Learning d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d324d6d3-6529-43b6-a5cf-bef1d948b33e · inbound
Break the Block: Dynamic-size Reasoning Blocks for Diffusion Large Language Models via Monotonic Entropy Descent with Reinforcement Learning d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b492a0da-1eb2-442e-9149-8b828358f3d6 · inbound
ReflectDrive-2: Reinforcement-Learning-Aligned Self-Editing for Discrete Diffusion Driving d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 127
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation aa7ebc25-e541-4844-a4ba-e03d2b854d86 · inbound
ReflectDrive-2: Reinforcement-Learning-Aligned Self-Editing for Discrete Diffusion Driving d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 127
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 71595bb9-060a-46a1-88b4-2766891b5934 · inbound
Continuous Latent Diffusion Language Model d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 110
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 13fbf4fb-6809-4a55-889c-425d04698e0b · inbound
dFlowGRPO: Rate-Aware Policy Optimization for Discrete Flow Models d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b2fa284c-7fbb-4230-b0b8-81281968b923 · inbound
TAD: Temporal-Aware Trajectory Self-Distillation for Fast and Accurate Diffusion LLM d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 8055c6b4-1250-471a-bf0a-91d9a373dc56 · inbound
Relative Score Policy Optimization for Diffusion Language Models d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 102
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c5e585f8-aaf9-4ed1-85a4-b7b2fe0f681b · inbound
Self-Distilled Trajectory-Aware Boltzmann Modeling: Bridging the Training-Inference Discrepancy in Diffusion Language Models d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c30e4d21-cab1-445d-8072-e61f546b1df7 · inbound
Self-Distilled Trajectory-Aware Boltzmann Modeling: Bridging the Training-Inference Discrepancy in Diffusion Language Models d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2ed526fd-673b-4177-b9cd-5df196a68392 · inbound
AIS: Adaptive Importance Sampling for Quantized RL d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 045966a3-6947-4a08-be20-0d284b7cb4ce · inbound
Sketch Then Paint: Hierarchical Reinforcement Learning for Diffusion Multi-Modal Large Language Models d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 82863feb-1c3b-4a03-8c7d-2eac0d71b230 · inbound
Roll Out and Roll Back: Diffusion LLMs are Their Own Efficiency Teachers d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 85089373-950e-48ae-a8f3-c1b02776d4e3 · inbound
Beyond Execution: Static-Analysis Rewards and Hint-Conditioned Diffusion RL for Code Generation d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e7825b23-071e-4e58-bd64-9b5d22ed667b · inbound
Elastic-dLLM: Position Preserving Context Compression and Augmentation of Diffusion LLMs d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation aa982a74-efe6-4c88-be30-2ad996f62d48 · inbound
Contrastive Distribution Matching for Amortized Sequential Monte Carlo in Discrete Diffusion d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b96407f0-617f-4c14-8342-e1a05269d207 · inbound
Reinforcement Learning from Denoising Feedback d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation effaf0ab-56ba-4012-9596-4bedd1f16d6f · inbound
GDSD: Reinforcement Learning as Guided Denoiser Self-Distillation for Diffusion Language Models d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d901172b-d729-4545-af30-2f793b5f2ce2 · inbound
Efficient Diffusion LLMs via Temporal-Spatial Parallel Decoding and Confidence Extrapolation d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e39905c6-0a80-403e-992b-40487d217b6b · inbound
dMoE: dLLMs with Learnable Block Experts d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a1c0a56b-9b80-4d2b-bd89-5292295bea58 · inbound
MaskForge: Structure-Aware Adaptive Attacks for Jailbreaking Diffusion Large Language Models d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation fdccb9c1-9013-4bf4-ba2b-1b9e00b442f7 · inbound
Dynamic Infilling Anchors for Format-Constrained Generation in Diffusion Large Language Models d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 5ea81465-96df-4296-bbc6-3c48fb0fb2c7 · inbound
Back on Track: Aligning Rewards and States for Reasoning in Diffusion Large Language Models d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation ed32f20b-8433-4803-96e0-2d4069d3355d · inbound
Re-evaluating Confidence Remasking in Masked Diffusion Language Models d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d2cde3bc-a340-434b-89c0-e95e3c97ec0c · inbound
Beyond Fully Random Masking: Attention-Guided Denoising and Optimization for Diffusion Language Models d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 0172605a-8bac-4417-a5f9-dfbc82e844bb · inbound
DiPOD: Diffusion Policy Optimization without Drifting Apart d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 885ff215-a1bb-4c53-9e2d-76a187028fef · inbound
Learning from the Self-future: On-policy Self-distillation for dLLMs d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 5da9227b-31af-4665-9fb3-b6ef591817f8 · inbound
JetSpec: Breaking the Scaling Ceiling of Speculative Decoding with Parallel Tree Drafting d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b93a533f-2f28-471f-9e81-03cdc95ea587 · inbound
Improved Large Language Diffusion Models d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 34db73ff-5bea-4beb-9c8e-d6ebdfcdd98e · inbound
TACG: Trajectory-Aware Commit Gating for Diffusion Language Model Decoding d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d68e3dd8-7361-4025-8fea-4b6ebb70d01c · inbound
Nemotron-Labs-Diffusion: A Tri-Mode Language Model Unifying Autoregressive, Diffusion, and Self-Speculation Decoding d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e5d9661a-e64f-4502-905a-142f127982ae · inbound
Concurrent Image Understanding and Generation: Self-Correcting Coupled Markov Jump Processes d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f14f2bef-06a2-4359-881b-ba19f5f745c7 · inbound
Hierarchical Domain Generalization d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1fdbb62-14a8-4a64-b749-caff434a2fec · inbound
Trace-Based On-Policy Distillation for Masked Diffusion Language Models d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9519f0b2-c9e8-42f2-af52-d98a6fc5a3be · inbound
From Interface to Inference: Eliciting Any-Order Inference from Any-Order Models d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e433f52-9bbe-44e5-b545-88135a8b4f41 · inbound
Escaping Confidence Trap: Evolutionary Decoding for Mathematical Reasoning in Diffusion LLMs d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dcbaa409-4c60-4516-91f6-2a0c9657e95d · inbound
AURORA-LM: Autoencoding Unified Representation for Continuous-Latent Diffusion Language Modeling d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.