Pith. sign in

Paper Citation Record · LEDGER

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models

As of 16 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 0 inbound Pith citation observations for arXiv:2509.02528.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.02528 v1

Coverage vector

measured 65 of 65 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T16:44:36.855313Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

65 of 65 outbound references displayed

  • verified exact0
  • verified fuzzy20
  • unresolved45
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ce35b655-a1a8-4db1-b4df-c9f628aeb06a · outbound

This paper cites A tail inequality for suprema of unbounded empirical processes with applications to markov chains.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models A tail inequality for suprema of unbounded empirical processes with applications to markov chains

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:44:37.866601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.574676Z digest=sha256:166f5060ca451e2159fdfa0b61796f204d6100062a6fda58cf44e7c06d02d66f

Observation a6ea8990-d4fe-4572-be22-84bc2b3ff7b3 · outbound

This paper cites an unresolved cited work.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:44:37.852251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.579570Z digest=sha256:bdcf971c39caccce78df4de8585a4f47348042c520c173e56b1827c2ec54ab0b

Observation 2e81c948-05bc-48b6-a9c7-f80e217370dd · outbound

This paper cites Surprising Negative Results for Generative Adversarial Tree Search.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Surprising Negative Results for Generative Adversarial Tree Search

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.584243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.584243Z digest=sha256:bed03bdcef7e15e3071fa9abe5d003326f707c8d50c3ed4ce9a0c4f0b9c1e21e

Observation bba7200c-4b76-4ce5-bbf4-44609ff0e3a6 · outbound

This paper cites an unresolved cited work.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:44:37.837493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.589195Z digest=sha256:9ae41fe0c55076466ed13b764577862fd8f91be317ea4feac64b2751d22c7be8

Observation 66c207b8-7aa6-4645-a53a-88af4e9fa20e · outbound

This paper cites Accelerating RL for LLM Reasoning with Optimal Advantage Regression.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Accelerating RL for LLM Reasoning with Optimal Advantage Regression

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.593554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.593554Z digest=sha256:65673f51964d0e4f66725e6fd485a5470913a9e91193652f8a6da3a8a020cb07

Observation 3edd4543-5a71-4fbf-a4af-38c8c0b488bf · outbound

This paper cites Training Diffusion Models with Reinforcement Learning.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Training Diffusion Models with Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.598281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.598281Z digest=sha256:8b39e26c1c9cb584d39089647c5c922042e61b1e8726efea9a80d00bbe65e7ab

Observation f5a455da-2a25-47b4-b2b4-3055f47e7556 · outbound

This paper cites Approximation variationnelle des probl \`e mes aux limites.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Approximation variationnelle des probl \`e mes aux limites

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:44:37.822868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.603337Z digest=sha256:bd85bf74b44ee60dad239283c7eb8488901375cf994a5e8978aeb27b53593fcb

Observation 5db587be-6eb7-48db-997a-5668af502f1a · outbound

This paper cites an unresolved cited work.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.607399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.607399Z digest=sha256:4001e93df30cdf2d452f3d25d81bd03612639f5f20ff3bc3a5fdb21bbafbc831

Observation 5b33d044-71e8-48e4-af55-04559caf6201 · outbound

This paper cites an unresolved cited work.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:44:37.798980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.611684Z digest=sha256:0814cb35f48747b9799eb168638bcd072d715f0697c39f5b04bc1c79f592258f

Observation 357fdbad-4e84-4e50-bb3d-e06777dadda4 · outbound

This paper cites Tail bounds via generic chaining.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Tail bounds via generic chaining

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:44:37.784130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.616019Z digest=sha256:36a5f34196f0cad001df429f9df75f2a37a25a131aef95ec45500eca0924714f

Observation aab51627-1e16-4a97-8e84-3dca7ef1dc5b · outbound

This paper cites Dhariwal and A.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Dhariwal and A

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:44:37.769689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.620527Z digest=sha256:49f90461375a17f7d8495ee67033e31f15a70e843be00c54c0773b0e59c4b8a4

Observation 24deabe8-76ca-41a4-99f5-e158d3960904 · outbound

This paper cites an unresolved cited work.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:44:37.753577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.624739Z digest=sha256:3c6f3ffbf6ea93ca3a7f297b8f070ca21eadf83aa323aa4dbc78c384b91fc0ba

Observation 1f32e9d0-15c2-45a8-ba95-835510272c6c · outbound

This paper cites Optimizing DDPM Sampling with Shortcut Fine-Tuning.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Optimizing DDPM Sampling with Shortcut Fine-Tuning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.629752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.629752Z digest=sha256:1ae0662feea3151d8fc5a0858816ea3dde2a688de7fce49a57642aa0628a9829

Observation d9804de7-6b84-4161-9bb0-0592b1e077f2 · outbound

This paper cites an unresolved cited work.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:44:37.738600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.634175Z digest=sha256:939782075516c827715968f3cd2f675a9348f3435b1c170bdbcc408d3d6336e0

Observation 16f0848d-8fed-4a6c-a665-7a39131a2566 · outbound

This paper cites Deep neural network approximation for high-dimensional parabolic Hamilton-Jacobi-Bellman equations.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Deep neural network approximation for high-dimensional parabolic Hamilton-Jacobi-Bellman equations

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.638288Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.638288Z digest=sha256:7c0f5be872e3f249ededeb07857881f6da186620e1be03991655b1f8f397f4da

Observation 0d6da94f-2dbe-468f-9296-fb9fcfe3e5e5 · outbound

This paper cites Reward-Directed Score-Based Diffusion Models via q-Learning.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Reward-Directed Score-Based Diffusion Models via q-Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.642599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.642599Z digest=sha256:5ac9f092167ce27258e9230bf8bf53d340a0010c79f3962b88affb3b1fb4d543

Observation efe4ebd1-85f6-4f42-8577-660cb3ee0856 · outbound

This paper cites an unresolved cited work.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:44:37.723557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.646957Z digest=sha256:420a689bc56317c326889d24a1ecb918afb3ab283125f05a56555fb4a653b377

Observation 5e7f336f-cef4-419c-8c8e-a94b77f526eb · outbound

This paper cites an unresolved cited work.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:44:37.706807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.651274Z digest=sha256:ec29b40ed0190de7a64b6191e0e0e24962f333a9c9339ab3b212c7bb2915fca9

Observation 86941e3d-e0ef-4513-8728-6c54adc717d2 · outbound

This paper cites Stochastic Control for Fine-tuning Diffusion Models: Optimality, Regularity, and Convergence.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Stochastic Control for Fine-tuning Diffusion Models: Optimality, Regularity, and Convergence

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.655600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.655600Z digest=sha256:ea5f8ed15961f7cd12bea75a735e3a21a46f5ef096cbe5d3676e45d5baf4f5b8

Observation 581b27d2-9b22-4ada-8b96-78cca4a8b197 · outbound

This paper cites Reproducibility of Benchmarked Deep Reinforcement Learning Tasks for Continuous Control.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Reproducibility of Benchmarked Deep Reinforcement Learning Tasks for Continuous Control

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.660336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.660336Z digest=sha256:7f8bea7f8418d4b94237e153afb0c95b796f88b90cf2b7b99ccdad72de814ea2

Observation 7ff8cc8d-fada-4d31-afbe-1c2f2bfe6a60 · outbound

This paper cites an unresolved cited work.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:44:37.690904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.664682Z digest=sha256:6ff2c3a7e8c87da61c2ff904a0839bbf4719e61b9116852486c9eaefdd11967e

Observation 056644f1-4f77-4a11-bcce-3ccfd8d8673e · outbound

This paper cites Jia and X.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Jia and X

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.668860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.668860Z digest=sha256:a99c16d6bdf5252088391b488b59bd8538d77570c173e74bbdcda8e732586441

Observation 348ebea6-7878-4707-a245-fc2007bb9851 · outbound

This paper cites Jia and X.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Jia and X

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.673059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.673059Z digest=sha256:afd860ec2fa6f9487b50c612da922d145f8a2f4eb143a37c76029175ed568c59

Observation 24f46d19-dba4-4178-bcee-82886277c09f · outbound

This paper cites Jia and X.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Jia and X

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.677418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.677418Z digest=sha256:3f8869d92b385c45c4b8ae6c832d9d03ef74c38efba7305b20723c2250a1d110

Observation 84d92b7d-6826-4887-bbd7-4d64bced1a9f · outbound

This paper cites an unresolved cited work.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:44:37.648077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.681658Z digest=sha256:a7dbf49d31007b72e7833deecb19cf6649c294cb83876f1a95d4ac4964db106e

Observation 86cfa3e9-2e88-447a-9a5c-ac0ebecdb789 · outbound

This paper cites Korshunova, N.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Korshunova, N

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:44:37.634586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.685931Z digest=sha256:cc73cb1ebc3a843f0e3a4c9d046709d1184f4fb20773121ba56914ac38619dff

Observation 9a58790a-e4fc-4fff-9dcc-634caceb5191 · outbound

This paper cites Kakade and J.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Kakade and J

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:44:37.619924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.690568Z digest=sha256:de01b02d422d880c5b4a4c5acb6b0511e8a5254b38c09d20cf5cca4add82354c

Observation 227bb8ea-8bc6-4517-b426-a2a0bb38f526 · outbound

This paper cites Koltchinskii.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Koltchinskii

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:44:37.603426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.694800Z digest=sha256:9c5295eb2c558db7a501db66bd9f570631d3af485f47b29a36e4e19d974f9b2c

Observation 42a7b6b8-ca5f-48cb-892d-cef824af496c · outbound

This paper cites an unresolved cited work.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:44:37.588155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.699081Z digest=sha256:71061dfde85e301ff8c286811b3369e3696a9b4064a57d4ab8379c983ce802fb

Observation 66dfb0aa-481a-4a23-92ab-a2b26db2bc0f · outbound

This paper cites Machine Learning For Elliptic PDEs: Fast Rate Generalization Bound, Neural Scaling Law and Minimax Optimality.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Machine Learning For Elliptic PDEs: Fast Rate Generalization Bound, Neural Scaling Law and Minimax Optimality

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.703386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.703386Z digest=sha256:37e754467ecda4f0ac35a9d67462318e725df773ac54da4b48d59af14527ea60

Observation 3a182fd1-1462-4625-bcbb-8be9fead9435 · outbound

This paper cites an unresolved cited work.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:44:37.573562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.707821Z digest=sha256:6d046de065bef0c35f3acd505580305874e4110ee1cc9f2dbfe28d0a4b511263

Observation 85f2a205-f404-4abe-b49f-e97c7cff4591 · outbound

This paper cites Learning subgaussian classes : Upper and minimax bounds.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Learning subgaussian classes : Upper and minimax bounds

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.712015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.712015Z digest=sha256:786787ae175441f443d78b712693978183063e118b26fbeec0f588c0254d68e0

Observation f38cec05-2977-4b70-91a6-479ccf2ed169 · outbound

This paper cites Estimates of the numerical density for stochastic differential equations with multiplicative noise.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Estimates of the numerical density for stochastic differential equations with multiplicative noise

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.716547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.716547Z digest=sha256:8385b24ba720c69d50ee093bc4ab57d84bd2a1ddb5132bbe0b02dbd550372050

Observation eacea762-4c8f-4b48-a435-b039fbcbacee · outbound

This paper cites Madelung.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Madelung

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:44:37.558758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.720978Z digest=sha256:4b31b9eae9da336765916821436e6686841b734e8659239266402cc151249806

Observation 4475f4d1-ef11-4fa7-ac43-d3afba11cd90 · outbound

This paper cites Munos and P.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Munos and P

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:44:37.544291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.725555Z digest=sha256:0acf07d03ad4084bac7dcc6d808d8edc9974fe2f686d04965244b5667bbab9b2

Observation fb479064-3351-45b3-9509-d31099310be8 · outbound

This paper cites Mendelson.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Mendelson

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.729788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.729788Z digest=sha256:8f79b7c5d2a7a7c32c046d40dd25485de94b68f8ca778d9ac2bdd76d08b67324

Observation 1710e4d7-edad-49d5-b291-03d653ba5764 · outbound

This paper cites Muhle-Karbe, J.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Muhle-Karbe, J

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:44:37.520532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.734279Z digest=sha256:d281d9065bc60f54b3c0c0cbca0b3a9a92e3331c3f6eaebb987733cb166d5710

Observation e4051267-c4d5-497c-8c65-eecdcd60cc37 · outbound

This paper cites Statistical guarantees for continuous-time policy evaluation: blessing of ellipticity and new tradeoffs.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Statistical guarantees for continuous-time policy evaluation: blessing of ellipticity and new tradeoffs

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.738417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.738417Z digest=sha256:8a809c03f529acf20112d470a7cdc11c8e35a429db76fab8f425421ab3209242

Observation 089c8cea-3fa2-487d-ade6-293c0721c48e · outbound

This paper cites Optimal oracle inequalities for projected fixed-point equations, with applications to policy evaluation.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Optimal oracle inequalities for projected fixed-point equations, with applications to policy evaluation

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:44:37.506171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.742846Z digest=sha256:231e7eca44f92e7226ebf746ec4407102c37d723ab6dfad6573f14f158e44983

Observation 7be2d523-daa4-4493-a051-62d319c03827 · outbound

This paper cites Menozzi, A.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Menozzi, A

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.746982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.746982Z digest=sha256:04fe564eb41215687a2a29eab95b9a8a0374b35caee7b64522071f08b8c831b0

Observation 0e4bb261-589d-4c82-bc20-338bd2824240 · outbound

This paper cites On Bellman equations for continuous-time policy evaluation I: discretization and approximation.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models On Bellman equations for continuous-time policy evaluation I: discretization and approximation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.751413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.751413Z digest=sha256:0bdfbbf6b8ab5dd35549c248459381c0c85b5da21b6eb0846888b38d563f452c

Observation d40c9c10-e1b2-49ce-a523-11bd20dc49ad · outbound

This paper cites Nemirovski, A.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Nemirovski, A

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:44:37.483121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.755747Z digest=sha256:fb2dc849c13ba735a3a6287f9d5909904ea82e9013fbc3eca079ae06bfafb9c2

Observation 1cf5d741-49f8-4b6f-b147-d7e3e7dc64ea · outbound

This paper cites The Malliavin calculus and related topics.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models The Malliavin calculus and related topics

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:44:37.469223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.760228Z digest=sha256:929b9ee61d08b74dd50efd0cd8a16ffd03c74fc45871c90b5d31f5437c1ebe64

Observation 47be84e9-3cdf-4645-afd0-945a359f1801 · outbound

This paper cites Ouyang, J.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Ouyang, J

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:44:37.454023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.764450Z digest=sha256:9e75433c5de9094bad5fb032bc6e5a7d0a11def96c85f259e589138e9397335b

Observation 9711309c-3e0d-4b1d-a1ac-fcfb2a64a428 · outbound

This paper cites an unresolved cited work.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:44:37.438840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.768422Z digest=sha256:0278b612d04f4089b43f7e1e11f4278dcd80270a32a6361be5ed6ff3cc3c01ad

Observation 91b4fd3d-bd46-4ed0-b514-dacffc647c89 · outbound

This paper cites Sirignano and K.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Sirignano and K

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:44:37.424838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.772563Z digest=sha256:54859292417761c2c35c887e53ffeff0ffd665a949c6fa8d55d17c9d53c2fdd2

Observation f5efafcc-9ee1-400b-b60d-7a2fc52a5f62 · outbound

This paper cites Score-Based Generative Modeling through Stochastic Differential Equations.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Score-Based Generative Modeling through Stochastic Differential Equations

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.776671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.776671Z digest=sha256:2d3b57f2741586be02f0990dd63215b0d30efa31116173f99629f40401b92e9b

Observation d685b4b0-985a-4518-99ea-b9695def74be · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.781087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.781087Z digest=sha256:1b73b4c5e4840a64e16701c6f814c4eb5a0983504ad811b8c3b52dc761254dba

Observation 85048588-d7ea-45fd-afb4-2a59d9356843 · outbound

This paper cites an unresolved cited work.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Unresolved cited work

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.785848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.785848Z digest=sha256:6c9239f4d245a9166b8ba0e099bc19173242ee476ff6a8ccdc57eceae9ec0145

Observation 2c9fda2c-941c-4e37-b2c2-7afdd0c9cf02 · outbound

This paper cites Theodorou, J.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Theodorou, J

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:44:37.411284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.790041Z digest=sha256:72d50943bd94e93074e2a0e2312f6c86efef501527604d90b59810d9a230f8b2

Observation 15bce076-6d95-4e1a-afc8-82eee8824747 · outbound

This paper cites an unresolved cited work.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:44:37.397552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.795153Z digest=sha256:8abe9fd4f2c9c71a39173a9913f0b9ee4418f3b399580ca7823c0963d995c169

Observation 94fc095e-107d-4a16-9724-9c243e65ea04 · outbound

This paper cites Tang and R.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Tang and R

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:44:37.384029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.799325Z digest=sha256:0903b7749b29c6904379e28ffeaa5ec2a3e499adcd50f244a24097dfdf11a57a

Observation 6ee942e2-c912-4809-a2e2-485bb3e2e01f · outbound

This paper cites Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.803558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.803558Z digest=sha256:6723549d2dd30b453aa7240169d8fd096f4437a0f1fdc783f90c857e22a9b8c4

Observation bd38fb5c-88b1-4662-a24a-be437b15d194 · outbound

This paper cites Feedback Efficient Online Fine-Tuning of Diffusion Models.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Feedback Efficient Online Fine-Tuning of Diffusion Models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.807712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.807712Z digest=sha256:53cda38b0508468184241815c5d3bc4042ee434e362fd691d30600f1a5ead5b4

Observation 2e624b8a-a421-4994-b9c7-48d723e40b17 · outbound

This paper cites Understanding Reinforcement Learning-Based Fine-Tuning of Diffusion Models: A Tutorial and Review.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Understanding Reinforcement Learning-Based Fine-Tuning of Diffusion Models: A Tutorial and Review

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.812472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.812472Z digest=sha256:ed6099911140dd4d606ad55dfe1e719d263be1f0a4f73f965757dcc79532eb12

Observation fbdc3ec1-920a-40dc-8ae6-028382bbec99 · outbound

This paper cites an unresolved cited work.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Unresolved cited work

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.816826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.816826Z digest=sha256:598289521480bd19f8ad1203f618e5729af307b8ce145ab8f770d35322afdfd6

Observation 45695990-f541-4903-b0cf-45e7ee65ed2e · outbound

This paper cites Weisz, P.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Weisz, P

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:44:37.360790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.821225Z digest=sha256:ecd90ff084d9cdbc3d80857ce0b045b0bb257ee7fb40459053c08164d847af8d

Observation dea01282-61f6-4be2-b549-28e660392cfc · outbound

This paper cites an unresolved cited work.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:44:37.345095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.825462Z digest=sha256:cafa6425e865b387f791327ed745a899d107d9c258adf88929e216a43e7922de

Observation 1dd41a41-61cc-4e4f-8b26-754abea3338e · outbound

This paper cites Xie and N.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Xie and N

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:44:37.331166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.829886Z digest=sha256:331d957b0d1c896444d5ab5c26d78ceb72dfe182e99b94abdf4fb29edba6437f

Observation d7f0b92b-6c5a-4521-8406-67beed504235 · outbound

This paper cites an unresolved cited work.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Unresolved cited work

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.833885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.833885Z digest=sha256:53d3ff63def272c628384b3aa7990c6c6d9c1eaa61113eed0e60d3c7731b4ace

Observation 6829bc90-f340-4ee5-89d3-ea9de10cd911 · outbound

This paper cites Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.837980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.837980Z digest=sha256:13dcd49fc11fdd2a77bd89b674c2073a4d8f306cad661b66860ce0409b0c6eeb

Observation c10bce38-5728-4ab9-a8eb-f6c8654d5aad · outbound

This paper cites an unresolved cited work.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Unresolved cited work

Reference 62

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:44:37.316649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.842390Z digest=sha256:419e4eccadf87d65c54a2aa4ddd2e4ef145d8a355f9c2a2c611b0057b70e75c3

Observation afce878f-7cf4-4d11-bce1-aaed7f9bd41e · outbound

This paper cites Fine-Tuning Language Models from Human Preferences.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Fine-Tuning Language Models from Human Preferences

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.846633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.846633Z digest=sha256:43b77fe0e8f9ee9f9cdc194d0a6867d2ede9257dc5f28b9d5d4f7b943a034535

Observation e63fec42-085c-47d9-85ea-aa283a46723c · outbound

This paper cites Ziemann, S.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Ziemann, S

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:44:37.302837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.851181Z digest=sha256:42d50144959ce1188d39397ff395d72ffb46703725a41f239e2ba49efcd8659d

Observation 469d4b30-c38b-4926-bffb-249dc77554b5 · outbound

This paper cites an unresolved cited work.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:44:37.289119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:44:36.855313Z digest=sha256:b366c5c58c968423528a2264284c3bd4fa3ea5b2e78dfc9a17a6a6f531857e66

Pith citing papers

No inbound Pith citation observations are available.