Pith. sign in

Paper Citation Record · LEDGER

Backpropagating Through Simulation: Analytic Policy Gradients for Sample and Learning Efficient Differentiable Continuous Control

As of 23 August 2026, this Paper Citation Record lists 20 of 20 outbound references and 0 inbound Pith citation observations for arXiv:2606.21525.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.21525 v1

Coverage vector

measured 20 of 20 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-26T14:20:33.205947Z

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

20 of 20 outbound references displayed

  • verified exact3
  • verified fuzzy0
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 907a6b65-1e2a-4917-87dd-5b6e12bb034e · outbound

This paper cites Proximal Policy Optimization Algorithms.

Backpropagating Through Simulation: Analytic Policy Gradients for Sample and Learning Efficient Differentiable Continuous Control Proximal Policy Optimization Algorithms

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-07-04T06:39:37.282473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-26T14:20:33.205947Z digest=sha256:bbdf12a2181f43d0bf2102450ee6ddf69b40eaa244c33fade8ab64675b353f82

Observation 3bd1908d-6a3b-4fb6-8d54-893da4107e23 · outbound

This paper cites an unresolved cited work.

Backpropagating Through Simulation: Analytic Policy Gradients for Sample and Learning Efficient Differentiable Continuous Control Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-26T14:20:33.205947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T14:20:33.205947Z digest=sha256:058ce133aa0847499d59b3c5c65c741a5f59981a3e2992893ac8bd38c93d6cd8

Observation 6aac06a4-2c89-4693-b4eb-6c986fa1e46a · outbound

This paper cites an unresolved cited work.

Backpropagating Through Simulation: Analytic Policy Gradients for Sample and Learning Efficient Differentiable Continuous Control Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-26T14:20:33.205947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T14:20:33.205947Z digest=sha256:62b6207e884cf8c64b814f7e01d8b0af3fb7a081627b7ecad6f784f4ef8fcf91

Observation d0e11fee-f720-4452-9eff-554aaa3f67b8 · outbound

This paper cites Chromatic Homotopy is Monoidally Algebraic at Large Primes.

Backpropagating Through Simulation: Analytic Policy Gradients for Sample and Learning Efficient Differentiable Continuous Control Chromatic Homotopy is Monoidally Algebraic at Large Primes

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-04T06:39:37.276279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-26T14:20:33.205947Z digest=sha256:20d19dc7c74049f3d2ec745db1a7ba8450907cb70cd529ddeb7997813d5c5c5a

Observation 40a4f155-aa69-4973-be21-05867c9145f5 · outbound

This paper cites an unresolved cited work.

Backpropagating Through Simulation: Analytic Policy Gradients for Sample and Learning Efficient Differentiable Continuous Control Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-26T14:20:33.205947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T14:20:33.205947Z digest=sha256:edf3f23d024b31def3377a6162fbafed99ce6c543fc80c7215f6e2821f5d7c65

Observation 7e62e20a-211c-4e75-9fa7-1ca41c85729e · outbound

This paper cites Mastering Diverse Domains through World Models.

Backpropagating Through Simulation: Analytic Policy Gradients for Sample and Learning Efficient Differentiable Continuous Control Mastering Diverse Domains through World Models

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-07-04T06:39:37.285158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-26T14:20:33.205947Z digest=sha256:9ff32df4bf38cf1f2a7c8cb6bdc70dd0fd1b05d9310765be57389a7e4d43f868

Observation 8b42bcb1-6afb-4a4d-a4f4-b594c48e04d4 · outbound

This paper cites an unresolved cited work.

Backpropagating Through Simulation: Analytic Policy Gradients for Sample and Learning Efficient Differentiable Continuous Control Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-26T14:20:33.205947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T14:20:33.205947Z digest=sha256:b54f45a3eb0d09f055f7274013fa242d9d897173dec938b721d24ce5d21034c8

Observation 0716d78f-a2fc-487d-bdd5-dff5cb1565b5 · outbound

This paper cites an unresolved cited work.

Backpropagating Through Simulation: Analytic Policy Gradients for Sample and Learning Efficient Differentiable Continuous Control Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-26T14:20:33.205947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T14:20:33.205947Z digest=sha256:ca639485c0fdc9cd06bd5dfd96baf0bf6a7b9ece3c77007d29444816e4609e5a

Observation 17f0a133-4017-4535-846d-4ec7557601a9 · outbound

This paper cites an unresolved cited work.

Backpropagating Through Simulation: Analytic Policy Gradients for Sample and Learning Efficient Differentiable Continuous Control Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-26T14:20:33.205947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T14:20:33.205947Z digest=sha256:1982a853d15778eda397adcc2aca850e0331b20a2c0e1c60144c37620e5e976c

Observation 8afa6b05-c228-4145-8464-d65a60a72b13 · outbound

This paper cites an unresolved cited work.

Backpropagating Through Simulation: Analytic Policy Gradients for Sample and Learning Efficient Differentiable Continuous Control Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-26T14:20:33.205947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T14:20:33.205947Z digest=sha256:cd1a82b0a0375491a29e6a6aacac2f6aef52773cc5a2e472ef519e26468ce6a5

Observation c5b64a82-a0d9-46d0-af05-f22cbdd51824 · outbound

This paper cites J., Simchowitz, M., Zhang, K., & Tedrake, R.

Backpropagating Through Simulation: Analytic Policy Gradients for Sample and Learning Efficient Differentiable Continuous Control J., Simchowitz, M., Zhang, K., & Tedrake, R

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-26T14:20:33.205947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T14:20:33.205947Z digest=sha256:a54505fc3494565e26fc95e99c0a1a0d3a87c58c393e230a0fe51cc5a2841584

Observation 7fe45bd8-c761-4246-8086-020d8749d0c6 · outbound

This paper cites an unresolved cited work.

Backpropagating Through Simulation: Analytic Policy Gradients for Sample and Learning Efficient Differentiable Continuous Control Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-26T14:20:33.205947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T14:20:33.205947Z digest=sha256:2b65cb89b188c0b20593269f05f759fddda587791755594bd6b7dcca6e2c0b1f

Observation 6bdc747e-037e-4e9a-bea3-674d58df19d8 · outbound

This paper cites an unresolved cited work.

Backpropagating Through Simulation: Analytic Policy Gradients for Sample and Learning Efficient Differentiable Continuous Control Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-26T14:20:33.205947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T14:20:33.205947Z digest=sha256:9cda006915a2e20c997cf0508a138bc8653f498b0059544dbccbc9de5c6ba9b3

Observation 79abbb9f-f16b-4ef1-9bfe-7cafb58d8abe · outbound

This paper cites an unresolved cited work.

Backpropagating Through Simulation: Analytic Policy Gradients for Sample and Learning Efficient Differentiable Continuous Control Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-26T14:20:33.205947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T14:20:33.205947Z digest=sha256:98efdcd4dca1ffc960d5b582ec64775929df6287e975f7855ae01b95dbb6b738

Observation c55b6024-98e8-43ac-bea1-6e4cc3da0ef0 · outbound

This paper cites an unresolved cited work.

Backpropagating Through Simulation: Analytic Policy Gradients for Sample and Learning Efficient Differentiable Continuous Control Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-26T14:20:33.205947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T14:20:33.205947Z digest=sha256:90f21b7d051dc3461dd6cff9700612ad8ea3409941aaf1d3f6bb18117a832ab9

Observation d9d3403b-ad82-42a1-b026-beb52232302a · outbound

This paper cites an unresolved cited work.

Backpropagating Through Simulation: Analytic Policy Gradients for Sample and Learning Efficient Differentiable Continuous Control Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-26T14:20:33.205947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T14:20:33.205947Z digest=sha256:bc9069be8a943f5012b7ab767474c050957cd5409ee033bb9b0bfd7c4e3b6599

Observation 0f2d1c84-56fc-49c7-95fe-7351fec531cd · outbound

This paper cites Brax -- A Differentiable Physics Engine for Large Scale Rigid Body Simulation.

Backpropagating Through Simulation: Analytic Policy Gradients for Sample and Learning Efficient Differentiable Continuous Control Brax -- A Differentiable Physics Engine for Large Scale Rigid Body Simulation

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T06:39:37.279781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-26T14:20:33.205947Z digest=sha256:852b063830e7b6508570371195b17ae01e2c1788d4031182de474be998af3b4e

Observation 9021b7a3-0341-49ab-b3ba-3f2ccc169d93 · outbound

This paper cites an unresolved cited work.

Backpropagating Through Simulation: Analytic Policy Gradients for Sample and Learning Efficient Differentiable Continuous Control Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-06-26T14:20:33.205947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T14:20:33.205947Z digest=sha256:a49dd72c811a074bf18feef51b9119b7d1726cc10a7d56558b7cda08c15646e6

Observation cc9684c6-0431-4f77-8f43-d47e55c01e80 · outbound

This paper cites an unresolved cited work.

Backpropagating Through Simulation: Analytic Policy Gradients for Sample and Learning Efficient Differentiable Continuous Control Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-26T14:20:33.205947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T14:20:33.205947Z digest=sha256:64bdd76d0499c1d5846fb1c8d0b096671a59e877a728f0df844b55863f059f07

Observation 92c0baed-3895-43a3-9ce5-ee5872f0ffe7 · outbound

This paper cites an unresolved cited work.

Backpropagating Through Simulation: Analytic Policy Gradients for Sample and Learning Efficient Differentiable Continuous Control Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-26T14:20:33.205947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T14:20:33.205947Z digest=sha256:8358d36c4f4304b9f502a58d74910b27d16ccc1255bdfd3037292519383bf217

Pith citing papers

No inbound Pith citation observations are available.