Pith. sign in

Paper Citation Record · LEDGER

Reinforcement Learning Applications

As of 16 August 2026, this Paper Citation Record lists 75 of 75 outbound references and 2 inbound Pith citation observations for arXiv:1908.06973.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1908.06973 v1

Coverage vector

measured 75 of 75 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T12:40:22.124497Z

measured 77 of 77 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T21:42:27.858298Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T12:51:23.223053Z

Reference resolution

75 of 75 outbound references displayed

  • verified exact0
  • verified fuzzy27
  • unresolved48
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8d5de21e-bc20-4786-9d34-b8ebae94d5e6 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:23.143881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.835878Z digest=sha256:a04a99759a21793cdaf9b0ec3e63e1930ae64066e802f76149a683e6918acc25

Observation b2fa2962-081f-4a39-9fed-bf52fec5d31c · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:23.027397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.840654Z digest=sha256:9f676d0dd8cd9df7a0187f36f5699cfeb934253f551bef40277c268dbb746459

Observation c8f3c4a1-d936-461f-bf39-694cc2d2a5c9 · outbound

This paper cites N., Boulanger, A., Powell, W.

Reinforcement Learning Applications N., Boulanger, A., Powell, W

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:23.014390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.844744Z digest=sha256:305aba3fdfab79b5196e0ca9892ad211025c1ba259769ac5f9fb3d2af4bafdd3

Observation 3d38c89d-47b9-4574-9e7d-b305f41e835b · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.999768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.848932Z digest=sha256:ec01b32ed0605a5bbe7e3d1b7e358f6b7cd4b6b20e29b2e4bafbdcafd8490d95

Observation 3e1bfc83-96f4-4175-a9d3-95943370e7a3 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.986506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.853340Z digest=sha256:00488792ee98c5f337d7e1837d5b4815a61d7a54804612e31a41e116a1f0a3c0

Observation 152e761e-ad3e-4853-86d5-e599c4a0fca2 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-14T12:40:21.857300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T12:40:21.857300Z digest=sha256:bf5e03f43b88731c9a9804050ad49bc6ca2916284979c88a9c42c9f991c2dfee

Observation fc92b3b2-e268-4b92-af2d-aed8742499f4 · outbound

This paper cites and Sandholm, T.

Reinforcement Learning Applications and Sandholm, T

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.964450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.861656Z digest=sha256:a462815e3b5c1b8d0fe4f2f7180b6cbfaa72ee55d18abce395cf28e726ab6b27

Observation 7a166ae5-5a6a-42ac-be3d-49f7cab8bcef · outbound

This paper cites and Murphy, S.

Reinforcement Learning Applications and Murphy, S

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.953087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.865315Z digest=sha256:92b053480e3504edeca82dbfcf099b7ae5ec67e99600a3baf795b5091e987ca9

Observation 87c8fb09-97a5-4a7f-89bc-893b93c06288 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.940712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.869656Z digest=sha256:901995ef2f45141c602af8253cb0c2d80a43613b6275cb09ff308bb89b46dfc5

Observation 49c9401e-c41b-4b55-a611-ce77a91723e0 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.929604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.873447Z digest=sha256:d549b49edbfd6d62c86176b405a393df70cef428715f26cbdc1c1cbc2662b26a

Observation 9d0bdea4-bb35-468b-a03c-55e8c85e00e0 · outbound

This paper cites D., Zoph, B., Mane, D., Vasudevan, V., and Le, Q.

Reinforcement Learning Applications D., Zoph, B., Mane, D., Vasudevan, V., and Le, Q

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.917531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.877832Z digest=sha256:c649e19afb57add88db2e458441fd2f6cc2a6d3ba8dc8991958d18dd4eda44c5

Observation 6c93710f-2311-4fb5-afa5-72068bf68879 · outbound

This paper cites B., Zhang, Y., Dilkina, B., and Song, L.

Reinforcement Learning Applications B., Zhang, Y., Dilkina, B., and Song, L

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.905808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.881878Z digest=sha256:407c226c2d1332821ed4a86371ea1a011e664253065575059c311dfa53a6612a

Observation 463046ca-9592-4bc5-b0c4-dd6c46d9a165 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.894293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.886878Z digest=sha256:98b24c35f822babd73a3d054154255fdac99937cb5d6b481dbc4df99a3ed69e6

Observation 611742e7-9d6e-4b6e-a341-c34bb13eaaff · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-14T12:40:21.890829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T12:40:21.890829Z digest=sha256:da12be46da8458313b3672997b1f2b7d67e19b9b1fa8979fafcb62e1859ffce3

Observation 83826547-c105-4102-b045-01edd474889e · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.875582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.894404Z digest=sha256:ce772d54b8da32f1f46ab176673b0038c95f002ef10823e7e35884c93ea1a9b9

Observation 13734b0a-b9f0-4ac9-a31a-2ff2d4068117 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.864220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.897820Z digest=sha256:eb23ef67fd3f772c39e991c371bee745b8b54bb4593868ea49fe136ec7887b56

Observation 9e379179-952b-410b-a831-bd7481080418 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.853237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.901526Z digest=sha256:26465c97bf7bb0a052b5d4ed00b724134a0111b4c6c05ede9830ed37d5065b6c

Observation ae226209-1cad-4794-ace5-5414851737a0 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.842992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.905515Z digest=sha256:01a215274d11d469de8ab4137434431c39d02ad89e8920f445bf8528c2f71560

Observation e878ab7e-2116-4779-a4b5-76269e6e3cb4 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.832353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.909206Z digest=sha256:b28dd0acab54e7c495e83941cc93e35047ab7d061e25d91ccbe883632d3ce976

Observation 7ecc0aad-5d68-4d38-93c7-938bff2ee876 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.821842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.913137Z digest=sha256:2295374f3671f49977910fd57c638ec8a9a2a375b117139f849f2cd8dab1ccd6

Observation 36f2ccea-793f-4dde-8d50-c82a6dd807bd · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.811486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.917458Z digest=sha256:d9a025953a009158ee93284c7d4ce994639421fcfebd2a0dc3b662fb19a8be96

Observation c947c9ac-8867-4827-80a5-20001064fc5e · outbound

This paper cites C., Wu, R., Jain, V., and Boutilier, C.

Reinforcement Learning Applications C., Wu, R., Jain, V., and Boutilier, C

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.801740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.921477Z digest=sha256:1079df4a9347d366d0f6d33a7ea6b1a3f260b761e07a50866ebf91181de5479b

Observation 1ab8381d-0174-47e9-87f0-4f049fe1c5a1 · outbound

This paper cites M., Dunning , I., Marris , L., Lever , G., Garcia Castaneda , A., Beattie , C., Rabinowitz , N.

Reinforcement Learning Applications M., Dunning , I., Marris , L., Lever , G., Garcia Castaneda , A., Beattie , C., Rabinowitz , N

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.791510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.925088Z digest=sha256:c3e4e167bc7b4c31fe8904e263abc9b8588ac6fdc58a7f72fe1056b199a59b79

Observation 23d2bfd1-cd5d-4dd7-a138-2e00258ce882 · outbound

This paper cites and Li, L.

Reinforcement Learning Applications and Li, L

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.781324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.928971Z digest=sha256:4ce30a914d2ab009d22a8c7271b0ee3f50a55bd16a847041c63eaf37a79cfb77

Observation 69f3b506-8e13-4a7c-9506-bbf10446b5bd · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.770671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.932389Z digest=sha256:80ceef8bcecc03e67cae388356621db5469611010ba8d9cab909c5615d628dd6

Observation b08465b3-7602-4be9-a3fb-718a85883f7f · outbound

This paper cites and Nevmyvaka, Y.

Reinforcement Learning Applications and Nevmyvaka, Y

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.759692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.936835Z digest=sha256:29e8659641257b9c6ae99d5799c057d3dafa6eabb70546841ae84fd32375b44e

Observation d9b26ee2-43d7-42e4-8430-f563b90a79e8 · outbound

This paper cites A., and Peters, J.

Reinforcement Learning Applications A., and Peters, J

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.748033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.940438Z digest=sha256:5b46eb384f5f22ceaa6ebcc872fa0c8a179ce4abd533c4e8a953b5e7cb05902a

Observation b4a20137-1cbb-459d-857e-4acbe07c9213 · outbound

This paper cites A., Badawi, O., Gordon, A.

Reinforcement Learning Applications A., Badawi, O., Gordon, A

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.737886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.944500Z digest=sha256:c37d23630d951fc172554dacd5dec910bd3a5f8f9992e98aa51664486eda8cc8

Observation 6f874d38-cc30-4629-b337-26c7ede96a1e · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.727565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.948206Z digest=sha256:b73da5e5013ba9b466148d6e4c90e67174110d77d8fc1481468e749b189d415d

Observation aee0027e-5860-46d2-af57-40112f5b114a · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.716518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.952701Z digest=sha256:09e40f4a4dfb086196e08d07a702fd66b6086a47ec98ce50c78e6cf82f8de05e

Observation 60879a6e-3d04-42ec-a2c0-6989f4a678c9 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.706584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.956712Z digest=sha256:18d74bfe53c7455f8f3e4c13ae0a926a993b185fc8e1931960a3f8391b4c7fbc

Observation 935d4ba6-33ff-412f-852c-452cece6ea81 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.696888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.960511Z digest=sha256:c9371a3e8522077cdd65da05f531578db83a4ff1f426a0072e6d8aa5f6ae1c73

Observation 551bbf30-c44b-449c-bd3d-1fe010d2b529 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.686384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.964371Z digest=sha256:dafeee2aadc2bafc0d2662d286a385eacfeb8a3dd7782e607a1cc66fa0f04a9a

Observation 69313164-da77-4117-9790-87853d0d9ae2 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.675488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.968244Z digest=sha256:a514e344b7a132540d1760d2b97f17927e0e6bd1838a6c935b71eca4c257d7ce

Observation e2dee02a-24b8-4600-933a-1dda564f02fd · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.664767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.972194Z digest=sha256:cb1d53dbde19bac7422c6db93bb1a99e1c0215b7f8110eb57b837743cbd98160

Observation fffd76cf-c271-44aa-93de-09c3eedb253e · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.653437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.975980Z digest=sha256:d6c5eaca3ff62800ba9c16def2d7479e415087d89a3b6e3106c4fb7d655e5b3b

Observation 28c74c6c-6127-4c28-acaa-c43baec39809 · outbound

This paper cites P., Hunt, J.

Reinforcement Learning Applications P., Hunt, J

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.641107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.980023Z digest=sha256:b7f59f28e3421f197980fc83fe8195d81b583b4f8b734ac70707adfb763f0ab8

Observation ef8a8d04-0d69-4e5e-96d4-94f23a44869f · outbound

This paper cites A., Doshi-Velez, F., and Brunskill, E.

Reinforcement Learning Applications A., Doshi-Velez, F., and Brunskill, E

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.627339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.983763Z digest=sha256:65cbd577e551d8bc3f57e274ab96e4580c0f5b36dff4e4fc5b4793ae0bade2fd

Observation 0ded4ecb-a1a6-4a69-b03e-f3bb0c25610d · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.613716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.987398Z digest=sha256:55c4960ec79a8d02823f7dbe8f0e06cac3de8aed2f9e7a2480c38aac96abb8e5

Observation 7bf2e4c9-47c9-4516-b49c-884eb358003c · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.601387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.991323Z digest=sha256:de38e5385115c5e594abd1111007c2e5e95a32ea0762b6b6995aaf509333620d

Observation da586111-243c-487d-bc91-704f8c62ef57 · outbound

This paper cites B., Meng, Z., and Alizadeh, M.

Reinforcement Learning Applications B., Meng, Z., and Alizadeh, M

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.588662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.995184Z digest=sha256:8d2a0fc332f896b39f28c32374aa583f2305a1dab106d114ad67ca2af7c3f0d9

Observation 9f1704ec-0574-4025-ae3b-4fabb80700f2 · outbound

This paper cites V., Steiner, B., Larsen, R., Zhou, Y., Kumar, N., and Mohammad Norouzi, Samy Bengio, J.

Reinforcement Learning Applications V., Steiner, B., Larsen, R., Zhou, Y., Kumar, N., and Mohammad Norouzi, Samy Bengio, J

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.575987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.999784Z digest=sha256:f5f3fd28d965b1b3d05758a745bd02fefd8aa98021750e4ace36b17e8fc366c0

Observation ff763e50-92fc-4f61-a738-69809879a51f · outbound

This paper cites P., Mirza, M., Graves, A., Harley, T., Lillicrap, T.

Reinforcement Learning Applications P., Mirza, M., Graves, A., Harley, T., Lillicrap, T

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.563652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.003855Z digest=sha256:18fd887c38ac2abad89fa9d28ff8b3ae61af349aac592fb286041323469fa7fc

Observation 90e254b6-6f00-4d6e-be6c-f8beb95fd22d · outbound

This paper cites A., Veness, J., Bellemare, M.

Reinforcement Learning Applications A., Veness, J., Bellemare, M

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.550798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.007717Z digest=sha256:8171261d869fe0646a8e0084b49115216c3cba699b4dd5133f156b62ec7ce256

Observation d5ff2226-252c-4e12-9497-7f5f3d972eae · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.538362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.011865Z digest=sha256:e3e6f7b36a6bb164178fb1dd9f4d3a460436e2122adc18ec32c728589155d4f5

Observation 11fceae3-5e8d-430c-acbc-16722dfaa10d · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.525880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.015521Z digest=sha256:ef77166f0ac470aa1e2b0bef04562294f9e5075d3a1bbbf185a4ced90aad33bf

Observation 669d6c9a-ae5e-40f6-a510-21d7c989deaa · outbound

This paper cites B., Abbeel, P., Levine, S., and van de Panne, M.

Reinforcement Learning Applications B., Abbeel, P., Levine, S., and van de Panne, M

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.513290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.019461Z digest=sha256:9ae7633aedd4f9ba20411c78f9c22235cb3d260f2127a8f10f7cdd449346817c

Observation d442b3c9-1a86-4a0c-ae89-23b10506291d · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.500751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.023429Z digest=sha256:98c6caa3f2492e4ddad1e9e7e96144a16ceb5923d82d9793e98e3e589795917c

Observation ae1a9668-2511-4862-b250-5379368c4d05 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.487743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.027261Z digest=sha256:72687a1dc32a89ceb17f7d21275f26dd2524b76aeb2d602623970ada050fe64b

Observation f82df0b7-0aec-49f2-b40c-428fc4906453 · outbound

This paper cites and Norvig, P.

Reinforcement Learning Applications and Norvig, P

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.474892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.030968Z digest=sha256:41ff4bcc7411c5f147f7c6e8e49b1bbdeba872b10d3bc2584d2f8744ea618cbd

Observation f1b96dd9-0b2a-4ed6-a5e6-80b9ef7d236e · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.462667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.034822Z digest=sha256:20f3f8064499f4f6a395ceb50218db2467295a78aeeb99068c9ee8784634b0cf

Observation 552ba320-9923-4924-bb6b-7f489d9860f8 · outbound

This paper cites I., and Abbeel, P.

Reinforcement Learning Applications I., and Abbeel, P

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.450402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.038541Z digest=sha256:d50cb5318ad60610d36dbc4af582f9f0c79992059a39576a0b1b4e0821bc8498

Observation 21c89e61-dc4c-4f16-ab7c-3984f0770d5b · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.437659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.042819Z digest=sha256:8afde96f6b6c447709b3525ebac3d5931662072fe0cf585876583dc07f12a494

Observation 14305707-00c4-41bf-934d-521c99fdd4b3 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.424606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.046595Z digest=sha256:ca9ec57f8cfd361553feea457d494083ee09f9491a0e7e9fc1e225ddc3920563

Observation e41f5094-d0de-4b34-9a1b-e5524106dcba · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.411163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.050187Z digest=sha256:c10c8854254f13f7da1f617133652b2369ab71f711edfec75b1e7ff8209f8657

Observation a9f79b30-77eb-4df2-b598-c197cebe5578 · outbound

This paper cites J., Guez, A., Sifre, L., Van Den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V., Lanctot, M., et al.

Reinforcement Learning Applications J., Guez, A., Sifre, L., Van Den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V., Lanctot, M., et al

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.398848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.053643Z digest=sha256:9ce5354a9ea7a4a7d5505323f13ae98387bd03f967c8922b9833b92e4e1ba03a

Observation 99ccd675-f600-4f5a-a442-aee690179335 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.385865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.057109Z digest=sha256:27aa84c125ad543f31e8bbd3507edbb404a19ea56b99b603aaec77d21934d5e6

Observation c45bd6ef-717d-4903-afb1-c216e5295f52 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.373349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.060456Z digest=sha256:e3500d4202fbfc87fd4c03e9bbd8a3e14b06364dd281b4a0ce23e94fb3feaea0

Observation 0c233aef-a378-4734-8269-85de0a9cd765 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.361675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.063965Z digest=sha256:faa7079e94428727791b3180b5f4af019912586d4d4ba141f20fc8da58cd43e7

Observation d66b9d3d-020b-4750-9be3-3a42f1c02730 · outbound

This paper cites P., Day, J., George, A.

Reinforcement Learning Applications P., Day, J., George, A

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.349653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.067707Z digest=sha256:fd477c5b6ed1b8166b8aeea56e745fad66c3f2376316828eb69a72867f7ecc94

Observation 33980519-d9a9-4f0e-953a-30d237173c7f · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.337719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.071620Z digest=sha256:6641f23c05ccd19443f2dacc4472386c83d744a17978e66fbd8ddff46dd0f7fb

Observation 83261320-9ac4-4fd1-b90c-7489df0920f7 · outbound

This paper cites S., Precup, D., and Singh, S.

Reinforcement Learning Applications S., Precup, D., and Singh, S

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.325787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.076311Z digest=sha256:7dec2ef5a03830ee10dcd8d0ce5fee6639c59d4f6bf2d0650ec698a6e4934e53

Observation bb7e01f6-400f-4e52-9cd2-d929dc086a4e · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.313181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.080103Z digest=sha256:e558ff20b9bf3ecd916eb1645a87abcdc01e6b7f41654d9873dec083a062b511

Observation 2fcf13ba-386d-4d24-9e08-104980e58208 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.301193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.083691Z digest=sha256:e44595de668326c94b4c5521044c501ec2e380b5b462b3f1f3af81b8b1315c80

Observation 790005f3-beaf-4f2c-a2a5-de67ddd9b014 · outbound

This paper cites S., and Ghavamzadeh, M.

Reinforcement Learning Applications S., and Ghavamzadeh, M

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.289421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.087511Z digest=sha256:09ceafbb9c65ca3624fe000cd95586249538255e2f683085d5a5c934b4018556

Observation df56106d-0007-4937-accd-f38a88acd0ca · outbound

This paper cites S., Theocharous, G., and Ghavamzadeh, M.

Reinforcement Learning Applications S., Theocharous, G., and Ghavamzadeh, M

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.276240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.091088Z digest=sha256:7a35f26a3f6a724d16e45a199548fb1a49bdeaca7381fbe801265c5f76a07efe

Observation 9501534a-3e46-4850-9eef-1a820ba3ab67 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.264108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.094763Z digest=sha256:5f10c37d0618bfe724e0fc26849025f50ee57b87c530247796741ab664ccb0bf

Observation 3228a9ae-9646-4456-8af9-ca74786d28d1 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 68

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.251524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.098400Z digest=sha256:10648a60440d358a3b780cddf268c3a6f7b9850af15e0bdf232d4473e609df2d

Observation dce5b22c-90df-41f0-8994-81c84effbbec · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 69

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.239243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.102357Z digest=sha256:b544c22765d55191a7169abec6649f622360f3504b3836e5027bb1910b0ff8ad

Observation 273a616b-81b6-4dbe-98b5-191f42ad5805 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 70

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.226372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.106008Z digest=sha256:97dadf5cab476706b7e547c7b50596b48336c7e54885858c31112412b46b0fb0

Observation d7424efb-79a5-4a34-8a27-6a47e3c6c0c7 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 71

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.212484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.109716Z digest=sha256:07534152d6b949e01e07b594aa878a853b6ed9f32885a4db082e03814f1165db

Observation d78c73af-a8cb-40b4-8c84-4a7c07acb80d · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 72

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.199497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.113277Z digest=sha256:6b235fdb29217ada42a532898d6611fe645d5ed6fe53610bef80d87b5bc1d168

Observation e7d24ac1-f203-4063-80dc-4f15aaa50656 · outbound

This paper cites J., Xie, X., and Li, Z.

Reinforcement Learning Applications J., Xie, X., and Li, Z

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.185969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.117055Z digest=sha256:ad988a61ef978182837f4e7b9b2287c5fc6f05c8551b5e28d30b133dde8e582f

Observation d0084f41-d087-455a-8be6-1f0cacb98c3b · outbound

This paper cites and Le, Q.

Reinforcement Learning Applications and Le, Q

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.172908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.120901Z digest=sha256:bfa1bf685cf17dfd4c0a8a1edadb7a282d45349eda481919c85ba482e3ecf9e1

Observation 5aa186e2-6904-47e4-a8a2-6047b715eacc · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 75

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.159610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.124497Z digest=sha256:3cbb10f8e6352bf448f49de3290f686b17a3aa49e9f05fbbfa07e7037ef4c507

Pith citing papers

Observation bdb219d2-1fad-41b0-82a3-33feeeeb6b2e · inbound

RLInspect: An Interactive Visual Approach to Assess Reinforcement Learning Algorithm cites this paper.

RLInspect: An Interactive Visual Approach to Assess Reinforcement Learning Algorithm Reinforcement Learning Applications

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T21:42:27.858298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:42:27.858298Z digest=sha256:630131ab14c836b70a8a7d948b4a89ec833ae5944fbeb9dc1ddddca261ae58f2

Observation 2037abcc-700d-4236-a301-d12e794391a3 · inbound

MDP modeling for multi-stage stochastic programs cites this paper.

MDP modeling for multi-stage stochastic programs Reinforcement Learning Applications

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:51:23.226190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-18T12:50:53.853410Z digest=sha256:dbcf91a0dca204939765a34a05a20758078f0b3e69bd671cbfdf0d6b30501c3c