Pith. sign in

Paper Citation Record · LEDGER

Reinforcement Learning Applications

As of 19 August 2026, this Paper Citation Record lists 75 of 75 outbound references and 2 inbound Pith citation observations for arXiv:1908.06973.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1908.06973 v1

Coverage vector

measured 75 of 75 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T12:40:22.124497Z

measured 77 of 77 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T21:42:27.858298Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T12:51:23.223053Z

Reference resolution

75 of 75 outbound references displayed

  • verified exact0
  • verified fuzzy27
  • unresolved48
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8d5de21e-bc20-4786-9d34-b8ebae94d5e6 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:23.143881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.835878Z digest=sha256:fa06f37787ce82235028426de8be33ef436c5cf04ee0f2640304350e1a44d472

Observation b2fa2962-081f-4a39-9fed-bf52fec5d31c · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:23.027397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.840654Z digest=sha256:964b01274e9d6477d89810d9a6128c278357b51551baf80258d88b6478d82a0b

Observation c8f3c4a1-d936-461f-bf39-694cc2d2a5c9 · outbound

This paper cites N., Boulanger, A., Powell, W.

Reinforcement Learning Applications N., Boulanger, A., Powell, W

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:23.014390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.844744Z digest=sha256:b5f23d54ae64aff8de87d817bd896de7c8523d31b42a9fcd93675c7b23bbe0fa

Observation 3d38c89d-47b9-4574-9e7d-b305f41e835b · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.999768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.848932Z digest=sha256:74eb3dc83b2951ee5ad66240ac83b5eca71210e85aa643136b171ee67caa3125

Observation 3e1bfc83-96f4-4175-a9d3-95943370e7a3 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.986506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.853340Z digest=sha256:72cf3486c659682b007a8b9e8d82668261ba89b554736f08650d9541836812ca

Observation 152e761e-ad3e-4853-86d5-e599c4a0fca2 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-14T12:40:21.857300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T12:40:21.857300Z digest=sha256:05128a41cfc75f764ef8ff3a13d3192e6a2cfe7d58972ae922ca926d4a5fc5b4

Observation fc92b3b2-e268-4b92-af2d-aed8742499f4 · outbound

This paper cites and Sandholm, T.

Reinforcement Learning Applications and Sandholm, T

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.964450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.861656Z digest=sha256:d0cb92718f3df42cbcb7b38cdefd1609928c7f018e86a0aa6d686e25a3ddad4e

Observation 7a166ae5-5a6a-42ac-be3d-49f7cab8bcef · outbound

This paper cites and Murphy, S.

Reinforcement Learning Applications and Murphy, S

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.953087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.865315Z digest=sha256:a7b5b66086500d7bbbdcf9f2a757677f84e39593ea9ad29921930b94edb70cd2

Observation 87c8fb09-97a5-4a7f-89bc-893b93c06288 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.940712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.869656Z digest=sha256:9d98f83df8f899da1097f7c09ce7d3e71e6bf120afcc96973fc7a4ec7e09783d

Observation 49c9401e-c41b-4b55-a611-ce77a91723e0 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.929604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.873447Z digest=sha256:c3cf8c3ecc22e4d6f00b15623c9221376e0fd5b605b5d3fb14f5d76a94aaaa68

Observation 9d0bdea4-bb35-468b-a03c-55e8c85e00e0 · outbound

This paper cites D., Zoph, B., Mane, D., Vasudevan, V., and Le, Q.

Reinforcement Learning Applications D., Zoph, B., Mane, D., Vasudevan, V., and Le, Q

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.917531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.877832Z digest=sha256:8bb63d8d954292a5bf0bf43b27bbd6f10640125ceae04be028a04d23e441af05

Observation 6c93710f-2311-4fb5-afa5-72068bf68879 · outbound

This paper cites B., Zhang, Y., Dilkina, B., and Song, L.

Reinforcement Learning Applications B., Zhang, Y., Dilkina, B., and Song, L

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.905808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.881878Z digest=sha256:4b3b9b321daf79762e2de1c84c6753a75f2191cb837554962ac08b579d7dcf96

Observation 463046ca-9592-4bc5-b0c4-dd6c46d9a165 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.894293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.886878Z digest=sha256:aed79a92c158228c2efdf66395670976bd367fbe00f7d0d11f61d200393bf8c6

Observation 611742e7-9d6e-4b6e-a341-c34bb13eaaff · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-14T12:40:21.890829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T12:40:21.890829Z digest=sha256:4ab5c58064b971542f4e786278a8744133e674860920dbe45a6104e32efd53e7

Observation 83826547-c105-4102-b045-01edd474889e · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.875582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.894404Z digest=sha256:1d910e6d0278ea2d081f20309f519da0c51b8ea9b366b321a505873f637c3f6d

Observation 13734b0a-b9f0-4ac9-a31a-2ff2d4068117 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.864220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.897820Z digest=sha256:bb188eb034b584a8897cab2dea4b14dc9d41faced2f6b1e891aa6de18c06a9b2

Observation 9e379179-952b-410b-a831-bd7481080418 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.853237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.901526Z digest=sha256:34e60be761668b905ef820268f98ad24e3d84faadfeb7e7d233fa7ff586b9d0a

Observation ae226209-1cad-4794-ace5-5414851737a0 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.842992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.905515Z digest=sha256:cf5c45d3064473899dd288b69c3e016ae3dc8fa5d4dcb919363e5afe0d932742

Observation e878ab7e-2116-4779-a4b5-76269e6e3cb4 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.832353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.909206Z digest=sha256:81f617fd5310ea159bc1db66d1993cc7053cd4118b7d0b2578b2cfc6cbc28196

Observation 7ecc0aad-5d68-4d38-93c7-938bff2ee876 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.821842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.913137Z digest=sha256:48a09d73b22f869323384d9f98a97c4602118b65fdcfa8dad384284c8ac7aaf9

Observation 36f2ccea-793f-4dde-8d50-c82a6dd807bd · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.811486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.917458Z digest=sha256:3e181c46b21b098394e1b751bcae0b83c00e6a3b066877a049fe93ce75207be7

Observation c947c9ac-8867-4827-80a5-20001064fc5e · outbound

This paper cites C., Wu, R., Jain, V., and Boutilier, C.

Reinforcement Learning Applications C., Wu, R., Jain, V., and Boutilier, C

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.801740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.921477Z digest=sha256:87764c3b6f33464a7a0102ac560e849cc017e127b079c6702eb55723efa858ce

Observation 1ab8381d-0174-47e9-87f0-4f049fe1c5a1 · outbound

This paper cites M., Dunning , I., Marris , L., Lever , G., Garcia Castaneda , A., Beattie , C., Rabinowitz , N.

Reinforcement Learning Applications M., Dunning , I., Marris , L., Lever , G., Garcia Castaneda , A., Beattie , C., Rabinowitz , N

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.791510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.925088Z digest=sha256:bb8ffb24ad34c0c064019b1aa8c0e1deb1cc560b88d6234664a2079c5cba580c

Observation 23d2bfd1-cd5d-4dd7-a138-2e00258ce882 · outbound

This paper cites and Li, L.

Reinforcement Learning Applications and Li, L

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.781324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.928971Z digest=sha256:3cee13feea09c119a2a8e6c22a459430146ef14730454dbcb2a7dc3c107ab0e3

Observation 69f3b506-8e13-4a7c-9506-bbf10446b5bd · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.770671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.932389Z digest=sha256:ccdcc5c2f93972d85c57c3f6db2a6c85008831b6247390368b20ba76d4421cba

Observation b08465b3-7602-4be9-a3fb-718a85883f7f · outbound

This paper cites and Nevmyvaka, Y.

Reinforcement Learning Applications and Nevmyvaka, Y

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.759692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.936835Z digest=sha256:54276b0a96718a1c06908bb439685d8e264bfe206b8d517fdfffa2590b78ea33

Observation d9b26ee2-43d7-42e4-8430-f563b90a79e8 · outbound

This paper cites A., and Peters, J.

Reinforcement Learning Applications A., and Peters, J

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.748033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.940438Z digest=sha256:3c5dc69ea7f5df7a7b1083448c8e32d668519455524d560fb4205b43d123dd58

Observation b4a20137-1cbb-459d-857e-4acbe07c9213 · outbound

This paper cites A., Badawi, O., Gordon, A.

Reinforcement Learning Applications A., Badawi, O., Gordon, A

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.737886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.944500Z digest=sha256:0824b9c07f167cef4379dc477a8ad577bcec0846dbb25b94b19a80e1de8b620d

Observation 6f874d38-cc30-4629-b337-26c7ede96a1e · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.727565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.948206Z digest=sha256:583d439070e8fcc11a9fb477454cb0a5b28b24f6ee92f12b61b520cb1f331457

Observation aee0027e-5860-46d2-af57-40112f5b114a · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.716518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.952701Z digest=sha256:df3b9ab38f30ef86fdf0d6be0aa3d72ef1669255d1366cf70b4950512dc96f98

Observation 60879a6e-3d04-42ec-a2c0-6989f4a678c9 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.706584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.956712Z digest=sha256:67f5dc59355ffbd104ab30b08419d998b521a9fada6e9f764c7e355e3ae91529

Observation 935d4ba6-33ff-412f-852c-452cece6ea81 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.696888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.960511Z digest=sha256:9c34cdddcaa55a55db090c595d1eb33de300d0690691761d1295138165d37039

Observation 551bbf30-c44b-449c-bd3d-1fe010d2b529 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.686384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.964371Z digest=sha256:c5ded4fb6d7f1b4d9a83a3916444567a89c9b00dd46150813dbff227dbefcf9f

Observation 69313164-da77-4117-9790-87853d0d9ae2 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.675488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.968244Z digest=sha256:b1023ef31a0db367c0fb90dc7caf514c204f323a70153cdaf5198305072b2db8

Observation e2dee02a-24b8-4600-933a-1dda564f02fd · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.664767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.972194Z digest=sha256:a56a1c14dde54f29e75c86013223564d635f2070911e5b646bbf7597593ae93d

Observation fffd76cf-c271-44aa-93de-09c3eedb253e · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.653437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.975980Z digest=sha256:f86d0954e4c035a7e4b43952c55717f6e2d5bb59dd9c70ba8a233e39ded3a25d

Observation 28c74c6c-6127-4c28-acaa-c43baec39809 · outbound

This paper cites P., Hunt, J.

Reinforcement Learning Applications P., Hunt, J

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.641107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.980023Z digest=sha256:b329fdbe95295b81e69ca90a2847b56e78ecc85e8769e7bde04c5f6e09eddef8

Observation ef8a8d04-0d69-4e5e-96d4-94f23a44869f · outbound

This paper cites A., Doshi-Velez, F., and Brunskill, E.

Reinforcement Learning Applications A., Doshi-Velez, F., and Brunskill, E

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.627339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.983763Z digest=sha256:a87d71158fb3cb3a88a91e7c780501afe711c5fc5870de8b509692a0b1e627a5

Observation 0ded4ecb-a1a6-4a69-b03e-f3bb0c25610d · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.613716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.987398Z digest=sha256:e4f35d1ea57ed21982a989d66853ea2ed7446e19f99c96b6d8d336195bd5f5fc

Observation 7bf2e4c9-47c9-4516-b49c-884eb358003c · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.601387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.991323Z digest=sha256:79d971865863f5b8ffec3caf1bce86428a26e042a527d9ac219fadf0f2913866

Observation da586111-243c-487d-bc91-704f8c62ef57 · outbound

This paper cites B., Meng, Z., and Alizadeh, M.

Reinforcement Learning Applications B., Meng, Z., and Alizadeh, M

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.588662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.995184Z digest=sha256:2ae68e8d5b4f484aa98d5cbbc20658d3db8646e4194a4a6c2e50d48ee4412c2a

Observation 9f1704ec-0574-4025-ae3b-4fabb80700f2 · outbound

This paper cites V., Steiner, B., Larsen, R., Zhou, Y., Kumar, N., and Mohammad Norouzi, Samy Bengio, J.

Reinforcement Learning Applications V., Steiner, B., Larsen, R., Zhou, Y., Kumar, N., and Mohammad Norouzi, Samy Bengio, J

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.575987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:21.999784Z digest=sha256:2ab16702f64d73af176cffe8350665012ccc05007d4abd486b20dc2739529d69

Observation ff763e50-92fc-4f61-a738-69809879a51f · outbound

This paper cites P., Mirza, M., Graves, A., Harley, T., Lillicrap, T.

Reinforcement Learning Applications P., Mirza, M., Graves, A., Harley, T., Lillicrap, T

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.563652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.003855Z digest=sha256:4f84219b7063d054bf78b41f1eca03b7620e57767e4bbd149d6194a731d20d25

Observation 90e254b6-6f00-4d6e-be6c-f8beb95fd22d · outbound

This paper cites A., Veness, J., Bellemare, M.

Reinforcement Learning Applications A., Veness, J., Bellemare, M

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.550798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.007717Z digest=sha256:1685a1cf82cd36df56da04aed99ca2654a8cb5f77698cc519ce5ad1d61369331

Observation d5ff2226-252c-4e12-9497-7f5f3d972eae · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.538362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.011865Z digest=sha256:99dcf4f6875e40ce41b536c09524e21055053bb77772204d07b71a8220f06a08

Observation 11fceae3-5e8d-430c-acbc-16722dfaa10d · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.525880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.015521Z digest=sha256:41d6987dd0745011945941d16e8ced58dfed36a0bff4e7b79eba095256c9b4e8

Observation 669d6c9a-ae5e-40f6-a510-21d7c989deaa · outbound

This paper cites B., Abbeel, P., Levine, S., and van de Panne, M.

Reinforcement Learning Applications B., Abbeel, P., Levine, S., and van de Panne, M

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.513290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.019461Z digest=sha256:a84a47e543ac2b575ec6fb8524ae329c96e9127fa04dd251372e83dabbf7cac7

Observation d442b3c9-1a86-4a0c-ae89-23b10506291d · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.500751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.023429Z digest=sha256:908235e8671721750eb1a0da2bcd43d2ea447f0f7c9d001f62f7ba2b6748572a

Observation ae1a9668-2511-4862-b250-5379368c4d05 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.487743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.027261Z digest=sha256:15ab7f40094adab4c7c126109aceeebea135265f17d3a19718d56616f66e2430

Observation f82df0b7-0aec-49f2-b40c-428fc4906453 · outbound

This paper cites and Norvig, P.

Reinforcement Learning Applications and Norvig, P

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.474892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.030968Z digest=sha256:e78e7e63782d6b354a2d850db90c3bde21f19a4b3b2d97fb9a1ea8af0d07c52f

Observation f1b96dd9-0b2a-4ed6-a5e6-80b9ef7d236e · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.462667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.034822Z digest=sha256:8d380cff49493d84178afd59e5c4c23276dee796b0f9efc21e773ba71781b002

Observation 552ba320-9923-4924-bb6b-7f489d9860f8 · outbound

This paper cites I., and Abbeel, P.

Reinforcement Learning Applications I., and Abbeel, P

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.450402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.038541Z digest=sha256:19ed86848953dc1d787446e9c91a09c31b05d222e69d483335c4334c1c4a9b8f

Observation 21c89e61-dc4c-4f16-ab7c-3984f0770d5b · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.437659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.042819Z digest=sha256:8d8d37c4fc4d91bf9498d6554d5d89f4e87920d3a3039cbe4a23219a69dedac8

Observation 14305707-00c4-41bf-934d-521c99fdd4b3 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.424606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.046595Z digest=sha256:83764779a2b6ff8fd3f45d50574805536761f0b522df318a90cc9fd1de4717e6

Observation e41f5094-d0de-4b34-9a1b-e5524106dcba · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.411163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.050187Z digest=sha256:304c5c5453836a999852bd8b066439bb186edb52d7a78590d16d418df07e8b82

Observation a9f79b30-77eb-4df2-b598-c197cebe5578 · outbound

This paper cites J., Guez, A., Sifre, L., Van Den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V., Lanctot, M., et al.

Reinforcement Learning Applications J., Guez, A., Sifre, L., Van Den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V., Lanctot, M., et al

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.398848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.053643Z digest=sha256:09e1905a301a92390964a3cb74ca4e71c7951b35d8700611a882da6ee4d2d1ae

Observation 99ccd675-f600-4f5a-a442-aee690179335 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.385865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.057109Z digest=sha256:3d3caa19c103bbc3c8e7b346ff19628fe50f1978576366eb52035b58ab4a724f

Observation c45bd6ef-717d-4903-afb1-c216e5295f52 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.373349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.060456Z digest=sha256:d8918f03fa5f2ed831cda08c8ce58197ad5f281d6f6ac88e1515133fce691241

Observation 0c233aef-a378-4734-8269-85de0a9cd765 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.361675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.063965Z digest=sha256:e5193b0c600a0668e1ee7098af6663b88a04833b9464ab12f7735fbb1f84a28f

Observation d66b9d3d-020b-4750-9be3-3a42f1c02730 · outbound

This paper cites P., Day, J., George, A.

Reinforcement Learning Applications P., Day, J., George, A

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.349653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.067707Z digest=sha256:2d1f3a5449c3aaedd145008131d85ee83171d3211efbf1603e8b6b6a3fa00b7a

Observation 33980519-d9a9-4f0e-953a-30d237173c7f · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.337719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.071620Z digest=sha256:fa86547548760a17b16b5996f63b4108d68bf31531077637f40279fb9d998eb7

Observation 83261320-9ac4-4fd1-b90c-7489df0920f7 · outbound

This paper cites S., Precup, D., and Singh, S.

Reinforcement Learning Applications S., Precup, D., and Singh, S

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.325787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.076311Z digest=sha256:a6df82a0c16de423c152dd621c151ae00cada7102adc47d558cf49ff81bdb013

Observation bb7e01f6-400f-4e52-9cd2-d929dc086a4e · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.313181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.080103Z digest=sha256:5e0092ed6a8c691b869e48ac363f4d67508a140e948d30fbf7fe123e493d105c

Observation 2fcf13ba-386d-4d24-9e08-104980e58208 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.301193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.083691Z digest=sha256:b42d12587c930a3bae733efee576b3ee4070dc836e0407e74b56964475162267

Observation 790005f3-beaf-4f2c-a2a5-de67ddd9b014 · outbound

This paper cites S., and Ghavamzadeh, M.

Reinforcement Learning Applications S., and Ghavamzadeh, M

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.289421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.087511Z digest=sha256:debfb03e660b8d3f880547dca549dfad63acd8cf1ee41b4df04735e1111b6766

Observation df56106d-0007-4937-accd-f38a88acd0ca · outbound

This paper cites S., Theocharous, G., and Ghavamzadeh, M.

Reinforcement Learning Applications S., Theocharous, G., and Ghavamzadeh, M

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.276240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.091088Z digest=sha256:00bf596c5636c594b2488622080c127d502e7e902149df82a1ecf2f7a44db4d1

Observation 9501534a-3e46-4850-9eef-1a820ba3ab67 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.264108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.094763Z digest=sha256:95a16a7deff46a1f7e97383408f463f916fea20a134c3f9268eb7dbbf6d7b850

Observation 3228a9ae-9646-4456-8af9-ca74786d28d1 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 68

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.251524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.098400Z digest=sha256:defa843ff7d88318d7e6a7258e98958467e4fd3b5cd15345c3f4368ed60f7860

Observation dce5b22c-90df-41f0-8994-81c84effbbec · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 69

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.239243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.102357Z digest=sha256:adacaf551cde9a1be1a0cb3522d5bb59b1b526fb9b320df784732bf121d892e3

Observation 273a616b-81b6-4dbe-98b5-191f42ad5805 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 70

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.226372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.106008Z digest=sha256:22083934002a98215c33ccfa7954e52134ea37c0c25657edb1a76b4b7335ac50

Observation d7424efb-79a5-4a34-8a27-6a47e3c6c0c7 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 71

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.212484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.109716Z digest=sha256:d40d74da32949bd242084f21fa326f730628c6b65543e6b4b64f294bb282a860

Observation d78c73af-a8cb-40b4-8c84-4a7c07acb80d · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 72

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.199497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.113277Z digest=sha256:c490b900494e7deea608a447873e1f9d947889d5362471f341c101c342fac651

Observation e7d24ac1-f203-4063-80dc-4f15aaa50656 · outbound

This paper cites J., Xie, X., and Li, Z.

Reinforcement Learning Applications J., Xie, X., and Li, Z

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.185969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.117055Z digest=sha256:d11e85d3225bb718c3f60b7b5cd803c5b8a4382bac3fc74eb1c29c88a4620599

Observation d0084f41-d087-455a-8be6-1f0cacb98c3b · outbound

This paper cites and Le, Q.

Reinforcement Learning Applications and Le, Q

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:40:22.172908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.120901Z digest=sha256:50e7fce5d44edcc39c3634231d58e139e582e7099286a8249da27e000326d5ce

Observation 5aa186e2-6904-47e4-a8a2-6047b715eacc · outbound

This paper cites an unresolved cited work.

Reinforcement Learning Applications Unresolved cited work

Reference 75

Resolution
unresolved
raw_fallback, observed 2026-08-14T12:40:22.159610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T12:40:22.124497Z digest=sha256:19a5d1c676384d422adbdf5dbed67260f9ef0a65f0a5961887a1fa0218bc5387

Pith citing papers

Observation bdb219d2-1fad-41b0-82a3-33feeeeb6b2e · inbound

RLInspect: An Interactive Visual Approach to Assess Reinforcement Learning Algorithm cites this paper.

RLInspect: An Interactive Visual Approach to Assess Reinforcement Learning Algorithm Reinforcement Learning Applications

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T21:42:27.858298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:42:27.858298Z digest=sha256:304fbf3be405ad944ff4903372f963233c13247bea4f8a57d776ef00f0f56df0

Observation 2037abcc-700d-4236-a301-d12e794391a3 · inbound

MDP modeling for multi-stage stochastic programs cites this paper.

MDP modeling for multi-stage stochastic programs Reinforcement Learning Applications

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:51:23.226190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-18T12:50:53.853410Z digest=sha256:1c0603fe0294213bc7dcaca157da82c5547e44e4a4845a6b0a6e7e08a01f7a87