Pith. sign in

Paper Citation Record · LEDGER

A View on Deep Reinforcement Learning in System Optimization

As of 16 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:1908.01275.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1908.01275 v3

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T15:21:50.582390Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

31 of 31 outbound references displayed

  • verified exact2
  • verified fuzzy11
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 00e02884-7e5a-4207-9575-025b7792fcc5 · outbound

This paper cites Placeto: Learning Generalizable Device Placement Algorithms for Distributed Machine Learning.

A View on Deep Reinforcement Learning in System Optimization Placeto: Learning Generalizable Device Placement Algorithms for Distributed Machine Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.428466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.428466Z digest=sha256:d79c0c8776acc1e3e084c3864f068d420e796dd92225bec622a1c31c4a02a8d9

Observation 6e28d589-6bc8-46f8-ab5f-3a54c1d15dc2 · outbound

This paper cites Speech recog- nition with deep recurrent neural networks.

A View on Deep Reinforcement Learning in System Optimization Speech recog- nition with deep recurrent neural networks

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:21:51.127842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.455907Z digest=sha256:67a43a479ad8970e6e270c47f8a46ee6e15dcfcbb6d50509dba45ee538e42f71

Observation 228359e0-e55e-4081-8ab8-5bb7f1b6de37 · outbound

This paper cites From Language to Programs: Bridging Reinforcement Learning and Maximum Marginal Likelihood.

A View on Deep Reinforcement Learning in System Optimization From Language to Programs: Bridging Reinforcement Learning and Maximum Marginal Likelihood

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.460868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.460868Z digest=sha256:8c6af587821abb208155f48d312fc45109cfaa6596dcaa3e9875d2cea98807e7

Observation 3910cf45-f41f-41a6-bafc-b28f6d51451f · outbound

This paper cites Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor.

A View on Deep Reinforcement Learning in System Optimization Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.465908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.465908Z digest=sha256:1fbe41d7f4094f2a2e708660d6347298b39f794b5415fca3d1e78d22c14e699d

Observation 22f86071-80f4-4f34-acde-68180cfe3e0f · outbound

This paper cites Autophase: Compiler phase-ordering for hls with deep reinforcement learn- ing.

A View on Deep Reinforcement Learning in System Optimization Autophase: Compiler phase-ordering for hls with deep reinforcement learn- ing

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:21:51.111257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.471075Z digest=sha256:f89f9cc2a80c788a7a3152f2889d59af224ca9a75d6977fb638f5a4973071409

Observation 44a93076-ad39-4c38-afd4-a6c552b3e9c3 · outbound

This paper cites Learning- based and data-driven tcp design for memory-constrained iot.

A View on Deep Reinforcement Learning in System Optimization Learning- based and data-driven tcp design for memory-constrained iot

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:21:51.079314Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.486859Z digest=sha256:43ed644122deb75d5860b8d84739f3413e8ebe37df121eb1e0d9f951368eaceb

Observation 68442d96-dc64-45ac-8e18-5cbbf335b371 · outbound

This paper cites RLlib: Abstractions for Distributed Reinforcement Learning.

A View on Deep Reinforcement Learning in System Optimization RLlib: Abstractions for Distributed Reinforcement Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.497388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.497388Z digest=sha256:ab1b537d3f93f99db877cbb61e9da9cbbc760a021660573bd3a2f24f8cc0d4cf

Observation f883fe50-ff45-4c88-81be-b640f2782c62 · outbound

This paper cites Neural Packet Classification.

A View on Deep Reinforcement Learning in System Optimization Neural Packet Classification

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-14T15:21:50.793436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.502401Z digest=sha256:8dd2fec8bc61cb48ea73eac70045a83a4c14ec6a0b017157a937dd47ac85775a

Observation 46656035-c6da-4682-83ac-1e2136a7d39c · outbound

This paper cites Neo: A Learned Query Optimizer.

A View on Deep Reinforcement Learning in System Optimization Neo: A Learned Query Optimizer

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.518247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.518247Z digest=sha256:076169b9d8c1287680342454cd7b80f72150185c18cec7e67053ed74fb05c6b3

Observation 65ba32d6-babe-45d0-b023-20fd4d9ac686 · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

A View on Deep Reinforcement Learning in System Optimization Playing Atari with Deep Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.523130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.523130Z digest=sha256:477f06663f04ea64422c1804991e3a122e2734d38c84d017cde62ece515bc9a8

Observation 3eade8a4-45eb-46f3-b4b5-7b286ecf3f1b · outbound

This paper cites P., Mirza, M., Graves, A., Lillicrap, T., Harley, T., Silver, D., and Kavukcuoglu, K.

A View on Deep Reinforcement Learning in System Optimization P., Mirza, M., Graves, A., Lillicrap, T., Harley, T., Silver, D., and Kavukcuoglu, K

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:21:51.029564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.528120Z digest=sha256:ae4334926bd521e5348a75a8c7dd78d0d0f0d434733afd72a9546abab050f65a

Observation 6549131c-32b3-4b7a-8ed2-e46873144e3f · outbound

This paper cites A Stochastic Approximation Approach for Foresighted Task Scheduling in Cloud Computing.

A View on Deep Reinforcement Learning in System Optimization A Stochastic Approximation Approach for Foresighted Task Scheduling in Cloud Computing

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-14T15:21:50.718986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.533011Z digest=sha256:39f9b683b96110666bbbe139389db4e86b2df106159f908128c0ae64f59d1706

Observation 8e44b00e-10fd-41b8-994a-7e5e2f08af8e · outbound

This paper cites Learning State Representations for Query Optimization with Deep Reinforcement Learning.

A View on Deep Reinforcement Learning in System Optimization Learning State Representations for Query Optimization with Deep Reinforcement Learning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.537780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.537780Z digest=sha256:6303900f7bdcf745c576c686d889659d37121f10ba92615a3792466b29e71604

Observation 0b3c0038-9114-44d9-9009-fc73b5eb699e · outbound

This paper cites Reinforced Genetic Algorithm Learning for Optimizing Computation Graphs.

A View on Deep Reinforcement Learning in System Optimization Reinforced Genetic Algorithm Learning for Optimizing Computation Graphs

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.542568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.542568Z digest=sha256:69a3170a80f177f1a0ea41f0d4ee31ee0d96eafb316115ae05a8386d9eff7984

Observation 30458130-2d91-47b1-9a74-eeeeb3009117 · outbound

This paper cites Semantic locality and context-based prefetching using reinforce- ment learning.

A View on Deep Reinforcement Learning in System Optimization Semantic locality and context-based prefetching using reinforce- ment learning

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:21:51.010966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.547836Z digest=sha256:dd20c5b4144bc1f7400b2f64e006f96d89ed35f0b7863a0532d187b0d513ec8b

Observation 257fb1fb-ca77-40c9-9dbc-e75de8397943 · outbound

This paper cites P., Obraczka, K., Burleigh, S., and Hirata, C.

A View on Deep Reinforcement Learning in System Optimization P., Obraczka, K., Burleigh, S., and Hirata, C

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:21:50.993127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.563366Z digest=sha256:3822da22954b5387cec403aadf5bde6a83f9ffcb49c69b94e2ca041034caefaa

Observation ebb2c584-f104-445b-b136-84ff63d22609 · outbound

This paper cites Machine learning in compiler optimization.

A View on Deep Reinforcement Learning in System Optimization Machine learning in compiler optimization

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:21:50.976831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.568188Z digest=sha256:cab1583d0650ac6dc7c90d77a34f002a007dbbad93c3a8b4574617e44fec3eef

Observation 8daf851d-f838-4a05-b827-3c4f25a8cabe · outbound

This paper cites Seq2SQL: Generating Structured Queries from Natural Language using Reinforcement Learning.

A View on Deep Reinforcement Learning in System Optimization Seq2SQL: Generating Structured Queries from Natural Language using Reinforcement Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.577434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.577434Z digest=sha256:8c46c56c3afe8b04cc48fd0579bf991ab2f1b395df0871c5a8dc4deff3f35071

Observation 6ef79119-f153-434c-99ea-3f3fb5711f02 · outbound

This paper cites A reinforcement learning approach to automatic error recovery.

A View on Deep Reinforcement Learning in System Optimization A reinforcement learning approach to automatic error recovery

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:21:50.943948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.582390Z digest=sha256:2e8cef3503b846345c8b2ac188956ea34fe69252d9ce4c0ed51afa0350c9832f

Observation 353e3208-c3f6-4513-a045-991b4a3c2e96 · outbound

This paper cites Proximal Policy Optimization Algorithms.

A View on Deep Reinforcement Learning in System Optimization Proximal Policy Optimization Algorithms

Reference 1999

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.558300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.558300Z digest=sha256:bbc60f376f9902aea1ed6635a716e190a5a6be5361a2bea910eb7fec74c6b614

Observation ce5c9204-5520-4d21-ab6d-857cba63bb6b · outbound

This paper cites Energy-efficient virtual machines consolidation in cloud data centers us- ing reinforcement learning.

A View on Deep Reinforcement Learning in System Optimization Energy-efficient virtual machines consolidation in cloud data centers us- ing reinforcement learning

Reference 2000

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:21:51.143986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.445112Z digest=sha256:5166fc334ed7aa52fe4658643612acb4cffbe29647f617bd39d841d32170d295

Observation dd34cad1-ffdd-4823-b017-c9bc726f552f · outbound

This paper cites Learning to Optimize Join Queries With Deep Reinforcement Learning.

A View on Deep Reinforcement Learning in System Optimization Learning to Optimize Join Queries With Deep Reinforcement Learning

Reference 2002

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.481128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.481128Z digest=sha256:a460c4a2babe8f4214e2837ea89e3b26f99e548ed5a703cc0c0390eb4d14e095

Observation dca263bd-c607-4c3f-84da-2ad2d4d02024 · outbound

This paper cites M., Pahl, C., Metzger, A., and Estrada, G.

A View on Deep Reinforcement Learning in System Optimization M., Pahl, C., Metzger, A., and Estrada, G

Reference 2008

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:21:51.095317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.476129Z digest=sha256:545ca0c94594c7b67dc40de2ae169750b57e6183a6dfab858e26bdccbb1f82bd

Observation 3f65f03d-e9c0-4e6b-8269-d58db559c0ac · outbound

This paper cites Iroko: A Framework to Prototype Reinforcement Learning for Data Center Traffic Control.

A View on Deep Reinforcement Learning in System Optimization Iroko: A Framework to Prototype Reinforcement Learning for Data Center Traffic Control

Reference 2011

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.553187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.553187Z digest=sha256:4c5409f58aba7bbeb2277803a30eada3d5a5e01d884f21ff6658b1a6f85daebd

Observation 621689b6-c168-4c04-97de-037d4b72e670 · outbound

This paper cites an unresolved cited work.

A View on Deep Reinforcement Learning in System Optimization Unresolved cited work

Reference 2012

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:21:50.960793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.572810Z digest=sha256:928e9a18439c7c0e8b4dde52bd493e5aeeec435834ab678de08892da2a312a19

Observation 85855dde-fd10-4204-9fbf-a17f6c1e966d · outbound

This paper cites A hierarchical framework of cloud resource allo- cation and power management using deep reinforcement learning.

A View on Deep Reinforcement Learning in System Optimization A hierarchical framework of cloud resource allo- cation and power management using deep reinforcement learning

Reference 2013

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:21:51.047557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.512996Z digest=sha256:53bf483f0c86fed5ee92af6d85b83a6dab74cbd31b9bfb3ada66504f9dd47fbe

Observation 5ce10ee3-c0b1-47f6-9238-b612a0de9e03 · outbound

This paper cites Horizon: Facebook's Open Source Applied Reinforcement Learning Platform.

A View on Deep Reinforcement Learning in System Optimization Horizon: Facebook's Open Source Applied Reinforcement Learning Platform

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.450088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.450088Z digest=sha256:2dcbe926ebf07776d4d3bc8f70752176a7a418bfa9b624fb7d302baa6a9cb350

Observation e67973b2-5a57-4a21-9d62-e75a6d1f27bf · outbound

This paper cites Neural Symbolic Machines: Learning Semantic Parsers on Freebase with Weak Supervision.

A View on Deep Reinforcement Learning in System Optimization Neural Symbolic Machines: Learning Semantic Parsers on Freebase with Weak Supervision

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.492154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.492154Z digest=sha256:e4bd097f22e506bce443bbdef60a7dcb44d8d3104aaa7f98dd1ede115f77e5f5

Observation aca0b995-cb4e-4dd7-9516-27df91712d13 · outbound

This paper cites Castro, P.

A View on Deep Reinforcement Learning in System Optimization Castro, P

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.434077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.434077Z digest=sha256:7aade408bcb371f7cdde683bb3a0f04b0b75bcd34b35431eb1b2f1b6bd75ae29

Observation fe4bf235-4093-4328-a5d4-39a4d09e56f5 · outbound

This paper cites Dopamine: A Research Framework for Deep Reinforcement Learning.

A View on Deep Reinforcement Learning in System Optimization Dopamine: A Research Framework for Deep Reinforcement Learning

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.439294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.439294Z digest=sha256:8a08fefd0f0cd51f9199b99bdb7a95bcd71c4d720a1c2edcf27d96ddca8c8b5a

Observation 0020236b-a2b5-42d6-9bfd-54c9f8b01aca · outbound

This paper cites Continuous control with deep reinforcement learning.

A View on Deep Reinforcement Learning in System Optimization Continuous control with deep reinforcement learning

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.507654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.507654Z digest=sha256:9099ced7a958695cc592f237dd4f340825228d69f7819ddc8acdf457dd885a90

Pith citing papers

No inbound Pith citation observations are available.