Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T16:56:51.117654Z
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 0 inbound Pith citation observations for arXiv:2501.12703.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T16:56:51.117654Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
30 of 30 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 851083a3-8cf4-4036-a67c-4e29d48ebf7c · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation Deep Reinforce- ment Learning for Robotic Manipulation—The State of the Art,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation fd00e4a1-9261-401e-8ac9-afcbafc6761a · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation Trust Region Policy Optimization,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a2b5a92c-33fd-4cda-b83c-189239dfae9c · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation Proximal Policy Optimization Algorithms
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53dc0fd6-3ec5-445f-83a7-88984a35e637 · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation High-Dimensional Continuous Control Using Generalized Advantage Estimation
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6c50502-6585-4964-a7e2-d8d07e4a37b2 · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation Adam: A Method for Stochastic Optimization
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c80dd969-9d66-46f8-89a2-01d0cddfaf3c · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation FIXAR: A Fixed-Point Deep Re- inforcement Learning Platform with Quantization-Aware Training and Adaptive Parallelism,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 370fcbf2-6990-4c08-9678-f851544824b3 · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation QuaRL: Quantization for Fast and Environmentally Sustainable Reinforcement Learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7fb7a99-274b-4858-96e1-152edea360ea · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation EnvPool: A Highly Parallel Reinforcement Learning En- vironment Execution Engine,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 55aa3610-82b8-4d41-a3b8-acc07e81cc40 · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation Accelerating Proximal Policy Optimization on CPU-FPGA Heterogeneous Platforms,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation cbdc85e0-98e7-4158-ada6-9c843614d871 · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation GPU-Accelerated Robotic Simulation for Distributed Reinforcement Learning,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f716cda3-6545-4b38-b74f-da7f535c7f6d · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation Accelerating Reinforcement Learning through GPU Atari Emulation,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b38a05ef-509b-443f-b3a2-95588fa65843 · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0b0eaba1-278a-4cbb-b6aa-c24862367dbc · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation Note on a Method for Calculating Corrected Sums of Squares and Products,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 96ab6666-9ee0-4b7c-9951-0a6590e6b8af · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation (2018) Understanding Normalization of Advantage Function in PPO [Online]
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 4400bab7-7e5d-4998-bc50-f3338952fc91 · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 179541f6-90d1-4dbe-aa06-777520f4d671 · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 3251f2e3-1dcd-48ee-b6c9-159998dde902 · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation Revisiting Deep Learning Paral- lelism: Fine-Grained Inference Engine Utilizing Online Arithmetic,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 701ae6d7-a28a-4dc8-81ed-d53969493647 · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation Enabling Mixed-Timing NoCs for FPGAs: Reconfigurable Synthesizable Synchronization FIFOs,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 51f16659-a3a6-4dad-a98a-4e167a2b5d37 · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation Reconfigurable Synthesizable Syn- chronization FIFOs,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 8aaf4c8c-1f35-48f8-a33e-c841b30cbe3d · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation Safe Overclocking of Tightly Coupled CGRAs and Processor Arrays using Razor,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 2cf4b782-1809-47fa-846e-672adc3adb67 · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation High-Throughput Synthesizable Synchronization FI- FOs for Mixed-Timing NoCs,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 67226e91-8ae9-4118-8765-1702d5debd09 · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation Interleaved Architectures for High-Throughput Synthesizable Synchronization FIFOs,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6703a6f9-c379-4be3-9295-5f438e3b8517 · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation Atalanta: A Bit is Worth a “Thousand
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a885b039-9f97-453a-8efe-d055e03434b6 · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation Synthesizable Synchronization FIFOs Utilizing the Asynchronous Pulse-Based Handshake Protocol,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 31363f23-6940-4a1b-9181-9ca4b59fa406 · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation Boveda: Building an On-Chip Deep Learning Memory Hierarchy Brick by Brick,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7fc6d1aa-d000-4ca5-8945-078d9b7bdb80 · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation Mokey: Enabling Narrow Fixed-Point Inference for Out-of-the-Box Floating-Point Transformer Models,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 31dadffc-fdde-4cf5-8a32-e4db56796dcf · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation Available: https://proceedings.mlsys.org/paper files/paper/ 2021/file/12a304a31e42dfefa21c82431e849124-Paper.pdf 9
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 54985f18-abc4-4ad2-8b2b-272fb6748f7b · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation Unresolved cited work
Reference 485
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1e9d03f2-3603-451b-93ba-60e62f257be1 · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation Available: https://www.jstor.org/stable/1266577
Reference 1962
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f6ad1393-60cf-4f1d-bc29-dc0df14e6709 · outbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.