Pith. sign in

Paper Citation Record · LEDGER

Physics-Informed Reward Machines

As of 21 August 2026, this Paper Citation Record lists 16 of 16 outbound references and 0 inbound Pith citation observations for arXiv:2508.14093.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.14093 v1

Coverage vector

measured 16 of 16 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T17:37:34.393318Z

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

16 of 16 outbound references displayed

  • verified exact3
  • verified fuzzy5
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6f9a8e0a-14ec-47f4-a953-c1bdae50b6ff · outbound

This paper cites Exploration-Exploitation in Constrained MDPs.

Physics-Informed Reward Machines Exploration-Exploitation in Constrained MDPs

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T17:37:34.343026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:37:34.343026Z digest=sha256:da6c66bff433f6ae23e1d78a2675b9748d7f2004376afaa28bd95a354cfdda44

Observation dca5cba7-30dd-4c3a-b991-858643da6feb · outbound

This paper cites Probably Approximately Correct MDP Learning and Control With Temporal Logic Constraints.

Physics-Informed Reward Machines Probably Approximately Correct MDP Learning and Control With Temporal Logic Constraints

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T17:37:34.348100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:37:34.348100Z digest=sha256:b835a1b8c8049919c0a4ec8db660969bb13c1a2fba012d0178b9f964b18368f7

Observation ea60b857-45d7-41bb-88d7-c9c7427b33cd · outbound

This paper cites Formal controller syn- thesis for continuous-space MDPs via model-free reinforcement learning.

Physics-Informed Reward Machines Formal controller syn- thesis for continuous-space MDPs via model-free reinforcement learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:37:34.624934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:37:34.374103Z digest=sha256:19247219494d73bf268a45477351e7b7baceb453682f9f8a555b8295d5010edd

Observation 925c596c-fe02-4af5-947a-88720bda8870 · outbound

This paper cites Bridging Physics-Informed Neural Networks with Reinforcement Learning: Hamilton-Jacobi-Bellman Proximal Policy Optimization (HJBPPO).

Physics-Informed Reward Machines Bridging Physics-Informed Neural Networks with Reinforcement Learning: Hamilton-Jacobi-Bellman Proximal Policy Optimization (HJBPPO)

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T17:37:34.383833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:37:34.383833Z digest=sha256:576281a008550b5d17fa69dab2ee056e2bd0e448e910662bd65dacaa9ba9ae2c

Observation c98ab608-bc42-47f5-88f1-0aca93ff3346 · outbound

This paper cites Active finite reward automaton inference and reinforcement learning using queries and counterexamples.

Physics-Informed Reward Machines Active finite reward automaton inference and reinforcement learning using queries and counterexamples

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:37:34.609457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:37:34.393318Z digest=sha256:31992171a3c5e1380bb3a749630ecca76c21d92d62b3ce731c3e730152210e1e

Observation b7f0c960-62a6-4061-9503-e85ec4073ba0 · outbound

This paper cites Reinforcement learning for temporal logic control synthesis with probabilistic satisfaction guarantees.

Physics-Informed Reward Machines Reinforcement learning for temporal logic control synthesis with probabilistic satisfaction guarantees

Reference 1996

Resolution
unresolved
no resolver link, observed 2026-08-15T17:37:34.358033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:37:34.358033Z digest=sha256:7ebb25ee72ea6a9f551716626c8aafe8d7bb7d4176f44f9c3b862c8a568d451c

Observation a0c990af-c85c-495b-b88c-814d09cd46a0 · outbound

This paper cites A Survey on Physics Informed Reinforcement Learning: Review and Open Problems.

Physics-Informed Reward Machines A Survey on Physics Informed Reinforcement Learning: Review and Open Problems

Reference 2000

Resolution
unresolved
no resolver link, observed 2026-08-15T17:37:34.327506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:37:34.327506Z digest=sha256:ae82112841fd7f01005dd5b4befe6856eb21d6829318f80e210400362fc1ccc4

Observation a8710dab-ae79-46a8-9368-693876445fb5 · outbound

This paper cites Efficient Reinforcement Learning in Probabilistic Reward Machines.

Physics-Informed Reward Machines Efficient Reinforcement Learning in Probabilistic Reward Machines

Reference 2006

Resolution
verified exact
local_arxiv, observed 2026-08-15T17:37:34.468216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:37:34.379130Z digest=sha256:c6d9fc214881ca4c709b9ab00dc72a7528b07cdd2e72bd12466244e897eabad3

Observation 6b973915-72d7-4e0c-8f14-415d818fd143 · outbound

This paper cites Continuous control with deep reinforcement learning.

Physics-Informed Reward Machines Continuous control with deep reinforcement learning

Reference 2013

Resolution
unresolved
no resolver link, observed 2026-08-15T17:37:34.368652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:37:34.368652Z digest=sha256:b70da237793cc99b4ef15d42b6fb1897bbc61a184621433b1f5ae4ade513a806

Observation 6e4eebab-56aa-44ed-a669-76e5a2cc58a2 · outbound

This paper cites PhysQ: A Physics Informed Reinforcement Learning Framework for Building Control.

Physics-Informed Reward Machines PhysQ: A Physics Informed Reinforcement Learning Framework for Building Control

Reference 2014

Resolution
verified exact
local_arxiv, observed 2026-08-15T17:37:34.505892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:37:34.353011Z digest=sha256:b04343151a6fc44fb8ffbe9db02f2d120795808ff43ff32b2afdd44678f84644

Observation 6148e142-cc46-4273-b43c-dc9cb513a1cc · outbound

This paper cites Learning Non-Markovian Reward Models in MDPs.

Physics-Informed Reward Machines Learning Non-Markovian Reward Models in MDPs

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-15T17:37:34.388699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:37:34.388699Z digest=sha256:76656279f71879954ad2b0d96a8a6e7ea8398a35ec10b9d4c46f14de0dbfcbaa

Observation 311df833-1bcb-496f-9ba4-5dba6fbfd00b · outbound

This paper cites Fast Online Exact Solutions for Deterministic MDPs with Sparse Rewards.

Physics-Informed Reward Machines Fast Online Exact Solutions for Deterministic MDPs with Sparse Rewards

Reference 2020

Resolution
verified exact
local_arxiv, observed 2026-08-15T17:37:34.559440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:37:34.337626Z digest=sha256:daddbf15893b6aaae98905d86b749302f9df1529fdb357bad0003b958c5b4070

Observation 1c09ba70-160b-46ba-b061-a852b014e6ca · outbound

This paper cites Physics informed intrinsic rewards in reinforcement learning.

Physics-Informed Reward Machines Physics informed intrinsic rewards in reinforcement learning

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:37:34.639791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:37:34.363563Z digest=sha256:20d27b6ec75b03808d5ac1c2a0f20d6a808775523e37ac32b8f736b96908d119

Observation f5548230-3b2e-495c-8ac5-77e8259fd1c8 · outbound

This paper cites Data-driven Construction of Finite Abstractions for Interconnected Systems: A Compositional Approach.

Physics-Informed Reward Machines Data-driven Construction of Finite Abstractions for Interconnected Systems: A Compositional Approach

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-15T17:37:34.317082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:37:34.317082Z digest=sha256:94789b648196d82b61c1c4946ebaddd067ea4e5cb7cd8e90b989102c0fe35b2c

Observation f435f4dc-4ff8-43c4-9c0b-60e13b11b16c · outbound

This paper cites Control synthesis from linear temporal logic specifications using model-free reinforcement learning.

Physics-Informed Reward Machines Control synthesis from linear temporal logic specifications using model-free reinforcement learning

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:37:34.663923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:37:34.333000Z digest=sha256:7973305fb9d75990160324d254eca0066e59ee135cba49ced3e74c9144fc1d7d

Observation ae0d34d5-82ce-44ed-bef3-b80459e3e912 · outbound

This paper cites Verification of Markov decision processes using learning algorithms.

Physics-Informed Reward Machines Verification of Markov decision processes using learning algorithms

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:37:34.679160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:37:34.322588Z digest=sha256:15965dc386cccd60e0cdbb9c3a16eff56a30c6e1c03e1b220eb11d922014544d

Pith citing papers

No inbound Pith citation observations are available.