Pith. sign in

Paper Citation Record · LEDGER

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines

As of 12 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2510.27329.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2510.27329 v2

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T07:22:05.850923Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

31 of 31 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 60ee38ad-b93b-4efd-8f98-9f12364ac080 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:01.494820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:01.494820Z digest=sha256:cbd22e852a544127773d0c2ecb0d23dd6b76b22f9f5d235b81fcbdd0d7fe1ced

Observation 30bda6a6-0968-4f15-8700-8ad14a231a66 · outbound

This paper cites write newline.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:01.586456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:01.586456Z digest=sha256:cef7f80c95398f2884b4631a7c5f123b3830d7794211661938985b94cb2c7699

Observation f31762de-a072-4e3b-b41b-3d34d19637e7 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:01.703857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:01.703857Z digest=sha256:1da65af20c58a61cfc385a589e39e76491d2d6b4b3c93509cdfe688c962506c3

Observation aadc17ee-3eb9-4797-96b6-2d16f121f947 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:01.849683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:01.849683Z digest=sha256:5f4c29e1cc5f7d4cdb3c9fed8b239a837b759c15b6b5b270790e47e407dae96f

Observation aef4d452-0092-4735-bcf7-9098b23185e2 · outbound

This paper cites T.; Klassen, T.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines T.; Klassen, T

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:02.084791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:02.084791Z digest=sha256:e333238a38a8e46eb024c25cb6e6baec2dc09e1b46b88a18b087be3eadbb5ccf

Observation 8121e214-7a63-4f95-8f5d-b2dc1bddcee6 · outbound

This paper cites C.; Di Nunzio, L.; Fazzolari, R.; Giardino, D.; Re, M.; and Span \`o , S.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines C.; Di Nunzio, L.; Fazzolari, R.; Giardino, D.; Re, M.; and Span \`o , S

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:02.205642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:02.205642Z digest=sha256:0627d4b553e4cacf1bc9de1ea1bd2a0ee6c23fbacd068766c3974ec205ec5d39

Observation 67150340-0d77-4d8d-87a6-f3759208b471 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:02.415694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:02.415694Z digest=sha256:a2fa0f87f8a6fa9280df3b3abfcbc43d550f38137f178343366523768ea5ac8f

Observation 8f39e8e8-14ef-4155-9bad-0c927ae9bbef · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:02.495764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:02.495764Z digest=sha256:cf2a9bb0df497a6bb293827c0c44469ab85d228fd875ef97e31ffb8509c28ca1

Observation aa8173e9-45f2-43ba-92cc-575d2b078851 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:02.660494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:02.660494Z digest=sha256:dcd7ae808f4599355703ee0602e52924efce60ec1742bbd3bc5eafd97fcce3d4

Observation ee9f5f3c-eb77-44be-9239-1f44cd14990a · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:02.818652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:02.818652Z digest=sha256:f7669750085a6ed9921b2a989e0c15ecf36b885848b933e8a1b8c832912c809f

Observation fc0c8680-36dd-4408-b954-6eef134a1d82 · outbound

This paper cites T.; Klassen, T.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines T.; Klassen, T

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:02.916971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:02.916971Z digest=sha256:480991653c675ddbbabab85847c881dbb2b00ce68ddc6253ad86eda0425adb75

Observation b6a00321-4b07-469f-ad4d-d5a62f03555c · outbound

This paper cites T.; and McIlraith, S.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines T.; and McIlraith, S

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:03.044745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:03.044745Z digest=sha256:058bb05db7fa5c7eec852616586405b22513444c122163a57ee01964c378d10c

Observation fc329aa7-9d90-40ef-b5ce-ddb038df35f0 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:03.193495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:03.193495Z digest=sha256:04517b38fff99baed1f828a3ecf8ae7ab385e589620118ea0ebfad076f9791b0

Observation 083b3ea7-e1c8-4d7d-aae4-d34ee71ce940 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:03.345637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:03.345637Z digest=sha256:ba725416c343cca6ceb6ca97eff95249dba093dbddea9cd748fdba8b6801964d

Observation 84981524-30f8-4d08-9227-ec83e1aa7956 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:03.512091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:03.512091Z digest=sha256:1474084aa8a2c2e9fb97bc980db7afb7e2593fe9218ecb91970a10bc1664de70

Observation b64f9050-e377-4f55-bedd-d05e37f11c32 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:03.617089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:03.617089Z digest=sha256:737340d56c9f8ba126fb961747a8f762cc20ae1d7edf3a0f8295321865a9ec60

Observation 92dd31f8-f461-4f07-a689-c3250bebce7d · outbound

This paper cites u ller, M.; Sch \.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines u ller, M.; Sch \

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:03.688661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:03.688661Z digest=sha256:c6a8776a76d0a6a3322eb2accca7d5179ca4b9a57b4bb19286d0a320b91b9d59

Observation a47a56e2-2645-4730-b5e7-c336bbeb6242 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:03.757028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:03.757028Z digest=sha256:c826c384360a22ccd3d7d84e50c6512a03d5b0b3dc4ea37d81f3a1efdfb2b96c

Observation d58323c4-33dc-4f76-b6ab-18a439dda81c · outbound

This paper cites Continuous control with deep reinforcement learning.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Continuous control with deep reinforcement learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:03.852431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:03.852431Z digest=sha256:c3250f636edc8bb07c2705209f4dfea71f977e65c921fb9a6f0ef4afae11380e

Observation 3bf465c4-6af0-4a52-8f38-abea83477fb3 · outbound

This paper cites Modular Lifelong Reinforcement Learning via Neural Composition.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Modular Lifelong Reinforcement Learning via Neural Composition

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:04.013797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:04.013797Z digest=sha256:c1ed5b86bf85f55121dd07e5df369b481892d2f5b124fe20e71f23c9a1a953da

Observation 75f463ab-039e-419b-9ae6-1fc17d476b98 · outbound

This paper cites A.; Veness, J.; Bellemare, M.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines A.; Veness, J.; Bellemare, M

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:04.220326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:04.220326Z digest=sha256:3c6fd1081767db12f46738e1d67ad0492dc8398dce9776f4fe1eb3806532366a

Observation d62438cc-deb9-4392-9457-6de9dd91da95 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:04.368793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:04.368793Z digest=sha256:d2bd1fab578986569c338bf825a961e207cc78f56b58c908e6d83a76bcb77a9d

Observation 088352b7-6e21-4fac-bd70-e1b01fc23d4c · outbound

This paper cites Y.; Harada, D.; and Russell, S.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Y.; Harada, D.; and Russell, S

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:04.512375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:04.512375Z digest=sha256:b1570ea1637d90eadf8f7c3b0b1eaa9b66becb54a6394e44b1c8846be657b5ce

Observation 0ca057a4-3a4d-40ce-b829-8733ff2092aa · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:04.654376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:04.654376Z digest=sha256:9cbf00c3dbc0fbc81d789b16d5478253d25a2cc6dc15bede87a89d89bc0d1b8c

Observation b9a5b734-d376-496d-856e-0e7f65fa0f4a · outbound

This paper cites N.; Wright, R.; Velasquez, A.; and Sinapov, J.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines N.; Wright, R.; Velasquez, A.; and Sinapov, J

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:04.859930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:04.859930Z digest=sha256:d4f20246c67d5ab8edad3b933d52541b8dd3a566af8cde1bb519cf4bdba86d8b

Observation 1c8d9118-9459-46d6-ac0f-2f204ff6216a · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:05.019402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:05.019402Z digest=sha256:7b7bd3051b0c2ce94f89c00d4a6d89d1bd4c6c304c4ea78880a551c042d83742

Observation 4bfa9be9-6269-4a47-b9d2-1010079e5e65 · outbound

This paper cites B.; Talbert, D.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines B.; Talbert, D

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:05.203791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:05.203791Z digest=sha256:bc16b06e150bf248ebd520d15a62e8385d9705102f9d47d0e9eb991006d8584c

Observation 7f36fb87-e2bb-49c3-933b-744a217b44ae · outbound

This paper cites S.; and Barto, A.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines S.; and Barto, A

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:05.367931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:05.367931Z digest=sha256:0c88a00f6511b6793f79f41f2942e7ad3dca225d20c764daefcade09de1e4cea

Observation 2a49a3c4-083a-4f66-a534-5d853fbb78c2 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:05.518457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:05.518457Z digest=sha256:046f5e7fd89a072639b2f8cf03c200f3a9f8844dcaa4a93a62b1a753c091de43

Observation 08e497c5-5e88-40c2-af1a-71c957a70d89 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:05.695724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:05.695724Z digest=sha256:54961ad53fe9d7a67e7de4b4281287c7b02317285956d8aaafe38e91edf36b6a

Observation d414bbdf-1bf4-4d45-86d6-aaa5d4020e3f · outbound

This paper cites J.; and Dayan, P.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines J.; and Dayan, P

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:05.850923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:05.850923Z digest=sha256:5d9ef573207a6dd417819a2b65198e4fa489d7b16ee8a9b16a11288b1d5f869d

Pith citing papers

No inbound Pith citation observations are available.