Pith. sign in

Paper Citation Record · LEDGER

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines

As of 8 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2510.27329.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2510.27329 v2

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T07:22:05.850923Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

31 of 31 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 60ee38ad-b93b-4efd-8f98-9f12364ac080 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:01.494820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:01.494820Z digest=sha256:8a9811ac721e6a1065e938c410e1c6aeafe67c9c31232985918c0d5c6034c37e

Observation 30bda6a6-0968-4f15-8700-8ad14a231a66 · outbound

This paper cites write newline.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:01.586456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:01.586456Z digest=sha256:0b3a1559569c37cb467710446d03303bee19624b8f4afbaf47233f3a6136c71b

Observation f31762de-a072-4e3b-b41b-3d34d19637e7 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:01.703857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:01.703857Z digest=sha256:40039ba43f1f8ed917fd0b281bd18f33225b157bac869bae1d2b94bdb17af52e

Observation aadc17ee-3eb9-4797-96b6-2d16f121f947 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:01.849683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:01.849683Z digest=sha256:162d04d815af5eea91dd96755f904a5bcc32b5db3e182e0b668932c51df325a0

Observation aef4d452-0092-4735-bcf7-9098b23185e2 · outbound

This paper cites T.; Klassen, T.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines T.; Klassen, T

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:02.084791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:02.084791Z digest=sha256:fb4a3bc3b7b5da43515941940f6fb510b8daf9274232898c8ce4e2fc87eae37e

Observation 8121e214-7a63-4f95-8f5d-b2dc1bddcee6 · outbound

This paper cites C.; Di Nunzio, L.; Fazzolari, R.; Giardino, D.; Re, M.; and Span \`o , S.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines C.; Di Nunzio, L.; Fazzolari, R.; Giardino, D.; Re, M.; and Span \`o , S

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:02.205642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:02.205642Z digest=sha256:f49c2242f2d47823f2e77d1b958f10b9f1f469327bc4158e4013cd67099f789c

Observation 67150340-0d77-4d8d-87a6-f3759208b471 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:02.415694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:02.415694Z digest=sha256:2a7db6831163fee2fc03bf8757c24de9a4850b83a85c73cd50f72364cbd15f81

Observation 8f39e8e8-14ef-4155-9bad-0c927ae9bbef · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:02.495764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:02.495764Z digest=sha256:5cf983263d1fb4e58fa68758ff1d7fb545ea1204d67bbf998cba50f02bc0d5a6

Observation aa8173e9-45f2-43ba-92cc-575d2b078851 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:02.660494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:02.660494Z digest=sha256:8ac998aa3b2ee145d298d16ec405885ad8cf421624ab6161fca7c32850afd0d4

Observation ee9f5f3c-eb77-44be-9239-1f44cd14990a · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:02.818652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:02.818652Z digest=sha256:4f5c3cd32d5ae56d72966ad34d4f8b32fb59c9c8b0dc218cd742731daf68a3a1

Observation fc0c8680-36dd-4408-b954-6eef134a1d82 · outbound

This paper cites T.; Klassen, T.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines T.; Klassen, T

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:02.916971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:02.916971Z digest=sha256:460c1eaa2a67c36b3c17375ea3604bb16bce6f6e110df4c966c7fb7e9c91ddc2

Observation b6a00321-4b07-469f-ad4d-d5a62f03555c · outbound

This paper cites T.; and McIlraith, S.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines T.; and McIlraith, S

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:03.044745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:03.044745Z digest=sha256:039ded7900ca65b01364b7157d23efbcc8584bbb46f138ccaecc82411bbe6659

Observation fc329aa7-9d90-40ef-b5ce-ddb038df35f0 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:03.193495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:03.193495Z digest=sha256:15fb7b08f6a451c4cab3b83ef78061c53e9bc2fd7b398094c1ca6a5135fe33e5

Observation 083b3ea7-e1c8-4d7d-aae4-d34ee71ce940 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:03.345637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:03.345637Z digest=sha256:f888eeb10a11f84eb3e3c08e8bc09d18eebfedf55c6299ae6e17ea14e2837d72

Observation 84981524-30f8-4d08-9227-ec83e1aa7956 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:03.512091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:03.512091Z digest=sha256:c25a49b7fcba31d933f0ebc24b6215a1cab4298ff027e9a49dccee328809bf84

Observation b64f9050-e377-4f55-bedd-d05e37f11c32 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:03.617089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:03.617089Z digest=sha256:1ee1d97499a4f98c3e3c55986087b17197d5fce11eea0ba685a92bd715bc8b5f

Observation 92dd31f8-f461-4f07-a689-c3250bebce7d · outbound

This paper cites u ller, M.; Sch \.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines u ller, M.; Sch \

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:03.688661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:03.688661Z digest=sha256:61ea88a51203dfce4f330ed60adeead967458ec712cdaba862ab529e54883410

Observation a47a56e2-2645-4730-b5e7-c336bbeb6242 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:03.757028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:03.757028Z digest=sha256:25cb385d8052c8fa1af17120ef3d7725b5241031205b989aeeba557061f65961

Observation d58323c4-33dc-4f76-b6ab-18a439dda81c · outbound

This paper cites Continuous control with deep reinforcement learning.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Continuous control with deep reinforcement learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:03.852431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:03.852431Z digest=sha256:7e5f827a0986e7814013583488bb7a38143ec599f86f292861127ff92c45c952

Observation 3bf465c4-6af0-4a52-8f38-abea83477fb3 · outbound

This paper cites Modular Lifelong Reinforcement Learning via Neural Composition.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Modular Lifelong Reinforcement Learning via Neural Composition

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:04.013797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:04.013797Z digest=sha256:fde4c438569c0467ac7f7846f54a8fe817413af094d2a26dd8e6012dada81a6e

Observation 75f463ab-039e-419b-9ae6-1fc17d476b98 · outbound

This paper cites A.; Veness, J.; Bellemare, M.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines A.; Veness, J.; Bellemare, M

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:04.220326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:04.220326Z digest=sha256:4f57f8810b7ad26552d556be48a32a3708b9845d5a15d48a867dfeb59963272b

Observation d62438cc-deb9-4392-9457-6de9dd91da95 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:04.368793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:04.368793Z digest=sha256:0b5e1883eeb501fd2fac8e3189050a5a12bf16b247560867c953a8f72f07e4e0

Observation 088352b7-6e21-4fac-bd70-e1b01fc23d4c · outbound

This paper cites Y.; Harada, D.; and Russell, S.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Y.; Harada, D.; and Russell, S

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:04.512375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:04.512375Z digest=sha256:2104f111f4900edc13c8a94d4e3c709ba9777b5074be78fef1ce18f8a8a3cddd

Observation 0ca057a4-3a4d-40ce-b829-8733ff2092aa · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:04.654376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:04.654376Z digest=sha256:5cca3b3caacc9f666d08e5650d0038c3e8b4aa7551764f7942e79f7a6f4502df

Observation b9a5b734-d376-496d-856e-0e7f65fa0f4a · outbound

This paper cites N.; Wright, R.; Velasquez, A.; and Sinapov, J.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines N.; Wright, R.; Velasquez, A.; and Sinapov, J

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:04.859930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:04.859930Z digest=sha256:aa188005a65b8d13ce0a0ee5758b1fa78fc51a15e7fcba23f1267b4b8a3a3d2f

Observation 1c8d9118-9459-46d6-ac0f-2f204ff6216a · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:05.019402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:05.019402Z digest=sha256:bd64db752b70ee2754279aef9043d280ba48d2ecbba212a46ff9076ec4ae986f

Observation 4bfa9be9-6269-4a47-b9d2-1010079e5e65 · outbound

This paper cites B.; Talbert, D.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines B.; Talbert, D

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:05.203791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:05.203791Z digest=sha256:34b09f3881575d61f6196fc8ab00d688ef75d55f5cea8e03719c55b9174aa5c9

Observation 7f36fb87-e2bb-49c3-933b-744a217b44ae · outbound

This paper cites S.; and Barto, A.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines S.; and Barto, A

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:05.367931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:05.367931Z digest=sha256:5ec2a40ea350f08e50c18239a1304e4c9aeab41f9d45b3075dd9f7e90ec9a22d

Observation 2a49a3c4-083a-4f66-a534-5d853fbb78c2 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:05.518457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:05.518457Z digest=sha256:a10ba306627f5658dbb3d74c0fe5475e7a9c9a1f55ff19d701b72b74053582af

Observation 08e497c5-5e88-40c2-af1a-71c957a70d89 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:05.695724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:05.695724Z digest=sha256:c3535cb9ad7943e142abf10520beb92d58630d05d05dbdb33fc44bf101ad9683

Observation d414bbdf-1bf4-4d45-86d6-aaa5d4020e3f · outbound

This paper cites J.; and Dayan, P.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines J.; and Dayan, P

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:05.850923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:05.850923Z digest=sha256:294ea2d164364f9e2c2664fbeee1455f670c59c3c2a0fb7a59fdb297bcf725d6

Pith citing papers

No inbound Pith citation observations are available.