Pith. sign in

Paper Citation Record · LEDGER

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines

As of 9 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2510.27329.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2510.27329 v2

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T07:22:05.850923Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

31 of 31 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 60ee38ad-b93b-4efd-8f98-9f12364ac080 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:01.494820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:01.494820Z digest=sha256:0567887edb77099afca1d50a8489432eb3a52c1b93101b1c646955e80463778c

Observation 30bda6a6-0968-4f15-8700-8ad14a231a66 · outbound

This paper cites write newline.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:01.586456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:01.586456Z digest=sha256:2fa9b89668112774072bdac413c568b675b2e208406d4425edac907ea36ce2e5

Observation f31762de-a072-4e3b-b41b-3d34d19637e7 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:01.703857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:01.703857Z digest=sha256:d0bb254606e745c042056e39607d4ef3a3e88982bea2bc99b98b993edd59fa3d

Observation aadc17ee-3eb9-4797-96b6-2d16f121f947 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:01.849683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:01.849683Z digest=sha256:22d71f48bfbe5097dacaba92152543c29a10b6f95a5167d9016940bc85f729cf

Observation aef4d452-0092-4735-bcf7-9098b23185e2 · outbound

This paper cites T.; Klassen, T.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines T.; Klassen, T

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:02.084791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:02.084791Z digest=sha256:78ed14f83c92e8cb6ff86e03ec41d7627ab4a7f958a5e4ea1844e3676237b6a0

Observation 8121e214-7a63-4f95-8f5d-b2dc1bddcee6 · outbound

This paper cites C.; Di Nunzio, L.; Fazzolari, R.; Giardino, D.; Re, M.; and Span \`o , S.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines C.; Di Nunzio, L.; Fazzolari, R.; Giardino, D.; Re, M.; and Span \`o , S

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:02.205642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:02.205642Z digest=sha256:72d05f333cfc8de53692fc30a0f79d140ac25ad2d54ff11c8ee6de63d18e24eb

Observation 67150340-0d77-4d8d-87a6-f3759208b471 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:02.415694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:02.415694Z digest=sha256:760ec6286d8aad6ac830095e1c6f105989c769c395a4b237dae48052c1bab7c5

Observation 8f39e8e8-14ef-4155-9bad-0c927ae9bbef · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:02.495764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:02.495764Z digest=sha256:040c97d4c7f8459294770ee50e1faf0ed74e02430835099be3e7740159eacdf5

Observation aa8173e9-45f2-43ba-92cc-575d2b078851 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:02.660494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:02.660494Z digest=sha256:f593ef45ee0a31b6722143cd08bd5dc56ac62b16c4ffd4990cb7bbf670e5f640

Observation ee9f5f3c-eb77-44be-9239-1f44cd14990a · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:02.818652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:02.818652Z digest=sha256:374c9a847ab2e38db7d7ee679194ff012284c24b506de0b318a61b36b56f71fe

Observation fc0c8680-36dd-4408-b954-6eef134a1d82 · outbound

This paper cites T.; Klassen, T.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines T.; Klassen, T

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:02.916971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:02.916971Z digest=sha256:44ee8279734c7067591c2458a2231b048d579807bbc11a01f00f33f201660b93

Observation b6a00321-4b07-469f-ad4d-d5a62f03555c · outbound

This paper cites T.; and McIlraith, S.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines T.; and McIlraith, S

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:03.044745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:03.044745Z digest=sha256:c1386800d5d8be98176ee07e0a5956f0e9c8c4b34df47eec504eaabfc4035623

Observation fc329aa7-9d90-40ef-b5ce-ddb038df35f0 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:03.193495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:03.193495Z digest=sha256:12f4c4d4f52c561858f8b3c51fcb6b294a6789d134aeb344884180a464ac9116

Observation 083b3ea7-e1c8-4d7d-aae4-d34ee71ce940 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:03.345637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:03.345637Z digest=sha256:6f74bd3062f943150f967ff678841b005999a879a43c9bd5792895b843662849

Observation 84981524-30f8-4d08-9227-ec83e1aa7956 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:03.512091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:03.512091Z digest=sha256:c0e6fb3fd4f2a1caade54250ffd97d19589874a4a2a75a84dd166dd8251a0a9e

Observation b64f9050-e377-4f55-bedd-d05e37f11c32 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:03.617089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:03.617089Z digest=sha256:a1704dec87e497c5d717caeea608294c847922e7b654305a4b56110f01b08e69

Observation 92dd31f8-f461-4f07-a689-c3250bebce7d · outbound

This paper cites u ller, M.; Sch \.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines u ller, M.; Sch \

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:03.688661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:03.688661Z digest=sha256:764342b9959abe629113f95f18fcbaaedadefb3d08e3dd60822ddc18793eded7

Observation a47a56e2-2645-4730-b5e7-c336bbeb6242 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:03.757028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:03.757028Z digest=sha256:c2ca54fcfc1b38f39978df90a855727bfe621551878336ee555905b00c713d53

Observation d58323c4-33dc-4f76-b6ab-18a439dda81c · outbound

This paper cites Continuous control with deep reinforcement learning.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Continuous control with deep reinforcement learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:03.852431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:03.852431Z digest=sha256:8961933dee7ac956a7b841a4d8491a88f70fce3da913c7308812bad1a1073346

Observation 3bf465c4-6af0-4a52-8f38-abea83477fb3 · outbound

This paper cites Modular Lifelong Reinforcement Learning via Neural Composition.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Modular Lifelong Reinforcement Learning via Neural Composition

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:04.013797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:04.013797Z digest=sha256:9ac98e93010b044b54a2ca6a018127121b27f649fd34afb9b6c9bfa03bdea1e9

Observation 75f463ab-039e-419b-9ae6-1fc17d476b98 · outbound

This paper cites A.; Veness, J.; Bellemare, M.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines A.; Veness, J.; Bellemare, M

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:04.220326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:04.220326Z digest=sha256:4542e205694c29abbc46d34b50067db2a4415c71341f413d2ad63d6614a104d0

Observation d62438cc-deb9-4392-9457-6de9dd91da95 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:04.368793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:04.368793Z digest=sha256:1de302e84b1c7f2d29741b72b62aeb01778e8c36840217479964e732f8c0d2cc

Observation 088352b7-6e21-4fac-bd70-e1b01fc23d4c · outbound

This paper cites Y.; Harada, D.; and Russell, S.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Y.; Harada, D.; and Russell, S

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:04.512375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:04.512375Z digest=sha256:2275acc40eeff1e17a2278d56681e1d914a6ba3bacdb00840f4395a8cc355ee4

Observation 0ca057a4-3a4d-40ce-b829-8733ff2092aa · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:04.654376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:04.654376Z digest=sha256:a863fed9780c12a3a9536463502e8fb47832a4650bb2bd801209bb73d76d5bc8

Observation b9a5b734-d376-496d-856e-0e7f65fa0f4a · outbound

This paper cites N.; Wright, R.; Velasquez, A.; and Sinapov, J.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines N.; Wright, R.; Velasquez, A.; and Sinapov, J

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:04.859930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:04.859930Z digest=sha256:20c59ae22c376bf57eb6b778ba1fd0fccce3a5dd2ceda0b79bf49c857ba18301

Observation 1c8d9118-9459-46d6-ac0f-2f204ff6216a · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:05.019402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:05.019402Z digest=sha256:67858d77eb47c1401619b3877229a6a8d7b52b3f2bfcab17ce1c75b3e8eddaac

Observation 4bfa9be9-6269-4a47-b9d2-1010079e5e65 · outbound

This paper cites B.; Talbert, D.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines B.; Talbert, D

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:05.203791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:05.203791Z digest=sha256:4964482300bd03c34165598fd182e5beab9db3e7def1ee5d0a891882ff68a38f

Observation 7f36fb87-e2bb-49c3-933b-744a217b44ae · outbound

This paper cites S.; and Barto, A.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines S.; and Barto, A

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:05.367931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:05.367931Z digest=sha256:20b770a5636465ca724a196d0ab35f04c1d640a50d9125c7fb918c3c908d37ec

Observation 2a49a3c4-083a-4f66-a534-5d853fbb78c2 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:05.518457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:05.518457Z digest=sha256:c70879d5768f9ca1d105387dc3e46e646c90d245bf45b6e7bcbf6b4639b014f7

Observation 08e497c5-5e88-40c2-af1a-71c957a70d89 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:05.695724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:05.695724Z digest=sha256:9d08c88147dc068e548c7f4883cb7f4a8dff8b1b65644773d7880d7e9b80b499

Observation d414bbdf-1bf4-4d45-86d6-aaa5d4020e3f · outbound

This paper cites J.; and Dayan, P.

Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines J.; and Dayan, P

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:05.850923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:05.850923Z digest=sha256:387201bfea04ea356e68741fee5ee98f9bc520b1a298aabae2575698e1a5bb0a

Pith citing papers

No inbound Pith citation observations are available.