Pith. sign in

Paper Citation Record · LEDGER

Success in Humanoid Reinforcement Learning under Partial Observation

As of 21 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2507.18883.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.18883 v1

Coverage vector

measured 21 of 21 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T18:11:10.930242Z

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

21 of 21 outbound references displayed

  • verified exact0
  • verified fuzzy6
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0ffdf3b3-c372-4e72-9bca-4fe0151eaa71 · outbound

This paper cites Cipriano, P.

Success in Humanoid Reinforcement Learning under Partial Observation Cipriano, P

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:11:11.421259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.819046Z digest=sha256:d3f578b1bd8135e01b9a3e1e2542b8d7db0d52805288cf1bce3ea3e35144e699

Observation 37690468-5bf9-4288-8a48-9be8925a60c7 · outbound

This paper cites Estevez, J.

Success in Humanoid Reinforcement Learning under Partial Observation Estevez, J

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:11:11.401806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.826042Z digest=sha256:42ee4672c2bfa4eefbfd563f7ddb16a33ee163d9a3df6fd96d7614b31bc9e4c1

Observation c1dfbadf-792d-4715-8fb1-fee5982e6045 · outbound

This paper cites an unresolved cited work.

Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:11:11.382550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.831258Z digest=sha256:a488894f9ff5168ff0827c2e26f66ba0888df7eddbb5b9e2d4ce51f70dc13fd9

Observation 3c434fbe-43c8-49a7-9291-b06c5a8fbd78 · outbound

This paper cites an unresolved cited work.

Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:11:11.363776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.836379Z digest=sha256:9636300ab18f9713f74c2557045dd3b4984b9057d008ac0c52ccd3b2b110d702

Observation 4067b1c3-8aea-4b2d-878f-a8fca9b17162 · outbound

This paper cites Bellman, Dynamic Programming (Princeton University Press, Princeton, NJ) (1957), intro- duces the formalism of Markov decision processes (MDPs) and the principle of optimality.

Success in Humanoid Reinforcement Learning under Partial Observation Bellman, Dynamic Programming (Princeton University Press, Princeton, NJ) (1957), intro- duces the formalism of Markov decision processes (MDPs) and the principle of optimality

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:11:11.346834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.842521Z digest=sha256:90be25d1c2f6d28a5ac28f554cb1324ef102d0b5a7bad0364ca2cc083133cfff

Observation 2bb8cd5e-fbac-41a8-9d65-c1e312c30cd5 · outbound

This paper cites an unresolved cited work.

Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:11:11.327533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.847978Z digest=sha256:8eef39d1be4a152068f776b377998d7b41676b40077c46a73ef7805275f37406

Observation 4d81ecb9-85bd-4e89-8303-950e820349c8 · outbound

This paper cites an unresolved cited work.

Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:11:11.309457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.853493Z digest=sha256:fafb7d4e6909727e26d935643fb18e2dce2cfb475ad2179178c6dc7aba092fac

Observation f7e5f1c5-5923-487e-8baa-e09ab719a048 · outbound

This paper cites an unresolved cited work.

Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:11:11.292533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.859193Z digest=sha256:ba2d13db1019b95ef3270d81bc46e9f5d0025bf9701d582af6d850de8722d23d

Observation d694c060-bbb5-4f71-964c-92cf6bb4486f · outbound

This paper cites an unresolved cited work.

Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:11:11.275683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.864080Z digest=sha256:00d7a5be388c0bbb1b4b7c8bc739968ff25be4da69e48d501301d262369189ee

Observation 33c6fef4-d456-4592-92b7-21f8aa5f728b · outbound

This paper cites an unresolved cited work.

Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:11:11.254858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.870049Z digest=sha256:9abe36b537dc0d29dd9dd942f0588ac73ef22ba2d29151828acafee9c2949ff9

Observation 0f36d1c8-13eb-48aa-8ea0-da1530f32a02 · outbound

This paper cites Hochreiter, J.

Success in Humanoid Reinforcement Learning under Partial Observation Hochreiter, J

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:11:11.237605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.875545Z digest=sha256:cd81fa12a7ae8ce7e6cc4f9472752066f492138227f4da43b8cbc6355c2df9fb

Observation f16242e5-0751-4c0b-b493-40561be136db · outbound

This paper cites an unresolved cited work.

Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:11:11.217435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.880471Z digest=sha256:9c75c8d6fadf8d96105cd9badbc90edc9fed20fa694157d306c207572f7e6b62

Observation d5a7c946-349e-41ad-b066-855291a92ca0 · outbound

This paper cites an unresolved cited work.

Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:11:11.194171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.886132Z digest=sha256:d812eb55c966b21cc8fed53fac4942f8ec9c5bd3afbc7d355dbdb48d67a32bad

Observation f7f490aa-2a40-420d-b0af-34528ce028f8 · outbound

This paper cites an unresolved cited work.

Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:11:11.172333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.891644Z digest=sha256:47a6dd1018a7986c8c363737d9e2715186adaf49fdae73c72956f07d3210767a

Observation 90d0627e-a914-4568-a179-8993aa403def · outbound

This paper cites Arcieri, et al., POMDP Inference and Robust Solution via Deep Reinforcement Learning: An Application to Railway Optimal Maintenance.

Success in Humanoid Reinforcement Learning under Partial Observation Arcieri, et al., POMDP Inference and Robust Solution via Deep Reinforcement Learning: An Application to Railway Optimal Maintenance

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:11:11.153183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.896574Z digest=sha256:314830e1b942a39b164155edf110d2b778d033a004b1866dfac3a33350d6b7a8

Observation 3ff6713b-873f-4e3a-abb9-738e86910c1c · outbound

This paper cites Lemmel, R.

Success in Humanoid Reinforcement Learning under Partial Observation Lemmel, R

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:11:11.133464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.901994Z digest=sha256:ba377d024d86e8252e4c3ef3ff96e237d67211acebdfcdb720094820509809fd

Observation a1ab6c8b-4f4a-4989-8c4c-9930e88ca20d · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

Success in Humanoid Reinforcement Learning under Partial Observation Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T18:11:10.906816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:11:10.906816Z digest=sha256:68f31bdfe210ff1fa9c2f7db8e7f2dfd26fdfe20be4022e6cb02bb9769ceedef

Observation 3f30e640-9f71-4dab-bb45-3cab7c529e1e · outbound

This paper cites an unresolved cited work.

Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:11:11.113614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.913179Z digest=sha256:d083c4ca968109a20f0a3f19fddfe984081824254a878952c218adad51860734

Observation c8306416-5b0e-4e2e-83ff-b8828f7659cf · outbound

This paper cites Decision Mamba: Reinforcement Learning via Sequence Modeling with Selective State Spaces.

Success in Humanoid Reinforcement Learning under Partial Observation Decision Mamba: Reinforcement Learning via Sequence Modeling with Selective State Spaces

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T18:11:10.919344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:11:10.919344Z digest=sha256:1494211e610a4c35480565022a58eb802b308b05c38cddb2f8d28e3ea593186f

Observation 118bf2cf-c6c2-43cb-936c-a8d46f7556d2 · outbound

This paper cites MAMBA: an Effective World Model Approach for Meta-Reinforcement Learning.

Success in Humanoid Reinforcement Learning under Partial Observation MAMBA: an Effective World Model Approach for Meta-Reinforcement Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T18:11:10.924642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:11:10.924642Z digest=sha256:21c7483933ac1273e6b7e3ff7dfa0a48bbacd1270336b5d83fff6b8f2fa55192

Observation 424d8515-7bf4-4613-b66e-754ae0e53ebd · outbound

This paper cites Towers, et al., Gymnasium (2023), doi:10.5281/zenodo.8127026,https://zenodo.org/ record/8127025.

Success in Humanoid Reinforcement Learning under Partial Observation Towers, et al., Gymnasium (2023), doi:10.5281/zenodo.8127026,https://zenodo.org/ record/8127025

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T18:11:10.930242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:11:10.930242Z digest=sha256:b59a58080eae4f1f6b477aef0c27326c16e016732585bf43c34cb242cdaab7a0

Pith citing papers

No inbound Pith citation observations are available.