Pith. sign in

Paper Citation Record · LEDGER

A View on Deep Reinforcement Learning in System Optimization

As of 15 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:1908.01275.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1908.01275 v3

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T15:21:50.582390Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

31 of 31 outbound references displayed

  • verified exact2
  • verified fuzzy11
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 00e02884-7e5a-4207-9575-025b7792fcc5 · outbound

This paper cites Placeto: Learning Generalizable Device Placement Algorithms for Distributed Machine Learning.

A View on Deep Reinforcement Learning in System Optimization Placeto: Learning Generalizable Device Placement Algorithms for Distributed Machine Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.428466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.428466Z digest=sha256:4d9ef635ff479642221756e20cc27f8d15f4ef911ac8b4c4f29cbe87a5abb26e

Observation 6e28d589-6bc8-46f8-ab5f-3a54c1d15dc2 · outbound

This paper cites Speech recog- nition with deep recurrent neural networks.

A View on Deep Reinforcement Learning in System Optimization Speech recog- nition with deep recurrent neural networks

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:21:51.127842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.455907Z digest=sha256:3a3acb2b087e7698697819a7400a8facf1ec3ccfb75bf17357e57faf84e93403

Observation 228359e0-e55e-4081-8ab8-5bb7f1b6de37 · outbound

This paper cites From Language to Programs: Bridging Reinforcement Learning and Maximum Marginal Likelihood.

A View on Deep Reinforcement Learning in System Optimization From Language to Programs: Bridging Reinforcement Learning and Maximum Marginal Likelihood

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.460868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.460868Z digest=sha256:c3c0d38554cf827d0ab74037cf472009d914aeff05c6f62c9dc5af698a13d2bf

Observation 3910cf45-f41f-41a6-bafc-b28f6d51451f · outbound

This paper cites Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor.

A View on Deep Reinforcement Learning in System Optimization Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.465908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.465908Z digest=sha256:7fb471c5cc34a03d40cf25b35921e64b3132d195133e6fcd8bcb51293b1fb48c

Observation 22f86071-80f4-4f34-acde-68180cfe3e0f · outbound

This paper cites Autophase: Compiler phase-ordering for hls with deep reinforcement learn- ing.

A View on Deep Reinforcement Learning in System Optimization Autophase: Compiler phase-ordering for hls with deep reinforcement learn- ing

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:21:51.111257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.471075Z digest=sha256:abfdc71f346f45f25c05b3edb2f46865ff4d9c7c520e1ee7d3c1ba08695ee64e

Observation 44a93076-ad39-4c38-afd4-a6c552b3e9c3 · outbound

This paper cites Learning- based and data-driven tcp design for memory-constrained iot.

A View on Deep Reinforcement Learning in System Optimization Learning- based and data-driven tcp design for memory-constrained iot

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:21:51.079314Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.486859Z digest=sha256:13727b73de76046a7f041b7de953cd5678deaef0d7925617bce9b538ec31dfc4

Observation 68442d96-dc64-45ac-8e18-5cbbf335b371 · outbound

This paper cites RLlib: Abstractions for Distributed Reinforcement Learning.

A View on Deep Reinforcement Learning in System Optimization RLlib: Abstractions for Distributed Reinforcement Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.497388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.497388Z digest=sha256:d7e901452ae6c8f72859ca4f808b7d079d8f18c96d659b73635b4342b81bfd0f

Observation f883fe50-ff45-4c88-81be-b640f2782c62 · outbound

This paper cites Neural Packet Classification.

A View on Deep Reinforcement Learning in System Optimization Neural Packet Classification

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-14T15:21:50.793436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.502401Z digest=sha256:0b2eafc358e0c2f4abb7c568930bc69897ea9ddbd8f33ca88341f4d104244f37

Observation 46656035-c6da-4682-83ac-1e2136a7d39c · outbound

This paper cites Neo: A Learned Query Optimizer.

A View on Deep Reinforcement Learning in System Optimization Neo: A Learned Query Optimizer

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.518247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.518247Z digest=sha256:2e1958f71728d89594e238262f46b365819edccad24e55fcb5f9364fd8dd659b

Observation 65ba32d6-babe-45d0-b023-20fd4d9ac686 · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

A View on Deep Reinforcement Learning in System Optimization Playing Atari with Deep Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.523130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.523130Z digest=sha256:740f27055daaf8fa7b2f0df3de9be67eed1d7d07185164a1c38fe628fbf1a14d

Observation 3eade8a4-45eb-46f3-b4b5-7b286ecf3f1b · outbound

This paper cites P., Mirza, M., Graves, A., Lillicrap, T., Harley, T., Silver, D., and Kavukcuoglu, K.

A View on Deep Reinforcement Learning in System Optimization P., Mirza, M., Graves, A., Lillicrap, T., Harley, T., Silver, D., and Kavukcuoglu, K

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:21:51.029564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.528120Z digest=sha256:e779d0bcdef094485f39087e890211b5bf108e9ba730c4cd6adc2d3b64bbde20

Observation 6549131c-32b3-4b7a-8ed2-e46873144e3f · outbound

This paper cites A Stochastic Approximation Approach for Foresighted Task Scheduling in Cloud Computing.

A View on Deep Reinforcement Learning in System Optimization A Stochastic Approximation Approach for Foresighted Task Scheduling in Cloud Computing

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-14T15:21:50.718986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.533011Z digest=sha256:7c497698efd613572b490aa483a4c296dc51d613813efab3ac47f42dda180432

Observation 8e44b00e-10fd-41b8-994a-7e5e2f08af8e · outbound

This paper cites Learning State Representations for Query Optimization with Deep Reinforcement Learning.

A View on Deep Reinforcement Learning in System Optimization Learning State Representations for Query Optimization with Deep Reinforcement Learning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.537780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.537780Z digest=sha256:ff450c2f9c3fa44e83e13632770dba92e1f457f159980a629198e99766082bb3

Observation 0b3c0038-9114-44d9-9009-fc73b5eb699e · outbound

This paper cites Reinforced Genetic Algorithm Learning for Optimizing Computation Graphs.

A View on Deep Reinforcement Learning in System Optimization Reinforced Genetic Algorithm Learning for Optimizing Computation Graphs

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.542568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.542568Z digest=sha256:a93f672c66deb37d3eda6704a58549b05d013a1d60f581d0d7a3acfd35cb8979

Observation 30458130-2d91-47b1-9a74-eeeeb3009117 · outbound

This paper cites Semantic locality and context-based prefetching using reinforce- ment learning.

A View on Deep Reinforcement Learning in System Optimization Semantic locality and context-based prefetching using reinforce- ment learning

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:21:51.010966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.547836Z digest=sha256:4e929d778f5ee1e88f786dd5e22ed32bd8ea95b14b873b2faec215f1bc2aec96

Observation 257fb1fb-ca77-40c9-9dbc-e75de8397943 · outbound

This paper cites P., Obraczka, K., Burleigh, S., and Hirata, C.

A View on Deep Reinforcement Learning in System Optimization P., Obraczka, K., Burleigh, S., and Hirata, C

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:21:50.993127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.563366Z digest=sha256:66c6b3b4c265d9a49bee3b82e6a4bf6279299d6a7016159342d200b307fac11c

Observation ebb2c584-f104-445b-b136-84ff63d22609 · outbound

This paper cites Machine learning in compiler optimization.

A View on Deep Reinforcement Learning in System Optimization Machine learning in compiler optimization

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:21:50.976831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.568188Z digest=sha256:6c63c96e70d7fc14948fbfcee6c3dff4d5a69100a2fbfc4b55ba772554e6207c

Observation 8daf851d-f838-4a05-b827-3c4f25a8cabe · outbound

This paper cites Seq2SQL: Generating Structured Queries from Natural Language using Reinforcement Learning.

A View on Deep Reinforcement Learning in System Optimization Seq2SQL: Generating Structured Queries from Natural Language using Reinforcement Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.577434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.577434Z digest=sha256:5b073318b050a813d52979f4a00f47de09a4cf69824917f34874e9be1c5d1e2f

Observation 6ef79119-f153-434c-99ea-3f3fb5711f02 · outbound

This paper cites A reinforcement learning approach to automatic error recovery.

A View on Deep Reinforcement Learning in System Optimization A reinforcement learning approach to automatic error recovery

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:21:50.943948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.582390Z digest=sha256:01666aa78c9ebe0044fb151f4cd1f23f8dd6fc4b7341acdcbf1e2dc1cb4eea2b

Observation 353e3208-c3f6-4513-a045-991b4a3c2e96 · outbound

This paper cites Proximal Policy Optimization Algorithms.

A View on Deep Reinforcement Learning in System Optimization Proximal Policy Optimization Algorithms

Reference 1999

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.558300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.558300Z digest=sha256:989ca1017cc027685ee18c5b43dcd8bfd191e2e622308ab8d804437eefeb482e

Observation ce5c9204-5520-4d21-ab6d-857cba63bb6b · outbound

This paper cites Energy-efficient virtual machines consolidation in cloud data centers us- ing reinforcement learning.

A View on Deep Reinforcement Learning in System Optimization Energy-efficient virtual machines consolidation in cloud data centers us- ing reinforcement learning

Reference 2000

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:21:51.143986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.445112Z digest=sha256:432d9216391f74331065743289e620247588e89c4ca04894479eacd253b9f2a4

Observation dd34cad1-ffdd-4823-b017-c9bc726f552f · outbound

This paper cites Learning to Optimize Join Queries With Deep Reinforcement Learning.

A View on Deep Reinforcement Learning in System Optimization Learning to Optimize Join Queries With Deep Reinforcement Learning

Reference 2002

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.481128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.481128Z digest=sha256:61a0f72a6527d5d2f70bd47f9276e018d93154e38c04a1f77fad561c1634b6d1

Observation dca263bd-c607-4c3f-84da-2ad2d4d02024 · outbound

This paper cites M., Pahl, C., Metzger, A., and Estrada, G.

A View on Deep Reinforcement Learning in System Optimization M., Pahl, C., Metzger, A., and Estrada, G

Reference 2008

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:21:51.095317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.476129Z digest=sha256:26aa5dd8f672f33247a5547f8ba88ff8100efc8b2e67c7d74d9908c04e398383

Observation 3f65f03d-e9c0-4e6b-8269-d58db559c0ac · outbound

This paper cites Iroko: A Framework to Prototype Reinforcement Learning for Data Center Traffic Control.

A View on Deep Reinforcement Learning in System Optimization Iroko: A Framework to Prototype Reinforcement Learning for Data Center Traffic Control

Reference 2011

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.553187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.553187Z digest=sha256:f3b556391c4cc748adcca2c086ccea07d8d1c3b7d9b0dd6cfad6085b5ce15b1e

Observation 621689b6-c168-4c04-97de-037d4b72e670 · outbound

This paper cites an unresolved cited work.

A View on Deep Reinforcement Learning in System Optimization Unresolved cited work

Reference 2012

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:21:50.960793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.572810Z digest=sha256:6f6fa42e5f3ba1ebe1f7824f705d0052d131be64c31591f8e8453dfc30d65690

Observation 85855dde-fd10-4204-9fbf-a17f6c1e966d · outbound

This paper cites A hierarchical framework of cloud resource allo- cation and power management using deep reinforcement learning.

A View on Deep Reinforcement Learning in System Optimization A hierarchical framework of cloud resource allo- cation and power management using deep reinforcement learning

Reference 2013

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:21:51.047557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T15:21:50.512996Z digest=sha256:d7318ff56714d379ec253cdb0959e8cb3e695f7994843e8bec54054f1263bcbf

Observation 5ce10ee3-c0b1-47f6-9238-b612a0de9e03 · outbound

This paper cites Horizon: Facebook's Open Source Applied Reinforcement Learning Platform.

A View on Deep Reinforcement Learning in System Optimization Horizon: Facebook's Open Source Applied Reinforcement Learning Platform

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.450088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.450088Z digest=sha256:c2518f9a55de21fd43d408a49a4446750258b0b75a2ab95cbd4ca1fc0dd85c16

Observation e67973b2-5a57-4a21-9d62-e75a6d1f27bf · outbound

This paper cites Neural Symbolic Machines: Learning Semantic Parsers on Freebase with Weak Supervision.

A View on Deep Reinforcement Learning in System Optimization Neural Symbolic Machines: Learning Semantic Parsers on Freebase with Weak Supervision

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.492154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.492154Z digest=sha256:7fd29ecbf8c4ab46699dff6b73b7d1a18a85559c534fa217382912391a07ea1a

Observation aca0b995-cb4e-4dd7-9516-27df91712d13 · outbound

This paper cites Castro, P.

A View on Deep Reinforcement Learning in System Optimization Castro, P

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.434077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.434077Z digest=sha256:38db12f051d5527cb41fbd3941954eeeb121c4a7ce280a02631ce009a0520b25

Observation fe4bf235-4093-4328-a5d4-39a4d09e56f5 · outbound

This paper cites Dopamine: A Research Framework for Deep Reinforcement Learning.

A View on Deep Reinforcement Learning in System Optimization Dopamine: A Research Framework for Deep Reinforcement Learning

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.439294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.439294Z digest=sha256:18250303dcb198fdd0e089c151e7449ba46a34e441f7d7d4e117fa0ad3f7d0f4

Observation 0020236b-a2b5-42d6-9bfd-54c9f8b01aca · outbound

This paper cites Continuous control with deep reinforcement learning.

A View on Deep Reinforcement Learning in System Optimization Continuous control with deep reinforcement learning

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-14T15:21:50.507654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:21:50.507654Z digest=sha256:cf74f2fc60abc6a0702e243833dcd8204d3bc8c1970d79b8f149891a210ff40c

Pith citing papers

No inbound Pith citation observations are available.