Pith. sign in

Paper Citation Record · LEDGER

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning

As of 23 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 0 inbound Pith citation observations for arXiv:2607.18830.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.18830 v1

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T14:15:14.358886Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

23 of 23 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7b4f6fc8-9410-457b-9a71-da84699f2aee · outbound

This paper cites Meta-Reinforcement Learning via Language Instructions.

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning Meta-Reinforcement Learning via Language Instructions

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T14:15:14.236250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:15:14.236250Z digest=sha256:d53a7c5eae1d86473a1c2509e20091d96deb34f89e6c330f3a410015e577a537

Observation 5b5ee942-2a51-4d7c-b136-015d75a6dce5 · outbound

This paper cites End to End Learning for Self-Driving Cars.

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning End to End Learning for Self-Driving Cars

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T14:15:14.242664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:15:14.242664Z digest=sha256:eb2e687310d05312dd1c00e1c8621c71cee7ee6b7f88849a4b35598cab6b0fa5

Observation 596c7fbd-8913-42be-8452-e72057e22bca · outbound

This paper cites BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning.

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T14:15:14.248277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:15:14.248277Z digest=sha256:374afe4ce580781e7fa631ee8003ed2d816846dd65cf40b3a847ca70f860ac39

Observation a5431748-7ba4-4dea-bd62-497f709f206d · outbound

This paper cites In: Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing.

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning In: Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T14:15:14.253534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:15:14.253534Z digest=sha256:c84e8c7ec8947be7695d6108a69c184c7e159d1a701b99caef16c4c1e7d6ef62

Observation 95d8c804-4203-4ebd-a002-5711f78515fd · outbound

This paper cites In: International conference on machine learning.

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning In: International conference on machine learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T14:15:14.259141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:15:14.259141Z digest=sha256:d36fd3131c17db56ce609d6174be9adf10eb4459a8b0e9db5c992fc2f1667942

Observation 861ae59b-623f-414f-b05b-ca740fe6ea56 · outbound

This paper cites Meta-Learning with Warped Gradient Descent.

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning Meta-Learning with Warped Gradient Descent

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T14:15:14.265099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:15:14.265099Z digest=sha256:cf6eb17e1c6c3b2d97162e7287c0ded18bc1f4efc21a85a31ec7f3ff84456025

Observation 6fbad931-15b4-4a91-b545-40b860f574e3 · outbound

This paper cites The Inter- national Journal of Robotics Research40(4-5), 698–721 (2021).

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning The Inter- national Journal of Robotics Research40(4-5), 698–721 (2021)

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T14:15:14.272617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:15:14.272617Z digest=sha256:370d07ad5424f6ce63dd3e47943d6c5096859133bc63938e6f1a265f3a1da263

Observation 8e2b09cd-8129-4b92-b1f5-c5451aa874e5 · outbound

This paper cites Journal of Machine Learning Research17(39), 1–40 (2016).

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning Journal of Machine Learning Research17(39), 1–40 (2016)

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T14:15:14.277992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:15:14.277992Z digest=sha256:7d27f1f285d353b53d81b72fa232e40e1f877badf3ae6e21ea1753a07a1ad3aa

Observation 3016345c-f4b2-4a56-9ea5-c90cb7d68d34 · outbound

This paper cites Statistics in medicine41(20), 4034–4056 (2022).

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning Statistics in medicine41(20), 4034–4056 (2022)

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T14:15:14.283442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:15:14.283442Z digest=sha256:8ceaf784bc81def0697d72a47947c23e72217c7e4c632acc0a81f80991398ef1

Observation 3ae61141-3e89-43a1-b1d5-6f8fc0c3aedf · outbound

This paper cites In: Proceedings of the 2022 Conference of the North American Chapter of the Association for Computational Linguistics (NAACL) (2022).

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning In: Proceedings of the 2022 Conference of the North American Chapter of the Association for Computational Linguistics (NAACL) (2022)

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T14:15:14.290319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:15:14.290319Z digest=sha256:34ae3566532bffe8e70b13fc6029d28091de0f0e3708af5a2119477ebf205d84

Observation 148eb2f1-99a8-41f1-a175-8180402f59ca · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning Playing Atari with Deep Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T14:15:14.295217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:15:14.295217Z digest=sha256:6c0b4c889a6a3587f2d0941df39d8aa9b29d1aaecf7275cdd6d94ba472ee436e

Observation 59d6dc82-4de6-4754-b94a-89cf4c718146 · outbound

This paper cites Nature518, 529–533 (2015).

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning Nature518, 529–533 (2015)

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T14:15:14.300469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:15:14.300469Z digest=sha256:af83099b23721e24054eeba03b61bc8e949fc5f9378dfd65c4ba1506b1ffc629

Observation 43e587c0-de4e-4c26-81df-a8cdccbf226e · outbound

This paper cites In: Advances in Neural Information Processing Systems (NeurIPS) (2018).

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning In: Advances in Neural Information Processing Systems (NeurIPS) (2018)

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T14:15:14.305749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:15:14.305749Z digest=sha256:23c96a3e4c058793fd3a3352402fbb5a528ad24d73d9532056603f1233f54405

Observation 593cda1a-2f2c-42c4-8fa7-872aeda69a00 · outbound

This paper cites Rapid Learning or Feature Reuse? Towards Understanding the Effectiveness of MAML.

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning Rapid Learning or Feature Reuse? Towards Understanding the Effectiveness of MAML

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T14:15:14.311571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:15:14.311571Z digest=sha256:7a258efe7a9b5fefa0f36b34897c1e09cf3275f341f8a4cf2267cd6c69862fc5

Observation 04b664d7-8abc-450f-b201-42b68969937d · outbound

This paper cites In: Proceedings of the 2019 conference on empirical methods in natural language processing and the 9th international joint conference on natural language processing (EMNLP-IJCNLP).

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning In: Proceedings of the 2019 conference on empirical methods in natural language processing and the 9th international joint conference on natural language processing (EMNLP-IJCNLP)

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T14:15:14.317280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:15:14.317280Z digest=sha256:bbb65f181fcef7527254e5d6a861796c628654c0afde7fdde123fcdf7c46c057

Observation 91fbff65-d82c-41e6-a0dc-c918a6dabd3c · outbound

This paper cites Meta-Learning with Latent Embedding Optimization.

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning Meta-Learning with Latent Embedding Optimization

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T14:15:14.322590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:15:14.322590Z digest=sha256:3cb7cb3e91a038a6b33fc9b90bebb9a5f4d88304771be7a439ed52aacb538033

Observation 92011d35-e887-4ad2-8c13-452e0df2e864 · outbound

This paper cites In: International conference on machine learning.

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning In: International conference on machine learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T14:15:14.328466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:15:14.328466Z digest=sha256:0804ad778507a0d45a22c9d5eb8305dcf3e659d6461d1c96d11cbfe7f1b9c3b5

Observation bc7776ec-1b3d-4631-967d-9d78f8305d7c · outbound

This paper cites High-Dimensional Continuous Control Using Generalized Advantage Estimation.

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning High-Dimensional Continuous Control Using Generalized Advantage Estimation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T14:15:14.333922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:15:14.333922Z digest=sha256:3d28a0d532c1dee9080e5d90555bdb3df43aee0380d02b3fabba6252ec1f9c77

Observation f5dbf621-7bea-47ab-8e31-36c33b5cc53c · outbound

This paper cites Nature529, 484–489 (2016).

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning Nature529, 484–489 (2016)

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T14:15:14.339225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:15:14.339225Z digest=sha256:c429fefdc2e50850d99ae10bd2f5f779d42b8904819265d6939515bfed88b30b

Observation 8463e505-486f-4750-a5f8-4fa6700086aa · outbound

This paper cites In: Pro- ceedings of the 2020 Conference on Robot Learning (CoRL) (2020).

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning In: Pro- ceedings of the 2020 Conference on Robot Learning (CoRL) (2020)

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T14:15:14.344080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:15:14.344080Z digest=sha256:91bd8b5680fa066a394635346551c02d853b6eb3a6cde1c529e43dd81073e1d2

Observation 630ead43-f794-4f13-9ee1-97582ca47b1e · outbound

This paper cites MIT Press, 2 edn.

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning MIT Press, 2 edn

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T14:15:14.348989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:15:14.348989Z digest=sha256:1a2f0e71d4972a15186ab02a3ff818e1878f0a59b763f58adac00e87ddd4c8ba

Observation 4ba8d989-6be6-4a42-afd3-60e6e0e68394 · outbound

This paper cites In: Proceedings of the 2022 IEEE International Conference on Robotics and Automation (ICRA) (2022).

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning In: Proceedings of the 2022 IEEE International Conference on Robotics and Automation (ICRA) (2022)

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T14:15:14.353930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:15:14.353930Z digest=sha256:326be98966ea768eb341ad63dea4755c19d99f3c36dac8781d1bedfc6a69661d

Observation 65fb7753-eeb9-4732-8735-095e87c5bee5 · outbound

This paper cites In: Proceedings of the 36th International Conference on Machine Learning (ICML) (2019).

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning In: Proceedings of the 36th International Conference on Machine Learning (ICML) (2019)

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T14:15:14.358886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:15:14.358886Z digest=sha256:2021ffd20915e975c0189c05dbce8b40fb8dcdd01c01af1a4b7b3196f4fbd9d2

Pith citing papers

No inbound Pith citation observations are available.