Pith. sign in

Paper Citation Record · LEDGER

BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:1810.08272.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1810.08272 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T20:18:58.240737Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T17:40:00.108219Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 540739da-ba30-414f-b151-62475d691526 · inbound

CraftAssist: A Framework for Dialogue-enabled Interactive Agents cites this paper.

CraftAssist: A Framework for Dialogue-enabled Interactive Agents BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-24T19:09:50.132244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-24T19:09:00.377577Z digest=sha256:420fa538f90691153e84fcefce946aea7ce020d01b2e2c4ed9f0e86ed0958a4c

Observation 603f1deb-4f5b-43a6-ae9a-839173ee58d7 · inbound

Why Build an Assistant in Minecraft? cites this paper.

Why Build an Assistant in Minecraft? BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-24T18:24:48.312953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-24T18:20:58.081505Z digest=sha256:7b7aaa11c1120721a52f93efc589675c550920c6c199c4d6471645189d1aa84d

Observation 842799ab-8624-438d-aa67-453bd2236cb3 · inbound

A Generalist Agent cites this paper.

A Generalist Agent BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T06:24:49.992877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T06:24:49.833638Z digest=sha256:9f9c55f0b16268279517e3e970574be7977db4ee8b824ac4cc71dfe80e2b1cdb

Observation 485a5663-62e4-4a79-a6e4-cb0cc128fa00 · inbound

Analyzing Adversarial Inputs in Deep Reinforcement Learning cites this paper.

Analyzing Adversarial Inputs in Deep Reinforcement Learning BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-24T03:43:50.404980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-24T03:40:04.265426Z digest=sha256:997fdcb376268bf6778a688b0ded3a9466b969fdfa7fc03dd2de4561ad096b20

Observation 809b04b4-4169-4548-80bd-e8bf0d814cc7 · inbound

Disentangling Exploration of Large Language Models by Optimal Exploitation cites this paper.

Disentangling Exploration of Large Language Models by Optimal Exploitation BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T20:18:58.240737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:18:58.240737Z digest=sha256:666157feeaa880922cce7141039c4585ca431af3967619cc52888431115fed63

Observation 65dd1c8a-886c-49e1-a6d8-9a71178eff5d · inbound

Large Language Models for Planning: A Comprehensive and Systematic Survey cites this paper.

Large Language Models for Planning: A Comprehensive and Systematic Survey BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:11:53.184655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:11:53.184655Z digest=sha256:5784e0b517d8c91f167e3196a9fd86b9ab16fb4340f1460cbe9e463b6555a950

Observation 407496b4-6860-4f83-b366-8e1563520047 · inbound

ChatPD: An LLM-driven Paper-Dataset Networking System cites this paper.

ChatPD: An LLM-driven Paper-Dataset Networking System BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T13:15:11.714556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:15:11.714556Z digest=sha256:2e0a70d8e950138b779bf40c55d6c7bbfc021e9f0511160b126ad312107717be

Observation ebaae906-8f3d-45cc-b8ea-90587c083b3a · inbound

Enhancing Decision-Making of Large Language Models via Actor-Critic cites this paper.

Enhancing Decision-Making of Large Language Models via Actor-Critic BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T10:55:44.269879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:55:44.269879Z digest=sha256:4f36644fbc7cfe92aff0309e2808168bbd2ea9abd3cf00539a4fae0de4e6d9bb

Observation 8bd8230d-eec6-49c6-ba37-2385a26eb049 · inbound

An Open-Source Software Toolkit & Benchmark Suite for the Evaluation and Adaptation of Multimodal Action Models cites this paper.

An Open-Source Software Toolkit & Benchmark Suite for the Evaluation and Adaptation of Multimodal Action Models BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:39.925877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:59:39.925877Z digest=sha256:03b8b6b940cd660e5a4202b00c1e1580921469136bbfaf39537639d74942435c

Observation eee8d759-5ce4-4cb6-abd5-57ce5b90f050 · inbound

Thinking Beyond Tokens: From Brain-Inspired Intelligence to Cognitive Foundations for Artificial General Intelligence and its Societal Impact cites this paper.

Thinking Beyond Tokens: From Brain-Inspired Intelligence to Cognitive Foundations for Artificial General Intelligence and its Societal Impact BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Reference 279

Resolution
unresolved
no resolver link, observed 2026-08-06T21:07:11.602044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:07:11.602044Z digest=sha256:ba6c2a0ca3fd8fa4728a887dfeee80da3e3e624766fa46c7875f85a0a4010df0

Observation fa740c9c-f980-4591-a60b-95ef6745e02f · inbound

MazeEval: A Benchmark for Testing Sequential Decision-Making in Language Models cites this paper.

MazeEval: A Benchmark for Testing Sequential Decision-Making in Language Models BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T13:40:01.294846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:40:01.294846Z digest=sha256:27f2107882fac9b69cddd858d0691f1f008b3476757ef294874ade0fcecc71f9

Observation ae23cd1e-948a-4017-843c-0e3b0346aa4f · inbound

WebFactory: Automated Compression of Foundational Language Intelligence into Grounded Web Agents cites this paper.

WebFactory: Automated Compression of Foundational Language Intelligence into Grounded Web Agents BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T16:50:10.658667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T16:50:02.527821Z digest=sha256:e4fc72a8d5e4b5ce74feb324f91f3f2e6472399bcbc1aed412a8d4d3bf6d1334

Observation 428bdca9-0931-45f9-b368-faffba6b31f4 · inbound

Self-Guided Plan Extraction for Instruction-Following Tasks with Goal-Conditional Reinforcement Learning cites this paper.

Self-Guided Plan Extraction for Instruction-Following Tasks with Goal-Conditional Reinforcement Learning BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T00:24:47.612613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T00:04:26.736202Z digest=sha256:9d808cab68091a404c13c4ba364bcbeaaeb8c157e6180d5166edd8946c06b21e

Observation b4b8b541-4680-4bed-af0d-05deee37b619 · inbound

SpecRLBench: A Benchmark for Generalization in Specification-Guided Reinforcement Learning cites this paper.

SpecRLBench: A Benchmark for Generalization in Specification-Guided Reinforcement Learning BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T21:51:30.333146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T04:01:31.928056Z digest=sha256:151cfd7dac9e71c39f0e6154b0d8da67236c310f05d2bec893fb9ba23864ac2f

Observation 4725f3e5-7403-40da-ab28-ac7b43d1d98c · inbound

Test-Time Deep Thinking to Explore Implicit Rules cites this paper.

Test-Time Deep Thinking to Explore Implicit Rules BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:54:38.193285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T11:52:15.163893Z digest=sha256:6070a63afbcf2dee523e570a290ef0cbd798cdcbba2ef529aaed8e91d091f6a1

Observation 6f805bff-4e57-42bb-9448-567f431b30cd · inbound

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback cites this paper.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T17:40:00.109677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:180b99c313be743acb985f286552d1c600799b3e9a85a134a7e799b16002931b

Observation 027843e5-3dc8-42fd-8b5a-7205ce5d6d85 · inbound

Spinning Straw into Gold: Relabeling LLM Agent Trajectories in Hindsight for Successful Demonstrations cites this paper.

Spinning Straw into Gold: Relabeling LLM Agent Trajectories in Hindsight for Successful Demonstrations BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-11T20:45:51.403759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T20:45:51.403759Z digest=sha256:8fefc4adbb9d37fcfa87490c7008ae5da27873f5ff6665f3f23346440118c812

Observation 2eac8b42-aadf-4898-840d-d8cb82ce6ee1 · inbound

Dual-Process Atomic Skill Learning: Decoupling Semantic Reasoning and Real-Time Control cites this paper.

Dual-Process Atomic Skill Learning: Decoupling Semantic Reasoning and Real-Time Control BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-14T10:23:25.158961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T10:23:25.158961Z digest=sha256:3ba29812d9876cccf518e42c8e692f4e1108b872034e718df3ef3903a558d8d1

Observation 596c7fbd-8913-42be-8452-e72057e22bca · inbound

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning cites this paper.

From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T14:15:14.248277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:15:14.248277Z digest=sha256:683d9a3946e4e3f8c40e0f56225d8f47029a45091737e9abb1adef54e8367475

Observation 408098c4-ecea-4f17-83b9-68e66d2cc408 · inbound

From Trajectories to Prefixes: Reusing Teacher Trajectories via Replayed Prefixes and Online Continuation cites this paper.

From Trajectories to Prefixes: Reusing Teacher Trajectories via Replayed Prefixes and Online Continuation BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:35.867517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:35.867517Z digest=sha256:279018d274fbfbc2d94de1bce8fa95cda55c7d6852a44fb9059de1cf4777dc58

Observation b631ade4-3698-4af9-bf15-705a5dac37a6 · inbound

Long-Horizon Embodied Decision-Making via Multimodal Memory Compression cites this paper.

Long-Horizon Embodied Decision-Making via Multimodal Memory Compression BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T00:13:23.118998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:13:23.118998Z digest=sha256:cae5ec1a35f46900a08462f2a1675c7425c6ae93e4c0e50823809ea6add1150f