Pith. sign in

Paper Citation Record · LEDGER

Learn How to Cook a New Recipe in a New House: Using Map Familiarization, Curriculum Learning, and Bandit Feedback to Learn Families of Text-Based Adventure Games

As of 15 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 2 inbound Pith citation observations for arXiv:1908.04777.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1908.04777 v3

Coverage vector

measured 21 of 21 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T13:35:58.270098Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-14T10:35:35.341385Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T04:57:26.039049Z

Reference resolution

21 of 21 outbound references displayed

  • verified exact0
  • verified fuzzy16
  • unresolved4
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 49fd27ff-1e88-4005-883e-1b3f3f224c7a · outbound

This paper cites Improved algorithms for lin- ear stochastic bandits.

Learn How to Cook a New Recipe in a New House: Using Map Familiarization, Curriculum Learning, and Bandit Feedback to Learn Families of Text-Based Adventure Games Improved algorithms for lin- ear stochastic bandits

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:35:58.740721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T13:35:58.163240Z digest=sha256:7e10ccf6fce88f7c2b502ce59d7e5f3cc9e9bc93f518ee5686ba9e2d824ab012

Observation 2e32097e-793e-48f2-b8af-37d2ea5d43b7 · outbound

This paper cites Using confidence bounds for exploitation-exploration trade-offs.

Learn How to Cook a New Recipe in a New House: Using Map Familiarization, Curriculum Learning, and Bandit Feedback to Learn Families of Text-Based Adventure Games Using confidence bounds for exploitation-exploration trade-offs

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:35:58.691411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T13:35:58.185299Z digest=sha256:228edcf984b73375af6bb42b71e9f41f063758ff4a41d5d54b7ede889702e919

Observation 25194306-23e7-4bd7-89f6-ac3234c9dd7a · outbound

This paper cites Curriculum learning.

Learn How to Cook a New Recipe in a New House: Using Map Familiarization, Curriculum Learning, and Bandit Feedback to Learn Families of Text-Based Adventure Games Curriculum learning

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:35:58.670521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T13:35:58.190755Z digest=sha256:83da085732864e5d0b40d62c0d15dd64b88d8dc9bffab56d74a25e29357d458d

Observation f3bc4b54-d1b0-4bad-a22b-84f5ad4177af · outbound

This paper cites What can you do with a rock? affordance extraction via word embeddings.

Learn How to Cook a New Recipe in a New House: Using Map Familiarization, Curriculum Learning, and Bandit Feedback to Learn Families of Text-Based Adventure Games What can you do with a rock? affordance extraction via word embeddings

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:35:58.653404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T13:35:58.200821Z digest=sha256:f0b8151abbbe7f2600e15417b7c6675eed4d1815c56f4645b9898fb28e03de44

Observation cb26437d-46f0-4357-b1ea-d22f84665e7a · outbound

This paper cites Text-based ad- ventures of the golovin AI agent.

Learn How to Cook a New Recipe in a New House: Using Map Familiarization, Curriculum Learning, and Bandit Feedback to Learn Families of Text-Based Adventure Games Text-based ad- ventures of the golovin AI agent

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:35:58.529862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T13:35:58.229028Z digest=sha256:8c900f66fe89fbf5abbbb54cfa766164ed11eb16f3f6b2da0ba3644e12e4c5b7

Observation 314396b1-7c90-4927-8566-38263dd1f001 · outbound

This paper cites Special feature zork: A computerized fantasy simulation game.

Learn How to Cook a New Recipe in a New House: Using Map Familiarization, Curriculum Learning, and Bandit Feedback to Learn Families of Text-Based Adventure Games Special feature zork: A computerized fantasy simulation game

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:35:58.509251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T13:35:58.235016Z digest=sha256:68bef3037fc76961e8c42cbb8f21c062dedeec2ce8a78db0cf2ad1ddd0f92586

Observation 030c3a27-1da6-4018-86a8-9a3ddd4b0a92 · outbound

This paper cites [Mnih et al., 2015] V olodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A.

Learn How to Cook a New Recipe in a New House: Using Map Familiarization, Curriculum Learning, and Bandit Feedback to Learn Families of Text-Based Adventure Games [Mnih et al., 2015] V olodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:35:58.471003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T13:35:58.247782Z digest=sha256:9f2237e19ee4dc5cf30d32699b2e8ffce1b43e7db098cd8a532b1d772186a9a9

Observation 08f12976-8c7d-4c51-a526-c74b562635ff · outbound

This paper cites Language understanding for text- based games using deep reinforcement learning.

Learn How to Cook a New Recipe in a New House: Using Map Familiarization, Curriculum Learning, and Bandit Feedback to Learn Families of Text-Based Adventure Games Language understanding for text- based games using deep reinforcement learning

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:35:58.456021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T13:35:58.251548Z digest=sha256:cb4b5b8b71a978ec91aa1261f6cc0276d5070a0db7714cb28a134f9be776fc16

Observation 1d146a91-aab8-48fa-8ae4-b2e2add60c97 · outbound

This paper cites Gumbel-max trick and weighted reservoir sampling,.

Learn How to Cook a New Recipe in a New House: Using Map Familiarization, Curriculum Learning, and Bandit Feedback to Learn Families of Text-Based Adventure Games Gumbel-max trick and weighted reservoir sampling,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:35:58.441073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T13:35:58.256587Z digest=sha256:1d798e2aafed82b07ad6c4f84618f92cfe3d5dbe5431fb4afc6609cf9a54966e

Observation 5914f598-89c6-45bd-90df-affea5c13473 · outbound

This paper cites Mankowitz, and Shie Mannor.

Learn How to Cook a New Recipe in a New House: Using Map Familiarization, Curriculum Learning, and Bandit Feedback to Learn Families of Text-Based Adventure Games Mankowitz, and Shie Mannor

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:35:58.420705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T13:35:58.270098Z digest=sha256:0bdc933d5a205633c2eec4f064560bc843c680706e5ceebfd6568b3c1d688cfa

Observation 34b3b0a9-6f41-45ed-8ab6-dc321124f405 · outbound

This paper cites Deep reinforce- ment learning for dialogue generation.

Learn How to Cook a New Recipe in a New House: Using Map Familiarization, Curriculum Learning, and Bandit Feedback to Learn Families of Text-Based Adventure Games Deep reinforce- ment learning for dialogue generation

Reference 1979

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:35:58.490191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T13:35:58.241690Z digest=sha256:bbd9e91bf024e44395348505d0b2cc670e4038dacaa4d86f540a51bee7d2ae7a

Observation ece17bb7-24c2-44d9-bfcd-4ca1a1347070 · outbound

This paper cites Kingma and Jimmy Ba.

Learn How to Cook a New Recipe in a New House: Using Map Familiarization, Curriculum Learning, and Bandit Feedback to Learn Families of Text-Based Adventure Games Kingma and Jimmy Ba

Reference 1997

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:35:58.554172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T13:35:58.223653Z digest=sha256:7be17d092fa7f8aeadab197b72f2feaa6cbb7ad1867495ae4387d88d5505fb4e

Observation 141ac34b-f589-49df-854f-89fa18b23b5e · outbound

This paper cites Playing Text-Adventure Games with Graph-Based Deep Reinforcement Learning.

Learn How to Cook a New Recipe in a New House: Using Map Familiarization, Curriculum Learning, and Bandit Feedback to Learn Families of Text-Based Adventure Games Playing Text-Adventure Games with Graph-Based Deep Reinforcement Learning

Reference 2003

Resolution
unresolved
no resolver link, observed 2026-08-14T13:35:58.173048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T13:35:58.173048Z digest=sha256:fee8f97a703b4fe51fa07399518a8842ce51efdb2209d0840f58a85e3a8b0c91

Reference 2009

Resolution
unresolved
no resolver link, observed 2026-08-14T13:35:58.195708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T13:35:58.195708Z digest=sha256:8f760e643554a338336fd7df1e1c122bf2d0f21de1ddaaa7968ae76bad534abf

Observation 53e11674-6300-420c-9af7-26d5bd709514 · outbound

This paper cites Biermann, and Philip M.

Learn How to Cook a New Recipe in a New House: Using Map Familiarization, Curriculum Learning, and Bandit Feedback to Learn Families of Text-Based Adventure Games Biermann, and Philip M

Reference 2011

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:35:58.720484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T13:35:58.168072Z digest=sha256:311913aacab3893eb7a20001f3292d2f242aa478b1425c1456a564c09ad59bb0

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-14T13:35:58.261101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T13:35:58.261101Z digest=sha256:afd118a69531d929100dc59004a49b6adce072554b70a68fcdbfa72c54355fc6

Observation 7f0ff4e1-31b5-4cac-8a0b-f0ec62c67aa7 · outbound

This paper cites Long short-term memory.

Learn How to Cook a New Recipe in a New House: Using Map Familiarization, Curriculum Learning, and Bandit Feedback to Learn Families of Text-Based Adventure Games Long short-term memory

Reference 2015

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:35:58.573496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T13:35:58.219020Z digest=sha256:87e9822bb1cb73d015b6a0442f55b07d76cf10d8a81fe391467615bde96acae7

Observation cec0376c-3d40-4a9b-9ada-2ef968e2b851 · outbound

This paper cites Distilling the knowledge in a neural net- work.

Learn How to Cook a New Recipe in a New House: Using Map Familiarization, Curriculum Learning, and Bandit Feedback to Learn Families of Text-Based Adventure Games Distilling the knowledge in a neural net- work

Reference 2016

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:35:58.599671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T13:35:58.214838Z digest=sha256:d74b8a8b9abe2ab3804a98d642e4de513fe8cb73e66d8a2ca1497e99bf5b2c04

Observation 4e2eeb51-9db4-4348-8024-aa7c5b850583 · outbound

This paper cites Deep reinforcement learning with a natural language ac- tion space.

Learn How to Cook a New Recipe in a New House: Using Map Familiarization, Curriculum Learning, and Bandit Feedback to Learn Families of Text-Based Adventure Games Deep reinforcement learning with a natural language ac- tion space

Reference 2017

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:35:58.628019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T13:35:58.207865Z digest=sha256:f735aeea04959692c3e52c711d0fa2c1430a2546cb836998785b8e06df6bf0fa

Reference 2018

Resolution
metadata mismatch
local_arxiv, observed 2026-08-14T13:35:58.380000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T13:35:58.180327Z digest=sha256:87ad506cfa5cdf00085375294cae950225aa94adb0cecae9faa7ea93d8a31703

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-14T13:35:58.265637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T13:35:58.265637Z digest=sha256:f29e0ad0f3980d02b1b98c6a5bd3e2a8914e2248737d1d6cb69f966ffdc153b9

Pith citing papers

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-14T10:35:35.341385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T10:35:35.341385Z digest=sha256:fb2051717b8ac9f16b001be64d436fdbb2ad2ce89cdc0e8787dac6c822138ffa

Observation 3545c3f8-e91f-4dfe-bd1a-23e547cb23a7 · inbound

Multi-Agent Language Models: Advancing Cooperation, Coordination, and Adaptation cites this paper.

Multi-Agent Language Models: Advancing Cooperation, Coordination, and Adaptation Learn How to Cook a New Recipe in a New House: Using Map Familiarization, Curriculum Learning, and Bandit Feedback to Learn Families of Text-Based Adventure Games

Reference 84

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:57:26.042534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T04:57:25.850667Z digest=sha256:d536a59b272723b69a7c537f4f305f1157638121fe24528ea37d594d2ab45aad