Pith. sign in

Paper Citation Record · LEDGER

JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2311.05997.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2311.05997 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:18:28.360265Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

6
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7f5ae896-6dc5-4d83-a399-e905b05d9976 · inbound

AppAgent: Multimodal Agents as Smartphone Users cites this paper.

AppAgent: Multimodal Agents as Smartphone Users JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-17T10:16:43.859010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T10:16:43.364787Z digest=sha256:3ca01ece5cd9ee6e4d122205a2806695a79c734382a9e67e031c1d67cb7d8da5

Observation 15526ff7-f592-46b4-97f6-d0f44ac0f41c · inbound

A Survey on the Memory Mechanism of Large Language Model based Agents cites this paper.

A Survey on the Memory Mechanism of Large Language Model based Agents JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 159

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:21:39.639525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T07:21:39.440092Z digest=sha256:5b17c331bb1cc690a9d37780ac4265b529f5faaba14ef7b9253054c66c9d3104

Observation 824b9835-d3f1-4f75-a8c2-726ef302f0a7 · inbound

Preference Goal Tuning: Post-Training as Latent Control for Frozen Policies cites this paper.

Preference Goal Tuning: Post-Training as Latent Control for Frozen Policies JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:22:44.348038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-23T08:20:05.898025Z digest=sha256:0263283e403b8a26a7b46f690bfd3e32e626717681ad605f36197c1a158023fc

Observation 7d1ff9b9-edd1-4000-a3c4-76583f6e71b5 · inbound

Conditional Multi-Stage Failure Recovery for Embodied Agents cites this paper.

Conditional Multi-Stage Failure Recovery for Embodied Agents JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:28.360265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:18:28.360265Z digest=sha256:6f3b83321dac441742ba40e1d419e7c1710e4f629c8038189a3c4e6e96fd1c9d

Observation 3a0e4777-3c75-4ded-818a-f0f47d392161 · inbound

SoK: Agentic Skills -- Beyond Tool Use in LLM Agents cites this paper.

SoK: Agentic Skills -- Beyond Tool Use in LLM Agents JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:19:31.125901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-14T23:19:31.024268Z digest=sha256:afa297beff3f81336aaf663e26c3581d96f7e92c88c0ebb6333740ec77999704

Observation 739e99f3-24a0-4af3-9f5c-d42b8a4c71a9 · inbound

GameWorld: Towards Standardized and Verifiable Evaluation of Multimodal Game Agents cites this paper.

GameWorld: Towards Standardized and Verifiable Evaluation of Multimodal Game Agents JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:46:07.697516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T17:57:36.038091Z digest=sha256:a6700aabcab51b12b0ee68ce285e18b9e512f1def2f231fe5be853636af12213

Observation 52b44462-e38f-410a-be05-86e6d43ddfeb · inbound

Dream-Cubed: Controllable Generative Modeling in Minecraft by Training on Billions of Cubes cites this paper.

Dream-Cubed: Controllable Generative Modeling in Minecraft by Training on Billions of Cubes JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:31:02.641583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T01:36:23.517745Z digest=sha256:4a2f22e776bf68f069cb54ea6bf3eef5663743231bd12f39523fe09dd6f77ea6

Observation 7c497296-ea37-4c23-811e-5fa4a4a2e53d · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:20:57.243431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T01:47:39.926540Z digest=sha256:61ad698dfbdfe61a48f0960d8626fa009e73d2dcd93e842a2d4a24bc8f1b0c10

Observation 3693cabc-54e7-4662-ad09-4ad65c97457a · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-20T23:19:15.144913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T23:15:44.550045Z digest=sha256:96bcc24a3f9924e711cbb1b79e35404cc4db6e20422d0e6354a22cb826df1a46

Observation fe68687f-7002-4018-be4d-a9fef70d153a · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:25:07.322279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T23:23:42.883286Z digest=sha256:3cd790712fc3f95646f5ec9bcb06b3768d7fe8c86310f5d575e2a72a7443a2b3

Observation 9a54de7f-832a-45f2-a6a5-5807ba3488a9 · inbound

Do Vision-Language-Models show human-like logical problem-solving capability in point and click puzzle games? cites this paper.

Do Vision-Language-Models show human-like logical problem-solving capability in point and click puzzle games? JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T02:02:05.547051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-13T01:58:39.476408Z digest=sha256:9ddbb3e7b9dd9b444a3cc3666260d99e86fd0a15b6bf873d601c6e6786b79309

Observation 38091461-5cb6-4fdf-9737-c8aeff9360c9 · inbound

SPIKE: An Adaptive Dual Controller Framework for Cost-Efficient Long-Horizon Game Agents cites this paper.

SPIKE: An Adaptive Dual Controller Framework for Cost-Efficient Long-Horizon Game Agents JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:33:12.266648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T10:32:26.668583Z digest=sha256:1d8bc9f3dac93d902fbc9c2ea1fc0044ac1cf2420af2e9dfad4082ee24e352f5

Observation 0ce76651-5009-4389-bde0-20a88a50eefc · inbound

HiMe: Hierarchical Embodied Memory for Long-Horizon Vision-Language-Action Control cites this paper.

HiMe: Hierarchical Embodied Memory for Long-Horizon Vision-Language-Action Control JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-12T02:24:21.020383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T02:24:21.020383Z digest=sha256:0648bea3ae4b48d1000c196bcbc79c6fda2f8b695e6b9077171d3125106f8587

Observation 1a894190-d2c8-4030-ae46-267294b59629 · inbound

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills cites this paper.

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 259

Resolution
unresolved
no resolver link, observed 2026-08-04T19:45:35.319727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:45:35.319727Z digest=sha256:cccd317e4944d69d5af13f586807aafa88ee30fd7c49f22be769d9e987ccd6a2