Pith. sign in

Paper Citation Record · LEDGER

PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 26 inbound Pith citation observations for arXiv:2411.00081.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.00081 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 26 of 26 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:42:30.589774Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T23:27:38.656546Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8a141125-b6a5-4702-8f69-3006b1a4fc4c · inbound

TeamCraft: A Benchmark for Multi-Modal Multi-Agent Systems in Minecraft cites this paper.

TeamCraft: A Benchmark for Multi-Modal Multi-Agent Systems in Minecraft PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T20:53:34.919498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:53:34.919498Z digest=sha256:ed4126293992b4a38c1893d64948b7e99db07744495df20c881d5c1612b3e5c9

Observation 5ffd778c-5ba8-43a5-82b1-5f5319a52ff6 · inbound

Effect of Adaptive Communication Support on LLM-powered Human-Robot Collaboration cites this paper.

Effect of Adaptive Communication Support on LLM-powered Human-Robot Collaboration PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T12:42:07.142284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:42:07.142284Z digest=sha256:69cbc3895d23c3af0a66a07dd2d97d29c83306bb62f5c8f719632870dcba5d8a

Observation ba249ccf-2780-4f95-bdfb-a67ccd12da2c · inbound

ViGiL3D: A Linguistically Diverse Dataset for 3D Visual Grounding cites this paper.

ViGiL3D: A Linguistically Diverse Dataset for 3D Visual Grounding PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T22:32:38.624156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:32:38.624156Z digest=sha256:c59e2e43dc5d0b3f89eeb2074e50abc82871fbe5cbe61f0bfbf440002ce84ed9

Observation 05665e30-b543-419c-8c3a-a6438eb26726 · inbound

PLANET: A Collection of Benchmarks for Evaluating LLMs' Planning Capabilities cites this paper.

PLANET: A Collection of Benchmarks for Evaluating LLMs' Planning Capabilities PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T11:42:30.589774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:42:30.589774Z digest=sha256:b5dc16373183fdbf2ec34e14eeb39570b11018d306073a6d20488970b51e6cee

Observation 90fa97ae-9aba-45c6-a045-acbff35d598e · inbound

Collaborating Action by Action: A Multi-agent LLM Framework for Embodied Reasoning cites this paper.

Collaborating Action by Action: A Multi-agent LLM Framework for Embodied Reasoning PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T10:32:50.464554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:32:50.464554Z digest=sha256:f1b0057a4253397090aea4d9d8f397ebfcd21fa0177712f0a0ced1b49b608d8a

Observation ae589b1e-e781-4a62-8341-cbddfde35355 · inbound

Large Language Models for Planning: A Comprehensive and Systematic Survey cites this paper.

Large Language Models for Planning: A Comprehensive and Systematic Survey PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:11:51.625209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:11:51.625209Z digest=sha256:748386329d726ceff600214db4221d6619134cb29f9d06029709db815443245f

Observation 24701b72-183c-4e40-8b69-5a8f9fe7bd56 · inbound

TextAtari: 100K Frames Game Playing with Language Agents cites this paper.

TextAtari: 100K Frames Game Playing with Language Agents PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T10:51:57.934910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:51:57.934910Z digest=sha256:2d988a485ae111ecb0f1c02a7d382b91731ec90b728f2da1f2149480cbc76164

Observation bc6a5c96-906c-4130-b6b0-394a37ef44cf · inbound

A Call for Collaborative Intelligence: Why Human-Agent Systems Should Precede AI Autonomy cites this paper.

A Call for Collaborative Intelligence: Why Human-Agent Systems Should Precede AI Autonomy PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:51:55.370889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:51:55.370889Z digest=sha256:f89481aeaf2e662d33f91bad956f7fb5e40bc57215053c4c800bf5ca79063f9e

Observation f9a3f08d-56b6-4d43-b786-a86940cb4b05 · inbound

Agent Identity Evals: Measuring Agentic Identity cites this paper.

Agent Identity Evals: Measuring Agentic Identity PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T14:57:06.679450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:57:06.679450Z digest=sha256:ee5ea8a31fd17dc2ef2ae818f6ee783ddd28c2fb07796c489777b957b436e61f

Observation 1493e9ca-6552-4e83-9fd1-9439732b98df · inbound

ProToM: Promoting Prosocial Behaviour via Theory of Mind-Informed Feedback cites this paper.

ProToM: Promoting Prosocial Behaviour via Theory of Mind-Informed Feedback PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T05:43:44.535671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T05:43:44.535671Z digest=sha256:148c19b92a402b81a915e7a3e989bc6b712d2728d108f7fac81a339d3faf40ac

Observation 3241174d-e951-4a3d-b54f-fcbe8d14bf2b · inbound

When Should Users Check? Modeling Confirmation Frequency inMulti-Step Agentic AI Tasks cites this paper.

When Should Users Check? Modeling Confirmation Frequency inMulti-Step Agentic AI Tasks PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-18T09:01:08.966964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-18T08:59:35.944554Z digest=sha256:8af5c1b8bb2792955c36c163eb625044d6c2c774b2ca76b10552358756659903

Observation 254c8083-b32e-4293-b3a9-3f54ad28a658 · inbound

A Survey on Evaluating Quality and Trustworthiness in LLM-Generated Data cites this paper.

A Survey on Evaluating Quality and Trustworthiness in LLM-Generated Data PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T08:15:14.592377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T08:15:14.592377Z digest=sha256:51257bec4ef2de317bb3974113b732cf109928bc7389bff7df6bc73fe152bb81

Observation d5a0eae3-e905-4563-add2-d0b6ca384061 · inbound

ST-BiBench: Benchmarking Multi-Stream Multimodal Coordination in Bimanual Embodied Tasks for MLLMs cites this paper.

ST-BiBench: Benchmarking Multi-Stream Multimodal Coordination in Bimanual Embodied Tasks for MLLMs PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-16T06:00:40.476019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-16T06:00:02.043029Z digest=sha256:6ea982c267cdc3b3da124160e57283a1e2f21d676ee3fb74574f560514ccba45

Observation a6524d9d-dde1-4cbf-b589-9d0f485b7339 · inbound

Any House Any Task: Scalable Long-Horizon Planning for Abstract Human Tasks cites this paper.

Any House Any Task: Scalable Long-Horizon Planning for Abstract Human Tasks PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T06:05:29.182557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:05:29.182557Z digest=sha256:40afc43d33f56c183e4ac8e92b5a9931d527802a8dd1c95c97f05206d0080ff5

Observation a9a19770-539a-445b-a314-c2ad4cfe4685 · inbound

LLM-WikiRace Benchmark: How Far Can LLMs Plan over Real-World Knowledge Graphs? cites this paper.

LLM-WikiRace Benchmark: How Far Can LLMs Plan over Real-World Knowledge Graphs? PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T22:26:34.782737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T22:26:34.782737Z digest=sha256:70c60f1e34724688d6d24ab0c6a98a93c8531e96f0ac008ae10395e737880b06

Observation 2dca667a-62c8-4a17-8ad6-20fc1932113a · inbound

EmbodiedGovBench: A Benchmark for Governance, Recovery, and Upgrade Safety in Embodied Agent Systems cites this paper.

EmbodiedGovBench: A Benchmark for Governance, Recovery, and Upgrade Safety in Embodied Agent Systems PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:06.677152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T15:21:20.231759Z digest=sha256:348a682b5cbf27dbfcdfd904a92ac5c24506ffa2286458aa8f588993e0685d79

Observation c86b7f68-ac2e-40a8-9164-ed2d551725d5 · inbound

PersonalHomeBench: Evaluating Agents in Personalized Smart Homes cites this paper.

PersonalHomeBench: Evaluating Agents in Personalized Smart Homes PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T06:45:11.853525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-15T06:41:17.586880Z digest=sha256:e431cdfd1e8924d9ed4b43aafc3807ff69db1b36916257cc37023c40925b418a

Observation 41c036e8-2977-41c2-99cd-339ad2ed5e92 · inbound

TaskGround: Structured Executable Task Inference for Full-Scene Household Reasoning cites this paper.

TaskGround: Structured Executable Task Inference for Full-Scene Household Reasoning PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:03:13.990666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-20T10:53:44.014259Z digest=sha256:f7491ec140d243f19f47275d6f1dfc8c3606672904630c6f7ff61b9a5d9c7dd7

Observation 97131752-cd1a-4e84-a37f-31322e583fd5 · inbound

Seeing Together: Multi-Robot Cooperative Egocentric Spatial Reasoning with Multimodal Large Language Models cites this paper.

Seeing Together: Multi-Robot Cooperative Egocentric Spatial Reasoning with Multimodal Large Language Models PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:38:12.529366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-20T10:36:20.388169Z digest=sha256:c742c082ccbb01b00c92b3e3c62e63000b61d82dad63ac37b600e04993d6bc85

Observation 3577e93f-7686-401b-9fca-57f7a887ab90 · inbound

PACT: Proactive Asking for Continual Task Assistance in Human-Robot Collaboration cites this paper.

PACT: Proactive Asking for Continual Task Assistance in Human-Robot Collaboration PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-30T13:44:41.298644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-30T13:34:46.547334Z digest=sha256:9ffb4d9488efa0d31d40bb23b3c5e2dd4d50351dd718df81621e9001a8474c84

Observation fe4cc417-286c-477a-a58a-84e8aebf2ca0 · inbound

AdaPlanBench: Evaluating Adaptive Planning in Large Language Model Agents under World and User Constraints cites this paper.

AdaPlanBench: Evaluating Adaptive Planning in Large Language Model Agents under World and User Constraints PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-12T15:06:31.394162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T15:06:31.394162Z digest=sha256:fd87e470fbddaa3ea1c3d29c3116ac5552fa31a1be4a1c394ddc8985fe68ef07

Observation b7900c57-6db3-472a-862d-15ca122da883 · inbound

LLawCo: Learning Laws of Cooperation for Modeling Embodied Multi-Agent Behavior cites this paper.

LLawCo: Learning Laws of Cooperation for Modeling Embodied Multi-Agent Behavior PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T19:43:55.332960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-29T04:38:04.136601Z digest=sha256:335ade312ce923820172b12f474f67524b40f38dcc31736862f483e18e4a1bb0

Observation f760ef2d-dd2f-4aa0-87a2-7522e22190f2 · inbound

CAIRN: Cross-Room 3D Scene Understanding with Topology-Aware Large Multimodal Models cites this paper.

CAIRN: Cross-Room 3D Scene Understanding with Topology-Aware Large Multimodal Models PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-08T02:44:27.711175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-07-08T02:44:00.608590Z digest=sha256:4788d70f5acd79fe128f88a488e35119654be5baf2d879c2ab27697824c107fb

Observation abb6cf32-2618-412d-9753-8afbc5da791a · inbound

CAIRN: Cross-Room 3D Scene Understanding with Topology-Aware Large Multimodal Models cites this paper.

CAIRN: Cross-Room 3D Scene Understanding with Topology-Aware Large Multimodal Models PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-14T16:00:13.133298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:00:13.133298Z digest=sha256:4801af0c686a92e203fb796801715e10f68c335b5f8a2b9b59e7dad1e5698272

Observation 9ad4e48c-1c38-47ba-80ed-fa30fb8871cc · inbound

CoMind: Understanding Collaborative Human Activity from Multiple Minds and Views cites this paper.

CoMind: Understanding Collaborative Human Activity from Multiple Minds and Views PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-07-10T23:27:38.672502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-07-10T23:20:12.870365Z digest=sha256:4d40ec2dc4753d0ee5f0c0c8565134168ae1e9038159a9b79d6967d60b81984e

Observation d959edc4-cab0-491d-80f4-c34d7d274ea8 · inbound

Long-Horizon Embodied Decision-Making via Multimodal Memory Compression cites this paper.

Long-Horizon Embodied Decision-Making via Multimodal Memory Compression PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T00:13:21.218440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:13:21.218440Z digest=sha256:69a600e912c3564a4b474dfb0214f23f87b3b7ac1a2c8f0e8f77bc46ba1d9c3e