Pith. sign in

Paper Citation Record · LEDGER

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs

As of 9 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 0 inbound Pith citation observations for arXiv:2506.10527.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.10527 v1

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:28:47.872578Z

measured 38 of 38 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

38 of 38 outbound references displayed

  • verified exact1
  • verified fuzzy13
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 13daddeb-bf98-4bd7-9a13-2efcb2f7d51f · outbound

This paper cites How far can transformers reason? the globality barrier and inductive scratchpad.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs How far can transformers reason? the globality barrier and inductive scratchpad

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:51.870479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:28:37.631385Z digest=sha256:3b0774f7b8fc7ab5303235fb4f778c135648320e31ac32014d9232add00ae3b6

Observation 4a180558-2c7b-4523-96ac-45160c5a1113 · outbound

This paper cites Graph of logic: Enhancing llm reasoning with graphs and symbolic logic.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Graph of logic: Enhancing llm reasoning with graphs and symbolic logic

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:51.648962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:28:39.992525Z digest=sha256:7c1790c7cd21a31e574d3a51dab7b4f75a15fe770c4c38b0ec8e049ece9c2761

Observation 5d0623af-db05-4bc4-9540-25aa8c057937 · outbound

This paper cites Claude 3.7 sonnet.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Claude 3.7 sonnet

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:51.345069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:28:40.109199Z digest=sha256:b61c54ba69c8de5e518917f41745b0f6f277f7d0cbd5cc9d3ea070e079abf8d0

Observation 333256f3-6047-47a1-ae20-4dd8f38e1ec4 · outbound

This paper cites Eureka: Evaluating and Understanding Large Foundation Models.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Eureka: Evaluating and Understanding Large Foundation Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:40.218979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:40.218979Z digest=sha256:ba38a22e557260b2a566ceb113ed89abcbc469fa7c36bac04119089f9a48ee63

Observation cffd9cab-db11-42e7-ada0-fde6f9ff23c5 · outbound

This paper cites The Llama 3 Herd of Models.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs The Llama 3 Herd of Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:40.339630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:40.339630Z digest=sha256:3e7702ac172aaafdecca0bf38553944ce5892d16bc23a2e8eb434f2b73d0ccde

Observation ab83d070-3c04-44dc-b9e6-6c70cba195da · outbound

This paper cites NPHardEval: Dynamic Benchmark on Reasoning Ability of Large Language Models via Complexity Classes.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs NPHardEval: Dynamic Benchmark on Reasoning Ability of Large Language Models via Complexity Classes

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:40.881171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:40.881171Z digest=sha256:63e44a169737a352a8af9373b8198bf1ce3e8f256f2d3817de9bfc225937091d

Observation 52e18b69-3525-43f3-9d0c-641f6da8c5a4 · outbound

This paper cites Magentic-One: A Generalist Multi-Agent System for Solving Complex Tasks.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Magentic-One: A Generalist Multi-Agent System for Solving Complex Tasks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:43.800298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:43.800298Z digest=sha256:424e139682e84302d6168cf0bf35ad70a02a7137556a69ef637875b532748b49

Observation e381ab0e-bd16-476e-afba-2cea2d894a18 · outbound

This paper cites Gemini 2.0 pro experimental.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Gemini 2.0 pro experimental

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:51.060732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:28:43.913151Z digest=sha256:4df2ab2238f172ea67eb6314e40227ca7f5b63ea0afdf738838723c863b644f2

Observation 6a919f3c-4495-4f5c-a3f8-3db8f25737dd · outbound

This paper cites Gemini 2.0 flash thinking.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Gemini 2.0 flash thinking

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:50.806928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:28:44.022394Z digest=sha256:dd65efb95bb7353f9944d933fe084806e5fc2871363a39b4ac07c29881d4e357

Observation cdfb6979-b996-405f-b957-0f4421b6c7fe · outbound

This paper cites Reinforced Self-Training (ReST) for Language Modeling.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Reinforced Self-Training (ReST) for Language Modeling

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:44.141478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:44.141478Z digest=sha256:c20f291e897b9756dfa03f75782a11d118be55f2cee842d05b76bceeed8507cf

Observation 1d341136-a6eb-4881-b45a-183389793d9e · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:44.268329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:44.268329Z digest=sha256:a7041e6f76ed53c8b74a3af764d0eba11d9ab8b41f912cebab3aa54bf977ec43

Observation 34041c10-3514-41ee-8637-d2c7787ecfc9 · outbound

This paper cites GPT-4o System Card.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs GPT-4o System Card

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:44.369152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:44.369152Z digest=sha256:77ca251bc5189e51b41cd63e514055ccbe8288f5e2b593755045fecedc513836

Observation e0e8ec66-628e-48b0-ab30-204a7afcd6f8 · outbound

This paper cites OpenAI o1 System Card.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs OpenAI o1 System Card

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:44.472582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:44.472582Z digest=sha256:ea5c85a11f6329a983e9b97550219378f16a541fb43fa17508bfaa851b53ea9a

Observation e6c594e3-5fc4-4736-a8c4-6c679e502ebf · outbound

This paper cites Finding all the elementary circuits of a directed graph.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Finding all the elementary circuits of a directed graph

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:50.551608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:28:44.662847Z digest=sha256:5ae1224f2e7281c2850f164ff84ae11484c30dafe8c6fc15360191c8b6321943

Observation ee54dd7e-7624-49af-8097-9be19ed5a087 · outbound

This paper cites Same task, more tokens: the impact of input length on the reasoning performance of large language models.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Same task, more tokens: the impact of input length on the reasoning performance of large language models

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:50.304512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:28:44.779665Z digest=sha256:4b583c86f214f28db0d16d2cb2f93877a963f7ddf7e909a45f26732974856f15

Observation 40a4ecee-aee7-46c8-9ade-57a6469fbfae · outbound

This paper cites Large Language Models for Supply Chain Optimization.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Large Language Models for Supply Chain Optimization

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:44.907012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:44.907012Z digest=sha256:4f4a04efd657ea155c8d3c01045c552f9b312e089c28bd23b155738aefcace74

Observation 859e10bb-1161-4366-bb42-3df6f0a3b072 · outbound

This paper cites Llms for relational reasoning: How far are we? In Proceedings of the 1st International Workshop on Large Language Models for Code, pp.\ 119--126, 2024.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Llms for relational reasoning: How far are we? In Proceedings of the 1st International Workshop on Large Language Models for Code, pp.\ 119--126, 2024

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:50.020611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:28:45.061476Z digest=sha256:5c397c02b82796b978fe138558b3dac31f2621f4dd4276d8397136100eaa7daf

Observation 75549691-995f-42ec-85b7-bf025908a967 · outbound

This paper cites Let's verify step by step.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Let's verify step by step

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:45.169542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:45.169542Z digest=sha256:7a9fd549fa190e47d1e26f4dd12f1d59e21c5c74f78a6f6c36af046954d0b3fa

Observation 0585ec8d-bfb5-4b8b-8c30-03a8a3682e69 · outbound

This paper cites ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:45.311682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:45.311682Z digest=sha256:520084a9cf16db602a6759fd20a0769578e94e449c7af70156c32f76c7260399

Observation 1198c1e5-a59f-4cd8-a408-54ed93a49fad · outbound

This paper cites Evaluating cognitive maps and planning in large language models with cogeval.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Evaluating cognitive maps and planning in large language models with cogeval

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:49.777154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:28:45.438885Z digest=sha256:df4aefe8de2499bdbf8c5fc5caa303dddbf77cb78321244eeacd0c9c2b3f68b3

Observation 71007d16-5174-49f2-af57-cf863ebb13bb · outbound

This paper cites Foster, and Michael W.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Foster, and Michael W

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:49.524885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:28:45.587781Z digest=sha256:acc3d31682a12c6e6eeb2de3ba6aac603b3104ad7f08f2a9b09b8b9631349470

Observation 8d9ed02d-68bb-4d41-9b87-63c757460c4f · outbound

This paper cites Openai o3-mini system card.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Openai o3-mini system card

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:45.681138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:45.681138Z digest=sha256:484001e75899d62e124ea1988a492d8d0b53092c7bc8b53d71a2b4e2d53307c2

Observation e493e0fb-5ddf-4608-b0c2-eb598dbf49d3 · outbound

This paper cites Logicbench: Towards systematic evaluation of logical reasoning ability of large language models.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Logicbench: Towards systematic evaluation of logical reasoning ability of large language models

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:49.303330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:28:45.808086Z digest=sha256:482f180e44d5e5e668d8b20ecb1cd47b3691a09e6f8ddb4ae72bb318795203e8

Observation eec52d23-b3ff-4bc4-a8f0-e17870037a28 · outbound

This paper cites Multi-LogiEval: Towards Evaluating Multi-Step Logical Reasoning Ability of Large Language Models.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Multi-LogiEval: Towards Evaluating Multi-Step Logical Reasoning Ability of Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:45.947400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:45.947400Z digest=sha256:5f0372569bb7e89a8b98900dfaa2d767de270912dbac905d32f590518215d8e6

Observation 9571dae9-3f7e-403e-a0e1-e884514fadfb · outbound

This paper cites Divide and translate: Compositional first-order logic translation and verification for complex logical reasoning.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Divide and translate: Compositional first-order logic translation and verification for complex logical reasoning

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:49.020000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:28:46.143857Z digest=sha256:a1864e421b977afc392a7dfc45cb3aadb4137ec5773fb2e579a6387929e3b6e3

Observation da054956-7561-4149-95b1-ce862f3536ca · outbound

This paper cites MoreHopQA: More Than Multi-hop Reasoning.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs MoreHopQA: More Than Multi-hop Reasoning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:46.289978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:46.289978Z digest=sha256:69888a2148bd9aa7d5e96257f012e61abba40636536ad00853fdfe66c2d96dbf

Observation 7cfadf7a-1629-496a-aadb-5c31cfd83f56 · outbound

This paper cites Causal Language Modeling Can Elicit Search and Reasoning Capabilities on Logic Puzzles.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Causal Language Modeling Can Elicit Search and Reasoning Capabilities on Logic Puzzles

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:46.408842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:46.408842Z digest=sha256:e80b8d6f975b017e6bce602b38d2c78c78e9850aa9522b581bb93e564a1e66e2

Observation d708a61d-0097-4c7d-8de8-c0b73add71d0 · outbound

This paper cites Musique: Multihop questions via single-hop question composition.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Musique: Multihop questions via single-hop question composition

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:46.557922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:46.557922Z digest=sha256:4f7380c317e63359326d31f6353d81a323fec6ae035fee87a1702c1069bed7ed

Observation b8a51838-5ab7-44b6-bc01-16188e1fbe78 · outbound

This paper cites Holy Grail 2.0: From Natural Language to Constraint Models.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Holy Grail 2.0: From Natural Language to Constraint Models

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:28:48.259984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:28:46.653749Z digest=sha256:b544a7d303431a293ec05a54f2bd32de23d8386d68f52cc396f5543d51243f06

Observation 21847219-af4a-477b-958a-87aa8001bb95 · outbound

This paper cites On the planning abilities of large language models-a critical investigation.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs On the planning abilities of large language models-a critical investigation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:46.787264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:46.787264Z digest=sha256:107acde5bc46a4e851b0198e8fbd903a24474a221ab40aa5fdab93b40aff9d62

Observation 16a779d1-6d61-4191-9907-fea88806053e · outbound

This paper cites LLMs Still Can't Plan; Can LRMs? A Preliminary Evaluation of OpenAI's o1 on PlanBench.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs LLMs Still Can't Plan; Can LRMs? A Preliminary Evaluation of OpenAI's o1 on PlanBench

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:46.939636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:46.939636Z digest=sha256:82b3065129972d969f02a34d5af540154b6c34d2c0c203bc5a73e54de2af8dbb

Observation 0a527187-c70f-4ec1-b20d-88839e383437 · outbound

This paper cites Is A Picture Worth A Thousand Words? Delving Into Spatial Reasoning for Vision Language Models.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Is A Picture Worth A Thousand Words? Delving Into Spatial Reasoning for Vision Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:47.081061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:47.081061Z digest=sha256:07c544bba3405884704b242f9f92a656a20056ec7b34366e1c6fac3353cab5cc

Observation eba4517b-6b66-496a-b150-5b5060f6f192 · outbound

This paper cites On The Planning Abilities of OpenAI's o1 Models: Feasibility, Optimality, and Generalizability.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs On The Planning Abilities of OpenAI's o1 Models: Feasibility, Optimality, and Generalizability

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:47.210808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:47.210808Z digest=sha256:de2f8efbf9538e00a1a8f0634a795ad8a3fa8b9acbb5602c640014f0c4d1263e

Observation d6923359-51ac-47db-a56c-e6e919cba353 · outbound

This paper cites Math-shepherd: Verify and reinforce llms step-by-step without human annotations.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Math-shepherd: Verify and reinforce llms step-by-step without human annotations

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:48.740642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:28:47.330230Z digest=sha256:a5ca1b3ffac2b28adf84c12e969106fc30063dd43460e595cb2888e4a63d0a02

Observation 2f3f4aa8-d39d-4e1d-832d-41bd705e4542 · outbound

This paper cites AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:47.511050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:47.511050Z digest=sha256:b837415ee147a9ae25e215f28abebdb1a6971d3a1270ac26c4215a4b6f0f7af5

Observation 4979fd67-1488-487f-acdc-b9661780569c · outbound

This paper cites Faithful Logical Reasoning via Symbolic Chain-of-Thought.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Faithful Logical Reasoning via Symbolic Chain-of-Thought

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:47.609701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:47.609701Z digest=sha256:d55247fe1df8d7bc0c145e769d855d06549695a9d4d0012539d5b43dc751a14a

Observation 5bae4fd2-961f-4349-bb38-18dbad1a73e3 · outbound

This paper cites HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:47.750383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:47.750383Z digest=sha256:b4253135219449887fb2bbf08dc893a99407beda6a89556a7edc5ff1462a69ee

Observation 3f1b9139-7109-4809-907e-8bb7be6e5fb6 · outbound

This paper cites NATURAL PLAN: Benchmarking LLMs on Natural Language Planning.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs NATURAL PLAN: Benchmarking LLMs on Natural Language Planning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:47.872578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:47.872578Z digest=sha256:7ddce6faa5499faf3741418603bf997a716e810c240291b8b1d77d85ab8fa009

Pith citing papers

No inbound Pith citation observations are available.