Pith. sign in

Paper Citation Record · LEDGER

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs

As of 21 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 0 inbound Pith citation observations for arXiv:2506.10527.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.10527 v1

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:28:47.872578Z

measured 38 of 38 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

38 of 38 outbound references displayed

  • verified exact1
  • verified fuzzy13
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 13daddeb-bf98-4bd7-9a13-2efcb2f7d51f · outbound

This paper cites How far can transformers reason? the globality barrier and inductive scratchpad.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs How far can transformers reason? the globality barrier and inductive scratchpad

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:51.870479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T04:28:37.631385Z digest=sha256:5153cb70039b0e781acc9c59289e84636db29e6bfabda6f0760ef129edc3aae5

Observation 4a180558-2c7b-4523-96ac-45160c5a1113 · outbound

This paper cites Graph of logic: Enhancing llm reasoning with graphs and symbolic logic.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Graph of logic: Enhancing llm reasoning with graphs and symbolic logic

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:51.648962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T04:28:39.992525Z digest=sha256:5036366eda17fa20ebf59a7d92b75a68e8d6ea2e3b9f15dd7db421d7b1472c18

Observation 5d0623af-db05-4bc4-9540-25aa8c057937 · outbound

This paper cites Claude 3.7 sonnet.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Claude 3.7 sonnet

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:51.345069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T04:28:40.109199Z digest=sha256:966c7053152e62032e78d71cadb669b9081bad1977a2598f3abffe68670a5ff6

Observation 333256f3-6047-47a1-ae20-4dd8f38e1ec4 · outbound

This paper cites Eureka: Evaluating and Understanding Large Foundation Models.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Eureka: Evaluating and Understanding Large Foundation Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:40.218979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:40.218979Z digest=sha256:31e3f9c62687584b82bc191d82485b4cce7afeb2f9bb3acf96bb2e570e250310

Observation cffd9cab-db11-42e7-ada0-fde6f9ff23c5 · outbound

This paper cites The Llama 3 Herd of Models.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs The Llama 3 Herd of Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:40.339630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:40.339630Z digest=sha256:2059c05335179335a25467ce0b27fa853ad61f7fef67244592f6674fc10a762d

Observation ab83d070-3c04-44dc-b9e6-6c70cba195da · outbound

This paper cites NPHardEval: Dynamic Benchmark on Reasoning Ability of Large Language Models via Complexity Classes.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs NPHardEval: Dynamic Benchmark on Reasoning Ability of Large Language Models via Complexity Classes

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:40.881171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:40.881171Z digest=sha256:b67911f7e49dab76049cba8a51f72c84f1db50bcf4614a08f536c70f0a3d0c90

Observation 52e18b69-3525-43f3-9d0c-641f6da8c5a4 · outbound

This paper cites Magentic-One: A Generalist Multi-Agent System for Solving Complex Tasks.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Magentic-One: A Generalist Multi-Agent System for Solving Complex Tasks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:43.800298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:43.800298Z digest=sha256:faf6fd012cea45d05d64673dd21ea2765990d6862a5bf8ad3a9332ff7b0d06f4

Observation e381ab0e-bd16-476e-afba-2cea2d894a18 · outbound

This paper cites Gemini 2.0 pro experimental.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Gemini 2.0 pro experimental

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:51.060732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T04:28:43.913151Z digest=sha256:7e35df0173ec57a47a6587e7928c327e9ee134cd8f9d63d59043e95059a526f4

Observation 6a919f3c-4495-4f5c-a3f8-3db8f25737dd · outbound

This paper cites Gemini 2.0 flash thinking.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Gemini 2.0 flash thinking

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:50.806928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T04:28:44.022394Z digest=sha256:abdc9795217a530c9cf017ad7b123e6c958393522a743263c41ad768f0b331bb

Observation cdfb6979-b996-405f-b957-0f4421b6c7fe · outbound

This paper cites Reinforced Self-Training (ReST) for Language Modeling.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Reinforced Self-Training (ReST) for Language Modeling

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:44.141478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:44.141478Z digest=sha256:1dae57639da2e344ab6d6e81a8b11767ba32591bf7eb2ef09c06a5847f4c5a52

Observation 1d341136-a6eb-4881-b45a-183389793d9e · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:44.268329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:44.268329Z digest=sha256:ad4fba445aa284dfcbd5a6e70b27fad05325adcb0f14bea3db0212574124ed9a

Observation 34041c10-3514-41ee-8637-d2c7787ecfc9 · outbound

This paper cites GPT-4o System Card.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs GPT-4o System Card

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:44.369152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:44.369152Z digest=sha256:5a9f15adcb883f26efd230ad261e52f01348378bab0aeb92c41fa7e8c89cc78d

Observation e0e8ec66-628e-48b0-ab30-204a7afcd6f8 · outbound

This paper cites OpenAI o1 System Card.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs OpenAI o1 System Card

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:44.472582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:44.472582Z digest=sha256:4a7aa5dc8d4724e05e1010aaebf2a08c5aab431e192b90aeaf24e3995c23e746

Observation e6c594e3-5fc4-4736-a8c4-6c679e502ebf · outbound

This paper cites Finding all the elementary circuits of a directed graph.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Finding all the elementary circuits of a directed graph

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:50.551608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T04:28:44.662847Z digest=sha256:4f5d8456218abce2d4ca4670e5ccf1ca9860615e951eb177b21001460057b6a5

Observation ee54dd7e-7624-49af-8097-9be19ed5a087 · outbound

This paper cites Same task, more tokens: the impact of input length on the reasoning performance of large language models.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Same task, more tokens: the impact of input length on the reasoning performance of large language models

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:50.304512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T04:28:44.779665Z digest=sha256:e50463d52b1af80312f2210209ebf0d9691f97a8bc373608b966cbadfcfa1588

Observation 40a4ecee-aee7-46c8-9ade-57a6469fbfae · outbound

This paper cites Large Language Models for Supply Chain Optimization.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Large Language Models for Supply Chain Optimization

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:44.907012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:44.907012Z digest=sha256:ddedcf970aae2da2b6a34b6ca5dd18c832cafd5ef8e142e99317ef7b11c996d7

Observation 859e10bb-1161-4366-bb42-3df6f0a3b072 · outbound

This paper cites Llms for relational reasoning: How far are we? In Proceedings of the 1st International Workshop on Large Language Models for Code, pp.\ 119--126, 2024.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Llms for relational reasoning: How far are we? In Proceedings of the 1st International Workshop on Large Language Models for Code, pp.\ 119--126, 2024

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:50.020611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T04:28:45.061476Z digest=sha256:2753e3c73dfb2cdf407b32f522ffc98188d4e80ea9ecac3ea44b1b49d0db6465

Observation 75549691-995f-42ec-85b7-bf025908a967 · outbound

This paper cites Let's verify step by step.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Let's verify step by step

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:45.169542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:45.169542Z digest=sha256:b171bebd5f762e9a21603480e6b9eefb310addc2b62932cba2be6b39233ff0cc

Observation 0585ec8d-bfb5-4b8b-8c30-03a8a3682e69 · outbound

This paper cites ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:45.311682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:45.311682Z digest=sha256:54a6db117a63497928d83db640bfb6cf1af080cb4ea42ee93cf9291f8e7b04fa

Observation 1198c1e5-a59f-4cd8-a408-54ed93a49fad · outbound

This paper cites Evaluating cognitive maps and planning in large language models with cogeval.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Evaluating cognitive maps and planning in large language models with cogeval

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:49.777154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T04:28:45.438885Z digest=sha256:123bf19e4bf1c770834eb59afea246b7020b3913c4c7cf05e90d8c7684667df9

Observation 71007d16-5174-49f2-af57-cf863ebb13bb · outbound

This paper cites Foster, and Michael W.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Foster, and Michael W

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:49.524885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T04:28:45.587781Z digest=sha256:f7bf058b03894d9c0f243597ddef548d89a971dbc1cdb05001099b746acde808

Observation 8d9ed02d-68bb-4d41-9b87-63c757460c4f · outbound

This paper cites Openai o3-mini system card.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Openai o3-mini system card

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:45.681138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:45.681138Z digest=sha256:14992e308b8aed6f85562c6c72873ae9cc662721a4cbdc49d1369291e7be6f9d

Observation e493e0fb-5ddf-4608-b0c2-eb598dbf49d3 · outbound

This paper cites Logicbench: Towards systematic evaluation of logical reasoning ability of large language models.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Logicbench: Towards systematic evaluation of logical reasoning ability of large language models

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:49.303330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T04:28:45.808086Z digest=sha256:3cd0588cb85d602c086fa23af03406dcc5ec24b1e9333aece59cadff1767daae

Observation eec52d23-b3ff-4bc4-a8f0-e17870037a28 · outbound

This paper cites Multi-LogiEval: Towards Evaluating Multi-Step Logical Reasoning Ability of Large Language Models.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Multi-LogiEval: Towards Evaluating Multi-Step Logical Reasoning Ability of Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:45.947400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:45.947400Z digest=sha256:b752e1733bd0dc893fc258059fc2e14c93b2b6d86e6c1b4ccb4ee0ca4fdd8f35

Observation 9571dae9-3f7e-403e-a0e1-e884514fadfb · outbound

This paper cites Divide and translate: Compositional first-order logic translation and verification for complex logical reasoning.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Divide and translate: Compositional first-order logic translation and verification for complex logical reasoning

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:49.020000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T04:28:46.143857Z digest=sha256:0c0ea2debee9e4a5015ecdc600e5f938357cb5ef44d5d610c8b0d14963d14ab8

Observation da054956-7561-4149-95b1-ce862f3536ca · outbound

This paper cites MoreHopQA: More Than Multi-hop Reasoning.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs MoreHopQA: More Than Multi-hop Reasoning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:46.289978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:46.289978Z digest=sha256:b2ca85721363d17fab960c23f60753396b42d0ac75f47f651547e3922939ff05

Observation 7cfadf7a-1629-496a-aadb-5c31cfd83f56 · outbound

This paper cites Causal Language Modeling Can Elicit Search and Reasoning Capabilities on Logic Puzzles.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Causal Language Modeling Can Elicit Search and Reasoning Capabilities on Logic Puzzles

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:46.408842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:46.408842Z digest=sha256:8c1475ccd54a46c2771253a5bfd9249339493fb6cfd515679bf20c55a3576492

Observation d708a61d-0097-4c7d-8de8-c0b73add71d0 · outbound

This paper cites Musique: Multihop questions via single-hop question composition.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Musique: Multihop questions via single-hop question composition

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:46.557922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:46.557922Z digest=sha256:97c3b7a52b083c48c5309214dc51b31a4f576d74d62b96118d2d31f74b6cf821

Observation b8a51838-5ab7-44b6-bc01-16188e1fbe78 · outbound

This paper cites Holy Grail 2.0: From Natural Language to Constraint Models.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Holy Grail 2.0: From Natural Language to Constraint Models

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:28:48.259984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T04:28:46.653749Z digest=sha256:e38f4985688e85f49a22d0c0ce9624c4e70822d8f99fdb5267e0b511a8c1f977

Observation 21847219-af4a-477b-958a-87aa8001bb95 · outbound

This paper cites On the planning abilities of large language models-a critical investigation.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs On the planning abilities of large language models-a critical investigation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:46.787264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:46.787264Z digest=sha256:2b7a0b9ebdfe990eace37fe013c213e922447286f268e85f424179fc8f3acc30

Observation 16a779d1-6d61-4191-9907-fea88806053e · outbound

This paper cites LLMs Still Can't Plan; Can LRMs? A Preliminary Evaluation of OpenAI's o1 on PlanBench.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs LLMs Still Can't Plan; Can LRMs? A Preliminary Evaluation of OpenAI's o1 on PlanBench

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:46.939636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:46.939636Z digest=sha256:a5a8724082803330307d474f8ea19a5d23d47249620b620f66ab7efae34cf236

Observation 0a527187-c70f-4ec1-b20d-88839e383437 · outbound

This paper cites Is A Picture Worth A Thousand Words? Delving Into Spatial Reasoning for Vision Language Models.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Is A Picture Worth A Thousand Words? Delving Into Spatial Reasoning for Vision Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:47.081061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:47.081061Z digest=sha256:d6bc4c38738e1ad840fa9eb3fb745e5544b04342376f85a9548d88d5362f6349

Observation eba4517b-6b66-496a-b150-5b5060f6f192 · outbound

This paper cites On The Planning Abilities of OpenAI's o1 Models: Feasibility, Optimality, and Generalizability.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs On The Planning Abilities of OpenAI's o1 Models: Feasibility, Optimality, and Generalizability

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:47.210808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:47.210808Z digest=sha256:92a749dee5979078cfc82154f71bb9a3cb3639ee4fd46748593ff7f57dadd00a

Observation d6923359-51ac-47db-a56c-e6e919cba353 · outbound

This paper cites Math-shepherd: Verify and reinforce llms step-by-step without human annotations.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Math-shepherd: Verify and reinforce llms step-by-step without human annotations

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:48.740642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T04:28:47.330230Z digest=sha256:9a2afd778d57faafbff2e254f2058babfefa30907c1f3012dfcfe5d77c58bee7

Observation 2f3f4aa8-d39d-4e1d-832d-41bd705e4542 · outbound

This paper cites AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:47.511050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:47.511050Z digest=sha256:8ba4648c24a09b73cd3e4abdb8cbf57a552c7932eb28d2e1a053f2ab7e1b9b93

Observation 4979fd67-1488-487f-acdc-b9661780569c · outbound

This paper cites Faithful Logical Reasoning via Symbolic Chain-of-Thought.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Faithful Logical Reasoning via Symbolic Chain-of-Thought

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:47.609701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:47.609701Z digest=sha256:41e2a33f925ada45f7bcaf00256a1e24eb57ff35c4f6dc4463045b741e5bbf95

Observation 5bae4fd2-961f-4349-bb38-18dbad1a73e3 · outbound

This paper cites HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:47.750383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:47.750383Z digest=sha256:3f9ed934d3443da4e03c80e3191f7da90f72dcd7124c7bba8219d10b8fd68496

Observation 3f1b9139-7109-4809-907e-8bb7be6e5fb6 · outbound

This paper cites NATURAL PLAN: Benchmarking LLMs on Natural Language Planning.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs NATURAL PLAN: Benchmarking LLMs on Natural Language Planning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:47.872578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:47.872578Z digest=sha256:de9830a50c391adebf38a4608b4efbcf72a34c8a3195e6b361dbac0118199b46

Pith citing papers

No inbound Pith citation observations are available.