Pith. sign in

Paper Citation Record · LEDGER

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization

As of 19 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 0 inbound Pith citation observations for arXiv:2506.01475.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.01475 v1

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:47:51.893563Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

49 of 49 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved49
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a36002a7-5bac-42bd-affb-97ac55fbee3c · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.991948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T11:47:00.510118Z digest=sha256:6756668bdc810f36e5365d62f500ed177fd09e2f2c84cdcfc745e52639dabf50

Observation 622a48e0-de72-472a-a0c6-f43cbd5a0d5a · outbound

This paper cites FireAct: Toward Language Agent Fine-tuning.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization FireAct: Toward Language Agent Fine-tuning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:00.567143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:00.567143Z digest=sha256:73f5d54d627e1defe2665c61ff4308a7b374ead5457e50e8211deefe1634b9a7

Observation cdb255ec-c282-429d-b500-d0974b2c4670 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.920232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T11:47:00.600539Z digest=sha256:d15f192a835d7dbe70e7d7b8bbc0ec771c3f3aa242d28955d200e605713216b6

Observation 09c23844-b8de-457a-8f49-b541a72ad153 · outbound

This paper cites The Llama 3 Herd of Models.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization The Llama 3 Herd of Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:00.634668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:00.634668Z digest=sha256:9d1560a37ffb61e0442ff8dc7e5683c0c15440f8e2e82bb67d7a7e9eb40addb1

Observation c9a98375-1970-4930-a02b-426f3af5ff02 · outbound

This paper cites AgentRefine: Enhancing Agent Generalization through Refinement Tuning.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization AgentRefine: Enhancing Agent Generalization through Refinement Tuning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:00.667874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:00.667874Z digest=sha256:50c36cc243c2e98a7ed7cc40d3e2cda71353aed3d7a2806c5ab218ddf349bfa1

Observation d1554dac-de26-4c17-a6d1-f9e56b6d969e · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:00.702676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:00.702676Z digest=sha256:4ecc07bcd8b20fad73b2e36bc368c4cbb902c09799db4a7e51e12f78fd295795

Observation 4f1f6fff-9af9-4d3e-b4ac-53ce6e465792 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:00.747128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:00.747128Z digest=sha256:82fd0565fecf39e8e333d6957e6b2d3b564b70ec18ebe4296400aad4edbcbe9f

Observation 2d03f915-c74c-48f6-80cf-1086413519b4 · outbound

This paper cites Understanding the planning of LLM agents: A survey.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Understanding the planning of LLM agents: A survey

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:00.789154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:00.789154Z digest=sha256:73b466c5da7954630d3762bdcb29138184a15dc0fb73acbe4a060237512ae513

Observation 145139b7-6a41-41b1-83c9-2d07506fb036 · outbound

This paper cites Mistral 7B.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Mistral 7B

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:49.031991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:49.031991Z digest=sha256:d11eef653181c1efff3def7a2c43855ee97b39393bc40b6ef30691d055602cf6

Observation 991364c6-71ba-430b-9574-4ba73a98a6cc · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:49.594789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:49.594789Z digest=sha256:4944ff41ba3f84e95197e80cefac5d87314366e2b6479ba01970379847a53aa1

Observation c8cd9652-6ac1-47a8-b7eb-d497f3294b5b · outbound

This paper cites Formal-LLM: Integrating Formal Language and Natural Language for Controllable LLM-based Agents.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Formal-LLM: Integrating Formal Language and Natural Language for Controllable LLM-based Agents

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:49.662536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:49.662536Z digest=sha256:de6dfdbf632b354d1019bffd3599d406dd1a9b12a3faad1fac2295cec004f0e6

Observation bdf620be-78d7-460a-8495-b6706a267e70 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.863579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T11:47:49.716341Z digest=sha256:190249de3db34aee82684eb67fba8962a1f05b25b4f18aaa34e16b373ce354a9

Observation ba4523f0-6713-44a6-b0b7-af57f0cb6422 · outbound

This paper cites LLM+P: Empowering Large Language Models with Optimal Planning Proficiency.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization LLM+P: Empowering Large Language Models with Optimal Planning Proficiency

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:49.774743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:49.774743Z digest=sha256:cbda689d9808869f4f11f7eaa23b4e4fbfb81f12a255398d00ef0900ceae221e

Observation 5af744e2-fe28-4738-b2fc-489e74b80782 · outbound

This paper cites BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:49.833363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:49.833363Z digest=sha256:1371a3e32a0095c68737ecf32a0c269177e7dac854235a2eceb546dc377e133b

Observation a65dd59a-a1b4-411f-b430-cb9d5ee88428 · outbound

This paper cites AgentLite: A Lightweight Library for Building and Advancing Task-Oriented LLM Agent System.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization AgentLite: A Lightweight Library for Building and Advancing Task-Oriented LLM Agent System

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:49.917199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:49.917199Z digest=sha256:48cd2613f8b9c534a7f93f1c25a2b6d0fedefda4934d7e611593cd0a501766a1

Observation 332c5315-02f6-4918-b037-9877450aa16c · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.843640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T11:47:49.986828Z digest=sha256:4ca1af6f05157ec76a4a7bf72de6e385bc40528b731f092d630bce7601809d49

Observation 2e3e4396-081c-46b7-96e1-774db6aa8875 · outbound

This paper cites Iterative Reasoning Preference Optimization.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Iterative Reasoning Preference Optimization

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:50.045904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:50.045904Z digest=sha256:1a51bc20d1a73b5e1c3987ffd007e0ed0987b8a5bd7fc993db3884cc9da51cfa

Observation 4745af46-8475-4974-8e1f-1c4c094354d4 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.820253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.078733Z digest=sha256:f7c7d8ad3901d954d2578c4b7a23a061e39261a14c10b7b8a148547515057c7a

Observation 350e6912-6345-47ef-9f5d-6cac0f3a2371 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.793120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.148555Z digest=sha256:25e91c627e89b17b87090123fe8f8171ea77b6658461b4c2d4e4e5990910f56d

Observation 5079c6a6-cdb6-41fd-b1ab-9f9ae3cc8a4c · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.769821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.207706Z digest=sha256:b055263456571904471f3eee6c652ec128a4c2f48a767ef1b3958451e9057dc5

Observation 931bd661-e97d-4f6f-967b-f24bcb3b0f54 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:50.267644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:50.267644Z digest=sha256:82ed93d2374465de081376c10db66a9239c9828091c232a5bb095cecb771c76a

Observation cea3d569-e049-4f60-a72b-a5d4fdeed13b · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.736404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.310634Z digest=sha256:b8276fce475cd214eb39dc20226c0aebe467af58c6da99f9f5d847d371ddb7f7

Observation 73379426-c257-4d6d-af78-5ded0d9fbf6f · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.708647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.357062Z digest=sha256:fb64ed1e84e8b87f28207477ca39bd7ffd55a523ac3846360dd47fac6f0be916

Observation 4a9cee47-1aa2-4991-a7b1-7859d5b8e63e · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.676944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.408001Z digest=sha256:0c6da42e1fcce27bc75e6ba9dd51658f379aea0ee8a4963881f677d5dd51e0aa

Observation 64644980-8ae8-475e-abf1-d0179d441fd0 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:50.456516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:50.456516Z digest=sha256:8a4d1832fdc358d8c3b829fd5f8e4bd81721a5367d505ab1598581c95f3461e2

Observation 23f5e70d-6c64-403e-bb6b-de896eb297d7 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.629814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.492430Z digest=sha256:6e666ed671df933d9d98d8f386f70572972de6a88f0ad59d32bfa34379039fa6

Observation 0fb23917-fac1-4484-b599-61a0918d3c2e · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.598303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.552291Z digest=sha256:bf4423fdf2e8c003af98f63fc8408ee9b0c480ccaef3361ad9628c42e8a00aeb

Observation 16145647-c582-436b-b3d2-29220fe5fe15 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.545699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.586993Z digest=sha256:9b3a1b16ac2ceb743b76b100ef9229ffc8fb7276ecefba8e5538d6101ccb6a92

Observation 71e2d4e4-b78d-4a1c-a11b-653df120ae42 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:50.631039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:50.631039Z digest=sha256:f4acd1339d3843747e0b73375a62844182903b0018106a32a41c2075c067efa4

Observation 2140e262-1074-482b-af60-d775f7fce3b1 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.517420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.667589Z digest=sha256:c9ddd529c113c6d816b17383aefb88270427d8a4cdb3fce20c89e3bc3a991cba

Observation 2446a275-e63f-4c1e-a0a0-90e6b32ed609 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.482921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.718877Z digest=sha256:1aa7b986f56d17fd85d2c0928c012c5941000b86a147eaeb664bd6cb44db5f7b

Observation 8e05df81-66a4-4122-be45-da0781c574ff · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.454199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.761020Z digest=sha256:f764f888925c3b22dcc135609cf27b40efa6b2a0fbde5af91e10cd62afa28cf3

Observation c4321bbf-c7f3-42b4-90a5-8aca749254dc · outbound

This paper cites Learning From Failure: Integrating Negative Examples when Fine-tuning Large Language Models as Agents.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Learning From Failure: Integrating Negative Examples when Fine-tuning Large Language Models as Agents

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:50.812929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:50.812929Z digest=sha256:933cfccf94b995043ef2a85b440bb0c5ee42eee40965d21499524e361a2f2ef1

Observation 850942b7-7b5d-49c8-bca6-142653ce5f88 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:50.865043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:50.865043Z digest=sha256:235b700820389d8fbda658b762395a011fb70a3471acb032e604b845a34246bb

Observation 427674e7-135b-4c20-add5-a291ce001290 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.414023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.899805Z digest=sha256:f7021305e5cf809451a69a40bb495a3e65a58fda6f5797dea45fb8aa0f83bfb8

Observation b920b238-47a8-41b8-a720-b299ae11feeb · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.386404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.938845Z digest=sha256:88a04f2dff749257ed2df8e0098d17466add276b73846de231cb0f032bbf0eba

Observation d8b74d7d-38b9-4532-b0f1-3b92aad717fa · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.359706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T11:47:51.013837Z digest=sha256:ea0688fafa305d1f4d27f4ecafda3608a573a16c5c6916c5906824e449de3664

Observation 6144216c-00b3-4443-bd3b-686053f31473 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:51.088342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:51.088342Z digest=sha256:a0415c4ef7edbaa4a4fdd5d21176cd36b19029ce22011006d22ea32a6182e57a

Observation a684049a-f95e-4239-8b00-b51f5d2c69a8 · outbound

This paper cites AgentGym: Evolving Large Language Model-based Agents across Diverse Environments.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization AgentGym: Evolving Large Language Model-based Agents across Diverse Environments

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:51.170109Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:51.170109Z digest=sha256:55d38c30915960a11318c515bbbdf9188af27fadcbecf8f1f516212a5414d15d

Observation 22913d76-7fc6-4849-b3bc-136441fc9911 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.327626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T11:47:51.261249Z digest=sha256:d5bef7941f58081e63a568ff5868184be272e399b309601c975053161d5c0e2c

Observation 7596d7b1-f3ef-428c-a474-82d1cdf0b5c4 · outbound

This paper cites Qwen2.5 Technical Report.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Qwen2.5 Technical Report

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:51.387390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:51.387390Z digest=sha256:a1dd2d63a580920105c3e4d01249adc1d626cf5f85a6131b3c67603ed19e73e0

Observation 825e833e-7986-4d95-8019-6472786c8528 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.306526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T11:47:51.487455Z digest=sha256:3b0632f8dec27eba50116092a3df32fd828e5bbca4a833aa43f2d077225bd5b7

Observation 9ceab20c-1743-4f48-b571-657861aa7ba1 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:51.572633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:51.572633Z digest=sha256:6dd96e03622520077e75084688010566511a042d69fec1dce5ee0e0832c9827c

Observation 04f4195d-9466-4ca0-941e-56761dc095d6 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:51.659712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:51.659712Z digest=sha256:5ce91d47361a0ae544e79e8515dc4a416507028219ce87f7d597db1fd0c67267

Observation 69a5f091-aa01-45de-a69e-8fbe41d9fae7 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:51.766466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:51.766466Z digest=sha256:14b5fa46046efc3871773398209100f7f64280fc3a5d991ec9a1687b5aad1f8d

Observation 370278cf-386c-497b-95ca-2d64b6804765 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:47:52.230431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T11:47:51.798750Z digest=sha256:d554d74a805813ea6cdca452fd38bbef0e815c88eceadceb2047f97cd134fc12

Observation bf651055-7644-4e73-97af-9589fc132d70 · outbound

This paper cites LaMMA-P: Generalizable Multi-Agent Long-Horizon Task Allocation and Planning with LM-Driven PDDL Planner.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization LaMMA-P: Generalizable Multi-Agent Long-Horizon Task Allocation and Planning with LM-Driven PDDL Planner

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:51.821700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:51.821700Z digest=sha256:dcd12223610c04f9d9b57b382588b6fa456ce1b6f647649a65d9570fd9abaadf

Observation 97feef45-44db-42a5-9660-e46edeba4a60 · outbound

This paper cites online" 'onlinestring :=.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization online" 'onlinestring :=

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:51.859589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:51.859589Z digest=sha256:b8131421241e3282015dc3f180d128b1d4e758f5cb9ec0cff8dd8ba37a4eddf3

Observation deed7873-ab9b-466f-b0bf-db31a38ac115 · outbound

This paper cites write newline.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization write newline

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:51.893563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:51.893563Z digest=sha256:0833643ec6bb829c80a8b7cb9d8f3a40f3703ba3e76675282ea4a5b02503f020

Pith citing papers

No inbound Pith citation observations are available.