Pith. sign in

Paper Citation Record · LEDGER

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning

As of 9 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 0 inbound Pith citation observations for arXiv:2506.08507.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.08507 v2

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:15:46.397717Z

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

47 of 47 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved42
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b2306a7d-840b-4c49-9282-2b69f62b89d4 · outbound

This paper cites GPT-4 Technical Report.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.177624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.177624Z digest=sha256:8c58b8ca5c1dc9db6cacd3daa8948d1a506370a4a038a426b2aebd77db1bb50e

Observation 17845117-7d98-45f7-9fa7-a0bcced6e1db · outbound

This paper cites Program Synthesis with Large Language Models.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Program Synthesis with Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.183179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.183179Z digest=sha256:f15c821b9c3345b707197f2a6d5b65f5d7639d72b4b9f9cad07066ebae9c74af

Observation 914c045a-0591-4866-a91a-0a4fc1262488 · outbound

This paper cites AutoAgents: A Framework for Automatic Agent Generation.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning AutoAgents: A Framework for Automatic Agent Generation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.188579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.188579Z digest=sha256:63d22e884cd803411f5d6ea066a201bf30a0c0493cfab3997b170c510dfaf35f

Observation 99538ea8-fa36-419e-9393-96eab618fd42 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Evaluating Large Language Models Trained on Code

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.193285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.193285Z digest=sha256:e960f8feb27ad648568a336d111ccbd851da0fe628bb76980aba6ce5d1ee1ac3

Observation 8fa7a4e5-96c7-4595-8a94-4e6f5be402b6 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Training Verifiers to Solve Math Word Problems

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.198600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.198600Z digest=sha256:9556347ccc0b4fa5fa17a54bedacb24355150f9dfd18207e50c80861bd9805fe

Observation fafd95a7-c04b-42c5-82f5-d63377ef481b · outbound

This paper cites Flow-DPO: Improving LLM Mathematical Reasoning through Online Multi-Agent Learning.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Flow-DPO: Improving LLM Mathematical Reasoning through Online Multi-Agent Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.203453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.203453Z digest=sha256:0289f62727ec6a7748ba305887e41c7367d6384184ba1b67c7e938296a6da747

Observation 46752285-6f86-4489-912c-1556515930a9 · outbound

This paper cites Improving factuality and reasoning in language models through multiagent debate.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Improving factuality and reasoning in language models through multiagent debate

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.209026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.209026Z digest=sha256:0298a8f71050005417fa84a2aa53d1fcd3f3c97fbacb3796369172811d796079

Observation 9cb85eeb-2b74-47bd-b360-0339b3ea2e5c · outbound

This paper cites Large Language Model based Multi-Agents: A Survey of Progress and Challenges.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Large Language Model based Multi-Agents: A Survey of Progress and Challenges

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.213954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.213954Z digest=sha256:8ade5f80379047f9632d049e591597e660e67248606922faf7512b344807efcb

Observation 06173c9d-0bef-4e06-913b-e9990739c465 · outbound

This paper cites Measuring Massive Multitask Language Understanding.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Measuring Massive Multitask Language Understanding

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.218245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.218245Z digest=sha256:4e4a8af80861d97140646498afe49aaeb7c4159d97aeb712355d77f6f346404a

Observation a5e52074-8336-4651-9be7-688ad3dff401 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Measuring Mathematical Problem Solving With the MATH Dataset

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.222738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.222738Z digest=sha256:2a44e90c6dbb6e6af636ce8b65fa5bd7f1e1e4d40cbfc1b4c326d3b3fa8091c1

Observation caafe439-f005-4cf5-97c2-fa0c359d0b74 · outbound

This paper cites MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.227271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.227271Z digest=sha256:dd3a48879817329a5487a761515728cfaa7b74246d8183f4b5315363d30259c0

Observation 08fae210-a4c8-4711-89fe-2efa0b18fd26 · outbound

This paper cites Automated Design of Agentic Systems.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Automated Design of Agentic Systems

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.231681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.231681Z digest=sha256:f028db2bc669821634c74769d7bf7684128bef594d63f244128b2177db933e70

Observation 9c6a1a9d-baff-4b6c-90ee-1f66fe0d349d · outbound

This paper cites Self-Organized Agents: A LLM Multi-Agent Framework toward Ultra Large-Scale Code Generation and Optimization.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Self-Organized Agents: A LLM Multi-Agent Framework toward Ultra Large-Scale Code Generation and Optimization

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.236177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.236177Z digest=sha256:d249cff60993fd7a88ac5619c9ddfd32d78c31a4eaf06fab9ab70c5c6b63bb8b

Observation 438e905b-3bd6-48c5-ba3c-0303f6b86fcd · outbound

This paper cites Reinforcement learning: A survey.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Reinforcement learning: A survey

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.241395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.241395Z digest=sha256:0cbb347f2b6ba0991164443bf2b01e44028d5d2728d810387728906b32039e42

Observation 77742588-3936-4750-8ead-8178dd7e8acd · outbound

This paper cites The Dawn of Natural Language to SQL: Are We Fully Ready?.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning The Dawn of Natural Language to SQL: Are We Fully Ready?

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.246281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.246281Z digest=sha256:a68e718eda61c72ff54895cbe5e490923e4f2583b38c71e69ebc3327b5ee53a8

Observation 785a8e32-aa35-43f0-9e6e-5a24d3735d4d · outbound

This paper cites CodeTree: Agent-guided Tree Search for Code Generation with Large Language Models.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning CodeTree: Agent-guided Tree Search for Code Generation with Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.251381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.251381Z digest=sha256:be128e95d3e21d1f2cb9c311594659ddf7ed1027979bd7faa60b38025f7dd7e2

Observation dcecd16d-815b-490e-ac94-88ef76c87f6f · outbound

This paper cites Personal LLM Agents: Insights and Survey about the Capability, Efficiency and Security.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Personal LLM Agents: Insights and Survey about the Capability, Efficiency and Security

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.256745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.256745Z digest=sha256:f60e3ef491e09fb893e79e56c08a67ae26c5422f20e3e841d18d29ee6b6fb8eb

Observation 2891d2bc-6913-48a1-8b03-b4a1c52651e7 · outbound

This paper cites Deep Reinforcement Learning: An Overview.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Deep Reinforcement Learning: An Overview

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.261724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.261724Z digest=sha256:db447cde885118dba04c4c675018d47cc6d702912db4438f2885785c122ee57e

Observation 0a85782b-b7fd-4d4d-9b6b-8e4d485ddd4d · outbound

This paper cites A Dynamic LLM-Powered Agent Network for Task-Oriented Agent Collaboration.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning A Dynamic LLM-Powered Agent Network for Task-Oriented Agent Collaboration

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.266505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.266505Z digest=sha256:858c8be9429a91bd39abf3813569c760b4ac90395989fade79828fac79d300fc

Observation 574c3976-111a-4a8b-aba3-93900e025b94 · outbound

This paper cites Large Language Model Agent: A Survey on Methodology, Applications and Challenges.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Large Language Model Agent: A Survey on Methodology, Applications and Challenges

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.271094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.271094Z digest=sha256:cd8da801198b0c0cac023fac717f6b72880d0948514d4969e28e78028af7a450

Observation c6a7b146-9077-46f3-ac73-8b87b1a0b477 · outbound

This paper cites Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.275717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.275717Z digest=sha256:79be03a4ca522670f71789f8d77e7a0a7093a65954ed604391457cf887d228d2

Observation aaba72c0-e0e2-4604-b289-bb60c9465f40 · outbound

This paper cites Gpt-4o mini: Advancing cost-efficient intelligence.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Gpt-4o mini: Advancing cost-efficient intelligence

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:15:47.039262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:15:46.280402Z digest=sha256:f645b2dcbbeea4c2d3edd73b45ac07b211b83ae8fa9f024d75fac209794b3826

Observation 15f84a89-e88d-486c-b381-cf2b65eddfa9 · outbound

This paper cites Gpqa: A graduate-level google-proof q&a benchmark.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Gpqa: A graduate-level google-proof q&a benchmark

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.284657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.284657Z digest=sha256:3facf01acb8b8bf30df34601d77e6436486a3f9bc922f3c30c84c8cc4bfd66a6

Observation d0441e11-3d83-42d4-ae56-577172fe9880 · outbound

This paper cites Code Generation with AlphaCodium: From Prompt Engineering to Flow Engineering.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Code Generation with AlphaCodium: From Prompt Engineering to Flow Engineering

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.289026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.289026Z digest=sha256:1128ed74ce936678a29189eaf3d3e89464d03261bb6e0e121c9322d7d59a7e40

Observation 9c009297-d3c7-4dea-aa5c-d99d7fc736c5 · outbound

This paper cites Proximal Policy Optimization Algorithms.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.294794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.294794Z digest=sha256:397653ee2cd61990d127074e7de5d915a3f1dc233a7eb0a1ec0f9b6ca7c18398

Observation 1fbed30d-3835-4c7b-a58a-e29281af6da5 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.299563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.299563Z digest=sha256:1e1c6158d43ca044ecff2dacd17cf4f3e468de6815813361b986a8126c6b0fe5

Observation 88d880d4-c7b5-4a44-ba30-ab6ce57a7306 · outbound

This paper cites Reflexion: Language agents with verbal reinforcement learning.Advances in Neural Information Processing Systems, 36:8634–8652, 2023.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Reflexion: Language agents with verbal reinforcement learning.Advances in Neural Information Processing Systems, 36:8634–8652, 2023

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.304267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.304267Z digest=sha256:3738bcd16a7b7a1c9e87ffea1429a4300ff4fff40dbb49743da715d1cdd8e56f

Observation a7fc2418-0f88-4bb5-bb36-b00c7f139ffc · outbound

This paper cites Llm- planner: Few-shot grounded planning for embodied agents with large language models.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Llm- planner: Few-shot grounded planning for embodied agents with large language models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.308865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.308865Z digest=sha256:30c4da43a5a4c77b94b21d30ff94a8c1165e809cd4a10f6857dda6a80e9fb808

Observation 6dfa02b4-64d2-4034-907e-24d1ff83c84a · outbound

This paper cites LLM-based Multi-Agent Reinforcement Learning: Current and Future Directions.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning LLM-based Multi-Agent Reinforcement Learning: Current and Future Directions

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.313451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.313451Z digest=sha256:9d2b15960bde0ff1c54add2dd371555ac1d45ed8009849f103817cf2afefaeda

Observation 6d19bffc-3658-4fb8-8914-5bf61c01349e · outbound

This paper cites Self-Consistency Improves Chain of Thought Reasoning in Language Models.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Self-Consistency Improves Chain of Thought Reasoning in Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.318627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.318627Z digest=sha256:d51a8c1457244864f94f39da30f5a4784bb5faa3b73bdd14671a2be4e74ba2a2

Observation 6da77bcb-433c-4cf7-84c1-d7085f949c87 · outbound

This paper cites Unleashing the Emergent Cognitive Synergy in Large Language Models: A Task-Solving Agent through Multi-Persona Self-Collaboration.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Unleashing the Emergent Cognitive Synergy in Large Language Models: A Task-Solving Agent through Multi-Persona Self-Collaboration

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.323581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.323581Z digest=sha256:b77aab1a6c3e582bbd8589f86b5400fcd60693d1530c42c2ddbf769421aa1040

Observation d9caf22b-906b-4baf-a4c5-b97fb0ac96cc · outbound

This paper cites Chain-of-Table: Evolving Tables in the Reasoning Chain for Table Understanding.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Chain-of-Table: Evolving Tables in the Reasoning Chain for Table Understanding

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.328070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.328070Z digest=sha256:7da8b9bd88409d3cf909dcee6a394a0972ec421855b920ccc2940abcac3bfba4

Observation 1c87dbdb-1f92-420f-96f3-cf7ed6400928 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.Advances in neural information processing systems, 35:24824–24837, 2022.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Chain-of-thought prompting elicits reasoning in large language models.Advances in neural information processing systems, 35:24824–24837, 2022

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:15:46.996492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:15:46.332652Z digest=sha256:e27d7a2a42b8d689b21ce7a28e1e3b6dd32ed54c6caca16f32ebd88f793cd58d

Observation 1582b37b-1c27-4af3-88f0-31f8fff4eb9e · outbound

This paper cites AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.336932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.336932Z digest=sha256:4061daacf40d8e3dd8632baf8ed466afb7047ab978fc50f13f90724be50e3551

Observation 5fb0a280-7783-4f11-889f-dd962fea3173 · outbound

This paper cites TravelPlanner: A Benchmark for Real-World Planning with Language Agents.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.341641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.341641Z digest=sha256:ecf5aa089b9d5972e18a84cb005c0c4f932f228df980d2b6f44e8a5453d4f3a5

Observation a470c992-a3a4-497c-b6e5-8127f8093c98 · outbound

This paper cites Auto-GPT for Online Decision Making: Benchmarks and Additional Opinions.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Auto-GPT for Online Decision Making: Benchmarks and Additional Opinions

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.346163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.346163Z digest=sha256:b5aba8733317f2dd8dfbc43db4dd44f68a5ae2356ceb1df86d4d632c4fe27218

Observation 47d0df47-f349-4cc8-9a84-7a390d432d61 · outbound

This paper cites React: Synergizing reasoning and acting in language models.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning React: Synergizing reasoning and acting in language models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.351234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.351234Z digest=sha256:29654199b7fb81ff92a7d67e518d9393cffdc042a00285b9069a956618faf0b8

Observation e4a104ff-2598-47a9-8b26-713f04bac4b0 · outbound

This paper cites MAS-GPT: Training LLMs to Build LLM-based Multi-Agent Systems.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning MAS-GPT: Training LLMs to Build LLM-based Multi-Agent Systems

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.355522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.355522Z digest=sha256:6b54efcd6a6c5ba0511988c8a5190d331adcccfb7985edaf710a58c9a299dd63

Observation 81e18702-0380-4c05-9387-8926490a336e · outbound

This paper cites MasRouter: Learning to Route LLMs for Multi-Agent Systems.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning MasRouter: Learning to Route LLMs for Multi-Agent Systems

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.359847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.359847Z digest=sha256:764c70e340768127a2167391747370b04b4ef9c713fbf98e26308835c7777184

Observation 11de2d7d-f878-4b39-9c02-a7075fe1f524 · outbound

This paper cites TableGPT: Towards Unifying Tables, Nature Language and Commands into One GPT.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning TableGPT: Towards Unifying Tables, Nature Language and Commands into One GPT

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.364482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.364482Z digest=sha256:27b8e575080dd7cf3bab5935de1df2ad9dbf623635fca53927b92e2b178993d5

Observation caf2a6f2-e40a-4fc9-afb3-9ac3e194fb70 · outbound

This paper cites Multi-agent Architecture Search via Agentic Supernet.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Multi-agent Architecture Search via Agentic Supernet

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.369706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.369706Z digest=sha256:8b4172fe31d32a5f3955b6cb97ffa8e6fa79ac8e99f98d2f1e378444e71f38f3

Observation eab37f38-d52f-44c3-a0c7-5614cae9a874 · outbound

This paper cites G-Designer: Architecting Multi-agent Communication Topologies via Graph Neural Networks.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning G-Designer: Architecting Multi-agent Communication Topologies via Graph Neural Networks

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.374277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.374277Z digest=sha256:359282233f7574c10e960f0078e228b493a05983d2ecd776fc36502ca1a04cdb

Observation 1f7bc27b-54e8-4f01-9eb1-b81910fa998c · outbound

This paper cites AFlow: Automating Agentic Workflow Generation.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning AFlow: Automating Agentic Workflow Generation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.379327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.379327Z digest=sha256:297f23281a80f49e370602a19036dee382c4ecbb6858b3f55e2857bc6fa56d65

Observation 53735eb0-40f9-4b71-892a-1ac96110671b · outbound

This paper cites Achieving> 97% on gsm8k: Deeply understanding the problems makes llms perfect reasoners.arXiv e-prints, pages arXiv–2404, 2024.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Achieving> 97% on gsm8k: Deeply understanding the problems makes llms perfect reasoners.arXiv e-prints, pages arXiv–2404, 2024

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:15:46.972608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:15:46.383758Z digest=sha256:609f1d9fe011eb45a83f56e81547a9c6aa64d39d7eda830763a83983be8555f6

Observation 6bd22a3a-090e-4673-aff8-1cbbccc2d43c · outbound

This paper cites Star-agents: Automatic data optimization with llm agents for instruction tuning.Advances in Neural Information Processing Systems, 37:4575–4597, 2024.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Star-agents: Automatic data optimization with llm agents for instruction tuning.Advances in Neural Information Processing Systems, 37:4575–4597, 2024

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:15:46.958212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:15:46.388287Z digest=sha256:5ac1ad9dcfabb4a90d63c211877f0b4927f0ea021133dcb6c59f84b21d3850e3

Observation ad330fb1-de3a-42de-910c-5564b69fdb7c · outbound

This paper cites Are Large Language Models Good Statisticians?.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Are Large Language Models Good Statisticians?

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.393258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.393258Z digest=sha256:53d6643b6a4eddce965467d52a67df179e0667ddd128d2386fa342eff4cd288c

Observation 754a0e84-7138-4a4d-862d-c7837060706f · outbound

This paper cites Gptswarm: Language agents as optimizable graphs.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Gptswarm: Language agents as optimizable graphs

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:15:46.942530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:15:46.397717Z digest=sha256:c25b726f64afd0e30720902fea9d10b2f124809fc875b3def48f0e38da756623

Pith citing papers

No inbound Pith citation observations are available.