Pith. sign in

Paper Citation Record · LEDGER

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning

As of 19 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2507.10251.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.10251 v1

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:42:31.969789Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

31 of 31 outbound references displayed

  • verified exact3
  • verified fuzzy11
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c25dabca-4b28-44ef-aca6-fc5d86beb1ce · outbound

This paper cites Amato, G.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Amato, G

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:42:32.263964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.900918Z digest=sha256:0c5b1ed52064a0729806d044c7c1bf05ec885587d257ddfc2237aa250c9a997f

Observation dd50fff5-d071-4772-8ada-085108eb3937 · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.256851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.903594Z digest=sha256:838e1fdf1566bb12c2f782690dee746ef88f6630faeb62cbd6daaaa747585b10

Observation 254b6afe-a17d-4c24-92fe-8eb5fd9e4f05 · outbound

This paper cites Christopher, D.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Christopher, D

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:42:32.249983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.905919Z digest=sha256:e261198202e57f2bea4beffe4410fb187f95ef8b8ca00abf974c10a9b450c3fa

Observation 52688c32-0c2d-4039-9037-fdbf6c6ac072 · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.243323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.908315Z digest=sha256:4ea0dfe86c7af2d58262a5dce5a18a99db2a0f1390af71533e3dd7cbfc0cc3db

Observation 87ac3d3e-fd49-4e89-ab9b-363cda374d13 · outbound

This paper cites A Hitchhiker's Guide to Statistical Comparisons of Reinforcement Learning Algorithms.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning A Hitchhiker's Guide to Statistical Comparisons of Reinforcement Learning Algorithms

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T17:42:31.910654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:42:31.910654Z digest=sha256:cd7f8f6206a80f453e921573a6ff7f6a41bce6a9242e345e57a1195c425f19c0

Observation 2b01d831-0594-4a3e-9116-86fce109ecf0 · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.236904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.913173Z digest=sha256:70d8855d6ca26765b9251ef1c661a8f1c33c6acb624eebc4bbd19475199988ea

Observation 4e0fe3d0-cf6a-485d-8a92-5a6d53d0bd10 · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.230238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.915660Z digest=sha256:cc47342afedc830afb16bd6fec439b78e2fe919df0e684fa56b1c82f5ea6030a

Observation 13632030-4f3a-4502-8f6c-3f8f5480527f · outbound

This paper cites Gupta, A.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Gupta, A

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:42:32.223213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.917968Z digest=sha256:f1f4b3ddefd47b4d6e5714f3c7cd6d9307e14f022068569c9068342504e98180

Observation 8f48adb4-8852-4c19-8d82-0f2ad4709543 · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.216227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.920152Z digest=sha256:3193adee828eab73ae2b37f6f6f9af99e9ece6d0e42e786da8ee7a10309d25a5

Observation 790eb0c5-57ee-453e-830b-a88b36812a67 · outbound

This paper cites Liu and J.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Liu and J

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:42:32.209501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.922269Z digest=sha256:7e6be4906b0f1aaa3034830d545f7a154e48f3ea9d025b5eb88a4ef8fe572028

Observation ed8414c3-967e-4153-93c1-d1cedc9678b3 · outbound

This paper cites Mahajan, T.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Mahajan, T

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:42:32.202893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.924564Z digest=sha256:ca267f9371381db5b9369179bc5649011f42bfdc3b987305f10da5df6b8814df

Observation 5a1119d7-cb7c-43da-96e7-f94fcee32f4e · outbound

This paper cites Marchesini, Y.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Marchesini, Y

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:42:32.195768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.926863Z digest=sha256:283dadd6bfec5b46736ac55497955a0ed2dc951763cb5f9f4a15e69ee4253ec0

Observation 67c0c2a2-3264-4785-9487-05e1d6858dab · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.189160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.929012Z digest=sha256:6644d5ab8bc82ab2a310b2de1e958bb0bb1a001dc89c4bd43b91b3d80e444ddf

Observation 82f40597-9b22-4d56-8372-52b0f7ac358f · outbound

This paper cites Omidshafiei, A.-A.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Omidshafiei, A.-A

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:42:32.182632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.931025Z digest=sha256:67ab8f6e2b3b159161618da391250d7f2d54b505e100aa7dc8655dae6087cb2d

Observation 49742869-cd09-48bf-a46f-e9c9af372feb · outbound

This paper cites Rashid, G.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Rashid, G

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:42:32.175255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.933128Z digest=sha256:97f751b0f44bf5185b0e87cbd5896d21ca5b621a92a04a49661c4c8ec2a0cbd4

Observation 4a7f6d14-3b7f-4e10-9d34-2c72b9679b40 · outbound

This paper cites Rashid, M.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Rashid, M

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:42:32.167625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.935270Z digest=sha256:e70678bef927b414c50f5f45ecae984b6dd0467a2f9400f0b6246dbdd627c75e

Observation 213fe555-0bf7-4a43-9fdf-ed13993412af · outbound

This paper cites Improved Memory-Bounded Dynamic Programming for Decentralized POMDPs.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Improved Memory-Bounded Dynamic Programming for Decentralized POMDPs

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:42:32.078285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.937356Z digest=sha256:eead57b233948b7d8326897a762525918271b39011b12ca6d4615592f54bbdae

Observation f70f9ac5-1bf7-4c74-89bb-bcd46b022d46 · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.159833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.939852Z digest=sha256:5e7f857b8a4c99dfc36564f27969ee38ae4feb8552c15a4fc042bebab80190b1

Observation 663f352d-5a3c-455a-8d73-f8f2d404acd0 · outbound

This paper cites Tuyls and G.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Tuyls and G

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:42:32.152705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.941942Z digest=sha256:31916d947262b95374c39090f85a00d4c145841d8d16c0defd1fd9757aed5716

Observation c27bcdf6-02d5-44d5-af66-852287f41051 · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 20

Resolution
verified exact
raw_fallback, observed 2026-08-06T17:42:32.067921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.943955Z digest=sha256:3d77334043e13a73a10c638dcf3b2e1a5c6cac59d231377d29a5c6d90823760d

Observation 37a69add-050c-4d38-a5dd-7115c7cca8c4 · outbound

This paper cites QPLEX: Duplex Dueling Multi-Agent Q-Learning.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning QPLEX: Duplex Dueling Multi-Agent Q-Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T17:42:31.946134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:42:31.946134Z digest=sha256:0abfa1d768c97f11fe92ed7d7e10423a4f0fe951db4e489d698fde84ad346351

Observation 8eaf9bc8-6a94-4169-b974-c3bc0d905bc9 · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.145781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.948359Z digest=sha256:c9302c8b31b679caf5c0f4e3170b441041ce852a21e54fa153e3f55578321e26

Observation 7adf4894-07da-4228-8204-e490ebb307f9 · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.139040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.950455Z digest=sha256:f406a3352f078c6c41678d3bf7a7f9e7176e25b3a0493e87ed994711d54c71a1

Observation 306e5843-c729-450b-854d-c68a2f77800e · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.131597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.952475Z digest=sha256:36eca7b11747720a2c4875c29fc03be90311d9d869755bf49c5756f5849cca97

Observation f4077a1d-1ad6-48a4-a211-96b6b5ffbbe6 · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.123834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.954837Z digest=sha256:8b3f598952d735b001a561474eddda7fd718d3a200740ad46ea158a911e2cfec

Observation 2d14c30f-4a84-43d3-a9f9-fb1a9a123e85 · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.116283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.957866Z digest=sha256:a71518bc0b963be6f5b8ba533c521f0bc0a90404d689fc4729e2561c07ed737d

Observation 250ed19e-5f67-42e3-af7c-257086c5631a · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.108364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.960123Z digest=sha256:0990ecc65d9d6b55114ba7b3d4e2050305b89eda6ad55caee2b89528630057da

Observation ccc3942c-aaf4-42c8-af05-541c2d1dae88 · outbound

This paper cites Qatten: A General Framework for Cooperative Multiagent Reinforcement Learning.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Qatten: A General Framework for Cooperative Multiagent Reinforcement Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T17:42:31.962434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:42:31.962434Z digest=sha256:6b97776bc85d152e2bd11f82e4a0247f403d3a9ff4efe1fbf01fefb707a28cbb

Observation df4d22d7-557a-4d19-8be0-0ee404368ebc · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.101132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.964784Z digest=sha256:d9995d292687e0e64b38790d0feef6fcd2c2246ef9217a67805427095537e75f

Observation d79e2eba-811b-4f73-8619-f2cc41268c39 · outbound

This paper cites Hierarchical Reinforcement Learning for Multi-agent MOBA Game.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Hierarchical Reinforcement Learning for Multi-agent MOBA Game

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:42:31.995475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.967163Z digest=sha256:b2c562a3b7ba9297c43eb92fb3b895cb75bc6518b81fa9dfbf7e6948ab3a16b7

Observation 5e005ef1-12af-435f-aa95-f5422e99adac · outbound

This paper cites go to tomato.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning go to tomato

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:42:32.093800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T17:42:31.969789Z digest=sha256:b3af12afb5e2d2b986bbd796a397ad723e630a5843a4ef0358e1f013da7974e3

Pith citing papers

No inbound Pith citation observations are available.