Pith. sign in

Paper Citation Record · LEDGER

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning

As of 8 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2507.10251.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.10251 v1

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:42:31.969789Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

31 of 31 outbound references displayed

  • verified exact3
  • verified fuzzy11
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c25dabca-4b28-44ef-aca6-fc5d86beb1ce · outbound

This paper cites Amato, G.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Amato, G

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:42:32.263964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.900918Z digest=sha256:315ee832530be2dd6fb4e0ef97ffd81e69de325b1ea24e6ea72979ae48025042

Observation dd50fff5-d071-4772-8ada-085108eb3937 · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.256851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.903594Z digest=sha256:4ecddee7234a541f1111a6a902eb5c84cf6c2fd5c054dcf81f4b887b34613070

Observation 254b6afe-a17d-4c24-92fe-8eb5fd9e4f05 · outbound

This paper cites Christopher, D.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Christopher, D

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:42:32.249983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.905919Z digest=sha256:74292fd364be3e7c682b04fb54f963e70619b93709b9ddb0cba4a49c1ef849b9

Observation 52688c32-0c2d-4039-9037-fdbf6c6ac072 · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.243323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.908315Z digest=sha256:76c0c8dfc35682594ad625c59d6d719cf107dc8eadf4820ccc89daf7c2838714

Observation 87ac3d3e-fd49-4e89-ab9b-363cda374d13 · outbound

This paper cites A Hitchhiker's Guide to Statistical Comparisons of Reinforcement Learning Algorithms.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning A Hitchhiker's Guide to Statistical Comparisons of Reinforcement Learning Algorithms

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T17:42:31.910654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:42:31.910654Z digest=sha256:74013c62024f0082470015b6868e8656350df5716008e5c0a6dd2977b160b413

Observation 2b01d831-0594-4a3e-9116-86fce109ecf0 · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.236904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.913173Z digest=sha256:305265ea62fefb0e336ad8a79ff3d3369ec696ffe67a4260bfa7115ccc7948e2

Observation 4e0fe3d0-cf6a-485d-8a92-5a6d53d0bd10 · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.230238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.915660Z digest=sha256:61f0df5cd557006ef344429214aa0cb6d28c470396622da8872f8435cac7d09f

Observation 13632030-4f3a-4502-8f6c-3f8f5480527f · outbound

This paper cites Gupta, A.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Gupta, A

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:42:32.223213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.917968Z digest=sha256:57fe0101c47a38a581006be319f115bc143ee294f795eb41d2b49cabaac8f0e0

Observation 8f48adb4-8852-4c19-8d82-0f2ad4709543 · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.216227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.920152Z digest=sha256:699e8bc355fc81ff730cda8d9e6b3344aa6b9a5d86f82fe144c233e50582c97e

Observation 790eb0c5-57ee-453e-830b-a88b36812a67 · outbound

This paper cites Liu and J.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Liu and J

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:42:32.209501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.922269Z digest=sha256:742a2a314a3c6d97bdc52bb67cf5ee401e464a62f07f5142ed8056b92f213ef5

Observation ed8414c3-967e-4153-93c1-d1cedc9678b3 · outbound

This paper cites Mahajan, T.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Mahajan, T

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:42:32.202893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.924564Z digest=sha256:b8f1e5a46a9cd7cc6c6b5ec116d5609b09c841f9afd07cf2dc2d40506385608f

Observation 5a1119d7-cb7c-43da-96e7-f94fcee32f4e · outbound

This paper cites Marchesini, Y.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Marchesini, Y

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:42:32.195768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.926863Z digest=sha256:5be61b78e55c96459d696fe4397cd558d253998ad63a1d873420b46e50d49043

Observation 67c0c2a2-3264-4785-9487-05e1d6858dab · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.189160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.929012Z digest=sha256:e5c4f5f593cb4d87f510b706050681472052b8a2208b19eb944dac6970faa52f

Observation 82f40597-9b22-4d56-8372-52b0f7ac358f · outbound

This paper cites Omidshafiei, A.-A.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Omidshafiei, A.-A

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:42:32.182632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.931025Z digest=sha256:5aa9b8fc39b6667c77fda9677d5a1bf8f707408d018aa27d85d80fa742811468

Observation 49742869-cd09-48bf-a46f-e9c9af372feb · outbound

This paper cites Rashid, G.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Rashid, G

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:42:32.175255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.933128Z digest=sha256:e184ac959276d0079e47eb4dbf7e2dc317cbd4a515a88e7151e4f7ed451a9378

Observation 4a7f6d14-3b7f-4e10-9d34-2c72b9679b40 · outbound

This paper cites Rashid, M.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Rashid, M

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:42:32.167625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.935270Z digest=sha256:4aeac3d4268e6e9bf05f93296888808279743a4e537f9b4e30095a532384c27a

Observation 213fe555-0bf7-4a43-9fdf-ed13993412af · outbound

This paper cites Improved Memory-Bounded Dynamic Programming for Decentralized POMDPs.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Improved Memory-Bounded Dynamic Programming for Decentralized POMDPs

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:42:32.078285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.937356Z digest=sha256:da9ce3f9379cd6b3ed18a91ebba07d8e6ff2c8930c84aea9e2f2873b9831c833

Observation f70f9ac5-1bf7-4c74-89bb-bcd46b022d46 · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.159833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.939852Z digest=sha256:eb501efe81c636f99620feada179bd3776a835ab2980f45cc4ec9e4a050c547d

Observation 663f352d-5a3c-455a-8d73-f8f2d404acd0 · outbound

This paper cites Tuyls and G.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Tuyls and G

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:42:32.152705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.941942Z digest=sha256:59cd730b0edad5101cdf1ebe0b556e7486808ad320b084c22bd0ffc0b7c5b48e

Observation c27bcdf6-02d5-44d5-af66-852287f41051 · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 20

Resolution
verified exact
raw_fallback, observed 2026-08-06T17:42:32.067921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.943955Z digest=sha256:3c2c15e720c11c41f1a9b0c9c8f60c9236720d965d62f4981db5acdff4add06c

Observation 37a69add-050c-4d38-a5dd-7115c7cca8c4 · outbound

This paper cites QPLEX: Duplex Dueling Multi-Agent Q-Learning.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning QPLEX: Duplex Dueling Multi-Agent Q-Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T17:42:31.946134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:42:31.946134Z digest=sha256:ce263c285a5a4507a389586bae9f8c3ef523873595e63a929b4a62e084776e2f

Observation 8eaf9bc8-6a94-4169-b974-c3bc0d905bc9 · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.145781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.948359Z digest=sha256:99a232f98f22241ea6255c7f08a0165a2c71b5b39a344df6121ee8af72368b43

Observation 7adf4894-07da-4228-8204-e490ebb307f9 · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.139040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.950455Z digest=sha256:fc53e59b1c3023d380d78e197179bdfa1109a9af8bcdcad63ee70aacbac41d80

Observation 306e5843-c729-450b-854d-c68a2f77800e · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.131597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.952475Z digest=sha256:3497211174bc0a29f3093f7e6eaa2044b88e92ad53deb4dc8fbfd2277788d55c

Observation f4077a1d-1ad6-48a4-a211-96b6b5ffbbe6 · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.123834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.954837Z digest=sha256:230a81fd2d7a910e04cc07b7e80a65895e16681e167e10d8dd9215bab5750170

Observation 2d14c30f-4a84-43d3-a9f9-fb1a9a123e85 · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.116283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.957866Z digest=sha256:0170ca6e3df83ac5dcf363127a5704387bc8d1c75a172ea15cb67865c83b75bb

Observation 250ed19e-5f67-42e3-af7c-257086c5631a · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.108364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.960123Z digest=sha256:5c8a41221a5c1016fa6df38a65e94d908a73971ad75f35397cff483aa309a632

Observation ccc3942c-aaf4-42c8-af05-541c2d1dae88 · outbound

This paper cites Qatten: A General Framework for Cooperative Multiagent Reinforcement Learning.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Qatten: A General Framework for Cooperative Multiagent Reinforcement Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T17:42:31.962434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:42:31.962434Z digest=sha256:97c12b68890dab26a71ceb1d00e2a2b4a308a3b53a62b8adf70c63738328011a

Observation df4d22d7-557a-4d19-8be0-0ee404368ebc · outbound

This paper cites an unresolved cited work.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:42:32.101132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.964784Z digest=sha256:1486f5aa4aa59ddbee493633c54ab2d3695d557129407ce5f25ca40790422141

Observation d79e2eba-811b-4f73-8619-f2cc41268c39 · outbound

This paper cites Hierarchical Reinforcement Learning for Multi-agent MOBA Game.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning Hierarchical Reinforcement Learning for Multi-agent MOBA Game

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:42:31.995475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.967163Z digest=sha256:631ac72706442cdf426727c8a533f2399c2155d6024da793737cac6356788a48

Observation 5e005ef1-12af-435f-aa95-f5422e99adac · outbound

This paper cites go to tomato.

ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning go to tomato

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:42:32.093800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:42:31.969789Z digest=sha256:2d46d0b4697a70aea434eab2cefee3397b2b3518eee4035f440b15daea680a14

Pith citing papers

No inbound Pith citation observations are available.