Pith. sign in

Paper Citation Record · LEDGER

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems

As of 15 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 0 inbound Pith citation observations for arXiv:2501.06554.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.06554 v1

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:02:18.838862Z

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

25 of 25 outbound references displayed

  • verified exact3
  • verified fuzzy15
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation eadc3631-0158-434b-89ca-ae7084e7a4f3 · outbound

This paper cites The Option-Critic Architecture.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems The Option-Critic Architecture

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:18.674067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:02:18.674067Z digest=sha256:1ae0dfc7f3d0f9d608c3fd2d3d241749f4b6bd14d96a02435423bc697424aaae

Observation 0fea0433-c780-4c95-9865-f9908bac73d3 · outbound

This paper cites Option-Critic in Cooperative Multi-agent Systems.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Option-Critic in Cooperative Multi-agent Systems

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:18.679710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:02:18.679710Z digest=sha256:8697b311b938685f5b6ad44e7767ef012cc68ca9a4dbcc81bad8eadf6e6b05ab

Observation f5ecb7a5-0afa-4c58-a1d6-afedcb1734fd · outbound

This paper cites Attention Option-Critic.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Attention Option-Critic

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:18.684622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:02:18.684622Z digest=sha256:87f5106c235561591a049f128ad4ec09d91616550fdba74ebedd3faa546a4e3c

Observation 95d6cceb-71fc-4029-b7cc-d00353e345d1 · outbound

This paper cites Dynamic planning in open-ended dialogue using reinforcement learning, 2022.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Dynamic planning in open-ended dialogue using reinforcement learning, 2022

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.303210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.694553Z digest=sha256:ea055dade5d08df5d804ea49a9a4345cf44f1348c07514b4d533d8570f968103

Observation 6c0a99da-a5ad-4ac9-a9e3-1e271965f52d · outbound

This paper cites Handbook on agent-oriented design processes.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Handbook on agent-oriented design processes

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.277212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.701957Z digest=sha256:bc49d0c94260a1ec34a79695a3b6811ccf8a295a6f900e326ba1516418551219

Observation 083f9d7f-3773-4241-bcaf-0371ef7eb66f · outbound

This paper cites Multi-agent deep reinforcement learning: a survey.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Multi-agent deep reinforcement learning: a survey

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.258403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.707714Z digest=sha256:dc3a6a7248845ea4bfc5bdecdb260668e9d4b23c0c1b792bfcc0193c9ec8174f

Observation c158bcc2-96ad-413d-a973-cabd98069dd6 · outbound

This paper cites Two-sided matching with firms' complementary preferences.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Two-sided matching with firms' complementary preferences

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-10T21:02:18.959877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.713199Z digest=sha256:ca134318766081554b76e2c1f4739485721403e0c9168d2ef1f07d745bf799f1

Observation b583ae01-4be9-4f64-a809-2756a559dd15 · outbound

This paper cites Emergence of division of labour in halictine bees: contributions of social interactions and behavioural variance.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Emergence of division of labour in halictine bees: contributions of social interactions and behavioural variance

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.235230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.722528Z digest=sha256:e562cb1aee9ec76082c48df7f9fa200d0eddd9eaf15c4b25f1b20509aa320a2b

Observation 83cdcd13-c859-4f58-88fb-bdaa6365972d · outbound

This paper cites Multi-agent deep reinforcement learning with type-based hierarchical group communication.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Multi-agent deep reinforcement learning with type-based hierarchical group communication

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.215155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.729786Z digest=sha256:1b1cf72696e26d4a29c90fba6065cdee0f1ac01caed805b57527f1c90146679a

Observation 090d9cbc-c6f9-4f6b-a59b-da3d7fd0a000 · outbound

This paper cites Multi-agent reinforcement learning as a rehearsal for decentralized planning.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Multi-agent reinforcement learning as a rehearsal for decentralized planning

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.190680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.737992Z digest=sha256:4179c07b371a13d2051ed00c5c1de89e64eb873b26e853bd730cf78a9155cc7c

Observation 431133cc-3620-43dd-a3e7-2becf05f7e7b · outbound

This paper cites Reinforcement learning-based joint user pairing and power allocation in mimo-noma systems.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Reinforcement learning-based joint user pairing and power allocation in mimo-noma systems

Reference 11

Resolution
verified exact
doi, observed 2026-08-10T21:02:18.918637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.744243Z digest=sha256:25b068007601586d36bfe781d0fee79764839e5813d5972a3b13b0bb52194179

Observation e289dd3b-78f5-4d0e-9483-30ed71682821 · outbound

This paper cites Role-based modeling for designing agent behavior in self-organizing multi-agent systems.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Role-based modeling for designing agent behavior in self-organizing multi-agent systems

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.168675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.748639Z digest=sha256:7cc969ac8ef74eef0f1ce6ec45a7330092b7fbb7db7ccb7c6c193e5419a322fa

Observation b0aa34ba-ace1-4d61-bd3e-c2063dff96d8 · outbound

This paper cites Jordan, and Zhuoran Yang.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Jordan, and Zhuoran Yang

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.151523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.754362Z digest=sha256:9893112e93372e383973dc2811bb2a98162419f9f1e9334206ec8f7760006fbd

Observation c1121493-a712-4bd5-9e3b-a86a99abd94b · outbound

This paper cites Optimal and approximate q-value functions for decentralized pomdps.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Optimal and approximate q-value functions for decentralized pomdps

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.135232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.760712Z digest=sha256:777bd02011f901d17662ce62abe09253bb8ceeeb0a93b82f7474d1280c2f0a70

Observation a4e77e8b-586f-4c33-a74f-83ea0aac2ac0 · outbound

This paper cites Hierarchical reinforcement learning: A comprehensive survey.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Hierarchical reinforcement learning: A comprehensive survey

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:18.768939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:02:18.768939Z digest=sha256:369541f1a37a1f972fdd38c67d87173334528436a381edfda600878303efc5b8

Observation 58fb5ca5-523b-4bcc-8d02-a0972aa9a9da · outbound

This paper cites Vast: Value function factorization with variable agent sub-teams.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Vast: Value function factorization with variable agent sub-teams

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:18.776671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:02:18.776671Z digest=sha256:c2c52e6be2e2ac45c83c167425e2ea0fb8ef1a34c43fceed5d44830048843913

Observation ba9c940b-af4e-41bf-b55a-d52e52d2de95 · outbound

This paper cites Advances in neural information processing systems 17: proceedings of the 2004 conference, volume 17.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Advances in neural information processing systems 17: proceedings of the 2004 conference, volume 17

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.103251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.784316Z digest=sha256:1e96216e84b62dad6dd207dbbe06931ab8add14ed54b1b23bfa02b6a9a39d9bd

Observation 986987ca-1a48-4a50-810a-63112b531ac5 · outbound

This paper cites Sutton, Doina Precup, and Satinder Singh.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Sutton, Doina Precup, and Satinder Singh

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:18.796890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:02:18.796890Z digest=sha256:ee6f4f0aeb14c007e5d5bc901213d9845822150e0dee7963b3b497a4d8d912fd

Observation c42a5f96-d91a-4728-b5e7-398ca9a73480 · outbound

This paper cites Effectiveness of gamified team competition as mhealth intervention for medical interns: a cluster micro-randomized trial.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Effectiveness of gamified team competition as mhealth intervention for medical interns: a cluster micro-randomized trial

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.084218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.802360Z digest=sha256:b96f51eb149d086dab24499fae0fe0286fac4b5df6f6412c843d8f47c1afb54a

Observation 25a0eddb-fe5e-48dd-b9dd-891f822cd454 · outbound

This paper cites Beaulieu, and Y.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Beaulieu, and Y

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.069932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.808135Z digest=sha256:cad156115d6282a52659449de1e2a409483a1327106c018ca6e1375d37a83626

Observation d7f37294-37b6-44e2-a88e-e1eb4c05c8e9 · outbound

This paper cites Adaptive dynamic bipartite graph matching: A reinforcement learning approach.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Adaptive dynamic bipartite graph matching: A reinforcement learning approach

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.055682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.813132Z digest=sha256:8fa5c107ef6b8ce0027ccd0c3e54b958867f72fcfbbcebfc183620fa5be0302f

Observation aa678bbc-d480-441f-a8c5-4eff6275f472 · outbound

This paper cites Hierarchical dominance structure and social organization in african elephants, loxodonta africana.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Hierarchical dominance structure and social organization in african elephants, loxodonta africana

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.036585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.821021Z digest=sha256:723d661aa792cae54c6cef74093af935fb90a8d4dfd6d861ff4e34a30208464d

Observation e7e48205-37c8-4852-b131-6c3c71240ba8 · outbound

This paper cites Large-scale order dispatch in on-demand ride-hailing platforms: A learning and planning approach.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Large-scale order dispatch in on-demand ride-hailing platforms: A learning and planning approach

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.022090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.826218Z digest=sha256:7179722e5ba7e3bdb8965576d1ba634eb192c2f939f79ce382260a4f2410ad6d

Observation 8e214a0b-6254-44c9-a47d-0ce9e1dca286 · outbound

This paper cites Deep Sets.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Deep Sets

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:18.831276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:02:18.831276Z digest=sha256:cedceda893a47ecea6663b046c45520f854c1f9f2371d326f24022917696ae22

Observation f0d3cc4c-a014-4406-b5e2-c81ee647b00b · outbound

This paper cites Reinforcement learning based local search for grouping problems.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Reinforcement learning based local search for grouping problems

Reference 25

Resolution
verified exact
doi, observed 2026-08-10T21:02:18.895206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.838862Z digest=sha256:9913e2dbea01e62dee91fb61400e2f061200c2e7bbdf222837c71beca3746dc4

Pith citing papers

No inbound Pith citation observations are available.