Pith. sign in

Paper Citation Record · LEDGER

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems

As of 15 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 0 inbound Pith citation observations for arXiv:2501.06554.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.06554 v1

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:02:18.838862Z

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

25 of 25 outbound references displayed

  • verified exact3
  • verified fuzzy15
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation eadc3631-0158-434b-89ca-ae7084e7a4f3 · outbound

This paper cites The Option-Critic Architecture.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems The Option-Critic Architecture

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:18.674067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:02:18.674067Z digest=sha256:1ae0dfc7f3d0f9d608c3fd2d3d241749f4b6bd14d96a02435423bc697424aaae

Observation 0fea0433-c780-4c95-9865-f9908bac73d3 · outbound

This paper cites Option-Critic in Cooperative Multi-agent Systems.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Option-Critic in Cooperative Multi-agent Systems

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:18.679710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:02:18.679710Z digest=sha256:8697b311b938685f5b6ad44e7767ef012cc68ca9a4dbcc81bad8eadf6e6b05ab

Observation f5ecb7a5-0afa-4c58-a1d6-afedcb1734fd · outbound

This paper cites Attention Option-Critic.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Attention Option-Critic

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:18.684622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:02:18.684622Z digest=sha256:87f5106c235561591a049f128ad4ec09d91616550fdba74ebedd3faa546a4e3c

Observation 95d6cceb-71fc-4029-b7cc-d00353e345d1 · outbound

This paper cites Dynamic planning in open-ended dialogue using reinforcement learning, 2022.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Dynamic planning in open-ended dialogue using reinforcement learning, 2022

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.303210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.694553Z digest=sha256:c94cbebc7eba98c9bf35eb8f322c59545d8c2a09b1082dbbef8a8fa5ecf4bed8

Observation 6c0a99da-a5ad-4ac9-a9e3-1e271965f52d · outbound

This paper cites Handbook on agent-oriented design processes.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Handbook on agent-oriented design processes

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.277212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.701957Z digest=sha256:a28797a7a8aa4caa9d6f85d155b67c539d073ee6ba3555e76d563927eedddb53

Observation 083f9d7f-3773-4241-bcaf-0371ef7eb66f · outbound

This paper cites Multi-agent deep reinforcement learning: a survey.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Multi-agent deep reinforcement learning: a survey

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.258403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.707714Z digest=sha256:dfbadb2157ae3c35c8e38f0619c33ed217e442bded77c92e9dc88f6ed199c800

Observation c158bcc2-96ad-413d-a973-cabd98069dd6 · outbound

This paper cites Two-sided matching with firms' complementary preferences.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Two-sided matching with firms' complementary preferences

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-10T21:02:18.959877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.713199Z digest=sha256:fe537e138527bb6d194ce7eaa15429746e95d376b2e11a246db7f59a1d8fbf6c

Observation b583ae01-4be9-4f64-a809-2756a559dd15 · outbound

This paper cites Emergence of division of labour in halictine bees: contributions of social interactions and behavioural variance.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Emergence of division of labour in halictine bees: contributions of social interactions and behavioural variance

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.235230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.722528Z digest=sha256:c922cdcca436987e3ae42a4a82e9e2ca1b9c249cd0d9eb1e616d1b3ff704d3eb

Observation 83cdcd13-c859-4f58-88fb-bdaa6365972d · outbound

This paper cites Multi-agent deep reinforcement learning with type-based hierarchical group communication.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Multi-agent deep reinforcement learning with type-based hierarchical group communication

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.215155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.729786Z digest=sha256:0a614374ca70f29456e1f62c47586f74865d88354a3c4a6dfcabf790a453adfe

Observation 090d9cbc-c6f9-4f6b-a59b-da3d7fd0a000 · outbound

This paper cites Multi-agent reinforcement learning as a rehearsal for decentralized planning.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Multi-agent reinforcement learning as a rehearsal for decentralized planning

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.190680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.737992Z digest=sha256:a3e6d19397c9b655b08b848385d2df6ab18c76965ab32870c9ec403fd61bdcc7

Observation 431133cc-3620-43dd-a3e7-2becf05f7e7b · outbound

This paper cites Reinforcement learning-based joint user pairing and power allocation in mimo-noma systems.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Reinforcement learning-based joint user pairing and power allocation in mimo-noma systems

Reference 11

Resolution
verified exact
doi, observed 2026-08-10T21:02:18.918637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.744243Z digest=sha256:416ed6f0fcc54d8fce8a846bc5499797366486b0af3e045d4c45d73c4153855e

Observation e289dd3b-78f5-4d0e-9483-30ed71682821 · outbound

This paper cites Role-based modeling for designing agent behavior in self-organizing multi-agent systems.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Role-based modeling for designing agent behavior in self-organizing multi-agent systems

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.168675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.748639Z digest=sha256:f8539968735cdffc355c2155e487d2108d9aabd83366968bb0adad01a95f7598

Observation b0aa34ba-ace1-4d61-bd3e-c2063dff96d8 · outbound

This paper cites Jordan, and Zhuoran Yang.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Jordan, and Zhuoran Yang

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.151523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.754362Z digest=sha256:0e960d6ce279687d389ac06e9c5e320e63e2ca2632013f5a6cf060846118e6c3

Observation c1121493-a712-4bd5-9e3b-a86a99abd94b · outbound

This paper cites Optimal and approximate q-value functions for decentralized pomdps.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Optimal and approximate q-value functions for decentralized pomdps

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.135232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.760712Z digest=sha256:981bc1ac65313f80e5d639bc6862f95606b423823b630e2cb31b7394856175c3

Observation a4e77e8b-586f-4c33-a74f-83ea0aac2ac0 · outbound

This paper cites Hierarchical reinforcement learning: A comprehensive survey.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Hierarchical reinforcement learning: A comprehensive survey

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:18.768939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:02:18.768939Z digest=sha256:369541f1a37a1f972fdd38c67d87173334528436a381edfda600878303efc5b8

Observation 58fb5ca5-523b-4bcc-8d02-a0972aa9a9da · outbound

This paper cites Vast: Value function factorization with variable agent sub-teams.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Vast: Value function factorization with variable agent sub-teams

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:18.776671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:02:18.776671Z digest=sha256:c2c52e6be2e2ac45c83c167425e2ea0fb8ef1a34c43fceed5d44830048843913

Observation ba9c940b-af4e-41bf-b55a-d52e52d2de95 · outbound

This paper cites Advances in neural information processing systems 17: proceedings of the 2004 conference, volume 17.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Advances in neural information processing systems 17: proceedings of the 2004 conference, volume 17

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.103251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.784316Z digest=sha256:36d98b69f9142c6ccf3b376c321b1f2609fa2f20378bb71c528f3b4023b2a6c0

Observation 986987ca-1a48-4a50-810a-63112b531ac5 · outbound

This paper cites Sutton, Doina Precup, and Satinder Singh.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Sutton, Doina Precup, and Satinder Singh

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:18.796890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:02:18.796890Z digest=sha256:ee6f4f0aeb14c007e5d5bc901213d9845822150e0dee7963b3b497a4d8d912fd

Observation c42a5f96-d91a-4728-b5e7-398ca9a73480 · outbound

This paper cites Effectiveness of gamified team competition as mhealth intervention for medical interns: a cluster micro-randomized trial.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Effectiveness of gamified team competition as mhealth intervention for medical interns: a cluster micro-randomized trial

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.084218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.802360Z digest=sha256:c46ac61881136fc7c51fc5df9271c5ab93b020028350b9f518b694acf8f9faf2

Observation 25a0eddb-fe5e-48dd-b9dd-891f822cd454 · outbound

This paper cites Beaulieu, and Y.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Beaulieu, and Y

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.069932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.808135Z digest=sha256:43b2157325e70dc1bfb4755299ef406e4ab231efbb64d9bb3e4bec894918bb30

Observation d7f37294-37b6-44e2-a88e-e1eb4c05c8e9 · outbound

This paper cites Adaptive dynamic bipartite graph matching: A reinforcement learning approach.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Adaptive dynamic bipartite graph matching: A reinforcement learning approach

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.055682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.813132Z digest=sha256:18fbe4997e8d274deaca6658ecf37dd4f9f7154b210c43bf29a3b53bb62114c0

Observation aa678bbc-d480-441f-a8c5-4eff6275f472 · outbound

This paper cites Hierarchical dominance structure and social organization in african elephants, loxodonta africana.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Hierarchical dominance structure and social organization in african elephants, loxodonta africana

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.036585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.821021Z digest=sha256:b0c61da5b99ee70cc439740407451aa7d09b3e3826bce0f0208572e1144fb503

Observation e7e48205-37c8-4852-b131-6c3c71240ba8 · outbound

This paper cites Large-scale order dispatch in on-demand ride-hailing platforms: A learning and planning approach.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Large-scale order dispatch in on-demand ride-hailing platforms: A learning and planning approach

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.022090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.826218Z digest=sha256:2a05fc5c5768dc96e8a9d4622912e9ae874486be3aae67b6bbc22277776cb33c

Observation 8e214a0b-6254-44c9-a47d-0ce9e1dca286 · outbound

This paper cites Deep Sets.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Deep Sets

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:18.831276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:02:18.831276Z digest=sha256:cedceda893a47ecea6663b046c45520f854c1f9f2371d326f24022917696ae22

Observation f0d3cc4c-a014-4406-b5e2-c81ee647b00b · outbound

This paper cites Reinforcement learning based local search for grouping problems.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Reinforcement learning based local search for grouping problems

Reference 25

Resolution
verified exact
doi, observed 2026-08-10T21:02:18.895206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.838862Z digest=sha256:767366645b551de5bee0dfcf62a97a0a7e4da97d6b84a64b82e838a4befc56db

Pith citing papers

No inbound Pith citation observations are available.