Pith. sign in

Paper Citation Record · LEDGER

Strategy-Following Multi-Agent Deep Reinforcement Learning Considering Control Strategies Provided to Other Agents

As of 8 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 0 inbound Pith citation observations for arXiv:2607.18719.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.18719 v1

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T14:36:01.708812Z

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

22 of 22 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c4d897d4-29c6-490d-933d-65ab9b10ade1 · outbound

This paper cites Learning to Understand Goal Specifications by Modelling Reward.

Strategy-Following Multi-Agent Deep Reinforcement Learning Considering Control Strategies Provided to Other Agents Learning to Understand Goal Specifications by Modelling Reward

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T14:35:58.826822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:35:58.826822Z digest=sha256:a65387f493f2820638943f3020827ef2c5d9ffe634840ebe64e3ebe6dc81b4e0

Observation a70a5ff2-aa3f-430b-a609-d656c87e30cf · outbound

This paper cites Ask Your Humans: Using Human Instructions to Improve Generalization in Reinforcement Learning.

Strategy-Following Multi-Agent Deep Reinforcement Learning Considering Control Strategies Provided to Other Agents Ask Your Humans: Using Human Instructions to Improve Generalization in Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T14:35:58.982586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:35:58.982586Z digest=sha256:5be97924b96c3dc6a2696bc65c9311923dfca5bb76fdae16b42b2395cbf4bfb5

Observation 59a9f9a9-83d8-49be-b80a-2a714e789019 · outbound

This paper cites Implicit quantile networks for distributional reinforcement learning,.

Strategy-Following Multi-Agent Deep Reinforcement Learning Considering Control Strategies Provided to Other Agents Implicit quantile networks for distributional reinforcement learning,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T14:35:59.112881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:35:59.112881Z digest=sha256:64b6548d9b0427324a311a0016bc00626dab1055894af20aae8730f291843c97

Observation b2633022-5b90-413b-b976-8475aff39cd3 · outbound

This paper cites Speaker- follower models for vision-and-language navigation,.

Strategy-Following Multi-Agent Deep Reinforcement Learning Considering Control Strategies Provided to Other Agents Speaker- follower models for vision-and-language navigation,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T14:35:59.280736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:35:59.280736Z digest=sha256:3a329ce8feb4ab3fe6db68f68f537f8a540b16e39ba2013f1fb20715e28e9517

Observation 4021501e-a9e8-4c9a-98de-ca07834f5eba · outbound

This paper cites Hierarchical program- triggered reinforcement learning agents for automated driving,.

Strategy-Following Multi-Agent Deep Reinforcement Learning Considering Control Strategies Provided to Other Agents Hierarchical program- triggered reinforcement learning agents for automated driving,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T14:35:59.426740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:35:59.426740Z digest=sha256:b16c62ea466ba071a90731612c704733e37265f3edb632a1ac40184d036c3dbe

Observation e071932b-aaae-4483-8464-7346bb70bca1 · outbound

This paper cites Cirl: Controllable imitative reinforcement learning for vision-based self-driving,.

Strategy-Following Multi-Agent Deep Reinforcement Learning Considering Control Strategies Provided to Other Agents Cirl: Controllable imitative reinforcement learning for vision-based self-driving,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T14:35:59.585344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:35:59.585344Z digest=sha256:ec2604ef308f3e39b9c22bb884833cd8e4d74d49355a0b6d3888f269f804a290

Observation 4bc79385-cb73-4ba4-a861-3c1b1bc472e7 · outbound

This paper cites Mapping Instructions to Actions in 3D Environments with Visual Goal Prediction.

Strategy-Following Multi-Agent Deep Reinforcement Learning Considering Control Strategies Provided to Other Agents Mapping Instructions to Actions in 3D Environments with Visual Goal Prediction

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T14:35:59.717941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:35:59.717941Z digest=sha256:78935763149e1f12481c4b97f3702dd53c27da87a48fee757c9dc4d8d31c032d

Observation 38e4d5d0-f80e-4f4a-946e-af85c2f85e83 · outbound

This paper cites Analysis of coordinated behavior structures with multi-agent deep reinforcement learning,.

Strategy-Following Multi-Agent Deep Reinforcement Learning Considering Control Strategies Provided to Other Agents Analysis of coordinated behavior structures with multi-agent deep reinforcement learning,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T14:35:59.850965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:35:59.850965Z digest=sha256:ea2f75502545d9eb99ad8e4764032559fffc9d26b5aeb7b2d1f5549b8f50d4ef

Observation 7b2af1ed-1193-494d-94db-78bc5344b516 · outbound

This paper cites Interpretability for conditional co- ordinated behavior in multi-agent reinforcement learning,.

Strategy-Following Multi-Agent Deep Reinforcement Learning Considering Control Strategies Provided to Other Agents Interpretability for conditional co- ordinated behavior in multi-agent reinforcement learning,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T14:35:59.968014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:35:59.968014Z digest=sha256:252e1d907da104020b1f97662bb1b0ab86148d1d900051eed241c37c73c54555

Observation e9f12557-d7f7-48d0-badb-15f5d24e0eb6 · outbound

This paper cites Strategy-following multi-agent deep reinforcement learning through external high-level instruction,.

Strategy-Following Multi-Agent Deep Reinforcement Learning Considering Control Strategies Provided to Other Agents Strategy-following multi-agent deep reinforcement learning through external high-level instruction,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T14:36:00.142659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:36:00.142659Z digest=sha256:8d79e576b0bafd9c47278ba9cdbf85a96f176101ad2f18dd174ce52f4d257378

Observation 5fec71b9-d086-4670-876d-db3aba18f3f6 · outbound

This paper cites an unresolved cited work.

Strategy-Following Multi-Agent Deep Reinforcement Learning Considering Control Strategies Provided to Other Agents Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T14:36:00.334122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:36:00.334122Z digest=sha256:1c2f4ff6889fd32fb94cf5be394ef6086991c072fd62e5b91f06f61dabdbcc04

Observation 4426bf42-7569-4684-8472-cc6802b87a8c · outbound

This paper cites EPOpt: Learning Robust Neural Network Policies Using Model Ensembles.

Strategy-Following Multi-Agent Deep Reinforcement Learning Considering Control Strategies Provided to Other Agents EPOpt: Learning Robust Neural Network Policies Using Model Ensembles

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T14:36:00.510306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:36:00.510306Z digest=sha256:5ca3088a471bf5202a23b590c1fc8f1d382120cfefabc7bcde24835215304a70

Observation 8d4bfb62-2df5-4099-b524-3473ed7501de · outbound

This paper cites A reduction of imitation learning and structured prediction to no-regret online learning,.

Strategy-Following Multi-Agent Deep Reinforcement Learning Considering Control Strategies Provided to Other Agents A reduction of imitation learning and structured prediction to no-regret online learning,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T14:36:00.635131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:36:00.635131Z digest=sha256:f697deb8e5c1a705d419980d8012043e2a4510de5cd849f8a608c8b36a38b968

Observation f29c1eb2-9923-48e6-9b69-b4c50c397dd7 · outbound

This paper cites The StarCraft Multi-Agent Challenge.

Strategy-Following Multi-Agent Deep Reinforcement Learning Considering Control Strategies Provided to Other Agents The StarCraft Multi-Agent Challenge

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T14:36:00.853580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:36:00.853580Z digest=sha256:38c838000fbe3a99f70d4ef4b7eff19eee5dc9b4e3a401b787aabaeb04e09f68

Observation f0fb071e-d947-4772-9e45-7ec5f919a554 · outbound

This paper cites Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving.

Strategy-Following Multi-Agent Deep Reinforcement Learning Considering Control Strategies Provided to Other Agents Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T14:36:00.966006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:36:00.966006Z digest=sha256:15bba230cbd293a4c11b45ed2029f4c416256e27151d414fe8f6a6d4c66c10da

Observation da24faac-f0cc-4686-8792-e6507573961c · outbound

This paper cites Task offloading and trajectory scheduling for uav-enabled mec networks: An madrl algorithm with prioritized experience replay,.

Strategy-Following Multi-Agent Deep Reinforcement Learning Considering Control Strategies Provided to Other Agents Task offloading and trajectory scheduling for uav-enabled mec networks: An madrl algorithm with prioritized experience replay,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T14:36:01.030417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:36:01.030417Z digest=sha256:4c0c411aaac0389b057d6f25d9218b11e1b5b4df49c009dbb6e3bbba17f76a58

Observation 457f68cf-5c46-487a-8daa-7713383c56f8 · outbound

This paper cites Program guided agent,.

Strategy-Following Multi-Agent Deep Reinforcement Learning Considering Control Strategies Provided to Other Agents Program guided agent,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T14:36:01.180288Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:36:01.180288Z digest=sha256:df6d5bf8b62ea51cd0330e72a28f751b71576fa4e74f3fb418efe3c79a19a96c

Observation bcbbf0e7-5517-4b61-bfca-8cffc81f6c96 · outbound

This paper cites Attention is all you need,.

Strategy-Following Multi-Agent Deep Reinforcement Learning Considering Control Strategies Provided to Other Agents Attention is all you need,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T14:36:01.310222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:36:01.310222Z digest=sha256:2c19c2d0ffb462b772e1c9a3b7d25e7342d0ee3ab75696d198b5f5fd2a7dcb16

Observation 409aac75-7301-4958-ac92-270275af07e9 · outbound

This paper cites Understanding natural language,.

Strategy-Following Multi-Agent Deep Reinforcement Learning Considering Control Strategies Provided to Other Agents Understanding natural language,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T14:36:01.389483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:36:01.389483Z digest=sha256:0ea789f3d3c1350588f7834a888912d9382e96c5d4a24991c11c1c69f174ca48

Observation 4514b892-c853-4ff2-9209-44fe0a5d6d96 · outbound

This paper cites Toward human-in-the-loop ai: Enhancing deep reinforcement learning via real-time human guidance for autonomous driving,.

Strategy-Following Multi-Agent Deep Reinforcement Learning Considering Control Strategies Provided to Other Agents Toward human-in-the-loop ai: Enhancing deep reinforcement learning via real-time human guidance for autonomous driving,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T14:36:01.464775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:36:01.464775Z digest=sha256:4de4e0c35b9c39d397be4e09132b82cc0c10ddb275d321100510ee3ff5ad2c40

Observation aaf0ee00-3225-4951-ae76-e3b2370b72df · outbound

This paper cites Program synthesis guided reinforcement learning for partially observed environments,.

Strategy-Following Multi-Agent Deep Reinforcement Learning Considering Control Strategies Provided to Other Agents Program synthesis guided reinforcement learning for partially observed environments,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T14:36:01.597241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:36:01.597241Z digest=sha256:5bc7f5755ef053f67a5820d07164581be9ebee58dd569b43200bf975658e3507

Observation 3c66c4dc-028f-4165-8d36-acdb504def6b · outbound

This paper cites Joint sensing and communication optimization in target-mounted stars-assisted vehicular networks: A madrl approach,.

Strategy-Following Multi-Agent Deep Reinforcement Learning Considering Control Strategies Provided to Other Agents Joint sensing and communication optimization in target-mounted stars-assisted vehicular networks: A madrl approach,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T14:36:01.708812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:36:01.708812Z digest=sha256:096b03b1be172978a53de400167e7d24a20162b026d1a8dd14d354d2e2ba8822

Pith citing papers

No inbound Pith citation observations are available.