Pith. sign in

Paper Citation Record · LEDGER

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour

As of 10 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 0 inbound Pith citation observations for arXiv:2502.06113.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.06113 v1

Coverage vector

measured 48 of 48 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T16:45:17.660092Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

48 of 48 outbound references displayed

  • verified exact1
  • verified fuzzy43
  • unresolved4
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation abb96bfa-0dbe-4932-a438-bd80b8f69325 · outbound

This paper cites It refers to the process of determining a path or set of movements, that ensures a robot or swarm thoroughly covers a given area or surface.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour It refers to the process of determining a path or set of movements, that ensures a robot or swarm thoroughly covers a given area or surface

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.416073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.433678Z digest=sha256:0d95694a005bc5199129259813fc55acdd71295158d4088fd2df962b258dff80

Observation 273f97ce-c745-4405-a28c-06df49f509e3 · outbound

This paper cites This environment can be either known or unknown.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour This environment can be either known or unknown

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.401502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.439182Z digest=sha256:770c2af05330d42bb65e344da2d28ca60e1328899f9735d6d4e92229ce47ebf5

Observation c7647a79-500a-433c-b29e-5cc1a80a49f8 · outbound

This paper cites Then, the epsilon-greedy method is chosen to improve exploration during training and facilitate the incorporation of the PSO with the first version of the MASAC.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Then, the epsilon-greedy method is chosen to improve exploration during training and facilitate the incorporation of the PSO with the first version of the MASAC

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.371369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.449539Z digest=sha256:6390855ab3111e561897591fe3960d6cd748f5888533fe27069f856e3e8c0330

Observation a4f81953-5d06-408d-a89c-8b12a926a371 · outbound

This paper cites One of the chosen metrics for the evaluation is the reward value.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour One of the chosen metrics for the evaluation is the reward value

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.355986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.454964Z digest=sha256:d3acd064cd1cf2575895e9e1fcd3a9eb7ad0dc0e1b8b79c9ac82c43933778958

Observation 87d4b32d-a796-47b5-a2a8-a8a1fa2f235a · outbound

This paper cites an unresolved cited work.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-08T16:45:18.341516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.459856Z digest=sha256:4cd9b599f5e6d83922aba763239ea589ea87cdd6203a657ea1c098894da26dc7

Observation 1d33d3ac-930a-4e54-ba50-e66f028b7dee · outbound

This paper cites Deep reinforcement learning for multiagent systems: A review of challenges, solutions, and applications,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Deep reinforcement learning for multiagent systems: A review of challenges, solutions, and applications,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.268937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.491223Z digest=sha256:9b339f5eea3a9c197505ec4fcfef7a6b2d5cb8dd89031dfcfcc5ba7f19abb408

Observation 1e84e4d5-fd69-4599-8d8e-463e6ec058bb · outbound

This paper cites Multi-Agent Generative Adversarial Interactive Self-Imitation Learning for AUV Formation Control and Obstacle Avoidance.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Multi-Agent Generative Adversarial Interactive Self-Imitation Learning for AUV Formation Control and Obstacle Avoidance

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-08T16:45:17.735834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.465692Z digest=sha256:a654f46c76872bb82622a35ac2d4049b796d6d5448067acea88838af01124b6e

Observation 982cb50b-fa8d-4473-b73c-e76256481ef9 · outbound

This paper cites A multi -agent framework with moos-ivp for autonomous underwater vehicles with sidescan sonar sensors,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour A multi -agent framework with moos-ivp for autonomous underwater vehicles with sidescan sonar sensors,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.326993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.470720Z digest=sha256:e76804be52ca153944b3689996f398298dd465ead7879b915781bcc9931a8555

Observation 0bb748c7-f6a9-4782-a540-ba2971c3937c · outbound

This paper cites A survey on swarm robotics for area coverage problem,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour A survey on swarm robotics for area coverage problem,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.313006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.475913Z digest=sha256:c5314c6c2e965ed314c9eb825dd77b76c2e91ae4ea7fca5b6cc67ffaea0ae24a

Observation 9dad8229-07fe-4680-9283-a23270629952 · outbound

This paper cites an unresolved cited work.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-08T16:45:18.298630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.481025Z digest=sha256:820394e100a16d91d5d038316cef8099381f0a604d07f5fe48d71efd20fc044c

Observation 8e87ce62-82ee-49ea-980e-15cf8f5defdd · outbound

This paper cites Buşoniu, R.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Buşoniu, R

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.283796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.486029Z digest=sha256:3c2b332ab29b994b328a1df578080022b88b6a1881398b6b5eaf5c002ad9c630

Observation 4062682d-c4bf-42ec-8fb8-c89367072fff · outbound

This paper cites Multi -agent reinforcement learning: A selective overview of theories and algorithms,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Multi -agent reinforcement learning: A selective overview of theories and algorithms,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.194350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.521597Z digest=sha256:3db1a71da4e83ad0f52b27160e9602140fc511b6fbb2f803641717c6b1636699

Observation 3d6f99fc-0e4d-4d20-b4ea-ec7b5953206e · outbound

This paper cites An empirical investigation of the challenges of real-world reinforcement learning.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour An empirical investigation of the challenges of real-world reinforcement learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T16:45:17.496518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:45:17.496518Z digest=sha256:88ff2840b9f96794ed3cbcda96e6ead08e434b1613d4964aa6ee192e29d404bb

Observation ecbd545d-e1ed-48d2-a25e-6b0b923ec27c · outbound

This paper cites Multi -robot path planning using an improved self -adaptive particle swarm optimization,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Multi -robot path planning using an improved self -adaptive particle swarm optimization,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.254248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.501870Z digest=sha256:b3ebe8d111d3355082f1972c31108a408a00019344ab394e04af78d4754f902c

Observation 61b184c1-99a3-4f71-84d3-1754267870f5 · outbound

This paper cites Group decisions in humans and animals: A survey,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Group decisions in humans and animals: A survey,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.238665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.506755Z digest=sha256:c6dc94f8ae50591ac3d176af5b60a5fe3149eb972a852570f2d71133ba30dd46

Observation 98b29bea-e15e-4585-8577-d1066e378b5c · outbound

This paper cites From animal collective behaviors to swarm robotic cooperation,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour From animal collective behaviors to swarm robotic cooperation,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.224268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.511594Z digest=sha256:d6eb5cf7dec60f82408280e961ef5b80f5d241120e1bf2a49f80eb0522dec647

Observation 71208a3d-ccbb-4bd7-b78d-b71a8e206246 · outbound

This paper cites Sample efficient multi-agent reinforcement learning with masked reconstruction,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Sample efficient multi-agent reinforcement learning with masked reconstruction,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.209389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.516727Z digest=sha256:a24cd5d68bed3174c2b1bb17c91fcbe404d949b359ff6ff5192b3ebb3889631f

Observation cf726d11-10d0-4005-ae16-30e368c59876 · outbound

This paper cites Multi-agent reinforcement learning: a critical survey,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Multi-agent reinforcement learning: a critical survey,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.106374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.550000Z digest=sha256:6013b9c28fb47e02bef105f723e6b977d192ca78e7c48c207739a8a5e723f284

Observation eb51cdbb-c60f-455f-a049-bb008d303251 · outbound

This paper cites A survey and critique of multiagent deep reinforcement learning,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour A survey and critique of multiagent deep reinforcement learning,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.179774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.526242Z digest=sha256:2a1a31f076a09e50e5c8a2adc67bb8124a2b486ac095238ef15492d4974aabb0

Observation 9bbe0531-54d1-4944-8e41-f8cce4951621 · outbound

This paper cites Value-based methods [22] estimate the value of each state or state -action pair to derive optimal policies by selecting actions that maximize the cumulative reward.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Value-based methods [22] estimate the value of each state or state -action pair to derive optimal policies by selecting actions that maximize the cumulative reward

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.386370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.444002Z digest=sha256:1b7bb4f86dbb0b4f25269abac549f748ca3878e2313803fee4ec1ab5190db947

Observation 4411e184-94d5-43f6-9fe7-321bfd02ef13 · outbound

This paper cites Multi-agent reinforcement learning: A review of challenges and applications,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Multi-agent reinforcement learning: A review of challenges and applications,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.164646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.530828Z digest=sha256:6cc52dc25e0f54d2c39ffd7eeca51c86b279616a064842efe82fce1845d19cd9

Observation 50b40735-6b42-4ec3-aadd-327bcc2eab2f · outbound

This paper cites Decentralised learning in systems with many, many strategic agents,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Decentralised learning in systems with many, many strategic agents,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.150351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.535528Z digest=sha256:8571dc12f82780f20ffe6fb40a4a7cd3dc3fc45ef2d20181dfc324e4e0f5a688

Observation 9cae0347-92ca-4334-8e15-edb8f0d13012 · outbound

This paper cites An efficient centralized multi-agent reinforcement learner for cooperative tasks,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour An efficient centralized multi-agent reinforcement learner for cooperative tasks,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.136007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.540381Z digest=sha256:e9a7bcf027b7028065ffcedd5649da88c0ab24f2e60a85374517e35241d070f4

Observation 7952996a-9bae-4e78-a30c-5605b8f36e03 · outbound

This paper cites Fully decentralized cooperative multi-agent reinforcement learning: A survey,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Fully decentralized cooperative multi-agent reinforcement learning: A survey,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.121258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.545146Z digest=sha256:cee4e01fae108f962aa09200a20f17b80b6f3c7c29cfa333c715f12be234dd2d

Observation d7568734-782b-465a-a9d2-7e59c2024b0f · outbound

This paper cites Entropy regularized actor-critic based multi-agent deep reinforcement learning for stochastic games,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Entropy regularized actor-critic based multi-agent deep reinforcement learning for stochastic games,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.092185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.554790Z digest=sha256:21b83f74d9d11175f464e4bc5efc83c546016afd82dadc4aef07b1065c8a3902

Observation a5ebed8b-25bf-4813-9c02-f66a98935604 · outbound

This paper cites Soft Actor-Critic Algorithms and Applications.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Soft Actor-Critic Algorithms and Applications

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-08T16:45:17.559339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:45:17.559339Z digest=sha256:c2cd2c947f756372784d2be4754e82adb4dd6274cf3c599d87f23bdb53cea31a

Observation 219ad0d0-9e7b-49e4-ab94-d283b730e6cb · outbound

This paper cites Evaluation of a deep reinforcement-learning-based controller for the control of an autonomous underwater vehicle,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Evaluation of a deep reinforcement-learning-based controller for the control of an autonomous underwater vehicle,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.077725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.564292Z digest=sha256:acdedd0df4dc99a795c4b1c1b65b0660eaefe50a17c98514787e326a1e2444a0

Observation ad5cc09f-6f48-4f54-9f36-aacab44778e0 · outbound

This paper cites Leveraging world model disentanglement in value-based multi- agent reinforcement learning,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Leveraging world model disentanglement in value-based multi- agent reinforcement learning,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.062131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.569154Z digest=sha256:e12e76c171a88c8eb4b3546c40e460e18652a508c02d8e1af0f9482d1d8f6c61

Observation b51f8442-ad12-438a-803a-580a3eb11b70 · outbound

This paper cites A collaborative multiagent reinforcement learning method based on policy gradient potential,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour A collaborative multiagent reinforcement learning method based on policy gradient potential,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.046399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.573472Z digest=sha256:5bafea0fc8d86c1eee0daa15a7078b075983d03500fcca6a60a5678881ea05b2

Observation c304391f-5f75-4daa-abf3-f7246bd797c8 · outbound

This paper cites Learning to walk via deep reinforcement learning,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Learning to walk via deep reinforcement learning,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.031577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.577934Z digest=sha256:f7d07b762e417d462174bbacae155ef718edba6927e78a999719674589aa1ebd

Observation a5156528-d142-4a9f-8676-e616a80d7efb · outbound

This paper cites Sampling efficient deep reinforcement learning through preference-guided stochastic exploration,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Sampling efficient deep reinforcement learning through preference-guided stochastic exploration,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.016412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.582158Z digest=sha256:860a26c8ecf1ef413f4e06721cef9167605854c99b1fb7b3894acfae22836c70

Observation 44ed5261-e358-4d88-b6cd-996022d156ab · outbound

This paper cites Heuristic-guided reinforcement learning,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Heuristic-guided reinforcement learning,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:18.001581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.586664Z digest=sha256:8a8bfdf26991eb2195e145c595a81ad00a64258dc5a8d59052321529cf345733

Observation a6c92d13-05da-48d7-8356-3eba463f5a95 · outbound

This paper cites Heuristically accelerated reinforcement learning: Theoretical and experimental results,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Heuristically accelerated reinforcement learning: Theoretical and experimental results,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:17.986376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.591170Z digest=sha256:7f0f4d16d9a7db0193fd72fc12ed5a6f03800eb4971af628b52e71ca84054192

Observation 0e201400-f711-4f32-8838-fde2336bf79c · outbound

This paper cites Transferring knowledge as heuristics in reinforcement learning: A case-based approach,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Transferring knowledge as heuristics in reinforcement learning: A case-based approach,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:17.971416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.595424Z digest=sha256:2011f246184bab5cc9406f71fc2129e7b2fa087311629826b49f34f535a8f1c4

Observation 64b3559c-c0cd-4759-9a9c-b0d4d9550d3c · outbound

This paper cites Heuristically Accelerated Reinforcement Learning by Means of Case-Based Reasoning and Transfer Learning,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Heuristically Accelerated Reinforcement Learning by Means of Case-Based Reasoning and Transfer Learning,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:17.957079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.599928Z digest=sha256:1f01168fdae7e55588f02e957eb6e721d957f6de559d4d62ece44ba222a28412

Observation 9b219dd5-5c92-402c-8db3-c7f6dac807c5 · outbound

This paper cites Ant system: optimization by a colony of cooperating agents,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Ant system: optimization by a colony of cooperating agents,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:17.941959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.604590Z digest=sha256:da49b4a9b06247fc8a40b9903191cd390143096ba29851ba2e9195c088ac7612

Observation 747fc2de-80cf-4c29-b5cc-7fcc64b9b29b · outbound

This paper cites The bees algorithm and mechanical design optimisation,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour The bees algorithm and mechanical design optimisation,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:17.926120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.609190Z digest=sha256:45d7b1c2663156dfeb4fde63f47415145a63e56b7010978ffcce310e3addc28b

Observation d00e7d7d-dd25-4594-a6a1-f147871ebddb · outbound

This paper cites Firefly algorithms for multimodal optimization,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Firefly algorithms for multimodal optimization,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:17.910718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.613713Z digest=sha256:92a2e30abfde28185d054a15680fa4bdd2822aa3c5de49817e743f44bcf59da9

Observation b3145215-8b85-4ce8-9a54-9a5c89c7ca37 · outbound

This paper cites Brain storm optimization algorithm,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Brain storm optimization algorithm,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:17.894455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.618424Z digest=sha256:b159f4ac6113a60c1eabbc778181793d7d09661f6a9fe8cce5b4a13a05dd1a51

Observation 17e15b61-98b5-490c-87ac-3e8251b864e3 · outbound

This paper cites Group search optimizer: An optimization algorithm inspired by animal searching behavior,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Group search optimizer: An optimization algorithm inspired by animal searching behavior,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:17.879904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.623267Z digest=sha256:28d1889d508d4a3961d96b0ef2222fa5cc88dd97a4f928fde2deef2099ad4661

Observation 1c4c3b39-6653-428d-9ba4-457f1302fefd · outbound

This paper cites Solving engineering design problems by social cognitive optimization,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Solving engineering design problems by social cognitive optimization,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:17.864951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.627931Z digest=sha256:ec076ec9708262558fcb2210069d667f65ac66f87fcb3709899b6e99c18cdf20

Observation 907fce4d-ec6b-4bee-bbc6-ad0d109cd0ef · outbound

This paper cites Particle swarm optimization,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Particle swarm optimization,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:17.849273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.632726Z digest=sha256:475554d04d637426a7efcdc9bb1968e02b77f118ce6bfe18f736955a61eb775c

Observation 35ff0bfd-42e2-49bb-b08f-4ce87f845130 · outbound

This paper cites Distributed 3-d path planning for multi-uavs with full area surveillance based on particle swarm optimization,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Distributed 3-d path planning for multi-uavs with full area surveillance based on particle swarm optimization,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:17.834147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.637226Z digest=sha256:dfe1d4432b50603a9779b92f35926d70cd14ac99cf18de9f5f3c3613360daf63

Observation 1068253a-6dc0-4ff8-87e7-86fe797f71a9 · outbound

This paper cites Particle swarm optimization algorithm and its applications: A systematic review,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Particle swarm optimization algorithm and its applications: A systematic review,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:17.817444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.641924Z digest=sha256:8c0267c1bc94a3a227f8a1e5e256980d1ceaf27e9ec92d53a57111827d80bb11

Observation 42c7d4af-ad0b-4a18-a990-5014cbe49ef1 · outbound

This paper cites Sim-to-real transfer of adaptive control parameters for auv stabilisation under current disturbance,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Sim-to-real transfer of adaptive control parameters for auv stabilisation under current disturbance,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:17.800829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.646377Z digest=sha256:ec1f642ee99421e579ea572f49b0e3f58d7eb2f183dc10a6a0d60a24290583ee

Observation 7c446ed9-fa46-435f-af65-88d8a34ebd33 · outbound

This paper cites A deeper look at experience replay,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour A deeper look at experience replay,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:17.785395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.651026Z digest=sha256:d430680a27d57de3f323494b54e96614d086bfd1b4c2042675e80fd2f9d77466

Observation e5923bd0-637d-4b4d-b9a4-2a304d0d1442 · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:17.768383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.655612Z digest=sha256:f106ddd41ddc66b910acb0545c23fa0d7fbe50dba5ce7423eb3e0c9d0212bd8a

Observation 842c66fc-812d-4e2c-9a54-2e812af0fe69 · outbound

This paper cites Learning adaptive control of a uuv using a bio-inspired experience replay mechanism,.

Towards Bio-inspired Heuristically Accelerated Reinforcement Learning for Adaptive Underwater Multi-Agents Behaviour Learning adaptive control of a uuv using a bio-inspired experience replay mechanism,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:45:17.751603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:45:17.660092Z digest=sha256:79459055e43e7aa671f8cf9c16188281482db77adf9403caddbdaa5d475ef23a

Pith citing papers

No inbound Pith citation observations are available.