Pith. sign in

Paper Citation Record · LEDGER

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain

As of 8 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 0 inbound Pith citation observations for arXiv:2506.06786.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.06786 v1

Coverage vector

measured 24 of 24 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:55:08.160720Z

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

24 of 24 outbound references displayed

  • verified exact1
  • verified fuzzy12
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7398286a-4956-4f95-a711-1972aa3f39ce · outbound

This paper cites Exploration in deep reinforcement learning: A survey,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Exploration in deep reinforcement learning: A survey,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:55:08.037952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:55:08.037952Z digest=sha256:9d7267546aef22df65b8ae37b69fa74bc1d078218cab60761d61a885ac9cab2d

Observation 2b7abad1-1726-4681-be06-c2e94e5c61fb · outbound

This paper cites Deep reinforcement learning for time-critical wilderness search and rescue using drones,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Deep reinforcement learning for time-critical wilderness search and rescue using drones,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:55:08.043659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:55:08.043659Z digest=sha256:e5b20dc1f9b834427ef9a91bfba513bb749ea3ee68ef6d71baeb050a71b8ee84

Observation 292fd82b-4056-49fd-bcbe-7899350d43ac · outbound

This paper cites MaxInfoRL: Boosting exploration in reinforcement learning through information gain maximization.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain MaxInfoRL: Boosting exploration in reinforcement learning through information gain maximization

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T05:55:08.048563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:55:08.048563Z digest=sha256:e1a2cecdf100a035734013cdde3606524c68d5095f4bf49d3c7d23ef61ecf5da

Observation e5e0854f-babb-4008-85b6-8929849cc418 · outbound

This paper cites Regret bounds for information-directed reinforcement learning,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Regret bounds for information-directed reinforcement learning,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:55:08.540988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:55:08.054122Z digest=sha256:722ce05f58354f384413158d1a1ec6d638d71a324e555466b1b94b6e47936a88

Observation adf17a1d-4b14-4273-a5f1-24b4f5ddc7b7 · outbound

This paper cites Selective exploration and information gathering in search and rescue using hierarchical learning guided by natural language input,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Selective exploration and information gathering in search and rescue using hierarchical learning guided by natural language input,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:55:08.523642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:55:08.059018Z digest=sha256:301d2b909fa95c1ed8c1ecfe7840aff15a5e515086e316935b7dbafe1ffbdfca

Observation 021747fd-36ab-4553-82ce-91559fef86b9 · outbound

This paper cites Adversar: Adversarial search and rescue via multi-agent reinforcement learning,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Adversar: Adversarial search and rescue via multi-agent reinforcement learning,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:55:08.506480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:55:08.064858Z digest=sha256:2e2b8eb309305293ad4b61c643db63fe08a74ed4cf1412c9f931190df494353b

Observation 7867432d-7291-4cf8-8fbc-71340a251fdd · outbound

This paper cites Target search and navigation in heterogeneous robot systems with deep reinforcement learning,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Target search and navigation in heterogeneous robot systems with deep reinforcement learning,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:55:08.490184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:55:08.070670Z digest=sha256:a7c6f113e431614c1c03d5dd52d5d77ef371469a2d5d1454db3b4fc877f65c06

Observation f7627ad9-0862-4d8c-8e21-2015ab9d841f · outbound

This paper cites Multi-robot cooperative target search based on distributed reinforcement learning method in 3d dynamic environments,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Multi-robot cooperative target search based on distributed reinforcement learning method in 3d dynamic environments,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:55:08.473171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:55:08.075594Z digest=sha256:657b627c3e95d7997a9f4e3b428d2f9714fcc014d225fc82e9a74745c858c4eb

Observation 399349a3-5f65-4f6e-8ac5-275f1bafbadf · outbound

This paper cites an unresolved cited work.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:55:08.080420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:55:08.080420Z digest=sha256:86c1a10fbb54b7a7508cab60c923e38f48c022bf2fb5979aa553ae68a1b8e9ba

Observation 9d35c641-da1d-4d39-933d-4eb8530a0265 · outbound

This paper cites Boltzmann exploration done right,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Boltzmann exploration done right,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:55:08.444604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:55:08.085052Z digest=sha256:ff912bcc1dce38df4e05053e7788290aabd1e50e62499f8af3f8298634918ebc

Observation 35cbcc6b-197f-486b-9597-9c78e7c6b97d · outbound

This paper cites Using confidence bounds for exploitation-exploration trade- offs,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Using confidence bounds for exploitation-exploration trade- offs,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:55:08.426789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:55:08.089768Z digest=sha256:284f6d7eb653c2419d406c0e872bc6a52b1f8118e14cd7b84900b9152afcf3bb

Observation 4a812954-355a-49bf-a755-6be8287c776d · outbound

This paper cites An empirical evaluation of thompson sam- pling,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain An empirical evaluation of thompson sam- pling,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:55:08.408573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:55:08.094471Z digest=sha256:0a82ceee7ef3dc67a34e5ab03218a93a6db24a33cdc13835e0304ca71c30299e

Observation b9813539-f0d5-48f3-8c5c-cada4c50e4c9 · outbound

This paper cites Curiosity-driven exploration by self-supervised prediction,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Curiosity-driven exploration by self-supervised prediction,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T05:55:08.099735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:55:08.099735Z digest=sha256:9b3648d5f8799e9caaaee7dd1b5e8e17544ba1025cd1eb4eb217325600589561

Observation 95e244b5-9f07-4d80-b95f-e0a892e61399 · outbound

This paper cites Unifying count-based exploration and intrinsic motiva- tion,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Unifying count-based exploration and intrinsic motiva- tion,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:55:08.104666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:55:08.104666Z digest=sha256:4c6cd254917f8f305d3196e0d7fb7baf1e875cdd36f0485782253ec5e9243c3b

Observation 1e91654f-7996-496d-9846-07d9ccb8f408 · outbound

This paper cites Surprise-Based Intrinsic Motivation for Deep Reinforcement Learning.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Surprise-Based Intrinsic Motivation for Deep Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T05:55:08.110056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:55:08.110056Z digest=sha256:94ac4602f80e9aad0904673e7d88ba0b8e3d7fe90c6df3ce14f0a52542e8e4ea

Observation 217eb0e2-bdb2-4011-aec0-5feafce2fd0a · outbound

This paper cites Exploration by Random Network Distillation.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Exploration by Random Network Distillation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:55:08.116114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:55:08.116114Z digest=sha256:ed5b96eac0e4895ba33c8a830d72564b74695996e0fa5011d5d60c9d7f4b99ca

Observation 20bf38b9-0336-400e-80d1-e7f705d683df · outbound

This paper cites Model-agnostic meta-learning for fast adaptation of deep networks,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Model-agnostic meta-learning for fast adaptation of deep networks,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:55:08.121560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:55:08.121560Z digest=sha256:ea3ba4912f7fe190225bda5f86d97e877d7d685c2f85c20d2e5f940a568abec8

Observation beca5d45-bc1d-40dd-93c9-efe68b3f5075 · outbound

This paper cites Lattimore and C.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Lattimore and C

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:55:08.358828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:55:08.126666Z digest=sha256:28a4886cb0a32ca54413cfb6e1dd57520c751f3506fae94a534ee7d7c7a05e01

Observation d2b74e39-e08d-469b-bce2-248c23810283 · outbound

This paper cites Hidden parameter markov decision processes: A semiparametric regression approach for discovering latent task parametrizations,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Hidden parameter markov decision processes: A semiparametric regression approach for discovering latent task parametrizations,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:55:08.341394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:55:08.131798Z digest=sha256:96a3a204a4686d6e445c5b079d48414541651e241ca4c16dacf35530fdb2aeff

Observation 9135ebbd-6a7a-4ac0-89ff-ded3ac6cd182 · outbound

This paper cites Contextual Markov Decision Processes.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Contextual Markov Decision Processes

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T05:55:08.137170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:55:08.137170Z digest=sha256:282044032ed532e3f9afdb3325c323f69d240393dc50ec926f562f0764e7c51f

Observation 5d9ee41b-2377-4147-b8d2-58f513e91a1e · outbound

This paper cites Reinforcement Learning in Presence of Discrete Markovian Context Evolution.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Reinforcement Learning in Presence of Discrete Markovian Context Evolution

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:55:08.207365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:55:08.142795Z digest=sha256:01628317643a3cb510f3df1a3651c969fbb9a77ad92982fe04662c1397e2e12e

Observation 3e2e32e3-2288-4089-91a9-9d2c56dae461 · outbound

This paper cites Context-aware dynamics model for generalization in model-based reinforcement learning,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Context-aware dynamics model for generalization in model-based reinforcement learning,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:55:08.323520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:55:08.149028Z digest=sha256:c127f8c10cd3b8bc4d242b29722756d77f6b8b3f9a88530d6d00d193d2d4e7ea

Observation 26648876-6b35-45f1-8482-0234814bf0cc · outbound

This paper cites Learning to optimize via information- directed sampling,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Learning to optimize via information- directed sampling,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:55:08.306067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:55:08.155554Z digest=sha256:d1d06cc571095c1dbdb17972307477c4decc263a558f7d2cf793c48137f49054

Observation 7f08c08f-3949-43d9-b9a4-7393359e6b6f · outbound

This paper cites Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:55:08.160720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:55:08.160720Z digest=sha256:51157ea4b2ad594de84078dc5078ee0cc336695e53a4843ce87eca8b651de415

Pith citing papers

No inbound Pith citation observations are available.