Pith. sign in

Paper Citation Record · LEDGER

Fairness in Reinforcement Learning with Bisimulation Metrics

As of 22 August 2026, this Paper Citation Record lists 78 of 78 outbound references and 1 inbound Pith citation observation for arXiv:2412.17123.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.17123 v2

Coverage vector

measured 78 of 78 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T05:52:34.651783Z

measured 79 of 79 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-22T07:48:56.042653Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T07:51:16.230069Z

Reference resolution

78 of 78 outbound references displayed

  • verified exact0
  • verified fuzzy49
  • unresolved28
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b2557ea1-fc6b-40be-9b1a-9b8e67a7a19d · outbound

This paper cites write newline.

Fairness in Reinforcement Learning with Bisimulation Metrics write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.382043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.382043Z digest=sha256:a3934d174a087690863ba70686fe0f6367a5c26831ff87b1419e843572f03803

Observation d75cb631-1cd8-4203-9a06-0cd63955e2db · outbound

This paper cites Design and control of soft robots using differentiable simulation.

Fairness in Reinforcement Learning with Bisimulation Metrics Design and control of soft robots using differentiable simulation

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.519834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.387041Z digest=sha256:7e12b70a84429b5bd379cebaa4e606fdfee89b10436120166a2707b3f4b2047b

Observation 1140f803-5ca3-4136-8694-12e48b358db1 · outbound

This paper cites Fairness and machine learning: Limitations and opportunities.

Fairness in Reinforcement Learning with Bisimulation Metrics Fairness and machine learning: Limitations and opportunities

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.508887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.390814Z digest=sha256:d5ea752197df65176bb3d6e24a7d2aa9fc4a045ae4cbe4f708b5adcce5b2b4a4

Observation d92d77cc-4c89-4ba2-a3d1-f755143adf47 · outbound

This paper cites Estimating or Propagating Gradients Through Stochastic Neurons for Conditional Computation.

Fairness in Reinforcement Learning with Bisimulation Metrics Estimating or Propagating Gradients Through Stochastic Neurons for Conditional Computation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.395095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.395095Z digest=sha256:919572866ecb85bd9200d90f0e818f27dc2e8beb7f2ce5dbe80119ee5e685f04

Observation 958b782e-6643-4718-842d-8fda2d1f8a2c · outbound

This paper cites My fair bandit: Distributed learning of max-min fairness with multi-player bandits.

Fairness in Reinforcement Learning with Bisimulation Metrics My fair bandit: Distributed learning of max-min fairness with multi-player bandits

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.498211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.399374Z digest=sha256:afbf45087baceef5d4f35644d765e29b71d2bb489847a5314820e030d03b3870

Observation 2f630c7f-1d57-4b3b-9e96-c38ad731dc0a · outbound

This paper cites an unresolved cited work.

Fairness in Reinforcement Learning with Bisimulation Metrics Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-11T05:52:35.487371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.403149Z digest=sha256:27e36076da0cf120c2798b711993516bbc5495ccf01745369c6d1950478fd8b5

Observation a016338f-dcd3-4ffa-9ac7-b5aecfb84d71 · outbound

This paper cites OpenAI Gym.

Fairness in Reinforcement Learning with Bisimulation Metrics OpenAI Gym

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.406782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.406782Z digest=sha256:f9b8352bdac24c0c0239325e7d933a48f53dd007f981e624b9c7b819ae70c867

Observation ed927710-a17e-4371-b44a-c2e08259c2d0 · outbound

This paper cites Scalable methods for computing state similarity in deterministic markov decision processes.

Fairness in Reinforcement Learning with Bisimulation Metrics Scalable methods for computing state similarity in deterministic markov decision processes

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.411004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.411004Z digest=sha256:40745d444fe256af5132d6cd0c8ffd73d9e5fdd8895deaec4cdfbf0ef1565b8a

Observation 77aa1244-33fe-4564-91bc-d374cabcef8f · outbound

This paper cites Scalable co-optimization of morphology and control in embodied machines.

Fairness in Reinforcement Learning with Bisimulation Metrics Scalable co-optimization of morphology and control in embodied machines

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.469699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.414402Z digest=sha256:5c397a5cf9d2e76f4c1fa88f66deba86e494c50b965a16089500a7088d534741

Observation de2e77ea-19df-4bd5-b535-f52411e0b393 · outbound

This paper cites Heuristic-guided reinforcement learning.

Fairness in Reinforcement Learning with Bisimulation Metrics Heuristic-guided reinforcement learning

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.458485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.417749Z digest=sha256:60571310a8ba1dece7afcc3fe9d86e47a8826161eb899055aaebe6d8a2129083

Observation e21549f5-21a5-4006-ba2f-3fb9de7ec36b · outbound

This paper cites Intrinsically motivated reinforcement learning.

Fairness in Reinforcement Learning with Bisimulation Metrics Intrinsically motivated reinforcement learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.421212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.421212Z digest=sha256:cde8cc58d3ab862fa34d27407516bde970336f281b888135cc52992d1bb3df5c

Observation 3e64b078-22e1-4c98-ac77-9b6795fa1a19 · outbound

This paper cites On welfare-centric fair reinforcement learning.

Fairness in Reinforcement Learning with Bisimulation Metrics On welfare-centric fair reinforcement learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.440364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.424773Z digest=sha256:ab2493838c3c5612903d8b22521a9907473cb7f89cfff634ac003a2365628103

Observation 018e90d6-d410-4237-a0ff-ad5a1d0622d5 · outbound

This paper cites Fairness is not static: deeper understanding of long term fairness via simulation studies.

Fairness in Reinforcement Learning with Bisimulation Metrics Fairness is not static: deeper understanding of long term fairness via simulation studies

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.428789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.428234Z digest=sha256:82eff5b7ac90c2f1429af15a97e9f6860723d71863dba35d5a06b25b602885be

Observation e17d4839-1240-42b8-bb31-575d68f51706 · outbound

This paper cites What hides behind unfairness? exploring dynamics fairness in reinforcement learning.

Fairness in Reinforcement Learning with Bisimulation Metrics What hides behind unfairness? exploring dynamics fairness in reinforcement learning

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.416730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.431763Z digest=sha256:f7dd4390a2aad17e968ee7d485379bd0099506dceeb816a73d3bac0e21c97be1

Observation e40f814f-0ff9-47e5-b51e-91c12749cf16 · outbound

This paper cites Desharnais, A.

Fairness in Reinforcement Learning with Bisimulation Metrics Desharnais, A

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.405812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.435162Z digest=sha256:524b3c2e55ec4d8792bd3e100a66f3dcdcd530e74cb36e945af0cce6d1d4f6d9

Observation 68539439-689e-4d93-8302-327f375fb32a · outbound

This paper cites Dynamic potential-based reward shaping.

Fairness in Reinforcement Learning with Bisimulation Metrics Dynamic potential-based reward shaping

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.394372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.438591Z digest=sha256:2be1bb23cba2cc0d6f0a886f23030df92112029efe4368f7dbb75e4093c925ab

Observation dfc113fb-f37c-4894-95bb-2eb286d7e179 · outbound

This paper cites Contextual bandits with concave rewards, and an application to fair ranking.

Fairness in Reinforcement Learning with Bisimulation Metrics Contextual bandits with concave rewards, and an application to fair ranking

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.383316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.441924Z digest=sha256:9c7e50eadc168429c8a3c4ee975cb778d1faf7e63df650e477f5dc616f20b758

Observation 6015a5da-0876-4b24-8387-c15e90fe5665 · outbound

This paper cites On the analysis of the (1+ 1) evolutionary algorithm.

Fairness in Reinforcement Learning with Bisimulation Metrics On the analysis of the (1+ 1) evolutionary algorithm

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.372085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.445407Z digest=sha256:fac14fd923fdc0df526589a1b7f939be958648c4729a84c569560d396572aef3

Observation a440ef3b-2204-4647-ade1-9342474773a2 · outbound

This paper cites Fairness through awareness.

Fairness in Reinforcement Learning with Bisimulation Metrics Fairness through awareness

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.448752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.448752Z digest=sha256:951bf5cd9455c8a6603557518577920de736a0f79e0c3408dec09f3d5e6e75db

Observation d24e34fc-e44e-4707-a555-4105ea8f3381 · outbound

This paper cites Stochastic spatio-temporal optimization for control and co-design of systems in robotics and applied physics.

Fairness in Reinforcement Learning with Bisimulation Metrics Stochastic spatio-temporal optimization for control and co-design of systems in robotics and applied physics

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.354290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.452385Z digest=sha256:2d6a656febeea2e82648f7f066f31f5191a5d2cc241afd92eca4d0783e254ac0

Observation c0f2d690-f4a4-4288-8536-557f9bb25d30 · outbound

This paper cites Fair lending implications of credit scoring systems.

Fairness in Reinforcement Learning with Bisimulation Metrics Fair lending implications of credit scoring systems

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.343313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.455745Z digest=sha256:140662412b09a5999068e7209ba4a668bf44350609f785aa6191dd4095bc53d4

Observation 45e30fdc-de03-4b4c-af91-ee7db85194ae · outbound

This paper cites Metrics for finite M arkov decision processes.

Fairness in Reinforcement Learning with Bisimulation Metrics Metrics for finite M arkov decision processes

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.332182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.459042Z digest=sha256:5ea40ab5cff6eed29d2518c587e1a88cd5c675e5e1b803bcacb401f0a293cd19

Observation 4298556d-ce92-44e1-b8ac-c7c767e4b4d0 · outbound

This paper cites Bisimulation metrics for continuous M arkov decision processes.

Fairness in Reinforcement Learning with Bisimulation Metrics Bisimulation metrics for continuous M arkov decision processes

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.321764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.462295Z digest=sha256:60b64c1cd48efbc4a2eb97ecb16da83ef9154a082fe157309ce39a1a13ccd7fc

Observation 027d792f-2519-4a5b-a7b0-ff92f5d1955c · outbound

This paper cites Fair off-policy learning from observational data.

Fairness in Reinforcement Learning with Bisimulation Metrics Fair off-policy learning from observational data

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.311397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.465500Z digest=sha256:765bd10c9b5d3e3251b03b367ed58ccfe7b7c2f5c5217d8a69350bb57e77288e

Observation 33a1f1b5-340e-45b6-898c-19c898d97b54 · outbound

This paper cites Potential based reward shaping for hierarchical reinforcement learning.

Fairness in Reinforcement Learning with Bisimulation Metrics Potential based reward shaping for hierarchical reinforcement learning

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.301066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.468866Z digest=sha256:32c5d7b2a62584c6aec5226969f4cb4fd497ea05f3f41579c06afa9ed1df1cb4

Observation efff686f-9f30-4338-aa4d-f01c18bd424a · outbound

This paper cites Equivalence notions and model minimization in M arkov decision processes.

Fairness in Reinforcement Learning with Bisimulation Metrics Equivalence notions and model minimization in M arkov decision processes

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.290675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.472121Z digest=sha256:258e9d93b0d91f19840d6fe425dc31131606964d7a9b23861e79d3b9d72f9cfc

Observation d5d8e10d-a046-4c28-85e7-263195f84c8f · outbound

This paper cites Bisimulation makes analogies in goal-conditioned reinforcement learning.

Fairness in Reinforcement Learning with Bisimulation Metrics Bisimulation makes analogies in goal-conditioned reinforcement learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.475324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.475324Z digest=sha256:71f9b9e0a00b4afc8845eb7f241ba71986a875ceaec682dbf24020d5b9132273

Observation 6547bbba-f87e-43fa-9b07-b068d42f1a7b · outbound

This paper cites Strategic classification.

Fairness in Reinforcement Learning with Bisimulation Metrics Strategic classification

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.273052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.479064Z digest=sha256:fcabcff2cac3d3b9db3e6ca54fce601ad44f94bad3cf8f45abb36a2a723f345b

Observation 2fe2f097-a439-428c-976c-ba772a438ad9 · outbound

This paper cites Equality of opportunity in supervised learning.

Fairness in Reinforcement Learning with Bisimulation Metrics Equality of opportunity in supervised learning

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.261694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.482689Z digest=sha256:18b5ff2d477e1152d2b8153ef078a7dd6510201139be483763de19269ecb7e63

Observation 604356bb-0226-43b6-9aa9-087180664b42 · outbound

This paper cites Fair algorithms for multi-agent multi-armed bandits.

Fairness in Reinforcement Learning with Bisimulation Metrics Fair algorithms for multi-agent multi-armed bandits

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.251213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.486134Z digest=sha256:b4c3294ba2fa91922cbd26be1bfe2b966952287fabc13f94a83af2e3e4154fa8

Observation ef851973-ec53-47eb-868d-0baaa048e5f0 · outbound

This paper cites Achieving long-term fairness in sequential decision making.

Fairness in Reinforcement Learning with Bisimulation Metrics Achieving long-term fairness in sequential decision making

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.240680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.489415Z digest=sha256:9ce5e6f14b2678effab41b325998012735b178ad9761d111d8de697c62ce337d

Observation f499aca9-5141-41b5-b1ca-3181cd72019a · outbound

This paper cites Striking a balance in fairness for dynamic systems through reinforcement learning.

Fairness in Reinforcement Learning with Bisimulation Metrics Striking a balance in fairness for dynamic systems through reinforcement learning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.229383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.492784Z digest=sha256:5b78b3a9e3a90c2564fe0d41baa89ea0fb7b670c853ae3d5a0dad3197b299c50

Observation 76555383-c82e-4a9f-aefc-482f616829df · outbound

This paper cites Chainqueen: A real-time differentiable physical simulator for soft robotics.

Fairness in Reinforcement Learning with Bisimulation Metrics Chainqueen: A real-time differentiable physical simulator for soft robotics

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.219037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.496342Z digest=sha256:f31dcefc9cf055dac6ede3b86b4eb053b81bf3a727bff277629762571e6172a2

Observation 7255bab3-3bef-4825-a380-df08522d1918 · outbound

This paper cites Learning to utilize shaping rewards: A new approach of reward shaping.

Fairness in Reinforcement Learning with Bisimulation Metrics Learning to utilize shaping rewards: A new approach of reward shaping

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.499741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.499741Z digest=sha256:bf45317c56b6829466535e5f81597489d5cec371f73ba3ba05d6b664ff15f499

Observation 37100529-5b97-43cf-a046-3f068638cb16 · outbound

This paper cites an unresolved cited work.

Fairness in Reinforcement Learning with Bisimulation Metrics Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-11T05:52:35.200638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.503030Z digest=sha256:5157581937dc0d6c91f2f27ad6dee333a291d267aa917c9b834dda14cdd7b508

Observation 0fca4401-47b8-45a1-8279-13415a90b445 · outbound

This paper cites Fairness in reinforcement learning.

Fairness in Reinforcement Learning with Bisimulation Metrics Fairness in reinforcement learning

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.189859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.506755Z digest=sha256:e9e65f8fd47ff7a9b9e28ede4b6aca1e956add76f77c27d76b2346ad8e5e8383

Observation 81e6d37e-d43d-447f-be37-819957396c2a · outbound

This paper cites Learning fairness in multi-agent systems.

Fairness in Reinforcement Learning with Bisimulation Metrics Learning fairness in multi-agent systems

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.178482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.510281Z digest=sha256:f7c8bd112658d6769719d50d47c7ed32f624c7fe09871ccf6687e25a6a5213c2

Observation f3e4f1ac-e558-41c6-951c-cbc01416d8df · outbound

This paper cites Fairness in learning: Classic and contextual bandits.

Fairness in Reinforcement Learning with Bisimulation Metrics Fairness in learning: Classic and contextual bandits

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.167654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.513673Z digest=sha256:cc7d187efbbf4eec39bf5e7b90cfcfd376e8d1848068f4edbd0207a8ce2d69b1

Observation 8eacde2d-cf0a-4a88-9d77-5dc8b05a3aa8 · outbound

This paper cites Achieving Fairness in Multi-Agent Markov Decision Processes Using Reinforcement Learning.

Fairness in Reinforcement Learning with Bisimulation Metrics Achieving Fairness in Multi-Agent Markov Decision Processes Using Reinforcement Learning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.516784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.516784Z digest=sha256:7e9fd7491178076c7217425381fa43876ae73740811874f6c3ee6d0c4b2ca2de

Observation 563b420d-55f7-4ec3-8a68-ad70a67c57d5 · outbound

This paper cites Stochastic hillclimbing as a baseline method for evaluating genetic algorithms.

Fairness in Reinforcement Learning with Bisimulation Metrics Stochastic hillclimbing as a baseline method for evaluating genetic algorithms

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.156941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.520683Z digest=sha256:8f4d9f8e9c3abdf47089b767e6b400e6f01a2ccdb721db6e172c3eaec73e4c07

Observation acebd42a-7c64-41bb-8af5-cade291f5878 · outbound

This paper cites Towards robust bisimulation metric learning.

Fairness in Reinforcement Learning with Bisimulation Metrics Towards robust bisimulation metric learning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.524048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.524048Z digest=sha256:9084995203924b24fce63c88a19745357356a866c9eb64e5cfa58553577c64c9

Observation 7f6c5455-b828-410c-be16-57942f3bd8b5 · outbound

This paper cites The long road to fairer algorithms.

Fairness in Reinforcement Learning with Bisimulation Metrics The long road to fairer algorithms

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.139615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.530887Z digest=sha256:01efdce7d0c4655d0efa1b70b0d6ddab31b1f89b870f718f96098e56a6bef173

Observation 2604f7b8-6b6e-4a8b-8cf5-474e457f3464 · outbound

This paper cites Larsen and Arne Skou.

Fairness in Reinforcement Learning with Bisimulation Metrics Larsen and Arne Skou

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.534330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.534330Z digest=sha256:6ae47cd34a56aa4a9c420937843ed451495686c8768338af46372367d37944e7

Observation ae2c3c62-91d6-40b1-8726-dbfe94c262cb · outbound

This paper cites Delayed impact of fair machine learning.

Fairness in Reinforcement Learning with Bisimulation Metrics Delayed impact of fair machine learning

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.128957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.537727Z digest=sha256:dbb4df4dde448ff927ad5adcca86f0a61737fee29865f407ca0828bef3b98359

Observation 672741c7-ca96-4c49-bdf5-2de0a428f434 · outbound

This paper cites Calibrated Fairness in Bandits.

Fairness in Reinforcement Learning with Bisimulation Metrics Calibrated Fairness in Bandits

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.540891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.540891Z digest=sha256:1a7397f879d12f1c2cba8c702e393a69cd5d24bd88a94c9362cb5866e2ba8e6f

Observation db2aa9f0-fb47-4abf-86c2-0208746b6fb8 · outbound

This paper cites Diffaqua: A differentiable computational design pipeline for soft underwater swimmers with shape interpolation.

Fairness in Reinforcement Learning with Bisimulation Metrics Diffaqua: A differentiable computational design pipeline for soft underwater swimmers with shape interpolation

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.117929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.544615Z digest=sha256:f1a8efd549b82901085c1f00c8c4894e1a404feecd13fb0732a12b9b91710b08

Observation cca406cb-314b-4baf-b432-62cb5f4153e4 · outbound

This paper cites Socially Fair Reinforcement Learning.

Fairness in Reinforcement Learning with Bisimulation Metrics Socially Fair Reinforcement Learning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.548025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.548025Z digest=sha256:7d6b64af180ed90628786afbf5516fdb877c88f0874556d292bd21dc383aeb8f

Observation 9a615218-8c20-46bf-b908-cdee165c86f0 · outbound

This paper cites Comparison of three methods for selecting values of input variables in the analysis of output from a computer code.

Fairness in Reinforcement Learning with Bisimulation Metrics Comparison of three methods for selecting values of input variables in the analysis of output from a computer code

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.106687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.551680Z digest=sha256:cb3c3a6ed9116a68d552c68883c5868e65ef6b7d9eec3d0f3d75095257d87171

Observation 6582a8f3-68ff-46ae-a918-7f75afd33470 · outbound

This paper cites A survey on bias and fairness in machine learning.

Fairness in Reinforcement Learning with Bisimulation Metrics A survey on bias and fairness in machine learning

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.555064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.555064Z digest=sha256:58a3ce91d31eaf5891d1e6ba5aabfe700a7c5acf0a018cc861b8e9f828c654be

Observation f1d2312d-9652-4df2-888a-583b489866db · outbound

This paper cites Offline contextual bandits with high probability fairness guarantees.

Fairness in Reinforcement Learning with Bisimulation Metrics Offline contextual bandits with high probability fairness guarantees

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.558261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.558261Z digest=sha256:73a88b336901eab1d83230feb4e74947ff65e06a53c22abafd452b667520b38a

Observation 6916ca90-db6d-4877-b74b-9deb9f5dfa02 · outbound

This paper cites The social cost of strategic classification.

Fairness in Reinforcement Learning with Bisimulation Metrics The social cost of strategic classification

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.083963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.561791Z digest=sha256:46996cb8ec49327acf957800617e784293f1610a0b089baf5e4d8168611c9917

Observation 9b18da54-cfb0-419d-9aee-ce878b9e3475 · outbound

This paper cites Human-level control through deep reinforcement learning.

Fairness in Reinforcement Learning with Bisimulation Metrics Human-level control through deep reinforcement learning

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.565003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.565003Z digest=sha256:b31a5e2d83d6e266578fdb066d467e29e023481c9b658dc999852dd80d8b092b

Observation 2b72bd8f-d4be-47ed-a13c-ea00c34449a9 · outbound

This paper cites Learning optimal fair policies.

Fairness in Reinforcement Learning with Bisimulation Metrics Learning optimal fair policies

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.066073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.568469Z digest=sha256:cefe78a9fd1eb66d67e283a03d498658408e6d72eabdf502c3769c7ed9196fe9

Observation b057f933-293e-4b24-b5f9-ca2669e7d8b8 · outbound

This paper cites Fairness and Sequential Decision Making: Limits, Lessons, and Opportunities.

Fairness in Reinforcement Learning with Bisimulation Metrics Fairness and Sequential Decision Making: Limits, Lessons, and Opportunities

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.571641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.571641Z digest=sha256:2f92e4fb476868026b7580494b434c5b7fb9813dce848809c6bcf101503c64fe

Observation 958d15ae-4d60-4f11-9a43-4abbf61f99cb · outbound

This paper cites Policy invariance under reward transformations: Theory and application to reward shaping.

Fairness in Reinforcement Learning with Bisimulation Metrics Policy invariance under reward transformations: Theory and application to reward shaping

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.054629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.575199Z digest=sha256:74bdd6509b88e5b0dd43285d2904c0bfa83dff7e42612a14caf54a6c3632ea9d

Observation 12fa267c-7704-47b1-855f-340b1c63aa49 · outbound

This paper cites Labelled Markov Processes.

Fairness in Reinforcement Learning with Bisimulation Metrics Labelled Markov Processes

Reference 57

Resolution
malformed identifier
no resolver link, observed 2026-08-11T05:52:34.578419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.578419Z digest=sha256:74f561e8d476ce761502b9d9d16a3a4c4f3030e205745cd15e674b096b834faf

Observation 227fa880-dd24-4096-9bf4-b2b443f24461 · outbound

This paper cites Discrimination-aware data mining.

Fairness in Reinforcement Learning with Bisimulation Metrics Discrimination-aware data mining

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.043060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.581682Z digest=sha256:4b197b57a9f0a0dea8663466bb4f5b8555ab25d82e111562fcc8e0fbd9b5df15

Observation 9c355279-28eb-4a24-a10f-d9db22404c35 · outbound

This paper cites Group fairness in reinforcement learning.

Fairness in Reinforcement Learning with Bisimulation Metrics Group fairness in reinforcement learning

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.032102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.584975Z digest=sha256:787ecee046da20ad7e5a5dd73400022f715e8c0547a78fffc78814b604830460

Observation 71d2a4bc-c032-444e-b40b-e1da08252dd2 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Fairness in Reinforcement Learning with Bisimulation Metrics Proximal Policy Optimization Algorithms

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.588391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.588391Z digest=sha256:1207b602ea21d95ad45c59882983aac5115266efa4ab071051150a8753869eb3

Observation 828ee0f1-16d0-481a-b435-90f21212f4b2 · outbound

This paper cites Learning fair policies in multi-objective (deep) reinforcement learning with average and discounted rewards.

Fairness in Reinforcement Learning with Bisimulation Metrics Learning fair policies in multi-objective (deep) reinforcement learning with average and discounted rewards

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.020478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.591948Z digest=sha256:97d255ab466c118d60cfb696ecab1c51344be0e7803cb400cf88ddd30038f505

Observation 02a43e17-3690-46f4-8432-626bdc27193d · outbound

This paper cites Intrinsically motivated reinforcement learning: An evolutionary perspective.

Fairness in Reinforcement Learning with Bisimulation Metrics Intrinsically motivated reinforcement learning: An evolutionary perspective

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.595051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.595051Z digest=sha256:e121286699cfcf2740b409ff64633a586de8868c2bdfdc3d39256217a1a89d30

Observation 7338ed5d-4294-4d10-b500-e383fd46f4b8 · outbound

This paper cites Reward design via online gradient ascent.

Fairness in Reinforcement Learning with Bisimulation Metrics Reward design via online gradient ascent

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:35.002760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.598261Z digest=sha256:8d69759de91073f14e605e36a711fc74f62afb6a02942e239acf1ad57d51964b

Observation c558dda3-d553-4a8c-be6d-f39219f12d20 · outbound

This paper cites Learning-in-the-loop optimization: End-to-end control and co-design of soft robots through learned deep latent representations.

Fairness in Reinforcement Learning with Bisimulation Metrics Learning-in-the-loop optimization: End-to-end control and co-design of soft robots through learned deep latent representations

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:34.992298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.601450Z digest=sha256:89348be7b3faa14a466e2c2adff2eeab13a13fcb76cbfa7364c1b4632392685e

Observation d8b741b7-5580-4702-9931-be0d861e5167 · outbound

This paper cites Co-learning of task and sensor placement for soft robotics.

Fairness in Reinforcement Learning with Bisimulation Metrics Co-learning of task and sensor placement for soft robotics

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:34.982079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.604699Z digest=sha256:e15e7d4cc32f451338e8fc40321c3678091a05415bddbb3e5daf9840d2ba0df1

Observation 1d33876f-b009-424a-959c-4010447de562 · outbound

This paper cites Terry, Ariel Kwiatkowski, John U.

Fairness in Reinforcement Learning with Bisimulation Metrics Terry, Ariel Kwiatkowski, John U

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.607949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.607949Z digest=sha256:f1e3c8103a8f2f001a2ee5bd9afc18ef0c3b7c2a5d754995dccb11852352912b

Observation 207663b5-a61c-4d51-aef4-423a0e17f5fc · outbound

This paper cites SoftZoo: A Soft Robot Co-design Benchmark For Locomotion In Diverse Environments.

Fairness in Reinforcement Learning with Bisimulation Metrics SoftZoo: A Soft Robot Co-design Benchmark For Locomotion In Diverse Environments

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.611080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.611080Z digest=sha256:f92179e07a9302fcf2220ee3f5963b874854b3baccc4920bf0499afdc552c108

Observation b0b87138-4e7a-4e60-a0e6-0a691051d6fd · outbound

This paper cites Curriculum-based co-design of morphology and control of voxel-based soft robots.

Fairness in Reinforcement Learning with Bisimulation Metrics Curriculum-based co-design of morphology and control of voxel-based soft robots

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:34.965107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.614665Z digest=sha256:92f695bcaa0c823dc730a95cbe41a2316f5f5a4e14317dd4563a5448d9a14398

Observation 607b0b26-0dde-4edf-ae08-a5305aeb46a3 · outbound

This paper cites Algorithms for fairness in sequential decision making.

Fairness in Reinforcement Learning with Bisimulation Metrics Algorithms for fairness in sequential decision making

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:34.953237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.617871Z digest=sha256:312d712aa29661963e3a3495f01b1d3dca68587f89d0ac9d902aab8b00da207c

Observation 6510edc4-28ee-47bd-8896-0459df7aa8ba · outbound

This paper cites Adapting static fairness to sequential decision-making: Bias mitigation strategies towards equal long-term benefit rate.

Fairness in Reinforcement Learning with Bisimulation Metrics Adapting static fairness to sequential decision-making: Bias mitigation strategies towards equal long-term benefit rate

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:34.940544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.620998Z digest=sha256:65d7ba33a7578f4634725f7695c0606c684ac57f513ffadaef13bd6dc4625f61

Observation a2b32714-2611-49f8-822c-c8b7daa03403 · outbound

This paper cites Long-Term Fairness with Unknown Dynamics.

Fairness in Reinforcement Learning with Bisimulation Metrics Long-Term Fairness with Unknown Dynamics

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.624290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.624290Z digest=sha256:36e6e7f066e753571081cfc43e6aaacedaee1e73111cc303414f55838b101f46

Observation 4041e368-afff-49ba-acd5-d3b6b9633331 · outbound

This paper cites Policy optimization with advantage regularization for long-term fairness in decision systems.

Fairness in Reinforcement Learning with Bisimulation Metrics Policy optimization with advantage regularization for long-term fairness in decision systems

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:34.929988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.627798Z digest=sha256:6535067f0a7bb0dba7418323890e60e60b67f59ee95dc44638bdf4f3c338c5c8

Observation a84f241b-a6f1-49b9-b7d6-b709696d4a68 · outbound

This paper cites Fair deep reinforcement learning with preferential treatment.

Fairness in Reinforcement Learning with Bisimulation Metrics Fair deep reinforcement learning with preferential treatment

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:34.919358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.630940Z digest=sha256:1217168b5204e1ea75844dbd71ccb9497dd749aaead7019d7c07a07f745a631c

Observation abd1e39f-b157-41d4-a758-4c5b1c2eb8b6 · outbound

This paper cites Learning invariant representations for reinforcement learning without reconstruction.

Fairness in Reinforcement Learning with Bisimulation Metrics Learning invariant representations for reinforcement learning without reconstruction

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:34.908328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.634205Z digest=sha256:fc2c6aca5a316fd52093b781ae3ca50f8783206b31e3071488bf131e3e1875bc

Observation ad3fa569-d953-48a6-a905-30fb8193401f · outbound

This paper cites Fairness in multi-agent sequential decision-making.

Fairness in Reinforcement Learning with Bisimulation Metrics Fairness in multi-agent sequential decision-making

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:52:34.896991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T05:52:34.637518Z digest=sha256:e79b053bb96353da7b85d195f75489eccfc8c8833535fbe3f5be1f091c7dbc28

Observation c90900d6-863e-49d1-bce5-01313327e6c3 · outbound

This paper cites On learning intrinsic rewards for policy gradient methods.

Fairness in Reinforcement Learning with Bisimulation Metrics On learning intrinsic rewards for policy gradient methods

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.640910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.640910Z digest=sha256:bd436de670b88260a6808a8b861c999511e664f5afbcef00aa3d9c738e1c2133

Observation e8f73d1e-b208-4498-8fe5-8cfa0dde2b89 · outbound

This paper cites @esa (Ref.

Fairness in Reinforcement Learning with Bisimulation Metrics @esa (Ref

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.644427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.644427Z digest=sha256:29ed3c8f9f5dc374387296bafd332b1375de0541cc03f6404bc8ec3760e1d62c

Observation 2f189ba7-fa3f-4e7f-91cb-ba2aa11a04e7 · outbound

This paper cites an unresolved cited work.

Fairness in Reinforcement Learning with Bisimulation Metrics Unresolved cited work

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.648305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.648305Z digest=sha256:ef13203111885212cf2c463054d8ac6f563cd3b21f8bf75ffd6927cd745f40d0

Observation 3abb6f3d-097f-4f6c-83d1-180ffcfc1537 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Fairness in Reinforcement Learning with Bisimulation Metrics Adam: A Method for Stochastic Optimization

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-11T05:52:34.651783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:52:34.651783Z digest=sha256:edb092b108710e5e61364ee59c39539389e31542c9d8846b60750dd5966236ed

Pith citing papers

Observation bc614d78-5ac2-4f14-a38f-210be06bfaaa · inbound

Long-term Fairness with Selective Labels cites this paper.

Long-term Fairness with Selective Labels Fairness in Reinforcement Learning with Bisimulation Metrics

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:51:16.233838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-22T07:48:56.042653Z digest=sha256:a645c3e20c10fe37c465651497711cadec4678023eaaa22dfbadefda1a5ec524