Pith. sign in

Paper Citation Record · LEDGER

A Survey of Reinforcement Learning For Economics

As of 10 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 1 inbound Pith citation observation for arXiv:2603.08956.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2603.08956 v6

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T18:34:55.781594Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-10T18:33:50.296933Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T00:25:50.044346Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact3
  • verified fuzzy0
  • unresolved39
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 71a340cb-2ad4-4f7f-af90-c4e18d27a645 · outbound

This paper cites Thompson Sampling for Dynamic Pricing.

A Survey of Reinforcement Learning For Economics Thompson Sampling for Dynamic Pricing

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:52.163280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:52.163280Z digest=sha256:1c9f60e4e48738e786b3da445125c18ba512c0dea07aa7e98655c2e43ad89249

Observation 6c5f1fdb-cba4-4d90-897b-6322d092428b · outbound

This paper cites Deep Reinforcement Learning from Self-Play in Imperfect-Information Games.

A Survey of Reinforcement Learning For Economics Deep Reinforcement Learning from Self-Play in Imperfect-Information Games

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:52.624543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:52.624543Z digest=sha256:9c243f3a8950517d76361305d9e344bef9e440598ef9b2769c8184edc6117bd5

Observation 927b2e5c-82e4-448e-bd24-903a4f5f2053 · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

A Survey of Reinforcement Learning For Economics Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:53.020591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:53.020591Z digest=sha256:51241cfce69c684c987515da09633bc5e80848f6142d019ec17668b8d6cd4180

Observation cc626079-6674-4405-84d6-62ddfef4638c · outbound

This paper cites Is q-learning minimax optimal? a tight sample complexity analysis.Operations Research, 72(1), 2024a.

A Survey of Reinforcement Learning For Economics Is q-learning minimax optimal? a tight sample complexity analysis.Operations Research, 72(1), 2024a

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:53.143016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:53.143016Z digest=sha256:07f75bde8d4a5e5ae075ec5e2264fc40c229585828f8096e7a018c3a564c6358

Observation 0f92ec82-d915-44e7-9f10-cd44bb304454 · outbound

This paper cites Clare Lyle, Mark Rowland, and Will Dabney.

A Survey of Reinforcement Learning For Economics Clare Lyle, Mark Rowland, and Will Dabney

Reference 23

Resolution
verified exact
doi, observed 2026-08-02T18:38:27.564433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-02T18:34:53.587179Z digest=sha256:788ed562d8de41e2af3d0b861a12cf59483cbfd4c7f3c8e8a30a1ef0480b39f3

Observation 4669983a-3fea-4362-8fcd-17d4fc19d61b · outbound

This paper cites Empirical Design in Reinforcement Learning.

A Survey of Reinforcement Learning For Economics Empirical Design in Reinforcement Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:53.949455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:53.949455Z digest=sha256:0c7dbdc9c88af58b9c7dd0df22f544f00c3e66deb1ba3a5fe22c845b96d1eef2

Observation 119e6291-060f-4239-901b-42e80bf3633a · outbound

This paper cites an unresolved cited work.

A Survey of Reinforcement Learning For Economics Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:54.061619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:54.061619Z digest=sha256:d7ea507b2c2e94ad2facb0a92d4cf2c481f7cd1109b3d43bdcd76bf05ae3cdb2

Observation 5f2cd22e-e456-4a26-8948-b6b7bd4decee · outbound

This paper cites Proximal Policy Optimization Algorithms.

A Survey of Reinforcement Learning For Economics Proximal Policy Optimization Algorithms

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:54.308125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:54.308125Z digest=sha256:66d7fa2b28e6f0f778e85d375751fcd9784646222afc1d9674791f4d06f8b625

Observation e7167921-8f51-4979-bdf3-c24f93511b12 · outbound

This paper cites Residual Policy Learning.

A Survey of Reinforcement Learning For Economics Residual Policy Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:54.495942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:54.495942Z digest=sha256:891a7a447439ba9c6ad8829fd9abf2109fe089e948243ebbeceb266003739ac8

Observation 64a11181-fc75-40e7-8191-253dfb935df7 · outbound

This paper cites Solving Large Imperfect Information Games Using CFR+.

A Survey of Reinforcement Learning For Economics Solving Large Imperfect Information Games Using CFR+

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:54.754901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:54.754901Z digest=sha256:406171d5fac8fefd135bc662f9c7a76ed7b7c48e35621c22d45373224e410dc5

Observation 1f82c125-154d-45af-aa97-ec4c8389fe68 · outbound

This paper cites Gymnasium: A Standard Interface for Reinforcement Learning Environments.

A Survey of Reinforcement Learning For Economics Gymnasium: A Standard Interface for Reinforcement Learning Environments

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:54.957047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:54.957047Z digest=sha256:992220c387819048a534e68af22ad009877813aaab0607d293d480afa9b5698e

Observation d5d4fc0e-c118-4390-83a9-0a0b68b9bc99 · outbound

This paper cites Near-Optimal Sample Complexity for Iterated CVaR Reinforcement Learning with a Generative Model.

A Survey of Reinforcement Learning For Economics Near-Optimal Sample Complexity for Iterated CVaR Reinforcement Learning with a Generative Model

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:55.252690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:55.252690Z digest=sha256:4af5b1c867d8a2103867f86081e3082319dee7c1b237e0bff5f80692e541b86e

Observation 6f426e85-277b-48ca-a7f0-571f93d4a738 · outbound

This paper cites Self-Rewarding Language Models.

A Survey of Reinforcement Learning For Economics Self-Rewarding Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:55.457454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:55.457454Z digest=sha256:c7d1f9e0bb927832eaf28e2c3cd6a76e0c207d0b908100128df531f8cbe37741

Observation 8bac77cc-6aac-4a84-ac73-212adbbfeb8b · outbound

This paper cites A Deeper Look at Experience Replay.

A Survey of Reinforcement Learning For Economics A Deeper Look at Experience Replay

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:55.617640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:55.617640Z digest=sha256:2f5f690efeeee6e341a402b68227d0d3de06e5d664bd3001dd9b40f46ac28fb7

Observation 1f67ffb0-3c62-4ccc-916e-d369ab71e3b1 · outbound

This paper cites Fine-Tuning Language Models from Human Preferences.

A Survey of Reinforcement Learning For Economics Fine-Tuning Language Models from Human Preferences

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:55.781594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:55.781594Z digest=sha256:79743a200f5f4b15367121716efba603dd3dc64a12d567bf774fdbeac6bdeaf9

Observation 98777ef0-3492-46fb-a102-4aa0c37129cb · outbound

This paper cites Santos and John Rust.

A Survey of Reinforcement Learning For Economics Santos and John Rust

Reference 1959

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:54.209538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:54.209538Z digest=sha256:cbb7fbfb344909dfdbc4ce26cb41121a1c066a661e38699b72ba004491185b71

Observation 3aa4ea9d-4aa4-4680-b70c-b827d8e4e9b5 · outbound

This paper cites RL with KL penalties is better viewed as Bayesian inference.

A Survey of Reinforcement Learning For Economics RL with KL penalties is better viewed as Bayesian inference

Reference 1960

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:52.804048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:52.804048Z digest=sha256:96c413aed811c6ca6e1a51966532d678b6527043e8c05300f79ed0b72b4c6225

Observation 22f541ec-4d3b-49de-87f0-d9c7e45661cd · outbound

This paper cites Classifying fermionic states via many-body correlation measures.

A Survey of Reinforcement Learning For Economics Classifying fermionic states via many-body correlation measures

Reference 1982

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:51.151228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:51.151228Z digest=sha256:53b8ebdeb59146be8535e605cc6568cdc1e1aeb52c4e06c8547ef555de1219c7

Observation 7bcf9018-48f8-4f49-94c2-872f28177592 · outbound

This paper cites Richard S.

A Survey of Reinforcement Learning For Economics Richard S

Reference 1988

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:54.651143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:54.651143Z digest=sha256:05059714727f64e2d80c7f1e58342e0b75826d09704b0a36e47418cd55a8bb87

Observation 91d8cf4c-8412-4a2f-9c12-9dd68b69c670 · outbound

This paper cites Markov games as a framework for multi-agent reinforcement learning.

A Survey of Reinforcement Learning For Economics Markov games as a framework for multi-agent reinforcement learning

Reference 1992

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:53.281062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:53.281062Z digest=sha256:803598a3aeab67d0458db8f58460e21fcc586efa9d695489ce17cac7b20d6526

Observation 07be33d8-1cc9-4968-873e-45541c0561bc · outbound

This paper cites First-order methods for Wasserstein distributionally robust MDP.

A Survey of Reinforcement Learning For Economics First-order methods for Wasserstein distributionally robust MDP

Reference 1993

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:52.317741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:52.317741Z digest=sha256:a74af1a2d4fcec415e797e4e1a8b13103b615c3ade6d861e3a906b5dd70329d3

Observation 00c241db-d497-4aa5-a9f6-3c4a36166254 · outbound

This paper cites an unresolved cited work.

A Survey of Reinforcement Learning For Economics Unresolved cited work

Reference 1994

Resolution
verified exact
doi, observed 2026-08-02T18:38:27.259490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-02T18:34:54.976781Z digest=sha256:32fd70586d5a4250455eb36cfa4cdc269de5ffbd73713ac1eb4c9cf21da08252

Observation 09a6e2d1-766a-4fb1-abcc-f68be9b8c1ed · outbound

This paper cites Learning to Solve Constraint Satisfaction Problems with Recurrent Transformer.

A Survey of Reinforcement Learning For Economics Learning to Solve Constraint Satisfaction Problems with Recurrent Transformer

Reference 1997

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:54.981380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:54.981380Z digest=sha256:a1636eab83865d8b9a619440153c34fa12c17eb6f7df441b70322167a2b21992

Observation f93d637a-29e8-444d-bf6e-0117c7a7071c · outbound

This paper cites Contextual Dynamic Pricing with Strategic Buyers.

A Survey of Reinforcement Learning For Economics Contextual Dynamic Pricing with Strategic Buyers

Reference 2001

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:53.382887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:53.382887Z digest=sha256:eca5a03814b077790b3e5d7e11041acbd5891768ea9e1cc2a4c1275170b7e575

Observation 1d0a0f84-62ca-4e8e-8e6d-8a31da115225 · outbound

This paper cites Core equality of real sequences.

A Survey of Reinforcement Learning For Economics Core equality of real sequences

Reference 2003

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:51.728721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:51.728721Z digest=sha256:a0ed7e34e5708e31490af7e771f8d330b3ddaa766d1ff752615bb3b00d339e9d

Observation c4c7435d-f718-4e02-8e66-00a09d950a12 · outbound

This paper cites Assessing Game Balance with AlphaZero: Exploring Alternative Rule Sets in Chess.

A Survey of Reinforcement Learning For Economics Assessing Game Balance with AlphaZero: Exploring Alternative Rule Sets in Chess

Reference 2006

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:54.851226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:54.851226Z digest=sha256:fcb0465a72d1d0f6e41eebe787af266ea5f5c71512bf93b576772179a34651f4

Observation 91e31647-01c9-4f35-8e62-6adecfc4f605 · outbound

This paper cites Peter Arcidiacono and Robert A.

A Survey of Reinforcement Learning For Economics Peter Arcidiacono and Robert A

Reference 2008

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:50.916350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:50.916350Z digest=sha256:52b73c85c376f769eee4ea81de107b6b26714516af61919d62f939d443a1a42b

Observation c587db8a-935b-4326-807c-d59b0b5ba6eb · outbound

This paper cites Deep Reinforcement Learning and the Deadly Triad.

A Survey of Reinforcement Learning For Economics Deep Reinforcement Learning and the Deadly Triad

Reference 2010

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:55.067775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:55.067775Z digest=sha256:e5cfd586ea97e97ada48ab9a45589a1a76ddfeff5a5a6a6fc79e64a7f6cbe1a8

Observation 6402ceee-ba27-4ed3-b4ff-d00b2dd5e9da · outbound

This paper cites Insight from the elliptic flow of identified hadrons measured in relativistic heavy-ion collisions.

A Survey of Reinforcement Learning For Economics Insight from the elliptic flow of identified hadrons measured in relativistic heavy-ion collisions

Reference 2011

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:55.353046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:55.353046Z digest=sha256:03492e6b9f6ce6f97063abc381d3519a235abde256825afa9046adfe37f80264

Observation 8e98d37c-c83d-4e75-a175-3f619ef22804 · outbound

This paper cites Fleming and William M.

A Survey of Reinforcement Learning For Economics Fleming and William M

Reference 2012

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:51.853335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:51.853335Z digest=sha256:dc5c04d33db460fe2ba9c2a5f8e99f5a9296a018f874202e9610c9de8515dc4c

Observation 15c6b402-2baf-4b60-8858-5b6bff1233a9 · outbound

This paper cites Asynchronous methods for deep reinforce- ment learning.

A Survey of Reinforcement Learning For Economics Asynchronous methods for deep reinforce- ment learning

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:53.696586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:53.696586Z digest=sha256:4caf81ff213c8cfe8c4e3f85f0af93919ae3387e50b24582896c281eef17226a

Observation cc756144-df6d-4d3c-914e-b05d665be10e · outbound

This paper cites Reinforcement Learning and Control as Probabilistic Inference: Tutorial and Review.

A Survey of Reinforcement Learning For Economics Reinforcement Learning and Control as Probabilistic Inference: Tutorial and Review

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:52.913040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:52.913040Z digest=sha256:c18ed37de7d9cb0ad49d8ad09e2cedaed85964d6d0c60cc7afa7165ab2886b15

Observation 446d8768-e16b-46ff-a6dd-6144eb50ec26 · outbound

This paper cites Music Source Separation with Band-split RNN.

A Survey of Reinforcement Learning For Economics Music Source Separation with Band-split RNN

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:50.651580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:50.651580Z digest=sha256:388cd6af717d7c972d4bcdf83252649fdf5b2b8f60f82b5756b036c285cbfb7e

Observation cecaf6e8-7df4-4e35-8a21-a56aa9cfac76 · outbound

This paper cites Off-policy deep reinforcement learning with- out exploration.

A Survey of Reinforcement Learning For Economics Off-policy deep reinforcement learning with- out exploration

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:51.982337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:51.982337Z digest=sha256:516fb8ab535161a9c4863b7c34c9fe5f4d9359ba2258c73c2de4a7cf0bb057ea

Observation 7a5f69ed-2c80-46e3-b39d-11e9649ee6c5 · outbound

This paper cites Policy Optimization for Constrained MDPs with Provable Fast Global Convergence.

A Survey of Reinforcement Learning For Economics Policy Optimization for Constrained MDPs with Provable Fast Global Convergence

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:53.495337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:53.495337Z digest=sha256:da1a3ab2505a9bab50b51cfe3e9c8758bfc9cfee9d37b11301c1d41c410d145c

Observation 35f2d2d3-0679-4b3b-becf-104b1167add1 · outbound

This paper cites Deep reinforcement learning: Emerging trends in macroe- conomics and future prospects.

A Survey of Reinforcement Learning For Economics Deep reinforcement learning: Emerging trends in macroe- conomics and future prospects

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:51.033340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:51.033340Z digest=sha256:aa98dafc8b1a23104a1f14ed8dac33fea6b0a6d4ec7285587e5c3b33f4748182

Observation 23280bfc-bf70-48a0-b34a-71d88ffa8a98 · outbound

This paper cites Generalizing across Temporal Domains with Koopman Operators.

A Survey of Reinforcement Learning For Economics Generalizing across Temporal Domains with Koopman Operators

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:51.455116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:51.455116Z digest=sha256:6799f5ecf1eda9e73757d9873d2f8be4ae5f335207a208d6470207b72e7f8056

Observation 560b2954-1697-4888-95b7-bb4c44c3845e · outbound

This paper cites Unifying causal reinforcement learning: Survey, taxonomy, algorithms and applications.arXiv preprint arXiv:2512.18135,.

A Survey of Reinforcement Learning For Economics Unifying causal reinforcement learning: Survey, taxonomy, algorithms and applications.arXiv preprint arXiv:2512.18135,

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:51.580275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:51.580275Z digest=sha256:c4da74ba4d066c9fe57f9b67e8922f7afe2e1e86efa6f9edf2175a0650149757

Observation 2f95a455-8d94-49ca-aa7f-ae9c74fce8f0 · outbound

This paper cites Lili Chen, Kevin Lu, Aravind Rajeswaran, Kimin Lee, Aditya Grover, Michael Laskin, Pieter Abbeel, Aravind Srinivas, and Igor Mordatch.

A Survey of Reinforcement Learning For Economics Lili Chen, Kevin Lu, Aravind Rajeswaran, Kimin Lee, Aditya Grover, Michael Laskin, Pieter Abbeel, Aravind Srinivas, and Igor Mordatch

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:51.271729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:51.271729Z digest=sha256:d6998bf4a0615a422a75ffe64de353a07b46df1c0df37f19d1fc5b16c7afac67

Observation 62fa3c21-54ae-41a9-b35c-980a993046a9 · outbound

This paper cites Strategic classifi- cation.

A Survey of Reinforcement Learning For Economics Strategic classifi- cation

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:52.497920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:52.497920Z digest=sha256:29a2e63add107f9abb0b17e283a9cb2c952b9adb825513808dcdd2c802679cc8

Observation 70ce1f1a-f9f1-4941-ab16-f50dfa0933c0 · outbound

This paper cites Jonas Mueller, Vasilis Syrgkanis, and Matt Taddy.

A Survey of Reinforcement Learning For Economics Jonas Mueller, Vasilis Syrgkanis, and Matt Taddy

Reference 2025

Resolution
verified exact
doi, observed 2026-08-02T18:38:27.393039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-02T18:34:53.848980Z digest=sha256:5bf8d120b4b2a75438a6e211e9cb23bc7c72a5372c7834213af60e4cafb06425

Observation 3e7c990f-f04f-48fe-b847-7e44fa88a3c9 · outbound

This paper cites Some asymptotic formulae for torsion in homotopy groups.

A Survey of Reinforcement Learning For Economics Some asymptotic formulae for torsion in homotopy groups

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:50.788053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:50.788053Z digest=sha256:21d074095043e99d1f39e62dee7bed98fc1b9c52489ed784f7cf63e1acacc257

Pith citing papers

Observation d15b34f2-af80-4e1c-8936-36870312e240 · inbound

The Theorems of Dr. David Blackwell and Their Contributions to Artificial Intelligence cites this paper.

The Theorems of Dr. David Blackwell and Their Contributions to Artificial Intelligence A Survey of Reinforcement Learning For Economics

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-07-28T02:23:27.763889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T18:33:50.296933Z digest=sha256:f111484bc2fcdc28310886a8db3c65d19bab43d3350991ff35767744871188e5