Pith. sign in

Paper Citation Record · LEDGER

Linear Mixture Distributionally Robust Markov Decision Processes

As of 8 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 0 inbound Pith citation observations for arXiv:2505.18044.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.18044 v1

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:44:16.530086Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

51 of 51 outbound references displayed

  • verified exact0
  • verified fuzzy43
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation dc186288-6cb8-4ced-95f5-554b82e2a0c9 · outbound

This paper cites Improved algorithms for linear stochastic bandits.Advances in Neural Information Processing Systems, 24, 2011.

Linear Mixture Distributionally Robust Markov Decision Processes Improved algorithms for linear stochastic bandits.Advances in Neural Information Processing Systems, 24, 2011

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:27.385792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:10.019563Z digest=sha256:9757a73288f832a3fd86613b59f56a902b872656490631e7e757a2e0e1499ac7

Observation bb835027-f4f9-4f05-93ca-afff8eca1798 · outbound

This paper cites Model-based rein- forcement learning with value-targeted regression.

Linear Mixture Distributionally Robust Markov Decision Processes Model-based rein- forcement learning with value-targeted regression

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:27.195035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:10.070077Z digest=sha256:1a7449a3d072fcbd7f8fda95f50526d8c840ae92c924610f82f8f4d9fd03e8f2

Observation 38a13ecd-1e1d-4364-8365-9ae7148b5933 · outbound

This paper cites Robust reinforcement learning using least squares policy iteration with provable performance guarantees.

Linear Mixture Distributionally Robust Markov Decision Processes Robust reinforcement learning using least squares policy iteration with provable performance guarantees

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:26.981351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:10.135324Z digest=sha256:994968e49c4aa436c35d33e2499a4d43391d2a873e598ce5cba456204b4b35c4

Observation 00956383-db3a-4030-bb26-9b74abc7e6f1 · outbound

This paper cites an unresolved cited work.

Linear Mixture Distributionally Robust Markov Decision Processes Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:44:26.789328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:10.254486Z digest=sha256:4384c6bf5207a418753b0f12c11e1cf929f7b00fc1864c84389f761496eea92d

Observation 1dd62c58-047c-4ffc-8b5b-baf919b0cf50 · outbound

This paper cites Provably efficient exploration in policy optimization.

Linear Mixture Distributionally Robust Markov Decision Processes Provably efficient exploration in policy optimization

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:26.541681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:10.322725Z digest=sha256:4d2747501f43ef36e4adae49a9d1508dc561c1de02f8fd9df3b5fc73ff3c8292

Observation 3f20851a-ba71-43a7-af1d-095e670e23a8 · outbound

This paper cites A Survey of Sim-to-Real Methods in RL: Progress, Prospects and Challenges with Foundation Models.

Linear Mixture Distributionally Robust Markov Decision Processes A Survey of Sim-to-Real Methods in RL: Progress, Prospects and Challenges with Foundation Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:10.442723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:10.442723Z digest=sha256:8e3d51aaacc0787ef65280e0396733a22703c08c274ab1ce92907a08ac0a4e5b

Observation 19a16ae0-0857-41d3-ae1b-e2da5c47ff67 · outbound

This paper cites Off-dynamics reinforcement learning: Training for transfer with domain classifiers.

Linear Mixture Distributionally Robust Markov Decision Processes Off-dynamics reinforcement learning: Training for transfer with domain classifiers

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:26.293774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:10.544737Z digest=sha256:61415d7f469602740e504a0faa14c0542b090e56ddc3d1b84700b4683cec53c0

Observation 54623bc2-edaa-4cb8-993d-1db9562a20cc · outbound

This paper cites Birkhauser Boston Inc., 1989.

Linear Mixture Distributionally Robust Markov Decision Processes Birkhauser Boston Inc., 1989

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:25.988409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:10.646720Z digest=sha256:7136817a564c51a1d5bc6cc05641f1c41734bb35ab775b30480a01aba48dcded

Observation 7ab5c91c-f782-4b02-9065-0bc0a743a499 · outbound

This paper cites Robust markov decision processes: Beyond rectangu- larity.Mathematics of Operations Research, 48(1):203–226, 2023.

Linear Mixture Distributionally Robust Markov Decision Processes Robust markov decision processes: Beyond rectangu- larity.Mathematics of Operations Research, 48(1):203–226, 2023

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:25.708718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:10.741252Z digest=sha256:0b443c97ad388dafa98ea5096e9e7963bbb8df870c64646cec4f9b5f7faeffe1

Observation 3fa3af11-b595-4805-931b-4d939d270587 · outbound

This paper cites Off-dynamics reinforcement learning via domain adaptation and reward augmented imitation.

Linear Mixture Distributionally Robust Markov Decision Processes Off-dynamics reinforcement learning via domain adaptation and reward augmented imitation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:25.473463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:10.836792Z digest=sha256:d741a9b05739622c7220343761abe1bd6de92d1c35f0b965dd13f9b49f40cdc3

Observation 8cf0cc27-d3e2-42cc-94fc-00e9988764f3 · outbound

This paper cites Kullback-leibler divergence constrained distributionally robust optimization.Available at Optimization Online, 1(2):9, 2013.

Linear Mixture Distributionally Robust Markov Decision Processes Kullback-leibler divergence constrained distributionally robust optimization.Available at Optimization Online, 1(2):9, 2013

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:25.114752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:10.941541Z digest=sha256:c5b1bfaca9fc350358461dff69aa5c03de4b25cbf54f06a39631a62024465d26

Observation 26f5d310-fae8-4dec-9df9-918a18143f24 · outbound

This paper cites Robust dynamic programming.Mathematics of Operations Research, 30(2): 257–280, 2005.

Linear Mixture Distributionally Robust Markov Decision Processes Robust dynamic programming.Mathematics of Operations Research, 30(2): 257–280, 2005

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:24.825016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:11.093164Z digest=sha256:c74f3a831b10c480a58a4398ff1bd518b94d5cb03e8fbd59d53dbd211de1e7db

Observation 1133a7aa-aac7-4269-9920-365395c561b4 · outbound

This paper cites Model-based reinforcement learning with value-targeted regression.

Linear Mixture Distributionally Robust Markov Decision Processes Model-based reinforcement learning with value-targeted regression

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:24.634036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:11.207448Z digest=sha256:a6b16c577a87d848d5032f769aca288912446e1263a426cd4052e46226dbd483

Observation 9681b5ac-3886-4ad9-a7e8-705f2973edb4 · outbound

This paper cites Reinforcement learning in robotics: A survey.

Linear Mixture Distributionally Robust Markov Decision Processes Reinforcement learning in robotics: A survey

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:24.415638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:11.315068Z digest=sha256:23ad71319574825e507a5724884f70d4a064f16ce45147a5b2fc1384fc3f964e

Observation a8f9b543-73ed-47b2-be81-feb4db80c716 · outbound

This paper cites The transferability approach: Crossing the reality gap in evolutionary robotics.IEEE Transactions on Evolutionary Computa- tion, 17(1):122–145, 2012.

Linear Mixture Distributionally Robust Markov Decision Processes The transferability approach: Crossing the reality gap in evolutionary robotics.IEEE Transactions on Evolutionary Computa- tion, 17(1):122–145, 2012

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:24.211608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:11.487182Z digest=sha256:1d46ada21ef7332c2368a307b3725c823905ddde13c4dbba8a4ad14b69c8ac77

Observation 537e3031-7b53-4ac4-84f6-ac234e66cc16 · outbound

This paper cites Improved algorithm for adversarial linear mixture mdps with bandit feedback and unknown transition.

Linear Mixture Distributionally Robust Markov Decision Processes Improved algorithm for adversarial linear mixture mdps with bandit feedback and unknown transition

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:23.992506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:11.642883Z digest=sha256:6381868b71fce6efd81e531505f1e69c5696b257886352b4483bd583cd64df56

Observation 9e4b0ec0-dad0-4e2b-8215-312c05ec1e7f · outbound

This paper cites Policy gradient algorithms for robust mdps with non-rectangular uncertainty sets.arXiv preprint arXiv:2305.19004, 2023.

Linear Mixture Distributionally Robust Markov Decision Processes Policy gradient algorithms for robust mdps with non-rectangular uncertainty sets.arXiv preprint arXiv:2305.19004, 2023

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:11.789167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:11.789167Z digest=sha256:6a2e0263866aa0d89652fea8daa24f97e021342d548e63cd99af6e43b5bb0a5a

Observation c0a9d7d5-2e5b-490a-9ce1-c4a8d716093d · outbound

This paper cites Distributionally robust off-dynamics reinforcement learning: Prov- able efficiency with linear function approximation.

Linear Mixture Distributionally Robust Markov Decision Processes Distributionally robust off-dynamics reinforcement learning: Prov- able efficiency with linear function approximation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:23.793901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:11.853527Z digest=sha256:6c45651ae6b603f851f6df9c6cb77c9e387eae7155b4e3fbc2802673d567123b

Observation 0dc062d7-014f-4736-813b-841708978fd2 · outbound

This paper cites Minimax optimal and computationally efficient algorithms for distributionally robust offline reinforcement learning.

Linear Mixture Distributionally Robust Markov Decision Processes Minimax optimal and computationally efficient algorithms for distributionally robust offline reinforcement learning

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:23.553624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:11.967284Z digest=sha256:fbcbff5603ef8002b50e7586c9dc6f4d98539cb3336c31dd72ed6faf0257f64a

Observation 63daded1-ac5d-49ab-99aa-5b63af275ba0 · outbound

This paper cites Upper and Lower Bounds for Distributionally Robust Off-Dynamics Reinforcement Learning.

Linear Mixture Distributionally Robust Markov Decision Processes Upper and Lower Bounds for Distributionally Robust Off-Dynamics Reinforcement Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:12.083195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:12.083195Z digest=sha256:73de6bc7fb76959985698aba08d16a7751a9b6b39661bf036ae49f6bbe196251

Observation 95e6eb6c-acaa-4371-9d67-69ea426d175c · outbound

This paper cites Distributionally robust reinforcement learning with interactive data collection: Fundamental hardness and near-optimal algorithms.

Linear Mixture Distributionally Robust Markov Decision Processes Distributionally robust reinforcement learning with interactive data collection: Fundamental hardness and near-optimal algorithms

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:23.385864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:12.199629Z digest=sha256:77f76033c9c5233556631d4736426cf48475e03c5f7cb05a3546ec4df5de14f4

Observation b85144d0-96dd-4589-b3c1-7a46c22f8c9f · outbound

This paper cites Distributionally Robust Offline Reinforcement Learning with Linear Function Approximation.

Linear Mixture Distributionally Robust Markov Decision Processes Distributionally Robust Offline Reinforcement Learning with Linear Function Approximation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:12.293596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:12.293596Z digest=sha256:faa99cbe52b48c3221d0294865466270d08fd26ae65caba89e733b01eee5a5e2

Observation 0c3ffa72-4d05-4171-ad13-ea9788ec4a94 · outbound

This paper cites Finite mixture models.Annual review of statistics and its application, 6(1):355–378, 2019.

Linear Mixture Distributionally Robust Markov Decision Processes Finite mixture models.Annual review of statistics and its application, 6(1):355–378, 2019

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:23.288836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:12.397984Z digest=sha256:4af00789b74160e00e7bbee53e2f689489fd4b03c4a684e4207a714bb4548c92

Observation 44bb74d5-08e3-42dd-ae8a-af04cf923cc2 · outbound

This paper cites A simplex method for function minimization.The computer journal, 7(4):308–313, 1965.

Linear Mixture Distributionally Robust Markov Decision Processes A simplex method for function minimization.The computer journal, 7(4):308–313, 1965

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:23.223487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:12.495205Z digest=sha256:1a04caffb419ec9b290d56655484b742fa666a491822c0dd6117795e18505191

Observation 67835c7f-4f71-4a3e-97e5-8b9fc42752ed · outbound

This paper cites Robust control of markov decision processes with uncertain transition matrices.Operations Research, 53(5):780–798, 2005.

Linear Mixture Distributionally Robust Markov Decision Processes Robust control of markov decision processes with uncertain transition matrices.Operations Research, 53(5):780–798, 2005

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:23.169848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:12.579883Z digest=sha256:d4064ef5eeba408363a5e62d6a407000fcd0f68c872b2beca333b126245039fb

Observation 839e7291-ab3a-497e-8e24-f86eb3b3790a · outbound

This paper cites Robustness in markov decision problems with uncertain transition matrices.Advances in Neural Information Processing Systems, 16, 2003.

Linear Mixture Distributionally Robust Markov Decision Processes Robustness in markov decision problems with uncertain transition matrices.Advances in Neural Information Processing Systems, 16, 2003

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:23.014790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:12.671049Z digest=sha256:30024b87fdea6ed14c65c074b816b5d79621f331e34e05a8bfcebe7fe14ac504

Observation 18d66843-81be-44cc-aa87-48dd55c772f0 · outbound

This paper cites Assessing Generalization in Deep Reinforcement Learning.

Linear Mixture Distributionally Robust Markov Decision Processes Assessing Generalization in Deep Reinforcement Learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:12.736776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:12.736776Z digest=sha256:245459e80cc15a0685dddd80c679e01fb6058b59bc903c4eb58dcba9d7c9bc28

Observation 4142db4e-26e7-4179-b4b9-ce97031e44e5 · outbound

This paper cites Bridging distributionally robust learning and offline rl: An approach to mitigate distribution shift and partial data coverage.

Linear Mixture Distributionally Robust Markov Decision Processes Bridging distributionally robust learning and offline rl: An approach to mitigate distribution shift and partial data coverage

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:22.701706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:12.810675Z digest=sha256:59dbaf0f730643a9286e7b4b69d4bbddee6c2aeaf628e2ab67c432ff437e4981

Observation e8fd1073-752e-49a0-92fe-9ccba380a9ab · outbound

This paper cites The infinite gaussian mixture model.Advances in Neural Information Processing Systems, 12, 1999.

Linear Mixture Distributionally Robust Markov Decision Processes The infinite gaussian mixture model.Advances in Neural Information Processing Systems, 12, 1999

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:22.533967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:12.922041Z digest=sha256:5883b54afb7d4838ca89b7260a20202124948f4644ab80214dd7544ecb0eefa0

Observation 45337095-b368-4d61-abe8-bec4486f2fed · outbound

This paper cites Gaussian mixture models.Encyclopedia of biometrics, 741(659-663): 3, 2009.

Linear Mixture Distributionally Robust Markov Decision Processes Gaussian mixture models.Encyclopedia of biometrics, 741(659-663): 3, 2009

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:22.342941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:12.982231Z digest=sha256:caeb812f11428493b9f97b4da00dd4a8b159a18a746d615b43911467de6b9b05

Observation e64b59e7-8913-4b22-a6da-19a224eb9698 · outbound

This paper cites Markovian decision processes with uncertain transition probabilities.Operations Research, 21(3):728–740, 1973.

Linear Mixture Distributionally Robust Markov Decision Processes Markovian decision processes with uncertain transition probabilities.Operations Research, 21(3):728–740, 1973

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:22.165116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:13.089462Z digest=sha256:af59fdfd905a5f9027980c3b07ee3bbafd17cb07b03b39d55039b8d6c3888ee6

Observation 942aebe0-2956-4962-aff9-008cd03e2f88 · outbound

This paper cites Distributionally robust model-based offline reinforcement learning with near-optimal sample complexity.Journal of Machine Learning Research, 25(200):1–91,.

Linear Mixture Distributionally Robust Markov Decision Processes Distributionally robust model-based offline reinforcement learning with near-optimal sample complexity.Journal of Machine Learning Research, 25(200):1–91,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:22.019192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:13.224367Z digest=sha256:ff93ff52bcd807965f80fd39582bf15864a3963cc7c1fe50d1a4b677a93b692d

Observation ac69338f-c324-45ef-b6b7-b78ecf2e27b3 · outbound

This paper cites The curious price of distributional robustness in reinforcement learning with a generative model.Advances in Neural Information Processing Systems, 36, 2024.

Linear Mixture Distributionally Robust Markov Decision Processes The curious price of distributional robustness in reinforcement learning with a generative model.Advances in Neural Information Processing Systems, 36, 2024

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:21.820136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:13.360333Z digest=sha256:a791344528a27920a98a21ce2f5c35bead6ca49f32289bf416c19af7f7b6e077

Observation 1bc9241f-01ad-4e82-9451-126cbac39427 · outbound

This paper cites Robust offline reinforcement learning with linearly structuredf-divergence regularization.arXiv preprint arXiv:2411.18612, 2024.

Linear Mixture Distributionally Robust Markov Decision Processes Robust offline reinforcement learning with linearly structuredf-divergence regularization.arXiv preprint arXiv:2411.18612, 2024

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:13.540609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:13.540609Z digest=sha256:6ea95963a890b2022f53ab83922b258dcda0bd793bb3fa0c164269674624f0ad

Observation b1cf2ff8-fe45-4251-b08c-791f50ab0925 · outbound

This paper cites Pessimistic model-based offline reinforcement learning under partial coverage.

Linear Mixture Distributionally Robust Markov Decision Processes Pessimistic model-based offline reinforcement learning under partial coverage

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:21.615475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:13.759936Z digest=sha256:80bcdc4acb9705f17383c82aac7100ec0155b3c9521c3829c0c08f3ea4f023d8

Observation 3e8d53b8-5bfb-4d31-816b-e06f9bef2976 · outbound

This paper cites Sample complexity of offline distributionally robust linear markov decision processes.

Linear Mixture Distributionally Robust Markov Decision Processes Sample complexity of offline distributionally robust linear markov decision processes

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:21.435252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:13.955582Z digest=sha256:3e812d56d69e98aeda5d71770935432d07c39a08d179c009eadd243c14161453

Observation 48e2cbd0-62f7-4b33-aa0d-c0db5afc2550 · outbound

This paper cites Return augmented decision transformer for off-dynamics reinforcement learning.arXiv preprint arXiv:2410.23450, 2024.

Linear Mixture Distributionally Robust Markov Decision Processes Return augmented decision transformer for off-dynamics reinforcement learning.arXiv preprint arXiv:2410.23450, 2024

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:14.180938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:14.180938Z digest=sha256:5be568adbea7f07116f6ae8f554df60dafcb3037cc0e4208724a27fb9454b0e3

Observation e85e83e1-f41d-410e-990b-32bc20e4b662 · outbound

This paper cites Robust markov decision processes.

Linear Mixture Distributionally Robust Markov Decision Processes Robust markov decision processes

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:21.282437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:14.366644Z digest=sha256:b632972f5422170c551acf47e19d9c4231aac11a2c3b903e4a1e89423f87df11

Observation 5cd48174-e78c-4762-9c2e-2fc32074bb3d · outbound

This paper cites Mutual alignment transfer learning.

Linear Mixture Distributionally Robust Markov Decision Processes Mutual alignment transfer learning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:21.148819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:14.512261Z digest=sha256:f422f5f82a2e8e68416c6b27ccf83f0d75c41ebcbcd710b8db9e550dba2f7be6

Observation f019cdc2-eb5f-4a4f-b270-8c78b598182f · outbound

This paper cites The robustness-performance tradeoff in markov decision processes.

Linear Mixture Distributionally Robust Markov Decision Processes The robustness-performance tradeoff in markov decision processes

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:20.851485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:14.695106Z digest=sha256:0b9ee37e0c19430e4de716e25d03cdd495918e1d5c444f867381e86c58d50d35

Observation 4c789318-dc3a-46c2-942a-78328245af70 · outbound

This paper cites Improved sample complexity bounds for distributionally robust reinforcement learning.

Linear Mixture Distributionally Robust Markov Decision Processes Improved sample complexity bounds for distributionally robust reinforcement learning

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:20.560793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:14.854406Z digest=sha256:b703f16675c08b33997d363cbb18f978b3d70d4dedc8c21ec27ee8d49c7b96e8

Observation 48fe2c9b-761c-4683-95c9-cbe1ca36377a · outbound

This paper cites Toward theoretical understandings of robust markov decision processes: Sample complexity and asymptotics.The Annals of Statistics, 50 (6):3223–3248, 2022.

Linear Mixture Distributionally Robust Markov Decision Processes Toward theoretical understandings of robust markov decision processes: Sample complexity and asymptotics.The Annals of Statistics, 50 (6):3223–3248, 2022

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:20.325566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:15.061034Z digest=sha256:8d744d1edc6738946962f81ea4c7efc7e812bfac328b0b480e1784cd4f485622

Observation 9ea4557c-3d4f-4c06-8821-0998a738ab90 · outbound

This paper cites Reward-free model-based reinforcement learning with linear function approximation.Advances in Neural Information Processing Systems, 34:1582–1593, 2021.

Linear Mixture Distributionally Robust Markov Decision Processes Reward-free model-based reinforcement learning with linear function approximation.Advances in Neural Information Processing Systems, 34:1582–1593, 2021

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:20.071385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:15.218794Z digest=sha256:dffa4439d0bb90c3bc606cfdba217e8ff9799e8816929603076da0e6f2f1edd7

Observation c9bd51f5-0c82-4c63-b83e-3b55b28179df · outbound

This paper cites Learning adversarial linear mixture markov decision processes with bandit feedback and unknown transition.

Linear Mixture Distributionally Robust Markov Decision Processes Learning adversarial linear mixture markov decision processes with bandit feedback and unknown transition

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:19.715374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:15.452113Z digest=sha256:6e83c0dfe1614293d5c6a08091c2b6705da3f7baf184e1e569742f921427d4cd

Observation 64543293-787a-4e7d-9421-ba5098cc83b2 · outbound

This paper cites Sim-to-real transfer in deep reinforcement learning for robotics: a survey.

Linear Mixture Distributionally Robust Markov Decision Processes Sim-to-real transfer in deep reinforcement learning for robotics: a survey

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:19.344128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:15.598520Z digest=sha256:db17d3190002a537c839ef3f591b646127240453bed790244d667a4ac15ba5e3

Observation c31db8ea-24ae-41f8-9c98-2f9ac1ce89a9 · outbound

This paper cites Nearly minimax optimal reinforcement learning for linear mixture markov decision processes.

Linear Mixture Distributionally Robust Markov Decision Processes Nearly minimax optimal reinforcement learning for linear mixture markov decision processes

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:18.984048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:15.708272Z digest=sha256:79be48d6a1d5f6f0d94f8353b2995e991ef0c7d5983573b97386a996b36e316f

Observation e07a3f5e-17ce-4ad3-bd72-8388f5f6787a · outbound

This paper cites Provably efficient reinforcement learning for discounted mdps with feature mapping.

Linear Mixture Distributionally Robust Markov Decision Processes Provably efficient reinforcement learning for discounted mdps with feature mapping

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:18.616950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:15.877799Z digest=sha256:2ab0fc39c275eaf63cb8a4bd1015c678e13ee621c43404d8c091656114ea42db

Observation 749cdfc5-2079-4a84-9f5c-0edfc4defa36 · outbound

This paper cites Natural actor-critic for robust reinforcement learning with function approximation.

Linear Mixture Distributionally Robust Markov Decision Processes Natural actor-critic for robust reinforcement learning with function approximation

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:18.244415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:16.093379Z digest=sha256:e2de6d48e79d73d82d7fcc286c7b850a227735e999528db3427b435ba5ccb43f

Observation e0055a81-ecff-4ada-b7a5-12dacbbec727 · outbound

This paper cites Finite-sample regret bound for distributionally robust offline tabular reinforcement learning.

Linear Mixture Distributionally Robust Markov Decision Processes Finite-sample regret bound for distributionally robust offline tabular reinforcement learning

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:17.858529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:16.285333Z digest=sha256:161733b1a727745bc7720cc10fb396af09007b5072ede71e5014a59373574930

Observation 832614ab-83b0-4b22-b97b-5e9906b73a12 · outbound

This paper cites Time- constrained robust mdps.Advances in Neural Information Processing Systems, 37:35574–35611,.

Linear Mixture Distributionally Robust Markov Decision Processes Time- constrained robust mdps.Advances in Neural Information Processing Systems, 37:35574–35611,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:17.527348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:16.401146Z digest=sha256:7cd1d2ba22c35c04fd58a985e4002875acba7a1717a643feae56a523559c93cf

Observation 5a065462-4302-4047-8c1a-f3386aceef79 · outbound

This paper cites A.1 Proof of Theorem 3.4 Proof.

Linear Mixture Distributionally Robust Markov Decision Processes A.1 Proof of Theorem 3.4 Proof

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:17.145142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:44:16.530086Z digest=sha256:8e9216e35f43384037369a220abb3e97f0ac7dac5fe5ba51173b32ad032803e3

Pith citing papers

No inbound Pith citation observations are available.