Pith. sign in

Paper Citation Record · LEDGER

Distributional Inverse Reinforcement Learning

As of 6 August 2026, this Paper Citation Record lists 67 of 67 outbound references and 0 inbound Pith citation observations for arXiv:2510.03013.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2510.03013 v4

Coverage vector

measured 67 of 67 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T12:42:58.448782Z

measured 67 of 67 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

67 of 67 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved67
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5b5898a2-9e6e-4207-8cff-c45924fdaa44 · outbound

This paper cites write newline.

Distributional Inverse Reinforcement Learning write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:55.390965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:55.390965Z digest=sha256:b2515291a22496beb66e0c0feb814f4513d960d99e6a0dd0a5e5149e1806f962

Observation 4b8df049-d540-41cd-a937-bab4baf90072 · outbound

This paper cites Apprenticeship learning via inverse reinforcement learning.

Distributional Inverse Reinforcement Learning Apprenticeship learning via inverse reinforcement learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:55.494235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:55.494235Z digest=sha256:c1c0cc486e2e5ab5377d73dbccd2101c450e434026c8a5c1209c907bcffaba53

Observation 951c8b6d-31ed-4378-b53e-f4530b549d7c · outbound

This paper cites A survey of inverse reinforcement learning: Challenges, methods and progress.

Distributional Inverse Reinforcement Learning A survey of inverse reinforcement learning: Challenges, methods and progress

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:55.585972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:55.585972Z digest=sha256:8b4d25844cd9d9535234d017c5a8089e18abddc3eb7d91266c4b2d1942d3e99f

Observation dc68dd71-a32f-46d9-9dfa-14bdeae8aa38 · outbound

This paper cites Dynamic inverse reinforcement learning for characterizing animal behavior.

Distributional Inverse Reinforcement Learning Dynamic inverse reinforcement learning for characterizing animal behavior

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:55.641301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:55.641301Z digest=sha256:eab00cb8cbc4141c43cbfff17b2a7cf7c100b878ccc2032af17d6410ed373457

Observation 987d6621-cb01-4a7e-8463-08ee09d332c3 · outbound

This paper cites The multivariate skew-normal distribution.

Distributional Inverse Reinforcement Learning The multivariate skew-normal distribution

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:55.700567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:55.700567Z digest=sha256:49e45ec7e57e571561b08495db53062be3abf3c1dab24b02659ffff8c3b0a8c4

Observation e240067a-e277-4e0d-aea1-92a7cebae757 · outbound

This paper cites Walking the Values in Bayesian Inverse Reinforcement Learning.

Distributional Inverse Reinforcement Learning Walking the Values in Bayesian Inverse Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:55.813564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:55.813564Z digest=sha256:6f5132991e9df84af435214ed84e7ad3887c5cbc8ff6bea79c2f5de3b2b1086a

Observation 6d5069a7-bdef-4b9c-95eb-8e9c8b03b40a · outbound

This paper cites A distributional perspective on reinforcement learning.

Distributional Inverse Reinforcement Learning A distributional perspective on reinforcement learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:55.940976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:55.940976Z digest=sha256:b24ca3bfe40c79eecd8d0273f8a387adc378497f5f1981a13221c431f472f72c

Observation a9ca4cd1-a0c0-4c5a-b060-f28417ab564d · outbound

This paper cites Variational inference: A review for statisticians.

Distributional Inverse Reinforcement Learning Variational inference: A review for statisticians

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:56.088085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:56.088085Z digest=sha256:8e907f75bef722808081755766886e7a8d662858eecfc097bb752df99c63dd36

Observation 3280a2e8-d2b7-4a9e-b2b6-f5f3d768c23a · outbound

This paper cites Scalable Bayesian Inverse Reinforcement Learning.

Distributional Inverse Reinforcement Learning Scalable Bayesian Inverse Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:56.164696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:56.164696Z digest=sha256:6083a5e5b7d9fe30b75fdc7a1f1838302a026626bddbe93ebfc66a90799a883b

Observation 87d522fa-a973-41ac-8739-c59118afa6d0 · outbound

This paper cites Eliciting risk aversion with inverse reinforcement learning via interactive questioning.

Distributional Inverse Reinforcement Learning Eliciting risk aversion with inverse reinforcement learning via interactive questioning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:56.276796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:56.276796Z digest=sha256:1c8b6fe64a7a91dd1ba799e58960829c463a8620fd9c5506d1a8cf278c273199

Observation 1251fc23-f23c-41da-b5ee-f62ade9979f4 · outbound

This paper cites Map inference for bayesian inverse reinforcement learning.

Distributional Inverse Reinforcement Learning Map inference for bayesian inverse reinforcement learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:56.402619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:56.402619Z digest=sha256:904e086d0c8ef14d3f3eb706070d0a05c8b71902e8c137267a563634261a0a13

Observation 8c67766b-f540-42ec-9332-be25c24175b9 · outbound

This paper cites Implicit quantile networks for distributional reinforcement learning.

Distributional Inverse Reinforcement Learning Implicit quantile networks for distributional reinforcement learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:56.493316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:56.493316Z digest=sha256:410d8016d6d87bc31523c88a4840d917066a0b33fc31ab405ad1f4c55ab922c8

Observation 82372248-f577-43f2-94c3-bff25775f54a · outbound

This paper cites Distributional reinforcement learning with quantile regression.

Distributional Inverse Reinforcement Learning Distributional reinforcement learning with quantile regression

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:56.694782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:56.694782Z digest=sha256:c0c7c7234ccbb9155f922f0336383547b33a0578955fdd19f6612c36c6ecf83f

Observation c0b8e8b4-6128-4a8e-b5c9-9fef05542847 · outbound

This paper cites Cortical substrates for exploratory decisions in humans.

Distributional Inverse Reinforcement Learning Cortical substrates for exploratory decisions in humans

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:56.902589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:56.902589Z digest=sha256:6e03109882e0cb004b37e167a45308ece4f2e88a688a4cfe3a0521cb3f0ede9f

Observation 05695a72-5de5-4f56-8617-f3a7ef553418 · outbound

This paper cites Nonuniform random variate generation.

Distributional Inverse Reinforcement Learning Nonuniform random variate generation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:57.072176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:57.072176Z digest=sha256:3c9f7bbbcbe7703d8d366ab6e0a6a3d17176abc2078a310da680a30f8d16cd95

Observation 6ca678e3-d3bd-4010-815b-0c40b3e63a09 · outbound

This paper cites Remarks on quantiles and distortion risk measures.

Distributional Inverse Reinforcement Learning Remarks on quantiles and distortion risk measures

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:57.227314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:57.227314Z digest=sha256:5470fff0780b8e59180922c607c5b9911d43a3286510404bde41baf09b2a61cc

Observation 71049599-d292-4904-8b18-aca7577d1680 · outbound

This paper cites Distributional soft actor-critic: Off-policy reinforcement learning for addressing value estimation errors.

Distributional Inverse Reinforcement Learning Distributional soft actor-critic: Off-policy reinforcement learning for addressing value estimation errors

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:57.331776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:57.331776Z digest=sha256:b5bc2e5e06e8ac6be777856d4f81d21228d691d2853678a0f5967dc13448ca45

Observation 6ff1f605-4b91-41a4-928e-48d7f8e0ade7 · outbound

This paper cites D4RL: Datasets for Deep Data-Driven Reinforcement Learning.

Distributional Inverse Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:57.503906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:57.503906Z digest=sha256:2f3c6d3626a4c05144af1a824d8c61a8371a289cb3aa686b983ad760f7876be1

Observation 191ba8eb-0a53-45de-96b4-86c9b0c801af · outbound

This paper cites Iq-learn: Inverse soft-q learning for imitation.

Distributional Inverse Reinforcement Learning Iq-learn: Inverse soft-q learning for imitation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:57.530407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:57.530407Z digest=sha256:624323141cde60aa4f067890ec2f945f05cce9a5f03f97478b297b1dbce45b48

Observation d84e4251-f711-41d5-a78a-566f5746e555 · outbound

This paper cites Probability: a graduate course, volume 200.

Distributional Inverse Reinforcement Learning Probability: a graduate course, volume 200

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:57.626102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:57.626102Z digest=sha256:d99cfc2cc674df9a890a9b26d8c68a684972a0eca46255856be07be70cd0d06c

Observation a9f5bb93-514b-4621-9b9d-60801c06df64 · outbound

This paper cites Rules for ordering uncertain prospects.

Distributional Inverse Reinforcement Learning Rules for ordering uncertain prospects

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:57.746691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:57.746691Z digest=sha256:28251ce41cc7cc649ac410faceac2699a3789dfc5a79c5f3077e672a67029344

Observation c2569331-cc4c-4bc4-9fd4-e6d1bfa48911 · outbound

This paper cites IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies.

Distributional Inverse Reinforcement Learning IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:57.841983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:57.841983Z digest=sha256:09b58b98bb213c6e28cd7fb63c7c0dfe4a2e48f3bbcb3fd2a54463e3fa905afe

Observation bcadbcff-cf03-4c10-884a-67ebf6fe7aca · outbound

This paper cites Introduction to real analysis, volume 280.

Distributional Inverse Reinforcement Learning Introduction to real analysis, volume 280

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:57.961326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:57.961326Z digest=sha256:ad3a4466e51fde7c9b4f55f61fd1820c67a55c7e640aced53fa72f4335f7e336

Observation 5adf489c-1b76-45a4-b6d5-39d8e54cfbca · outbound

This paper cites Generative adversarial imitation learning.

Distributional Inverse Reinforcement Learning Generative adversarial imitation learning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.070477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.070477Z digest=sha256:a8771d3afd3b7c5095e20eddba18923af2dce7f58480b258e6301048fb1d987d

Observation 5a7e6729-f9ef-450e-8d2c-57da76d2abf3 · outbound

This paper cites A bayesian approach to generative adversarial imitation learning.

Distributional Inverse Reinforcement Learning A bayesian approach to generative adversarial imitation learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.230803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.230803Z digest=sha256:5044376135586759b1448d2cdf48df6a1a2c0486919ff198a059bb193161a5a7

Observation bb73d1d7-8978-4d1d-b438-4bb3022ba912 · outbound

This paper cites Rize: Regularized imitation learning via distributional reinforcement learning.

Distributional Inverse Reinforcement Learning Rize: Regularized imitation learning via distributional reinforcement learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.278539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.278539Z digest=sha256:6f49bbc9164f777bac0417fea0bbe28a8531ebca9fc987a4e19622c01e072347

Observation c51944cb-860f-43b7-bdd3-aa00e00e696b · outbound

This paper cites Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors.

Distributional Inverse Reinforcement Learning Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.307441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.307441Z digest=sha256:6c34992221e62b1e103a3a45f43a04cbeedd94d47cbfeca660e4be673ada7055

Observation 0eb30226-101e-49d8-b78c-4e8d28c65da1 · outbound

This paper cites Imitation Learning via Off-Policy Distribution Matching.

Distributional Inverse Reinforcement Learning Imitation Learning via Off-Policy Distribution Matching

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.311175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.311175Z digest=sha256:6261b97e0816412227cafbad22b91072220c2e5f468485bc79812ed4ffcaff20

Observation 4a2b943b-36dd-43c2-8c59-3d1ea16fc9af · outbound

This paper cites Offline Reinforcement Learning with Implicit Q-Learning.

Distributional Inverse Reinforcement Learning Offline Reinforcement Learning with Implicit Q-Learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.314862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.314862Z digest=sha256:92e26a9ab096fd028d86250fae31ab18dd4427bc49034e1e9c89a95a717d8a3d

Observation 96537912-f523-42aa-9d93-d40580eef4d7 · outbound

This paper cites Risk-sensitive generative adversarial imitation learning.

Distributional Inverse Reinforcement Learning Risk-sensitive generative adversarial imitation learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.318874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.318874Z digest=sha256:132f7faf9ae5aabc91f8c00397b98702ff5d19f08ef0c93c67a61e8e2a7b7246

Observation 522bfb5c-f96f-4eea-875c-cdeae8180990 · outbound

This paper cites A tutorial on energy-based learning.

Distributional Inverse Reinforcement Learning A tutorial on energy-based learning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.322564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.322564Z digest=sha256:9eec16ee5608bf70df058a3b13499d2e18f1b06f757df64c8a2540be69d51c1d

Observation 6274dfce-ac41-4ba8-970c-be7613c3b777 · outbound

This paper cites Risk-sensitive mpcs with deep distributional inverse rl for autonomous driving.

Distributional Inverse Reinforcement Learning Risk-sensitive mpcs with deep distributional inverse rl for autonomous driving

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.325721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.325721Z digest=sha256:6f91589bc24f51c76e03a3c859206baf7675452bbdc3800743a36b0d4f4d87b7

Observation 38eaa1e7-0b35-4894-b588-06dcbbc25883 · outbound

This paper cites Nonlinear inverse reinforcement learning with gaussian processes.

Distributional Inverse Reinforcement Learning Nonlinear inverse reinforcement learning with gaussian processes

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.330053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.330053Z digest=sha256:6d58b678d350ad9b41a04a0051b2d6f72c60d1cfaa0d0165cc51b860c800f770

Observation a6701269-4e89-4611-9d51-e494dc725a07 · outbound

This paper cites Internally rewarded reinforcement learning.

Distributional Inverse Reinforcement Learning Internally rewarded reinforcement learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.333477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.333477Z digest=sha256:2246115e6a811658bdc58941418d5aa71b0d54b0e48371ad908d5a267a7b1332

Observation 789e4226-b552-4357-ae0f-9ff5342a97de · outbound

This paper cites Bellman Diffusion: Generative Modeling as Learning a Linear Operator in the Distribution Space.

Distributional Inverse Reinforcement Learning Bellman Diffusion: Generative Modeling as Learning a Linear Operator in the Distribution Space

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.337145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.337145Z digest=sha256:4f9a657e0b94c2b62605aa201ade110ad2cc5f3f9ddb451cb1a745aef2d62aa4

Observation 83032962-50b1-42ca-af09-87ba7654cf0c · outbound

This paper cites Distributional reinforcement learning for risk-sensitive policies.

Distributional Inverse Reinforcement Learning Distributional reinforcement learning for risk-sensitive policies

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.341059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.341059Z digest=sha256:bd845e7395222b44dbab853d5fb554ac2d31a6f447dfe7c794f2ebde500b4d46

Observation b1e489fe-d107-4c75-b4da-8e70b1bf39c2 · outbound

This paper cites Kernel Density Bayesian Inverse Reinforcement Learning.

Distributional Inverse Reinforcement Learning Kernel Density Bayesian Inverse Reinforcement Learning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.344896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.344896Z digest=sha256:6d89847e85d6c14b142e83c469e8d1b10b80842cc2ee225e98ef1185d9f40a43

Observation 9cd77b32-3602-4bbc-8a2a-6df66c756a64 · outbound

This paper cites Spontaneous behaviour is structured by reinforcement without explicit reward.

Distributional Inverse Reinforcement Learning Spontaneous behaviour is structured by reinforcement without explicit reward

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.348525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.348525Z digest=sha256:d0ca7dfdd30e6bbb39cf924666f02197b0683ad11f80847bdbe4e98df9bb6e6d

Observation fc515a0d-1ff0-4141-97ee-87b63a27c736 · outbound

This paper cites Spontaneous behaviour is structured by reinforcement without explicit reward.

Distributional Inverse Reinforcement Learning Spontaneous behaviour is structured by reinforcement without explicit reward

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.351966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.351966Z digest=sha256:f95eb8d6f95690eb0714db78b855be290194cb2382feae4b8a17dcce25e3393f

Observation b780f1f3-fe50-484b-9146-888e506328f0 · outbound

This paper cites The kolmogorov-smirnov test for goodness of fit.

Distributional Inverse Reinforcement Learning The kolmogorov-smirnov test for goodness of fit

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.355441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.355441Z digest=sha256:24cc1a4bfd72923c7c9e57fc7b54556bfeb8e7bc9426436b077acf4c57e3131e

Observation 89935910-fef2-4b54-9ee2-bb176b05306b · outbound

This paper cites Foraging for foundations in decision neuroscience: insights from ethology.

Distributional Inverse Reinforcement Learning Foraging for foundations in decision neuroscience: insights from ethology

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.358680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.358680Z digest=sha256:e9c47a04f3a87b4ccb7e2ddfea83115738cf4a40552543b0672253d7d8fc8089

Observation fd7e5dd7-16e3-44a6-8a66-6f40131ed568 · outbound

This paper cites f-irl: Inverse reinforcement learning via state marginal matching.

Distributional Inverse Reinforcement Learning f-irl: Inverse reinforcement learning via state marginal matching

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.361919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.361919Z digest=sha256:a814493cf4fc4c33df36e0393d8f83b847e516fefc10b95dbeaf35304c9d5764

Observation 49e421da-f209-41ef-a317-006242c67d6c · outbound

This paper cites Bayesian inverse reinforcement learning.

Distributional Inverse Reinforcement Learning Bayesian inverse reinforcement learning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.365160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.365160Z digest=sha256:e56786999b49605e727923e19fc977909b33453582b831e393f65727006b28cd

Observation d10c9de8-520a-44ea-81d5-ef6ce727d9f7 · outbound

This paper cites Optimization of conditional value-at-risk.

Distributional Inverse Reinforcement Learning Optimization of conditional value-at-risk

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.368395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.368395Z digest=sha256:e1b024fe30690d40ab2a7a324831ee122f78606fdf1f9b3e1b4f14b95035a61d

Observation 962fc82b-049f-43bb-a3db-974dcc98e7af · outbound

This paper cites Driving with style: Inverse reinforcement learning in general-purpose planning for automated driving.

Distributional Inverse Reinforcement Learning Driving with style: Inverse reinforcement learning in general-purpose planning for automated driving

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.372070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.372070Z digest=sha256:537e2ab3e2492d04ca29c8066a96b3e5e89b17a23547d4161e55309aeb04801d

Observation bd848cbc-33f6-462e-bffb-0fe808aa7fd6 · outbound

This paper cites Learning risk-aware quadrupedal locomotion using distributional reinforcement learning.

Distributional Inverse Reinforcement Learning Learning risk-aware quadrupedal locomotion using distributional reinforcement learning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.375329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.375329Z digest=sha256:a48f273a9da9f4ac495f04455cd5189ad4b528407201d07f7fd5a3bc3fede033

Observation d70fa356-7c2a-4479-9e7b-87088047875d · outbound

This paper cites A neural substrate of prediction and reward.

Distributional Inverse Reinforcement Learning A neural substrate of prediction and reward

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.378371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.378371Z digest=sha256:dcc086ad5a45ace5ad9157f5bca1d71cfe8bd07eb0fd283ac1687424a47cf1e0

Observation e934e22a-4172-48e9-ad35-ff0ea398474a · outbound

This paper cites Distortion risk measures in portfolio optimization.

Distributional Inverse Reinforcement Learning Distortion risk measures in portfolio optimization

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.381868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.381868Z digest=sha256:9b54382682aa2e2ac0eb82f1a5fd59c0f665bc8c9bfd979f36615379af9c946a

Observation f8643bf4-bbad-422a-8b03-dda969f203f9 · outbound

This paper cites Risk-sensitive inverse reinforcement learning via semi-and non-parametric methods.

Distributional Inverse Reinforcement Learning Risk-sensitive inverse reinforcement learning via semi-and non-parametric methods

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.385441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.385441Z digest=sha256:0ebb9367f39663c36abbcb1cf60b2cb53369eab90905fd4ae6ab7e7505c85de9

Observation 694dc24c-9400-4ba9-baa9-422b04a29993 · outbound

This paper cites Reinforcement learning: An introduction, volume 1.

Distributional Inverse Reinforcement Learning Reinforcement learning: An introduction, volume 1

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.389074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.389074Z digest=sha256:781532424f5edc478985f623dcbeb4c7bdd4c0228b57b953419da00069ee8f0f

Observation ae7feb3d-cf74-45cc-9a8f-b91b17477862 · outbound

This paper cites Risk-Averse Offline Reinforcement Learning.

Distributional Inverse Reinforcement Learning Risk-Averse Offline Reinforcement Learning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.392711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.392711Z digest=sha256:026528e036689ff2e7067155b5149d5d1a45bb55ae93a87db83ca06fefe74c1a

Observation f284b05b-1aed-4ac9-a934-981632a98f3f · outbound

This paper cites Inverse reinforcement learning algorithms and features for robot navigation in crowds: an experimental comparison.

Distributional Inverse Reinforcement Learning Inverse reinforcement learning algorithms and features for robot navigation in crowds: an experimental comparison

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.396647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.396647Z digest=sha256:0c241927e6e901bff74767d2a7d4fc0b93752c7ea8e4d14e3a6306df91173e09

Observation 619a0652-a124-43d8-ac85-72d12e9b5eed · outbound

This paper cites A bayesian approach to robust inverse reinforcement learning.

Distributional Inverse Reinforcement Learning A bayesian approach to robust inverse reinforcement learning

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.399778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.399778Z digest=sha256:94aa18efd6d4be282563f966f4474a2e8d1bd0cb5b748215222ef5943557e819

Observation b54de15a-532a-4048-895d-f47826a0041c · outbound

This paper cites Foundations of multivariate distributional reinforcement learning.

Distributional Inverse Reinforcement Learning Foundations of multivariate distributional reinforcement learning

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.402913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.402913Z digest=sha256:b7eae5f2e96ab26928daf909a8be8e0f9419a04dd20a180f71e6e19d8a3fa22e

Observation ac59373e-e0cd-40cf-a6f2-1b0b3e05b5d1 · outbound

This paper cites Inverse reinforcement learning with the average reward criterion.

Distributional Inverse Reinforcement Learning Inverse reinforcement learning with the average reward criterion

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.406388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.406388Z digest=sha256:c79c00a9dfdd8a2d16f8f9506f38be328557b72e65429445804275b55868f28e

Observation 928c7e42-6f67-4705-b4b8-226eb39208a6 · outbound

This paper cites Infer and adapt: Bipedal locomotion reward learning from demonstrations via inverse reinforcement learning.

Distributional Inverse Reinforcement Learning Infer and adapt: Bipedal locomotion reward learning from demonstrations via inverse reinforcement learning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.410677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.410677Z digest=sha256:4857e070a04635abec20dc02b56279b907ece7b6554b488cf6ea7c20c8ea4dad

Observation 21d38730-2082-4924-9641-d4821001d9b2 · outbound

This paper cites Efficient sampling-based maximum entropy inverse reinforcement learning with application to autonomous driving.

Distributional Inverse Reinforcement Learning Efficient sampling-based maximum entropy inverse reinforcement learning with application to autonomous driving

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.414728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.414728Z digest=sha256:ada17b6ce376cccb603f5c94e43da9b6e1acc48e41aa9a703e5bae32d957d1bf

Observation be72741b-a6bd-4959-b66a-c2d68dd9bf25 · outbound

This paper cites Maximum Entropy Deep Inverse Reinforcement Learning.

Distributional Inverse Reinforcement Learning Maximum Entropy Deep Inverse Reinforcement Learning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.418262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.418262Z digest=sha256:4cf6b4150b6ce0fd86b885c21ef494dee4050e4791cfa7030bd51322a8c7443c

Observation 4b45876f-d777-4d36-a0ab-a76728734313 · outbound

This paper cites Modeling, learning, perception, and control methods for deformable object manipulation.

Distributional Inverse Reinforcement Learning Modeling, learning, perception, and control methods for deformable object manipulation

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.421838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.421838Z digest=sha256:e83691a6c0d947d4ea9dabbb365f30734087154a70a1b233a85a4113cff6d6a2

Observation c6610b64-fb91-4bf9-b361-f978d26a6849 · outbound

This paper cites Maximum-likelihood inverse reinforcement learning with finite-time guarantees.

Distributional Inverse Reinforcement Learning Maximum-likelihood inverse reinforcement learning with finite-time guarantees

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.425216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.425216Z digest=sha256:efc8bd4777eeb99b926b671729de543f92f6d981dd870954ec760b011e29e2b2

Observation 370e9883-f575-4580-9d4d-3c98b22aa114 · outbound

This paper cites When demonstrations meet generative world models: A maximum likelihood framework for offline inverse reinforcement learning.

Distributional Inverse Reinforcement Learning When demonstrations meet generative world models: A maximum likelihood framework for offline inverse reinforcement learning

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.428599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.428599Z digest=sha256:e95d01fd5a41976f84d8978bb7ff9f848339e01231cf56425027804f9f1c5f57

Observation d71cef4b-e355-4f82-8641-1f296cff776e · outbound

This paper cites From Demonstrations to Rewards: Alignment Without Explicit Human Preferences.

Distributional Inverse Reinforcement Learning From Demonstrations to Rewards: Alignment Without Explicit Human Preferences

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.431814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.431814Z digest=sha256:1cecaa24abb1473633081fd3b87bc8add23af17d2e493a67786496e17344155c

Observation 41948553-1110-4682-9151-3bb609fa59b0 · outbound

This paper cites Maximum entropy inverse reinforcement learning.

Distributional Inverse Reinforcement Learning Maximum entropy inverse reinforcement learning

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.435351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.435351Z digest=sha256:5c046427a2fe14d91fe6b37cdb2d73ec5f3c6ae4a6a56f6b4e69e2c3199c6f23

Observation b22c4cda-33a9-4827-a484-d24b276a9057 · outbound

This paper cites Modeling interaction via the principle of maximum causal entropy.

Distributional Inverse Reinforcement Learning Modeling interaction via the principle of maximum causal entropy

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.438443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.438443Z digest=sha256:86d720e30742122d3ffffb325a4c101e46ea3dbf698241eecba0c00b963eb408

Observation ee23053d-dee5-4620-84c1-22874c44a4ba · outbound

This paper cites @esa (Ref.

Distributional Inverse Reinforcement Learning @esa (Ref

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.441299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.441299Z digest=sha256:5413f1c02895f40dfd25c054ea9a61b3dee739d953308feaf34038e569bd6d4a

Observation d7ffbf34-29ce-49cd-8d9f-9c684f698892 · outbound

This paper cites an unresolved cited work.

Distributional Inverse Reinforcement Learning Unresolved cited work

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.445479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.445479Z digest=sha256:c0efa07fd0ea8d173e238713bb48119867e28f76fcaae6d017e967154213f868

Observation ac7d397b-a0ae-4487-a0cf-968152f5c08f · outbound

This paper cites an unresolved cited work.

Distributional Inverse Reinforcement Learning Unresolved cited work

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.448782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.448782Z digest=sha256:7426581a7720b4cb07f39b6e1c7c10ce87f0a1e19948acbbb3cacdc4dafa602c

Pith citing papers

No inbound Pith citation observations are available.