Pith. sign in

Paper Citation Record · LEDGER

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors

As of 15 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 3 inbound Pith citation observations for arXiv:2501.12633.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.12633 v3

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T17:05:29.071278Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T12:42:58.307441Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T19:23:53.741950Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact2
  • verified fuzzy20
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e31c651a-b48a-4060-8030-3f71ccc54dd5 · outbound

This paper cites an unresolved cited work.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:28.935285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:28.935285Z digest=sha256:294717ac42ed6b96316636c0f0ab643d7e159738546129adb735ac7c04c8693a

Observation 2babc4da-bd78-4ce4-a79a-dd5003996cba · outbound

This paper cites an unresolved cited work.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-10T17:05:29.851954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.939335Z digest=sha256:f8973d9d8d27bc8dae75b4c88103825292dd1a135cdd03df3806e47e66539b13

Observation 06928f22-99a3-4fb0-955e-5b7b5fe850c1 · outbound

This paper cites Ashwood, Nicholas A.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Ashwood, Nicholas A

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:28.942657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:28.942657Z digest=sha256:bf1d9a4dc9913f526492f6992fb8524fec7ad1ee22ac4866b0ce457c1f10ae24

Observation 890e67a3-8726-4d3a-afb5-bf6a446529a2 · outbound

This paper cites an unresolved cited work.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-10T17:05:29.841725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.946216Z digest=sha256:702a718167ef54a8d0fb00e823a027a085e3bf51fe183ca57fc08cc461def369

Observation f8c7b837-2bf0-49fe-ab1f-a1f58ea797e5 · outbound

This paper cites Reinforcement learning with lstm in non-markovian tasks with longterm dependencies.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Reinforcement learning with lstm in non-markovian tasks with longterm dependencies

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.831416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.949756Z digest=sha256:74612c0a1b4c974a7be9f861fd1499ae0dd3ce97712dbcdb71305b036c2fed72

Observation 1282d7ee-2698-4720-bfb5-35e74d3ac728 · outbound

This paper cites Option-aware adversarial inverse reinforcement learning for robotic control.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Option-aware adversarial inverse reinforcement learning for robotic control

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:28.953251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:28.953251Z digest=sha256:cf579a2030ebf44b9f96a1d710a5229dfd199220af4f5f104bdeb11939d8936d

Observation 43de4f6b-770f-4ca3-9b3f-37463e1cea94 · outbound

This paper cites Learning robust rewards with adverserial inverse reinforcement learning.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Learning robust rewards with adverserial inverse reinforcement learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:28.957197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:28.957197Z digest=sha256:f01528cdcceb26fa76c91da866afd89fb303e53b6c926571719050688e49ab6f

Observation dffccea0-607c-405b-ae09-dcaaaf4276f6 · outbound

This paper cites Iq-learn: Inverse soft-q learning for imitation.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Iq-learn: Inverse soft-q learning for imitation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.814666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.960380Z digest=sha256:dea50ca47ae19663875357b3c5f3e4f7f9b8f8951108ceec0fcf10fe31d5d360

Observation 6f25f3a1-887d-4b1c-a0c9-9960b0861ff1 · outbound

This paper cites Reinforcement learning with deep energy-based policies.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Reinforcement learning with deep energy-based policies

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:28.963293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:28.963293Z digest=sha256:cdcacb0eb720076fc94315c3a4c29cde2748548975c087e2ad11c030ee4929a7

Observation 5c6dff0c-d7d9-4947-8b4f-64ba44489e1d · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.798644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.966469Z digest=sha256:826f8a9023390598e44c106e6878adc5507239bc69529380747b9845848e6fde

Observation 427144c5-0c2f-43ce-88b5-91edca1dbbf3 · outbound

This paper cites Area-specificity and plasticity of history-dependent value coding during learning.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Area-specificity and plasticity of history-dependent value coding during learning

Reference 11

Resolution
verified exact
doi, observed 2026-08-10T17:05:29.111183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.969606Z digest=sha256:a8d4d3a7861a0912626c36c9866173522d9f9b746750074a96d3920934d2ea39

Observation c5c50c21-94f8-4111-8cc6-0c453c3277bd · outbound

This paper cites Deep recurrent q-learning for partially observable mdps.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Deep recurrent q-learning for partially observable mdps

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.788619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.972933Z digest=sha256:decf51ac71038985c5e675330a40c78bf71761960952956fd12cc0584819a193

Observation fa6ed0f0-96c9-453a-8637-935da7e403ff · outbound

This paper cites Can ai predict animal movements? filling gaps in animal trajectories using inverse reinforcement learning.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Can ai predict animal movements? filling gaps in animal trajectories using inverse reinforcement learning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.778941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.975881Z digest=sha256:321e4536d1d3f88bbd9e786af945f289c935744fc37cb99cac851c1c97bb918b

Observation 2f2eb347-f708-452c-b819-0ec3daf43d64 · outbound

This paper cites Vime: Variational information maximizing exploration.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Vime: Variational information maximizing exploration

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.768677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.978951Z digest=sha256:03eec62b0a40723b0a1d7b153acff0bfa4b4d2fc0d816da4677b5484e99c7ba5

Observation 4103dd3f-95b1-4c01-ac9c-6248f1fafbc8 · outbound

This paper cites The what, how, and why of naturalistic behavior.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors The what, how, and why of naturalistic behavior

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:28.982112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:28.982112Z digest=sha256:1d94ce5793ef6c6679e5eb8ff2d21980aa3d06259a39854aa296603d633dc58a

Observation 9d79c6b3-84ee-4c1a-94cb-793bb61daef1 · outbound

This paper cites Recurrent switching linear dynamical systems.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Recurrent switching linear dynamical systems

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:28.984970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:28.984970Z digest=sha256:34adc52a2a9d1fc3f58577f1feccc932b787f924871ce92402f5faccaf89b4f9

Observation 1de4615a-6cf6-4980-a309-d20ee000f3e8 · outbound

This paper cites Spontaneous behaviour is structured by reinforcement without explicit reward.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Spontaneous behaviour is structured by reinforcement without explicit reward

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.758259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.988119Z digest=sha256:b66bb4a9677cdd6ee009a44d07e66af15db40778ee6134657665d42cb0f8cad9

Observation 897886ba-5686-4cec-9104-bde448a1cdd1 · outbound

This paper cites Neural mechanisms underlying the temporal organization of naturalistic animal behavior.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Neural mechanisms underlying the temporal organization of naturalistic animal behavior

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.748647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.991419Z digest=sha256:300bc1f09c101a06675cd7443e6e6dc361cfe07a232b3958657bbf2d65427935

Observation 8ba32143-773c-44da-b6c7-1827de1a3ed3 · outbound

This paper cites Ng and Stuart J.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Ng and Stuart J

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.737675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.994406Z digest=sha256:b6aae7cea7de3b9bb9020b0712817a79c21bf28282ad256e14d059229e7057fc

Observation e0baeb92-cfd2-447d-8c1a-7ee6aad6784e · outbound

This paper cites Inverse reinforcement learning with locally consistent reward functions.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Inverse reinforcement learning with locally consistent reward functions

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.727171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.997453Z digest=sha256:856132419859a1252555d50e09f140788282f77475711ffe1baea7004f969dae

Observation 0ab91280-6518-42cc-8831-eaf8884f6c9b · outbound

This paper cites Neural Map: Structured Memory for Deep Reinforcement Learning.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Neural Map: Structured Memory for Deep Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.000160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.000160Z digest=sha256:087f3cd342ebeaca8f65dc90f74c6510133d3724ff450254733d69b2f5a879db

Observation ba0d4822-8f84-4231-aa2a-2259be80926c · outbound

This paper cites Inverse reinforcement learning of bird flocking behavior.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Inverse reinforcement learning of bird flocking behavior

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.717394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.003408Z digest=sha256:5a757a176a7e74c171b41676a804fcdcd0f38d80e698af5a951791734cd25a18

Observation ea2be7f1-1628-401d-88af-fd0ee16ec21e · outbound

This paper cites Mice in a labyrinth show rapid learning, sudden insight, and efficient exploration.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Mice in a labyrinth show rapid learning, sudden insight, and efficient exploration

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.006505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.006505Z digest=sha256:f743e2fbec04325375a67102869359b3e83b3603c0fbdabf23f6c0f264629711

Observation dd59d97c-f7dc-4b11-a245-a941a9a686f6 · outbound

This paper cites Obtaining reward functions of rats using inverse reinforcement learning.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Obtaining reward functions of rats using inverse reinforcement learning

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.707238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.009760Z digest=sha256:e739000dd9e22170b371946db2f191d893ed4ea4a5778b46017b473fbd5bb42e

Observation a564e214-214c-4827-a054-5d47c3763555 · outbound

This paper cites Active sensing with predictive coding and uncertainty minimization.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Active sensing with predictive coding and uncertainty minimization

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.012758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.012758Z digest=sha256:58863de9d516fefd75b2c1474752fcb8840fbdd216acfc3383df7dd549fa8a62

Observation 299d3dd6-e37f-4c77-9a1c-ba5ec95589c8 · outbound

This paper cites Latent Variable Models for Characterizing the Dynamic Structure Underlying Complex Behaviors.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Latent Variable Models for Characterizing the Dynamic Structure Underlying Complex Behaviors

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.697431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.016014Z digest=sha256:a4cb0ac8be8e92e2aa6c9cf773f4b113588346cc4363ee58d301b39d27369109

Observation 9bd69652-71cb-4681-b458-9cec253884de · outbound

This paper cites Bayesian nonparametric inverse reinforcement learning for switched markov decision processes.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Bayesian nonparametric inverse reinforcement learning for switched markov decision processes

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.687576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.019844Z digest=sha256:ab24d8b5efe2d7927b08e3e0a030de22efa4764fb3f5010bc94f86bed0cee077

Observation ed2aa13e-5143-443e-be81-8636624f3247 · outbound

This paper cites Dyna, an integrated architecture for learning, planning, and reacting.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Dyna, an integrated architecture for learning, planning, and reacting

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.022764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.022764Z digest=sha256:2bc301129c18b3fa7383fd6e4ae02e6ca65d3fed0548c104b9d439c9146a24fc

Observation adfcf920-13b8-418e-8918-e4843a75b2b6 · outbound

This paper cites Keypoint-moseq: parsing behavior by linking point tracking to pose dynamics.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Keypoint-moseq: parsing behavior by linking point tracking to pose dynamics

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.670162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.025723Z digest=sha256:527a73e4408e909d164c7fc9d7d4515ffbc883f7ac13f836047e06675ad0be09

Observation 1f5fe048-d2d1-414f-9316-5b995af711ea · outbound

This paper cites Mapping sub-second structure in mouse behavior.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Mapping sub-second structure in mouse behavior

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.659566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.028562Z digest=sha256:9596d44a88d63cbf196278c01b67a730feadbb166cd6be1728fe607b7dedd694

Observation 46462f06-c652-4c83-b02d-3fb2003331b5 · outbound

This paper cites an unresolved cited work.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.031667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.031667Z digest=sha256:6fc002317008054e5764ccfc15e1c979950a5bda61abcc2080b0d213e3fac36e

Observation ba47349e-543f-4d1f-96c3-9fda901104bf · outbound

This paper cites Inverse reinforcement learning with the average reward criterion.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Inverse reinforcement learning with the average reward criterion

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.650153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.034608Z digest=sha256:f189b1dbd51c995da19e3735e09dc609ec7b5beebeb65c4ec9473bc580232f58

Observation 1b85156d-480c-4efa-a8a7-02b44159334a · outbound

This paper cites Imitating Language via Scalable Inverse Reinforcement Learning.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Imitating Language via Scalable Inverse Reinforcement Learning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.037500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.037500Z digest=sha256:2171ff011ae1a78d934286a190a618e76d2a907c064b44c72aa6e40a53f513cc

Observation 4c744ff2-8032-44f2-a197-3d9a3014733f · outbound

This paper cites Identification of animal behavioral strategies by inverse reinforcement learning.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Identification of animal behavioral strategies by inverse reinforcement learning

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.640503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.041100Z digest=sha256:dce1ae13b74ccb64eccf05c0fd2759314f750251e10c9c19ecdc1117488b3949

Observation 464930bb-32d5-420a-957f-c511a27d0462 · outbound

This paper cites Maximum-likelihood inverse reinforcement learning with finite-time guarantees.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Maximum-likelihood inverse reinforcement learning with finite-time guarantees

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.630114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.043980Z digest=sha256:935e09f70e69b9826fe66ff92e68b1dbd0a72bd7665e9f2c069b1a23fbf5b9d9

Observation 4c32f061-0b69-4036-be9b-1301f034a06c · outbound

This paper cites Multi-intention inverse q-learning for interpretable behavior representation.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Multi-intention inverse q-learning for interpretable behavior representation

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.618872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.046980Z digest=sha256:84eaae0a4c5c1388285fac10161b3bf74a3f3e527a6a899b2e1f47f3649735cb

Observation c4a94482-262b-48d2-8fef-da36566aea0a · outbound

This paper cites Ziebart, Andrew Maas, J.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Ziebart, Andrew Maas, J

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.050233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.050233Z digest=sha256:cd99b28dfb2279a646eebeec7f2b6237508f262bda6c5808ab4785840645f26d

Observation 4c2a85c1-0248-42ef-9331-fea43efe26bb · outbound

This paper cites Ziebart, J.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Ziebart, J

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.053428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.053428Z digest=sha256:eed5198ff8f8dc7fedd0c341ab97a762267dce67decac9a5dc825f19ff52c0a7

Observation b2a863eb-0719-42b0-8ee7-863662020543 · outbound

This paper cites write newline.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors write newline

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.056307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.056307Z digest=sha256:e0dad7f9aef7757b573f9c210c769a055037ba4662a1682dd93e63ce8aedf97d

Observation d3d62832-d692-48eb-9628-5bce9ccc937d · outbound

This paper cites @esa (Ref.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors @esa (Ref

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.060869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.060869Z digest=sha256:aa52cf43a8a1f4f67e1f39d4bff885c92b52933af5b67a189db0ae8af2c3cb5d

Observation 1d8df172-0eec-49d6-ad23-ebccad2ed9a3 · outbound

This paper cites an unresolved cited work.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.064277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.064277Z digest=sha256:567f17cf3e89ab9de7c05dc7ce4876c7e7f649f1980b34e4e1746714931c9312

Observation 7db36db6-7191-4c21-a088-56f2739c9be3 · outbound

This paper cites an unresolved cited work.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Unresolved cited work

Reference 42

Resolution
verified exact
raw_fallback, observed 2026-08-10T17:05:29.219493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.071278Z digest=sha256:743d2b354339692cb9bd5a1cce9e86633e3d7f240aa9985c7dbae0f9661307ae

Pith citing papers

Observation c51944cb-860f-43b7-bdd3-aa00e00e696b · inbound

Distributional Inverse Reinforcement Learning cites this paper.

Distributional Inverse Reinforcement Learning Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.307441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.307441Z digest=sha256:4d6a9f977a2b41077b0d00df1ae1dd5651dc896037430a5768fb74b9399f5821

Observation 5dd3d431-9932-4ed4-acfd-6a61b135a375 · inbound

Improving Zero-Shot Offline RL via Behavioral Task Sampling cites this paper.

Improving Zero-Shot Offline RL via Behavioral Task Sampling Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:41:18.778357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-07T16:27:15.347522Z digest=sha256:1d451229b4a89315e3f58e567505970f7393c138b0b4822db57b307a8e2785b2

Observation da554326-c531-43e7-bf51-d82c82b18576 · inbound

Probabilistic Recurrent Intention Switching Model cites this paper.

Probabilistic Recurrent Intention Switching Model Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T19:23:53.744266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-29T19:22:57.970141Z digest=sha256:6f2fcc55f365f57dd34601014a47a861e4078ad8240e87472670dea8e908c29c