Pith. sign in

Paper Citation Record · LEDGER

Offline Learning of Controllable Diverse Behaviors

As of 24 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 1 inbound Pith citation observation for arXiv:2504.18160.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.18160 v1

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:28:16.326058Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T12:36:12.177892Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

40 of 40 outbound references displayed

  • verified exact0
  • verified fuzzy24
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 46dc3da8-6009-4c4a-af39-8d0f03021353 · outbound

This paper cites Tenenbaum, Tommi S.

Offline Learning of Controllable Diverse Behaviors Tenenbaum, Tommi S

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.818801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.194332Z digest=sha256:430c7791505d1f30ef2725e3353018cb10e86a35ef51d8277f655b47b0b64dd6

Observation d94f85c7-da1e-4e97-a018-6624f513a8f1 · outbound

This paper cites Testing, validation, and verification of robotic and autonomous systems: A systematic review.

Offline Learning of Controllable Diverse Behaviors Testing, validation, and verification of robotic and autonomous systems: A systematic review

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.198338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.198338Z digest=sha256:f87c99e2ea62b0edfcb4360198f4ee483873380507192d76b164ec2f685f2c12

Observation 75694176-20d2-4b57-97ba-ffd5c4ceadcc · outbound

This paper cites Champandard.

Offline Learning of Controllable Diverse Behaviors Champandard

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.809260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.202103Z digest=sha256:08f54c58ba388b93a074131f96dd7d149405ca0315eb8c223146c94cfe19135c

Observation 29d0b07a-2bb0-4d81-b0aa-c2745a9a9d21 · outbound

This paper cites Decision transformer: Reinforcement learning via sequence modeling, 2021.

Offline Learning of Controllable Diverse Behaviors Decision transformer: Reinforcement learning via sequence modeling, 2021

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.798417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.205553Z digest=sha256:27c863a4795ac2330f44663aee23b04b48f0e5a6725875264af997a875b19949

Observation 677b6835-8134-4494-9728-7abd941bf429 · outbound

This paper cites Diffusion policy: Visuomotor policy learning via action diffusion, 2024.

Offline Learning of Controllable Diverse Behaviors Diffusion policy: Visuomotor policy learning via action diffusion, 2024

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.209477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.209477Z digest=sha256:798a533ecad4e09ea56ecaf0a0a27cafd61848d8e57f9f3beb8b09381385bc52

Observation 806fdde5-3a9b-4be7-94bb-00ce85872741 · outbound

This paper cites Implicit behavioral cloning, 2021.

Offline Learning of Controllable Diverse Behaviors Implicit behavioral cloning, 2021

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.212886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.212886Z digest=sha256:8d5a16d4475c89d5d2dba9649813f175ef87a1e267668b9deeb27efa5926d698

Observation 0953c5cd-3a7c-4463-9fac-f45d6a91b049 · outbound

This paper cites Off-policy deep reinforcement learning without exploration, 2019.

Offline Learning of Controllable Diverse Behaviors Off-policy deep reinforcement learning without exploration, 2019

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.216548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.216548Z digest=sha256:9c642f392d611e6941449c2c4dde0145ddc67a9e7261219f5e09f2d3aa6a276b

Observation a09e9538-4725-4bbc-bf54-38fb43b40b65 · outbound

This paper cites Multi-modal imitation learning from unstructured demonstrations using generative adversarial nets, 2017.

Offline Learning of Controllable Diverse Behaviors Multi-modal imitation learning from unstructured demonstrations using generative adversarial nets, 2017

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.767964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.219985Z digest=sha256:5ae0bed501a32f755218a5bef2243a7cbfc10c5e507684d56a0bf965f2372586

Observation 30759a24-2d69-4c66-af40-f82e2d2d7d57 · outbound

This paper cites Generative adversarial imitation learning, 2016.

Offline Learning of Controllable Diverse Behaviors Generative adversarial imitation learning, 2016

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.756042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.223419Z digest=sha256:a19049f7a27d829b89803bf56909b32ede9b070108fea3e0f413bc61f29fadc4

Observation 3c6a584c-cdd4-4cee-bdb4-f0ab0482104b · outbound

This paper cites ArtificialIntelligenceforGames.

Offline Learning of Controllable Diverse Behaviors ArtificialIntelligenceforGames

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.743286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.226687Z digest=sha256:f7d0fe6a6aee3e5a54602263d96603534a324e33b50fbc29b382111303fffb79

Observation 41bd42a6-83e8-49e5-8437-3633733ccdb3 · outbound

This paper cites Offline reinforcement learning as one big sequence modeling problem.

Offline Learning of Controllable Diverse Behaviors Offline reinforcement learning as one big sequence modeling problem

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.229950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.229950Z digest=sha256:a6ce4ff384547c20dcb00298bfa03eee7ef348bab4121d8531c7c14d61c1d44c

Observation 96968399-8ea5-41d2-9ad7-6c69e81f4142 · outbound

This paper cites Tenenbaum, and Sergey Levine.

Offline Learning of Controllable Diverse Behaviors Tenenbaum, and Sergey Levine

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.722532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.233360Z digest=sha256:162542f4b205a0f8339b5586a61e7b1e05bf2f28c5b518d5aadedfc0b057a6f4

Observation 70ecc02c-ea3e-41e2-8b8f-3e232cb958cd · outbound

This paper cites Towards diverse behaviors: A benchmark for imitation learning with human demonstrations, 2024.

Offline Learning of Controllable Diverse Behaviors Towards diverse behaviors: A benchmark for imitation learning with human demonstrations, 2024

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.710539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.236628Z digest=sha256:0d8fbdffa70d3f7e51b9ba730802e45be8f4fbf447bf39b5652074b32b82057a

Observation aa1b326c-bf1a-4638-9289-7951864c66e1 · outbound

This paper cites Offline reinforcement learning with implicit q-learning, 2021.

Offline Learning of Controllable Diverse Behaviors Offline reinforcement learning with implicit q-learning, 2021

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.239819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.239819Z digest=sha256:0aac398eaf81919748a0448f228d72729aaf97090f4d3419f320c9af09fc6f44

Observation 81c57388-98f0-4765-9981-4bf24cf02e82 · outbound

This paper cites Conservative q-learning for offline reinforcement learning, 2020.

Offline Learning of Controllable Diverse Behaviors Conservative q-learning for offline reinforcement learning, 2020

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.690906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.242854Z digest=sha256:ad55c5f130b4a108a0764b00293e5f8e1ffea55279c2c5c9a5d76538d76aec00

Observation 7ace76f2-8dca-4388-99b4-7c1dd7ef7e5f · outbound

This paper cites When should we prefer offline reinforcement learning over behavioral cloning?, 2022.

Offline Learning of Controllable Diverse Behaviors When should we prefer offline reinforcement learning over behavioral cloning?, 2022

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.678245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.246014Z digest=sha256:c9f73067201c13072bd9ffac2c0be71fc92f0dff7a6bf2f0f8ea16c3e8fa5fa0

Observation b7a876c2-4689-4c94-b62a-cb78e193429d · outbound

This paper cites Human-level ai’s killer application: Interactive computer games.

Offline Learning of Controllable Diverse Behaviors Human-level ai’s killer application: Interactive computer games

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.249328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.249328Z digest=sha256:84a49e77570f2f6eca4e4880e4c3a08570e880c6fb4f5c2fac813ab7fb4d728b

Observation 7417b1a5-a62f-4145-9588-54006e4cd578 · outbound

This paper cites Cic: Contrastive intrinsic control for unsupervised skill discovery, 2022.

Offline Learning of Controllable Diverse Behaviors Cic: Contrastive intrinsic control for unsupervised skill discovery, 2022

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.667157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.252605Z digest=sha256:b93075dc771514b17dd80db7c614e2b89b3071cfa8d9b3d156911c6c8a66f420

Observation fd3eb4a2-5eb3-4cbe-9aa3-c833f28ac8ee · outbound

This paper cites Infogail: Interpretable imitation learning from visual demonstrations, 2017.

Offline Learning of Controllable Diverse Behaviors Infogail: Interpretable imitation learning from visual demonstrations, 2017

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.653531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.255737Z digest=sha256:65e0ea671956d6c0b1b8672680bdd0444ce922f7a984b5b68f1cbb8ca660fc1f

Observation be3534d8-099c-419e-9ff4-8f7b4e999e2d · outbound

This paper cites What matters in learning from offline human demonstrations for robot manipulation, 2021.

Offline Learning of Controllable Diverse Behaviors What matters in learning from offline human demonstrations for robot manipulation, 2021

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.642839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.259013Z digest=sha256:e002a08970c0eb2e08cb5b1d83ed00a4c11870aa828d2286c2eb0125cffff116

Observation 99228e55-36e6-4003-b741-8d1e1ef89f24 · outbound

This paper cites Stylized offline reinforcement learning: Extracting diverse high-quality behaviors from heterogeneous datasets.

Offline Learning of Controllable Diverse Behaviors Stylized offline reinforcement learning: Extracting diverse high-quality behaviors from heterogeneous datasets

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.631729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.262412Z digest=sha256:a3d81ce8a06a4dbff7abaead243c7b8c22db2101ceee3af288454d3e556260c8

Observation 97acf20e-55e4-4883-a99d-d8a839750e9e · outbound

This paper cites Behavioral Mathematics for Game AI.

Offline Learning of Controllable Diverse Behaviors Behavioral Mathematics for Game AI

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.618992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.265486Z digest=sha256:1b03b525ea597cc4bbf06f3a7cb6900e9cafab5299c52f223a1b3c79174b299d

Observation 3dc502c1-1d7e-411f-90db-0f42da0ffacb · outbound

This paper cites Three states and a plan: The a.i.

Offline Learning of Controllable Diverse Behaviors Three states and a plan: The a.i

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.606400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.268571Z digest=sha256:b3462d4336f0f6420d7814e4bfee61767550c6589f8c41c46c77ce78e68bfaeb

Observation b54f62dd-7fe2-4ce8-8496-4a741e9a815e · outbound

This paper cites Imitating human behaviour with diffusion models, 2023.

Offline Learning of Controllable Diverse Behaviors Imitating human behaviour with diffusion models, 2023

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.271870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.271870Z digest=sha256:32155024f7934236d874d555ada44a6ba823641476bb966fb5d23c7c588e679d

Observation 4bf5aac0-a10a-412f-af32-3338b3a11e68 · outbound

This paper cites Pomerleau.

Offline Learning of Controllable Diverse Behaviors Pomerleau

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.585525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.275279Z digest=sha256:1ee0da812dcda4844c35be2743e91841ff239f4253d536d14cb2a12d4978b55e

Observation 75bce0af-0add-40f5-817a-b847e25a177f · outbound

This paper cites Goal-conditioned imitation learning using score-based diffusion policies, 2023.

Offline Learning of Controllable Diverse Behaviors Goal-conditioned imitation learning using score-based diffusion policies, 2023

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.278429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.278429Z digest=sha256:8bc9b5348071a9e9100578058a6cdc5d2a47eef3b30e0838e58cbf54baf2d3d4

Observation a62b62cb-de44-46b8-aeda-95dd382d1fa7 · outbound

This paper cites Artificial Intelligence: A Modern Approach.

Offline Learning of Controllable Diverse Behaviors Artificial Intelligence: A Modern Approach

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.564457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.281800Z digest=sha256:04c69195b7489ddeb703dd7fba5e1dd33a4054c24551c001c719e7df278440ea

Observation d12c0e44-e2a1-43ff-a5b9-e4f346195735 · outbound

This paper cites Behavior transformers: Cloning k modes with one stone, 2022.

Offline Learning of Controllable Diverse Behaviors Behavior transformers: Cloning k modes with one stone, 2022

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.284915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.284915Z digest=sha256:f11ca2e4c21fa8b670a3e7fc3636ef4ab779726fcb014c344c68558ba0460c59

Observation 3295ec91-7198-4626-afaf-8f8ea7743146 · outbound

This paper cites A mathematical theory of communication.

Offline Learning of Controllable Diverse Behaviors A mathematical theory of communication

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.543040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.288357Z digest=sha256:44b447dbdfaa4564357bf8e215c8baa39c22211ce78a19751f21115b5df45b57

Observation fd4e0919-0d39-4389-8edc-e607fc787cd2 · outbound

This paper cites Diverse behavior is what game ai needs: Generating varied human-like playing styles using evolutionary multi-objective deep reinforcement learning, 2020.

Offline Learning of Controllable Diverse Behaviors Diverse behavior is what game ai needs: Generating varied human-like playing styles using evolutionary multi-objective deep reinforcement learning, 2020

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.531677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.291853Z digest=sha256:845f3e902907c91135b9161115214e838a97c04e87f3d5a12139e92d848b4337

Observation 3c9df94a-e6ec-4e6b-b5d5-cade8256946b · outbound

This paper cites Skill decision transformer, 2023.

Offline Learning of Controllable Diverse Behaviors Skill decision transformer, 2023

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.520841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.294913Z digest=sha256:2229db2fb8e4cdcfc2bc00ac33a1ec3343ad682b761a91225a68f5c00381f2f9

Observation c16693e5-94de-4bcc-936d-9d34e356db7f · outbound

This paper cites A note on the evaluation of generative models, 2016.

Offline Learning of Controllable Diverse Behaviors A note on the evaluation of generative models, 2016

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.298184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.298184Z digest=sha256:79613d6cb92b1cd815e54b1b40238de642a5e7349246f8624fbae168bbd0bfbc

Observation 3a41362b-78d4-49a6-997a-e68ed080dc1d · outbound

This paper cites Braviner, Panteha Naderian, Chris J.

Offline Learning of Controllable Diverse Behaviors Braviner, Panteha Naderian, Chris J

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.503524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.301529Z digest=sha256:a0d6e052eaa8d2fac380fdcd23fc8d4f8069009bfc01d2b01a3e41496c755e1e

Observation 153f1c95-98b0-444d-b0f0-23afce40b45d · outbound

This paper cites Robust imitation of diverse behaviors, 2017.

Offline Learning of Controllable Diverse Behaviors Robust imitation of diverse behaviors, 2017

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.492499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.304586Z digest=sha256:c7c7e44b9d0ce55d1db72b5f6315eb657bac5e704e0b4557e87fbb760a7eebb5

Observation f13976e7-4bca-4154-992d-164f467843ae · outbound

This paper cites Diverse policies recovering via pointwise mutual information weighted imitation learning.

Offline Learning of Controllable Diverse Behaviors Diverse policies recovering via pointwise mutual information weighted imitation learning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.481318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.307723Z digest=sha256:718ce13669d7af9211e92b52abdbf7104fb51b3a22f056a9ada0f34cb7f8cff4

Observation d71bee41-b59d-4b80-bb59-e21ed30f1f03 · outbound

This paper cites Zhao, Vikash Kumar, Sergey Levine, and Chelsea Finn.

Offline Learning of Controllable Diverse Behaviors Zhao, Vikash Kumar, Sergey Levine, and Chelsea Finn

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.311023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.311023Z digest=sha256:c05f2194bbe023a0fc2abb3c07b8ad886014eb7d828d4d6e1b8b918af2d5c3d9

Observation 242579c5-806e-48f2-87b1-d953bae54ea9 · outbound

This paper cites write newline.

Offline Learning of Controllable Diverse Behaviors write newline

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.314234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.314234Z digest=sha256:fe57096e004ea72a5fece2a070d906a969f321a63a187e2d6bd8f0a4e652d8e0

Observation 9088d9d2-0943-4710-952c-867cb64beb36 · outbound

This paper cites @esa (Ref.

Offline Learning of Controllable Diverse Behaviors @esa (Ref

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.318479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.318479Z digest=sha256:8c47beabd9456d389f553382de4e59452469e216b326c696bd144d71cb8925c1

Observation 41ed44f3-f1f2-4c2d-900e-f3e37ace2b32 · outbound

This paper cites an unresolved cited work.

Offline Learning of Controllable Diverse Behaviors Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.322335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.322335Z digest=sha256:435c91133d094253c63d5265ad0382153c24963c653fff0b14e2a2aa81e08f68

Observation 9decf6ca-2c54-469c-84c5-1e1332b4bb65 · outbound

This paper cites For robotics, learning from human experts allows to reach human-level performance without any controller hard coding or expensive interaction with simulated or real environments.

Offline Learning of Controllable Diverse Behaviors For robotics, learning from human experts allows to reach human-level performance without any controller hard coding or expensive interaction with simulated or real environments

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.326058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.326058Z digest=sha256:30b0a60815bda196d88add6bf6c3a0bb712a7d8cde92746544c4e2092f46322e

Pith citing papers

Observation 62ae9cd7-9a10-4c82-ba49-9080037c237c · inbound

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play cites this paper.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Offline Learning of Controllable Diverse Behaviors

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:12.177892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:12.177892Z digest=sha256:e08a5714c84ef22392f64e8d7510102e9155218c952ff6fbf2a413a771ab49e0