Pith. sign in

Paper Citation Record · LEDGER

Student-Informed Teacher Training

As of 22 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 2 inbound Pith citation observations for arXiv:2412.09149.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.09149 v2

Coverage vector

measured 46 of 46 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T17:19:47.254748Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:36:14.419285Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T00:59:58.944954Z

Reference resolution

46 of 46 outbound references displayed

  • verified exact0
  • verified fuzzy27
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1b4b5ca8-a8b8-468f-9807-7ac7190e832a · outbound

This paper cites On the role of the action space in robot manipulation learning and sim-to-real transfer.

Student-Informed Teacher Training On the role of the action space in robot manipulation learning and sim-to-real transfer

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:54.076988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:45.091043Z digest=sha256:6b66268b58aa1d7d43b2d15d8e5480a79968f1a40ff20b7d01cda9dd0e2a3945

Observation ad098a11-0045-4715-8564-2c42450ea739 · outbound

This paper cites Data generation as sequential decision making.

Student-Informed Teacher Training Data generation as sequential decision making

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:53.924897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:45.134768Z digest=sha256:ee259e85308e3f414392aa51308a56dbd4eb9f0d147d9d1f22e7fccacde9793d

Observation 4b169bcc-73e3-42b6-80b3-877272e0762e · outbound

This paper cites A framework for behavioural cloning.

Student-Informed Teacher Training A framework for behavioural cloning

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:53.734764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:45.176231Z digest=sha256:5f99342fe25a347bac0736cba137673da67f941a68ab1239265b8759d8ed8905

Observation a2763ca7-050f-4357-8472-32520a08aa68 · outbound

This paper cites Openai gym, 2016.

Student-Informed Teacher Training Openai gym, 2016

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T17:19:45.234821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:19:45.234821Z digest=sha256:561e66363c7f1444aa5b74a2ad128d445fc0ab71c5d80b27717e83bd85151076

Observation e6c79ca6-7631-47f5-b25b-1022de1ec1ce · outbound

This paper cites Soloparkour: Constrained reinforcement learning for visual locomotion from privileged experience.

Student-Informed Teacher Training Soloparkour: Constrained reinforcement learning for visual locomotion from privileged experience

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:53.434759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:45.278164Z digest=sha256:d22b514f44f0dd0185cace96ca61966c2a4cfd2e8fac7f20ad4a474654f860b1

Observation 46fcd4e3-2cd0-4ae3-9624-2bbc733b5b83 · outbound

This paper cites Learning by cheating.

Student-Informed Teacher Training Learning by cheating

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:53.255729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:45.324929Z digest=sha256:f1868bb21ab9b732407fc8070e1c4b6b2d9f66e394351cb20187d82b2dbf7db1

Observation 1d0f962d-38b1-4d7a-9f40-5941185c01e1 · outbound

This paper cites Diffusion policy: Visuomotor policy learning via action diffusion.

Student-Informed Teacher Training Diffusion policy: Visuomotor policy learning via action diffusion

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:53.074751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:45.361032Z digest=sha256:44908a94ef65763f701387747e195d919b3335097abeea40434efbca563fa348

Observation da9585ea-5bfd-412d-8d65-31e5bda06849 · outbound

This paper cites Scaling cross-embodied learning: One policy for manipulation, navigation, locomotion and aviation.

Student-Informed Teacher Training Scaling cross-embodied learning: One policy for manipulation, navigation, locomotion and aviation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:52.894760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:45.394766Z digest=sha256:7e08b23c68b03800f80a208b9c913d2225732d2fef3fce68126e86a6ed818bd6

Observation b9254128-60b7-4f5a-b954-4bca264b5c10 · outbound

This paper cites Deep whole-body control: Learning a unified policy for manipulation and locomotion.

Student-Informed Teacher Training Deep whole-body control: Learning a unified policy for manipulation and locomotion

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:52.704746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:45.434851Z digest=sha256:5f14a7177971907306157a6b70cbaefd42c6a4e8cbc295680dc0be77eea451dc

Observation 56033bf4-31bd-47d7-859e-774797ca09ee · outbound

This paper cites Generative adversarial nets.

Student-Informed Teacher Training Generative adversarial nets

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T17:19:45.494938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:19:45.494938Z digest=sha256:06ba3a4305dfd8ea67d711f1ef35214b901b7afa80c52246d49cfba35a916816

Observation 0b8ec5ff-bd9d-42ae-98cc-e152e13c42a7 · outbound

This paper cites Designing skill-compatible AI : Methodologies and frameworks in chess.

Student-Informed Teacher Training Designing skill-compatible AI : Methodologies and frameworks in chess

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:52.373977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:45.544828Z digest=sha256:51ac1a60ff46f13a454ea2f01fcb12aa7e8389cd0b80e9e14a97f6874d032779

Observation 781e952c-90c8-4d4f-837d-646998d4dd48 · outbound

This paper cites Bridging the sim-to-real gap from the information bottleneck perspective.

Student-Informed Teacher Training Bridging the sim-to-real gap from the information bottleneck perspective

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:52.192724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:45.594755Z digest=sha256:b35aacb8e4f96a6befde5db98ff941207bf7fcca4753010700c909176e590465

Observation 3cf318f2-4e08-46d9-95b7-9ecfa41de9de · outbound

This paper cites Generative adversarial imitation learning.

Student-Informed Teacher Training Generative adversarial imitation learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T17:19:45.645086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:19:45.645086Z digest=sha256:7d3191a1a84bbad8a1e1b396ed2e9201aef990947dcf14e4cf7b680f491bbfc2

Observation 071664ce-ca06-4d03-831a-1c8f9a387764 · outbound

This paper cites Hu, James Springer, Oleh Rybkin, and Dinesh Jayaraman.

Student-Informed Teacher Training Hu, James Springer, Oleh Rybkin, and Dinesh Jayaraman

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:51.903475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:45.705021Z digest=sha256:05b32aee2a0e6f824beeceda888b17e35f65b1895793ad21298658507d31261e

Observation 1e76b09d-9425-40e3-aecc-c3759933c2c3 · outbound

This paper cites Champion-level drone racing using deep reinforcement learning.

Student-Informed Teacher Training Champion-level drone racing using deep reinforcement learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T17:19:45.764751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:19:45.764751Z digest=sha256:faa10102329cfb9993b28fdf35ddfa9aa81809b4ca63f9ad0c3d8def2fb3df04

Observation 8df427f2-9207-408b-9649-fb4ce55be0e5 · outbound

This paper cites Adam: A method for stochastic optimization.

Student-Informed Teacher Training Adam: A method for stochastic optimization

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T17:19:45.814884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:19:45.814884Z digest=sha256:2d26914ea5ab891a1b56c24b92fb085026d4ca3506c08f6ce882a4de02f5bb84

Observation 9a2d48e1-70d7-48b4-8762-d85e15dbb76e · outbound

This paper cites Multi-stage cable routing through hierarchical imitation learning.

Student-Informed Teacher Training Multi-stage cable routing through hierarchical imitation learning

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:51.524754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:45.864887Z digest=sha256:861d113184dc24f91013d41ac4d6291f420e38ad868d8a54ec18561c5ba0bd6c

Observation d40616de-54d7-4d3a-b579-741441cef8b6 · outbound

This paper cites rl-games: A high-performance framework for reinforcement learning.

Student-Informed Teacher Training rl-games: A high-performance framework for reinforcement learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T17:19:45.914762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:19:45.914762Z digest=sha256:6fb8b87dbc4a5c7601c2e99384fe4ac9f14611199a0e8378c69336a3be2e8755

Observation 897e1b74-775f-4a79-be53-f299be6eb8a5 · outbound

This paper cites Human-level control through deep reinforcement learning.

Student-Informed Teacher Training Human-level control through deep reinforcement learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T17:19:45.954736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:19:45.954736Z digest=sha256:949218e73c190b5661426d43d76b44cb3b14ac7bd39d1e3a697ba3d711548f26

Observation 1538c22a-a674-41c5-b613-a1fbf22a10f1 · outbound

This paper cites Leveraging fully observable policies for learning under partial observability.

Student-Informed Teacher Training Leveraging fully observable policies for learning under partial observability

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:51.164749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:45.982549Z digest=sha256:6820d3235579bc19b9296da11383b070c8e5a33057dba0f6483d80206b8ecfd3

Observation 55f2fcbd-bb5b-4c67-9c6c-b21658e8e427 · outbound

This paper cites an unresolved cited work.

Student-Informed Teacher Training Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T17:19:46.024750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:19:46.024750Z digest=sha256:508381961867a1063c48a6d10374033d73b606a264f18fe2b1fbf81b538242f1

Observation 78889d21-9a1d-472f-b107-973e72c4871d · outbound

This paper cites An algorithmic perspective on imitation learning.

Student-Informed Teacher Training An algorithmic perspective on imitation learning

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:50.889055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:46.054750Z digest=sha256:de532aaac5891db813c240041f5db3dc7f8c24be33d94cbd56e947ec44d977b7

Observation 52404c5c-df70-429e-9664-672874be8432 · outbound

This paper cites Can increasing input dimensionality improve deep reinforcement learning? In International conference on machine learning, pp.\ 7424--7433.

Student-Informed Teacher Training Can increasing input dimensionality improve deep reinforcement learning? In International conference on machine learning, pp.\ 7424--7433

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:50.738721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:46.084786Z digest=sha256:a73a9a1e187f297d9ccb569ec7ef42dbaf1491c340f203d387d5bd5346f13fd4

Observation 3d7b6623-0904-4e05-a672-ba48fcdbc57f · outbound

This paper cites Automatic differentiation in pytorch, 2017.

Student-Informed Teacher Training Automatic differentiation in pytorch, 2017

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:50.581530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:46.114750Z digest=sha256:626e2cdbcf1286687dcd5bf9439d9b0dcb547f7a452241cad0d88593211d4d05

Observation a53a6ac4-2c3e-4223-a07e-13738870e498 · outbound

This paper cites Curiosity-driven exploration by self-supervised prediction.

Student-Informed Teacher Training Curiosity-driven exploration by self-supervised prediction

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T17:19:46.165050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:19:46.165050Z digest=sha256:f0eca6047e617be71f636e8cfbcf5cbc6f0ef4d862201a0a568a496edb010319

Observation 6db48e52-65b5-4d0e-b5ec-2cda546bfa10 · outbound

This paper cites Asymmetric actor critic for image-based robot learning.

Student-Informed Teacher Training Asymmetric actor critic for image-based robot learning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:50.354990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:46.215076Z digest=sha256:34a8bcf76f2bc21e04270212c5054bf0e566cf54cc1dec7349e95db3b07beeab

Observation 28cc79ac-472c-400d-9a2a-2127558f66bb · outbound

This paper cites Real-world humanoid locomotion with reinforcement learning.

Student-Informed Teacher Training Real-world humanoid locomotion with reinforcement learning

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:50.219105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:46.273231Z digest=sha256:54374726bb04eb843e782723540be041ce9cd1291e8883e568dd849372011e2f

Observation c9d4e944-f2db-457b-a5a3-c5d9bc2fe86c · outbound

This paper cites Stable-baselines3: Reliable reinforcement learning implementations.

Student-Informed Teacher Training Stable-baselines3: Reliable reinforcement learning implementations

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T17:19:46.324838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:19:46.324838Z digest=sha256:1d02499c6806df13a44569cfe3a1a2c525a1b3e8c4b4cf7c4ec6262c747358ea

Observation 6076b84c-5c43-4b9f-bd27-1c6e19363795 · outbound

This paper cites Reinforcement and Imitation Learning via Interactive No-Regret Learning.

Student-Informed Teacher Training Reinforcement and Imitation Learning via Interactive No-Regret Learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T17:19:46.384751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:19:46.384751Z digest=sha256:4f67412fe727f7203738edab629add4a77b54205bde663c54f30fd445eb86153

Observation 090557b2-6bfb-49d0-86c9-c6ad39e3b0dd · outbound

This paper cites A reduction of imitation learning and structured prediction to no-regret online learning.

Student-Informed Teacher Training A reduction of imitation learning and structured prediction to no-regret online learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T17:19:46.434755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:19:46.434755Z digest=sha256:113b3e74caa7e794e540b410007dbbb0470d00534c0730374a7635f15b4ce153

Observation 64119360-2e65-4566-a577-0cd6d24dda80 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Student-Informed Teacher Training Proximal Policy Optimization Algorithms

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T17:19:46.495336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:19:46.495336Z digest=sha256:2d007d92ab9afaf80c48e2ce9efc30881ad25072ac9571d4c08173e0420f0ab4

Observation 19dd418f-6ba4-4a42-b87f-f99482c13da3 · outbound

This paper cites Tgrl: An algorithm for teacher guided reinforcement learning.

Student-Informed Teacher Training Tgrl: An algorithm for teacher guided reinforcement learning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:49.864839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:46.544750Z digest=sha256:53eb4e835b5ce17f473006675f18eb012883b236912ae1ca38a7a15e23abe5b6

Observation 588bfa44-c4d2-4728-813b-c0541e5a3429 · outbound

This paper cites Mastering the game of go with deep neural networks and tree search.

Student-Informed Teacher Training Mastering the game of go with deep neural networks and tree search

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T17:19:46.625739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:19:46.625739Z digest=sha256:5b1b7b64a965871a8f6078bab2543894f2f742ccefa9427dece51d8b23d6b955

Observation 91ce2587-c16e-4087-a25b-25804349cc1c · outbound

This paper cites Flightmare: A flexible quadrotor simulator.

Student-Informed Teacher Training Flightmare: A flexible quadrotor simulator

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:49.601665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:46.675218Z digest=sha256:beeafe031e8682733edf03b93a85be15642991089784fc974ca1aba3bb81048b

Observation 961391ee-bd1e-4fc3-9205-c07a901a9a4c · outbound

This paper cites Sequence model imitation learning with unobserved contexts.

Student-Informed Teacher Training Sequence model imitation learning with unobserved contexts

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:49.442258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:46.711515Z digest=sha256:89b26efb77ebcdbfd556e02fccf15f4c7cc9dde9e564842b0124ed7e311257c3

Observation eb1d5689-d0fe-465d-a6b6-cf233275ead8 · outbound

This paper cites A reduction from apprenticeship learning to classification.

Student-Informed Teacher Training A reduction from apprenticeship learning to classification

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:49.306221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:46.745612Z digest=sha256:74d8245bae465206796d4ea8016666ff8529f216dca908ba0d9280021e6b420c

Observation 8bc498b8-850d-4c4d-b9cf-94b06620f0fd · outbound

This paper cites Octo: An open-source generalist robot policy.

Student-Informed Teacher Training Octo: An open-source generalist robot policy

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:49.187834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:46.795965Z digest=sha256:f41396aadad08e09da96acd4afdb41023f3668a54095a9472c037bcea5502faf

Observation a4be21ce-fa01-42be-90ec-a29119d0305c · outbound

This paper cites Grandmaster level in starcraft ii using multi-agent reinforcement learning.

Student-Informed Teacher Training Grandmaster level in starcraft ii using multi-agent reinforcement learning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T17:19:46.864841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:19:46.864841Z digest=sha256:70c2241ed26606f710cdbf01d851a7e6fa0886a7b2a21eb15d2cd601c4fa8f61

Observation ec6cdaef-d8d9-4bec-ab69-dc99c71078c1 · outbound

This paper cites Impossibly good experts and how to follow them.

Student-Informed Teacher Training Impossibly good experts and how to follow them

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:48.934839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:46.914773Z digest=sha256:3c4d9b33bed320325a4c506c12935eeeb5a34d4934d88745cc4a1dc480502bb4

Observation 1b9b29c4-2197-4597-8326-eb0b05e8e8d4 · outbound

This paper cites Robust asymmetric learning in pomdps.

Student-Informed Teacher Training Robust asymmetric learning in pomdps

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:48.788448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:46.974751Z digest=sha256:8d401b3789d63704538579d9c262f5bdb2bcf7aca81850b96c0a993fe1d48ac6

Observation 5eec6154-c674-4369-8222-85a8cae30300 · outbound

This paper cites Bridging the imitation gap by adaptive insubordination.

Student-Informed Teacher Training Bridging the imitation gap by adaptive insubordination

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:48.604883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:47.024759Z digest=sha256:bf31d2e65c194b2465dfe6755431126b9f7945c31826965cb828615d5caff594

Observation d1ab7e8d-fefc-4351-8ac6-6d3edcbd42bf · outbound

This paper cites Error bounds of imitating policies and environments.

Student-Informed Teacher Training Error bounds of imitating policies and environments

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:19:48.443210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T17:19:47.055447Z digest=sha256:990fa8a9161abebefac216ff7bec6cd92a38cb1b2177f5e68bc8de5b7974ebda

Observation 6d512e1f-dac4-465a-850b-e80e0c4f45e7 · outbound

This paper cites write newline.

Student-Informed Teacher Training write newline

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T17:19:47.104750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:19:47.104750Z digest=sha256:51d122fb7ef4f1f23c949ea5e92599205158b1422521b2db4b8c4d7f3f921a63

Observation e7cafc18-2d98-4ffb-8953-13a48db1ca0e · outbound

This paper cites @esa (Ref.

Student-Informed Teacher Training @esa (Ref

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T17:19:47.154736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:19:47.154736Z digest=sha256:1c280a2bfd7632872d3f006ffa3f0b587f5cae2b74fc77ada90d42390c94cc06

Observation 7eca8665-f795-4fef-b991-203bbc1a6b78 · outbound

This paper cites an unresolved cited work.

Student-Informed Teacher Training Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T17:19:47.204750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:19:47.204750Z digest=sha256:6a5c1a156030dc118e806f240bddcb96a6274f357736f12c867a9afe2f0613dd

Observation e550cca8-d219-405e-9353-c852120634c1 · outbound

This paper cites KL-Div Gradient.

Student-Informed Teacher Training KL-Div Gradient

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T17:19:47.254748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:19:47.254748Z digest=sha256:0da22ed31a8c0256d6e31086f8339ae1b3310c8589a0693c68001f6d543ea8bd

Pith citing papers

Observation dffe9795-2301-4f17-867a-023842cb3762 · inbound

Distilling Realizable Students from Unrealizable Teachers cites this paper.

Distilling Realizable Students from Unrealizable Teachers Student-Informed Teacher Training

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T21:36:14.419285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:36:14.419285Z digest=sha256:b933784aaff84a6155ba63f93af7460835a499f3029ee738c4ab92c4bee497c1

Observation 1054dace-c1ec-4079-a190-8a8c09fd0dd1 · inbound

Light-Loco-Parkour: Versatile Perceptive Whole-Body Locomotion via Multi-Skill Distillation cites this paper.

Light-Loco-Parkour: Versatile Perceptive Whole-Body Locomotion via Multi-Skill Distillation Student-Informed Teacher Training

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-08-05T00:59:59.036623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T00:59:58.080577Z digest=sha256:cbffb2f98585bb11d40e8173f82749cecc0fb4bc026e61226a49f169be910538