Pith. sign in

Paper Citation Record · LEDGER

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies

As of 17 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 0 inbound Pith citation observations for arXiv:2506.14162.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.14162 v1

Coverage vector

measured 34 of 34 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:24:26.960195Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

34 of 34 outbound references displayed

  • verified exact0
  • verified fuzzy27
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b2e05dca-4bb0-494a-88b4-74b3c57f7952 · outbound

This paper cites write newline.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T00:24:24.188306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:24:24.188306Z digest=sha256:8b5f8c679d70fc04faa8d8c450ce3430348a6a9723158325a453798fa1c60aa2

Observation fdafc41b-37d2-495b-8aa3-428532357391 · outbound

This paper cites Recursive program synthesis.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Recursive program synthesis

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:32.350067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:24.254476Z digest=sha256:da6fcab0f3a4ca9a7964faddce3ab32371bee70de8179a5ab36f18ef73c02d0c

Observation 0b763c18-8921-4c08-ad4b-66a29ad62d6b · outbound

This paper cites Unveiling options with neural network decomposition.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Unveiling options with neural network decomposition

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:32.037426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:24.333427Z digest=sha256:f847e37a617666b547ada6d05991f10ecce8cf90ca28cbb8763b578de000e247

Observation 14fe9c8b-824a-4237-92e7-99dd372cf9ec · outbound

This paper cites Qwen technical report, 2023.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Qwen technical report, 2023

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:31.898822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:24.397393Z digest=sha256:e0786dfcd30d643754201ae20dd24468147637cd5c229616d4338b3b0df7acb0

Observation c891efc3-6e21-4b61-bd9e-09d5cb8e8b9c · outbound

This paper cites Verifiable reinforcement learning via policy extraction.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Verifiable reinforcement learning via policy extraction

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:31.670160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:24.455899Z digest=sha256:88de54c5a1421b37e1541a401029c837a1f09c118ae4d7d1b4d777f527d99382

Observation 6f7e1649-ef5e-4dcf-a19a-dd4c94e9b9ef · outbound

This paper cites Look where you look! saliency-guided q-networks for generalization in visual reinforcement learning.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Look where you look! saliency-guided q-networks for generalization in visual reinforcement learning

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:31.518960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:24.511312Z digest=sha256:ee9b23eb7aefce376e8994ca18af85e38f60e43caddadd5edb85b6176eed36f2

Observation c0928049-af5e-4fa8-a75c-86986db7d231 · outbound

This paper cites Olausson, Lionel Wong, Gabriel Grand, Joshua B.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Olausson, Lionel Wong, Gabriel Grand, Joshua B

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:31.369063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:24.568290Z digest=sha256:c9f7df10beaeb2c781d17eca6d2e380f8e9d2b672721dd9e1c2b2a970169cd83

Observation 3d2f82a4-d236-4da1-a742-c013e89aba84 · outbound

This paper cites Babble: Learning better abstractions with e-graphs and anti-unification.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Babble: Learning better abstractions with e-graphs and anti-unification

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:31.173696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:24.623900Z digest=sha256:483ea45ea6ac788abc6edf60529615a04d731b7daaa872309da49b64ff61b645

Observation f3e66895-f08b-47e6-88da-df02c52173dc · outbound

This paper cites Empirical evaluation of gated recurrent neural networks on sequence modeling, 2014.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Empirical evaluation of gated recurrent neural networks on sequence modeling, 2014

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T00:24:24.682449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:24:24.682449Z digest=sha256:d9fc729a4a8b21ddbb99ed3a701e38e2490a6aeca76765dec80e5e346194bf65

Observation ab7354ee-ee7f-491c-a0bc-a74a46a0dda8 · outbound

This paper cites Dijkstra.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Dijkstra

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:24:24.783231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:24:24.783231Z digest=sha256:12c6755b940a3effe64374248b642b476d40635f7a238b094808c098050f96f3

Observation d48d8739-afd1-41ed-b373-15dcd0bdbc56 · outbound

This paper cites Dreamcoder: growing generalizable, interpretable knowledge with wake?sleep bayesian program learning.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Dreamcoder: growing generalizable, interpretable knowledge with wake?sleep bayesian program learning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:30.986298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:24.831681Z digest=sha256:2a4f2f33e043139e5fc458f236e23c9d3b523730343e959e867c7ddeaed09f4b

Observation 12478ca4-927b-41c2-88df-905531747906 · outbound

This paper cites Taylor, A.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Taylor, A

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:30.828607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:24.881831Z digest=sha256:c992eac810bcaf11d543d260d0ff9e960d62d071dca8b08c1b231c14a6d1f23f

Observation 1bf2dd75-2e31-412e-925e-f24bdada3e56 · outbound

This paper cites Long short-term memory.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Long short-term memory

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:30.673485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:25.014625Z digest=sha256:6b7fd8e5e8eeacbbe09c10baf7ea838944d85c8565110f29a2ca7f8569c5a11f

Observation 536ec4aa-9c9a-4c30-91cf-0bdcf6b9f3b1 · outbound

This paper cites Synthesizing programmatic policies that inductively generalize.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Synthesizing programmatic policies that inductively generalize

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:30.524241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:25.090304Z digest=sha256:06d98ec73535bcd8659fa4594e2e1d973930d497c0fa6a40f75027699f7fd0a3

Observation c070fefd-9fd5-4432-801c-e969b4043981 · outbound

This paper cites Inferring algorithmic patterns with stack-augmented recurrent nets.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Inferring algorithmic patterns with stack-augmented recurrent nets

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:30.336928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:25.170492Z digest=sha256:d8a49e9e408c65fcfbe0466f553ab54d9ae5bb4e42d6ba13b6f9aa482c537864

Observation 7a56048a-8b13-4405-83a1-b4a242eb4120 · outbound

This paper cites Kingma and Max Welling.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Kingma and Max Welling

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:30.154175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:25.257096Z digest=sha256:89a8c4359d472543d04cfa21faed55a86601b804f816596761521892097af5a3

Observation 15a664e3-51d0-4399-a195-3acb48f4dc9c · outbound

This paper cites Lillicrap, Jonathan J.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Lillicrap, Jonathan J

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T00:24:25.350457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:24:25.350457Z digest=sha256:d469ed6436237fd057527a340360f46087cb0ea4d0c649866ff723999d779a8a

Observation c016d48f-9aa2-48a3-acd7-6628d1deeb7a · outbound

This paper cites Rubinstein, and Yohai Gat.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Rubinstein, and Yohai Gat

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:30.005284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:25.479684Z digest=sha256:7cc781282c2ee8514a688bd20c4cd487ad02dd539fe5ab02c6414f31408409bd

Observation 224e344c-4365-4974-9340-c7f56ba7493a · outbound

This paper cites Human-level control through deep reinforcement learning.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Human-level control through deep reinforcement learning

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:29.877806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:25.552952Z digest=sha256:061b60a3b135ef636b079377b8cbb16d0ec25c8afe56340bdf4cae1c9e633dde

Observation d7db9feb-16fc-4d64-a77a-c4b21e20c2aa · outbound

This paper cites Palmarini, Christopher G.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Palmarini, Christopher G

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:29.723542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:25.657125Z digest=sha256:c34aae320418826f99171eaaad2c986c4b2de909853a35c007cbaacf86f984eb

Observation 50144bc9-2ff7-426b-954d-0e69c816b9e5 · outbound

This paper cites Programmatic reinforcement learning without oracles.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Programmatic reinforcement learning without oracles

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:29.538860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:25.735850Z digest=sha256:d17191c6d3e09c89c0a6072998f017311fa63d35aa46703b8055951352570906

Observation b6bc77e5-f0b0-419d-9d5b-40b1314792a8 · outbound

This paper cites Synthesizing libraries of programs with auxiliary functions.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Synthesizing libraries of programs with auxiliary functions

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:29.358860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:25.851549Z digest=sha256:8653485f4419dcdb2e7751f5762ae5a5bd4c3d72d80cd60afab6cc60d0ed91a6

Observation 6504eb1e-9e05-4fdb-bc7e-1155b9add056 · outbound

This paper cites Pawan Kumar, Emilien Dupont, Francisco J.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Pawan Kumar, Emilien Dupont, Francisco J

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:29.146352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:25.953863Z digest=sha256:df573cf66a3605688da06a3248dcdc0a6a45a9001c547af3693133ff16cdc8af

Observation 8e812504-1a83-45d2-a8a1-b5a80a9609b1 · outbound

This paper cites A reduction of imitation learning and structured prediction to no-regret online learning.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies A reduction of imitation learning and structured prediction to no-regret online learning

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:28.960074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:26.043133Z digest=sha256:5e30243cf7dde77e8e1b406b5e0c7529ca6737b41de9a0a85b68a02ed314f483

Observation e8f8a462-93c9-4c27-86f2-13c01bceaf97 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Proximal Policy Optimization Algorithms

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T00:24:26.131603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:24:26.131603Z digest=sha256:4979a505041b9664b9e53da6c5c423921bb2d60764ba5d62eaa9680cad81fd79

Observation bc4f9ab6-fcc9-4521-8974-70fab293be19 · outbound

This paper cites Siegelmann and Eduardo D.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Siegelmann and Eduardo D

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:28.760277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:26.236979Z digest=sha256:d86941043b75bd4786ec2d567a9a0ef97927c030ebdc0d5ec8f03814cb9eb553

Observation 7575e484-c8f9-45cb-aaab-b2921e57e7ff · outbound

This paper cites Siegelmann and Eduardo D.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Siegelmann and Eduardo D

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:28.498717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:26.328571Z digest=sha256:bbffcaa04bc1c5d4cda64e7f8b1eda47c6ceb93c0b0ba3a441b6d9c94b7c72b1

Observation 33596d3e-63e0-42e6-b948-62df59513989 · outbound

This paper cites an unresolved cited work.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:24:28.249927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:26.426924Z digest=sha256:2b490e6526fe57ac799c2778de62aa0efb59acc1e805cbeb6ac960cc4e954c99

Observation 66fe22cf-15ff-43bd-831d-d795793f9263 · outbound

This paper cites an unresolved cited work.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:24:28.070023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:26.491726Z digest=sha256:8aaaca4e2e5a8f5339ef1b6f2f562974cd74e05a77600b19204fbd02be56e95d

Observation f6f8d18d-6dfb-45d4-a4ea-df2db9e87bd9 · outbound

This paper cites Deshmukh, Sela Mador-Haim, Milo M.K.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Deshmukh, Sela Mador-Haim, Milo M.K

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:27.925595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:26.554607Z digest=sha256:827b552e91d96c2ab5a4bceae5483ac89e1fa35090f9d5612cf09feb187a5004

Observation a5115db3-76ad-43eb-b8eb-e2e00e3d9a77 · outbound

This paper cites Programmatically interpretable reinforcement learning.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Programmatically interpretable reinforcement learning

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:27.763126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:26.690608Z digest=sha256:8ad7ba77de2665af56fff4e229524aa624a75d73ecb5f3ce61f788d0f2893d87

Observation 8f06d682-8153-4594-9e0a-43bdd1ede863 · outbound

This paper cites Imitation-projected programmatic reinforcement learning.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Imitation-projected programmatic reinforcement learning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:27.572256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:26.798040Z digest=sha256:ae2c3cae969d973811767a6a74a67916d4c90c0855d4f1b16712593e0a480751

Observation 5b048b05-b984-441c-93c2-9940990be8e4 · outbound

This paper cites On the practical computational power of finite precision rnns for language recognition.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies On the practical computational power of finite precision rnns for language recognition

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:27.392607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:26.878300Z digest=sha256:ae32b78a24fdb051ec7cfd3a42b7a7d901c34a9f07d333a49d85052ea5e266a0

Observation 314ae24f-e9ba-43b2-9239-45322674bac9 · outbound

This paper cites Torcs, the open racing car simulator.

Common Benchmarks Undervalue the Generalization Power of Programmatic Policies Torcs, the open racing car simulator

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:24:27.227135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T00:24:26.960195Z digest=sha256:8acaad1aa88e73d32489596d908b1bc3ac0b4f7ffd8ea7501cba6b1f34993e98

Pith citing papers

No inbound Pith citation observations are available.