Pith. sign in

Paper Citation Record · LEDGER

Teacher-Aware Evolution of Heuristic Programs from Learned Optimization Policies

As of 6 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 0 inbound Pith citation observations for arXiv:2605.10634.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.10634 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-12T05:33:08.169350Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

14 of 14 outbound references displayed

  • verified exact0
  • verified fuzzy14
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 850ac782-11fd-4b63-a36d-ae92a886a523 · outbound

This paper cites Verifiability via iterative policy extraction.

Teacher-Aware Evolution of Heuristic Programs from Learned Optimization Policies Verifiability via iterative policy extraction

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T11:11:31.308650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T05:33:08.169350Z digest=sha256:81a7667d16264d59745376fc23c62b393f90cdbf900b465697db7ba5bd3e25b7

Observation efc51311-1532-4bf0-95ef-aa735ae734b0 · outbound

This paper cites Le, Mohammad Norouzi, and Samy Bengio.

Teacher-Aware Evolution of Heuristic Programs from Learned Optimization Policies Le, Mohammad Norouzi, and Samy Bengio

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T11:11:31.319409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T05:33:08.169350Z digest=sha256:89c4e09d52dd67c9bbeed6af6dd1db6eedfbbd15de7f266f6adeaae4f15124ff

Observation 5d34d461-b269-4ecc-bb69-fd15b579bfa8 · outbound

This paper cites Attention, learn to solve routing problems! In Proceedings of the International Conference on Learning Representations (ICLR 2019).

Teacher-Aware Evolution of Heuristic Programs from Learned Optimization Policies Attention, learn to solve routing problems! In Proceedings of the International Conference on Learning Representations (ICLR 2019)

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T11:11:31.282898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T05:33:08.169350Z digest=sha256:3063509a8a2833d0ee5a998a1ad902e429043cca916d582fc0ff119d1e995f48

Observation 0679628c-2d79-4321-9007-cd92d4f2e951 · outbound

This paper cites Pomo: Policy optimization with multiple optima for reinforcement learning.

Teacher-Aware Evolution of Heuristic Programs from Learned Optimization Policies Pomo: Policy optimization with multiple optima for reinforcement learning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T11:11:31.314563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T05:33:08.169350Z digest=sha256:e9ca1478e5a86c699e3c2fcd82075b05aa3fd5b6d6b0451e6aaee5164f781b7a

Observation 4beb4d66-7d16-44e4-a414-e754283c9edd · outbound

This paper cites Evolution of heuristics: Towards efficient automatic algorithm design using large language models.

Teacher-Aware Evolution of Heuristic Programs from Learned Optimization Policies Evolution of heuristics: Towards efficient automatic algorithm design using large language models

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T11:11:31.337752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T05:33:08.169350Z digest=sha256:9b7f5366a1e2f8c9bf8d4bc4c5812c2359000fd7d1508a7718acdec18c3c59e3

Observation 3f32e41c-0f64-469b-bb01-c9cd68f942e5 · outbound

This paper cites Large language models as evolutionary optimizers.

Teacher-Aware Evolution of Heuristic Programs from Learned Optimization Policies Large language models as evolutionary optimizers

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T11:11:31.294665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T05:33:08.169350Z digest=sha256:c680362e8ddf0984719e57f439ac2c07dac7d84f271d54b4ed641482a1cd02aa

Observation a271185e-04a1-42dc-a09b-d4d34499d46c · outbound

This paper cites Adjustable robust reinforcement learning for online 3d bin packing.

Teacher-Aware Evolution of Heuristic Programs from Learned Optimization Policies Adjustable robust reinforcement learning for online 3d bin packing

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T11:11:31.289384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T05:33:08.169350Z digest=sha256:45ae12af329b0e0d4689e6dfe8003bce223e1057a37e38e33d9bbf7ac2edeba4

Observation a4c01855-39a4-4a88-bead-ef87964d6173 · outbound

This paper cites Neuro-symbolic program synthesis.

Teacher-Aware Evolution of Heuristic Programs from Learned Optimization Policies Neuro-symbolic program synthesis

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T11:11:31.299942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T05:33:08.169350Z digest=sha256:7a86bbbb519b2b211c6644190af1518760f0d7a4c5d0427d8872cc2797091cdb

Observation f756406d-6037-493f-a258-e6bef9ad4c66 · outbound

This paper cites Rusu, Sergio Gomez Colmenarejo, Caglar Gulcehre, Guillaume Desjardins, James Kirk- patrick, Razvan Pascanu, V olodymyr Mnih, Koray Kavukcuoglu, and Raia Hadsell.

Teacher-Aware Evolution of Heuristic Programs from Learned Optimization Policies Rusu, Sergio Gomez Colmenarejo, Caglar Gulcehre, Guillaume Desjardins, James Kirk- patrick, Razvan Pascanu, V olodymyr Mnih, Koray Kavukcuoglu, and Raia Hadsell

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T11:11:31.343224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T05:33:08.169350Z digest=sha256:fa7ff400567b7fc896f904af7c07060831898651411eead1fc51df1ce0251025

Observation 07001b91-5c2e-4bd9-b76d-a531d15649f8 · outbound

This paper cites Programmatically interpretable reinforcement learning.

Teacher-Aware Evolution of Heuristic Programs from Learned Optimization Policies Programmatically interpretable reinforcement learning

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T11:11:31.304183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T05:33:08.169350Z digest=sha256:eef86abe7b20a854cce5366bf90abe993ceb1545d8650c29777be85512c9cd9b

Observation 6b4df8f9-c977-4658-90fb-996563bd2b3f · outbound

This paper cites Efficient heuristics generation for solving combinatorial optimization problems using large language models.

Teacher-Aware Evolution of Heuristic Programs from Learned Optimization Policies Efficient heuristics generation for solving combinatorial optimization problems using large language models

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T11:11:31.331377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T05:33:08.169350Z digest=sha256:9d18d40c20253fea71e73f32fc965072714ab7269fa468d8707aa0351d9ff6ab

Observation b209f907-7e82-4628-9659-ed8383c82ed1 · outbound

This paper cites Reevo: Large language models as hyper-heuristics with reflective evolution.

Teacher-Aware Evolution of Heuristic Programs from Learned Optimization Policies Reevo: Large language models as hyper-heuristics with reflective evolution

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T11:11:31.324533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T05:33:08.169350Z digest=sha256:f9d8d1046318425e7d52520b8b2e3f50bba2fe0e31c3235de627eeb7393e5d91

Observation e30bb3ea-6d1a-4bc5-8286-e1f167a893f7 · outbound

This paper cites Learning to dispatch for job shop scheduling via deep reinforcement learning.

Teacher-Aware Evolution of Heuristic Programs from Learned Optimization Policies Learning to dispatch for job shop scheduling via deep reinforcement learning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T11:11:31.267456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T05:33:08.169350Z digest=sha256:690030dc249cf74318e20bc4f2e46ebd4024137ff444ec168eb40ebe2a9b0752

Observation cfad38b5-ba61-429f-ad74-9b0faef60138 · outbound

This paper cites Monte carlo tree search for comprehen- sive exploration in llm-based automatic heuristic design.

Teacher-Aware Evolution of Heuristic Programs from Learned Optimization Policies Monte carlo tree search for comprehen- sive exploration in llm-based automatic heuristic design

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T11:11:31.276198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T05:33:08.169350Z digest=sha256:2812dbc0f9a493102726c64423fafe942b5ece20876ae2a8deb6dbd1914fe834

Pith citing papers

No inbound Pith citation observations are available.