Pith. sign in

Paper Citation Record · LEDGER

In-Context Learning as Implicit Policy Gradient

As of 7 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 0 inbound Pith citation observations for arXiv:2607.23153.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.23153 v1

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T03:32:24.383801Z

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

47 of 47 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved46
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5033dc2a-7c41-4b21-bfb1-a41edc07e80f · outbound

This paper cites Advances in neural information processing systems , volume=.

In-Context Learning as Implicit Policy Gradient Advances in neural information processing systems , volume=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.145593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.145593Z digest=sha256:9b49f6ccedd11252b5cb969238f48c95542270f56178b1fb8e5ef14d85a500dd

Observation 3b04400d-0cf7-42c5-b304-b66e84f3aeaa · outbound

This paper cites International Conference on Machine Learning , pages=.

In-Context Learning as Implicit Policy Gradient International Conference on Machine Learning , pages=

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.152289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.152289Z digest=sha256:227fb80334b403b4c9cd407d428b1e71a550c42942d4eef919e1df5c2a068540

Observation 6da39227-a068-4dc8-922e-393712d7ec0e · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

In-Context Learning as Implicit Policy Gradient Advances in Neural Information Processing Systems , volume=

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.157301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.157301Z digest=sha256:21d74163a7cea840a74708ded0fc29a8bfb2612c4b605de3238a89ac7921a6d9

Observation f4f272ef-77fd-4193-b2bc-0c627450b6a0 · outbound

This paper cites Large Language Models as Optimizers.

In-Context Learning as Implicit Policy Gradient Large Language Models as Optimizers

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.162820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.162820Z digest=sha256:2f6877469c552554fae7f24ea189354067c4dc01868c778c093615ed75a94a17

Observation 36fe7e6e-dbee-4fb7-a77c-98bed0a65d6e · outbound

This paper cites Findings of the Association for Computational Linguistics: ACL 2023 , pages=.

In-Context Learning as Implicit Policy Gradient Findings of the Association for Computational Linguistics: ACL 2023 , pages=

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.168207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.168207Z digest=sha256:4a6f70583e574862c71d524f884070c9a3535bc24664fb732706d9d89d4bcd2d

Observation 8b735784-7495-4890-9dcf-f9fdc8c3ce48 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

In-Context Learning as Implicit Policy Gradient Advances in Neural Information Processing Systems , volume=

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.172884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.172884Z digest=sha256:a8032d5b3572b825fedb67a618f4d69bf0f73b6c9605c8c60fa2c2bc66c8fecc

Observation 202835aa-0c13-4d4e-ad2c-1b915909362f · outbound

This paper cites International Conference on Learning Representations , year=.

In-Context Learning as Implicit Policy Gradient International Conference on Learning Representations , year=

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.177857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.177857Z digest=sha256:d895e62856f12f32d3a83eeca03b936d77eb2747c8febaad4f6458f42c31af20

Observation 3f72c1b8-abff-4bce-9bd7-dd874486f9f5 · outbound

This paper cites International Conference on Learning Representations , year=.

In-Context Learning as Implicit Policy Gradient International Conference on Learning Representations , year=

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.183464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.183464Z digest=sha256:e0e70c2d3940340b08f332f3eb5f7948ae78b6d6f0f5676f9d362eccf7ff56a5

Observation 9a8584b1-8675-475e-82af-d065057cb640 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

In-Context Learning as Implicit Policy Gradient Advances in Neural Information Processing Systems , volume=

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.188567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.188567Z digest=sha256:0e35df773578d20b0c4af1bcff24f120ea548e5173d7cd3bf938168bdc87a6dc

Observation 83fcc5f2-4629-4a53-a95d-ca611f35ea69 · outbound

This paper cites In-context Learning and Induction Heads.

In-Context Learning as Implicit Policy Gradient In-context Learning and Induction Heads

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.193341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.193341Z digest=sha256:ed594942f09c6527b3d81cb273961fe3e077cd9315234359162cfd1e8b24d0b0

Observation f1032574-869d-4cae-9ad4-c77cbf3668c8 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

In-Context Learning as Implicit Policy Gradient Advances in Neural Information Processing Systems , volume=

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.198200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.198200Z digest=sha256:d97c7d47881f8096eb6bc4aec917f1660178915e9afdc2dc8445189825a65cd8

Observation 2d81438d-1dfb-466b-8601-64ccbed63264 · outbound

This paper cites International Conference on Machine Learning , year=.

In-Context Learning as Implicit Policy Gradient International Conference on Machine Learning , year=

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.202731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.202731Z digest=sha256:4f729b30e7d4d0ed927dea28e148236c2b436e832c62d29f8f5d24506f10f77a

Observation 725d43e2-a0dd-469a-a5fa-e052fa326dfc · outbound

This paper cites International Conference on Machine Learning , pages=.

In-Context Learning as Implicit Policy Gradient International Conference on Machine Learning , pages=

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.207556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.207556Z digest=sha256:8964a73dfb0fade91363e6db906e9931ae316c927d6deedc8e552fb1b5b53221

Observation a1bb02dd-2614-4e7b-81e8-90b976f618f8 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

In-Context Learning as Implicit Policy Gradient Advances in Neural Information Processing Systems , volume=

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.212560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.212560Z digest=sha256:304612de9b3e443ea66f312ae76deeac52bcd767709f6e7e2c00d4ed64c974a6

Observation f8a94a0f-a5fa-48d2-90ad-2f26252e6d29 · outbound

This paper cites International conference on machine learning , pages=.

In-Context Learning as Implicit Policy Gradient International conference on machine learning , pages=

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.217332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.217332Z digest=sha256:c59d745a876a41a93aa7da5ead2f841f95a29e1798150eab17749d0b1304aa34

Observation 405a149d-67e1-422b-853c-6dfe8619358e · outbound

This paper cites Proximal Policy Optimization Algorithms.

In-Context Learning as Implicit Policy Gradient Proximal Policy Optimization Algorithms

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.222698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.222698Z digest=sha256:60ec5936b6b64517126ad0e85a5e6031fd4feef049bd9b9b0d357a57684dadc4

Observation 32f46e90-0747-489c-bd00-bd240c34fe4e · outbound

This paper cites 2025 , month=.

In-Context Learning as Implicit Policy Gradient 2025 , month=

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.228317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.228317Z digest=sha256:682f7b14feb8839f2c2d8d95af530d03dc5a7401b377e08091d173a48e1a828b

Observation 1af7114e-c0e4-4557-be3d-c5e0b9d651db · outbound

This paper cites Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities.

In-Context Learning as Implicit Policy Gradient Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.233779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.233779Z digest=sha256:04523714b1ed4d8b97702e5bb686704da19f7833478f3faf706a0cca5d91a545

Observation 55291bae-0ea5-428f-b6c1-8baa1a951b10 · outbound

This paper cites GPT-4o System Card.

In-Context Learning as Implicit Policy Gradient GPT-4o System Card

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.238826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.238826Z digest=sha256:18493764fef6b31534e59262854492ec91c654a041265231e43e8870c9c1ebc4

Observation 09aa6a92-592f-432b-b5de-1d94360bc237 · outbound

This paper cites Transformer Feed-Forward Layers Build Predictions by Promoting Concepts in the Vocabulary Space.

In-Context Learning as Implicit Policy Gradient Transformer Feed-Forward Layers Build Predictions by Promoting Concepts in the Vocabulary Space

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.244739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.244739Z digest=sha256:db413d4e0d8d4a9521bbca80e77ceee9289593411662f21c5627862debf9bc93

Observation c4e311bc-2f0a-46ba-96b9-18213219def8 · outbound

This paper cites Dissecting Recall of Factual Associations in Auto-Regressive Language Models.

In-Context Learning as Implicit Policy Gradient Dissecting Recall of Factual Associations in Auto-Regressive Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.249935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.249935Z digest=sha256:59e6a8cacde5a380d4ca2086b48aadf1cf2ab1473478d73da1da62488048c918

Observation 6c8d8249-20cc-4ad0-bcf0-744a03cc3632 · outbound

This paper cites Machine learning , volume=.

In-Context Learning as Implicit Policy Gradient Machine learning , volume=

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.255551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.255551Z digest=sha256:4feb1d0faccd0a7fa6ac456fb36ad62935b3601dbab119c583e7f4061120f3cb

Observation 7104a18a-cfc7-4ef1-8703-ccbef8eadee7 · outbound

This paper cites arXiv e-prints , pages=.

In-Context Learning as Implicit Policy Gradient arXiv e-prints , pages=

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.260605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.260605Z digest=sha256:6d8308df7c9e602f231f2bfe5352cb50197f14f8d743d88adfdca4a63a2d30b8

Observation e1a8e865-1f52-4a46-9308-7023386108c2 · outbound

This paper cites Olmo 3.

In-Context Learning as Implicit Policy Gradient Olmo 3

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.264936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.264936Z digest=sha256:7507f006eb2005b07d0bda117685ed1d3674e26917e4143a6510a70c59f66877

Observation 49728a35-695d-496f-9fc5-1466779576cc · outbound

This paper cites 2025 , eprint=.

In-Context Learning as Implicit Policy Gradient 2025 , eprint=

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.269825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.269825Z digest=sha256:d68d84812f7f52602e045453dc04a3b572b208127c3cb41169dca86f7b83fdbd

Observation 8158f2cf-9993-4511-9051-036f16d233c8 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

In-Context Learning as Implicit Policy Gradient Advances in Neural Information Processing Systems , volume=

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.274382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.274382Z digest=sha256:8d79893e3dfbc3b231e59fdadab3c06d538d5e4eeeec5783bcafb2ed50695bb1

Observation 9d99406b-c498-4949-beca-47d058c0aac5 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

In-Context Learning as Implicit Policy Gradient Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.279757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.279757Z digest=sha256:0ec837ba5590930007d1aac36ef904088bfb0cc39b24560d19f47dc6db387a96

Observation 79b704e0-73fa-4128-9e00-11b8d075ee4d · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

In-Context Learning as Implicit Policy Gradient Training Verifiers to Solve Math Word Problems

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.285115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.285115Z digest=sha256:3fb8650ddf4a102c8e89aed1482029ddcd8320d49c90861e6978a4253055cfbc

Observation 48fc6866-17ec-46fb-9948-4daa26ea4602 · outbound

This paper cites DROP : A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs.

In-Context Learning as Implicit Policy Gradient DROP : A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.291842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.291842Z digest=sha256:2e3322defb7a8f546103e4d0bc1e12383c374ab4770c7da40459cee648ec6c5f

Observation 0cea5548-3399-4279-b710-89631107ce0e · outbound

This paper cites an unresolved cited work.

In-Context Learning as Implicit Policy Gradient Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.297221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.297221Z digest=sha256:419443031b17c097304b031bd0f405d199856acd53c7afe6452e78973a98f507

Observation f16e94fc-e17e-4562-8b6c-c1ca32efcb07 · outbound

This paper cites Think Outside the Policy: In-Context Steered Policy Optimization.

In-Context Learning as Implicit Policy Gradient Think Outside the Policy: In-Context Steered Policy Optimization

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.302286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.302286Z digest=sha256:6fe56e4ecdbd5c43d2877c362ac5da0eeec3f404406afa5ccfe661d5aa9f6168

Observation a180430d-b52b-437c-ad62-b77fef06a1bf · outbound

This paper cites Sampling-based Pseudo-Likelihood for Membership Inference Attacks.

In-Context Learning as Implicit Policy Gradient Sampling-based Pseudo-Likelihood for Membership Inference Attacks

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.307229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.307229Z digest=sha256:68969fe9e0a70d7e33080a2373a6e39353a09891fc94c04f2cba83d8f0385dfa

Observation 78fceb19-54cb-4b28-a7db-a7b7cf3ffb6c · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

In-Context Learning as Implicit Policy Gradient Advances in Neural Information Processing Systems , volume=

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.311803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.311803Z digest=sha256:59a5358de10dec10b06ed5689db7bd806e7182581fdff9201c6f408d032492c7

Observation 6c072540-6ae8-43ba-b8ce-eed402c33c25 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

In-Context Learning as Implicit Policy Gradient Advances in Neural Information Processing Systems , volume=

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.317246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.317246Z digest=sha256:152ceb555797324a4a7790a82f2a56045723f7767bd70b9582ee28717a803adf

Observation 31c19998-ce13-4563-ada0-1b66fce1514d · outbound

This paper cites arXiv preprint arXiv:2410.05362 , year=.

In-Context Learning as Implicit Policy Gradient arXiv preprint arXiv:2410.05362 , year=

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.321701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.321701Z digest=sha256:111306519c5f30fa31a4126cc68980281384f772b7d87b3495f580b4581ed19c

Observation af40d2c9-bf3e-4225-bb26-51f576d838c9 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

In-Context Learning as Implicit Policy Gradient Advances in Neural Information Processing Systems , volume=

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.326578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.326578Z digest=sha256:d72d3ad769922098bbeaa901ef72c7c8e5fd81a6bafe2c76a69f52864ade439a

Observation 9086b9e9-18de-4924-aacf-47a16b72ceb1 · outbound

This paper cites The Fourteenth International Conference on Learning Representations , year=.

In-Context Learning as Implicit Policy Gradient The Fourteenth International Conference on Learning Representations , year=

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.331401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.331401Z digest=sha256:a12f7e3629b9e97322f3405112fa39fb0a4163a8ac1e40a5ec03f5497a5e4d92

Observation 4b9fba77-ec17-48d7-94aa-a41df64141f9 · outbound

This paper cites Online Learning Defense against Iterative Jailbreak Attacks via Prompt Optimization.

In-Context Learning as Implicit Policy Gradient Online Learning Defense against Iterative Jailbreak Attacks via Prompt Optimization

Reference 38

Resolution
verified exact
doi, observed 2026-08-01T03:34:05.146592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-01T03:32:24.336474Z digest=sha256:6c18ab0ccdd2947d964d345876d24824cdc3e4ea664a55cccf9e78d67c0efe50

Observation 21ba238a-dcb3-4405-81a6-75a643ccf4df · outbound

This paper cites arXiv preprint arXiv:2601.06884 , year=.

In-Context Learning as Implicit Policy Gradient arXiv preprint arXiv:2601.06884 , year=

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.341314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.341314Z digest=sha256:0c33a835f06f812b4b6938d15ba37c4e6fb5022bf74d6688029b25b1c3f5258c

Observation f5737e52-8d78-4006-a1d1-f7470c72ab4a · outbound

This paper cites RLP rompt: Optimizing Discrete Text Prompts with Reinforcement Learning.

In-Context Learning as Implicit Policy Gradient RLP rompt: Optimizing Discrete Text Prompts with Reinforcement Learning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.347218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.347218Z digest=sha256:eaf0ba21e110bbdd5bf939f115289aa9b8fb562897eedd12401838ebde899165

Observation 67f41f14-4a98-4b62-b9c0-c45daf1baf41 · outbound

This paper cites TEMPERA: Test-Time Prompting via Reinforcement Learning.

In-Context Learning as Implicit Policy Gradient TEMPERA: Test-Time Prompting via Reinforcement Learning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.352368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.352368Z digest=sha256:d3a9e0d1d3730d04a44ed09d8ce2cd0738896af8ec37271e0f06395c371a3f8d

Observation 69ce3b4a-81ba-494b-9d7b-32f2a120318c · outbound

This paper cites The eleventh international conference on learning representations , year=.

In-Context Learning as Implicit Policy Gradient The eleventh international conference on learning representations , year=

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.357810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.357810Z digest=sha256:a9646f48bc8de546b99ea6bb26aa413bf6b0337860c4e3c5691a18f1cf2a2240

Observation f5de7818-0b14-4f61-9385-f104a69f5966 · outbound

This paper cites 2025 IEEE Conference on Secure and Trustworthy Machine Learning (SaTML) , pages=.

In-Context Learning as Implicit Policy Gradient 2025 IEEE Conference on Secure and Trustworthy Machine Learning (SaTML) , pages=

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.362786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.362786Z digest=sha256:a6d30f6f900e6a9a21acd99486ad187ed4d18abb517cc1b380e17ec95f196ef3

Observation 5782760c-9aba-4dde-8f7c-036797e72914 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

In-Context Learning as Implicit Policy Gradient Advances in Neural Information Processing Systems , volume=

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.367909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.367909Z digest=sha256:ebe0f6982a006e1ea7275464ccda37f0add16215e49631af6ac283f35a2816ca

Observation ce8b9c8c-ff4b-40b7-a6fb-5ce1fb71ab7e · outbound

This paper cites Findings of the Association for Computational Linguistics: ACL 2026 , pages=.

In-Context Learning as Implicit Policy Gradient Findings of the Association for Computational Linguistics: ACL 2026 , pages=

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.373654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.373654Z digest=sha256:421e0247bac1034f85a0cb5a4c491caca0f7fe184d0a92a9dba39dea62d6e570

Observation d281de99-ac6e-4a72-99a1-8c1c266a321c · outbound

This paper cites Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=.

In-Context Learning as Implicit Policy Gradient Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.378910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.378910Z digest=sha256:3f9b461fede09f3cc9f2d2222403795ba9f73cc9dc8d7d8942edfacb9694881c

Observation 90f3d2ce-4667-4f59-8109-5be21015de5d · outbound

This paper cites International Conference on Learning Representations , volume=.

In-Context Learning as Implicit Policy Gradient International Conference on Learning Representations , volume=

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-01T03:32:24.383801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:32:24.383801Z digest=sha256:34d9fd3aa5aae4e0a6400ae8b7461b099148d9274c21f532946872a8233cdb8f

Pith citing papers

No inbound Pith citation observations are available.