Pith. sign in

Paper Citation Record · LEDGER

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability

As of 21 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2608.03933.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.03933 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T05:28:30.133653Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact8
  • verified fuzzy3
  • unresolved15
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e30aae58-dcf6-4a5f-84b5-89737d1cea5b · outbound

This paper cites Deep Reinforcement Learning and Mean-Variance Strategies for Responsible Portfolio Optimization.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability Deep Reinforcement Learning and Mean-Variance Strategies for Responsible Portfolio Optimization

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-05T05:28:33.033333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T05:28:26.945078Z digest=sha256:8a5d401b6c5020789c1faccd4e9e17e90084dd524f6de000cd151c57b2b3c130

Observation 78506cbb-40fe-44cc-bd8b-e9c7c81c9906 · outbound

This paper cites 1957.Dynamic Programming.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability 1957.Dynamic Programming

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:28:34.074772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T05:28:27.082210Z digest=sha256:9a695560badba0db2dad5b26e2d61e3210984b29daff44f30c00d55c028dc9d5

Observation b39ca46d-224e-4cb4-bdc2-db845c0089a2 · outbound

This paper cites an unresolved cited work.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-05T05:28:33.840049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T05:28:27.162128Z digest=sha256:3c80be717ee26df21bd759214ad04f7ed00cd7301827b37804537f9afff199fd

Observation 5d1595a0-c618-47a6-a752-f30147354af7 · outbound

This paper cites Carroll and Miles S.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability Carroll and Miles S

Reference 4

Resolution
verified exact
doi, observed 2026-08-05T05:28:31.372265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T05:28:27.238955Z digest=sha256:315cd4fb8371db69e10ed5609c6201907a603b41f6cb461172c637d875dbace2

Observation 12da4b16-efd5-4b73-8e20-f9b528e2265a · outbound

This paper cites Cocco, Francisco J.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability Cocco, Francisco J

Reference 5

Resolution
verified exact
doi, observed 2026-08-05T05:28:31.065479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T05:28:27.346383Z digest=sha256:d987eddb7dbc1fa8cfdd6989545b2181ed29fb4384f5d86d68ba5cbb3b899310

Observation 33a6bbd0-664d-4754-8105-cd59c98a19ac · outbound

This paper cites Goodman, and Jonathan A.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability Goodman, and Jonathan A

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T05:28:27.448464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:28:27.448464Z digest=sha256:85ab076a4836d8a9d0fe04ace719796332bc27e72447db246b7472c770b14277

Observation 69e01304-0528-456f-be1f-264426d3d2bb · outbound

This paper cites an unresolved cited work.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability Unresolved cited work

Reference 7

Resolution
malformed identifier
no resolver link, observed 2026-08-05T05:28:27.563407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:28:27.563407Z digest=sha256:e806101ea83c370931595e1462677fa129a2468b22807840a471883e6ad379ad

Observation 14096051-be24-489a-a3f7-7ce755dcb1b3 · outbound

This paper cites 2016.Deep Learning.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability 2016.Deep Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T05:28:27.667435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:28:27.667435Z digest=sha256:2cec71ffa631cbf8faf46b119f70236540afdd3b7e7989c94894b2fe4512bc8b

Observation c5f41f95-c8b5-4260-8446-1d20a0a1d841 · outbound

This paper cites Deep Learning Approximation for Stochastic Control Problems.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability Deep Learning Approximation for Stochastic Control Problems

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T05:28:27.758490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:28:27.758490Z digest=sha256:e6aa4c97e3c103a6daad3973e16104585f5183cd312e429f104cf0a59d908063

Observation 3b8da8b9-ae0d-44ec-9f9a-48486a4ae3ed · outbound

This paper cites an unresolved cited work.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T05:28:27.895092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:28:27.895092Z digest=sha256:d19eed0a8a586449772ecb5bcaa4f0ab52209c842c39b8009543c07913fa8781

Observation 55fed3da-31f0-4d0e-ad88-c88fdc38e268 · outbound

This paper cites an unresolved cited work.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability Unresolved cited work

Reference 11

Resolution
malformed identifier
raw_fallback, observed 2026-08-05T05:28:33.659016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T05:28:27.980573Z digest=sha256:7a964f1c39e0a6762ad7832e71ca2bbd86a7250a41fca3340301e73f3e00ad19

Observation 75e156f9-a4fb-425f-9b15-20871083ef87 · outbound

This paper cites A Deep Reinforcement Learning Framework for the Financial Portfolio Management Problem.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability A Deep Reinforcement Learning Framework for the Financial Portfolio Management Problem

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T05:28:28.114830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:28:28.114830Z digest=sha256:bb49f8b0dfc088b60b6e09447cf4fcd5aeb075a470d1b672b826e3fb9980a3a9

Observation ff087b80-086b-4456-b210-68023587e5d4 · outbound

This paper cites an unresolved cited work.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T05:28:28.201141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:28:28.201141Z digest=sha256:973c1d4f9e11627c69f62ea691eb461e73a63af0b1aa940a1126a312bfece99b

Observation 65f39e83-77de-43b5-80a8-e531c627bf1a · outbound

This paper cites Optimistic Bull or Pessimistic Bear: Adaptive Deep Reinforcement Learning for Stock Portfolio Allocation.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability Optimistic Bull or Pessimistic Bear: Adaptive Deep Reinforcement Learning for Stock Portfolio Allocation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T05:28:28.301772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:28:28.301772Z digest=sha256:6884a54646dc67593391cfd75988629e172bc2cb40acbc62e3adc41e608ac79b

Observation aea4af67-f011-4877-a25e-60836eaae1a3 · outbound

This paper cites Continuous control with deep reinforcement learning.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability Continuous control with deep reinforcement learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T05:28:28.409082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:28:28.409082Z digest=sha256:330842f351a9315bbe2ba9dcd2d2f2ef8b07ae8986a6e3902d134b78f2df7f54

Observation 2f0ee181-b328-4b30-ab27-ea6da17241d0 · outbound

This paper cites an unresolved cited work.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability Unresolved cited work

Reference 16

Resolution
verified exact
doi, observed 2026-08-05T05:28:30.737679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T05:28:28.547867Z digest=sha256:0809767f630907b8d475dbe7e338647473b866d18f015779ac9bccf77e0d3371

Observation e54f719c-f058-46aa-ac68-7e1f353aa32b · outbound

This paper cites an unresolved cited work.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability Unresolved cited work

Reference 17

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-05T05:28:32.715072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T05:28:28.625830Z digest=sha256:df135fe70e3dcb8a3f680efc23e796fb3705c0eb9fe98a880210223fde4e34ac

Observation 4da032e5-401c-4a10-b82c-087e7e09ee86 · outbound

This paper cites an unresolved cited work.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T05:28:28.731378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:28:28.731378Z digest=sha256:67969c8c717d37eb44f4ca761bc8ad0db7e3a7328955e0542f7f4d0503f08d3e

Observation d4943357-6cbd-4306-b9bd-e43a02b1ef46 · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability Playing Atari with Deep Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T05:28:28.820951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:28:28.820951Z digest=sha256:cad0754bc319c528f5e7205b36522f9de82b5e50e0a865834dac954f3a000b86

Observation 59d15008-6443-4f19-a4bf-7a5eb674ca4d · outbound

This paper cites an unresolved cited work.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T05:28:28.902877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:28:28.902877Z digest=sha256:b9cf8ce5718ee73fc63735172f205ce2d59a7d4a11959d418d771f8a42a7b7fe

Observation 43b90f71-5c30-4868-b4e2-5b6f0ccdb08d · outbound

This paper cites A Machine Learning Algorithm for Finite-Horizon Stochastic Control Problems in Economics.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability A Machine Learning Algorithm for Finite-Horizon Stochastic Control Problems in Economics

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-05T05:28:32.071279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T05:28:29.016853Z digest=sha256:63dbe9d4905297b7843c671d86ef966efd86327d0dd6a0290eab4bbcce337605

Observation 96c7ef28-7a70-4b93-8bd9-22ac91ede638 · outbound

This paper cites 2023.Foundations of Reinforcement Learning with Applications in Finance.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability 2023.Foundations of Reinforcement Learning with Applications in Finance

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:28:33.522792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T05:28:29.110658Z digest=sha256:13b301214b5a469c036d869afe1f533af0fd02ffebfab8247c1a8c983869eaf5

Observation 7ee1ccd3-8730-49f6-b229-96f9beef05de · outbound

This paper cites Learning from zero: how to make consumption-saving decisions in a stochastic environment with an AI algorithm.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability Learning from zero: how to make consumption-saving decisions in a stochastic environment with an AI algorithm

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-05T05:28:31.712352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T05:28:29.245022Z digest=sha256:07987e67bf6ca5a4195752e8e086779e5ae13357884d9c36bde56b7c067cbe65

Observation 8418ba68-1aa9-4110-9ae5-cc20050457c1 · outbound

This paper cites an unresolved cited work.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T05:28:29.403841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:28:29.403841Z digest=sha256:de9a002e63ccebf61b05497b5e8af17c263e08456bb4462b33498a82e678d57d

Observation 911ed933-4f4d-466a-86a4-ba44bab84cfc · outbound

This paper cites an unresolved cited work.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-05T05:28:33.364260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T05:28:29.618705Z digest=sha256:ba93a19f22cf0a46e6331d29e43c74ccff6f3e13623a5e5a6968d797eb26a81b

Observation f70cd350-547d-45ad-9976-aac9e8efadb0 · outbound

This paper cites Sutton and Andrew G.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability Sutton and Andrew G

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:28:33.227643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T05:28:29.773462Z digest=sha256:4a3bdc24100652dcaa90cc6cf1013dc27c0426ad50e2a7ca8ccdc9ec68938f10

Observation b60168a6-42d6-4121-ab8a-9b8c4066731a · outbound

This paper cites an unresolved cited work.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability Unresolved cited work

Reference 27

Resolution
verified exact
doi, observed 2026-08-05T05:28:30.380758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T05:28:29.981245Z digest=sha256:e1de466d417102f5b924afeab19775d032ce620aa1fb3c5ddf0ac47e83208ce9

Observation 11180c39-69dd-4634-93fa-4db260ce1870 · outbound

This paper cites Williams.

Simulation-Based Neural Policies for Portfolio Choice: Architecture, Training, and Interpretability Williams

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T05:28:30.133653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:28:30.133653Z digest=sha256:876517807e61f44dc2e4e149fccf82a70b6a611c7c4685359854919743b2788c

Pith citing papers

No inbound Pith citation observations are available.