Pith. sign in

Paper Citation Record · LEDGER

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization

As of 19 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 0 inbound Pith citation observations for arXiv:2505.12759.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.12759 v1

Coverage vector

measured 46 of 46 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:30:48.008630Z

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

46 of 46 outbound references displayed

  • verified exact0
  • verified fuzzy7
  • unresolved39
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d1414144-2152-4168-bfcb-ebaa2d8044b4 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.616750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.818185Z digest=sha256:f6e37588cf4361ac3e9c1297bf1b5a5d6a8b72e0c4ba0888d92ca622575c3b29

Observation 7d9b98de-1ef1-4056-a0f5-ae3c984574ae · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.604617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.822600Z digest=sha256:3a9e143586674bcde87c21f0aecd498f2cb32600be87f19fe4735594669ffd10

Observation d3cf46b7-6f1a-492b-9f22-d260faf2b4ad · outbound

This paper cites Deep Reinforcement Learning for Active High Frequency Trading.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Deep Reinforcement Learning for Active High Frequency Trading

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T20:30:47.826795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:30:47.826795Z digest=sha256:270d26119bb993966f92993ec8b236b0ad3a5ebab9185c2307b6a00b15e544a7

Observation fd616618-f30d-4ed9-a0fb-39f173de0dcb · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.592851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.831378Z digest=sha256:2ba08f0757e636f9a237499dc7b62c52099f78d239e4da50db17c6518f8451ff

Observation 42f99dbc-d39f-47d0-8588-a0534af751da · outbound

This paper cites J., Torn \'e , R.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization J., Torn \'e , R

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:30:48.581900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.835243Z digest=sha256:7e1fc26439d9635037dd3a2c49967bafb148dfeab91497dc3d4931122dac1cdb

Observation fab88295-dc35-4e80-aedb-fc423c3290b5 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.569448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.839203Z digest=sha256:d159e68a23d4436a0e0d6283bcf4e9d9cb86256d409776a7cca875704df20ce6

Observation 5be321d1-acc6-465f-aea9-fd5f76a275c3 · outbound

This paper cites L., Sutskever, I., and Abbeel, P.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization L., Sutskever, I., and Abbeel, P

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:30:48.556261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.843230Z digest=sha256:4b7a01cbae41984b9417b80d7ac7a4ac48c6cf15e9d2be1ea730e3abcba9bb07

Observation aac2ae63-c55f-4fb3-8671-98724dd733e3 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.540930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.846813Z digest=sha256:a3282eb2cd6975c31815665e5404714d375b60bd73691d9ee76e65d03c6d0685

Observation 5145849a-4b26-4f1c-8b4c-9dff31afd5ab · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T20:30:47.850552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:30:47.850552Z digest=sha256:14bddbece72ce75f7bb2ab8896e71f7a012b06193995be98693f9cf78529a8c0

Observation e0d40e2c-e6e7-4096-9058-4a6db3f24271 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.519551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.854196Z digest=sha256:148a5425ee7fb5ca93daa854555a815e2ac0fa025dd2d5be9aaea415cfb72c69

Observation 6aa10ef7-a70f-4096-9b10-e33af0cfe990 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.506735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.858379Z digest=sha256:cf144d621b55efb9b552e5667028b46141215e2d77afcdd92517b3ba8b8ae520

Observation e10debd4-7eed-4d39-8741-a3d872fd9a57 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.495503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.862287Z digest=sha256:df8d989bba7d4566adf69514ff99b0b444db2e8082be32c79676a18bc82b2b92

Observation 878cae8a-7ff1-4d0b-9b65-cd45d492a754 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.484498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.866408Z digest=sha256:568f765891a68b08b97c94281293fc23a8c0f9193b85339ab6e6ad6fcb8d2827

Observation 7d56223c-fff3-47bb-8b3d-af3f6ffb1a73 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.471652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.870555Z digest=sha256:a61f10d56a342d30ab84604fb38816952a681d37e4b9714331b00ed364368c8f

Observation 6ff1d268-1543-48d7-a5ec-5c6b5457a531 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.459738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.874719Z digest=sha256:d5b81632b2cae468ba876e90baba016a205210ef36237c3a94a131b0991cafe9

Observation 65181769-8a1f-4f5a-ab49-523deba1fcde · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T20:30:47.878459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:30:47.878459Z digest=sha256:5dc80a5be1bc134eaa025ef3d06b11a3663675e6878dc21d66d4c60aefb81f04

Observation be7c0f1f-e0f6-48ff-b34f-7dd14597da02 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.439186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.882071Z digest=sha256:2172ef370f0fb083159396da7334b21af67916d614c350901133861ac3dda69c

Observation a9997ac8-a5fd-4ecc-bbf1-8631297aa3db · outbound

This paper cites Meta reinforcement learning as task inference.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Meta reinforcement learning as task inference

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T20:30:47.885697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:30:47.885697Z digest=sha256:7ff5cf3a77c9b3b6b31703a1f8077d3e76ebaa8ddcefa1c884a5a54e6809e9ad

Observation b47400a4-fd62-43b3-abc1-cd421321a837 · outbound

This paper cites and Kim, H.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization and Kim, H

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:30:48.427340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.889741Z digest=sha256:f4fd1e9f3cd7118abed6f10bc8cc55bbbe5359ff5bac0e6b9e86faa7a951df9d

Observation a4563d8d-9614-439d-8c74-4c2848d092a7 · outbound

This paper cites Offline Reinforcement Learning with Implicit Q-Learning.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Offline Reinforcement Learning with Implicit Q-Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T20:30:47.894793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:30:47.894793Z digest=sha256:431c81ca868528ad78998e5d0bd564f72699668ea15117b3f2db0f77e5964b72

Observation e15bc488-2b09-4f4e-9db5-8e32023409cc · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.414071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.900181Z digest=sha256:57f794743bb8116b4d65ff5ba35a11337848154133008e31793a91efff769872

Observation a3142cc2-b03f-4e45-87a4-a9b840f6fb9d · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.400983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.905327Z digest=sha256:fb4621fc09e60329bfbf5f472d69b1e46198778275b8c70b5bd0fd5a6ca33a5a

Observation c2099106-82e3-40f0-a573-80999b4ebf6c · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.388452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.909311Z digest=sha256:1adf3d4429959a2b2d703939154bf7f9bd83f4339d539773f557f3dfb60fe3ce

Observation 4a26a00c-f134-487f-a1de-4d13ae290357 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.375710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.915066Z digest=sha256:d0ae1c464a9d6f6c240c6a8f03d0959f1f1da90f02ca43e0201a70950ac1912a

Observation 410d7879-1d78-4a9f-9c0c-5751829b42ae · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.363691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.920484Z digest=sha256:93d26430d0ea530e3ea000b6371d1f51e376b69396faa3879c1557b2fe73c73e

Observation 0e25b83d-9551-430c-8fca-02d6ea980c6e · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.352163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.924504Z digest=sha256:cf12bfdf6a08d443fb57e725c669e8bb6486a26bd5d8f4afd66ae44922ec6910

Observation 5e7bfdef-6d50-43b3-96a3-e26f215141cc · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.339641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.929044Z digest=sha256:6cd02172c483b58e9f641ae5e3258c388be780fdcc558fa2c20236231084e9c5

Observation 33fe92a2-5378-453b-834d-8e911d438a7f · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.327406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.935173Z digest=sha256:aad7ef6240c89311656512c2737fdb19b009a394d500906d688aaadd19ec7928

Observation c94f7c40-e1fc-44a3-978a-84a1f153c69f · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.314829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.939722Z digest=sha256:04962aeae50abc936100fa0384bcd355a6ced5cad4a84446c6e0069d147469d0

Observation 06b5d353-7602-4682-858b-45c20b9064f2 · outbound

This paper cites B., Levine, S., and Finn, C.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization B., Levine, S., and Finn, C

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:30:48.300344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.943840Z digest=sha256:119055a9de05cbdb4ccdd087fd8a739a95a23204aa9389cabce2baae11076e49

Observation e151072d-89e1-46d1-8d12-d6a037e62884 · outbound

This paper cites S., Abbeel, P., Levine, S., and Finn, C.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization S., Abbeel, P., Levine, S., and Finn, C

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:30:48.286863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.948140Z digest=sha256:6d1f6a145891d986f4cd1a97990f3271da74f694c9f933c401427b688f8b1eaf

Observation b333851b-008d-4123-b157-e4b427d013d7 · outbound

This paper cites H., Nair, A.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization H., Nair, A

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:30:48.273713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.952430Z digest=sha256:625d10f8cb0a83b830fa8b0123614b9368f16e03d5b7424fbc79a6c7fb3b6b03

Observation 81d91329-8e77-41b5-925d-2421000d1312 · outbound

This paper cites Meta Reinforcement Learning with Latent Variable Gaussian Processes.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Meta Reinforcement Learning with Latent Variable Gaussian Processes

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T20:30:47.956265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:30:47.956265Z digest=sha256:3fbed43b5d8bef550160bcea2217d7fee1aba00ac2126ad10a89e08228a66e0e

Observation 581f20a3-f69e-432e-b623-b4f20c5564d8 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.262059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.961192Z digest=sha256:cbcd9e1fd4b2b435d5e32ebee3e4f44630226021b181354062848622a6254eee

Observation 5873c74f-4994-4b0b-b6f4-49b639f65287 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.248008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.964925Z digest=sha256:8bd9da2a966dfa717bd4b8ae9bca6b91bc3e0de69deabf3f7c226bce92fe6ac5

Observation 141df753-1438-4a8e-b3d8-3ed450f1f522 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.234201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.969043Z digest=sha256:93b3369cc2cfcf8f4413bd910d1fe0518adbe5e24fb0fad05f9b57a42af7bcfb

Observation e61db1d1-ef4a-486b-bd03-3bfc10794084 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.222829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.972451Z digest=sha256:10604c6ae83add3c87517edf4a2afc029f179795e4a388e9defa1534137e7000

Observation f6e537c7-d651-4893-8195-ea83363031af · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.209334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.976255Z digest=sha256:c705b27b450a9085bdda2aaeef871e3a3c7f0cd4e8c6c806d87d60c2d48433a0

Observation 5a22a7e5-f885-42fa-ab71-da9813076c08 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.194833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.979787Z digest=sha256:9cf00c9a8c656be0bd2585e2bbfb746af58f227b6d827494233b386bf0819e53

Observation 47c2fa11-644f-41e8-9433-8be9a5bad0d1 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.181143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.983368Z digest=sha256:8804d5bbd5c2fe330345d2cbef06cd906966970250f2e456c8b9f89177d66579

Observation 5958fe27-a25b-434c-ab87-b0fb79a71ad5 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.167429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.986940Z digest=sha256:1f0a468acde11adec2bc874e86cad7293a8354ff02f023f1747199a3d4775d75

Observation e904b8ce-173c-4bb5-adbc-fb73f761ae3c · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.154382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.990557Z digest=sha256:bd22f3c54bcb5afb75f05f1f42d02212868957d7d765b202ac9f7339c3c81e7c

Observation 08d3fc73-6e22-4cd3-a958-cec115089eab · outbound

This paper cites and Cohen, S.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization and Cohen, S

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:30:48.141416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.994457Z digest=sha256:74a5c60a1fd30561417cdc0440ecdb7cbb6b5f3fc2b1cde8186091b39498ac87

Observation 4b541ba5-c967-4ada-8d22-86753c70c356 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.129024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.998919Z digest=sha256:e3ee6058f8dc5fbf5d2e9ad74eebd096a9cc8ae105030641032d249bb18fc668

Observation 59c3c86a-1001-42af-b119-67eaf67a3414 · outbound

This paper cites ENOTO: Improving Offline-to-Online Reinforcement Learning with Q-Ensembles.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization ENOTO: Improving Offline-to-Online Reinforcement Learning with Q-Ensembles

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T20:30:48.003571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:30:48.003571Z digest=sha256:781b44ce16d32111ea6d778e95088b03b896a2afa7f5c132239330c1ae8c7a3b

Observation 4203a027-64bd-4053-804e-6d42fd215126 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.115919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T20:30:48.008630Z digest=sha256:ecc372a1999703d8713c399c97eee4ad2843eeb24b02c9f390cb2908130fd108

Pith citing papers

No inbound Pith citation observations are available.