Pith. sign in

Paper Citation Record · LEDGER

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization

As of 19 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 0 inbound Pith citation observations for arXiv:2505.12759.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.12759 v1

Coverage vector

measured 46 of 46 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:30:48.008630Z

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

46 of 46 outbound references displayed

  • verified exact0
  • verified fuzzy7
  • unresolved39
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d1414144-2152-4168-bfcb-ebaa2d8044b4 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.616750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.818185Z digest=sha256:2d9623e61bf4e07eabbbda159c92355bfeb3a0579ac9131625a40400adca3fe7

Observation 7d9b98de-1ef1-4056-a0f5-ae3c984574ae · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.604617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.822600Z digest=sha256:53f04670c9759703d83b18d9a191cf99ba8e2cda23759af85600bcd936c878f8

Observation d3cf46b7-6f1a-492b-9f22-d260faf2b4ad · outbound

This paper cites Deep Reinforcement Learning for Active High Frequency Trading.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Deep Reinforcement Learning for Active High Frequency Trading

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T20:30:47.826795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:30:47.826795Z digest=sha256:270d26119bb993966f92993ec8b236b0ad3a5ebab9185c2307b6a00b15e544a7

Observation fd616618-f30d-4ed9-a0fb-39f173de0dcb · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.592851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.831378Z digest=sha256:1eab2d67418b42dc80e70e43d47be6ba47e89e37e8c1e76087bf8aa0967452d5

Observation 42f99dbc-d39f-47d0-8588-a0534af751da · outbound

This paper cites J., Torn \'e , R.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization J., Torn \'e , R

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:30:48.581900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.835243Z digest=sha256:40c8b169e076f14d5c4f233eb20761cdb88a7f5cd6368c8467f23ccbe152123c

Observation fab88295-dc35-4e80-aedb-fc423c3290b5 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.569448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.839203Z digest=sha256:ced33431132b20a04699727373e0c88e30b101c845938a0ca9880bb0649355e2

Observation 5be321d1-acc6-465f-aea9-fd5f76a275c3 · outbound

This paper cites L., Sutskever, I., and Abbeel, P.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization L., Sutskever, I., and Abbeel, P

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:30:48.556261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.843230Z digest=sha256:c9773eacf17c27e05970a73a9861c582d05a063806f90b74804ea1b797f03d59

Observation aac2ae63-c55f-4fb3-8671-98724dd733e3 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.540930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.846813Z digest=sha256:978c8879711671b9c5d00eb6cb32f1c9704ddb9eedfb2bd4a5df1c24d15763e0

Observation 5145849a-4b26-4f1c-8b4c-9dff31afd5ab · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T20:30:47.850552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:30:47.850552Z digest=sha256:14bddbece72ce75f7bb2ab8896e71f7a012b06193995be98693f9cf78529a8c0

Observation e0d40e2c-e6e7-4096-9058-4a6db3f24271 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.519551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.854196Z digest=sha256:845b66d402e2e94b9db508304388436c6d01579ce0b1012e38a7494ba5070423

Observation 6aa10ef7-a70f-4096-9b10-e33af0cfe990 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.506735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.858379Z digest=sha256:6e64fe0b395a7e36f9201d4ec2e6caaf262e047d6a2ca1f84bf565165e5c86b6

Observation e10debd4-7eed-4d39-8741-a3d872fd9a57 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.495503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.862287Z digest=sha256:f7e54ed1bd22c7d2f4cf3670633c7c81e2960aa9bb8183242c3dd000e2213d16

Observation 878cae8a-7ff1-4d0b-9b65-cd45d492a754 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.484498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.866408Z digest=sha256:2145abdf997f67dd76e4ffbbe7228be4c057f38a782fab1ed5f7f40f878a3175

Observation 7d56223c-fff3-47bb-8b3d-af3f6ffb1a73 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.471652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.870555Z digest=sha256:098c46f33fc25a0a9073f3794b328d038f48aba62413d62005fa424a2d3266a9

Observation 6ff1d268-1543-48d7-a5ec-5c6b5457a531 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.459738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.874719Z digest=sha256:51e202d2500d41337426d83599a499351a4f264abba86f0c14b203fe2c819648

Observation 65181769-8a1f-4f5a-ab49-523deba1fcde · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T20:30:47.878459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:30:47.878459Z digest=sha256:5dc80a5be1bc134eaa025ef3d06b11a3663675e6878dc21d66d4c60aefb81f04

Observation be7c0f1f-e0f6-48ff-b34f-7dd14597da02 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.439186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.882071Z digest=sha256:a04c6bb39742ed5aef72b861b50da873651e5153d7f91502ff62c1e996e294f8

Observation a9997ac8-a5fd-4ecc-bbf1-8631297aa3db · outbound

This paper cites Meta reinforcement learning as task inference.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Meta reinforcement learning as task inference

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T20:30:47.885697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:30:47.885697Z digest=sha256:7ff5cf3a77c9b3b6b31703a1f8077d3e76ebaa8ddcefa1c884a5a54e6809e9ad

Observation b47400a4-fd62-43b3-abc1-cd421321a837 · outbound

This paper cites and Kim, H.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization and Kim, H

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:30:48.427340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.889741Z digest=sha256:fa241ea82adf9788485fcf48e4460e2e3f4a5c4386833433e9a3436e15e6f53e

Observation a4563d8d-9614-439d-8c74-4c2848d092a7 · outbound

This paper cites Offline Reinforcement Learning with Implicit Q-Learning.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Offline Reinforcement Learning with Implicit Q-Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T20:30:47.894793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:30:47.894793Z digest=sha256:431c81ca868528ad78998e5d0bd564f72699668ea15117b3f2db0f77e5964b72

Observation e15bc488-2b09-4f4e-9db5-8e32023409cc · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.414071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.900181Z digest=sha256:85523792ef30e5867f2bb76bdb353082878a6d2b3ed70514ab214df238ee3ab2

Observation a3142cc2-b03f-4e45-87a4-a9b840f6fb9d · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.400983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.905327Z digest=sha256:ba74847f3fc33f8818b018bdeecd77431ae630382c9e7dd4ede24072a2782dd9

Observation c2099106-82e3-40f0-a573-80999b4ebf6c · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.388452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.909311Z digest=sha256:a731ee6a2d9eb77215456ce360f75e2998147ee37d607c1055a53d6c70457bcb

Observation 4a26a00c-f134-487f-a1de-4d13ae290357 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.375710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.915066Z digest=sha256:9c13564a82ddafcf9dbd113143dc48ce6cf204a7fde32477581260d00275d0af

Observation 410d7879-1d78-4a9f-9c0c-5751829b42ae · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.363691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.920484Z digest=sha256:bceb4a81e91cf1b20b27a4d2f447d2d212de7e7a93a5e97a93ade7f1bc3b4138

Observation 0e25b83d-9551-430c-8fca-02d6ea980c6e · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.352163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.924504Z digest=sha256:28e96bdec87f2cf11c72d8476895e2c39ee080bffc8608bd20dacf9a1033775d

Observation 5e7bfdef-6d50-43b3-96a3-e26f215141cc · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.339641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.929044Z digest=sha256:580b33ba565917d6df73e68bd2c1ca61908a92aa3807fc7e498800b88348b4d6

Observation 33fe92a2-5378-453b-834d-8e911d438a7f · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.327406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.935173Z digest=sha256:5180dccf2bf040c61614b11ce43ca5f8e254b21f0348b0607be742b5ba18fc5e

Observation c94f7c40-e1fc-44a3-978a-84a1f153c69f · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.314829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.939722Z digest=sha256:1ed368d993a198e4235f4a5796d92e5cd6e14366a2843d5bc159b940342bcf4f

Observation 06b5d353-7602-4682-858b-45c20b9064f2 · outbound

This paper cites B., Levine, S., and Finn, C.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization B., Levine, S., and Finn, C

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:30:48.300344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.943840Z digest=sha256:2c327ad874203fc80aa0ae31768f760af978ffeeaad1048f3348383c2dd8376e

Observation e151072d-89e1-46d1-8d12-d6a037e62884 · outbound

This paper cites S., Abbeel, P., Levine, S., and Finn, C.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization S., Abbeel, P., Levine, S., and Finn, C

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:30:48.286863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.948140Z digest=sha256:2e1be164c3f6c86848a719c68b4d4a773160feac39c8af7f88ea5f28e04e658e

Observation b333851b-008d-4123-b157-e4b427d013d7 · outbound

This paper cites H., Nair, A.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization H., Nair, A

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:30:48.273713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.952430Z digest=sha256:1c5361f68ca0e4d87516f410030256a004d71625faf6db704af404c8ed0f2b12

Observation 81d91329-8e77-41b5-925d-2421000d1312 · outbound

This paper cites Meta Reinforcement Learning with Latent Variable Gaussian Processes.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Meta Reinforcement Learning with Latent Variable Gaussian Processes

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T20:30:47.956265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:30:47.956265Z digest=sha256:3fbed43b5d8bef550160bcea2217d7fee1aba00ac2126ad10a89e08228a66e0e

Observation 581f20a3-f69e-432e-b623-b4f20c5564d8 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.262059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.961192Z digest=sha256:ee55a64309a8cb994b46d4f2cffc8373ef659fa8fc384e1fee336b873c0d5dc3

Observation 5873c74f-4994-4b0b-b6f4-49b639f65287 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.248008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.964925Z digest=sha256:3fd4b867ce5d6832167520850986690a114531a5a86ec53da4bf588cd1d80905

Observation 141df753-1438-4a8e-b3d8-3ed450f1f522 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.234201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.969043Z digest=sha256:6051baeebb64fa20c7d13f28a34c3ae99faf4c690ddd8cca0e07785add9b98ee

Observation e61db1d1-ef4a-486b-bd03-3bfc10794084 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.222829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.972451Z digest=sha256:b8163477ae80e3d6fea537ae3410f7707a84f7e67d4e6b6e3290d2ea9881a04c

Observation f6e537c7-d651-4893-8195-ea83363031af · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.209334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.976255Z digest=sha256:b06422bb44902d72f71d3b0939c7f2e18a4dfe3da2d1be4dd54dc136f9632174

Observation 5a22a7e5-f885-42fa-ab71-da9813076c08 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.194833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.979787Z digest=sha256:2be11b35fb34ecfd73cb23b0b53b190eb67d817ca51b801250ea3c8c485cb066

Observation 47c2fa11-644f-41e8-9433-8be9a5bad0d1 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.181143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.983368Z digest=sha256:94a0635aa15752559234b052bdc05b1b90853228760777917a7c8a1be270868d

Observation 5958fe27-a25b-434c-ab87-b0fb79a71ad5 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.167429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.986940Z digest=sha256:368b4aa4f9a0361b31491a4a908ab8870a4a33f26324062ae157326aeaad5c86

Observation e904b8ce-173c-4bb5-adbc-fb73f761ae3c · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.154382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.990557Z digest=sha256:c3904f45d65feb5aff9f4442c90c8e25bc1ce557d499bae979b45ad7799d47f8

Observation 08d3fc73-6e22-4cd3-a958-cec115089eab · outbound

This paper cites and Cohen, S.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization and Cohen, S

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:30:48.141416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.994457Z digest=sha256:7d1586b8ccb3182dc6c1e087b67e9c360417621738394a0b25dc98dc3d3dd361

Observation 4b541ba5-c967-4ada-8d22-86753c70c356 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.129024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:47.998919Z digest=sha256:83ad9d36099b6010a0eca15258aaabe897b084f2fccd4c454e52acf630c82887

Observation 59c3c86a-1001-42af-b119-67eaf67a3414 · outbound

This paper cites ENOTO: Improving Offline-to-Online Reinforcement Learning with Q-Ensembles.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization ENOTO: Improving Offline-to-Online Reinforcement Learning with Q-Ensembles

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T20:30:48.003571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:30:48.003571Z digest=sha256:781b44ce16d32111ea6d778e95088b03b896a2afa7f5c132239330c1ae8c7a3b

Observation 4203a027-64bd-4053-804e-6d42fd215126 · outbound

This paper cites an unresolved cited work.

Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:30:48.115919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:30:48.008630Z digest=sha256:6da09e5e5544f9d6d5b58f8fd1b180d9ea9d763be3ec09964bafdd947577f34b

Pith citing papers

No inbound Pith citation observations are available.