Pith. sign in

Paper Citation Record · LEDGER

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability

As of 7 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2505.23857.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.23857 v2

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:00:50.297031Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

36 of 36 outbound references displayed

  • verified exact0
  • verified fuzzy29
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ee0928d1-5450-4824-8cb1-c4cca182d8af · outbound

This paper cites Outracing champion gran turismo drivers with deep reinforcement learning,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Outracing champion gran turismo drivers with deep reinforcement learning,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:58.276582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:46.693038Z digest=sha256:3fd62599aea59eb2dd6866883fb6b426ef2119af5e32b67109d24a84b0238a90

Observation 787d3297-e4d8-4e5d-9163-85c5022bfe9a · outbound

This paper cites Perceiving the world: Question-guided reinforcement learning for text-based games,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Perceiving the world: Question-guided reinforcement learning for text-based games,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:58.044516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:46.802798Z digest=sha256:95a894c43d6d72e78f0ee5c959b3591ca6dfc6d6e9008f8f38b7c99643f28e02

Observation 3700d2d0-d447-40bb-8dd9-84e69ec476c8 · outbound

This paper cites Deep reinforcement learning in health- care and bio-medical applications,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Deep reinforcement learning in health- care and bio-medical applications,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:57.782473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:46.914626Z digest=sha256:e91f8c8b8ecbf93f8aa4c29ab232d464e1e40dfe5a69b95c22e6493b7befbfae

Observation d1f0340e-21e3-4e01-a9cd-875eeb1b51de · outbound

This paper cites Deep reinforcement learning in radiation therapy planning optimization: A comprehensive review,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Deep reinforcement learning in radiation therapy planning optimization: A comprehensive review,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:57.515201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:47.002939Z digest=sha256:167e34beffd8323d123be08ec98e05272488bd0e955b9253b7b176456dea4301

Observation e8796b3e-719e-4147-8566-b3f304acd6ff · outbound

This paper cites Magnetic control of tokamak plasmas through deep reinforcement learning,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Magnetic control of tokamak plasmas through deep reinforcement learning,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:57.214044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:47.105465Z digest=sha256:a539b7de78922f8269a4fa085810f5b248fa812db30009e56dd035e7b79d3bb3

Observation 612b66ac-8085-491b-ab01-7c23a58ebf16 · outbound

This paper cites Reinforcement learning for decision-making and control in power systems,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Reinforcement learning for decision-making and control in power systems,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:57.100933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:47.192450Z digest=sha256:e84775a6fbb095d1cb422916a992fdbabd2511f93a67ab60a33ff252159d22f0

Observation 25134737-fe53-4046-8510-6efc80983b3d · outbound

This paper cites Reinforcement learning with long short-term memory,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Reinforcement learning with long short-term memory,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:56.935703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:47.318792Z digest=sha256:1e5c7b2ec2350462d75814dfa5b04caabf654fb18d225dc63ae3b17bc18b4b3f

Observation 1c813bee-709a-4a04-8910-346eda6952fc · outbound

This paper cites Deep recurrent q-learning for partially observable mdps,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Deep recurrent q-learning for partially observable mdps,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:56.671191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:47.426053Z digest=sha256:34498b8e533328710ff1ab7e1fd63873100a5c98f3c193ef9d9bf4c7c06164b0

Observation f9a1636e-1115-4150-b901-67b8c5bc996d · outbound

This paper cites Attention is all you need,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Attention is all you need,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:56.482078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:47.548687Z digest=sha256:c22589139ce86792c0f6877525bda98c54995560b8f7c8c7f215016630cad578

Observation 93c61bfe-d5f4-4e00-8e33-ebfe93463c0d · outbound

This paper cites Decision transformer: Reinforcement learning via sequence modeling,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Decision transformer: Reinforcement learning via sequence modeling,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:56.202824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:47.679746Z digest=sha256:7ca329320e7b244d8873bdbd69ed9f4f27ee5fa7b2c387b0b31dde44a725d40a

Observation 1def55a1-4992-4de2-a9af-c0b7aab7be50 · outbound

This paper cites Online decision transformer,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Online decision transformer,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:55.971803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:47.813180Z digest=sha256:08407fd86a46bbe027cf25cf8d37618dbac8a3011030bfa7e9759fd8b7df3c93

Observation 06446c79-5413-42ee-a72e-54df07bcc829 · outbound

This paper cites Trajectory transformer: Model-based reinforcement learning with long-term dependencies,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Trajectory transformer: Model-based reinforcement learning with long-term dependencies,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:55.676674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:47.930098Z digest=sha256:957406bb31b3c7317c77724f58c4d8330f74e1fbf7a968a14fc517c09488ddb0

Observation 06ed75b3-7397-4c66-b1a0-905fdb64cad9 · outbound

This paper cites Optimal control of markov processes with incomplete state information,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Optimal control of markov processes with incomplete state information,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:55.408801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:48.023453Z digest=sha256:b357994aafac3e7e9bcc427675e26bebcb45615c039118c9d65efaa868f333e1

Observation a416ac10-d75e-4701-9591-b4d99b42f385 · outbound

This paper cites Multi- agent rollout and policy iteration for pomdp with application to multi- robot repair problems,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Multi- agent rollout and policy iteration for pomdp with application to multi- robot repair problems,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:55.081488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:48.117648Z digest=sha256:181eb7b85703fcc1f27887665cef9a3d77211d0ab138b2fc9c87da7e4e30afb8

Observation 980d06a0-93f1-46b0-8080-a46c11894433 · outbound

This paper cites Controlling contact-rich manipulation under partial observability,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Controlling contact-rich manipulation under partial observability,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:54.827760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:48.245135Z digest=sha256:32c60d0b009fcc749a880875728bb4ef199ac47dee3dbc5695b5c5826484ee67

Observation 1ac0fda2-bb84-4d20-b8f2-ca37bfe00c90 · outbound

This paper cites Magic: Learning macro-actions for online pomdp planning,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Magic: Learning macro-actions for online pomdp planning,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:54.530242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:48.364342Z digest=sha256:dfedbf7fc977c901551e3e5a4394a8c07700f11806ac1421810c9d752f1d5802

Observation 1ac94790-f526-4394-b40d-b91cc5ab4816 · outbound

This paper cites Optimizing active surveil- lance for prostate cancer using partially observable markov decision processes,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Optimizing active surveil- lance for prostate cancer using partially observable markov decision processes,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:54.272813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:48.451457Z digest=sha256:e8f0d3d463ed929ed979c63938efde17d4fb7efee5b05b441015d7d819b0cf83

Observation 82fe0bda-cff5-44d2-a44b-57e113b991b6 · outbound

This paper cites Diagnostic policies optimization for chronic diseases based on pomdp model,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Diagnostic policies optimization for chronic diseases based on pomdp model,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:53.980464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:48.549375Z digest=sha256:b3d06ef33d74c902fc34ddcc50fed18191f7b252612995c3918b0f23b8d14306

Observation a9cf74e5-acff-4bfd-9039-d452f373a715 · outbound

This paper cites Planning and acting in partially observable stochastic domains,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Planning and acting in partially observable stochastic domains,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T13:00:48.673350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:00:48.673350Z digest=sha256:834d9d979fabdb038b97fdc1d79817146666fd42c0a82bd521eebfc4ecdbe0a2

Observation 03faa789-ee6f-412e-8ece-635e2626ae33 · outbound

This paper cites Finding structure in time,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Finding structure in time,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T13:00:48.767588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:00:48.767588Z digest=sha256:6ae8809d17017f226085c8ea2601a624ca8e4de60d6e12afe7f98e352835b447

Observation c026bac9-5a73-4688-b88d-2ccd7dc3ea67 · outbound

This paper cites Deep Recurrent Q-Learning for Partially Observable MDPs.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Deep Recurrent Q-Learning for Partially Observable MDPs

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T13:00:48.834145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:00:48.834145Z digest=sha256:f21ff06ce963ee106dbce93a66c99bdd78d8dba2b7e0149cd1e0573ee597e2c0

Observation 3f246ff6-32ab-4de7-b26f-cc04966336d6 · outbound

This paper cites Long short-term memory,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Long short-term memory,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T13:00:48.927033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:00:48.927033Z digest=sha256:44a1b7e0683dcec42ac92e8099f4a2b0391205eb0e331e18e18e08eb5a1866f1

Observation 1158ee86-f78e-4150-b2d9-c65ef2d035bf · outbound

This paper cites Human-level control through deep reinforcement learning,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Human-level control through deep reinforcement learning,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T13:00:49.028677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:00:49.028677Z digest=sha256:e3022fec20c051f5d1c6e557a923400cdc4e6607c254e2abc69083e8c504705d

Observation fd299bc2-33a4-4dbd-b520-42e05fef0675 · outbound

This paper cites Recurrent determin- istic policy gradient method for bipedal locomotion on rough terrain challenge,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Recurrent determin- istic policy gradient method for bipedal locomotion on rough terrain challenge,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:53.738085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:49.133496Z digest=sha256:5bda0d9b524ac6ed26a2607d8fd96c3b72a4df2519b0fcb9f6a3bb4ddae3db36

Observation 25cf2b11-0e80-4300-be83-112202c4cea7 · outbound

This paper cites Recurrent soft actor critic reinforcement learning for demand response problems,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Recurrent soft actor critic reinforcement learning for demand response problems,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:53.499128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:49.253527Z digest=sha256:4561883317b124af408b09236d70d46e91d958452be8117fa4edb7edaf8cad61

Observation 8dc248fd-84b5-4920-a38b-e5e51796b581 · outbound

This paper cites Memory-based deep reinforcement learning for pomdps,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Memory-based deep reinforcement learning for pomdps,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:52.965017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:49.325360Z digest=sha256:17c9dd0b0e1d3cc0cfdbf35b6b0e3cfebf612b00593b3435ea0eabe809397ed7

Observation e5e61561-e27a-45bb-82a2-7cdd380b3d24 · outbound

This paper cites Addressing function ap- proximation error in actor-critic methods,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Addressing function ap- proximation error in actor-critic methods,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:52.114131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:49.382397Z digest=sha256:c4abfc29a5da9fdc3d1ba774098ffa502ec2f5e8288e80f7c54c6898afbcba60

Observation d87b4ba8-4666-4b32-9376-d4908df20c88 · outbound

This paper cites Recurrent model-free RL can be a strong baseline for many POMDPs,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Recurrent model-free RL can be a strong baseline for many POMDPs,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:51.891029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:49.445900Z digest=sha256:5c8cbe2a485fa5727b77ed0c267962e60bb4c2910667d323548a4699ab4f5650

Observation e8aa4f39-2325-4c08-a84a-14c8350b98b1 · outbound

This paper cites Ode-based recurrent model-free reinforcement learning for pomdps,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Ode-based recurrent model-free reinforcement learning for pomdps,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:51.663513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:49.540563Z digest=sha256:db4ad7e12b8a093b4bc8411da88340b4500ad92c1c4799c63d09061661292242

Observation 5885b47e-04a4-4382-b76e-44fe516dc020 · outbound

This paper cites Efficient recurrent off-policy rl requires a context-encoder-specific learning rate,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Efficient recurrent off-policy rl requires a context-encoder-specific learning rate,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:51.428197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:49.615798Z digest=sha256:ed4ac4437d3134427f12c43608c1a30cb1ffd1f0683bafd21a50e6dc83f190a8

Observation 94bc0942-4711-4676-ab77-67e683241ee3 · outbound

This paper cites The moving horizon estimation concept,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability The moving horizon estimation concept,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:51.257691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:49.721558Z digest=sha256:6f0b58ca83dc84b4c4292a8773927f1c1f800932f33b9f74a61d80955cebabc5

Observation 0eb34a92-e2a9-43e5-bfe7-c90ad9d335c2 · outbound

This paper cites an unresolved cited work.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T13:00:49.828362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:00:49.828362Z digest=sha256:fbe481bd01d4d7d9c35628a44e8f6780db6749b36333fe7067d4736bd70c45e9

Observation bade8861-cc21-45c7-b86a-e1fed2ee2fb5 · outbound

This paper cites Xception: Deep learning with depthwise separable con- volutions,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Xception: Deep learning with depthwise separable con- volutions,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:51.044199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:49.941340Z digest=sha256:a5c78a0b054de80c7e5ee1c4167757930b9d40f9d410623cc2a00eb2147fe080

Observation a175d284-c13c-4de4-a929-76aa000de7b5 · outbound

This paper cites Thin mobilenet: An enhanced mobilenet architecture,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Thin mobilenet: An enhanced mobilenet architecture,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:50.809998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:50.078357Z digest=sha256:0dfce43bc4525988ba72ad8b34a66921bb0ba2222c4c1fa9cf5548b71fd26ec9

Observation 7864e1a8-ca95-400d-81fe-c7983382935d · outbound

This paper cites Gymnasium,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Gymnasium,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:00:50.549173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:00:50.191643Z digest=sha256:5636f33fb5be025d3fa163e21bc8c56fa1b1dce0465db42ab11cca19e045da56

Observation 5462c8ff-cf8e-477c-852e-295f4576d4e3 · outbound

This paper cites Bert: Pre-training of deep bidirectional transformers for language understanding,.

A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Bert: Pre-training of deep bidirectional transformers for language understanding,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T13:00:50.297031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:00:50.297031Z digest=sha256:fcddc037faab4f0d6ae6115c96d0931a134b195e102cee820bd0bb91d2368035

Pith citing papers

No inbound Pith citation observations are available.