Pith. sign in

Paper Citation Record · LEDGER

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning

As of 22 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 0 inbound Pith citation observations for arXiv:2506.14125.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.14125 v1

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:01:08.600180Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

49 of 49 outbound references displayed

  • verified exact1
  • verified fuzzy37
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5be75a55-5d9c-4f6e-a65e-ce76b800917e · outbound

This paper cites Constrained policy optimization.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Constrained policy optimization

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.256897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.396381Z digest=sha256:1ed110e5d603f4d5e470a1d5d9d9114fd1d43c733f8643c66d153ecf4c397216

Observation 05f4c639-5402-4d0c-9e39-eb0f1b9d5967 · outbound

This paper cites Safe reinforcement learning via shielding, 2017.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Safe reinforcement learning via shielding, 2017

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.243510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.401174Z digest=sha256:32e573926c9085c50e0bfc11566f208a1971c18c8e3b0514c7169743b3837d6e

Observation 1d364716-2cde-46eb-9fbd-d48e60177199 · outbound

This paper cites Asymptotic properties of constrained markov decision processes.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Asymptotic properties of constrained markov decision processes

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.229000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.405566Z digest=sha256:3213172a2af79800eecffc76ed2e670f1aacc50009d9cddb69441d8c98abb1e2

Observation 4c69894e-990d-4215-be53-24a5235ab0b3 · outbound

This paper cites Deep reinforcement learning for demand response in distribution networks.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Deep reinforcement learning for demand response in distribution networks

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.215622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.410077Z digest=sha256:e4d02a49f12a360d24a5dc5818f4ceb330762385b0fb9bb4642ab6b03ce66bc8

Observation 7e3fd9b3-e9a8-407e-a3c0-638221a2a80f · outbound

This paper cites Dynamic allocations for multi-product distribution.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Dynamic allocations for multi-product distribution

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.202202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.414520Z digest=sha256:620a5e06078cb5647773e8107bee5e51ed12da470c582669ecbe8fa33cd538cd

Observation ba30aaeb-4974-4d4b-90e1-479690a6b11b · outbound

This paper cites Deliveries in an inventory/routing problem using stochastic dynamic programming.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Deliveries in an inventory/routing problem using stochastic dynamic programming

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.188578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.418841Z digest=sha256:71c9935539bc3947650a7bb34ed8660e9f328bb650498ac43b7956fc41d0a7bd

Observation 0e48722e-60d5-4a7f-be11-af9e8544ea56 · outbound

This paper cites Resource constrained deep reinforcement learning.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Resource constrained deep reinforcement learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.175242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.423605Z digest=sha256:0a82f748a0f23996112c952bab0d84c9c385313d040f7ff50f89759e379c85bb

Observation 7910ef68-b089-4069-b9cd-2697d104341e · outbound

This paper cites Duality between density function and value function with applications in constrained optimal control and Markov Decision Process.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Duality between density function and value function with applications in constrained optimal control and Markov Decision Process

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-15T20:01:08.759910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.427952Z digest=sha256:133833d256a9a6260db7125bf553633290ca105c3ea3e29b7f3f3913486911a7

Observation 12714306-cb76-499b-9d89-0877225c6f86 · outbound

This paper cites A tutorial on kernel density estimation and recent advances.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning A tutorial on kernel density estimation and recent advances

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.432595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.432595Z digest=sha256:6357cc248bf8302caefe1e95bee34ad22115c6bd360ac95cd659e97e54bb881e

Observation 7995c69f-1974-48c2-9d0b-468dcc0d49a7 · outbound

This paper cites Supervised fuzzy reinforcement learning for robot navigation.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Supervised fuzzy reinforcement learning for robot navigation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.153865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.436678Z digest=sha256:11b486aa9a35c7349efedd72ba6a19fe58b261a9e64cc2738e6a80a5ceda5e2a

Observation 41d4d7a9-e90b-488a-b1c9-a3cff33419b5 · outbound

This paper cites A comprehensive survey on safe reinforcement learning.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning A comprehensive survey on safe reinforcement learning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.142245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.440861Z digest=sha256:54b884471854d9ecc3500ae1975a87aab2dd7eb34420e9252539828127fa1a30

Observation ba56ffb1-40e2-42e6-9208-98fb0ce15ca4 · outbound

This paper cites Fuzzy q-learning.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Fuzzy q-learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.130738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.445200Z digest=sha256:ac16f08455c9e573e5b372251705025716596f68c325969887d0a0c090919ffa

Observation de379194-4561-4849-b674-b63d3bf68da1 · outbound

This paper cites Deep reinforcement learning for robotic manipulation with asynchronous off-policy updates.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Deep reinforcement learning for robotic manipulation with asynchronous off-policy updates

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.119336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.449306Z digest=sha256:40291f9ede415cdece3797f7673b2bc68b34ed737dabc84bfd06d9dea7de712e

Observation 14b093ad-91a2-433a-9ab2-e4d41e255575 · outbound

This paper cites A Review of Safe Reinforcement Learning: Methods, Theory and Applications.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning A Review of Safe Reinforcement Learning: Methods, Theory and Applications

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.453440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.453440Z digest=sha256:2405952d7a25619709ee356dced80749021cf6c441b9ab997751f8ffb5e0b16d

Observation 88062e48-5473-4e2f-9530-f51adaf64026 · outbound

This paper cites Learning to Walk in the Real World with Minimal Human Effort.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Learning to Walk in the Real World with Minimal Human Effort

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.458090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.458090Z digest=sha256:6e00586636f2083c0920a43cad42199d4b8eafa1e6fc1b1fbe9b5a802e28d471

Observation e3590007-a40a-422f-b7af-b26cb61e6b52 · outbound

This paper cites Hierarchical reinforcement learning for scarce medical resource allocation with imperfect information.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Hierarchical reinforcement learning for scarce medical resource allocation with imperfect information

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.107394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.462602Z digest=sha256:5d384b9bf0e9350fbdd638d1442189dd5a2f7a484048feef08ef52f405a0094c

Observation 5e5fb747-f99f-4985-bc58-65a0d267616a · outbound

This paper cites Logically-Constrained Reinforcement Learning.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Logically-Constrained Reinforcement Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.466900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.466900Z digest=sha256:00ece10fc5f164075bb48c3e7b79b2a8a05b128f0de428a86f5a7525c91ee280

Observation 82d539d9-c5d4-48c0-8671-88ab9ef72d1a · outbound

This paper cites Deep reinforcement learning with temporal logics.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Deep reinforcement learning with temporal logics

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.094339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.471573Z digest=sha256:e4b1e06441f73dd278247de8c2b18b082a4b4dc4d85ddc31a33c8325f5c59a47

Observation 89b97076-1a54-444e-a2d5-24564f2123b0 · outbound

This paper cites Achieving sustainable supply chains through energy justice.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Achieving sustainable supply chains through energy justice

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.081515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.475627Z digest=sha256:19d237351182290f2ca8fb43f875387c07a92873858d002ed804e79bb31733ca

Observation 202b9f2b-0152-4143-9156-db1170c6accc · outbound

This paper cites Line: Logical query reasoning over hierarchical knowledge graphs.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Line: Logical query reasoning over hierarchical knowledge graphs

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.068577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.479956Z digest=sha256:c53253bd8343ccbd02276762d8e8a6f4f084f2d650b36df292e705826f607331

Observation 8cdda88b-b8b6-45c6-9a1a-92f9d296e6da · outbound

This paper cites America's strategy to secure the supply chain for a robust clean energy transition.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning America's strategy to secure the supply chain for a robust clean energy transition

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.055482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.484102Z digest=sha256:9b362d9e830c0d1831c4c46d26d378904871df8f8ba29e453457113851b53210

Observation 2084936b-8d76-41c1-99e3-3d9a89687a1e · outbound

This paper cites Community-based operations research.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Community-based operations research

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.042539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.488583Z digest=sha256:5b63a9493a788a68ce9ac0bbb2d2840754f5c860165aa7951bc7b1f7377d3e10

Observation 302212df-a0a1-43b2-b284-e203a09e1a88 · outbound

This paper cites On bayesian index policies for sequential resource allocation.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning On bayesian index policies for sequential resource allocation

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.029257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.492684Z digest=sha256:eedfaa230280b6f42a26bcceefb6ad7925a20967e34b959e04e77673e319dbe9

Observation 578cea17-3554-4b64-aeaa-41ceb3b279b5 · outbound

This paper cites Between steps: Intermediate relaxations between big-m and convex hull formulations.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Between steps: Intermediate relaxations between big-m and convex hull formulations

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.016571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.496825Z digest=sha256:c2f0eaafdb5400a61b698aaf99301099e61a93a9478cd97dda2e541a1e2613e0

Observation c825ee03-52fb-44b7-bb5f-73b2dfd2e703 · outbound

This paper cites Augmenting Neural Networks with First-order Logic.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Augmenting Neural Networks with First-order Logic

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.500522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.500522Z digest=sha256:41144d4924453507df316922ef6dc030b766c83796a2be84f519aee0c943ccc5

Observation 6264bf78-637d-498b-91e9-5fc8dca917e0 · outbound

This paper cites Deep Reinforcement Learning for Efficient and Fair Allocation of Health Care Resources.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Deep Reinforcement Learning for Efficient and Fair Allocation of Health Care Resources

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.504544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.504544Z digest=sha256:4b8b33ac1a63762219dbeea4a3480dcfb0a399cad5cab2b7838723c768e3cf44

Observation a607d748-7a27-4367-9403-1aa91efb3936 · outbound

This paper cites Sequential resource allocation for nonprofit operations.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Sequential resource allocation for nonprofit operations

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.002610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.508841Z digest=sha256:65fcf150c0493a5627c46fd11f8c31bc5c4ebd2492f380717d28694b6f372dc7

Observation de61594b-3e94-43e1-b4d1-acbcbb0dfcc0 · outbound

This paper cites Clara: A constrained reinforcement learning based resource allocation framework for network slicing.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Clara: A constrained reinforcement learning based resource allocation framework for network slicing

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.990043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.513144Z digest=sha256:002c56b97ba1f9fcda16c9a087a4af4c736db6812c3e4cec29323e9378c5f769

Observation e973397e-f8a5-414d-9a06-5aebaba24179 · outbound

This paper cites Adaptive sequential surveillance with network and temporal dependence.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Adaptive sequential surveillance with network and temporal dependence

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.978102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.517342Z digest=sha256:a45b2b342d3a7bb7b3b4e706d18d8b2e989e74fb3cd2092158844536de11f9fc

Observation 80454cf3-ccc6-4a8d-b12b-b88fdd8c7b20 · outbound

This paper cites Ethical resource allocation in policing: Why policing requires a different approach from healthcare.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Ethical resource allocation in policing: Why policing requires a different approach from healthcare

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.966400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.521491Z digest=sha256:5d43fcc87e5900e17d9332bced7fedf882dd7dfb7f913e9b06cfc93b58ac0e62

Observation acc5a0ff-d3ef-48f5-ba61-29a15f77d610 · outbound

This paper cites A primal dual formulation for deep learning with constraints.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning A primal dual formulation for deep learning with constraints

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.954186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.525778Z digest=sha256:060c5a45d179ce3db58ec861f4275f1cbbb69818ad99cef07dce9551490bb6a9

Observation b7567450-0cc4-47cc-a7f3-979a69769425 · outbound

This paper cites Deep reinforcement learning approach for capacitated supply chain optimization under demand uncertainty.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Deep reinforcement learning approach for capacitated supply chain optimization under demand uncertainty

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.942113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.529994Z digest=sha256:a488d8edfaa1cd44883295ab60be38163ae444608facc309e47db1233ca8b2b0

Observation 92c40bbb-04a0-47f4-9bd7-c59d718e438c · outbound

This paper cites Combining fuzzy logic and reinforcement learning for resource management in edge computing.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Combining fuzzy logic and reinforcement learning for resource management in edge computing

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.928679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.534280Z digest=sha256:bc4b9bb392a3b47fd0a34838feebae1d4035f4dc803be70ce82080e845768736

Observation 1dfc4132-eedb-4c26-8214-c3b69eec8098 · outbound

This paper cites Fairness of the distribution of public medical and health resources.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Fairness of the distribution of public medical and health resources

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.915694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.538381Z digest=sha256:5080712d042849d30b04bf82c57cec0576e53da3ef4cace4e31632cbccbac646

Observation 27eb12f5-e496-440d-a687-bebfd19efc97 · outbound

This paper cites Density constrained reinforcement learning.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Density constrained reinforcement learning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.902621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.542300Z digest=sha256:f840f32af7e0cb76e86094f02be762578f026d1e6d441ec29e95af987b83bf69

Observation a0b9784c-736f-40ce-88d7-21c4d9e15602 · outbound

This paper cites A dual to lyapunov's stability theorem.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning A dual to lyapunov's stability theorem

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.889428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.547582Z digest=sha256:f7f162051f15b6193c2b815499c1db021e0ec7194b8facb4aac38b53594bd3b4

Observation 38398143-cf84-4274-b7c1-9790afab83c2 · outbound

This paper cites Benchmarking Safe Exploration in Deep Reinforcement Learning.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Benchmarking Safe Exploration in Deep Reinforcement Learning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.876651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.551759Z digest=sha256:e1efb763cbd075655059780b24f919c170463420e6aad91172c14782025da998

Observation 2f1b0077-17c1-4454-9540-4155843d11e5 · outbound

This paper cites Query2box: Reasoning over Knowledge Graphs in Vector Space using Box Embeddings.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Query2box: Reasoning over Knowledge Graphs in Vector Space using Box Embeddings

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.555712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.555712Z digest=sha256:3c5543fb70415902224b6fb5ceaf450efc8562f17916d807210af2095339ad23

Observation 133b02cf-616d-496b-90e0-9e10b66e26d4 · outbound

This paper cites Apprenticeship learning using linear programming.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Apprenticeship learning using linear programming

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.863196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.560069Z digest=sha256:b28f0ccfd4d25f2254ec9236ecfaa6082692de9f26e90a2ac6840865c7eae306

Observation c9d9f0ca-6b92-4be6-b98f-a1c9cd5c8665 · outbound

This paper cites The impacts of the covid-19 traffic light system on staff in tertiary education in new zealand.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning The impacts of the covid-19 traffic light system on staff in tertiary education in new zealand

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.848622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.564114Z digest=sha256:34718fcf6b27beb9cbc317f3acb742b47dde50c03cb18b3e4b5a93a742492588

Observation fe46283b-7b35-440a-9ffd-8a00c8395dcc · outbound

This paper cites Reward Constrained Policy Optimization.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Reward Constrained Policy Optimization

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.568016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.568016Z digest=sha256:154632aff1b8f8bf9ce714661b367ff3dc3428a887a83c8df3804429851603e8

Observation a1bfcaee-d873-4fde-ad75-fd784cc5df72 · outbound

This paper cites Improved big-m reformulation for generalized disjunctive programs.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Improved big-m reformulation for generalized disjunctive programs

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.833735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.572065Z digest=sha256:c6e30ddc4907e46345855da3168a2b416e176dab02b11f6bbdd20c5731df455f

Observation 9222da3e-c6b5-4756-a7ef-3f0fdf7b288c · outbound

This paper cites Disjunctive programming techniques for the optimization of process systems with discontinuous investment costs- multiple size regions.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Disjunctive programming techniques for the optimization of process systems with discontinuous investment costs- multiple size regions

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.820024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.576006Z digest=sha256:6fbf5f8c9d3e67ebcd59d1d8e0bebbee0c60b065777fde04c49bce7d7feebc37

Observation ea06e07e-2147-4ac1-b132-57ad133fc53d · outbound

This paper cites Dynamic shielding for reinforcement learning in black-box environments, 2022.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Dynamic shielding for reinforcement learning in black-box environments, 2022

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.806624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.580100Z digest=sha256:6c7674562e61f5d5544df8373d7a848620cfb2cbf803ad83a2584a1ba0c93646

Observation d1e43745-4fb0-4697-8826-476cac56d1cc · outbound

This paper cites Off-Policy Primal-Dual Safe Reinforcement Learning.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Off-Policy Primal-Dual Safe Reinforcement Learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.584153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.584153Z digest=sha256:195f591d4395dd89b1779fc566928d171972b7609d6644e89e877e9eabfd63fc

Observation 311564d5-b7e6-4960-aa54-f66b851479f6 · outbound

This paper cites Projection-Based Constrained Policy Optimization.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Projection-Based Constrained Policy Optimization

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.588304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.588304Z digest=sha256:67a2e5aa1ddee0b577e6c58f42f015d725575088de824573ac0eaef0043cc792

Observation eb4687ca-97ae-4ace-86ec-c1b1c18c5817 · outbound

This paper cites Learning density-based correlated equilibria for markov games.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Learning density-based correlated equilibria for markov games

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.794529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.592377Z digest=sha256:ae6725f4b0c5c6bddede1df5fe00308a71d49f2958e760e8ef6a7aa3c391fc4c

Observation 96cf58b7-de99-4668-9312-c989ccffde74 · outbound

This paper cites The energy injustice of hydropower: Development, resettlement, and social exclusion at the hongjiang and wanmipo hydropower stations in china.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning The energy injustice of hydropower: Development, resettlement, and social exclusion at the hongjiang and wanmipo hydropower stations in china

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.782022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.596286Z digest=sha256:a0a09b4f97ef0b6fde771a262572cb5d50ed1292e97c8648d8a7732858d53464

Observation 824a5941-4552-4a73-bb4b-6e728d9cc9b9 · outbound

This paper cites write newline.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning write newline

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.600180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.600180Z digest=sha256:b37ec337b5428f774d6f0d1bfe4db1b4bde9158cf3c112e90887a275b570b208

Pith citing papers

No inbound Pith citation observations are available.