Pith. sign in

Paper Citation Record · LEDGER

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning

As of 16 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 0 inbound Pith citation observations for arXiv:2506.14125.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.14125 v1

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:01:08.600180Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

49 of 49 outbound references displayed

  • verified exact1
  • verified fuzzy37
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5be75a55-5d9c-4f6e-a65e-ce76b800917e · outbound

This paper cites Constrained policy optimization.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Constrained policy optimization

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.256897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.396381Z digest=sha256:aed00eefbaab99a4935caf100d2040763ea4548e27d42496c589d11609a9835f

Observation 05f4c639-5402-4d0c-9e39-eb0f1b9d5967 · outbound

This paper cites Safe reinforcement learning via shielding, 2017.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Safe reinforcement learning via shielding, 2017

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.243510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.401174Z digest=sha256:3fb7f3e25ed9a8752e24cb8081f03b03581c87cebdbbe3a27a6dc01bfd8858c5

Observation 1d364716-2cde-46eb-9fbd-d48e60177199 · outbound

This paper cites Asymptotic properties of constrained markov decision processes.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Asymptotic properties of constrained markov decision processes

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.229000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.405566Z digest=sha256:b015695d021bb1a54e5ead4d0eb66c2380adabc9d5ab1af4c5a2ef31030edb7f

Observation 4c69894e-990d-4215-be53-24a5235ab0b3 · outbound

This paper cites Deep reinforcement learning for demand response in distribution networks.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Deep reinforcement learning for demand response in distribution networks

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.215622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.410077Z digest=sha256:48372f1dd2887b291baad30a4e399598071e6c8de4ba3165839771a02e18dce8

Observation 7e3fd9b3-e9a8-407e-a3c0-638221a2a80f · outbound

This paper cites Dynamic allocations for multi-product distribution.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Dynamic allocations for multi-product distribution

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.202202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.414520Z digest=sha256:121cf3e0bfe471f9e37ee800efc534db7f51ed8cdb41ca00d61d96a85a926032

Observation ba30aaeb-4974-4d4b-90e1-479690a6b11b · outbound

This paper cites Deliveries in an inventory/routing problem using stochastic dynamic programming.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Deliveries in an inventory/routing problem using stochastic dynamic programming

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.188578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.418841Z digest=sha256:255ed9c37f5c50ec3b0bbad5373b078b916aa717d29700e60f5a410bffdc752c

Observation 0e48722e-60d5-4a7f-be11-af9e8544ea56 · outbound

This paper cites Resource constrained deep reinforcement learning.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Resource constrained deep reinforcement learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.175242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.423605Z digest=sha256:f9efaee3b161d3cda0c861a87b05d12c0cdbd9a9e0db8719731ba943a09e3ab6

Observation 7910ef68-b089-4069-b9cd-2697d104341e · outbound

This paper cites Duality between density function and value function with applications in constrained optimal control and Markov Decision Process.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Duality between density function and value function with applications in constrained optimal control and Markov Decision Process

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-15T20:01:08.759910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.427952Z digest=sha256:c8181e5fbddb0788f512a5df9068b7bc648788f8c791b86baa459b0e38b669b0

Observation 12714306-cb76-499b-9d89-0877225c6f86 · outbound

This paper cites A tutorial on kernel density estimation and recent advances.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning A tutorial on kernel density estimation and recent advances

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.432595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.432595Z digest=sha256:6357cc248bf8302caefe1e95bee34ad22115c6bd360ac95cd659e97e54bb881e

Observation 7995c69f-1974-48c2-9d0b-468dcc0d49a7 · outbound

This paper cites Supervised fuzzy reinforcement learning for robot navigation.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Supervised fuzzy reinforcement learning for robot navigation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.153865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.436678Z digest=sha256:68e99d0acc3b0f2fc81d3e94f24af42c2738fef419a738b8888f5985251a70a9

Observation 41d4d7a9-e90b-488a-b1c9-a3cff33419b5 · outbound

This paper cites A comprehensive survey on safe reinforcement learning.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning A comprehensive survey on safe reinforcement learning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.142245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.440861Z digest=sha256:f61b6f0db0c2940f81e4134515f4f0cf1772cedcda7dd07b79e086c9ad9f2925

Observation ba56ffb1-40e2-42e6-9208-98fb0ce15ca4 · outbound

This paper cites Fuzzy q-learning.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Fuzzy q-learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.130738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.445200Z digest=sha256:84ae72c464ceb9c6c62708967a7cb6686da17d9324f7d8f4649779ece63d331e

Observation de379194-4561-4849-b674-b63d3bf68da1 · outbound

This paper cites Deep reinforcement learning for robotic manipulation with asynchronous off-policy updates.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Deep reinforcement learning for robotic manipulation with asynchronous off-policy updates

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.119336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.449306Z digest=sha256:aa88e964bdd4faeed4996d69e90aa7bf13926cb1fcdae7f23927fea4dae2cef5

Observation 14b093ad-91a2-433a-9ab2-e4d41e255575 · outbound

This paper cites A Review of Safe Reinforcement Learning: Methods, Theory and Applications.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning A Review of Safe Reinforcement Learning: Methods, Theory and Applications

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.453440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.453440Z digest=sha256:961f9dafad72f2c5282e4971d877c02675498ff72638635a2102661d851c17fe

Observation 88062e48-5473-4e2f-9530-f51adaf64026 · outbound

This paper cites Learning to Walk in the Real World with Minimal Human Effort.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Learning to Walk in the Real World with Minimal Human Effort

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.458090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.458090Z digest=sha256:bca547d479876525c0c4e9b04d4169e3c15990ab5db5cd249f415fd485c9ba15

Observation e3590007-a40a-422f-b7af-b26cb61e6b52 · outbound

This paper cites Hierarchical reinforcement learning for scarce medical resource allocation with imperfect information.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Hierarchical reinforcement learning for scarce medical resource allocation with imperfect information

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.107394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.462602Z digest=sha256:dae6a96d3a388087a88e37b441e7f2c043ac882010af05f998fc593c02ab0f9a

Observation 5e5fb747-f99f-4985-bc58-65a0d267616a · outbound

This paper cites Logically-Constrained Reinforcement Learning.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Logically-Constrained Reinforcement Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.466900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.466900Z digest=sha256:00ece10fc5f164075bb48c3e7b79b2a8a05b128f0de428a86f5a7525c91ee280

Observation 82d539d9-c5d4-48c0-8671-88ab9ef72d1a · outbound

This paper cites Deep reinforcement learning with temporal logics.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Deep reinforcement learning with temporal logics

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.094339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.471573Z digest=sha256:4545dcc525c2d58d86ea072b10a6ca667b4522a20230063f8378b58e7696e18a

Observation 89b97076-1a54-444e-a2d5-24564f2123b0 · outbound

This paper cites Achieving sustainable supply chains through energy justice.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Achieving sustainable supply chains through energy justice

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.081515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.475627Z digest=sha256:9e3d058c82c2786286433806fdddee144c18fcead797cf95445400d98bf06a6d

Observation 202b9f2b-0152-4143-9156-db1170c6accc · outbound

This paper cites Line: Logical query reasoning over hierarchical knowledge graphs.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Line: Logical query reasoning over hierarchical knowledge graphs

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.068577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.479956Z digest=sha256:015fad3346738f148f998031150662df0a16cec485636344401797b40731b044

Observation 8cdda88b-b8b6-45c6-9a1a-92f9d296e6da · outbound

This paper cites America's strategy to secure the supply chain for a robust clean energy transition.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning America's strategy to secure the supply chain for a robust clean energy transition

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.055482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.484102Z digest=sha256:43da9bc55419383a9f5a7be2adc838b2d98c97f65a1231091ab424b6bd7a23d5

Observation 2084936b-8d76-41c1-99e3-3d9a89687a1e · outbound

This paper cites Community-based operations research.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Community-based operations research

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.042539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.488583Z digest=sha256:fdbf88ca6ffccb43ec88d6167281f63e205903c7f2fe7023f91c92eb00454d55

Observation 302212df-a0a1-43b2-b284-e203a09e1a88 · outbound

This paper cites On bayesian index policies for sequential resource allocation.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning On bayesian index policies for sequential resource allocation

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.029257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.492684Z digest=sha256:7f05908cc4109c42e4e2e84a848fe8919a4f51d3a4b7d427b7ccfa983f5b93a7

Observation 578cea17-3554-4b64-aeaa-41ceb3b279b5 · outbound

This paper cites Between steps: Intermediate relaxations between big-m and convex hull formulations.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Between steps: Intermediate relaxations between big-m and convex hull formulations

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.016571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.496825Z digest=sha256:95c8dde876df4d1272caf1c9f937f2d5ad6a6b97b59fe73cbfadfc54c0d0a738

Observation c825ee03-52fb-44b7-bb5f-73b2dfd2e703 · outbound

This paper cites Augmenting Neural Networks with First-order Logic.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Augmenting Neural Networks with First-order Logic

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.500522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.500522Z digest=sha256:41144d4924453507df316922ef6dc030b766c83796a2be84f519aee0c943ccc5

Observation 6264bf78-637d-498b-91e9-5fc8dca917e0 · outbound

This paper cites Deep Reinforcement Learning for Efficient and Fair Allocation of Health Care Resources.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Deep Reinforcement Learning for Efficient and Fair Allocation of Health Care Resources

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.504544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.504544Z digest=sha256:be99b422c9127c314e80837cd960c0b04ebfa9cba85b36a05d293c147126b31f

Observation a607d748-7a27-4367-9403-1aa91efb3936 · outbound

This paper cites Sequential resource allocation for nonprofit operations.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Sequential resource allocation for nonprofit operations

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.002610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.508841Z digest=sha256:a18c49acbfae52313aa93cc002d5df9a5191aeb71bc805040564a35da5f22ca1

Observation de61594b-3e94-43e1-b4d1-acbcbb0dfcc0 · outbound

This paper cites Clara: A constrained reinforcement learning based resource allocation framework for network slicing.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Clara: A constrained reinforcement learning based resource allocation framework for network slicing

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.990043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.513144Z digest=sha256:28e021ce69445b2b41c0024a2a7b012252cd6df0dc9eaf3a9539d030bc5a7949

Observation e973397e-f8a5-414d-9a06-5aebaba24179 · outbound

This paper cites Adaptive sequential surveillance with network and temporal dependence.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Adaptive sequential surveillance with network and temporal dependence

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.978102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.517342Z digest=sha256:aa703380079fde8c386ea604a60ebcf0aa3343ee8eef388a8eebd467e08c0702

Observation 80454cf3-ccc6-4a8d-b12b-b88fdd8c7b20 · outbound

This paper cites Ethical resource allocation in policing: Why policing requires a different approach from healthcare.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Ethical resource allocation in policing: Why policing requires a different approach from healthcare

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.966400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.521491Z digest=sha256:9ecd3aa9fe70fa74189b0fb82bc18f078d7e12aeedac967200cee189c9a0d987

Observation acc5a0ff-d3ef-48f5-ba61-29a15f77d610 · outbound

This paper cites A primal dual formulation for deep learning with constraints.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning A primal dual formulation for deep learning with constraints

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.954186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.525778Z digest=sha256:8b2b23eb87b3a29478c6f0444c20591533fc685cb5c4d7b55a7dc47eae05bfd6

Observation b7567450-0cc4-47cc-a7f3-979a69769425 · outbound

This paper cites Deep reinforcement learning approach for capacitated supply chain optimization under demand uncertainty.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Deep reinforcement learning approach for capacitated supply chain optimization under demand uncertainty

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.942113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.529994Z digest=sha256:2dafab6f56409b498b3f837151cead5bd08391350500d2115377763e6763c0ea

Observation 92c40bbb-04a0-47f4-9bd7-c59d718e438c · outbound

This paper cites Combining fuzzy logic and reinforcement learning for resource management in edge computing.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Combining fuzzy logic and reinforcement learning for resource management in edge computing

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.928679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.534280Z digest=sha256:4929df5b3fcfb5b582a5d5d09e4c511d1f265055f170202dc4975fc7278837d9

Observation 1dfc4132-eedb-4c26-8214-c3b69eec8098 · outbound

This paper cites Fairness of the distribution of public medical and health resources.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Fairness of the distribution of public medical and health resources

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.915694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.538381Z digest=sha256:35781acd638b6b2baff3a9a06338b812fc1c77eda9dbf653e02835c56db2065d

Observation 27eb12f5-e496-440d-a687-bebfd19efc97 · outbound

This paper cites Density constrained reinforcement learning.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Density constrained reinforcement learning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.902621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.542300Z digest=sha256:afc0c6537e4f55e6d02b64d2fa99651a2b9ef6a0e6a321fb7a6ce2cc3569f7f1

Observation a0b9784c-736f-40ce-88d7-21c4d9e15602 · outbound

This paper cites A dual to lyapunov's stability theorem.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning A dual to lyapunov's stability theorem

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.889428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.547582Z digest=sha256:ae96b4b823119fa1023e91a3defee4c09fb95725d0352ea1f9ed710e6c11377b

Observation 38398143-cf84-4274-b7c1-9790afab83c2 · outbound

This paper cites Benchmarking Safe Exploration in Deep Reinforcement Learning.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Benchmarking Safe Exploration in Deep Reinforcement Learning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.876651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.551759Z digest=sha256:2810726d4a66f83ff77a1822748c49f09fccf6bbcfcf53a5b3075bcbbc5346f6

Observation 2f1b0077-17c1-4454-9540-4155843d11e5 · outbound

This paper cites Query2box: Reasoning over Knowledge Graphs in Vector Space using Box Embeddings.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Query2box: Reasoning over Knowledge Graphs in Vector Space using Box Embeddings

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.555712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.555712Z digest=sha256:3c5543fb70415902224b6fb5ceaf450efc8562f17916d807210af2095339ad23

Observation 133b02cf-616d-496b-90e0-9e10b66e26d4 · outbound

This paper cites Apprenticeship learning using linear programming.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Apprenticeship learning using linear programming

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.863196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.560069Z digest=sha256:df31c2fa18e34e27a5f0271207e2744381bb0b41804e901d5e9d97f06833f9c2

Observation c9d9f0ca-6b92-4be6-b98f-a1c9cd5c8665 · outbound

This paper cites The impacts of the covid-19 traffic light system on staff in tertiary education in new zealand.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning The impacts of the covid-19 traffic light system on staff in tertiary education in new zealand

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.848622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.564114Z digest=sha256:65dc1ec42482ce4c75299f1f9b9fe4f9a0baf0c537014f179ff7fca5220e0f55

Observation fe46283b-7b35-440a-9ffd-8a00c8395dcc · outbound

This paper cites Reward Constrained Policy Optimization.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Reward Constrained Policy Optimization

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.568016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.568016Z digest=sha256:154632aff1b8f8bf9ce714661b367ff3dc3428a887a83c8df3804429851603e8

Observation a1bfcaee-d873-4fde-ad75-fd784cc5df72 · outbound

This paper cites Improved big-m reformulation for generalized disjunctive programs.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Improved big-m reformulation for generalized disjunctive programs

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.833735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.572065Z digest=sha256:6fa90a310c9e0fda1947d913984711fb65ebd075f083fa1b12a070355b819b9c

Observation 9222da3e-c6b5-4756-a7ef-3f0fdf7b288c · outbound

This paper cites Disjunctive programming techniques for the optimization of process systems with discontinuous investment costs- multiple size regions.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Disjunctive programming techniques for the optimization of process systems with discontinuous investment costs- multiple size regions

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.820024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.576006Z digest=sha256:7d5c6a519de3747d81d2deb72e7f84f4dd606864193f6257e4c97e0000e6ab78

Observation ea06e07e-2147-4ac1-b132-57ad133fc53d · outbound

This paper cites Dynamic shielding for reinforcement learning in black-box environments, 2022.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Dynamic shielding for reinforcement learning in black-box environments, 2022

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.806624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.580100Z digest=sha256:f4edf741b178b4d1a8612d17d441478fc822fdfd50152f1a953f03f80a0ba86a

Observation d1e43745-4fb0-4697-8826-476cac56d1cc · outbound

This paper cites Off-Policy Primal-Dual Safe Reinforcement Learning.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Off-Policy Primal-Dual Safe Reinforcement Learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.584153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.584153Z digest=sha256:87844fe49c10cd7cf871fe4f07dfd013dd604a3a9fa18a2879e8246bc1af89c2

Observation 311564d5-b7e6-4960-aa54-f66b851479f6 · outbound

This paper cites Projection-Based Constrained Policy Optimization.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Projection-Based Constrained Policy Optimization

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.588304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.588304Z digest=sha256:aa402a33c98569014d790dfac823463365462cabbfb27fa8eddf707efbdf9ced

Observation eb4687ca-97ae-4ace-86ec-c1b1c18c5817 · outbound

This paper cites Learning density-based correlated equilibria for markov games.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Learning density-based correlated equilibria for markov games

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.794529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.592377Z digest=sha256:63b66de7b205335ea88c7dbb2ee03173777ecaa017e5adcabd6552cd0892eeba

Observation 96cf58b7-de99-4668-9312-c989ccffde74 · outbound

This paper cites The energy injustice of hydropower: Development, resettlement, and social exclusion at the hongjiang and wanmipo hydropower stations in china.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning The energy injustice of hydropower: Development, resettlement, and social exclusion at the hongjiang and wanmipo hydropower stations in china

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.782022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.596286Z digest=sha256:29c31c0cdd7c6fc2b9eb2add37010e221e650b4701e6b9d7ab7a43cdde59b09d

Observation 824a5941-4552-4a73-bb4b-6e728d9cc9b9 · outbound

This paper cites write newline.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning write newline

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.600180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.600180Z digest=sha256:b37ec337b5428f774d6f0d1bfe4db1b4bde9158cf3c112e90887a275b570b208

Pith citing papers

No inbound Pith citation observations are available.