Pith. sign in

Paper Citation Record · LEDGER

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report

As of 17 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 1 inbound Pith citation observation for arXiv:2501.11136.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.11136 v1

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T18:41:32.541911Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-13T12:50:02.212145Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

40 of 40 outbound references displayed

  • verified exact3
  • verified fuzzy32
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f0b0cc1a-6804-4042-84b7-e2ff4689ea69 · outbound

This paper cites Queueing Network Controls via Deep Reinforcement Learning,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Queueing Network Controls via Deep Reinforcement Learning,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.572517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.283220Z digest=sha256:4e47ddbf8ff5506d68f958e8b219d1a2cbcb7f64fbfedad7748b87bf83c91556

Observation b94b5740-a008-4a0f-931f-1fd8e85709fc · outbound

This paper cites Reinforcement learning in queues,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Reinforcement learning in queues,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.553813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.289976Z digest=sha256:2668d353e08cc5bf08ad439e10ec1004048ce807d6c4dd0e64cf2b5281b87378

Observation ad3f3fb4-a5cc-4df2-9263-a521ab76f925 · outbound

This paper cites Queue-Learning: A Reinforcement Learning Approach for Providing Quality of Service,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Queue-Learning: A Reinforcement Learning Approach for Providing Quality of Service,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.531770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.295865Z digest=sha256:58bbc354177a68995eab56b6fd334980851c3f8cd768ad98da44e61ed0e87915

Observation aac3882e-c12c-452f-bc1e-1d147d1c8958 · outbound

This paper cites Deep Reinforcement Learning for Smart Queue Management,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Deep Reinforcement Learning for Smart Queue Management,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.512264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.302477Z digest=sha256:d4c7ee2cf05ce595ad0cb6d83b19e73f42ac32536c0239562d1d65ba20771c58

Observation 752032a3-346b-446b-98a1-620a526fae69 · outbound

This paper cites Intervention-Assisted Policy Gradient Methods for Online Stochastic Queuing Network Optimization: Technical Report.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Intervention-Assisted Policy Gradient Methods for Online Stochastic Queuing Network Optimization: Technical Report

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-08-10T18:41:32.753342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.308346Z digest=sha256:357a3b6b7b96e7c79d7f4a00b5b43f58a781db9c75c5dbcb36ea1e1ef13494ee

Observation b713f99c-b954-49f6-9459-db6807bd3bcb · outbound

This paper cites Approximation theory of the MLP model in neural net- works,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Approximation theory of the MLP model in neural net- works,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.493501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.315605Z digest=sha256:4d8822917064de026a4504eca51b97d73d580fa402de6f13138ff2bdd8f61052

Observation 680c0180-b7c5-49b2-9d3d-57bffb4e1966 · outbound

This paper cites Scheduling Algorithms for Minimizing Age of Information in Wireless Broadcast Networks with Random Arrivals: The No-Buffer Case.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Scheduling Algorithms for Minimizing Age of Information in Wireless Broadcast Networks with Random Arrivals: The No-Buffer Case

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-10T18:41:32.721137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.322186Z digest=sha256:0c12a45651f18fc8ad9d9f8f9c77fa028fe2c1e35ad8e417bc25c2e072707c61

Observation d2dc76bc-445b-4889-93af-bebd3ae250b4 · outbound

This paper cites Dynamic server allocation to parallel queues with randomly varying connectivity,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Dynamic server allocation to parallel queues with randomly varying connectivity,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.474295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.330392Z digest=sha256:94f6a93865dc37bb3486462677b4b94434cd121bc520e4e577b2560608305e26

Observation 6f57e04c-3fe5-41ed-a98b-328e3ff1ce0a · outbound

This paper cites Stability and Asymptotic Optimality of Generalized MaxWeight Policies,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Stability and Asymptotic Optimality of Generalized MaxWeight Policies,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.455153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.336474Z digest=sha256:ac36449a781f8e31fe277929538d1a7815b45877cfeee2cf0bf461481175d42d

Observation efaf5b56-3f74-47e1-a0d8-54e099129f54 · outbound

This paper cites Minimizing the Age of Information in Wireless Networks with Stochastic Arrivals,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Minimizing the Age of Information in Wireless Networks with Stochastic Arrivals,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.421670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.342562Z digest=sha256:63c2781f481d4d2563b01a3f2f3766cdb75d3be111b1ab1f7e265b74c48da812

Observation 362e7a4f-5e50-42b6-8b1b-8e649316779b · outbound

This paper cites Tracking MaxWeight: Optimal Control for Partially Observable and Controllable Networks,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Tracking MaxWeight: Optimal Control for Partially Observable and Controllable Networks,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.402215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.347459Z digest=sha256:aa7bfa62545328ba0c1da717850636257cd9cc861ec3adcfff5a0e0704619d69

Observation 4df6132d-8393-4016-9556-5c3d1f21cb20 · outbound

This paper cites MaxWeight scheduling in a generalized switch: State space collapse and workload minimization in heavy traffic,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report MaxWeight scheduling in a generalized switch: State space collapse and workload minimization in heavy traffic,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.376714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.355890Z digest=sha256:eecfab7cac52708f9923486bdde5ef1ae26249b2f9c17bb63e9dfa5050f2e5ef

Observation 02a4ba55-1ccc-4894-890f-a8601d680622 · outbound

This paper cites An analysis of the join the shortest queue (JSQ) policy,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report An analysis of the join the shortest queue (JSQ) policy,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.353067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.361931Z digest=sha256:57911cfaa009d564739d14b979ca0df636d24e33b8d1ef53ef56aaedd125677d

Observation d3cf1b9d-ae52-47f8-8188-6e76966038ce · outbound

This paper cites Stochastic Network Optimization with Application to Com- munication and Queueing Systems,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Stochastic Network Optimization with Application to Com- munication and Queueing Systems,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.335308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.368583Z digest=sha256:510974f1c0093be8b5d87e1dcb4a0ba022899b3c85420e26bbfa6c1590b294f8

Observation ae15885e-2e2d-4f58-9c64-248e9446748c · outbound

This paper cites Restless bandits: activity allocation in a changing world,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Restless bandits: activity allocation in a changing world,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.318447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.374339Z digest=sha256:5d1c17afab5441235e4ebd8c78bbbb305b612f57dfd5b43bab5af99371eb1f01

Observation f78fbe7c-09f5-49bb-a8c4-f7759bce5624 · outbound

This paper cites Dynamic priority allocation via restless bandit marginal productivity indices,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Dynamic priority allocation via restless bandit marginal productivity indices,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.291826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.380870Z digest=sha256:02eee7d78d69365667a22284354e68c14557a4d9d61882d225387a2eba4db31c

Observation bff8e695-decb-4690-a684-993164296799 · outbound

This paper cites Whit- tle’s index policy for a multi-class queueing system with convex holding costs,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Whit- tle’s index policy for a multi-class queueing system with convex holding costs,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.270034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.386276Z digest=sha256:acd1215d8fa5633efad010e1a813080a88d9869c7068274ff9dab1c57a0ee8dd

Observation 98c0a9de-148f-4951-80de-a8eae26c160f · outbound

This paper cites Congestion control of TCP flows in Internet routers by means of index policy,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Congestion control of TCP flows in Internet routers by means of index policy,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.243558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.392312Z digest=sha256:e679a637e09bf684a0e78a66694feaef09440970e9846bd13f42cc6c206470e5

Observation 92535104-44fc-42fc-89eb-906537382e77 · outbound

This paper cites Indexability of Restless Bandit Problems and Optimality of Whittle Index for Dynamic Multichannel Access,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Indexability of Restless Bandit Problems and Optimality of Whittle Index for Dynamic Multichannel Access,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.222731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.398235Z digest=sha256:4dbc67f716560e2b88e294ffe6d0ce37497acdacc98d773dc6a6757d63e0a05f

Observation 64fb8c22-1dd0-4df4-b24b-3572efe7d6fd · outbound

This paper cites A unifying computations of Whittle’s Index for Markovian bandits,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report A unifying computations of Whittle’s Index for Markovian bandits,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.200534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.406430Z digest=sha256:14e0341fcd4b3f083bc1dbd5e529a03a591c1d92e02ab8a9501bacf09efeed66

Observation 587454cc-cc2c-4658-80d6-749ae2cc4c19 · outbound

This paper cites Optimal control of a queueing system with two heterogeneous servers,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Optimal control of a queueing system with two heterogeneous servers,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.178362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.412195Z digest=sha256:aa1cd6805001b5e7ab7195f7ac8a14b7076e168629d532122ef63ad55b366e6b

Observation 11aace9d-cd2c-4d04-8415-14af7b71fddb · outbound

This paper cites A simple proof of the optimality of a threshold policy in a two-server queueing system,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report A simple proof of the optimality of a threshold policy in a two-server queueing system,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.156219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.417178Z digest=sha256:16f5dfcfabdc46ae48f399a05909c8c6c67d51367bab6170215e476d34596bd5

Observation b6e1c794-9337-4f1d-9a79-54a444e61274 · outbound

This paper cites Extension of the optimality of the threshold policy in heterogeneous multiserver queueing systems,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Extension of the optimality of the threshold policy in heterogeneous multiserver queueing systems,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.135852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.426124Z digest=sha256:147e486cfe123826e0c8b158a263de28b2b0640527b4736ef33f9a5b584e5ea9

Observation dec16675-64b9-467a-9720-3cdf98188033 · outbound

This paper cites Monotone Control of Queueing Systems with Heterogeneous Servers,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Monotone Control of Queueing Systems with Heterogeneous Servers,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.113874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.432313Z digest=sha256:3c3779ecc5858819606ba5e3a775cc02fab6d5f978c71309947a0735ddd9c722

Observation 499cb480-10d0-4586-98a0-c728e2eaa74e · outbound

This paper cites an unresolved cited work.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-10T18:41:33.091542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.438737Z digest=sha256:651fe9f23371075948ce1bf5d028daa78eac41cd19606e1dea6aa35d642a46a3

Observation 18857e94-4ee9-4a5c-9845-139e65c515db · outbound

This paper cites Average Cost Optimal Stationary Policies in Infinite State Markov Decision Processes with Unbounded Costs,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Average Cost Optimal Stationary Policies in Infinite State Markov Decision Processes with Unbounded Costs,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.061073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.444936Z digest=sha256:0af38e2d9f5b733dc5acbc5fc09507dfb9d1cc14c1ea4c641d9c9a64a73c0674

Observation 83f7999e-c1ce-40d5-a1ae-e9242cc3db28 · outbound

This paper cites an unresolved cited work.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-10T18:41:33.036842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.450780Z digest=sha256:3af74b45946482034067992daf28dbc2ca84fbe4332e211d0d03210e96ecad45

Observation 5677feb2-10cf-4ba0-b2bb-83a533fd043a · outbound

This paper cites Proximal Policy Optimization Algorithms.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Proximal Policy Optimization Algorithms

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T18:41:32.463691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:41:32.463691Z digest=sha256:7e38328f7fcdc14a85a9ad10aaf7d7c26706321d867a3a35d924a82c36d93731

Observation 38b5460f-305d-45e2-b98e-55da1756330e · outbound

This paper cites A Dissection of Overfitting and Generalization in Continuous Reinforcement Learning,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report A Dissection of Overfitting and Generalization in Continuous Reinforcement Learning,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.010974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.474879Z digest=sha256:a9f3d03bf0dd0c185964a902cf3c8a0163256d260fa9813f817b669fca90a91d

Observation 0a4b2f77-ec85-43d2-abc2-bcbda093953c · outbound

This paper cites A Survey of Zero-shot Generalisation in Deep Reinforcement Learning.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report A Survey of Zero-shot Generalisation in Deep Reinforcement Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T18:41:32.481954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:41:32.481954Z digest=sha256:a51225b16326b5f222c422a15158e733b3d9aab2fecff8826c372d0d4d81cfe3

Observation 4cca26f9-0d7e-4721-af68-d70354ad00ad · outbound

This paper cites Quantifying Generalization in Reinforcement Learning,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Quantifying Generalization in Reinforcement Learning,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:32.990143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.488187Z digest=sha256:1a3f27ad5e3cb3c2b5672a2b1386b5d13cc08049db9e198d8e130e232b0a7187

Observation 96078e02-3631-4aa5-a1fd-5e1ffc46856f · outbound

This paper cites Neuroevolution of self-interpretable agents,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Neuroevolution of self-interpretable agents,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:32.961108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.493985Z digest=sha256:3847298212733b94f9dce781ddd4a3558aa3287dfbbed6147bd394f2dba89fb7

Observation f2f9bf7f-e24e-4ebb-bbb8-d062c8f74d25 · outbound

This paper cites The Sensory Neuron as a Transformer: Permutation-Invariant Neural Networks for Reinforcement Learning.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report The Sensory Neuron as a Transformer: Permutation-Invariant Neural Networks for Reinforcement Learning

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-08-10T18:41:32.629516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.499838Z digest=sha256:bdfb58fe8429c0ac407751fd5eb1b47c5815c59c8879d106b692eda0b7e92628

Observation 57b98c88-9c07-4770-988c-72c2c38c560c · outbound

This paper cites Unsupervised Visual Attention and Invariance for Reinforcement Learning,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Unsupervised Visual Attention and Invariance for Reinforcement Learning,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:32.937935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.505623Z digest=sha256:a517e2faa8a40bd08fead4ab95d63da574c260fa9b5e01e8e2374a62d440566b

Observation 9896f5d2-ea73-42d5-b17c-9008958f0025 · outbound

This paper cites Deep reinforcement learning with relational inductive biases,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Deep reinforcement learning with relational inductive biases,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:32.914959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.511121Z digest=sha256:feb38f0c6f3804ac0773d4465c8cd8ca62efae8bfeeba3074163b006fdc5b4a4

Observation 7e9539ea-2e76-4d5d-9e87-359d74ef06a9 · outbound

This paper cites DEEP REINFORCEMENT LEARNING WITH RELATIONAL INDUCTIVE BIASES,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report DEEP REINFORCEMENT LEARNING WITH RELATIONAL INDUCTIVE BIASES,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:32.886922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.517130Z digest=sha256:87a6e82a26665d8ea76b48046e00a7dd7d2533b6bb2dcc44cf05051999011d27

Observation a064ee0a-beef-4b28-a12f-dfb319f94c61 · outbound

This paper cites Neuro-algorithmic Policies Enable Fast Combinatorial Generalization,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Neuro-algorithmic Policies Enable Fast Combinatorial Generalization,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:32.846276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.523146Z digest=sha256:9571fb34752cae435067e1ee262648112e571f282e8b53d31ffb9522efd24933

Observation 0f2ba949-9d1a-40e0-a332-954b7cdf8098 · outbound

This paper cites Scalable Monotonic Neural Networks,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Scalable Monotonic Neural Networks,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:32.820684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.528670Z digest=sha256:30f068233f5b7e762153e49cc0a23a8992cadf78bc42cfc8ca09784c4a1ec214

Observation 11258812-fce2-485f-a6b8-72eb20f4bcac · outbound

This paper cites Bounded activation functions for enhanced training stability of deep neural networks on visual pattern recognition problems,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Bounded activation functions for enhanced training stability of deep neural networks on visual pattern recognition problems,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:32.783428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.534941Z digest=sha256:3fb98ebe7101c3e1f4234a414f7485903424516a6032bf97ce230015de2c98b0

Observation f4181248-e694-4e06-a8bf-a9a0350572ff · outbound

This paper cites Certified Monotonic Neural Networks.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Certified Monotonic Neural Networks

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T18:41:32.541911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:41:32.541911Z digest=sha256:16ef8d081d8479e72e15d63fe26ab12755dd4b17779a860daf579963d5b252ad

Pith citing papers

Observation a26e53d4-3619-4abb-be75-4ada0bd4ff8b · inbound

Multi-Robot Multi-Queue Control via Exhaustive Assignment Actor-Critic Learning cites this paper.

Multi-Robot Multi-Queue Control via Exhaustive Assignment Actor-Critic Learning A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-13T12:50:02.212145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T12:50:02.212145Z digest=sha256:5bfe198f60c8e7fd82123390e61fa7378e5524b5a81a56ec76d6f773992a3e58