Pith. sign in

Paper Citation Record · LEDGER

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report

As of 17 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 1 inbound Pith citation observation for arXiv:2501.11136.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.11136 v1

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T18:41:32.541911Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-13T12:50:02.212145Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

40 of 40 outbound references displayed

  • verified exact3
  • verified fuzzy32
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f0b0cc1a-6804-4042-84b7-e2ff4689ea69 · outbound

This paper cites Queueing Network Controls via Deep Reinforcement Learning,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Queueing Network Controls via Deep Reinforcement Learning,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.572517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.283220Z digest=sha256:428d62115d478103971eec5c175934bb49ca27f20b644d3600c42b609176b504

Observation b94b5740-a008-4a0f-931f-1fd8e85709fc · outbound

This paper cites Reinforcement learning in queues,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Reinforcement learning in queues,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.553813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.289976Z digest=sha256:233871c768badb44cf68fd07673365b5118ffc35bdbeecd8b35f9e7344effeb4

Observation ad3f3fb4-a5cc-4df2-9263-a521ab76f925 · outbound

This paper cites Queue-Learning: A Reinforcement Learning Approach for Providing Quality of Service,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Queue-Learning: A Reinforcement Learning Approach for Providing Quality of Service,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.531770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.295865Z digest=sha256:df74e6edf346c075b3f6f6f74a5f6bc5561c1674666921aa88c18b421d2a37c9

Observation aac3882e-c12c-452f-bc1e-1d147d1c8958 · outbound

This paper cites Deep Reinforcement Learning for Smart Queue Management,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Deep Reinforcement Learning for Smart Queue Management,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.512264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.302477Z digest=sha256:b9a4fb296ed0a76c7606750f470c7ae9005a3da561e200c6b0150fe813e11665

Observation 752032a3-346b-446b-98a1-620a526fae69 · outbound

This paper cites Intervention-Assisted Policy Gradient Methods for Online Stochastic Queuing Network Optimization: Technical Report.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Intervention-Assisted Policy Gradient Methods for Online Stochastic Queuing Network Optimization: Technical Report

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-08-10T18:41:32.753342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.308346Z digest=sha256:724c66cd207f72820d929df863de5bbdc5fdc75ff0a96a0d33c53ad86472f4ae

Observation b713f99c-b954-49f6-9459-db6807bd3bcb · outbound

This paper cites Approximation theory of the MLP model in neural net- works,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Approximation theory of the MLP model in neural net- works,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.493501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.315605Z digest=sha256:c6935c45c41a3a791e37bc8aa70049905bdba82911f36e5ceba6acb14434b03b

Observation 680c0180-b7c5-49b2-9d3d-57bffb4e1966 · outbound

This paper cites Scheduling Algorithms for Minimizing Age of Information in Wireless Broadcast Networks with Random Arrivals: The No-Buffer Case.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Scheduling Algorithms for Minimizing Age of Information in Wireless Broadcast Networks with Random Arrivals: The No-Buffer Case

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-10T18:41:32.721137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.322186Z digest=sha256:c4ee851493bda9ba8d056b6dffd47ff1963752f10a24406b2c47b22a7fdca2f4

Observation d2dc76bc-445b-4889-93af-bebd3ae250b4 · outbound

This paper cites Dynamic server allocation to parallel queues with randomly varying connectivity,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Dynamic server allocation to parallel queues with randomly varying connectivity,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.474295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.330392Z digest=sha256:977140b6e59aa19e0cbd31d065956f7186de09de6dd8fae1e298f78bf44e55d7

Observation 6f57e04c-3fe5-41ed-a98b-328e3ff1ce0a · outbound

This paper cites Stability and Asymptotic Optimality of Generalized MaxWeight Policies,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Stability and Asymptotic Optimality of Generalized MaxWeight Policies,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.455153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.336474Z digest=sha256:be2b814a6831245139b4f8bd1bef254956060544ce012ec928a133604c440629

Observation efaf5b56-3f74-47e1-a0d8-54e099129f54 · outbound

This paper cites Minimizing the Age of Information in Wireless Networks with Stochastic Arrivals,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Minimizing the Age of Information in Wireless Networks with Stochastic Arrivals,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.421670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.342562Z digest=sha256:cefa0094adc35fb34197f3fbd9575b8d1974a36bffd18d2abb792bce14b5d506

Observation 362e7a4f-5e50-42b6-8b1b-8e649316779b · outbound

This paper cites Tracking MaxWeight: Optimal Control for Partially Observable and Controllable Networks,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Tracking MaxWeight: Optimal Control for Partially Observable and Controllable Networks,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.402215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.347459Z digest=sha256:c23b173ef715c0101f3e6216c5e247c1851029fad7bfe7c1980690de4fc45f68

Observation 4df6132d-8393-4016-9556-5c3d1f21cb20 · outbound

This paper cites MaxWeight scheduling in a generalized switch: State space collapse and workload minimization in heavy traffic,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report MaxWeight scheduling in a generalized switch: State space collapse and workload minimization in heavy traffic,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.376714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.355890Z digest=sha256:f396d1bece03f99b6420704d5b12e63bc722ad83d5173a20e16bcddc76aa0869

Observation 02a4ba55-1ccc-4894-890f-a8601d680622 · outbound

This paper cites An analysis of the join the shortest queue (JSQ) policy,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report An analysis of the join the shortest queue (JSQ) policy,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.353067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.361931Z digest=sha256:f686404efb23151b7ced747d9f17eb34ae67ed94f84ba3eff5c4e24a16d1ee17

Observation d3cf1b9d-ae52-47f8-8188-6e76966038ce · outbound

This paper cites Stochastic Network Optimization with Application to Com- munication and Queueing Systems,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Stochastic Network Optimization with Application to Com- munication and Queueing Systems,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.335308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.368583Z digest=sha256:22aeddcb8b84978b54dc8a705521ed635cb8a4d0f0584f219ef25135109db337

Observation ae15885e-2e2d-4f58-9c64-248e9446748c · outbound

This paper cites Restless bandits: activity allocation in a changing world,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Restless bandits: activity allocation in a changing world,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.318447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.374339Z digest=sha256:cd04c1c8aeed66abdcac716cd801f68fee14a41ed5ac8b2083a56ab2113668bb

Observation f78fbe7c-09f5-49bb-a8c4-f7759bce5624 · outbound

This paper cites Dynamic priority allocation via restless bandit marginal productivity indices,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Dynamic priority allocation via restless bandit marginal productivity indices,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.291826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.380870Z digest=sha256:4597b36d9a5e28cc70907b14c17736bf842daf649e232655cd03294bf3c0a118

Observation bff8e695-decb-4690-a684-993164296799 · outbound

This paper cites Whit- tle’s index policy for a multi-class queueing system with convex holding costs,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Whit- tle’s index policy for a multi-class queueing system with convex holding costs,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.270034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.386276Z digest=sha256:e720ab65ec92b3ec9d8f41b7f8b7fe92faaa0316a3b1478b74f8f0095cd5dea3

Observation 98c0a9de-148f-4951-80de-a8eae26c160f · outbound

This paper cites Congestion control of TCP flows in Internet routers by means of index policy,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Congestion control of TCP flows in Internet routers by means of index policy,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.243558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.392312Z digest=sha256:84e19e1490c9677f7093f6a0072fd2365b84c1a473b97dc44d72f3997dbe6099

Observation 92535104-44fc-42fc-89eb-906537382e77 · outbound

This paper cites Indexability of Restless Bandit Problems and Optimality of Whittle Index for Dynamic Multichannel Access,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Indexability of Restless Bandit Problems and Optimality of Whittle Index for Dynamic Multichannel Access,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.222731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.398235Z digest=sha256:4e308aa11d54895648486c7329c546f3e3a32534489f855856ec4bf43adfcb56

Observation 64fb8c22-1dd0-4df4-b24b-3572efe7d6fd · outbound

This paper cites A unifying computations of Whittle’s Index for Markovian bandits,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report A unifying computations of Whittle’s Index for Markovian bandits,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.200534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.406430Z digest=sha256:11c1f4fb09119061dd85a6c211e10c3674c4be25fc7fc42dccbcbe6ea5e1c21e

Observation 587454cc-cc2c-4658-80d6-749ae2cc4c19 · outbound

This paper cites Optimal control of a queueing system with two heterogeneous servers,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Optimal control of a queueing system with two heterogeneous servers,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.178362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.412195Z digest=sha256:52e69adb309e37c7eabbc08cf9167b0a315a840c13240c6600f3b0073dcdd08f

Observation 11aace9d-cd2c-4d04-8415-14af7b71fddb · outbound

This paper cites A simple proof of the optimality of a threshold policy in a two-server queueing system,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report A simple proof of the optimality of a threshold policy in a two-server queueing system,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.156219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.417178Z digest=sha256:5cfc0ab6f6807fd096961867fb113c2f7cfc01c20bf23e79c6f515e799335d3e

Observation b6e1c794-9337-4f1d-9a79-54a444e61274 · outbound

This paper cites Extension of the optimality of the threshold policy in heterogeneous multiserver queueing systems,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Extension of the optimality of the threshold policy in heterogeneous multiserver queueing systems,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.135852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.426124Z digest=sha256:5ecc842854a36b6bc5c14e6d28cdf9aca0c798866b4ab4a416dc35c825ec7fd2

Observation dec16675-64b9-467a-9720-3cdf98188033 · outbound

This paper cites Monotone Control of Queueing Systems with Heterogeneous Servers,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Monotone Control of Queueing Systems with Heterogeneous Servers,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.113874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.432313Z digest=sha256:98654e377d191dd1ce99c93f3378d47b0e29075dda19b5c1ffdeb97297a2f5bd

Observation 499cb480-10d0-4586-98a0-c728e2eaa74e · outbound

This paper cites an unresolved cited work.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-10T18:41:33.091542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.438737Z digest=sha256:6d61b47ba5cbf1844fe30c6d5add0ea812a3354233abc065a99f9ed08265afca

Observation 18857e94-4ee9-4a5c-9845-139e65c515db · outbound

This paper cites Average Cost Optimal Stationary Policies in Infinite State Markov Decision Processes with Unbounded Costs,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Average Cost Optimal Stationary Policies in Infinite State Markov Decision Processes with Unbounded Costs,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.061073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.444936Z digest=sha256:20879e54c4dcd92530b65a8fb52ea95d506a084922f40e97de133ded616bd5cc

Observation 83f7999e-c1ce-40d5-a1ae-e9242cc3db28 · outbound

This paper cites an unresolved cited work.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-10T18:41:33.036842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.450780Z digest=sha256:6acac40c047ee8d2b2376ec3f4691b2b2660de4457e4debf4acb59458ea115ef

Observation 5677feb2-10cf-4ba0-b2bb-83a533fd043a · outbound

This paper cites Proximal Policy Optimization Algorithms.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Proximal Policy Optimization Algorithms

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T18:41:32.463691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:41:32.463691Z digest=sha256:136faaa23255b255e133a93451d171507791e3a480abce2efcb6f5b6c517541b

Observation 38b5460f-305d-45e2-b98e-55da1756330e · outbound

This paper cites A Dissection of Overfitting and Generalization in Continuous Reinforcement Learning,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report A Dissection of Overfitting and Generalization in Continuous Reinforcement Learning,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.010974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.474879Z digest=sha256:46c15e4dfb47eaacf33862dc41e5e5f58114fe0aa60efdd2f1c8825e8d014176

Observation 0a4b2f77-ec85-43d2-abc2-bcbda093953c · outbound

This paper cites A Survey of Zero-shot Generalisation in Deep Reinforcement Learning.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report A Survey of Zero-shot Generalisation in Deep Reinforcement Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T18:41:32.481954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:41:32.481954Z digest=sha256:4973b7224776bf22d9814b5bf145ec79e76856f4900bd5eb0bd9f1fc02d57344

Observation 4cca26f9-0d7e-4721-af68-d70354ad00ad · outbound

This paper cites Quantifying Generalization in Reinforcement Learning,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Quantifying Generalization in Reinforcement Learning,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:32.990143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.488187Z digest=sha256:9d8c28eabc1b5c0f4bb14b59a6baec0c678427a69b0e295e5a675e60e8ecac36

Observation 96078e02-3631-4aa5-a1fd-5e1ffc46856f · outbound

This paper cites Neuroevolution of self-interpretable agents,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Neuroevolution of self-interpretable agents,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:32.961108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.493985Z digest=sha256:e1e5bebc940a2bd5c2439b5f6da5e3421b5835aa43ca19f0e493b21036b45d0d

Observation f2f9bf7f-e24e-4ebb-bbb8-d062c8f74d25 · outbound

This paper cites The Sensory Neuron as a Transformer: Permutation-Invariant Neural Networks for Reinforcement Learning.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report The Sensory Neuron as a Transformer: Permutation-Invariant Neural Networks for Reinforcement Learning

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-08-10T18:41:32.629516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.499838Z digest=sha256:ed3aab8d433db82c77d9402f20a8574859fcf033efcb683af8fdf14df7912fd3

Observation 57b98c88-9c07-4770-988c-72c2c38c560c · outbound

This paper cites Unsupervised Visual Attention and Invariance for Reinforcement Learning,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Unsupervised Visual Attention and Invariance for Reinforcement Learning,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:32.937935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.505623Z digest=sha256:643960312b1c0a2a3e05cad7144c5d0b1e518cf546a4400b54a8275107c1dbab

Observation 9896f5d2-ea73-42d5-b17c-9008958f0025 · outbound

This paper cites Deep reinforcement learning with relational inductive biases,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Deep reinforcement learning with relational inductive biases,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:32.914959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.511121Z digest=sha256:da03f47375decd58b94ef8404b1a19c32da3c1127e3e4331bc0500b682deabf1

Observation 7e9539ea-2e76-4d5d-9e87-359d74ef06a9 · outbound

This paper cites DEEP REINFORCEMENT LEARNING WITH RELATIONAL INDUCTIVE BIASES,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report DEEP REINFORCEMENT LEARNING WITH RELATIONAL INDUCTIVE BIASES,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:32.886922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.517130Z digest=sha256:eb0ed738047919b555e5e64cb03d475d07642877606e4825b35ae0e0644c084e

Observation a064ee0a-beef-4b28-a12f-dfb319f94c61 · outbound

This paper cites Neuro-algorithmic Policies Enable Fast Combinatorial Generalization,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Neuro-algorithmic Policies Enable Fast Combinatorial Generalization,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:32.846276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.523146Z digest=sha256:ba604395d44c882457b9e9a305afdf8bc8c9abcb0dc97b412dcd3ad44886cbc6

Observation 0f2ba949-9d1a-40e0-a332-954b7cdf8098 · outbound

This paper cites Scalable Monotonic Neural Networks,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Scalable Monotonic Neural Networks,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:32.820684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.528670Z digest=sha256:b8791bb45bca3ca7e84713384767e1be22a1b9a00a2125c944085e756613caf4

Observation 11258812-fce2-485f-a6b8-72eb20f4bcac · outbound

This paper cites Bounded activation functions for enhanced training stability of deep neural networks on visual pattern recognition problems,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Bounded activation functions for enhanced training stability of deep neural networks on visual pattern recognition problems,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:32.783428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:41:32.534941Z digest=sha256:f83a02e03d460ccd16f737256fa550e7f13d99fc0f679ed24235ddecff1c9054

Observation f4181248-e694-4e06-a8bf-a9a0350572ff · outbound

This paper cites Certified Monotonic Neural Networks.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Certified Monotonic Neural Networks

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T18:41:32.541911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:41:32.541911Z digest=sha256:8a303405bcb32c44c348bfed39fac99a49a6463c1cf938b1e3ba62371a0aa863

Pith citing papers

Observation a26e53d4-3619-4abb-be75-4ada0bd4ff8b · inbound

Multi-Robot Multi-Queue Control via Exhaustive Assignment Actor-Critic Learning cites this paper.

Multi-Robot Multi-Queue Control via Exhaustive Assignment Actor-Critic Learning A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-13T12:50:02.212145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T12:50:02.212145Z digest=sha256:8ccfb3c6a854c137caee00aa21bfec3f41369fefdda8d47c3a7a86e439b1f45d