Pith. sign in

Paper Citation Record · LEDGER

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report

As of 17 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 1 inbound Pith citation observation for arXiv:2501.11136.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.11136 v1

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T18:41:32.541911Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-13T12:50:02.212145Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

40 of 40 outbound references displayed

  • verified exact3
  • verified fuzzy32
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f0b0cc1a-6804-4042-84b7-e2ff4689ea69 · outbound

This paper cites Queueing Network Controls via Deep Reinforcement Learning,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Queueing Network Controls via Deep Reinforcement Learning,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.572517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.283220Z digest=sha256:85f810e0a0e43d64b14e97272b5cc0386bd579fbd00d5357e6d8231725032cca

Observation b94b5740-a008-4a0f-931f-1fd8e85709fc · outbound

This paper cites Reinforcement learning in queues,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Reinforcement learning in queues,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.553813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.289976Z digest=sha256:842c3b72c4ad3a3e21648b411bd870691ed9dec00cd13af8ca4c5e02df64472d

Observation ad3f3fb4-a5cc-4df2-9263-a521ab76f925 · outbound

This paper cites Queue-Learning: A Reinforcement Learning Approach for Providing Quality of Service,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Queue-Learning: A Reinforcement Learning Approach for Providing Quality of Service,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.531770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.295865Z digest=sha256:91289ea00a43d1733b4ff00df43efae4d2aadb90cb80885e6a228b4f692895b9

Observation aac3882e-c12c-452f-bc1e-1d147d1c8958 · outbound

This paper cites Deep Reinforcement Learning for Smart Queue Management,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Deep Reinforcement Learning for Smart Queue Management,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.512264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.302477Z digest=sha256:20288d92147289b856bf36f2e03847fb06f8a22923c8dab9e8fdd8294e835178

Observation 752032a3-346b-446b-98a1-620a526fae69 · outbound

This paper cites Intervention-Assisted Policy Gradient Methods for Online Stochastic Queuing Network Optimization: Technical Report.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Intervention-Assisted Policy Gradient Methods for Online Stochastic Queuing Network Optimization: Technical Report

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-08-10T18:41:32.753342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.308346Z digest=sha256:70ad3204c87dc564ffcb6b93ebb4ed6df14819b4f03a79e58b942f340d3e1dde

Observation b713f99c-b954-49f6-9459-db6807bd3bcb · outbound

This paper cites Approximation theory of the MLP model in neural net- works,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Approximation theory of the MLP model in neural net- works,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.493501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.315605Z digest=sha256:45731963967b428f6a9c6e0774978fd31b1571a3a0ef19cfeee55d75b6688d69

Observation 680c0180-b7c5-49b2-9d3d-57bffb4e1966 · outbound

This paper cites Scheduling Algorithms for Minimizing Age of Information in Wireless Broadcast Networks with Random Arrivals: The No-Buffer Case.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Scheduling Algorithms for Minimizing Age of Information in Wireless Broadcast Networks with Random Arrivals: The No-Buffer Case

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-10T18:41:32.721137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.322186Z digest=sha256:75d9fd76f633a0a24ce7a15d99f4971429366fed769a616401c42fe72735dbdc

Observation d2dc76bc-445b-4889-93af-bebd3ae250b4 · outbound

This paper cites Dynamic server allocation to parallel queues with randomly varying connectivity,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Dynamic server allocation to parallel queues with randomly varying connectivity,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.474295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.330392Z digest=sha256:2266892881cbbe2ec19753720c81259a620f23cf3d31d20205c66ebc56e21406

Observation 6f57e04c-3fe5-41ed-a98b-328e3ff1ce0a · outbound

This paper cites Stability and Asymptotic Optimality of Generalized MaxWeight Policies,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Stability and Asymptotic Optimality of Generalized MaxWeight Policies,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.455153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.336474Z digest=sha256:6023d148e305888bd014f9fd1582cc18c5c5856b4baaf00756bd71b5799e394b

Observation efaf5b56-3f74-47e1-a0d8-54e099129f54 · outbound

This paper cites Minimizing the Age of Information in Wireless Networks with Stochastic Arrivals,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Minimizing the Age of Information in Wireless Networks with Stochastic Arrivals,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.421670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.342562Z digest=sha256:f43390474fc2a20eba42fcc85deddddf9b0b8deaf632f515f9aa4a9ffb2a531e

Observation 362e7a4f-5e50-42b6-8b1b-8e649316779b · outbound

This paper cites Tracking MaxWeight: Optimal Control for Partially Observable and Controllable Networks,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Tracking MaxWeight: Optimal Control for Partially Observable and Controllable Networks,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.402215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.347459Z digest=sha256:90a3a1507726fdbd1d6308a07a9314e4d729312b9775633ea83481d1c6e3cfce

Observation 4df6132d-8393-4016-9556-5c3d1f21cb20 · outbound

This paper cites MaxWeight scheduling in a generalized switch: State space collapse and workload minimization in heavy traffic,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report MaxWeight scheduling in a generalized switch: State space collapse and workload minimization in heavy traffic,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.376714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.355890Z digest=sha256:1f43e1591d278846996b7be4a4b68587c958b40bfafda882d99f268fc37b0e17

Observation 02a4ba55-1ccc-4894-890f-a8601d680622 · outbound

This paper cites An analysis of the join the shortest queue (JSQ) policy,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report An analysis of the join the shortest queue (JSQ) policy,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.353067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.361931Z digest=sha256:7a1dac767875adfc80168f0a9810e193f17660a0eba34dd6e11338be60089639

Observation d3cf1b9d-ae52-47f8-8188-6e76966038ce · outbound

This paper cites Stochastic Network Optimization with Application to Com- munication and Queueing Systems,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Stochastic Network Optimization with Application to Com- munication and Queueing Systems,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.335308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.368583Z digest=sha256:4297aecfe498c07b8da5a4bb0212f914892b92efe72f64c88bd10cb3b151cdf5

Observation ae15885e-2e2d-4f58-9c64-248e9446748c · outbound

This paper cites Restless bandits: activity allocation in a changing world,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Restless bandits: activity allocation in a changing world,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.318447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.374339Z digest=sha256:a1b4a072cdb01a490f3b4142f8ee6aab7109d8416351f2ebfd6d54ef0035106a

Observation f78fbe7c-09f5-49bb-a8c4-f7759bce5624 · outbound

This paper cites Dynamic priority allocation via restless bandit marginal productivity indices,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Dynamic priority allocation via restless bandit marginal productivity indices,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.291826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.380870Z digest=sha256:d3e46302d72448ccf6d7f3b080c4c15dcf7690f20fa11bad14c5a3e4392e3e89

Observation bff8e695-decb-4690-a684-993164296799 · outbound

This paper cites Whit- tle’s index policy for a multi-class queueing system with convex holding costs,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Whit- tle’s index policy for a multi-class queueing system with convex holding costs,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.270034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.386276Z digest=sha256:df46f377f0e6cb6db9afa299fb0a524420add9a6dbaf7e6736976f00544b8b76

Observation 98c0a9de-148f-4951-80de-a8eae26c160f · outbound

This paper cites Congestion control of TCP flows in Internet routers by means of index policy,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Congestion control of TCP flows in Internet routers by means of index policy,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.243558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.392312Z digest=sha256:d33d9bc595e4562d2655c286fe06d5e4a260cdd82f72388a1048f36f600d9786

Observation 92535104-44fc-42fc-89eb-906537382e77 · outbound

This paper cites Indexability of Restless Bandit Problems and Optimality of Whittle Index for Dynamic Multichannel Access,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Indexability of Restless Bandit Problems and Optimality of Whittle Index for Dynamic Multichannel Access,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.222731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.398235Z digest=sha256:100fc8988a6e47232afcb8d5ae78c9f46d0b0b0b4aaf82d60430709047cd9731

Observation 64fb8c22-1dd0-4df4-b24b-3572efe7d6fd · outbound

This paper cites A unifying computations of Whittle’s Index for Markovian bandits,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report A unifying computations of Whittle’s Index for Markovian bandits,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.200534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.406430Z digest=sha256:02e28dab6c60afca248ba4d0c2bf0ee8aef7050f7f7a11f5450815818bebdd05

Observation 587454cc-cc2c-4658-80d6-749ae2cc4c19 · outbound

This paper cites Optimal control of a queueing system with two heterogeneous servers,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Optimal control of a queueing system with two heterogeneous servers,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.178362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.412195Z digest=sha256:b01ae94485697abb7e9c0b63456a6d4ad803f9e43911f2d78b111cc3694ebb73

Observation 11aace9d-cd2c-4d04-8415-14af7b71fddb · outbound

This paper cites A simple proof of the optimality of a threshold policy in a two-server queueing system,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report A simple proof of the optimality of a threshold policy in a two-server queueing system,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.156219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.417178Z digest=sha256:1ba9a2f56df1773bad613d722acf2fed7e81b0fe6b9977121d6764ca09eb8b0d

Observation b6e1c794-9337-4f1d-9a79-54a444e61274 · outbound

This paper cites Extension of the optimality of the threshold policy in heterogeneous multiserver queueing systems,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Extension of the optimality of the threshold policy in heterogeneous multiserver queueing systems,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.135852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.426124Z digest=sha256:bcb6aa80ac0e8673c9ef004aef2b3c0af35e2db0c7b2b0d90948a05ef0ff1585

Observation dec16675-64b9-467a-9720-3cdf98188033 · outbound

This paper cites Monotone Control of Queueing Systems with Heterogeneous Servers,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Monotone Control of Queueing Systems with Heterogeneous Servers,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.113874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.432313Z digest=sha256:f3f234e098ecfef3edf3a57932836923b205732ca19e3eb05657db3a8c9a6896

Observation 499cb480-10d0-4586-98a0-c728e2eaa74e · outbound

This paper cites an unresolved cited work.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-10T18:41:33.091542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.438737Z digest=sha256:be3e34d73d8f36951990f2056f4a8f9d0d5788dd03906b34e67092c691b17127

Observation 18857e94-4ee9-4a5c-9845-139e65c515db · outbound

This paper cites Average Cost Optimal Stationary Policies in Infinite State Markov Decision Processes with Unbounded Costs,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Average Cost Optimal Stationary Policies in Infinite State Markov Decision Processes with Unbounded Costs,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.061073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.444936Z digest=sha256:4e92d83aae150d064fccfed0d85cc6e1b1d4dd577e4643c33861676df6aa8908

Observation 83f7999e-c1ce-40d5-a1ae-e9242cc3db28 · outbound

This paper cites an unresolved cited work.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-10T18:41:33.036842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.450780Z digest=sha256:c259d21b340b49c67fb617f7be32051d516306b7854259497f543e743845fd86

Observation 5677feb2-10cf-4ba0-b2bb-83a533fd043a · outbound

This paper cites Proximal Policy Optimization Algorithms.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Proximal Policy Optimization Algorithms

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T18:41:32.463691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:41:32.463691Z digest=sha256:b5e69d378e35e36f6bab13e487256d243f2a72c5749f9ccd378b7cb267cddaf0

Observation 38b5460f-305d-45e2-b98e-55da1756330e · outbound

This paper cites A Dissection of Overfitting and Generalization in Continuous Reinforcement Learning,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report A Dissection of Overfitting and Generalization in Continuous Reinforcement Learning,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:33.010974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.474879Z digest=sha256:c07688bb05a4c23ad6917a32e4dd49ede5b2b2607fdcc4254420a40982928589

Observation 0a4b2f77-ec85-43d2-abc2-bcbda093953c · outbound

This paper cites A Survey of Zero-shot Generalisation in Deep Reinforcement Learning.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report A Survey of Zero-shot Generalisation in Deep Reinforcement Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T18:41:32.481954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:41:32.481954Z digest=sha256:1960b0778faa343b15c743c76d047c79c0827a07f24d07251f8690b99d6477fd

Observation 4cca26f9-0d7e-4721-af68-d70354ad00ad · outbound

This paper cites Quantifying Generalization in Reinforcement Learning,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Quantifying Generalization in Reinforcement Learning,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:32.990143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.488187Z digest=sha256:5a0f7441355848f928f13f1de3d5e8c79b5af09de088f25f3f9b8a3990bfec08

Observation 96078e02-3631-4aa5-a1fd-5e1ffc46856f · outbound

This paper cites Neuroevolution of self-interpretable agents,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Neuroevolution of self-interpretable agents,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:32.961108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.493985Z digest=sha256:e2b7a41a9e51235242ca9c80d096db2ebd84cea279090191bb11fbc3795a1c72

Observation f2f9bf7f-e24e-4ebb-bbb8-d062c8f74d25 · outbound

This paper cites The Sensory Neuron as a Transformer: Permutation-Invariant Neural Networks for Reinforcement Learning.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report The Sensory Neuron as a Transformer: Permutation-Invariant Neural Networks for Reinforcement Learning

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-08-10T18:41:32.629516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.499838Z digest=sha256:2068ff10a702663b1b871520a9885fd58937963de3bd4f60457d6e4d3b7f16f8

Observation 57b98c88-9c07-4770-988c-72c2c38c560c · outbound

This paper cites Unsupervised Visual Attention and Invariance for Reinforcement Learning,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Unsupervised Visual Attention and Invariance for Reinforcement Learning,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:32.937935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.505623Z digest=sha256:c616dcfb58edf5e6368841c2db6742ae1a24c6b91c58a8a32712aae566dd98d9

Observation 9896f5d2-ea73-42d5-b17c-9008958f0025 · outbound

This paper cites Deep reinforcement learning with relational inductive biases,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Deep reinforcement learning with relational inductive biases,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:32.914959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.511121Z digest=sha256:6ebd4a853df7d7b73996e2eb0c48c91d1e2c8f15522e01734335ba43c6515bf9

Observation 7e9539ea-2e76-4d5d-9e87-359d74ef06a9 · outbound

This paper cites DEEP REINFORCEMENT LEARNING WITH RELATIONAL INDUCTIVE BIASES,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report DEEP REINFORCEMENT LEARNING WITH RELATIONAL INDUCTIVE BIASES,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:32.886922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.517130Z digest=sha256:d67b018a6c89566c4d2f6c03110d67672aa1c87e02404b362e4ae2fb6a870fa7

Observation a064ee0a-beef-4b28-a12f-dfb319f94c61 · outbound

This paper cites Neuro-algorithmic Policies Enable Fast Combinatorial Generalization,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Neuro-algorithmic Policies Enable Fast Combinatorial Generalization,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:32.846276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.523146Z digest=sha256:6390734b0949a7092e6bfc6a12a103faa55e39e0a832ceabb72da77c2f707e02

Observation 0f2ba949-9d1a-40e0-a332-954b7cdf8098 · outbound

This paper cites Scalable Monotonic Neural Networks,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Scalable Monotonic Neural Networks,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:32.820684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.528670Z digest=sha256:d618865f0be3ea281adfaaabbdf3aafaf81a2d14e94613cec64c8d94068e3a92

Observation 11258812-fce2-485f-a6b8-72eb20f4bcac · outbound

This paper cites Bounded activation functions for enhanced training stability of deep neural networks on visual pattern recognition problems,.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Bounded activation functions for enhanced training stability of deep neural networks on visual pattern recognition problems,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T18:41:32.783428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T18:41:32.534941Z digest=sha256:323f9b0c288d20e8831b9446fbd4d59230fd4a5f60197ad34a4c96525b145c10

Observation f4181248-e694-4e06-a8bf-a9a0350572ff · outbound

This paper cites Certified Monotonic Neural Networks.

A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report Certified Monotonic Neural Networks

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T18:41:32.541911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:41:32.541911Z digest=sha256:1289652ca24cd988d18f094cc700ef2a8ac9a05cdf21361f0cc5930ed30479f5

Pith citing papers

Observation a26e53d4-3619-4abb-be75-4ada0bd4ff8b · inbound

Multi-Robot Multi-Queue Control via Exhaustive Assignment Actor-Critic Learning cites this paper.

Multi-Robot Multi-Queue Control via Exhaustive Assignment Actor-Critic Learning A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-13T12:50:02.212145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T12:50:02.212145Z digest=sha256:bbd044cf7a4444afacc5c0247575851036adc189191e04d265f068ee1a101e9f