Pith. sign in

Paper Citation Record · LEDGER

Latent Spherical Flow Policy for Reinforcement Learning with Combinatorial Actions

As of 5 August 2026, this Paper Citation Record lists 11 of 11 outbound references and 0 inbound Pith citation observations for arXiv:2601.22211.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2601.22211 v2

Coverage vector

measured 11 of 11 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T06:53:14.015388Z

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

11 of 11 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 41be6fb3-8e07-4e82-9e14-8525a8a2aba0 · outbound

This paper cites an unresolved cited work.

Latent Spherical Flow Policy for Reinforcement Learning with Combinatorial Actions Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T06:53:13.310018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:53:13.310018Z digest=sha256:a59bfd04715bd877d34abe51c86ba498b303d2dba8e95b8010384add90dffb9f

Observation 9b73690d-616f-4ff5-b036-bf7341fa8a84 · outbound

This paper cites an unresolved cited work.

Latent Spherical Flow Policy for Reinforcement Learning with Combinatorial Actions Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T06:53:13.488156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:53:13.488156Z digest=sha256:041e9ca387dddb075b1eba786bb4a9b91b4241c716733fae6de9d8c927359351

Observation 306bfc9e-1f4b-4218-8dc0-c49469b15538 · outbound

This paper cites Poganˇci´c, M.

Latent Spherical Flow Policy for Reinforcement Learning with Combinatorial Actions Poganˇci´c, M

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T06:53:13.067294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:53:13.067294Z digest=sha256:5f592bfd5eb56a0edf7204e3f3ad9a8d3f863f2f8852af86a891bb4ca33209c3

Observation 68f61645-9c4f-47ab-8769-2e31f30e66e2 · outbound

This paper cites OnceX t(v)∈ {0,1}, it remains fixed for the rest of the episode.

Latent Spherical Flow Policy for Reinforcement Learning with Combinatorial Actions OnceX t(v)∈ {0,1}, it remains fixed for the rest of the episode

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T06:53:14.015388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:53:14.015388Z digest=sha256:9657c346c6dc6c377925ca9550c54e4808a40791660d79221e28214b18e8973e

Observation 594d51a4-cb0f-407e-8f7c-0cef19aaf14d · outbound

This paper cites Proof of Theorem 3.5.Recall the smoothed Bellman operator in Eq.

Latent Spherical Flow Policy for Reinforcement Learning with Combinatorial Actions Proof of Theorem 3.5.Recall the smoothed Bellman operator in Eq

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T06:53:13.585889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:53:13.585889Z digest=sha256:628b86bf9a8627005982494792e2acdb1e0010669c3bc4a373d3e0db90410fab

Observation 201a3b81-ccee-40e1-b59b-7d7e776ce690 · outbound

This paper cites an unresolved cited work.

Latent Spherical Flow Policy for Reinforcement Learning with Combinatorial Actions Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T06:53:13.741852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:53:13.741852Z digest=sha256:751cbf1d84cbe07a13a756490b2b3901921cf97bbb31c1b87a7e9637ac85bed8

Observation 8897b223-7fa8-47f1-83f0-bf1886e96ce4 · outbound

This paper cites an unresolved cited work.

Latent Spherical Flow Policy for Reinforcement Learning with Combinatorial Actions Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T06:53:13.792049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:53:13.792049Z digest=sha256:a42ebc13732c5fd12f97fdacf968e6bf0a573dfdb1fd4d9535f61848196902d5

Observation 3ca55a59-0eb9-4e65-a4fa-8195611bb224 · outbound

This paper cites The immediate reward at time tis Rt = X v∈At r(Yv).

Latent Spherical Flow Policy for Reinforcement Learning with Combinatorial Actions The immediate reward at time tis Rt = X v∈At r(Yv)

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T06:53:13.903083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:53:13.903083Z digest=sha256:9a931c6642e02d0b4810539460d925e653c5e751eb3c3c3a95a31294b7464225

Observation c2ba938c-553c-429c-908a-ba0c0096ddcb · outbound

This paper cites Define the strict-optimality cone Ka := c∈R m :c ⊤a<c ⊤a′ for alla ′ ∈ A \ {a}.

Latent Spherical Flow Policy for Reinforcement Learning with Combinatorial Actions Define the strict-optimality cone Ka := c∈R m :c ⊤a<c ⊤a′ for alla ′ ∈ A \ {a}

Reference 2004

Resolution
unresolved
no resolver link, observed 2026-08-03T06:53:13.224156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:53:13.224156Z digest=sha256:2f285e180f49bedd6590733cf2bc0cd70301cd15423a166dee7aaf148552b53a

Observation d07b044f-25ac-4f1c-88f2-8959263c01fc · outbound

This paper cites TorchRL: A data-driven decision-making library for PyTorch.

Latent Spherical Flow Policy for Reinforcement Learning with Combinatorial Actions TorchRL: A data-driven decision-making library for PyTorch

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-03T06:53:12.919775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:53:12.919775Z digest=sha256:4251a365b6d7da86c9515644f2dfbb13da2f88cefaee62b275df8a6bdbe5977d

Observation bba57258-5cac-4815-bcff-19ad5145b34d · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Latent Spherical Flow Policy for Reinforcement Learning with Combinatorial Actions Adam: A Method for Stochastic Optimization

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-03T06:53:12.982267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:53:12.982267Z digest=sha256:aa97eb12a9bdc023f22fa7d6e83dc7fdd4292e0ad1467dafec984bf3aa043ab0

Pith citing papers

No inbound Pith citation observations are available.