Pith. sign in

Paper Citation Record · LEDGER

Active Learning for Stochastic Contextual Linear Bandits

As of 9 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 0 inbound Pith citation observations for arXiv:2605.24803.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.24803 v1

Coverage vector

measured 19 of 19 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-30T11:49:35.327709Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

19 of 19 outbound references displayed

  • verified exact3
  • verified fuzzy11
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 75a53ea7-5779-4f74-9f99-ddf14ebf28d9 · outbound

This paper cites A contextual-bandit algorithm for mo- bile context-aware recommender system.

Active Learning for Stochastic Contextual Linear Bandits A contextual-bandit algorithm for mo- bile context-aware recommender system

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T08:36:06.948629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T11:49:35.327709Z digest=sha256:abd089d270f0438623757a81e40034ee9216b917ae9452fd751ec81f0e05a206

Observation 88780afa-3224-4f39-8a10-a179370278c3 · outbound

This paper cites A short note on learning discrete distributions.

Active Learning for Stochastic Contextual Linear Bandits A short note on learning discrete distributions

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T11:54:38.394622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T11:49:35.327709Z digest=sha256:687fd2d1ba4acea00836e2dbc1ec98de0d73cc108cf504bbbe6a55bf7e250f2b

Observation 91328711-5186-447a-aa88-0448c01b8788 · outbound

This paper cites Active Preference Optimization for Sample Efficient RLHF.

Active Learning for Stochastic Contextual Linear Bandits Active Preference Optimization for Sample Efficient RLHF

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:54:38.397551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T11:49:35.327709Z digest=sha256:cf4bbd759dfebddf0b94a05eafa532272170efcb6a84fc4d4a4a1d01b85accf1

Observation 9b2dd146-3d6a-41ea-bb85-30046c1e08c1 · outbound

This paper cites Simple Regret Minimization for Contextual Bandits.

Active Learning for Stochastic Contextual Linear Bandits Simple Regret Minimization for Contextual Bandits

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T11:54:38.400425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T11:49:35.327709Z digest=sha256:0d6046dcd457ea49bf2e82ec142031a54cc359d11dc42504b9e0e687d8c8ddb8

Observation 2eda639d-6f0c-453d-8c1f-0f79d47eaf7c · outbound

This paper cites Clarabel: An interior-point solver for conic programs with quadratic objectives.

Active Learning for Stochastic Contextual Linear Bandits Clarabel: An interior-point solver for conic programs with quadratic objectives

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T11:54:38.408760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T11:49:35.327709Z digest=sha256:3332fd117b8ea50916cd247f42d89fd7949e77a95a6354fcacd61bab69034230

Observation 3bc4a6cf-c6a8-492e-815c-42e7a19b3f1d · outbound

This paper cites A faster interior point method for semidefinite programming.

Active Learning for Stochastic Contextual Linear Bandits A faster interior point method for semidefinite programming

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T08:36:06.972406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T11:49:35.327709Z digest=sha256:3e751531e501631e02748b7a4557909d0c7c841744d83c26662428af08ab763f

Observation d1a59aeb-53ec-4b2b-883e-5e93ee7abd7b · outbound

This paper cites Near-optimal Policy Identification in Active Reinforcement Learning.

Active Learning for Stochastic Contextual Linear Bandits Near-optimal Policy Identification in Active Reinforcement Learning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:54:38.405912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T11:49:35.327709Z digest=sha256:cc8a4de2daf71672c2ec3e8bbb9e14f2b50cc6af21f7b6e7cb7c90c880aabe51

Observation 0310cb40-3e92-40ed-970d-e43a98e0add8 · outbound

This paper cites Sample-Efficient Alignment for LLMs.

Active Learning for Stochastic Contextual Linear Bandits Sample-Efficient Alignment for LLMs

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:54:38.403214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T11:49:35.327709Z digest=sha256:5b5abf9cdea5b22654ca257f5738b9c61297474d08aa02daf0fb9a0f45a4d37e

Observation a9961935-fb89-4c50-8cd1-be5f5f5a10fe · outbound

This paper cites Then, E x∼p max a∈A ∥ϕ(x, a)∥2 Σ−1 ˆw ≤4C B + 32d·tv(p,ˆp).

Active Learning for Stochastic Contextual Linear Bandits Then, E x∼p max a∈A ∥ϕ(x, a)∥2 Σ−1 ˆw ≤4C B + 32d·tv(p,ˆp)

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T08:36:06.967389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T11:49:35.327709Z digest=sha256:772c64f31f41219d020d6cc83cb02b2579e50e2dd42810b31e57d7d5bf8b6803

Observation 72ef5a53-4395-4567-a6a7-6d74c2d4cde3 · outbound

This paper cites an unresolved cited work.

Active Learning for Stochastic Contextual Linear Bandits Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-07-09T08:36:06.969750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T11:49:35.327709Z digest=sha256:d847bbb7da70ad320be0c4c1603fbefd6f8bf48a8fdaba62c5ebf1ae1e5a02de

Observation 14fd74d7-c0de-429b-8c96-a9ec422baa0e · outbound

This paper cites solution concept.

Active Learning for Stochastic Contextual Linear Bandits solution concept

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T08:36:06.974609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T11:49:35.327709Z digest=sha256:da1dddf8b30f9303b6ba925559414053d37b3049f0be52da55fdfdf4b00672f8

Observation 780dbc30-b990-405f-b292-ce5b398de7ac · outbound

This paper cites (a) Linear bandits (b) Active learning (e.g., for regression) (c) Passive context sampling for SCLBs (d) Active context sampling for SCLBs Figure 4: Four learning paradigms.

Active Learning for Stochastic Contextual Linear Bandits (a) Linear bandits (b) Active learning (e.g., for regression) (c) Passive context sampling for SCLBs (d) Active context sampling for SCLBs Figure 4: Four learning paradigms

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T08:36:06.961459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T11:49:35.327709Z digest=sha256:b53457c181f08c15cbae99574256510ced2e3940c8c744ad4433b2ae1a5c5249

Observation 482b9661-b4c4-4c34-930d-70c12aecd492 · outbound

This paper cites Thus our result is minimax optimal, up to polylog factors.

Active Learning for Stochastic Contextual Linear Bandits Thus our result is minimax optimal, up to polylog factors

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T08:36:06.958726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T11:49:35.327709Z digest=sha256:8efb777b5a94d4f302a337b6d0b70eaee94d2990da391b983b212325da7bb87b

Observation 5b975061-7109-4a51-a16d-3c791cb0f14d · outbound

This paper cites (2022b) on an instance-dependent basis.

Active Learning for Stochastic Contextual Linear Bandits (2022b) on an instance-dependent basis

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T08:36:06.963682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T11:49:35.327709Z digest=sha256:e15454f8cc71f6055081b90213cdcff0962161e49e405e34bde9e7de4cbb1c9d

Observation f63a7dbe-ff67-47cd-834d-a8a93eb148e6 · outbound

This paper cites Planner-SamplerZanette et al.

Active Learning for Stochastic Contextual Linear Bandits Planner-SamplerZanette et al

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T08:36:06.956376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T11:49:35.327709Z digest=sha256:fb0964fec42f97be74920d0e8d0ca6fec37e092e3c0e5697933eaa4a6f124f15

Observation 88ac9197-5229-44e1-8ec6-d1e5c533cc09 · outbound

This paper cites MOSEK is a software package for solving structured optimizations such as SDPs.

Active Learning for Stochastic Contextual Linear Bandits MOSEK is a software package for solving structured optimizations such as SDPs

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T08:36:06.954050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T11:49:35.327709Z digest=sha256:4003709a91d0e75cd81ad20e89693b309fd1852334ae13e5fff190ce14eaa718

Observation d5744fdf-b185-409c-99ff-eb1979deb3ce · outbound

This paper cites Baselines require significantly more samples for the average regret to match that of Active-SCLB-Empirical.

Active Learning for Stochastic Contextual Linear Bandits Baselines require significantly more samples for the average regret to match that of Active-SCLB-Empirical

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T08:36:06.951445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T11:49:35.327709Z digest=sha256:34e1b708f55b1a255434edb021b0533d15b830b3c74c1c5df10c2aed3bd740bf

Observation a9e2a916-fdae-4807-851b-12a6300eff05 · outbound

This paper cites Now, (B−1/2AB−1/2)−1 ⪯IimpliesB 1/2A−1B1/2 ⪯I.

Active Learning for Stochastic Contextual Linear Bandits Now, (B−1/2AB−1/2)−1 ⪯IimpliesB 1/2A−1B1/2 ⪯I

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T08:36:06.976810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T11:49:35.327709Z digest=sha256:212a0f4b8f93420b8e8eb18484b23bec282a6a4c6c7edc195ee1a864881f71da

Observation fd883185-8083-436f-ba6f-5949e59bfc1d · outbound

This paper cites an unresolved cited work.

Active Learning for Stochastic Contextual Linear Bandits Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-07-09T08:36:06.945747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T11:49:35.327709Z digest=sha256:eab03ee6f459c0e589be18ca37998104d6308cc1d17280b21fbbb3e528fe172d

Pith citing papers

No inbound Pith citation observations are available.