Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:50:45.634287Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2505.23673.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:50:45.634287Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
61 of 61 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation cbb2eea2-eefa-4813-a6ce-abe16523ff2b · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f25d9bbb-539f-4ffa-a943-5380cdd1e2ee · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Online learning for linearly parametrized control problems
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f63f9162-3596-4bc2-b4e7-b14c5fb6ed7d · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Reducing dueling bandits to cardinal bandits
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c916829b-9808-4ebb-a0ce-b4bda5cf7a20 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds S., Hu, W., Li, Z., Salakhutdinov, R
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 661e7698-f079-457d-91a0-dcbd759a411c · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds R., Daulton, S., Letham, B., Wilson, A
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3c8de591-d8e2-458b-9040-58d0cb36bd85 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Preference-based online learning with dueling bandits: A survey
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3b08b2bb-24a7-42e8-af6a-df979b24cc66 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Stochastic contextual dueling bandits under linear stochastic transitivity models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c5f9d829-ad58-45fd-b147-a251fe5e750a · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Mat \'e rn gaussian processes on riemannian manifolds
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1cbd7aa1-5dfe-4a1d-9f66-859d3c207769 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Unresolved cited work
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1b04aac-25ff-468c-a4f1-279ce8a31e28 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds A Tutorial on Bayesian Optimization of Expensive Cost Functions, with Application to Active User Modeling and Hierarchical Reinforcement Learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb9496c5-f672-4075-a016-e37658ba9289 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Instructzero: Efficient instruction optimization for black-box large language models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 72f068b0-e951-4dec-aabc-ef40b7e314d5 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Human-in-the-loop: Provably efficient preference-based reinforcement learning with general function approximation
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd4f5627-73e6-42a9-9a40-4be5c0d094e4 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3a2067e6-4446-4226-8812-ca1d198dd0cc · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds and Steinwart, I
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bf1ae0a8-8f2e-4f09-9963-fb340e6a3851 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6dd47c75-eda3-44a6-93c1-9020fbd86d56 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds E., Slivkins, A., and Zoghi, M
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 279f94d4-b194-4864-b290-d74173a76ca7 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 90cc3a77-2405-4f42-bf4c-0d4f49d85df7 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Improved optimistic algorithms for logistic bandits
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3fd1dca9-b811-4db7-8ba6-edaa73054cf9 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds A Tutorial on Bayesian Optimization
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18fce1e8-13d2-40ae-9586-1d9ec0b24e46 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds R., Pleiss, G., Bindel, D., Weinberger, K
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b5164a3c-46a1-4518-9082-846da787c71c · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1cc4fddb-d77a-422f-bbb7-62bba9674992 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds L., and Thomaz, A
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f82de9f4-ae72-47bb-a88c-9b04475e04bf · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds and Yang, X.-S
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dad41594-2365-4926-9a1e-5a488c4bb929 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds R., Schonlau, M., and Welch, W
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 807202dd-e4de-4525-b1ca-f3f9eeb24320 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Feel-good thompson sampling for contextual dueling bandits
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 70eed8b9-bb8e-4b6d-b810-cfc21206ee71 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds and Scarlett, J
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6fa7c559-fe37-473b-916b-3147b64858b7 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c52c829c-9edc-4326-9bd9-17ec706ccd74 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3e05f0f4-440a-4c96-815a-b223275df31c · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Sample Efficient Preference Alignment in LLMs via Active Exploration
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05359119-0546-450b-85ca-779f84c2d013 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Kernelized offline contextual dueling bandits
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5a862fc4-f9e7-4c39-a3ff-b6c87c6ebf2d · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Functions of positive and negative type, and their connection with the theory of integral equations
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1ce9876e-4bd5-4bb8-86be-301ac54aa385 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Projective preferential bayesian optimization
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 849b6473-3f9d-4ee8-883c-119846cf8701 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Dueling posterior sampling for preference-based reinforcement learning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8646b46f-3fbe-4102-89a5-ea926abb3064 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Training language models to follow instructions with human feedback
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f7d64a3-69ac-49b3-a6b5-ab2d59de08c1 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Bandits with preference feedback: A stackelberg game perspective
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation caa9fe46-d605-42c5-bc2a-bcebf3c6eac5 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Scikit-learn: Machine learning in P ython
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 166d52dd-32ae-4492-b81d-5df8ea1d1138 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Optimal algorithms for stochastic contextual preference bandits
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7d52213e-1a68-489b-9e55-ee72fe4bb353 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds and Krishnamurthy, A
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 889bb251-af6e-403a-9647-950bd1ee4e92 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Dueling rl: Reinforcement learning with trajectory preferences
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1c144d77-19f3-46d8-b247-76c1035bb965 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds A domain-shrinking based bayesian optimization algorithm with order-optimal regret performance
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c2c9467f-1cca-4af2-bb75-07d3faab1062 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Lower bounds on regret for noisy gaussian process bandit optimization
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b123e8fe-ebfc-4bcf-940c-41a2afa6489c · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Unresolved cited work
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3758a7e0-2e92-465b-bbf7-cdeb81572b1c · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds P., and De Freitas, N
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23b44e2a-b881-4f5f-981d-8fbcc8ea02c5 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds M., and Seeger, M
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e8d65b47-20a1-4c62-9db7-57bebf343634 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Towards practical preferential bayesian optimization with skew gaussian processes
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a985fd11-9b1a-4b58-a2eb-372ec3ca4b69 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Unresolved cited work
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 08434648-ed7d-40f3-bf5d-6b6810b15521 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Optimal order simple regret for gaussian process bandits
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b9381b72-7c92-4d46-9a8c-8c9188bc71d4 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds On information gain and regret bounds in gaussian process bandits
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4d154dbc-42cf-4bea-8214-ebb4695eacbe · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Finite-time analysis of kernelised contextual bandits
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b91b39b6-17ce-4ffc-a485-0509836099e9 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Unresolved cited work
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4fd5c7cf-d33e-4877-9957-135aea590e2a · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds High-dimensional probability: An introduction with applications in data science, volume 47
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25dd1c45-fc65-4544-b0e1-c48f8c1ee64d · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Unresolved cited work
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6e71dfec-1ae2-40e0-934d-9d0c8eee30b4 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds and Sun, W
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 92e305af-7512-4a3b-85d2-659758cb44ae · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Principled preferential bayesian optimization
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bb3d5c7e-ab48-4452-b719-d08bd5b063cb · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Zeroth order non-convex optimization with dueling-choice bandits
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f840f36d-5b10-4be3-8f71-ee7fb1ae6126 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Preference-based reinforcement learning with finite-time guarantees
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9f09a41f-a9d6-47d5-91eb-6c09adb7e6d7 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds and Joachims, T
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f90a452d-7f1b-4882-ade5-a0cca8fe9794 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds The k-armed dueling bandits problem
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 084eb858-f460-4be7-b7e6-405729726a10 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds D., and Sun, W
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9d3dc523-5eee-4577-b38a-cb2a093b2b88 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Relative upper confidence bound for the k-armed dueling bandit problem
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 243d5c50-0041-43b5-9901-9c4cb9c5cdb0 · outbound
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds Mergerucb: A method for large-scale online ranker evaluation
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
No inbound Pith citation observations are available.