Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-20T15:11:27.420642Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2605.17458.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-20T15:11:27.420642Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
61 of 61 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a9fedcd5-885f-4b89-9401-fb3949f8fc33 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks ACM Transactions on Intelligent Systems and Technology (TIST) , volume=
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5c22b0ff-da08-49cf-ba32-4440e293b21b · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Text Classification via Large Language Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 1ec1cfed-debd-4058-9314-ec0097dafa5f · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c24bfd67-cff3-4acc-ada6-3799022325d8 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 80637b89-26ae-434c-8530-9107c17bd162 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Advances in neural information processing systems , volume=
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a02fdab5-897a-4da5-ac31-337df667b012 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Advances in neural information processing systems , volume=
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 9b13c925-b33e-4547-870a-7ac46b865433 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks RLHF Workflow: From Reward Modeling to Online RLHF
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d7288a0e-8e87-45c6-9e52-3a4533071513 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks arXiv preprint arXiv:2505.23349 , year=
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 10ea430b-06fd-4acc-bec5-679ffd246169 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 012cbd38-7fc4-4258-b46a-b3141717f4c1 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks RoBERTa: A Robustly Optimized BERT Pretraining Approach
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 9d765d9c-add7-4e97-aa74-572e5aa3775e · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Qwen3 Technical Report
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d10e3172-8a52-4dfc-a9f6-d7ec72f70bca · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Proceedings of the IEEE international conference on computer vision , pages=
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 624bd122-18b4-4191-accd-761bd2a20e88 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Advances in neural information processing systems , volume=
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f895def9-9d77-415d-80ac-53c514cd8ebe · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Advances in neural information processing systems , volume=
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c365df9d-b4b1-4af0-8059-92d5f2883edd · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks IEEE Transactions on Neural Networks and Learning Systems , volume=
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 9402ece9-f7a3-4897-962c-1270472b8a92 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Conference on robot learning , pages=
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation fa5a7ed4-633c-4ef9-b3b5-6aacde0ecc8f · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Advances in Neural Information Processing Systems , volume=
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d16bec14-85bc-43f0-b0c9-76cd313e0301 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 7971edb9-0539-41e6-901f-60b34b10ef65 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks IEEE Transactions on Instrumentation and Measurement , year=
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 31b96833-b075-4a68-bbbb-d445b86f4158 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Fine-Tuning Language Models from Human Preferences
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation bc06323c-f7c4-4934-8bd5-ded631f70959 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks 2025 IEEE/ACM International Workshop on Deep Learning for Testing and Testing for Deep Learning (DeepTest) , pages=
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation fe6cba06-e4ff-4dc1-8acc-ad59bddcf975 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Preference learning , pages=
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 90b08339-ce5d-4041-b8c6-325a449b5aa5 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks 2021 IEEE International Conference on Big Data (Big Data) , pages=
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 64a83656-d067-46e3-b9ac-56a5d0edae43 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks International Conference on Machine Learning , pages=
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation bf36eb17-74c9-4736-a62f-643be0d2bfd3 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Proceedings of the 2013 conference on empirical methods in natural language processing , pages=
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a4536758-faa6-4355-8e41-137def431398 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks A Survey on Progress in LLM Alignment from the Perspective of Reward Design
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 093818c7-712a-4e6c-b676-5150f60c4e46 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Meta-radiology , volume=
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 7f6d0971-12b5-448b-882a-d9e4bac8ea56 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Transactions of the Association for Computational Linguistics , volume=
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 1d2de6e3-9ac4-4b9e-b162-ae723f6a3d70 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Proceedings of the third international workshop on paraphrasing (IWP2005) , year=
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 43fe05e4-725e-493c-bc31-52d9006a94e7 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Advances in neural information processing systems , volume=
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 8ae1e34f-41fa-47ca-a4a4-d277e250d077 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Proceedings of the 2018 conference on empirical methods in natural language processing , pages=
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 65443843-4f7c-43ca-8f83-4642aadd71d3 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Advances in neural information processing systems , volume=
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5af75729-1ec7-4afb-9170-96e70514c5e3 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks 2020 IEEE 27th International Conference on Software Analysis, Evolution and Reengineering (SANER) , pages=
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 1becff8f-9dd0-4366-bdd2-412d320427cb · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 09f4800a-8d6d-4d61-b5aa-ac4ac1bdcfee · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks OPT: Open Pre-trained Transformer Language Models
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 8552a3f6-10a2-4eba-9e8d-713e246efb5b · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Journal of machine learning research , volume=
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 1b2b73dc-e2ef-40c1-b990-5f5e56378f59 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks CodeBERT: A Pre-Trained Model for Programming and Natural Languages
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation cbcef752-5ebc-4f9a-bbf7-ceb32ddc8bc9 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks CodeT5: Identifier-aware Unified Pre-trained Encoder-Decoder Models for Code Understanding and Generation
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d877232b-af3a-468a-97b2-e813e785673e · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Proceedings of the 2023 conference on empirical methods in natural language processing , pages=
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 949f82bf-56e4-41c5-a6dd-3050f64bf2f3 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks CodeGen: An Open Large Language Model for Code with Multi-Turn Program Synthesis
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 911035af-2bd1-44bb-a982-09a79cb2a1b2 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Proceedings of the 2020 conference on empirical methods in natural language processing: system demonstrations , pages=
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 80146d0c-a01c-4b6b-8e2a-33ba52af6694 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Ieee Access , volume=
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation edc67d5b-6660-483a-9e06-eb396f2cadb6 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks International conference on machine learning , pages=
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ae3863ea-f21e-4080-99cf-78a60a083ad6 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Proceedings of the 2018 EMNLP workshop BlackboxNLP: Analyzing and interpreting neural networks for NLP , pages=
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 086c5b8f-e20f-407d-bd4a-585640bbf872 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and Generation
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 14972fec-6d40-49f4-9dc5-469deeed9fd0 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Decoupled Weight Decay Regularization
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation eb812d02-efee-4a3d-928b-54e85711dd3f · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 439ca1c9-db11-4adb-a0b6-399285b01c57 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing , pages=
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 226ff9cb-35bc-40a2-8ca3-e8184ca1845d · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks ACM Transactions on Knowledge Discovery from Data (TKDD) , volume=
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a8a188e3-d691-4f47-8715-774229786dfd · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks ACM Computing Surveys , issn =
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 48e48c96-9ed2-4443-8ce8-cb46c82a46ab · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks A Stability Analysis of Fine-Tuning a Pre-Trained Model
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 134094df-a128-4114-b62e-dc477c9eadb9 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Open Problems and Fundamental Limitations of Reinforcement Learning from Human Feedback
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e1955c49-22bf-4dad-b21b-a970a5d56f05 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks AI Alignment through Reinforcement Learning from Human Feedback? Contradictions and Limitations
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b619a5bc-a6e5-418b-9dd0-8f52e20e6a53 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Proximal Policy Optimization Algorithms
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0ba35602-f6db-424b-a9f4-63b7946581f0 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Aho and Jeffrey D
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 39022c8a-ea3b-4154-a72b-caf0e52c7221 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Unresolved cited work
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 37d126ce-6dd4-48f2-9816-4b45fe5571c3 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Chandra and Dexter C
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 04c7e943-35dc-4c70-ae39-6ea91f313273 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Scalable training of
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4eb78d1e-7e24-4906-b26d-99c9464038be · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Unresolved cited work
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a6ce0d35-20f3-46d3-9883-703869d1f4e0 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Tetreault , title =
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 147ec98a-d4d5-49ea-b7ce-960332c5df45 · outbound
ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks A Framework for Learning Predictive Structures from Multiple Tasks and Unlabeled Data , Volume =
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
No inbound Pith citation observations are available.