Pith. sign in

Paper Citation Record · LEDGER

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks

As of 13 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2605.17458.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.17458 v1

Coverage vector

measured 61 of 61 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-20T15:11:27.420642Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

61 of 61 outbound references displayed

  • verified exact5
  • verified fuzzy38
  • unresolved3
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch15

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a9fedcd5-885f-4b89-9401-fb3949f8fc33 · outbound

This paper cites ACM Transactions on Intelligent Systems and Technology (TIST) , volume=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks ACM Transactions on Intelligent Systems and Technology (TIST) , volume=

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.587288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:6bef39e28b8e9fcdc0d1b011fcf4648d4cc0668e5a3f500085ae0de01bb19d7e

Observation 5c22b0ff-da08-49cf-ba32-4440e293b21b · outbound

This paper cites Text Classification via Large Language Models.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Text Classification via Large Language Models

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T15:13:24.951372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:289173bcceec57dd8a5dd471144372d94f1964cfde5e0e90ae88139199df5cec

Observation 1ec1cfed-debd-4058-9314-ec0097dafa5f · outbound

This paper cites Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.589496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:df74c772cc2f989a6ff66315129eb60edfb010206a11148daf185d24adee43d8

Observation c24bfd67-cff3-4acc-ada6-3799022325d8 · outbound

This paper cites Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.685015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:cfcd45490ad7be3e340459cbf4b778feb656c0e5925c85c286dff472c729826c

Observation 80637b89-26ae-434c-8530-9107c17bd162 · outbound

This paper cites Advances in neural information processing systems , volume=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Advances in neural information processing systems , volume=

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.680903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:b834874d206db245700c05e9773b52d2539214384467ecfcf4aea498f1bed362

Observation a02fdab5-897a-4da5-ac31-337df667b012 · outbound

This paper cites Advances in neural information processing systems , volume=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Advances in neural information processing systems , volume=

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.682953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:610683c3a4fbb91d393ea0a7570561ba6a5a7c5239a362daf134d09a30fa0ad6

Observation 9b13c925-b33e-4547-870a-7ac46b865433 · outbound

This paper cites RLHF Workflow: From Reward Modeling to Online RLHF.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks RLHF Workflow: From Reward Modeling to Online RLHF

Reference 7

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T15:13:24.960452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:30ec9d21a54be6e6bf42eea7546d4f991359b2bb28146a00e9830eed96e2be1b

Observation d7288a0e-8e87-45c6-9e52-3a4533071513 · outbound

This paper cites arXiv preprint arXiv:2505.23349 , year=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks arXiv preprint arXiv:2505.23349 , year=

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-20T15:13:25.003198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:4820d1017b008dfb75226ea79f6b8d34cb2da4ef5c3bd9acf153188f00493a57

Observation 10ea430b-06fd-4acc-bec5-679ffd246169 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T15:13:25.006116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:45928fd0e2c3d56e6616b50312252f2221977570a8f6e7328a29956cc1b73dc1

Observation 012cbd38-7fc4-4258-b46a-b3141717f4c1 · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T15:13:25.000060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:f1a473d458d992b16038724b84ae76e51a97d16fc1a4ab74e3ffaff7ffdf93a8

Observation 9d765d9c-add7-4e97-aa74-572e5aa3775e · outbound

This paper cites Qwen3 Technical Report.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Qwen3 Technical Report

Reference 11

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T15:13:24.969131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:ffb519bacca6b489e0c854e85ac220beda7e111fd97f98338ed6d7e23d8a5ab5

Observation d10e3172-8a52-4dfc-a9f6-d7ec72f70bca · outbound

This paper cites Proceedings of the IEEE international conference on computer vision , pages=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Proceedings of the IEEE international conference on computer vision , pages=

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.671948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:f905e78a9e1f75bf276e2cd00ab3c6eafc802a3a852f30dc0ff8df862fefe759

Observation 624bd122-18b4-4191-accd-761bd2a20e88 · outbound

This paper cites Advances in neural information processing systems , volume=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Advances in neural information processing systems , volume=

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.674054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:78499187751dabb0c6fe7e83d509b94462525352035e51c85b25697a213023f3

Observation f895def9-9d77-415d-80ac-53c514cd8ebe · outbound

This paper cites Advances in neural information processing systems , volume=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Advances in neural information processing systems , volume=

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.664877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:12ced9c0b1ad7f8482ab8276e8c749f9bc8085d3248eff096dfdaa8180d986c3

Observation c365df9d-b4b1-4af0-8059-92d5f2883edd · outbound

This paper cites IEEE Transactions on Neural Networks and Learning Systems , volume=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks IEEE Transactions on Neural Networks and Learning Systems , volume=

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.667050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:a1d83a9654c8253c7bce867abc72b4e0d481e5af82dfe0712ded94c4a6e04a61

Observation 9402ece9-f7a3-4897-962c-1270472b8a92 · outbound

This paper cites Conference on robot learning , pages=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Conference on robot learning , pages=

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.662408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:85712c62f1974ce79e1fa4bab5cb5a0f4a0a498f36f6274699655cb09cb90c57

Observation fa5a7ed4-633c-4ef9-b3b5-6aacde0ecc8f · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Advances in Neural Information Processing Systems , volume=

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.669556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:7398c5b44795d7e8253b97a87553db5af6085f45bc94f329a990c7337c9c69e3

Observation d16bec14-85bc-43f0-b0c9-76cd313e0301 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.676202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:d91c517be145d76f662553c4ed6de5805ef9848754a9fdc944363844bdeeecfb

Observation 7971edb9-0539-41e6-901f-60b34b10ef65 · outbound

This paper cites IEEE Transactions on Instrumentation and Measurement , year=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks IEEE Transactions on Instrumentation and Measurement , year=

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.646925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:886d886dd6d1982d595de82b65c9877496d6f57ce88bb3904a390002d4ed73f0

Observation 31b96833-b075-4a68-bbbb-d445b86f4158 · outbound

This paper cites Fine-Tuning Language Models from Human Preferences.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Fine-Tuning Language Models from Human Preferences

Reference 20

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T15:13:24.978579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:13ea84178b8b39078d65b3f5930ee211fd7b1b4abefd8263ff9bf456ff3e33eb

Observation bc06323c-f7c4-4934-8bd5-ded631f70959 · outbound

This paper cites 2025 IEEE/ACM International Workshop on Deep Learning for Testing and Testing for Deep Learning (DeepTest) , pages=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks 2025 IEEE/ACM International Workshop on Deep Learning for Testing and Testing for Deep Learning (DeepTest) , pages=

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.649461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:5eef280e002c3bef22dd54a0e4d51d8bbd4fef3b0c3c75b68bcda6526df1f833

Observation fe6cba06-e4ff-4dc1-8acc-ad59bddcf975 · outbound

This paper cites Preference learning , pages=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Preference learning , pages=

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.642691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:60b2e322c66759d45a9295b001b1997bc8c5cce7464b7ab444f636d6d3812f69

Observation 90b08339-ce5d-4041-b8c6-325a449b5aa5 · outbound

This paper cites 2021 IEEE International Conference on Big Data (Big Data) , pages=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks 2021 IEEE International Conference on Big Data (Big Data) , pages=

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.640446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:61c0d687d39ca17cebb571538d438e132b7bb2acc0bb9470f9164a298a8eb4d4

Observation 64a83656-d067-46e3-b9ac-56a5d0edae43 · outbound

This paper cites International Conference on Machine Learning , pages=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks International Conference on Machine Learning , pages=

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.644666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:049c7707a679a263281062c1bde2e22fa09a345963e2d78f12d848b5cb4a934a

Observation bf36eb17-74c9-4736-a62f-643be0d2bfd3 · outbound

This paper cites Proceedings of the 2013 conference on empirical methods in natural language processing , pages=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Proceedings of the 2013 conference on empirical methods in natural language processing , pages=

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.652023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:b5ac35b5ce8297f92581ad466432c56853f11d1fd556589de5ea0992a9b8f294

Observation a4536758-faa6-4355-8e41-137def431398 · outbound

This paper cites A Survey on Progress in LLM Alignment from the Perspective of Reward Design.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks A Survey on Progress in LLM Alignment from the Perspective of Reward Design

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-20T15:13:24.997088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:c26e6a5c5cb8f01ef3c54e013982f707b4788175cd91c7e44e23b9c3c647b910

Observation 093818c7-712a-4e6c-b676-5150f60c4e46 · outbound

This paper cites Meta-radiology , volume=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Meta-radiology , volume=

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.633663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:0699eaa271eb507970f6be7c63a400f7967b432cd78406679127c4b79f9e3589

Observation 7f6d0971-12b5-448b-882a-d9e4bac8ea56 · outbound

This paper cites Transactions of the Association for Computational Linguistics , volume=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Transactions of the Association for Computational Linguistics , volume=

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.635906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:82cd9380d31eaff4841e25d398184e911706027c3bab7c1e14119ec6345d4271

Observation 1d2de6e3-9ac4-4b9e-b162-ae723f6a3d70 · outbound

This paper cites Proceedings of the third international workshop on paraphrasing (IWP2005) , year=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Proceedings of the third international workshop on paraphrasing (IWP2005) , year=

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.629419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:a569f53ee42ba0aed9903556e7462fd125a7a7648de99ad7bb884217330a7170

Observation 43fe05e4-725e-493c-bc31-52d9006a94e7 · outbound

This paper cites Advances in neural information processing systems , volume=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Advances in neural information processing systems , volume=

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.627089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:89aac1af1d4643d188104d673199d4e7e7907c4746f783fb62f9fb120b380434

Observation 8ae1e34f-41fa-47ca-a4a4-d277e250d077 · outbound

This paper cites Proceedings of the 2018 conference on empirical methods in natural language processing , pages=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Proceedings of the 2018 conference on empirical methods in natural language processing , pages=

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.631613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:87e44b2416ea2b8d9c9ac5359ffdd7cc6db9d30a070eb70952c22724a434194d

Observation 65443843-4f7c-43ca-8f83-4642aadd71d3 · outbound

This paper cites Advances in neural information processing systems , volume=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Advances in neural information processing systems , volume=

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.638324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:b11dda5d1217063240d9f27a1375c471556df86f2010760c30aaf4e3b5892e91

Observation 5af75729-1ec7-4afb-9170-96e70514c5e3 · outbound

This paper cites 2020 IEEE 27th International Conference on Software Analysis, Evolution and Reengineering (SANER) , pages=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks 2020 IEEE 27th International Conference on Software Analysis, Evolution and Reengineering (SANER) , pages=

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.659691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:e7166bafb8098d4637352cf7e51065c940d10e07efe1a7792b4ec8c487e3253d

Observation 1becff8f-9dd0-4366-bdd2-412d320427cb · outbound

This paper cites an unresolved cited work.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-05-20T15:13:25.678577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:375ce9b5da96cd25a92ed2a9f6e7f6e6d803167497ee785e522f30e7936fd5d8

Observation 09f4800a-8d6d-4d61-b5aa-ac4ac1bdcfee · outbound

This paper cites OPT: Open Pre-trained Transformer Language Models.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks OPT: Open Pre-trained Transformer Language Models

Reference 35

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T15:13:24.984212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:6c0bcf55bcdaf32922a1a952ef92740bfaef086c22bd435bd36bcdeded3ff0ea

Observation 8552a3f6-10a2-4eba-9e8d-713e246efb5b · outbound

This paper cites Journal of machine learning research , volume=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Journal of machine learning research , volume=

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.622406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:36d04f5a61a947cd8fcaff751f6f4cd150d8b4b7c80a5e54d4977d3ce7eb6775

Observation 1b2b73dc-e2ef-40c1-b990-5f5e56378f59 · outbound

This paper cites CodeBERT: A Pre-Trained Model for Programming and Natural Languages.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks CodeBERT: A Pre-Trained Model for Programming and Natural Languages

Reference 37

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T15:13:24.993587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:38aa997d5eeec430843006b65a34a42d5f5dbc7609db795df087d9cb8d613e1d

Observation cbcef752-5ebc-4f9a-bbf7-ceb32ddc8bc9 · outbound

This paper cites CodeT5: Identifier-aware Unified Pre-trained Encoder-Decoder Models for Code Understanding and Generation.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks CodeT5: Identifier-aware Unified Pre-trained Encoder-Decoder Models for Code Understanding and Generation

Reference 38

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T15:13:24.987411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:9b10ee9037c2e09df9ce7601c0e648d5b31ea87baf75a61fc77d85c3d33435d0

Observation d877232b-af3a-468a-97b2-e813e785673e · outbound

This paper cites Proceedings of the 2023 conference on empirical methods in natural language processing , pages=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Proceedings of the 2023 conference on empirical methods in natural language processing , pages=

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.619962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:05927a96dd85fed3552deebebea0bff6eb41560095c14ccb61a7395bc4019636

Observation 949f82bf-56e4-41c5-a6dd-3050f64bf2f3 · outbound

This paper cites CodeGen: An Open Large Language Model for Code with Multi-Turn Program Synthesis.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks CodeGen: An Open Large Language Model for Code with Multi-Turn Program Synthesis

Reference 40

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T15:13:24.954832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:87228f10068c8af6ffc45a5f55dcd53023b02eaaf3404ef66bc902a9d4a47cf8

Observation 911035af-2bd1-44bb-a982-09a79cb2a1b2 · outbound

This paper cites Proceedings of the 2020 conference on empirical methods in natural language processing: system demonstrations , pages=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Proceedings of the 2020 conference on empirical methods in natural language processing: system demonstrations , pages=

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.615597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:19f61cbe38b636e6f7a1e7cbdad4eda75915d9b425f3c17c4f567bd4a7510001

Observation 80146d0c-a01c-4b6b-8e2a-33ba52af6694 · outbound

This paper cites Ieee Access , volume=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Ieee Access , volume=

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.613200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:cc287a5ef98096ec2b66802211f3920c6e4e0a8ae2cbbe27182113075a57d39a

Observation edc67d5b-6660-483a-9e06-eb396f2cadb6 · outbound

This paper cites International conference on machine learning , pages=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks International conference on machine learning , pages=

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.617932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:46f557b4b92672e690d8a34e149182a4a775c2b43f69e0992ba22bd310044681

Observation ae3863ea-f21e-4080-99cf-78a60a083ad6 · outbound

This paper cites Proceedings of the 2018 EMNLP workshop BlackboxNLP: Analyzing and interpreting neural networks for NLP , pages=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Proceedings of the 2018 EMNLP workshop BlackboxNLP: Analyzing and interpreting neural networks for NLP , pages=

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.610522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:b28ab4d3547b9cb17fe88c0d1cc82f8c4f1bc8881aa8ce4fb473622fe2f310cd

Observation 086c5b8f-e20f-407d-bd4a-585640bbf872 · outbound

This paper cites CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and Generation.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and Generation

Reference 45

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T15:13:24.957683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:6d55912cdfb8bbdbe7e218b3704693dbc4b22d60e00db3a78408044669addca6

Observation 14972fec-6d40-49f4-9dc5-469deeed9fd0 · outbound

This paper cites Decoupled Weight Decay Regularization.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Decoupled Weight Decay Regularization

Reference 46

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T15:13:24.990296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:b3fb598833d07e753bd56f04ddaeb93165934ec820ed88d6b97f8843b72276e9

Observation eb812d02-efee-4a3d-928b-54e85711dd3f · outbound

This paper cites SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training

Reference 47

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T15:13:24.972089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:b2d954b0e16b4c05dd09f4265d634c36069a5cd358c61027afd2d97788436601

Observation 439ca1c9-db11-4adb-a0b6-399285b01c57 · outbound

This paper cites Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing , pages=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing , pages=

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.624573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:67507f690c034d59f203e60995b5c3f4af01f6bd35e51691b4012c3d346b5e7c

Observation 226ff9cb-35bc-40a2-8ca3-e8184ca1845d · outbound

This paper cites ACM Transactions on Knowledge Discovery from Data (TKDD) , volume=.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks ACM Transactions on Knowledge Discovery from Data (TKDD) , volume=

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.605408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:bd42efd7a4415b3c3a34048512f5d65fa4992b318f47246b5bfa1e33605bda52

Observation a8a188e3-d691-4f47-8715-774229786dfd · outbound

This paper cites ACM Computing Surveys , issn =.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks ACM Computing Surveys , issn =

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.603135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:c66e54ab9fd1d0a84aaef168b226931bf78fe21008cb62f91cbbd7509cc15919

Observation 48e48c96-9ed2-4443-8ce8-cb46c82a46ab · outbound

This paper cites A Stability Analysis of Fine-Tuning a Pre-Trained Model.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks A Stability Analysis of Fine-Tuning a Pre-Trained Model

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-20T15:13:24.963500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:bd5694e919ee2e5559d2da21a2d1d4b59ebfae7a17088243a5b850728db374c1

Observation 134094df-a128-4114-b62e-dc477c9eadb9 · outbound

This paper cites Open Problems and Fundamental Limitations of Reinforcement Learning from Human Feedback.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Open Problems and Fundamental Limitations of Reinforcement Learning from Human Feedback

Reference 52

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T15:13:24.966375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:6fae0c3395e777bf22f92c81a846b393016c4ac142503b0d548dd81d537443ad

Observation e1955c49-22bf-4dad-b21b-a970a5d56f05 · outbound

This paper cites AI Alignment through Reinforcement Learning from Human Feedback? Contradictions and Limitations.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks AI Alignment through Reinforcement Learning from Human Feedback? Contradictions and Limitations

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-20T15:13:24.975593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:d9751ca6184a90cf586c2d94857122898bdf13abc3a3639f52a75bc53e7cea14

Observation b619a5bc-a6e5-418b-9dd0-8f52e20e6a53 · outbound

This paper cites Proximal Policy Optimization Algorithms.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Proximal Policy Optimization Algorithms

Reference 54

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T15:13:24.981176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:f8eaae8ff79ab04ca34bb09514aa1fca1b33fc2a76e6c0120de86464b7c9e721

Observation 0ba35602-f6db-424b-a9f4-63b7946581f0 · outbound

This paper cites Aho and Jeffrey D.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Aho and Jeffrey D

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.607999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:6d6fabe421101f3ef01d3658b69bd3bea5705f2075c34a752e0f658e1a90c5de

Observation 39022c8a-ea3b-4154-a72b-caf0e52c7221 · outbound

This paper cites an unresolved cited work.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-05-20T15:13:25.598694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:327b694e8b7d13efed1be8d60a70c93a1f9fbcbfd30ee1ab1755f6c5f2013419

Observation 37d126ce-6dd4-48f2-9816-4b45fe5571c3 · outbound

This paper cites Chandra and Dexter C.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Chandra and Dexter C

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-20T15:13:24.751418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:47b15a59bb027ccf5c573d626158a888d0672544de490f33ba39438f6943c1f8

Observation 04c7e943-35dc-4c70-ae39-6ea91f313273 · outbound

This paper cites Scalable training of.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Scalable training of

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.601090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:707bfae07e5851e1af95eab6217f9eb4f44300ab26cafff031a7239bd969c3f3

Observation 4eb78d1e-7e24-4906-b26d-99c9464038be · outbound

This paper cites an unresolved cited work.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-05-20T15:13:25.593945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:4a2d1bb942615d0fac657be9aad8532f8b561dd7205675bc14f6a47f74f555c1

Observation a6ce0d35-20f3-46d3-9883-703869d1f4e0 · outbound

This paper cites Tetreault , title =.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks Tetreault , title =

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.591601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:2f286027d6f27e1621eec70642d8134fe93a0a2e3a3bd52dae8018598ca919b7

Observation 147ec98a-d4d5-49ea-b7ce-960332c5df45 · outbound

This paper cites A Framework for Learning Predictive Structures from Multiple Tasks and Unlabeled Data , Volume =.

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks A Framework for Learning Predictive Structures from Multiple Tasks and Unlabeled Data , Volume =

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T15:13:25.596444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T15:11:27.420642Z digest=sha256:c0b4c9c60150fce13d81a5d512e7bdbcf421e9bdae717ccb9657da7a98d7d8cd

Pith citing papers

No inbound Pith citation observations are available.