Pith. sign in

Paper Citation Record · LEDGER

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models

As of 7 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 0 inbound Pith citation observations for arXiv:2506.02726.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02726 v1

Coverage vector

measured 56 of 56 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:21:43.077882Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

56 of 56 outbound references displayed

  • verified exact3
  • verified fuzzy16
  • unresolved36
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 35378862-893b-4597-b95b-05288e3f5877 · outbound

This paper cites Large language models: A survey of their development, capabilities, and applications.Knowledge and Information Systems, pages 1–56, 2024.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Large language models: A survey of their development, capabilities, and applications.Knowledge and Information Systems, pages 1–56, 2024

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:47.731516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:21:38.818556Z digest=sha256:d6b2fb6bf468f07b101e9eb0eb6324b47adb9e693a8a5ce0bdf566cbb9cba2b2

Observation d7bbc86f-e931-45c9-8953-57ef7582e0ec · outbound

This paper cites Nlp for social good: A survey of challenges, opportunities, and responsible deployment.arXiv preprint arXiv:2505.22327, 2025.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Nlp for social good: A survey of challenges, opportunities, and responsible deployment.arXiv preprint arXiv:2505.22327, 2025

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:38.860198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:38.860198Z digest=sha256:cb75440e10f8e52ac0e7bc103784394a6052674b01cc34bbaadfa036f5018f30

Observation bb19e456-97f4-405b-aae1-7bae18f537a2 · outbound

This paper cites Current applications and challenges in large language models for patient care: a systematic review.Communications Medicine, 5(1):26, 2025.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Current applications and challenges in large language models for patient care: a systematic review.Communications Medicine, 5(1):26, 2025

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:47.478778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:21:38.949163Z digest=sha256:6aef74ad8109629e3a6408b389c70b49a1b0a9971c4b30a388c491dc32f178d0

Observation 57b394cb-018e-46af-be82-05bfc3318205 · outbound

This paper cites A Survey on Large Language Models with some Insights on their Capabilities and Limitations.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models A Survey on Large Language Models with some Insights on their Capabilities and Limitations

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:39.055557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:39.055557Z digest=sha256:3cd4c18e255db2a3d2ed4a43f45f76a86616ba776def4d2ab7f173e15226a267

Observation 44ab70e2-1c85-4c56-a568-e9ce0a935689 · outbound

This paper cites Direct preference optimization: Your language model is secretly a reward model.Advances in Neural Information Processing Systems, 36:53728–53741, 2023.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Direct preference optimization: Your language model is secretly a reward model.Advances in Neural Information Processing Systems, 36:53728–53741, 2023

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:39.094280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:39.094280Z digest=sha256:1a99a35b0605120b8ccd6d723dc4e20d31118c501b6ab93a2c1183f4753ea514

Observation 74c35b06-33bc-4769-a41d-ee2412b33a4e · outbound

This paper cites Reinforcement Learning Enhanced LLMs: A Survey.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Reinforcement Learning Enhanced LLMs: A Survey

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:39.153862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:39.153862Z digest=sha256:402c7126289f87a13408a901da6be32552adbb8a09b581d46cdf476f426fe003

Observation a381681a-455a-4ea2-9fb1-fabb69e82004 · outbound

This paper cites LLMs for Explainable AI: A Comprehensive Survey.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models LLMs for Explainable AI: A Comprehensive Survey

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:39.242160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:39.242160Z digest=sha256:50e85d647e283fb5671826e079767b2374caa110a56226c1fc8a8e45037dd0f0

Observation ba7746db-7ab2-49c0-9c0d-d2fe3e7f91eb · outbound

This paper cites Explainability for large language models: A survey.ACM Transactions on Intelligent Systems and Technology, 15(2):1–38, 2024.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Explainability for large language models: A survey.ACM Transactions on Intelligent Systems and Technology, 15(2):1–38, 2024

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:47.331458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:21:39.392138Z digest=sha256:9d60fe7d561eeeab067b7038c6d02bd63cc1c8ffed274518913f30f210853960

Observation ab4988a7-bbac-49a4-8d43-ae72784b60c7 · outbound

This paper cites Understand what llm needs: Dual preference alignment for retrieval-augmented generation.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Understand what llm needs: Dual preference alignment for retrieval-augmented generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:39.447414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:39.447414Z digest=sha256:9c0b99937417e0aa532ee358102f61d874541a304bb75815581d0a2231d09297

Observation 4f2b9230-9f4f-41fa-b65c-682f64177050 · outbound

This paper cites Chain of preference optimization: Improving chain-of-thought reasoning in llms.Advances in Neural Information Processing Systems, 37:333–356, 2024.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Chain of preference optimization: Improving chain-of-thought reasoning in llms.Advances in Neural Information Processing Systems, 37:333–356, 2024

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:39.545398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:39.545398Z digest=sha256:feadb6fc365efb5c050ef37f93a011f9c74e967e25f6b018c0a86266e7a0d3d1

Observation 5faf9fb3-54ce-4de8-99ab-631bf7637757 · outbound

This paper cites Context-DPO: Aligning Language Models for Context-Faithfulness.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Context-DPO: Aligning Language Models for Context-Faithfulness

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:39.613828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:39.613828Z digest=sha256:507112ea4fda78814f6abb5203f226cbbbc81090527a7f80eb396690d58f237a

Observation bb9969f1-74a8-4879-9709-705b98e488d6 · outbound

This paper cites PA-RAG: RAG Alignment via Multi-Perspective Preference Optimization.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models PA-RAG: RAG Alignment via Multi-Perspective Preference Optimization

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:21:44.178173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:21:39.749604Z digest=sha256:876bdd5a1221e14ddeea9b3897cad5eabcd67000b2bf3128570be4a2ae7da33e

Observation 26010888-1271-453c-9889-bb799a52a1a3 · outbound

This paper cites Knowpo: Knowledge-aware preference optimization for controllable knowledge selection in retrieval- augmented language models.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Knowpo: Knowledge-aware preference optimization for controllable knowledge selection in retrieval- augmented language models

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:47.090260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:21:39.821944Z digest=sha256:f14f251f35311b463ad2e0be15a417ae02a0e911fae266846fc4159171c1f2e5

Observation 9d28fa35-cf80-4ed3-9bc1-053bd929396d · outbound

This paper cites Iterative reasoning preference optimization.Advances in Neural Information Processing Systems, 37:116617–116637, 2024.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Iterative reasoning preference optimization.Advances in Neural Information Processing Systems, 37:116617–116637, 2024

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:39.897385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:39.897385Z digest=sha256:b524662f88b75e569c10de2e2887d84eaa5126a311b87d2a9b4992485151968b

Observation 903b4620-bbfa-4740-b989-c9aa6a4f3c2d · outbound

This paper cites Self-Training with Direct Preference Optimization Improves Chain-of-Thought Reasoning.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Self-Training with Direct Preference Optimization Improves Chain-of-Thought Reasoning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:40.015790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:40.015790Z digest=sha256:0c9f7df5d1e3ed26ab532fecd187c75eb183b7e30db2d642daa7148b36b827d4

Observation a6cab298-93e4-432d-a3cd-31d582c665e6 · outbound

This paper cites Solving Maxwell's Equations.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Solving Maxwell's Equations

Reference 16

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T11:21:44.000302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:21:40.090748Z digest=sha256:e8bb31ec0f9ffc90d960072f9f71ed82d70317bbfcc67f7b1aaef04b66939123

Observation a9a06361-e0ee-4ac3-abe2-d72c8f9d7519 · outbound

This paper cites A Survey on Knowledge-Oriented Retrieval-Augmented Generation.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models A Survey on Knowledge-Oriented Retrieval-Augmented Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:40.158290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:40.158290Z digest=sha256:c310abc006264a985c75cd7ff0e52b2e7c0924a4627a37b0ba9c0a7494e5ebf3

Observation 23df82cc-59c2-4a25-bebe-450afeb4def6 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.Advances in neural information processing systems, 35:24824–24837, 2022.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Chain-of-thought prompting elicits reasoning in large language models.Advances in neural information processing systems, 35:24824–24837, 2022

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:40.284051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:40.284051Z digest=sha256:05b0faeef752d215dc776a0e04b7f9dc2705572ac51c4b6ff94d3f09d66d638c

Observation fa508398-e9cd-4894-bbe0-055d2e765f73 · outbound

This paper cites A Survey on Post-training of Large Language Models.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models A Survey on Post-training of Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:40.383980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:40.383980Z digest=sha256:1f9454541e10e2d7b1b2e60aded31c36aef3713d12e0803c02915483477f6b74

Observation fdf48b54-8869-47c6-bf9a-03f922f803a9 · outbound

This paper cites A Survey on Personalized Alignment -- The Missing Piece for Large Language Models in Real-World Applications.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models A Survey on Personalized Alignment -- The Missing Piece for Large Language Models in Real-World Applications

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:40.479121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:40.479121Z digest=sha256:a18951703fd27759bab709cd153c0791d3571a9c429bc689fb5815776bce8578

Observation 9b0ec985-fb2e-429e-bbeb-c23793eb36c8 · outbound

This paper cites MM-RLHF: The Next Step Forward in Multimodal LLM Alignment.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models MM-RLHF: The Next Step Forward in Multimodal LLM Alignment

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:40.615790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:40.615790Z digest=sha256:0aed72096d56f5cbf0476fc6a61c2324e8edcd8a985d3f53c500a11f5870dafa

Observation 0204eac1-f9af-4dc2-881c-4b43e76c53fe · outbound

This paper cites Asynchronous RLHF: Faster and More Efficient Off-Policy RL for Language Models.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Asynchronous RLHF: Faster and More Efficient Off-Policy RL for Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:40.684605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:40.684605Z digest=sha256:6872d4a00c5eb6d1f97d8e874a7103cb8f1a8f71ab6b31d921ef7978ee67f3c0

Observation 0434a918-f900-4b85-946d-c03d54180b4b · outbound

This paper cites RLHS: Mitigating Misalignment in RLHF with Hindsight Simulation.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models RLHS: Mitigating Misalignment in RLHF with Hindsight Simulation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:40.801007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:40.801007Z digest=sha256:5f3fca6a1dab6bc0ffe4a70cbc23408ff7c4a18ea0c49f4ea5ae09d8635cdce7

Observation 82e0992e-7473-4faf-b2b3-07506c9ccda4 · outbound

This paper cites REINFORCE++: Stabilizing Critic-Free Policy Optimization with Global Advantage Normalization.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models REINFORCE++: Stabilizing Critic-Free Policy Optimization with Global Advantage Normalization

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:40.927940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:40.927940Z digest=sha256:e34b4694d00af8b83a794b849fae5fd861fdbb1ea6d2800cc427f93442ef79a8

Observation 12271706-1c52-45c2-9b33-393f1be2538e · outbound

This paper cites Exploratory Preference Optimization: Harnessing Implicit Q*-Approximation for Sample-Efficient RLHF.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Exploratory Preference Optimization: Harnessing Implicit Q*-Approximation for Sample-Efficient RLHF

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:41.016055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:41.016055Z digest=sha256:dba8cd4540f345b055da05066effed55b6c94d4736706a49e6b13977cb582432

Observation fad7e606-d01f-4894-84a3-b0e44b075c91 · outbound

This paper cites an unresolved cited work.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:21:46.854469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:21:41.058198Z digest=sha256:7965286102bb309ad85e0a39cd805b972d2f6e5e5809faeb941a524b8af0dfb5

Observation bbe81107-5f30-4627-ad40-7fa6c65256cc · outbound

This paper cites Safer-Instruct: Aligning Language Models with Automated Preference Data.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Safer-Instruct: Aligning Language Models with Automated Preference Data

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:41.130792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:41.130792Z digest=sha256:6779925df2f35dbd2b64ebf4436ca043441501e813f8298edd664250588c3fd4

Observation 710738de-a91c-4664-8a43-71f3c1dc08e5 · outbound

This paper cites Self-Boosting Large Language Models with Synthetic Preference Data.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Self-Boosting Large Language Models with Synthetic Preference Data

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:41.210733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:41.210733Z digest=sha256:4fa6b66dea5ed8766d2eae39e50df3b4d444a062ae6791771527208e40fcd963

Observation a28d5641-84b1-4e89-a0e7-5e53d514116e · outbound

This paper cites A comprehensive review of large language models: issues and solutions in learning environments.Discover Sustainability, 6(1):27, 2025.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models A comprehensive review of large language models: issues and solutions in learning environments.Discover Sustainability, 6(1):27, 2025

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:46.524500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:21:41.269690Z digest=sha256:f75dc06356ab120a48218d521ad3ba3b088479ac1c9666d175cf722ee4570a03

Observation 148f7fa5-3df2-48f8-93f9-b611bc3a5f75 · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Constitutional AI: Harmlessness from AI Feedback

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:41.322841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:41.322841Z digest=sha256:5dee9ccfee66f341b2d270e08ac9c378b5f28409df4f991a7b8186b663352f50

Observation cdf321d2-69d2-4864-8466-81151cfdede6 · outbound

This paper cites Zhang and J.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Zhang and J

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:46.337621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:21:41.393033Z digest=sha256:848adb62fabc0bbef84da1b37e120393c57d233a47b309c671caf9c549f4be37

Observation 3535564f-345b-4253-9118-ffd667784e46 · outbound

This paper cites Retrieval-Augmented Generation for Large Language Models: A Survey.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Retrieval-Augmented Generation for Large Language Models: A Survey

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:41.484804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:41.484804Z digest=sha256:fd5ffa43caf9986e7f4ce2cab34b7a2fee226bbb7e7866f3029b1739e390f533

Observation f5a0a599-c7ff-48f5-93b9-bc6aa98de06c · outbound

This paper cites an unresolved cited work.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:21:46.013978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:21:41.550399Z digest=sha256:720f4ea759ac218528b8240ca2c5667f54772efb387b45f6b4286c53a59faefc

Observation 4a09417b-8f90-4f45-8b86-de503559f55b · outbound

This paper cites Wang and Y.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Wang and Y

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:45.870818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:21:41.601069Z digest=sha256:4dedef702cd11e9836f1e8a34d324754f466710ad8cd4614bfc0149f2d08e37d

Observation a92a8db8-8f5d-4ba3-9e81-588b0fa43f84 · outbound

This paper cites Feng and L.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Feng and L

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:45.725340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:21:41.653470Z digest=sha256:56e6f797c737e292f19bf652af5c700ba38876bf66d42c8654c49f7bee117fcc

Observation aedea759-c1d4-422f-bc74-d0649ca28f24 · outbound

This paper cites Retrieval-Augmented Generation with Graphs (GraphRAG).

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Retrieval-Augmented Generation with Graphs (GraphRAG)

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:41.722040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:41.722040Z digest=sha256:c24d9fa74cdcdb34e3642bf3e592061b673f5655704f167030aa733690cb9558

Observation a8369b8b-1fca-444f-92c3-4f022cc3f304 · outbound

This paper cites Synthetic Data Generation & Multi-Step RL for Reasoning & Tool Use.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Synthetic Data Generation & Multi-Step RL for Reasoning & Tool Use

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:41.804315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:41.804315Z digest=sha256:8506a53381b9671d6069acdf9521aeea1ed2f19e316d8fe4e79fca8321311124

Observation a465f5ba-829c-4790-8c3e-56fc59df58af · outbound

This paper cites Enhancing chain of thought prompting in large language models via reasoning patterns.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Enhancing chain of thought prompting in large language models via reasoning patterns

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:45.550033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:21:41.874814Z digest=sha256:57467a46c54445423ec37880a0838f5b99bbe6cd45c61ad82ac2da1e4d41be4a

Observation 09bcebd1-1f60-478e-90b7-79b246d1c87f · outbound

This paper cites Transformers Provably Solve Parity Efficiently with Chain of Thought.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Transformers Provably Solve Parity Efficiently with Chain of Thought

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:41.931755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:41.931755Z digest=sha256:920583224f5f80104077653657b1e0e71cf00d5863b5a18e3f57451f31a694dc

Observation 43ec6d05-cebe-4146-a919-e602bba05855 · outbound

This paper cites Chain of Draft: Thinking Faster by Writing Less.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Chain of Draft: Thinking Faster by Writing Less

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:41.999082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:41.999082Z digest=sha256:e5c67879402a7131797385130201af7d13f5381d1996843996a177bd6fc2e00f

Observation 4f1fe121-7f15-46fa-906c-5d85df9b3985 · outbound

This paper cites Tool learning with large language models: A survey.Frontiers of Computer Science, 19(8):198343, 2025.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Tool learning with large language models: A survey.Frontiers of Computer Science, 19(8):198343, 2025

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:42.046502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:42.046502Z digest=sha256:fd563335e303d75b04f9e14be58bd1278fde58fc2c33e09afc30ea8fcfca2c51

Observation b3030e9e-68e9-480a-9bf1-70470f0c9d15 · outbound

This paper cites Making Large Language Models Better Reasoners with Alignment.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Making Large Language Models Better Reasoners with Alignment

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:42.079243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:42.079243Z digest=sha256:15d57d72ba908c19f86004f3c1d24d38b91d131a4dd64859ad424f50da8fa50f

Observation e9913d1b-478a-4d3d-b1c3-a4700efdfbb5 · outbound

This paper cites PORT: Preference Optimization on Reasoning Traces.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models PORT: Preference Optimization on Reasoning Traces

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:21:43.674287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:21:42.154602Z digest=sha256:f23f6661328e7fb35a80c2441d24ec75aee6b6e51a47e2ad64a275da974c7877

Observation e709ee68-ea43-4f99-907c-ab3a2df57a1d · outbound

This paper cites Beyond Chain-of-Thought: A Survey of Chain-of-X Paradigms for LLMs.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Beyond Chain-of-Thought: A Survey of Chain-of-X Paradigms for LLMs

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:42.204519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:42.204519Z digest=sha256:1aa6d46cf95515caba2ef9fd5aa834d7f308b651414839bf36abe04e47f560aa

Observation 22ae8ed5-37c1-4c8a-95fe-75f8c167d131 · outbound

This paper cites Preference tree optimization: Enhancing goal- oriented dialogue with look-ahead simulations.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Preference tree optimization: Enhancing goal- oriented dialogue with look-ahead simulations

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:45.372913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:21:42.271188Z digest=sha256:17f4a5e4ab43ba1e57c8317dc33d8192e54a91045fdb49d9d645a5b962651be6

Observation f98649cf-13d2-4964-9fa3-e4f9334cfae4 · outbound

This paper cites Large language models in traditional chinese medicine: A scoping review.Journal of Evidence-Based Medicine, 18(1):e12658, 2025.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Large language models in traditional chinese medicine: A scoping review.Journal of Evidence-Based Medicine, 18(1):e12658, 2025

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:45.203522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:21:42.326058Z digest=sha256:456115d208c9a1ec97dc192621d46b33c7eb2ab3ff41899ddeed8b2c1801527d

Observation 06d4c3d7-50f9-45b8-aae0-cfddaa4ce002 · outbound

This paper cites Tcmchat: A generative large language model for traditional chinese medicine.Pharmacological Research, 210:107530, 2024.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Tcmchat: A generative large language model for traditional chinese medicine.Pharmacological Research, 210:107530, 2024

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:45.046715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:21:42.411645Z digest=sha256:4e584ca09f01ac6332a57366e3b6c1ed39843cdd55d97271332aff247a8744e4

Observation e99d14b3-39d3-4263-8b86-4882bb6c85ab · outbound

This paper cites Biancang: A traditional chinese medicine large language model.arXiv preprint arXiv:2411.11027, 2024.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Biancang: A traditional chinese medicine large language model.arXiv preprint arXiv:2411.11027, 2024

Reference 48

Resolution
verified exact
raw_fallback, observed 2026-08-07T11:21:43.446236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:21:42.480389Z digest=sha256:d9202a207de9d880eecab244f454ee1d264725ecf9a64f8d6afcadd29bb95655

Observation 17d4fd0d-80d0-4afe-b56d-8bcd6eb95937 · outbound

This paper cites Qibo: A Large Language Model for Traditional Chinese Medicine.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Qibo: A Large Language Model for Traditional Chinese Medicine

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:42.548438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:42.548438Z digest=sha256:78f77013c1fd4ef8d0af0a824810ff2209f7345c91116505f37cd23d11d18dee

Observation 2ea30630-77fc-4016-a820-90eff300a36a · outbound

This paper cites Ai-powered lawyering: Ai reasoning models, retrieval augmented generation, and the future of legal practice.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Ai-powered lawyering: Ai reasoning models, retrieval augmented generation, and the future of legal practice

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:44.913465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:21:42.619817Z digest=sha256:60cb83ef7d6b94b1ebfeb75510fa848c5126e7c61c021bde940d2ab1c74b41f4

Observation 645540cd-f7b2-4593-be27-28e96ffafad7 · outbound

This paper cites A comprehensive review on financial explainable ai.Artificial Intelligence Review, 58(6):1–49, 2025.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models A comprehensive review on financial explainable ai.Artificial Intelligence Review, 58(6):1–49, 2025

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:44.785971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:21:42.699536Z digest=sha256:2444e8a2d754c120125f44a28a12bc229499c20ad3eaf1971059e47eeefc78fe

Observation 0595f1a3-41d5-4292-913c-e2cba6ba271f · outbound

This paper cites Findings of the association for computational linguistics: Eacl 2024.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Findings of the association for computational linguistics: Eacl 2024

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:44.608922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:21:42.746394Z digest=sha256:5acac80e860ece7528d0a97e4928aa11c9b540003c0e73d307601190187aa790

Observation 673dff5e-42c1-4845-97dc-2fade4c69f74 · outbound

This paper cites ALFA: Aligning LLMs to Ask Good Questions A Case Study in Clinical Reasoning.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models ALFA: Aligning LLMs to Ask Good Questions A Case Study in Clinical Reasoning

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:42.818394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:42.818394Z digest=sha256:91c6c35dba1a63a73a00be5291ba02d5698ee3cd78bb8bb51a78924ff489f1f6

Observation 6dfbdf6e-383e-4290-b84e-a2609ada25ee · outbound

This paper cites Shennong-tcm: A traditional chinese medicine large language model.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Shennong-tcm: A traditional chinese medicine large language model

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:44.433484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:21:42.914094Z digest=sha256:d7c4403cb3b689fb76815b1460dedca09e680a0d195f9a5674aea5a3c8d253c2

Observation aada8a85-a2e7-412f-bb92-483f6a257300 · outbound

This paper cites Rouge: A package for automatic evaluation of summaries.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Rouge: A package for automatic evaluation of summaries

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:42.986811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:42.986811Z digest=sha256:5b16e1c3574747fa77985eaf197f20cbea31c3209d280434c2da613bc28a0454

Observation 5f72547f-2c79-4385-95ec-873c5c2c23f2 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Bleu: a method for automatic evaluation of machine translation

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:43.077882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:43.077882Z digest=sha256:2d7c9912df4b48edfd367d0699c49fbded16e78bab2a130d3d1849132e4def96

Pith citing papers

No inbound Pith citation observations are available.