Pith. sign in

Paper Citation Record · LEDGER

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization

As of 7 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 0 inbound Pith citation observations for arXiv:2605.11974.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.11974 v1

Coverage vector

measured 59 of 59 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-13T07:32:58.404947Z

measured 59 of 59 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

59 of 59 outbound references displayed

  • verified exact12
  • verified fuzzy28
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch19

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0362c18f-461b-491d-9018-31192185d143 · outbound

This paper cites GPT-4o System Card.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization GPT-4o System Card

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:37:29.974011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:602288695489e1027c926ecfdf485cd931f426dc1da259559bbec07da8f91f0f

Observation 06e54a4e-ca37-4498-9d21-88d01532a308 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:37:29.963624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:2d59f88d6de493202adf471823b6d33bd59c074fcd5e07f98dcfc45184a867a0

Observation ae8ccd63-c0bc-4c16-9804-15bab25078da · outbound

This paper cites Qwen3 Technical Report.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Qwen3 Technical Report

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:37:29.968523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:6dfc7f49e5720f319be3682f80e02caafc398767f57afb41e86943bd777750c9

Observation d1d65bf8-1161-4fc0-8ce0-56e0b2fb919f · outbound

This paper cites Proceedings of the 47th international ACM SIGIR conference on research and development in information retrieval , pages=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Proceedings of the 47th international ACM SIGIR conference on research and development in information retrieval , pages=

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.323535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:5be689d306bea335b5ea28d6e4e132465bc53d3749041e0ed0016f7c21c7b801

Observation eee5ce83-11cb-4c95-b8dd-35a5018dc367 · outbound

This paper cites Proceedings of the AAAI conference on artificial intelligence , pages=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Proceedings of the AAAI conference on artificial intelligence , pages=

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.335650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:fba843fd08e04253777ddf09e0fae79fd7a6004b63779280b97c0b5286d57d32

Observation 959e19b4-e66d-497e-93b9-591aa06dbd47 · outbound

This paper cites Proceedings of the 2025 ACM Conference on Fairness, Accountability, and Transparency , pages=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Proceedings of the 2025 ACM Conference on Fairness, Accountability, and Transparency , pages=

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.327706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:fa4bcb90efcf87356e8e34af48fc89ab95758d51c1a3ab8ce2205d3a0d6e713b

Observation fafc6f86-061f-4ac5-b3ff-97c189fdd996 · outbound

This paper cites arXiv preprint arXiv:2506.17188 , year=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization arXiv preprint arXiv:2506.17188 , year=

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:37:29.957813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:17fd084e4418683d78a2fa2e2ddc66ff4923fd7dceb28aa6959c9088502e7650

Observation 2f94ca3c-12f3-4bfd-a427-0c5ca5b2ec5b · outbound

This paper cites Proceedings of the 17th ACM International Conference on Web Search and Data Mining , pages=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Proceedings of the 17th ACM International Conference on Web Search and Data Mining , pages=

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.331615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:2930661500a30fc630872581df914a2dbd8be4d829fe3e1df3ba2d670d0bb15d

Observation 2a949fc7-ad61-42f2-bd1f-849f64e4e844 · outbound

This paper cites Domaino1s: Guiding LLM Reasoning for Explainable Answers in High-Stakes Domains.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Domaino1s: Guiding LLM Reasoning for Explainable Answers in High-Stakes Domains

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:37:29.985874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:cf6db0c1489d5168dbbbddf5ffeedea233a16215c68b6c44f4db1fa60e3dba38

Observation 7c2004d1-4194-45b3-8fb0-c54057773778 · outbound

This paper cites Annual Meeting of the Association for Computational Linguistics , year=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Annual Meeting of the Association for Computational Linguistics , year=

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.318946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:6127c20ff6655b231a9db2053e45508c9a9481814a03078f808224828beb481c

Observation 0185bf1e-103b-4537-89a5-767e8531119e · outbound

This paper cites Transactions of the Association for Computational Linguistics , year=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Transactions of the Association for Computational Linguistics , year=

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.312155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:41858be1055bd3513b00a8e5879c5c3f6fe00df8e1433366353ad5eb2a462f69

Observation 1ab875a2-f7f7-4d64-8f5e-8f021281e532 · outbound

This paper cites ArXiv , year=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization ArXiv , year=

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.315375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:5c4bb90ab0285792ee207630e502506916e6596a107ac56776c8ed1a600c2361

Observation 8ecd1e62-f088-4899-9c2d-8180c9e85dd6 · outbound

This paper cites ArXiv , year=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization ArXiv , year=

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.303993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:f63e7ea3d7162e4f5dc03b097c268b0c86167093acf2d2aaba929a2e63cd0f7d

Observation 95920722-c4a1-4df3-b8d9-5399024d2278 · outbound

This paper cites Order Matters: Exploring Order Sensitivity in Multimodal Large Language Models.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Order Matters: Exploring Order Sensitivity in Multimodal Large Language Models

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:37:29.979718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:a4702eb15373f7d252fdbef0279ce49584b9d7a091394f039eb11357238cad5c

Observation 04f98399-2412-448f-8d27-a2456fa272a5 · outbound

This paper cites ArXiv , year=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization ArXiv , year=

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.308442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:14930b45fd8effee2149f68156c3ae1c11ef4758089201f524a0c6df97737d64

Observation c027ac73-8602-49dd-8ff5-a588af6921a9 · outbound

This paper cites An Information-theoretic Approach to Prompt Engineering Without Ground Truth Labels.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization An Information-theoretic Approach to Prompt Engineering Without Ground Truth Labels

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:37:29.991486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:bd5c40e3ca838e7f7e5246fff34453e848eee6e447bbb4419db8fc7b1d617b37

Observation 35476a24-3d20-4d64-8c2f-f1ee1f11bbf5 · outbound

This paper cites Annual Meeting of the Association for Computational Linguistics , year=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Annual Meeting of the Association for Computational Linguistics , year=

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.241263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:149dd592b350dc64098522872d055b7764d7ff1e5f6a8eee7fcec8e993af51d9

Observation 1c3fe5e4-05d7-40a0-b807-d2bb03a72065 · outbound

This paper cites ArXiv , year=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization ArXiv , year=

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.270262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:4b8fd807e043f3ccc0fbdb55aec5ba4c61574deeb876f2adfc791390a50c4252

Observation 23cfb539-1254-4c46-b3bb-7c27938673dc · outbound

This paper cites ArXiv , year=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization ArXiv , year=

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.257913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:c48fe717706c6f19f1d6ddabbfc773af2622ae3fc7d3f148399c9b2398b64af5

Observation 1cd9193f-6189-4db2-909a-a3b1fd9fbf99 · outbound

This paper cites 2024 , note =.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization 2024 , note =

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.253722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:b58f08ec01a75396dfd1e087dce29d3fb2c6a5c1b64f10ecc830531366915052

Observation f091f349-633b-4d69-96d5-c6ec09f55695 · outbound

This paper cites ArXiv , year=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization ArXiv , year=

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.227942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:8449ff6203259a709ab3634b6c75a45288d8f30afc33a3bae04650f21e679a57

Observation d7b8a243-be82-4b97-b640-ee808be6c8bc · outbound

This paper cites Findings of the Association for Computational Linguistics: EMNLP 2024 , pages=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Findings of the Association for Computational Linguistics: EMNLP 2024 , pages=

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.223339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:2ac8585c81deb220b369f39fc13f572878f271d3b4d6e390edc2be5570700c57

Observation 39277bbd-3b76-4088-baf2-022cbeaede7b · outbound

This paper cites Proceedings of the AAAI Symposium Series , number=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Proceedings of the AAAI Symposium Series , number=

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.274001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:77c0bcde2ebbf2ddcba01a6160f66b9ba57e946c1016d80578c0a0cc7a15748b

Observation 4c646333-71cf-4841-a174-ab96ca8692e3 · outbound

This paper cites Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing , pages=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing , pages=

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.262203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:2b1cb1fc800801e0f6f2fc5fc2e5eca83d9bc88179abfb271f461a2ec55075a5

Observation fefbe4fd-a195-47d1-b00c-f0533076146c · outbound

This paper cites Mitigate Position Bias in Large Language Models via Scaling a Single Dimension.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Mitigate Position Bias in Large Language Models via Scaling a Single Dimension

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:37:29.913435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:bc5d58059a426ac5879e3f7d32838d05281aabb53f6555ebe00b9dcba9b5eafd

Observation 6aa3cdd2-2321-4ec5-a7cf-ae64d63c28de · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Advances in Neural Information Processing Systems , volume=

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.291111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:667ece4a9092cfe5a43d4c420ae10f1581ac668b85e9a28eec831782d136bb1d

Observation eee701a0-3d6c-46f1-8b6d-884e3077443f · outbound

This paper cites Unveiling Selection Biases: Exploring Order and Token Sensitivity in Large Language Models.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Unveiling Selection Biases: Exploring Order and Token Sensitivity in Large Language Models

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:37:29.831887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:8d1731bbeae7e644f4b6160e6b5db189dd0b4fcc75534e550855507aca8edc09

Observation 07c141d9-e3fb-4af2-8bdf-d341136cebe4 · outbound

This paper cites 2024 IEEE International Conference on Web Services (ICWS) , pages=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization 2024 IEEE International Conference on Web Services (ICWS) , pages=

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.287213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:e742f3752c9ce5b0c2b0612ea51e960785479944e21884c7f80a852101b2a2bc

Observation f37073eb-81db-4968-a4f2-ef937aa54969 · outbound

This paper cites GraphSOS: Graph Sampling and Order Selection to Help LLMs Understand Graphs Better.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization GraphSOS: Graph Sampling and Order Selection to Help LLMs Understand Graphs Better

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:37:29.818585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:2831aedcfa078f10338df1fec634a889e65aa288d597b3c1eafce4df041814e0

Observation cb156072-3568-49d0-abda-e6c0d8b6697e · outbound

This paper cites Set-LLM: A Permutation-Invariant LLM.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Set-LLM: A Permutation-Invariant LLM

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:37:29.928879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:8d275732c2011c85739557fac47fa52d603b732106346dee24b9efd2c6dec8df

Observation a18cb4a0-36a1-4e26-aac4-0a47b6e80192 · outbound

This paper cites ArXiv , year=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization ArXiv , year=

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.295049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:f44345e6aba790ca0169ad36f836a9389d258f0de43b95e5436bf848bc548fed

Observation 42ea5fed-51e8-49d2-afe6-ee69b4fb1a02 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 32

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:37:29.836589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:18f735f3b87713ee2b51ce6b34382b0f3fb18008dfcfdce008ffa5a7a4f9bb4c

Observation d95be7c4-4c29-43f2-a4ad-0bafadbe4b76 · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-05-13T07:37:29.923782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:124520ae8e57eca40fce4dab514656b8a7425276256c84f619b321eafb45d58a

Observation 9076d9d2-869e-42e8-baf6-73e554382b6c · outbound

This paper cites InternLM-Math: Open Math Large Language Models Toward Verifiable Reasoning.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization InternLM-Math: Open Math Large Language Models Toward Verifiable Reasoning

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:37:29.887905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:21894625a10b0ab8338ab66a997b7b197dd8364db4de356b3e198ff2a0638d2c

Observation 0b7c421d-921c-4f56-b5c0-8d242e25aee3 · outbound

This paper cites Qwen2.5-Coder Technical Report.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Qwen2.5-Coder Technical Report

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-05-13T07:37:29.933416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:07689fbf7e9a25b17d4f703ced2a94e67f2782db414e6ff16a6207560ca37002

Observation 1f28dbfd-d9dd-415e-8e08-3788e98bfeb9 · outbound

This paper cites Preference Optimization for Reasoning with Pseudo Feedback.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Preference Optimization for Reasoning with Pseudo Feedback

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:37:29.881455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:878be547cebc0a032ab173d3136a40d35f426e046fa7778713d2f1779c55ee78

Observation d9c77c58-691a-41e5-a802-469db297c38c · outbound

This paper cites ArXiv , year=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization ArXiv , year=

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.237308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:944ed15722a9c59b374fcbedfd999148596d22edd8cddcdd19c7589b1158638a

Observation 8cc6f019-6d8f-482f-924f-841fba0a9cdf · outbound

This paper cites AceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization AceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:37:29.841604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:d4278a1a9e2b397be83e5c5534e3b39c950c7e32fd5829b74376338b3da8f75f

Observation 903285d5-f30e-477b-99f6-3badcbb1d3b0 · outbound

This paper cites CodeDPO: Aligning Code Models with Self Generated and Verified Source Code.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization CodeDPO: Aligning Code Models with Self Generated and Verified Source Code

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:37:29.939329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:f1898b190438736c8a95db3c8a79348e886e4e44e2e8fdbc526e4c84152c55ed

Observation c9042bc7-b01f-43a6-91d0-859b5e513e6c · outbound

This paper cites Proximal Policy Optimization Algorithms.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Proximal Policy Optimization Algorithms

Reference 40

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:37:29.893688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:6e8281f20d11e5d591aeb591026c11aa216feca74b0ee47103c251e0d1f1c1d4

Observation 9147c6af-295b-4b83-831d-307ff1f6eeeb · outbound

This paper cites Advances in neural information processing systems , volume=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Advances in neural information processing systems , volume=

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.282944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:433d363ce431b9c7d6488285482726b09a67836e9f2f1d3a0aec38092f231a31

Observation 618675be-7fab-41c9-ab5b-b7dad2daad01 · outbound

This paper cites Advances in neural information processing systems , volume=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Advances in neural information processing systems , volume=

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.233080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:459bbec330fbfd1fefb78da07a58d8123a127a273424c314d9dddbd840f9bd63

Observation 104d1f22-4915-4274-ac35-a8d19d9c7e2b · outbound

This paper cites Direct Preference Optimization with an Offset.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Direct Preference Optimization with an Offset

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:37:29.874766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:fe08109b494e667279681149db9eb72cdf9b7919149f2d7a59e41ca312a45317

Observation 3dd2b738-438b-4692-8fc8-673e0334ace8 · outbound

This paper cites Step-level Value Preference Optimization for Mathematical Reasoning.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Step-level Value Preference Optimization for Mathematical Reasoning

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:37:29.952562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:7188a09388dd5999dcc2731ba53326495187a7cacdd9db95792189307ed36de3

Observation 754e0a15-3f02-4279-b195-0293cbe81b07 · outbound

This paper cites Self-Rewarding Language Models.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Self-Rewarding Language Models

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T12:01:42.779186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:f860b17b6be6b44b1aff00b9be3b7deb69aa8aaba741accbfb83e914ee21dc2e

Observation d37aecc7-4f51-4543-9e9a-8999ad36b06b · outbound

This paper cites The Alignment Ceiling: Objective Mismatch in Reinforcement Learning from Human Feedback.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization The Alignment Ceiling: Objective Mismatch in Reinforcement Learning from Human Feedback

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:37:29.854732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:d91e07207cc702d96ca48ce2a4a45d285b2014c1bd85d8fbb37e1a82a554c799

Observation 9e9cfa9b-6513-47ce-a9ae-ac8768437ab2 · outbound

This paper cites Secrets of RLHF in Large Language Models Part II: Reward Modeling.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Secrets of RLHF in Large Language Models Part II: Reward Modeling

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:37:29.945673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:e8b5c25780369122afa84da3dff7af108a824645d6fc968927186aac2e7d2ea9

Observation 46547bf7-8b40-4470-a8f0-60745dcae52a · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 48

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:37:29.861770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:bbb373d089b7fcc0f28312b4dc870659340361ae01d8eaa15e31e53dccd72cfd

Observation 2b9c0382-4228-465e-8d9e-8b8b7bad606c · outbound

This paper cites Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information

Reference 49

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:37:29.848712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:b143ce3666bc0d46f7270751a939694fa65436fd734b61a7708359741af7258b

Observation 4439da91-bd72-4035-8751-93095a40737e · outbound

This paper cites Neural-Symbolic Solver for Math Word Problems with Auxiliary Tasks.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Neural-Symbolic Solver for Math Word Problems with Auxiliary Tasks

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:37:29.868070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:d291ed2af473db24eb174dbe85783ce66bada5a2ec48eab00eae015190ac5c23

Observation 0a941f00-88e0-443b-9a98-151094ac51f8 · outbound

This paper cites Proceedings of the 2013 conference on empirical methods in natural language processing , pages=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Proceedings of the 2013 conference on empirical methods in natural language processing , pages=

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.266220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:36aac8c0170ee8011891e24c99ba6cad3bd2440d1a42633554cae68e1d3e9860

Observation 54900de6-464b-4bdd-bf27-8f128ab2e243 · outbound

This paper cites Know What You Don't Know: Unanswerable Questions for SQuAD.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Know What You Don't Know: Unanswerable Questions for SQuAD

Reference 52

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:37:29.908455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:e415d6f23166e236ff865b746e9db0c3c308e7c235ab5653b7c16191aaa92f44

Observation 088a2967-6062-4c44-b660-a3a96c2e2756 · outbound

This paper cites Eliminating Position Bias of Language Models: A Mechanistic Approach.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Eliminating Position Bias of Language Models: A Mechanistic Approach

Reference 53

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:37:29.918096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:c49bd55205664cd9f7d324d186ed13a0c636cefa4f13ea88822b6101ecee0c9b

Observation f5af5d75-332c-4192-8ec7-1c89e6281408 · outbound

This paper cites Qwen2.5: A Party of Foundation Models , url =.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Qwen2.5: A Party of Foundation Models , url =

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.277998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:ffe7a59905b573eedb16d88fc4aac9729c8fad00dd0e9cd29ea78f152fa36006

Observation 67240abf-fa2e-4ba2-ae38-52d51a44508a · outbound

This paper cites In-Context Learning with Long-Context Models: An In-Depth Exploration.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization In-Context Learning with Long-Context Models: An In-Depth Exploration

Reference 55

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:37:29.901629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:1ac5a6725a7d363c4e9a61b42534b98fe9bcb6d52a616e35908711238b3ff42b

Observation 0d1d92ee-cd46-43f4-9788-2c0648c4691c · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Advances in Neural Information Processing Systems , volume=

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.299642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:445dec702b132e74fdf74f9245d8f1f4dab71af781a48be4f132dc5be95bd7d8

Observation df55ac3c-4b56-4a94-8271-19d8d5c51517 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Advances in Neural Information Processing Systems , volume=

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.245085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:f829235e719480ca66217b10f631d0f9a8133f4ea5155db00c77c4f06a6eaf9e

Observation efcac845-91b7-4687-8860-e397f9c909fb · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Advances in Neural Information Processing Systems , volume=

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.249789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:07503b93ef52ef27e5d7f97c61c2a9907237d75428c684a9525832e2a38e44d1

Observation 1243ead2-91b6-4eb9-9fc9-a88278b240b1 · outbound

This paper cites Arm-thinker: Reinforcing multimodal generative reward models with agentic tool use and visual reasoning.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Arm-thinker: Reinforcing multimodal generative reward models with agentic tool use and visual reasoning

Reference 59

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:37:29.824883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:4fb85849cb69204c2e250f68d0096b1be1e5f701fe75f408400bba33e6ef35f5

Pith citing papers

No inbound Pith citation observations are available.