Pith. sign in

Paper Citation Record · LEDGER

LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 47 inbound Pith citation observations for arXiv:2410.02884.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.02884 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 47 of 47 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:48:20.054861Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T16:29:57.287817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation de4aa7cf-ee15-4fc3-93f8-e136e69a60a5 · inbound

Search, Verify and Feedback: Towards Next Generation Post-training Paradigm of Foundation Models via Verifier Engineering cites this paper.

Search, Verify and Feedback: Towards Next Generation Post-training Paradigm of Foundation Models via Verifier Engineering LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 152

Resolution
unresolved
no resolver link, observed 2026-08-12T18:28:49.564529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:28:49.564529Z digest=sha256:e5d2ea442d2fcb622de87674bed9de4d9a00bc638802a3e2b991b3d947e723c0

Observation ecf2f00d-1a7d-41ad-b28f-6ea92b15a06a · inbound

Enhancing LLM Reasoning with Reward-guided Tree Search cites this paper.

Enhancing LLM Reasoning with Reward-guided Tree Search LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T18:18:28.820471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:18:28.820471Z digest=sha256:f493140763e69f18e9aba4c5cf1e3e1b03a601db5749f8a32f29f7257198fe6a

Observation 4a67b251-f41c-42ba-a47d-927695d4d923 · inbound

Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision cites this paper.

Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-12T13:02:57.358278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:02:57.358278Z digest=sha256:8d0c67553b6336e683b603f4c0c836f3e7672bc288d386e68f7076d05b600518

Observation 4c064ee6-a811-4e7e-87f3-006990c4a0b1 · inbound

Critic-V: VLM Critics Help Catch VLM Errors in Multimodal Reasoning cites this paper.

Critic-V: VLM Critics Help Catch VLM Errors in Multimodal Reasoning LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-12T11:27:33.403459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:27:33.403459Z digest=sha256:6c2a845fd56d3131180b6befed50434670ccf3e9fca9f005c484d731e6c89909

Observation 9a61a44d-5e3c-421a-b790-2fa1395b028c · inbound

Beyond Examples: High-level Automated Reasoning Paradigm in In-Context Learning via MCTS cites this paper.

Beyond Examples: High-level Automated Reasoning Paradigm in In-Context Learning via MCTS LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-12T11:13:42.575001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:13:42.575001Z digest=sha256:31941702c7b2f9446a26e56544b55473ea9dadd13fea1dc553a8ac9b7a5c69ef

Observation b932dfb7-4bab-4c32-8a26-883597f0c263 · inbound

o1-Coder: an o1 Replication for Coding cites this paper.

o1-Coder: an o1 Replication for Coding LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:29.085430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T10:11:29.085430Z digest=sha256:495e0604388f63e751f82d00a5c41e4ca244a48ee6c58732825890971204754c

Observation c196a863-f5ca-4cad-813a-6f1ed1e12eb4 · inbound

Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning Systems cites this paper.

Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning Systems LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:35:31.446927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-18T00:35:31.375020Z digest=sha256:3009e4db4bf4f4212eb4b2f856b905149960c412c3ec646e7ae6a17850a4aae1

Observation 0c836edf-5459-4b19-ab91-1b5d66732b5f · inbound

RAG-Star: Enhancing Deliberative Reasoning with Retrieval Augmented Verification and Refinement cites this paper.

RAG-Star: Enhancing Deliberative Reasoning with Retrieval Augmented Verification and Refinement LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T13:42:49.743901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:42:49.743901Z digest=sha256:83d308a45f382599ffe83f34fe048057c34eb50499dc49e8d937aeeef046d83a

Observation f4a67402-147c-403f-8278-11d586f7925e · inbound

Think&Cite: Improving Attributed Text Generation with Self-Guided Tree Search and Progress Reward Modeling cites this paper.

Think&Cite: Improving Attributed Text Generation with Self-Guided Tree Search and Progress Reward Modeling LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T11:55:30.569744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:55:30.569744Z digest=sha256:7a34566c433f831cc099c7a75d913186e2605a8a8bcc4dd8dbf78dd3a6b6b147

Observation bc8882e3-3828-4ee9-a73b-c4d5ca976e0a · inbound

HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs cites this paper.

HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-15T12:36:50.157350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T12:36:50.060335Z digest=sha256:908e076f0a8c19577c6cc4512986791cfecc6ffd14dc462d6cc8417613a94549

Observation 36a2a22f-dce7-4fa1-8790-b4e62d1ae012 · inbound

BoostStep: Boosting mathematical capability of Large Language Models via improved single-step reasoning cites this paper.

BoostStep: Boosting mathematical capability of Large Language Models via improved single-step reasoning LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T21:57:50.825944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:57:50.825944Z digest=sha256:b478640f14316f0de4d724835d4324530dfd424eeb11491a0896d7e84672411e

Observation ea5fe782-2f0d-4b24-ba71-c408cb908a1d · inbound

Search-o1: Agentic Search-Enhanced Large Reasoning Models cites this paper.

Search-o1: Agentic Search-Enhanced Large Reasoning Models LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 77

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:36:27.643883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-13T17:36:27.515468Z digest=sha256:3fa669ded53a7e895bbe3afc3d7521e701618849cf761c9fbe8234f14e201941

Observation 1cd4adb3-21b7-4b9a-82ee-e6d3d2c29b27 · inbound

Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models cites this paper.

Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 187

Resolution
verified exact
arxiv_id, observed 2026-05-15T21:20:59.427455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T21:20:59.128986Z digest=sha256:865b3bd35e3df89a6d11c7474e96655dc4ebc18812e67666ddb5706bfbe3475d

Observation edc7cdbd-51a5-4b1c-9f31-baec7e99413d · inbound

AirRAG: Autonomous Strategic Planning and Reasoning Steer Retrieval Augmented Generation cites this paper.

AirRAG: Autonomous Strategic Planning and Reasoning Steer Retrieval Augmented Generation LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:46.417078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T19:27:46.417078Z digest=sha256:091edfa1c37a92943e137831e1e1479ab4280ac602df407c0c21b4e1bc4f4d9b

Observation 4bfc6eeb-f2ca-4590-b52d-89aa034c0b8f · inbound

Large Language Model-Enhanced Multi-Armed Bandits cites this paper.

Large Language Model-Enhanced Multi-Armed Bandits LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 2012

Resolution
unresolved
no resolver link, observed 2026-08-09T16:35:28.447005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T16:35:28.447005Z digest=sha256:5825fc1843e3c5423adc3452c5223b725b44b9f93ffa86a1da88fd76be7d3db7

Observation cf00df94-b20b-4b68-8a09-efa45ca0b981 · inbound

Satori: Reinforcement Learning with Chain-of-Action-Thought Enhances LLM Reasoning via Autoregressive Search cites this paper.

Satori: Reinforcement Learning with Chain-of-Action-Thought Enhances LLM Reasoning via Autoregressive Search LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T11:57:48.815737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:57:48.815737Z digest=sha256:03316b3d2c0858f6183b25e27ac305e2daa02c9d682b104daf56056186d6f8b1

Observation b6ff821c-16ab-4b30-9cdc-9bb1c1ebb1ff · inbound

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning cites this paper.

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 246

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:32:32.542825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-23T04:30:38.804702Z digest=sha256:733c6a0b9479219915f74829388c24c891759dea6dfcff84d574c4ef92813775

Observation d8588eec-ff8a-4a69-9abf-5cc2cfbe793b · inbound

Agentic Reasoning: A Streamlined Framework for Enhancing LLM Reasoning with Agentic Tools cites this paper.

Agentic Reasoning: A Streamlined Framework for Enhancing LLM Reasoning with Agentic Tools LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T22:03:46.492505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T22:03:46.492505Z digest=sha256:12b24e65818f7849e4585213f11b4830269a5483329c53ce5a8c0eeb96065e11

Observation 170a4518-75c1-4145-ac98-90b41a6f8266 · inbound

Holistically Guided Monte Carlo Tree Search for Intricate Information Seeking cites this paper.

Holistically Guided Monte Carlo Tree Search for Intricate Information Seeking LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T21:40:26.985271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:40:26.985271Z digest=sha256:346ed93e25ebd0d359744a78175d49d9ebacaab46fb343a1d1cefb3515beac5c

Observation 407f26e5-679b-458a-860b-e946dfad7d99 · inbound

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition cites this paper.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.631264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.631264Z digest=sha256:d3095842d891c23f4ed164f3b63935746cff4be9db7fe76941ddc1592498dc1c

Observation cb1b0403-8a74-4fa9-bec1-c619e07898e1 · inbound

From System 1 to System 2: A Survey of Reasoning Large Language Models cites this paper.

From System 1 to System 2: A Survey of Reasoning Large Language Models LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 232

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:36:24.378838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-13T01:36:23.845366Z digest=sha256:9691d319df94e75df7f48b3f5c60c90ca380d353186713585b6aa404b5f850af

Observation af48fcaf-14c9-4ccc-b398-0d066deade89 · inbound

MathFlow: Enhancing the Perceptual Flow of MLLMs for Visual Mathematical Problems cites this paper.

MathFlow: Enhancing the Perceptual Flow of MLLMs for Visual Mathematical Problems LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-22T22:57:13.313394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-22T22:55:34.238427Z digest=sha256:7b6a6ded39704e243dc5a94d5171380314e13669c7acca45559a25a28ad747f6

Observation bd1a5786-fbb5-42d5-b9c1-36175921f33d · inbound

OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles cites this paper.

OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 88

Resolution
verified exact
arxiv_id, observed 2026-05-19T06:59:03.329513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T06:59:03.112252Z digest=sha256:f95b6d18eceb44375ad9bd26190fabcee6ab5a9de7916ac7bc937e2c6dd5b5e6

Observation bc6b0951-13c5-416e-a26c-64c85a88fa62 · inbound

Advances and Challenges in Foundation Agents: From Brain-Inspired Intelligence to Evolutionary, Collaborative, and Safe Systems cites this paper.

Advances and Challenges in Foundation Agents: From Brain-Inspired Intelligence to Evolutionary, Collaborative, and Safe Systems LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 92

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:42:10.437330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-22T21:39:49.832151Z digest=sha256:a94bfbfdd70953d25324d87a39322f3d89df94536c02ea55bc249ecc6d5b4226

Observation fd995c89-9839-4fb7-9549-9e722caa5ab1 · inbound

A Survey of Slow Thinking-based Reasoning LLMs using Reinforced Learning and Inference-time Scaling Law cites this paper.

A Survey of Slow Thinking-based Reasoning LLMs using Reinforced Learning and Inference-time Scaling Law LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 159

Resolution
unresolved
no resolver link, observed 2026-08-16T00:48:20.054861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:48:20.054861Z digest=sha256:0dc1a119a38ed3e485973a85550def1c6f868194c2d409219191e72e1f58c91e

Observation 7eb35742-43c2-473f-a627-9be68b112feb · inbound

Chain-of-Thought Tokens are Computer Program Variables cites this paper.

Chain-of-Thought Tokens are Computer Program Variables LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T23:23:02.189328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:23:02.189328Z digest=sha256:105e85a8d9955503239bbb30d3ec150ee2750493436c7b5d80def06996639762

Observation 1e219a06-2a95-42f5-a6fe-d0d55e9030d0 · inbound

CRPE: Expanding The Reasoning Capability of Large Language Model for Code Generation cites this paper.

CRPE: Expanding The Reasoning Capability of Large Language Model for Code Generation LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T21:22:19.025191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:22:19.025191Z digest=sha256:615036daa2389a5f1b2374411c446a7ec88ddfb9236b73d48becbe604a98bdf8

Observation 2b61f84a-557a-49f5-8e22-56a820d7d41e · inbound

J1: Exploring Simple Test-Time Scaling for LLM-as-a-Judge cites this paper.

J1: Exploring Simple Test-Time Scaling for LLM-as-a-Judge LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-15T20:50:10.767033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:50:10.767033Z digest=sha256:9a95db5e687134c748a460ea1b4488933578a62ed11df7a3b204f287e7469029

Observation 2b8f735d-2799-4145-b54a-06a8b8c257a7 · inbound

Enhancing Large Language Models with Reward-guided Tree Search for Knowledge Graph Question and Answering cites this paper.

Enhancing Large Language Models with Reward-guided Tree Search for Knowledge Graph Question and Answering LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:09.418765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:09.418765Z digest=sha256:3bd69c13d2ac48deeb7c865067720ea5bd51383c835bd32a651a425381d008ce

Observation 3dc80ecb-14b2-40a2-a3be-0fd6924aef7c · inbound

EquivPruner: Boosting Efficiency and Quality in LLM-Based Search via Action Pruning cites this paper.

EquivPruner: Boosting Efficiency and Quality in LLM-Based Search via Action Pruning LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:06:02.552762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:06:02.552762Z digest=sha256:eb6f040cb42a7076bd83b3e31b5c2a84248bd93126c36da4e6d28b6db93a7137

Observation b23991e7-dce8-4463-9a82-445e0a635091 · inbound

Reward Model Generalization for Compute-Aware Test-Time Reasoning cites this paper.

Reward Model Generalization for Compute-Aware Test-Time Reasoning LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:42:27.539519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:42:27.539519Z digest=sha256:8aeab58f550c773fd40dc53d9d0db8e95b525a755b82070db80a0d4df1838ade

Observation e10c0273-71c9-40c8-b89f-07d805541b8d · inbound

More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models cites this paper.

More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:36.471077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:36.471077Z digest=sha256:f69bec922a794477742acf7b2dadfeab83618fd5d23e0d660ed0ed9ebe73315c

Observation 25b11409-beeb-43db-9ca3-f480f5916501 · inbound

Advancing Multimodal Reasoning via Reinforcement Learning with Cold Start cites this paper.

Advancing Multimodal Reasoning via Reinforcement Learning with Cold Start LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T13:12:58.712192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:12:58.712192Z digest=sha256:a989a3b5c28036abb23e204588facef6ffb9c37c6a79b02d0ff9259b461b0904

Observation 9ebaadc9-87c8-41a0-ba69-f23f61aa0c81 · inbound

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM cites this paper.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.857251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.857251Z digest=sha256:be96bffcc1b77b9cbae9728b62dfa4a0420b6864463252ab120c86f32a0b09ac

Observation ab44c3fb-adf0-49fd-91b6-4a4f2c95a478 · inbound

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL cites this paper.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.401314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.401314Z digest=sha256:38816f372c166e6ac222ed1055770a09034b12e1b602351048a26c20e1f7c793

Observation f8eaf83f-44f6-4818-8360-306b7170690b · inbound

SELT: Self-Evaluation Tree Search for LLMs with Task Decomposition cites this paper.

SELT: Self-Evaluation Tree Search for LLMs with Task Decomposition LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:38:34.047050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:38:34.047050Z digest=sha256:ee20bd24a63009c4b69407623078f427fdd7a957d3ed96457f6a171486909397

Observation 898d3920-df61-443f-a8bb-61b92375818a · inbound

A Survey on Large Language Models for Mathematical Reasoning cites this paper.

A Survey on Large Language Models for Mathematical Reasoning LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 106

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:47.574551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:14:47.574551Z digest=sha256:c0daa691374413f38991002effb72f142b5950221d1d3e5af20c489d34b862b2

Observation f87f5801-3e66-45b7-a293-8186d484b3df · inbound

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism cites this paper.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:23.349125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:23.349125Z digest=sha256:70a291fb5f7cee0eec05937a60792b485486e5a9f77eb0a6e30fbeffbbf1d04b

Observation e9c90b97-f75b-49c0-b883-4ae79b60d37e · inbound

Boosting LLM's Molecular Structure Elucidation with Knowledge Enhanced Tree Search Reasoning cites this paper.

Boosting LLM's Molecular Structure Elucidation with Knowledge Enhanced Tree Search Reasoning LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T21:59:13.738456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:59:13.738456Z digest=sha256:4e4933232b91ca51b9fd81641590b4da8b8c747a8f842294d5c73fa583071309

Observation ffeeb267-0a6e-4780-8502-2ca705872d90 · inbound

Building Task Bots with Self-learning for Enhanced Adaptability, Extensibility, and Factuality cites this paper.

Building Task Bots with Self-learning for Enhanced Adaptability, Extensibility, and Factuality LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 207

Resolution
unresolved
no resolver link, observed 2026-08-05T15:38:55.248368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:38:55.248368Z digest=sha256:85d2a85ad8022d1838c02db7894f3a5dbdeda1c5c66ac56c3971e9851855c6c6

Observation f2bd868a-d220-45f5-ab20-d92fb7837cc5 · inbound

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models cites this paper.

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T10:39:00.088338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:39:00.088338Z digest=sha256:78f2698ba61798e8d2b909de8d1f2da2f85f81d3b80fb3977fb5d86852157470

Observation b4439750-3acd-4337-bf61-23ea4695e8ba · inbound

Plan Then Action:High-Level Planning Guidance Reinforcement Learning for LLM Reasoning cites this paper.

Plan Then Action:High-Level Planning Guidance Reinforcement Learning for LLM Reasoning LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T12:52:28.323140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:52:28.323140Z digest=sha256:a40f0736f0bb825d525f1b2323615892b58fec323d7f12450c0d69afde232063

Observation 8602e035-1add-4b8e-b107-eb26c08f2193 · inbound

TaoSR-AGRL: Adaptive Guided Reinforcement Learning Framework for E-commerce Search Relevance cites this paper.

TaoSR-AGRL: Adaptive Guided Reinforcement Learning Framework for E-commerce Search Relevance LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-04T10:53:10.440907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:53:10.440907Z digest=sha256:15820cee7c8a71a82cb3fc0c5377296a95806e5da9a1aa1703287855fc21fa15

Observation d58a5c61-c0a5-4142-9673-253d1e27afb8 · inbound

SIGMA: Search-Augmented On-Demand Knowledge Integration for Agentic Mathematical Reasoning cites this paper.

SIGMA: Search-Augmented On-Demand Knowledge Integration for Agentic Mathematical Reasoning LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T06:57:16.504758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:57:16.504758Z digest=sha256:fd3057ba71c91f0bad11715833e8c7c3ed59575137ede94a5d29e254e848040f

Observation 886e57de-74cb-493d-b40b-984eca64f435 · inbound

Mathematical Reasoning via Intervention-Based Time-Series Causal Discovery Using LLMs as Concept Mastery Simulators cites this paper.

Mathematical Reasoning via Intervention-Based Time-Series Causal Discovery Using LLMs as Concept Mastery Simulators LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:55:57.590779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-11T02:05:48.483640Z digest=sha256:018259e815f4a56c21a8c3c085cccc056196950fbdd18b0f05ce7575c482cca6

Observation 13c9f9ad-3931-456f-ae39-c1d6ec93214e · inbound

Latent Visual States for Efficient Multimodal Reasoning cites this paper.

Latent Visual States for Efficient Multimodal Reasoning LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-07-04T16:29:57.289397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-26T00:38:11.619574Z digest=sha256:63446e4a382d3813eacaecf944bb3304b94cd81e9162135f49e098aad1860281

Observation f4f5434f-73dc-467b-bdf1-89cbd12eebc4 · inbound

ThinkRetrieve: Retrieval-Augmented Reasoning Traces for Test-Time Scaling cites this paper.

ThinkRetrieve: Retrieval-Augmented Reasoning Traces for Test-Time Scaling LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 132

Resolution
unresolved
no resolver link, observed 2026-08-12T14:10:45.510784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:10:45.510784Z digest=sha256:2120a63e4c12e85d6973b321c62721df9837b120ecf62f9afe3c3e75998181e3