Pith. sign in

Paper Citation Record · LEDGER

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

As of 4 August 2026, this Paper Citation Record lists 100 of 274 outbound references and 95 inbound Pith citation observations for arXiv:2412.21187.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.21187 v2

Coverage vector

measured 100 of 274 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-13T15:51:29.022336Z

measured 195 of 195 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 95 of 95 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T13:51:50.210041Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T11:37:03.272167Z

Reference resolution

100 of 274 outbound references displayed

  • verified exact11
  • verified fuzzy69
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch18

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9ee45d81-8b36-467a-b626-adbc982b2386 · outbound

This paper cites Graph of thoughts: Solving elaborate problems with large language models.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Graph of thoughts: Solving elaborate problems with large language models

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:51:30.423853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:ea5166dca3c9d1dfde3040804b602214b27dbb379861432b0860041634266710

Observation a5a8c85f-007f-4312-b530-284b87c955f2 · outbound

This paper cites Learning How Hard to Think: Input-Adaptive Allocation of LM Computation.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Learning How Hard to Think: Input-Adaptive Allocation of LM Computation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:51:29.476618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:c7f685a445cee1cc3879c42f81b2a2b91bffb44e55ad5cd4ba0d5821403447dd

Observation a82ac9ad-c835-45a5-adda-b384f2fbc359 · outbound

This paper cites Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.096184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:9b641ac87d2ab59a72f98e19b171c40de3fbf1e06966c45afa563f90d97d6759

Observation b69b1fb7-9971-48b1-a4c3-37c08eac5d9e · outbound

This paper cites Improving factuality and reasoning in language models through multiagent debate.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Improving factuality and reasoning in language models through multiagent debate

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.100719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:823c3b9ff280def346333a4dc1942678750d4fb696e04498c143c5410e9ae3ae

Observation 586081ee-5b3e-45bf-ad1f-8250f70378c0 · outbound

This paper cites Think before you speak: Training language models with pause tokens.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Think before you speak: Training language models with pause tokens

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.110972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:859c2c8c1406a3b6f80174c964e33a0bc8a22bd9adac7087a3439b396e15c294

Observation 33b695fe-d50f-47d4-833f-ad1d67139b82 · outbound

This paper cites Training Large Language Models to Reason in a Continuous Latent Space.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Training Large Language Models to Reason in a Continuous Latent Space

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-05-13T15:51:29.470372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:8854791332bf1659f1657522d63c1ae15ca05f97cae8b2ad65e2f1bc218c1378

Observation b530ca2f-6aeb-4a28-82a4-05c44c23aea4 · outbound

This paper cites Improving minimum bayes risk decoding with multi-prompt.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Improving minimum bayes risk decoding with multi-prompt

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.115733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:248e329ab97610b3b434cbcbbc60f927d77b0843b8f1d2e45fc561a7776d7098

Observation f75f5ce6-e16c-4d70-97bd-445f37791658 · outbound

This paper cites Measuring mathematical problem solving with the MATH dataset.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Measuring mathematical problem solving with the MATH dataset

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.244905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:370b90af3b3202af44ab2fc15e41d4eb5c68bb8a1d3ee1482c3b27ec6de0a0fd

Observation 5050a9be-bdb1-456f-8f62-8d6f91990e9f · outbound

This paper cites Large language models are reasoning teachers.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Large language models are reasoning teachers

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.362239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:c0c2d01cfe4c718dd300f3ce57e19c0023062b8c15a7c9e3bf144fbf3e1a0841

Observation 6b4fef62-f032-42ad-b76e-2a4ddb55f5fa · outbound

This paper cites When can llms actually correct their own mistakes? a critical survey of self-correction of llms.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs When can llms actually correct their own mistakes? a critical survey of self-correction of llms

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.249376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:311a1b8e79bd401ea64ace577d60abe2ada7f5cf807d36a71b4b74350604d0bc

Observation d211d0ab-3fa3-4e98-8375-7b0ab98d682f · outbound

This paper cites Args: Alignment as reward-guided search.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Args: Alignment as reward-guided search

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.339625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:d5db64b3cab3bd248d5b330b7df5dec32cda8bfdfe98a82b487d2cfe69075cc1

Observation ad1e4531-7843-493d-a8ed-d10c1a64ae53 · outbound

This paper cites Prover-Verifier Games improve legibility of LLM outputs.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Prover-Verifier Games improve legibility of LLM outputs

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:51:29.224911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:6ad1602bd76ca38dc385e634311df869c9a5b4d769fd494d33fac61f9aa405ce

Observation 9d9156f8-2f5f-439a-819a-376bc7bce139 · outbound

This paper cites Large language models are zero-shot reasoners.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Large language models are zero-shot reasoners

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.280955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:05ab05d6015487408f86a9943b96c24433b34f5fb697ae1f4050c6adaca972d9

Observation 25a54340-cadc-407d-8916-f191197b2cc8 · outbound

This paper cites Escape sky-high cost: Early-stopping self-consistency for multi-step reasoning.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Escape sky-high cost: Early-stopping self-consistency for multi-step reasoning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.285336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:bea9c022e6d688e02b07b45cf4a9880443a557cf272e8950d3ff9e14a4afb501

Observation 4f6a8c78-69e4-4722-8f1a-799123695d08 · outbound

This paper cites Let's verify step by step.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Let's verify step by step

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.137014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:6fd5297513a87e095b2d265e88a00efc098cecb8272f93aa5628a3f77b6cce99

Observation e4934237-0cef-47d4-8266-8f599cd14eab · outbound

This paper cites Adaptive Inference-Time Compute: LLMs Can Predict if They Can Do Better, Even Mid-Generation.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Adaptive Inference-Time Compute: LLMs Can Predict if They Can Do Better, Even Mid-Generation

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:51:29.352729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:a4e2045f3a077370b0e5d4cc1756d8a8daf89cd8d48282974dc0f9b620a57cb4

Observation f0a6d509-3014-468c-9337-4d972aa6203f · outbound

This paper cites Simpo: Simple preference optimization with a reference-free reward.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Simpo: Simple preference optimization with a reference-free reward

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.320530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:9893f43189ee3678a11c7d286d35f91d819b2bf959d683be173054f50b3c82a7

Observation 30bae8e9-2145-4e89-8614-4f3ce73f7cad · outbound

This paper cites A diverse corpus for evaluating and developing english math word problem solvers.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs A diverse corpus for evaluating and developing english math word problem solvers

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.324566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:1e2f5d87f80385296912efec208401900d2289447df93255ba636fedee121fec

Observation 66abe065-91bc-4a2d-aa51-1e088dffdff6 · outbound

This paper cites Learning to reason with llms.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Learning to reason with llms

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.179065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:5ae306d5aee6d943c427cd340dd8a43b8dc09e04f911c68ed42b810651f5712a

Observation f6653a67-87d9-4f82-81b2-48e04ea81b76 · outbound

This paper cites Iterative reasoning preference optimization.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Iterative reasoning preference optimization

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.214092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:86773a062e35578e872bda107e014a8079d174fa3572af3bf2102a3ce93abc6d

Observation 56c47e10-9a76-48ca-a2da-a47684b71373 · outbound

This paper cites Qwq: Reflect deeply on the boundaries of the unknown, November 2024.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Qwq: Reflect deeply on the boundaries of the unknown, November 2024

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.303556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:7136352a34ed3d2e36e975b85de218f65228065723dff6582cf481e0fbefce3c

Observation a826e907-bf6d-402d-bc35-b0842da71dca · outbound

This paper cites Direct preference optimization: Your language model is secretly a reward model.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Direct preference optimization: Your language model is secretly a reward model

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.149229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:2975cd5af79046cf8d1d53a7e5c5cbc812845754ec1343190143c095d398d566

Observation 6e5de956-15cf-4c2e-a392-74dbf51a7f70 · outbound

This paper cites Alphazero-like tree-search can guide large language model decoding and training.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Alphazero-like tree-search can guide large language model decoding and training

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.332456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:b2943ab26b03bc117b842e43f674eb794b9923cb28f896ffc317d5488aaec214

Observation 2a677e1f-e64f-4444-a8e3-cc9fd47e445d · outbound

This paper cites Make Every Penny Count: Difficulty-Adaptive Self-Consistency for Cost-Efficient Reasoning.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Make Every Penny Count: Difficulty-Adaptive Self-Consistency for Cost-Efficient Reasoning

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:51:29.314914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:d05890c3cf1684209a8c5a2aeccfd46950568b03534f5e2fd51da1b94c2c4d52

Observation 0c34e016-5a13-4101-b6d0-12f1a395a105 · outbound

This paper cites Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.086898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:d7dd3d9ee6fe73f6465148c4eeefc35889207ce15e654f17fa5fa39ae2a34dd3

Observation a3863dbc-a729-469b-9b23-f9b32efc8625 · outbound

This paper cites Finetuned language models are zero-shot learners.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Finetuned language models are zero-shot learners

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.014146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:69fb4c8e2a4c4881a5c9d1f480ac20118c8e9fc2498f3dab06a73f95299536c9

Observation f34497f5-95ee-44cf-aece-159e4e25f39e · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Chain-of-thought prompting elicits reasoning in large language models

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.066139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:35281191aced4f4831c3b607e2d5274d889a442e0262dcca04c277f425ceaf56

Observation b6fad250-a192-4fd4-94d9-92487d5ebc7e · outbound

This paper cites Examining inter-consistency of large language models collaboration: An in-depth analysis via debate.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Examining inter-consistency of large language models collaboration: An in-depth analysis via debate

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.039500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:b61f7c1e1217e21b337005b31359cf2fd33a4302e7058aed27af229e0a1593da

Observation baaa20c3-a4f1-41ea-81c8-f5e7922817b0 · outbound

This paper cites Tree of thoughts: Deliberate problem solving with large language models.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Tree of thoughts: Deliberate problem solving with large language models

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.033888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:9bc8a1126176b61985d8f373556b98c00293977d74a696a8b8a505d0ea2f2405

Observation 37f61655-6c42-4345-ac46-1cdf04ae4ed5 · outbound

This paper cites Star: Bootstrapping reasoning with reasoning.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Star: Bootstrapping reasoning with reasoning

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.074586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:d68d497956d47b99b8e2ee0b91e9e9e36fe8e6b68c8eca6bc2dd53780dcc8aaa

Observation 2eae992a-d7dd-440e-a2cf-01206399f82f · outbound

This paper cites Automatic Curriculum Expert Iteration for Reliable LLM Reasoning.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Automatic Curriculum Expert Iteration for Reliable LLM Reasoning

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:51:29.418568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:c97e37fa52707224f2b37e60d98e62c1e7a5d0d7680fbb8a99042324e6247299

Observation 155dd51e-d0a0-497c-b0d0-808696ba0b50 · outbound

This paper cites International Conference on Learning Representations , year=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs International Conference on Learning Representations , year=

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.028486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:c1b36f9ee3d97ee5d43cbf2b6f39c875ddaab5be311a58e3d1a5924298c3b3b1

Observation 25f8ee6f-850c-4eb8-868e-332c27673ca3 · outbound

This paper cites Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics , year=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics , year=

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.045166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:5cd31346bc346499e4d660380656acb909f501d5d8befaf0432076203bfbc367

Observation df154638-28ea-4070-8a0d-282e4f14cd5b · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Advances in Neural Information Processing Systems (NeurIPS) , year=

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.050486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:051f342297cf05e7e326b808f63587daf4df203f0f002cd30a46cf8ec3cf1aec

Observation f16ca39c-4922-4f1a-8d85-2d147308172e · outbound

This paper cites Measuring Mathematical Problem Solving With the.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Measuring Mathematical Problem Solving With the

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.070414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:ef29a9fadb43700255fb5ba91276e7f3b936e11f51561d6166c052c1726943a0

Observation ff484b00-7452-4be2-90c2-0a731645bbc3 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Advances in Neural Information Processing Systems , volume=

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.081786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:11f9d79f28eadc9b62a84611981c253123b22a19be4915f1f2bd4cc7beb61f41

Observation 3b857212-b140-46e7-8d81-1baf41fcb6bf · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 54

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T15:51:29.394213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:4f74d81f5925bba47800d870643e61b0eb239158eb4ab9f33170c2500f0f77a3

Observation 457be864-3850-4bbb-a7ce-21edc9aa5793 · outbound

This paper cites European conference on machine learning , pages=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs European conference on machine learning , pages=

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.060463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:dfa7438fb788753a2c69119ae19417b14ccc212fa0c2a0b4dc7ec9a51a73e349

Observation c63381a4-f347-431f-a4b0-132267a5cf76 · outbound

This paper cites Computer , volume=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Computer , volume=

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.022418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:20497a36194e55c7219238b9f861ee6d86185cafe60e875dc9cd26c399ded51a

Observation 863d0472-bdc9-4316-be06-18bafa19f135 · outbound

This paper cites QwQ: Reflect Deeply on the Boundaries of the Unknown , url =.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs QwQ: Reflect Deeply on the Boundaries of the Unknown , url =

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.091675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:8f385eeecb1093b3270d34b9e2299db5c8ed3d429a3867bf1ffe9441a48906e7

Observation 14958ba0-fdb0-431c-9524-5a0e854c9da8 · outbound

This paper cites Qwen2 Technical Report.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Qwen2 Technical Report

Reference 58

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T15:51:29.449430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:5889a6a30672d86cf1fa09de60172c9be30e3844b00a6434cd7e9647e972d274

Observation d4a2473a-819e-4971-862e-6c57eeeafb17 · outbound

This paper cites 2024 , howpublished =.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs 2024 , howpublished =

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.055246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:2322ae7206e5bd0a5c35c47570571f185dc7676d40f67e727718e9b536bc495f

Observation cfe1a6af-393e-4e2e-b8f2-dee6f9727076 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 60

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T15:51:29.231425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:1282205937464cfcf57d26b377b0c520cd0dcbc5295701f2e82bee1287ccad55

Observation acd0e23c-a802-4c3e-889e-062bfdc57522 · outbound

This paper cites 2025 , url=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs 2025 , url=

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.183652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:8eb9c35a8516507a55be541c6ff7af438dd7dfe794bbd77f30de40c8a9d9abe7

Observation aa6368ed-1ee6-433c-a983-9a82f3ba253e · outbound

This paper cites 2005 , publisher=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs 2005 , publisher=

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.276509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:a5ea18ec665798ff89a8e7edf293cc4a8c776ab36b964325255b13c203576672

Observation 2548ef33-0f81-4ba7-8990-79cbfdaf57eb · outbound

This paper cites Occam's razor --- Wikipedia , The Free Encyclopedia.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Occam's razor --- Wikipedia , The Free Encyclopedia

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.132858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:7249ad011a9c5c42607c310805b160bb9fea88cd90b32381bbe6e5c57b1a5174

Observation c8e71b08-3227-4b64-b00a-9808e52c1a6d · outbound

This paper cites Algorithmic information theory --- Wikipedia , The Free Encyclopedia.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Algorithmic information theory --- Wikipedia , The Free Encyclopedia

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.263068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:0f576bd034d301f9ff382dd268be03ea6940a93c832e0915fb22996532ba1c0e

Observation 9040a724-5f41-4517-be7c-a05b91c4894d · outbound

This paper cites Minimum description length --- Wikipedia , The Free Encyclopedia.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Minimum description length --- Wikipedia , The Free Encyclopedia

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.347242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:ad665106d17add519f40a36b4bb5047ae7c24a291a070a6018192bbdd5356788

Observation 5d34ea4e-1cbf-4bde-b045-22057f11c085 · outbound

This paper cites CoRR , year=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs CoRR , year=

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.123571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:ace3ed02c541888db1cd991fc4c9c9ac6433f1996e4e97df07c2747ed6d31b17

Observation 0d748800-820d-457d-abf4-c814ea78cda9 · outbound

This paper cites 2024 , editor =.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs 2024 , editor =

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.316032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:c9b153c85b90f96c4b981ca088fc710a297828fa8ed9b84e1c65c0e06a792f8a

Observation bc4504ed-2426-490d-bac0-788cc9702230 · outbound

This paper cites Toward Self-Improvement of LLMs via Imagination, Searching, and Criticizing.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Toward Self-Improvement of LLMs via Imagination, Searching, and Criticizing

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:51:29.496507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:92348cee8e443321e54f1b863df2119b7589806bff681e7003d95c3040138625

Observation 9d18b4be-ecda-4ecb-abcb-f36efa44f133 · outbound

This paper cites 2024 , eprint=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs 2024 , eprint=

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.209678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:a82359f18890822ff0408793ba61eed0b1cabffb4992c516811b94e64f152a3a

Observation 5a0d80dd-873a-4727-b871-8b48a37d874d · outbound

This paper cites The Twelfth International Conference on Learning Representations , year=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs The Twelfth International Conference on Learning Representations , year=

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.254470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:92a7fa0a942ccba165f52a277f05adf476bd80553d4f8b763c185d5e41ffb416

Observation 114f5dd6-6d22-4571-9e0c-2b118ca0aa12 · outbound

This paper cites Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models

Reference 71

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T06:38:37.536547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:87dcde8dc6a120ca245dd6e5fd3c0ad403f867d1c8f8b10bf0eacb70b3528877

Observation 29985c99-166d-49c3-9ac2-c18dbd9e702c · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 72

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T15:51:29.273092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:49a1b801d8716e88950fdbd0bb9de806a09fc37966041c94de1b6b292268b91f

Observation afdf5663-32e6-4087-a5a3-52335bf41784 · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 73

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T15:51:29.284994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:67fab3acca5ae5044440271295d3ed66f096950da6a86a054ffbcc199f588a4c

Observation 1fb14e3a-9f2b-4f97-9a3d-0421d5036db0 · outbound

This paper cites Making Language Models Better Reasoners with Step-Aware Verifier.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Making Language Models Better Reasoners with Step-Aware Verifier

Reference 74

Resolution
verified exact
doi, observed 2026-05-13T15:51:29.182712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:ede3b51be0411ccdbc02b2854daa9a00cb7e6c1cac18d256f776ad5e2bde95e4

Observation 6f1df464-2883-4c90-9032-7f45bf55b909 · outbound

This paper cites Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.308067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:b7082d82b0c16743bdc9272478a10aadff92b60f0741b2bc1e254443d9ddff01

Observation cfe1276c-e95d-435b-8e78-a67bf88eadf4 · outbound

This paper cites Let's reward step by step: Step-Level reward model as the Navigators for Reasoning.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Let's reward step by step: Step-Level reward model as the Navigators for Reasoning

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:51:29.346680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:b8c4e379465852d4f5c44e26712a5d59886b18f6c2139eace75ac2b44c3c6e75

Observation 36d35aee-61a0-47ff-a0b7-7bfb4e7d1930 · outbound

This paper cites Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.198030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:620695008cc7c3c7601413f026b303caa42336e9936aa1a4b6dfc77144024e07

Observation d2a81a8f-fc16-4ec8-a5f0-e84c9e337a3a · outbound

This paper cites Improve Mathematical Reasoning in Language Models by Automated Process Supervision.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Improve Mathematical Reasoning in Language Models by Automated Process Supervision

Reference 78

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T20:53:45.995574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:b7dc2a1a9d92826e1db8182e3017c95a0c9996ba44d935a26e26813f52424e7b

Observation 5bb0c464-576e-4261-a2fe-4de3dfcd1e71 · outbound

This paper cites Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.358393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:a93a6f2bc00cec3c27d84d65f20e8c8e792c65f93d52d86288264bfee07c1d31

Observation b38d08de-bd9c-4d94-b58e-c1a59bffde19 · outbound

This paper cites C ritic B ench: Benchmarking LLM s for Critique-Correct Reasoning.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs C ritic B ench: Benchmarking LLM s for Critique-Correct Reasoning

Reference 80

Resolution
verified exact
doi, observed 2026-05-13T15:51:29.188352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:9759a8c65f91cd41c3dc6c86e06a20d2d98a36761e651fbcf511fa5f04c75955

Observation 5e73516e-3f9f-4fa4-91f2-75b242b20117 · outbound

This paper cites 2024 , editor =.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs 2024 , editor =

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.343436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:db256518e4cfa54dec597f0c86099c52fac87433b87bcd0de7ca02b4aaeadbcd

Observation 8b48864d-5bf9-4675-9fed-d64c95645dca · outbound

This paper cites Solving math word problems with process- and outcome-based feedback.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Solving math word problems with process- and outcome-based feedback

Reference 82

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T15:51:29.436100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:8dfccfd9e97e85a8dd5c6cbe12461d5b0fb136e33792eeb4316053136947122d

Observation e7ffcc27-666a-477a-b215-50647216c90d · outbound

This paper cites V-STaR: Training Verifiers for Self-Taught Reasoners.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs V-STaR: Training Verifiers for Self-Taught Reasoners

Reference 83

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T15:51:29.464238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:3db58c73652f8ff56725597006aa4c12a48e03ce50d5f7b9b32844fe97d24750

Observation 0e71e76a-94cb-4b5e-8f92-4aafb362f1a6 · outbound

This paper cites 2023 , eprint=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs 2023 , eprint=

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.128219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:06285b83e8462363ee4b56512e39e32d90794555d8defeb614c545c610a0e5f4

Observation 047eaa0a-80d8-49c0-89ce-69bfc9dfa299 · outbound

This paper cites The 4th Workshop on Mathematical Reasoning and AI at NeurIPS'24 , year=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs The 4th Workshop on Mathematical Reasoning and AI at NeurIPS'24 , year=

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.157733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:609e6f0a27cbe18afbc1aac13ba97466d2b8e4f84ba4f20ec1a8737dccd7a832

Observation a3d0e407-6e83-4150-b956-d5847f0b492e · outbound

This paper cites 2023 , url=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs 2023 , url=

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.187395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:c5abeec49a610acf06245fdc61e95c972041610943a5491040b04cf70c26831e

Observation 9113e954-2cad-4381-b676-407c0075bc48 · outbound

This paper cites European conference on machine learning , pages=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs European conference on machine learning , pages=

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:52:57.105326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:891752da18c9b4f198c75e119dc9876fa7a6726d98329d2818530475e81da991

Observation 547c4f5f-6ffd-45f1-9c91-fac1ba3bd1f3 · outbound

This paper cites Step-level Value Preference Optimization for Mathematical Reasoning.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Step-level Value Preference Optimization for Mathematical Reasoning

Reference 88

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T15:51:29.518054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:384bc71ce415b84b8372b25e9f09872dea1a0d52e844ebd7f5b77dee666311a3

Observation 13aaf7dd-b274-48a4-b322-3b491ec8eaf1 · outbound

This paper cites 2023 , html =.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs 2023 , html =

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:51:29.522751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:a294de1de7f2599929b08782673e55ee4640db02f1f39b56bb76f7d6c363d722

Observation 2c35be5d-6b58-4778-ab8a-7cbffc8aa8fd · outbound

This paper cites Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing , pages=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing , pages=

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:51:29.527515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:4d98fac6da89fc98c4a4eff3ee89073d644da9f151fc6ec95ec8eea259a7ef72

Observation eac58ad2-1686-4039-a336-b549b21fd3a9 · outbound

This paper cites Proceedings of the 2024 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers) , pages=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Proceedings of the 2024 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers) , pages=

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:51:29.531775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:fd67309bf2218558b7e89ddc405f637494fab6b0c260931c1f861524de721161

Observation d3372bee-37cc-41de-a8a1-703ccd74582d · outbound

This paper cites Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:51:29.535876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:e56c5837db01f590bee361f91f123eaf4ca96a136e0dad1c4606929d7a277c2f

Observation 8693e1fd-4e00-4f4c-bae0-f54d4c194a2d · outbound

This paper cites Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing , pages=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing , pages=

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:51:29.540222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:73cc4fb456c191363f038207df7bb6bd1891eab56d045c14fc722281e6a9c2b0

Observation 962f8ef8-e6aa-421e-b6a2-b32236d6b4a7 · outbound

This paper cites Journal of Machine Learning Research , volume=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Journal of Machine Learning Research , volume=

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:51:29.544544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:b7e716ef0a56315ff8e8f4a724cb547eb72a75a9e5bcff6e8803ccc21b18b1c2

Observation 02d6c0ff-be42-4875-b49f-8ed641d757af · outbound

This paper cites Rationale-Augmented Ensembles in Language Models.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Rationale-Augmented Ensembles in Language Models

Reference 95

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T15:51:29.303262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:784c52011c0c791f8fdcfd024e20d12a82410263215c5bed11cc0a6bb5f0b28a

Observation 0ff8f2dc-c5b6-4e7f-87be-69bdf58cabf3 · outbound

This paper cites OPT: Open Pre-trained Transformer Language Models.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs OPT: Open Pre-trained Transformer Language Models

Reference 96

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T15:51:29.308740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:af852b211296524fc9922e049dab7ea45044a1bf9da2e695f8ed4430e4cfb946

Observation e4de4033-b5d8-4961-bfc6-2b64425f309e · outbound

This paper cites an unresolved cited work.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Unresolved cited work

Reference 97

Resolution
unresolved
raw_fallback, observed 2026-05-13T15:51:29.549083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:428b0b4005c55b5331baeb099c7f12b7c72c111aecb3b4c2c64102f59ecd56c5

Observation 3eed9caa-d62a-4996-8955-6aaad255b6c0 · outbound

This paper cites The Llama 3 Herd of Models.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs The Llama 3 Herd of Models

Reference 98

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T15:51:29.332870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:1053b7600e883f7bc10dd38bd9bbe4e0e149c90a65c4ddf1849034190b15898a

Observation 432a4108-a021-480a-8b91-6ffc96570c86 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs LLaMA: Open and Efficient Foundation Language Models

Reference 99

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T15:51:29.339442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:79fa61af914d30ce8065732211e16110bdc9c49c1c6098401f5ab7cb4d692bff

Observation fc520a34-3e4b-4b20-93a6-58ae4ff3a287 · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , volume=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Proceedings of the AAAI Conference on Artificial Intelligence , volume=

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:51:29.554101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:c2c431ce2399204ff3fc0d3479f19a4e6df973a30dedbb15f50b03c9e266b7e4

Observation 0100b7b9-3115-495b-a1e6-a5027a97c429 · outbound

This paper cites The Eleventh International Conference on Learning Representations , year=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs The Eleventh International Conference on Learning Representations , year=

Reference 101

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:51:29.558794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:e2488ef924c4e270f71fe50b8958de2964a2356809ed73d4399fa6d8ff9916bc

Observation 8dc2b75a-e0b7-407e-b709-fa0184805202 · outbound

This paper cites International Conference on Machine Learning , pages=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs International Conference on Machine Learning , pages=

Reference 102

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:51:29.563099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:f163d0b13c73a78f9bbe1003e133260fc0eea8307da20209b7238720fcc35e42

Observation c5ceb1d3-fed3-4542-a26e-2cd07ee2b8d1 · outbound

This paper cites Boosted Prompt Ensembles for Large Language Models.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Boosted Prompt Ensembles for Large Language Models

Reference 103

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T15:51:29.382787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:07087b1ae2a0f779773fadb05ed0d75889013eadb977bde7d23ab922c044c173

Observation 9761b210-ae15-4637-a584-93897472d4bf · outbound

This paper cites Findings of the Association for Computational Linguistics: NAACL 2024 , pages=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Findings of the Association for Computational Linguistics: NAACL 2024 , pages=

Reference 104

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:51:29.567585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:1356f4839a887d678f9104957f234b88aa1dbff4dd6c96a067b06d7b2ba90197

Observation 748cca91-3f04-4dbb-b0b7-f1d96ffa009e · outbound

This paper cites The Eleventh International Conference on Learning Representations , year=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs The Eleventh International Conference on Learning Representations , year=

Reference 105

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:51:29.573733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:85307157c31da63a336e6ef844658b7047bdda8d25727c7d7164aa88d6c34f2b

Observation 240c9002-83cc-4348-b1c2-b8b9aac1eb21 · outbound

This paper cites The Eleventh International Conference on Learning Representations , year=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs The Eleventh International Conference on Learning Representations , year=

Reference 106

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:51:29.578779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:5e7c0525cc36ba5a79d211662a0be1cbc92c5ce9130a74d354c023d70b2904c1

Observation 1c35a64b-456f-419b-b021-b26e6d90ccc0 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 107

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T15:51:29.430255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:e5c431148121ff426043361849a0b7d472c846e7492d5d35ce705adf4b6b92ea

Observation e22cacbf-e5e9-4193-be3a-3f426a67ab89 · outbound

This paper cites ArXiv , year=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs ArXiv , year=

Reference 108

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:51:29.583361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:5fa710f68fb57259b04efd8287dad823abba7807e37955ab4dd85d7feccef2c3

Observation 4e8c8842-483f-4032-a463-587fe27780c2 · outbound

This paper cites an unresolved cited work.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Unresolved cited work

Reference 109

Resolution
unresolved
raw_fallback, observed 2026-05-13T15:51:29.587383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:20dc4bbaa55dd7478c4b39d8ce0ebe1ace0bc39176af92c63aa8dbd85a8502b2

Observation 0b43aa93-e4e6-438d-889b-98c032c70d21 · outbound

This paper cites Language Models are Few-Shot Learners , url =.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Language Models are Few-Shot Learners , url =

Reference 110

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:51:29.591945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:00d7042ed2b7ed049441a5fa20e9988ab35da07bbc4dc43033d846a042b62b17

Observation 019c6e5a-9d40-4035-8c5d-7ab5d0bab11c · outbound

This paper cites The Eleventh International Conference on Learning Representations , year=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs The Eleventh International Conference on Learning Representations , year=

Reference 111

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:51:29.595539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:3c22d8decc7ce1d8634901d1fc51991a72ae25f3daaf05300de7fe9116d2d140

Observation 6697f781-2796-4b39-bb25-9755005b6248 · outbound

This paper cites 2020 , eprint=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs 2020 , eprint=

Reference 112

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:51:29.598916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:2b55faa31919abb8c402ac4977a149f3500724e6acaa16cdbfaeca306f7ec430

Observation 5d53c28a-99b2-4957-9038-6c130c5f8852 · outbound

This paper cites Advances in neural information processing systems , volume=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Advances in neural information processing systems , volume=

Reference 113

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:51:29.602692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:c988640a70402dfef5f0385f9942fd011b446660494153c667c299d800a81708

Observation 19bead19-f08d-4ee0-b983-ae7d91540515 · outbound

This paper cites Scaling Language Models: Methods, Analysis & Insights from Training Gopher.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Scaling Language Models: Methods, Analysis & Insights from Training Gopher

Reference 114

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T15:51:29.502469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:8d2ff6c6f48eba61cdf86fb25f0d079a1326336d4111bae363bf32ec4b4535a2

Observation 95d34cb4-cc34-4d1c-b590-c1067d1ebeb8 · outbound

This paper cites Using DeepSpeed and Megatron to Train Megatron-Turing NLG 530B, A Large-Scale Generative Language Model.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs Using DeepSpeed and Megatron to Train Megatron-Turing NLG 530B, A Large-Scale Generative Language Model

Reference 115

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T15:51:29.509284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:71e54fc45328e7f29bf699715c10bdbc5714ccf3ef26a279febc9ea5fc4ff12b

Observation 6903032c-61f8-4146-98db-fadeb0e1547c · outbound

This paper cites The 4th Workshop on Mathematical Reasoning and AI at NeurIPS'24 , year=.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs The 4th Workshop on Mathematical Reasoning and AI at NeurIPS'24 , year=

Reference 116

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T15:51:29.606644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:02e9f336b833e25569a1efc9cca92db3e87c02a923eecc7b18096bcf52b45cfd

Observation dbb066f1-98e7-43fe-bf8a-e72f3c266edb · outbound

This paper cites CritiqueLLM: Towards an Informative Critique Generation Model for Evaluation of Large Language Model Generation.

Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs CritiqueLLM: Towards an Informative Critique Generation Model for Evaluation of Large Language Model Generation

Reference 117

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:51:29.214794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T15:51:29.022336Z digest=sha256:c64e3468a6cb6d858e7c4e99d1c3cb23a81ebcd374f513f98e30f9335aeaaf1f

Pith citing papers

Observation 436c7b94-c4af-4879-ad9e-1465bf410e3b · inbound

Model Merging in LLMs, MLLMs, and Beyond: Methods, Theories, Applications and Opportunities cites this paper.

Model Merging in LLMs, MLLMs, and Beyond: Methods, Theories, Applications and Opportunities Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-17T22:16:04.537901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-17T22:16:04.386706Z digest=sha256:4dfd1b16b3165054fa8555aeca4c8347c5645fb0770d82a8c44abc5579a0f6aa

Observation 44de7fbf-0299-4516-bfcb-a7de09b49d24 · inbound

How Does A Text Preprocessing Pipeline Affect Ontology Matching? cites this paper.

How Does A Text Preprocessing Pipeline Affect Ontology Matching? Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 59

Resolution
verified exact
local_arxiv, observed 2026-05-23T17:15:43.563104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-23T17:15:03.291348Z digest=sha256:eda288d857e4b1e5d17b27f20eb5ce62d2a938b18ebb6e29314f8107921f9932

Observation e9f0f70c-a7dd-4908-904b-d89ee30a0c47 · inbound

Muon is Scalable for LLM Training cites this paper.

Muon is Scalable for LLM Training Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 76

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T15:51:30.429014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-11T23:02:51.656353Z digest=sha256:f67eb7b9faaf5db3041cc1e6a9673f7d5e7d88f5e9391dfc972859e868042ca2

Observation e3bc4185-047a-42bf-ac1c-952489925e9c · inbound

From System 1 to System 2: A Survey of Reasoning Large Language Models cites this paper.

From System 1 to System 2: A Survey of Reasoning Large Language Models Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 124

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:51:30.429014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T01:36:23.845366Z digest=sha256:5a2ffbe8a6e449fad432160d818cad9c19992fce093c7a66c8dff3092e8b1b68

Observation 693e5b07-7cca-48f0-b816-fb3b36404d09 · inbound

L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning cites this paper.

L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T00:19:22.196152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T00:19:22.140009Z digest=sha256:6b00c2a99a2089e6d9bcbb023c73400c9c255e3af3cea0fa291628b3462b547b

Observation 831eaf7d-79f9-45d1-a9da-bcfe9eb204d0 · inbound

Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models cites this paper.

Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 106

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:51:30.429014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T08:40:40.910461Z digest=sha256:b6ae39a2fbf48b847cf4649ad21a7c9bddd6d0a36305486a58568302cf8bee13

Observation 0b736d04-d478-4032-8438-3cfd848fabc1 · inbound

Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models cites this paper.

Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-14T01:29:57.056442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T01:29:56.480020Z digest=sha256:aef30d1cde31c053b1c0b33ecdbe95b0a46066f4efc7420b84098587b408fa0a

Observation b520749e-ca3f-4cc1-b3c4-9a1c407015fe · inbound

A Survey of Scaling in Large Language Model Reasoning cites this paper.

A Survey of Scaling in Large Language Model Reasoning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-22T21:22:09.006698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:35467d7387a93e0d85e95a986d82af8d8c68802a4770ea928a2a88f5ef9b946a

Observation 2d5c5678-3006-4d48-a05d-14ba7ab3a382 · inbound

DeepMath-103K: A Large-Scale, Challenging, Decontaminated, and Verifiable Mathematical Dataset for Advancing Reasoning cites this paper.

DeepMath-103K: A Large-Scale, Challenging, Decontaminated, and Verifiable Mathematical Dataset for Advancing Reasoning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T10:31:04.759798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T10:31:04.728005Z digest=sha256:675f7a044cd076b866eebc8f4c4cc2e1054085fc9fba48e5fe1b8dc0b61ff4db

Observation 18113728-5cd6-4d11-a8cc-ff2aba59d8c7 · inbound

RetroInfer: A Vector Storage Engine for Scalable Long-Context LLM Inference cites this paper.

RetroInfer: A Vector Storage Engine for Scalable Long-Context LLM Inference Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:51:30.429014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T15:59:04.724780Z digest=sha256:6c30c65b68fdf8d7a844b6e418eaf0933037e3a1593c97c9185fb378d791d394

Observation 25037d45-2a3b-4bf7-a7cd-4b3cc4fb16ba · inbound

The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity cites this paper.

The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-05-15T16:10:31.502228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T16:10:31.440921Z digest=sha256:8d70962b8796307254cc56fa3ea80603a55b2c55ce69a01f129bc06f7dddef3e

Observation 0963d3f6-4cb9-4b03-abc0-7e9d40548ba6 · inbound

Do LLMs Overthink Basic Math Reasoning? Benchmarking the Accuracy-Efficiency Tradeoff in Language Models cites this paper.

Do LLMs Overthink Basic Math Reasoning? Benchmarking the Accuracy-Efficiency Tradeoff in Language Models Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-19T06:27:07.266123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-19T06:25:12.799097Z digest=sha256:d8e2aaf6700fab22ce5fd23e720109bfae88bcec5cb04d807448f4f0898d5e49

Observation 5b72617c-b04c-4f3b-9fd0-4dae57f3093f · inbound

What Factors Affect LLMs and RLLMs in Financial Question Answering? cites this paper.

What Factors Affect LLMs and RLLMs in Financial Question Answering? Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-19T05:07:04.480569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T05:06:43.406540Z digest=sha256:e7d41af67e738bef8ff2f3ddea8b45fe75f46333ae5ea43a1bed265f9ffed0c1

Observation 745e2af3-aa0d-479a-8f8f-37d80999a48c · inbound

ReasonCache: Accelerating Large Reasoning Model Serving through KV Cache Sharing cites this paper.

ReasonCache: Accelerating Large Reasoning Model Serving through KV Cache Sharing Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-19T03:17:00.860664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-19T03:14:05.109011Z digest=sha256:2992d34ac1c4e853950d9a12e91ba383cceb29ba9120a3147801fec7d75848ed

Observation cf5af67b-42cc-4d27-80d0-f9153b14c7c4 · inbound

Pruning Long Chain-of-Thought of Large Reasoning Models via Small-Scale Preference Optimization cites this paper.

Pruning Long Chain-of-Thought of Large Reasoning Models via Small-Scale Preference Optimization Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-05-18T22:31:53.133174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T22:26:52.748349Z digest=sha256:b4eda66e5953c9d9e69c3b3aafe6f1e8a3236bc195d8a76e01ede287698f97a1

Observation 8022d966-77f7-4dea-9770-8976a97247bc · inbound

A Survey of Reinforcement Learning for Large Reasoning Models cites this paper.

A Survey of Reinforcement Learning for Large Reasoning Models Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 71

Resolution
verified exact
local_arxiv, observed 2026-05-18T00:02:25.113802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-18T00:02:24.352947Z digest=sha256:32272e758c0ad9038458556efde0f32780693b85a88e3e673fd060061a4eabd2

Observation 90fd3298-5194-42b6-bcda-65084ef934f4 · inbound

GrACE: A Generative Approach to Better Confidence Elicitation and Efficient Test-Time Scaling in Large Language Models cites this paper.

GrACE: A Generative Approach to Better Confidence Elicitation and Efficient Test-Time Scaling in Large Language Models Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-05-18T17:56:42.066613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T17:54:18.515174Z digest=sha256:50e87f1a2b33336f886c400441886b4a48d7de72f7080d1aa874fa7f70316d28

Observation 864194ea-9311-43c9-857d-92a0fb357eca · inbound

Early Stopping Chain-of-thoughts in Large Language Models cites this paper.

Early Stopping Chain-of-thoughts in Large Language Models Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-21T22:34:24.065099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T22:33:03.394914Z digest=sha256:89053cf0eb3f0ff5c2afc92cb8e74a800f0d6fa4255a2fe1a020edfdf6c06d08

Observation a1d07536-eda7-4e9c-b93f-4a5bbdd0edfc · inbound

AdaThink-Med: Optimizing Inference-Time Compute for Medical Reasoning via Uncertainty Quantification cites this paper.

AdaThink-Med: Optimizing Inference-Time Compute for Medical Reasoning via Uncertainty Quantification Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-04T13:51:50.210041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:51:50.210041Z digest=sha256:6a17f9ef011830adaec46166608de8855dc2fe0317efa301ae350543352a8749

Observation ed608508-2dda-4a40-ba25-11a1edd24914 · inbound

Thinking Sparks!: Emergent Attention Heads in Reasoning Models During Post Training cites this paper.

Thinking Sparks!: Emergent Attention Heads in Reasoning Models During Post Training Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:31:24.945425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-18T13:28:32.093512Z digest=sha256:e17aab258dd7961aa671c8ff96dcf83f95732456e976b9fd067e8868ebdaef09

Observation 5431cba1-235a-4696-8b7d-e14820a28d9b · inbound

Entropy After </Think> for reasoning model early exiting cites this paper.

Entropy After </Think> for reasoning model early exiting Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:52:35.593295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:51:58.579048Z digest=sha256:6f6107a5615810b306e9d54c4d7e13c52240e7542a371f2e054a5df5ede1cdac

Observation c1aef93d-48c4-4c08-8cd8-0f00c393fb43 · inbound

Don't Pass@k: A Bayesian Framework for Large Language Model Evaluation cites this paper.

Don't Pass@k: A Bayesian Framework for Large Language Model Evaluation Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-05-18T10:06:13.833638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T10:04:39.223895Z digest=sha256:a22629f269cdee5affe271714188af8f0c9e82c271d6d68ed9360c5f345eac5f

Observation 50a72a1a-5b9f-475a-ba9c-dd249d4f83f4 · inbound

Investigating Thinking Behaviours of Reasoning-Based Language Models for Social Bias Mitigation cites this paper.

Investigating Thinking Behaviours of Reasoning-Based Language Models for Social Bias Mitigation Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T07:01:01.710438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T06:56:33.427852Z digest=sha256:3c7f37ce920c3fb21d11bed9c0cf368788063ae5692af703210c29bf3d59e845

Observation 366c5ef6-a61d-48e7-b371-f3c4ef7245c2 · inbound

When Models Outthink Their Safety: Unveiling and Mitigating Self-Jailbreak in Large Reasoning Models cites this paper.

When Models Outthink Their Safety: Unveiling and Mitigating Self-Jailbreak in Large Reasoning Models Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-18T05:05:55.572601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T05:02:52.176629Z digest=sha256:0c93cf543b3b9198f2c057eaebbfc2cacb87d2f877b34dbb36277e25b23e1864

Observation f1f5401f-2aaa-4042-967c-d4b895f50141 · inbound

TeaRAG: A Token-Efficient Agentic Retrieval-Augmented Generation Framework cites this paper.

TeaRAG: A Token-Efficient Agentic Retrieval-Augmented Generation Framework Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T23:33:43.810560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:33:43.810560Z digest=sha256:c4e21deafdaa924da4801adb1d3a2091933203710c615d7b209f4988a42ceb89

Observation 84552ad8-5377-4238-8945-a525deb40985 · inbound

Can Large Language Models Reason About Complex Execution Paths? An Empirical Study on Python cites this paper.

Can Large Language Models Reason About Complex Execution Paths? An Empirical Study on Python Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T20:51:06.005534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:51:06.005534Z digest=sha256:c35a706c535dee1c701e88a364635381d1d542c8edc6199bd95390e32d18c327

Observation ac5ae07d-7a07-4554-bcb2-15a2561fdcd8 · inbound

Thinking-Based Non-Thinking: Solving the Reward Hacking Problem in Training Hybrid Reasoning Models via Reinforcement Learning cites this paper.

Thinking-Based Non-Thinking: Solving the Reward Hacking Problem in Training Hybrid Reasoning Models via Reinforcement Learning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-03T11:57:55.221525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T11:57:55.221525Z digest=sha256:419b8f97c32130e155c56982013db0cdb3c35eff962afa827a8f59c57103e6a1

Observation 2e3fb94c-dff1-4e25-826c-9e4c31c1ab45 · inbound

Mid-Think: Training-Free Intermediate-Budget Reasoning via Token-Level Triggers cites this paper.

Mid-Think: Training-Free Intermediate-Budget Reasoning via Token-Level Triggers Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T11:16:00.357318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T11:16:00.357318Z digest=sha256:30338034657aed8b94dfb3c9851049a9ee2948bb7b0f9691b6601ace1367495b

Observation 75a2e224-6419-4b46-bc06-804320ec30b6 · inbound

ConPress: Learning Efficient Reasoning from Multi-Question Contextual Pressure cites this paper.

ConPress: Learning Efficient Reasoning from Multi-Question Contextual Pressure Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T05:43:52.986454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:43:52.986454Z digest=sha256:b2ad739c173910dc86f53aeed34228a00fadf5eccd045ba448f7910fd89037ae

Observation e64b3f27-fee6-4d19-b736-82243a0f6af2 · inbound

Towards Efficient Large Language Reasoning Models via Extreme-Ratio Chain-of-Thought Compression cites this paper.

Towards Efficient Large Language Reasoning Models via Extreme-Ratio Chain-of-Thought Compression Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-21T13:44:11.397533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T13:43:51.127429Z digest=sha256:385a26eb6cac99b9ab0fc01dbb0bcff7fe1daed29753730fdfacb4b2c5abae05

Observation 6f823bb8-2644-4fb1-b4b4-beb4b53345dd · inbound

Towards Efficient Large Language Reasoning Models via Extreme-Ratio Chain-of-Thought Compression cites this paper.

Towards Efficient Large Language Reasoning Models via Extreme-Ratio Chain-of-Thought Compression Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T03:25:21.948273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:25:21.948273Z digest=sha256:3e0f6d41a8af85dc581b97690b56e1a21d709cbe60ba2b6291e3ce93a0e72df3

Observation e7a0d7e1-dc7c-4f37-a560-d39c6e4da41c · inbound

Reward Modeling for Reinforcement Learning-Based LLM Reasoning: Design, Challenges, and Evaluation cites this paper.

Reward Modeling for Reinforcement Learning-Based LLM Reasoning: Design, Challenges, and Evaluation Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T03:04:43.464508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:04:43.464508Z digest=sha256:9513aa9b22c8ea2761096006d0da806c31559ce711626fc4fad54da277f7fca4

Observation a1a3dee6-b9d1-4a8f-9ff3-6cb51e75f1d4 · inbound

Compress the Easy, Explore the Hard: Difficulty-Aware Entropy Regularization for Efficient LLM Reasoning cites this paper.

Compress the Easy, Explore the Hard: Difficulty-Aware Entropy Regularization for Efficient LLM Reasoning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T20:44:43.851066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:44:43.851066Z digest=sha256:cf396d9e37d98b4fd99e127e093ded8e34231f9f485d28eed0803a2fc42e0aa0

Observation 2f4b4cd8-2414-4cb1-9810-d7b6d3b87201 · inbound

SHAPE: Stage-aware Hierarchical Advantage via Potential Estimation for LLM Reasoning cites this paper.

SHAPE: Stage-aware Hierarchical Advantage via Potential Estimation for LLM Reasoning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T15:51:30.429014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T19:14:10.609406Z digest=sha256:4e692f36503f76d08ab144cba5eb29f311f8d9455344eb77f448b44838a03dc3

Observation 5e3b2393-90a8-4ea3-84f6-6ad3824b90c3 · inbound

HiRO-Nav: Hybrid ReasOning Enables Efficient Embodied Navigation cites this paper.

HiRO-Nav: Hybrid ReasOning Enables Efficient Embodied Navigation Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:51:30.429014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:01:13.350219Z digest=sha256:2833797d77d3b066cc5c63b80e20493a257c71506eb2a812d7e396b599b6a275

Observation fd1cdbc2-6549-4d00-a72c-22b41e9da700 · inbound

From $P(y|x)$ to $P(y)$: Investigating Reinforcement Learning in Pre-train Space cites this paper.

From $P(y|x)$ to $P(y)$: Investigating Reinforcement Learning in Pre-train Space Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:51:30.429014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-10T12:50:57.603403Z digest=sha256:10109d8d97f3d89b637da8457768d5f3402b2c99803a2bab34b52ca7899b01a0

Observation 90538cf9-5209-448f-8a3d-92f55f7635f2 · inbound

When LLMs Stop Following Steps: A Diagnostic Study of Procedural Execution in Language Models cites this paper.

When LLMs Stop Following Steps: A Diagnostic Study of Procedural Execution in Language Models Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T15:51:30.429014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-09T18:51:59.402906Z digest=sha256:bae6837e2d96051f2b2978115c5fbe85b2c420692a79175d8ea07105809152af

Observation 3eff0351-bd3b-47cd-91b5-f7bc04aea742 · inbound

Free Energy-Driven Reinforcement Learning with Adaptive Advantage Shaping for Unsupervised Reasoning in LLMs cites this paper.

Free Energy-Driven Reinforcement Learning with Adaptive Advantage Shaping for Unsupervised Reasoning in LLMs Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 267

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T15:51:30.429014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-10T16:58:10.013475Z digest=sha256:b2d572025eafb21cb1e3f4b26c3d63d23995606c5d69638871b175ca1928376a

Observation f1348b64-79f3-43dc-a37e-392736cb957c · inbound

Adapt to Thrive! Adaptive Power-Mean Policy Optimization for Improved LLM Reasoning cites this paper.

Adapt to Thrive! Adaptive Power-Mean Policy Optimization for Improved LLM Reasoning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 252

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T15:51:30.429014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-10T16:51:19.555272Z digest=sha256:3e7aab0d19f2e8ba8252a81bb8660ebac933dc597dca0ac0c26d4f4645713545

Observation b26bdaa9-614d-41dc-9a8d-e555650010ef · inbound

Decodable but Not Corrected by Fixed Residual-Stream Linear Steering: Evidence from Medical LLM Failure Regimes cites this paper.

Decodable but Not Corrected by Fixed Residual-Stream Linear Steering: Evidence from Medical LLM Failure Regimes Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:51:30.429014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T11:42:16.090162Z digest=sha256:b3a9f09bb28aa27d414a8de42684481804a5feb1604e2775b9ef99682e4a6bb5

Observation d2f50608-8daf-48c3-a6e5-47efe4965920 · inbound

Post Reasoning: Improving the Performance of Non-Thinking Models at No Cost cites this paper.

Post Reasoning: Improving the Performance of Non-Thinking Models at No Cost Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 120

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T15:51:30.429014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T10:19:08.451445Z digest=sha256:745aa6c76a58156a18782172bdca953f3687d00dedd872360587a9767c2362db

Observation 73f16b18-c3a0-4fc1-a695-7f2e66c9d171 · inbound

How Well Do LLMs Perform on the Simplest Long-Chain Reasoning Tasks: An Empirical Study on the Equivalence Class Problem cites this paper.

How Well Do LLMs Perform on the Simplest Long-Chain Reasoning Tasks: An Empirical Study on the Equivalence Class Problem Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:51:30.429014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-11T01:29:33.453354Z digest=sha256:e1865e45e87625512d8ef51051583abed46ddfd251ee0efb8969d2380498ef55

Observation 5e516f3c-ad6d-4590-99d6-7dcaf053f4fe · inbound

Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training cites this paper.

Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:51:30.429014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-11T01:12:20.864362Z digest=sha256:db54999ab76c0a2a0ecfb4fedeee72645c1903b1be2cbaa07bfd2ad73d446fd0

Observation c2209b53-a6bc-4a2d-a730-e83e623e6f39 · inbound

Hint Tuning: Less Data Makes Better Reasoners cites this paper.

Hint Tuning: Less Data Makes Better Reasoners Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:51:30.429014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T00:54:13.146373Z digest=sha256:44410f2713d90994bb4dd8d26dc4ddb6bfbe93ec656323675f962e052079c035

Observation 6bca4fc8-d654-42fa-b39b-c3fc9b1aabcf · inbound

Hint Tuning: Less Data Makes Better Reasoners cites this paper.

Hint Tuning: Less Data Makes Better Reasoners Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-06-30T23:35:07.066238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T23:34:43.785312Z digest=sha256:cbc764d748fa5c32ff8e1c04dbd74b07c99afe060c3a946039619d888391a666

Observation 05e092c1-6179-4af5-88eb-c8fd5ad07e34 · inbound

Reasoning Compression with Mixed-Policy Distillation cites this paper.

Reasoning Compression with Mixed-Policy Distillation Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:51:30.429014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T01:29:02.387691Z digest=sha256:ca548758adb3fc2e063f61feeccbe871fe28852f63ed774f9cb6dce10c080fdc

Observation 3cfffa00-abe6-4b12-b40f-8abed7bd1180 · inbound

LEAD: Length-Efficient Adaptive and Dynamic Reasoning for Large Language Models cites this paper.

LEAD: Length-Efficient Adaptive and Dynamic Reasoning for Large Language Models Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:51:30.429014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T02:30:42.407934Z digest=sha256:f495880903349837a39806b40aa2cf21cad98e0e813e15343dee569a51604c79

Observation 38cb4bb4-9401-4c8b-b7de-372479f51b60 · inbound

The Gordian Knot for VLMs: Diagrammatic Knot Reasoning as a Hard Benchmark cites this paper.

The Gordian Knot for VLMs: Diagrammatic Knot Reasoning as a Hard Benchmark Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T15:51:30.429014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T04:38:31.924896Z digest=sha256:3f9c0b9541461ba37a57d50d73e2f894edfdd6e689e2cec857b391f4209296b5

Observation ef4e279b-e0cd-407a-a133-65916b51379a · inbound

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration cites this paper.

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:51:30.429014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T03:51:52.375703Z digest=sha256:7f4c50c3b24af520b246bc7cf2c24ed8aa514f139d4f906bfe23fbf622c9ebc4

Observation 0e28c31d-9d9f-4256-9987-a2f959d1d642 · inbound

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration cites this paper.

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-15T05:15:03.254338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T05:11:32.053440Z digest=sha256:d60a617a5b5a3fdca4f75ab616aa6e4a6de2419eb65be394d9840c92ceebe850

Observation f78d95bb-3636-4e04-b5e2-0210b358f48c · inbound

Efficient LLM Reasoning via Variational Posterior Guidance with Efficiency Awareness cites this paper.

Efficient LLM Reasoning via Variational Posterior Guidance with Efficiency Awareness Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:51:30.429014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T06:28:06.957054Z digest=sha256:ee443e443e438e22777e353a4cb74b6c78b77f2fcfd9ef8ab5a1bb45c516acf8

Observation b29a99f7-e1db-4f31-8a66-c9ff017c51fa · inbound

Nice Fold or Hero Call: Learning Budget-Efficient Thinking for Adaptive Reasoning cites this paper.

Nice Fold or Hero Call: Learning Budget-Efficient Thinking for Adaptive Reasoning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:51:30.429014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T01:31:21.377700Z digest=sha256:ddd919c9d5bbde733c6aea6f0437f1d476c4367be0dac939ece14a98473ed57a

Observation 5926c518-4292-405d-81c1-9830725be20b · inbound

Taming the Thinker: Conditional Entropy Shaping for Adaptive LLM Reasoning cites this paper.

Taming the Thinker: Conditional Entropy Shaping for Adaptive LLM Reasoning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T06:43:05.949277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-20T06:40:06.103206Z digest=sha256:e1e4311276485f10addd192b7d9689514250459a2000534855f562785a2d2e28

Observation 9aee66c1-e371-4964-b4fe-f8b74419ad36 · inbound

Efficient Agentic Reasoning Through Self-Regulated Simulative Planning cites this paper.

Efficient Agentic Reasoning Through Self-Regulated Simulative Planning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-22T06:34:40.919640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T06:33:36.846345Z digest=sha256:4b7c0d616db730b84a7c2fe67a31017da6038a6d984d0c8e58741a2af36f1b78

Observation 0a02b1c5-b8fe-49a1-9680-69799799f52b · inbound

How Much Thinking is Enough? Quantifying and Understanding Redundancy in LLM Reasoning cites this paper.

How Much Thinking is Enough? Quantifying and Understanding Redundancy in LLM Reasoning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-05T10:20:57.258709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-05T10:18:18.717871Z digest=sha256:6e141078a6dc775d1e0c18ed607731a5bb29cd820f19af81f877a77d31750ab4

Observation 8060d12e-c67e-492b-9a55-6fa94c59d864 · inbound

SLAT: Segment-Level Adaptive Trimming for Efficient CoT Reasoning cites this paper.

SLAT: Segment-Level Adaptive Trimming for Efficient CoT Reasoning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-06-28T22:32:44.410660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T22:27:30.783923Z digest=sha256:03c6500adbee51dd3cec556f5d2788f47c479be56635f569c74a37c082726bbf

Observation 4e6f284c-6585-479a-968d-43f74d219917 · inbound

What Am I Missing? Question-Answering as Hidden State Probing cites this paper.

What Am I Missing? Question-Answering as Hidden State Probing Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:36:08.870161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T22:17:50.790267Z digest=sha256:1f054690aa174b8e23f60a782ee43867dc688dee6c1001de1b43cc8093594ac9

Observation 26018240-5578-4e66-b3c4-cf99b11cc970 · inbound

Quantized Reasoning Models Think They Need to Think Longer, but They Do Not cites this paper.

Quantized Reasoning Models Think They Need to Think Longer, but They Do Not Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 49

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T19:16:00.131404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-28T23:05:00.401365Z digest=sha256:99c295964e0ef416ee1e0d320faeb5de1e84d69db91e778f15928ff43aec9ecb

Observation ee435a10-b645-4731-adeb-32302b7fd11c · inbound

Thinking Economically: A Hierarchical Framework for Adaptive-Complexity Reasoning in LLMs cites this paper.

Thinking Economically: A Hierarchical Framework for Adaptive-Complexity Reasoning in LLMs Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-06-28T17:12:25.278472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-28T17:05:48.244094Z digest=sha256:d4fe7ade11babfc02b2fdbb2f691f31df72747076d54715d2940a43a7f5f6e7f

Observation eb538a92-71c9-45b3-9a8d-57e266fb8f69 · inbound

Rethinking the Role of Positional Encoding: Sliding-Window Transformers without PE Remain Turing Complete cites this paper.

Rethinking the Role of Positional Encoding: Sliding-Window Transformers without PE Remain Turing Complete Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 84

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T22:06:16.216462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-28T15:48:48.046003Z digest=sha256:0dfc83f782677649ca86b7f27478e34b58166f05fad69947172e4f3e74c66cc3

Observation ecf39b90-91c5-446d-8494-7c299ce59999 · inbound

Adaptive Latent Agentic Reasoning cites this paper.

Adaptive Latent Agentic Reasoning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 30

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T23:26:22.338374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-28T14:24:22.855486Z digest=sha256:ace5c9e9e867d9b927b594c79cc80d794cb3fe89111cccb9408cfec4e6d253c7

Observation a08c83bf-c2b8-453c-9c63-8599e7c17ba4 · inbound

ThoughtFold: Folding Reasoning Chains via Introspective Preference Learning cites this paper.

ThoughtFold: Folding Reasoning Chains via Introspective Preference Learning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T03:26:28.695641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T10:07:16.700499Z digest=sha256:0a9f6a911a62b8a6ebc453c68d838f0a504a6977ce5a8fc78f6b87daec60157d

Observation 6ce6c997-8adf-4283-8de5-16114529662d · inbound

DyCon: Dynamic Reasoning Control via Evolving Difficulty Modeling cites this paper.

DyCon: Dynamic Reasoning Control via Evolving Difficulty Modeling Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-02T16:57:10.065388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-27T22:16:01.457276Z digest=sha256:576c969e695bc603f2f616bc62144151ff5158b57bd22a37589327ff35946eb2

Observation af67ccb2-de81-46b9-9367-4f965dee2995 · inbound

KCSAT-ML: Probing Reasoning Models with Nationwide-Cohort Human Difficulty cites this paper.

KCSAT-ML: Probing Reasoning Models with Nationwide-Cohort Human Difficulty Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-03T05:47:40.988008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-27T13:07:59.509660Z digest=sha256:dde443e8c99acc3475bcb05858fece66fc348e186f0ca1a106a57195873332af

Observation beb068ec-e980-48a5-80ec-7047440f825e · inbound

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes cites this paper.

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-06-27T13:00:56.032249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-27T12:59:51.091008Z digest=sha256:d8c73424cbf6abb767608ca76637ba384bf46aaceceb60cbd036326c67529b83

Observation 862276b2-b364-45f8-b7ca-cb77b510bef0 · inbound

Demystifying Hidden-State Recurrence: Switchable Latent Reasoning with On-Policy Reinforcement Learning cites this paper.

Demystifying Hidden-State Recurrence: Switchable Latent Reasoning with On-Policy Reinforcement Learning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-07-03T13:58:22.094525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-27T07:20:38.534666Z digest=sha256:3432c83d764c4036684587b53c0c688d233eb390860c1c3655ebf9d0c79288cf

Observation 663f9e9e-441c-44fb-bc1b-568910edaadf · inbound

ReSum: Synergizing LLM Reasoning and Summarization with Reinforcement Learning cites this paper.

ReSum: Synergizing LLM Reasoning and Summarization with Reinforcement Learning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-07-03T15:08:33.438834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-27T06:39:34.199607Z digest=sha256:f1e4b6ac68b6b5976253f699703045d32ea6a477a0712ababe29dcef405b2eeb

Observation abc466d0-f8ee-44c8-9f70-7163e6e20f7f · inbound

ReSum: Synergizing LLM Reasoning and Summarization with Reinforcement Learning cites this paper.

ReSum: Synergizing LLM Reasoning and Summarization with Reinforcement Learning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T02:12:17.134825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:12:17.134825Z digest=sha256:4156030b87b0cd07d0520d9e0f41953a56171d7ce36d55e575cb481f63ad4781

Observation 5007a34d-e995-4c72-96a1-62b99cb59e6d · inbound

Human vs Machine Mathematical Difficulty on Project Euler: An Experimental Analysis cites this paper.

Human vs Machine Mathematical Difficulty on Project Euler: An Experimental Analysis Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-04T08:19:44.166279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-26T11:55:19.015777Z digest=sha256:c3cf2e14a41bd77ad243e5dcd3ca5adc2ab5a8f4447014a858a273bcb724d390

Observation 90c68f56-1964-47bd-a181-93b7ac64475b · inbound

Cognitive Episodes in LLM Reasoning Traces Enable Interpretable Human Item Difficulty Prediction cites this paper.

Cognitive Episodes in LLM Reasoning Traces Enable Interpretable Human Item Difficulty Prediction Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-07-01T17:15:51.572709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-29T03:58:32.372896Z digest=sha256:572dced38c7f13c477cc170185a8895995d394b9703ebf5c063a67de07bc7bd7

Observation f5008faa-62ec-443d-aff8-1c480a625916 · inbound

Cognitive Episodes in LLM Reasoning Traces Enable Interpretable Human Item Difficulty Prediction cites this paper.

Cognitive Episodes in LLM Reasoning Traces Enable Interpretable Human Item Difficulty Prediction Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-14T17:12:24.565155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T17:12:24.565155Z digest=sha256:9b0521cc041942103d8f214d4a8af597be064fa487fa17d369909f8743633e75

Observation 78606c6c-f03a-4bcf-a8b0-f7e1691ad09f · inbound

LASER: Load-Aware Serving with Early-Exit for Reasoning LLMs at the Edge cites this paper.

LASER: Load-Aware Serving with Early-Exit for Reasoning LLMs at the Edge Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T11:55:43.669264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-01T03:07:28.653313Z digest=sha256:486d36986956ef1266abc5675961ad3698d3fb84744d98fd9f4c899e8ab28eb3

Observation 7d675b7d-5fdd-4051-a058-69c98975c14e · inbound

Addressing Over-Refusal in LLMs with Competing Rewards cites this paper.

Addressing Over-Refusal in LLMs with Competing Rewards Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 56

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T08:55:35.573638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-07-01T06:59:12.695984Z digest=sha256:9f386745c479e75f317981fbda6beb36025c8bbba4f4d0cc9bdb7f305cbcfa9a

Observation e8397435-0acb-47ed-8348-42d1edf95c38 · inbound

Know When to Stop: Segment-Level Credit Assignment for Reducing Overthinking cites this paper.

Know When to Stop: Segment-Level Credit Assignment for Reducing Overthinking Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-07-02T13:36:58.565666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-07-02T13:31:40.676016Z digest=sha256:af3675ed92d3ea72f092ce87d95739667e07761ff06aa4bfc5f8e7eb79f8e33c

Observation c166bee5-0ec9-4610-8be8-1f24edbb36ac · inbound

CAT: Confidence-Adaptive Thinking for Efficient Reasoning of Large Reasoning Models cites this paper.

CAT: Confidence-Adaptive Thinking for Efficient Reasoning of Large Reasoning Models Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-02T13:26:58.150947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-07-02T13:22:27.566432Z digest=sha256:3d23b30146919717dbcc002f36428bdc8c0921f3f6ccf9998ea84e3f97934e70

Observation 3d315fa4-5780-4b62-893f-c98aa9ba972c · inbound

Overthink-Triggered Slowdown Attacks on LVLM-Based Robotic Systems cites this paper.

Overthink-Triggered Slowdown Attacks on LVLM-Based Robotic Systems Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-07-03T19:58:53.670189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-03T19:52:11.018335Z digest=sha256:370747b6997be642dbc4e778c6809e0d212bdbb3eb737b8ad04a9c1738dc022e

Observation 02bc67ff-f3b3-4b93-a3ed-52f0f843a441 · inbound

Multi-Turn On-Policy Distillation with Prefix Replay cites this paper.

Multi-Turn On-Policy Distillation with Prefix Replay Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 152

Resolution
unresolved
no resolver link, observed 2026-07-11T13:53:36.775836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T13:53:36.775836Z digest=sha256:931f6f14fbdccf8a331ef2494793397c4a9067994a8c2d5758af60f9c4a58c40

Observation eb8dfb3c-3efb-4997-bc0d-9d7dec82d343 · inbound

Multi-Turn On-Policy Distillation with Prefix Replay cites this paper.

Multi-Turn On-Policy Distillation with Prefix Replay Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 153

Resolution
unresolved
no resolver link, observed 2026-08-02T08:40:49.250422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T08:40:49.250422Z digest=sha256:4b29f78e313d5695119b91a1a2d9990538314648ae23006a5810f4e2e469e471

Observation e4a74322-5fed-487d-b959-b5ec55bedc0d · inbound

A Temporal Reasoning Benchmarking Framework for LRMs via Difficulty-controlled and Dynamic Test Generation cites this paper.

A Temporal Reasoning Benchmarking Framework for LRMs via Difficulty-controlled and Dynamic Test Generation Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-11T13:34:00.838394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T13:34:00.838394Z digest=sha256:c7b39602cfa2fc595bf0fe7ceb2913477528d05ffc4a9a8b46673c8dad4f70dd

Observation ce115964-a637-4993-afd0-b9a35b5529d0 · inbound

Recursive Self-Improvement in AI: From Bounded Self-Refinement to Autonomous Research Loops cites this paper.

Recursive Self-Improvement in AI: From Bounded Self-Refinement to Autonomous Research Loops Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 105

Resolution
verified exact
local_arxiv, observed 2026-07-09T03:45:55.671091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-09T03:36:57.168246Z digest=sha256:0eb8feb6024960a856587b92c38c332b5f2b44a74005cab2a43dc6ed08ba3fdf

Observation 8b4461bc-13ee-49a2-89a6-2bda15b70083 · inbound

Agon: Competitive Cross-Model RL with Implicit Rival Grading of Reasoning cites this paper.

Agon: Competitive Cross-Model RL with Implicit Rival Grading of Reasoning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-07-09T02:35:53.856386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-09T02:29:12.366018Z digest=sha256:4150c3a86e0ca3d9a6abe113e225ab5127f3d475c60fd3ba4212c0f745b54db1

Observation 3b3d824b-072e-4093-bcfc-d3eace0ddefc · inbound

A First-Principles Theory of Slow Thinking and Active Perception cites this paper.

A First-Principles Theory of Slow Thinking and Active Perception Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-07-10T11:37:03.273952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-10T11:32:24.374377Z digest=sha256:c088f2cd1ec52469089eb7ec892b1f6f778a89cc025d0c8dfe5c82a11360a7ae

Observation adc877ab-5072-4981-9a05-dded96d433c4 · inbound

Different Teachers, Different Capabilities: Sub-1B On-Device Distillation for Structured Text Enrichment cites this paper.

Different Teachers, Different Capabilities: Sub-1B On-Device Distillation for Structured Text Enrichment Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-07-10T10:27:02.470943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-10T10:20:46.103409Z digest=sha256:58b9b2d4f9f453a8bb16daf7ac21cf1e5afef0442bc2a9e2f9c27332e7f52a2f

Observation f26eff8c-d703-4335-b55c-57b0b5d6f5ae · inbound

Length Penalties Make Chain-of-Thought Less Monitorable cites this paper.

Length Penalties Make Chain-of-Thought Less Monitorable Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 51

Resolution
unresolved
no resolver link, observed 2026-07-14T15:45:54.532529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T15:45:54.532529Z digest=sha256:6a62052b6478a8247d069b092d30fe2642ccf773294a34f5dc6979ae8bce796f

Observation f0eede3b-42d2-458f-aaa4-5aef9c0bb16d · inbound

Length Penalties Make Chain-of-Thought Less Monitorable cites this paper.

Length Penalties Make Chain-of-Thought Less Monitorable Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-02T08:06:10.770248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T08:06:10.770248Z digest=sha256:97a0024a96bc0c984f3c449ea5cf91b26f389ccb3eba8dba7e4e72372f4f4b53

Observation 3efc6f17-67e3-4b1e-a489-4700e721431e · inbound

Length Penalties Make Chain-of-Thought Less Monitorable cites this paper.

Length Penalties Make Chain-of-Thought Less Monitorable Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T04:30:27.191039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T04:30:27.191039Z digest=sha256:f180b2faac1e37691a8e6ca48563540b6a7f2cb33fd1f1b06e7a955ed6144d15

Observation 6179f191-9e5f-4067-9111-a0c035bd03e0 · inbound

Deep Interaction: An Efficient Human-AI Interaction Method for Large Reasoning Models cites this paper.

Deep Interaction: An Efficient Human-AI Interaction Method for Large Reasoning Models Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T03:01:40.938291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:01:40.938291Z digest=sha256:94e3fdef04f92ea1db055c9164b4ab0f2ef9cc5f8d1c59206b97d2f68ec4d26b

Observation df3ad554-d3a8-4833-97ba-d1865c1704a3 · inbound

Constrained Path Reasoning: Measuring When Committed Stages Earn Their Cost cites this paper.

Constrained Path Reasoning: Measuring When Committed Stages Earn Their Cost Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-01T18:43:47.572850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:43:47.572850Z digest=sha256:f400d9c26e6658dcfd23af3ebf33fe83197cf2bf13cae7cdda63f8fe102de7f6

Observation ee97af5e-1e52-49f2-9ed2-181159fc33f9 · inbound

Copy Less, Ground More: Overcoming Repetitive Copying in Long-Context Reasoning via Evidence-Aware Reinforcement Learning cites this paper.

Copy Less, Ground More: Overcoming Repetitive Copying in Long-Context Reasoning via Evidence-Aware Reinforcement Learning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-01T12:46:02.067733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:46:02.067733Z digest=sha256:8046be55477456f6f4d66f6e8de68ceb6ae55e2bed6b8eadac114c9bcf5c0dab

Observation b9247485-7d83-4127-8e3d-a6f5433680ab · inbound

Copy Less, Ground More: Overcoming Repetitive Copying in Long-Context Reasoning via Evidence-Aware Reinforcement Learning cites this paper.

Copy Less, Ground More: Overcoming Repetitive Copying in Long-Context Reasoning via Evidence-Aware Reinforcement Learning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-03T01:57:52.525886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:57:52.525886Z digest=sha256:0036570003040eae6f96b51d2427eac26656950e79522eee2de4e94199645908

Observation 4682635f-18ea-4e05-af52-36b22b0d716f · inbound

EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization cites this paper.

EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-01T11:17:40.200772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:17:40.200772Z digest=sha256:87fc3bcf19afedc45e964fbe81ef9e4e44e4e9672f177d517bf7f15b81b26244

Observation 6d8d6272-b4ec-4c35-b0ca-259c2c3fc1f2 · inbound

Test-Time Scaling via Error Localization cites this paper.

Test-Time Scaling via Error Localization Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-01T07:28:22.862807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:28:22.862807Z digest=sha256:19479df63aac9627da080efe326cb3d55563a797f34b87c2e77a04f1950d1b2d

Observation 45c9b49a-fb37-4dc3-a9ec-b206294a963a · inbound

Reasoning or Memorization: Can LLMs Understand and Generate Chinese Xiehouyu Riddles? cites this paper.

Reasoning or Memorization: Can LLMs Understand and Generate Chinese Xiehouyu Riddles? Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 66

Resolution
unresolved
no resolver link, observed 2026-07-30T22:05:20.643617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-30T22:05:20.643617Z digest=sha256:2b128c3d4119abc26dea8dff0c514070556cc4697203a551c09303fcca7bf295

Observation a1071cd9-3170-4130-b6c8-a714a4d03a1a · inbound

BLADE: Boundary-Expanded and Layer-Adaptive Dynamic Exit for Efficient LLM Reasoning cites this paper.

BLADE: Boundary-Expanded and Layer-Adaptive Dynamic Exit for Efficient LLM Reasoning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:38.604879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:38.604879Z digest=sha256:56992b4c56ecbadc38c5d9caa7aeb0869bba0c8b05439c6ad57c9e974f197327

Observation 4d07f6bb-1c7a-47b5-a714-0b7a0317d413 · inbound

Knowing When to Quit: Diagnosing and Training LLMs to Abort Futile Reasoning cites this paper.

Knowing When to Quit: Diagnosing and Training LLMs to Abort Futile Reasoning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-03T11:36:49.121574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T11:36:49.121574Z digest=sha256:91dc23a32071ff1b537fad94c8ebd920a872fc660b65ca4f48a546792c01611d