Pith. sign in

Paper Citation Record · LEDGER

Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 48 inbound Pith citation observations for arXiv:2408.06195.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2408.06195 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 48 of 48 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T22:37:07.216959Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

3
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5f060018-ae75-420e-b3c4-0b743773ab45 · inbound

Enhancing Reasoning through Process Supervision with Monte Carlo Tree Search cites this paper.

Enhancing Reasoning through Process Supervision with Monte Carlo Tree Search Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T22:37:07.216959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:37:07.216959Z digest=sha256:db1fc4c317940cc87167752682c11cceb44be7b3dbe916414ad2a9d976d2f399

Observation 9052ef75-c24a-4c31-8ea2-2a9a9c835db6 · inbound

Monte Carlo Tree Search for Comprehensive Exploration in LLM-Based Automatic Heuristic Design cites this paper.

Monte Carlo Tree Search for Comprehensive Exploration in LLM-Based Automatic Heuristic Design Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:29.365118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T20:26:29.365118Z digest=sha256:8af6b3ffa729a6a294f94c458c34561bb20ac487eae8ebf781682b263f01ee81

Observation 17cb9031-d2af-4fc8-9172-818cf15e232a · inbound

AirRAG: Autonomous Strategic Planning and Reasoning Steer Retrieval Augmented Generation cites this paper.

AirRAG: Autonomous Strategic Planning and Reasoning Steer Retrieval Augmented Generation Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:46.293230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T19:27:46.293230Z digest=sha256:f11931be0faab283da09e6e140b772ada4a3f1d6569cd0e327fcba13b7c4f54f

Observation ab5d8d34-3358-42f0-b686-3791ed038cdb · inbound

Domaino1s: Guiding LLM Reasoning for Explainable Answers in High-Stakes Domains cites this paper.

Domaino1s: Guiding LLM Reasoning for Explainable Answers in High-Stakes Domains Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T15:13:43.276912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:13:43.276912Z digest=sha256:297fdd5306106155220ca8c4eed9bd987d7ca8e233c92101dcc35983df6d028d

Observation 5f881308-29a5-4563-ba0e-b23af1b6e46f · inbound

Scaling Inference-Efficient Language Models cites this paper.

Scaling Inference-Efficient Language Models Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-10T00:43:29.080074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:43:29.080074Z digest=sha256:f2196d3ab874dbf95a014fff5573c9ef8119959c0b096a63a3b47bda3857346b

Observation 7af0daca-458b-4d09-bdab-3bd5a8633d16 · inbound

Reward-Guided Speculative Decoding for Efficient LLM Reasoning cites this paper.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.286290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.286290Z digest=sha256:61160be321d7bd342c81380dfae118e1478e672c588426d49d29f26781159f1e

Observation a8f88da3-8542-46a5-a734-23bf4496ad72 · inbound

Efficient Multi-Agent System Training with Data Influence-Oriented Tree Search cites this paper.

Efficient Multi-Agent System Training with Data Influence-Oriented Tree Search Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:07:30.522145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-23T04:06:23.521344Z digest=sha256:d508f24ad26e64085e339abdf80c0352469c5c9635746ba454df8c6f74b065ee

Observation ca61eb43-cea9-453d-a8af-06a79f3f8914 · inbound

LongDPO: Unlock Better Long-form Generation Abilities for LLMs via Critique-augmented Stepwise Information cites this paper.

LongDPO: Unlock Better Long-form Generation Abilities for LLMs via Critique-augmented Stepwise Information Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-09T13:25:52.058726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T13:25:52.058726Z digest=sha256:339624943a1a89362cc77aaa2931409ac6e707ca2d7957c89a9e7e3c71a70a1c

Observation f07620ed-7ca1-4ac4-9df6-786b49367384 · inbound

Satori: Reinforcement Learning with Chain-of-Action-Thought Enhances LLM Reasoning via Autoregressive Search cites this paper.

Satori: Reinforcement Learning with Chain-of-Action-Thought Enhances LLM Reasoning via Autoregressive Search Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-09T11:57:48.773155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:57:48.773155Z digest=sha256:fd52edbc5faab365e95dfee6e36c8272017e6e7c63ae41529d9ba2039c148fba

Observation 8b4943d0-f8c2-4144-99f0-d7676e972f86 · inbound

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition cites this paper.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.496098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.496098Z digest=sha256:52b2fc5cec713762c40778f867267e98733079627410783ad9d36edab268ca1c

Observation 45588ba6-0ae0-450f-92ae-55e8bd3face8 · inbound

Policy Guided Tree Search for Enhanced LLM Reasoning cites this paper.

Policy Guided Tree Search for Enhanced LLM Reasoning Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-09T11:20:31.845722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T11:20:31.845722Z digest=sha256:824d8c3e551d91c130ae670bde34563ab5ae340c9d63aa098f454c2ebaa46ed2

Observation e80d78c0-8be0-42a8-9da8-0ce835730b0d · inbound

From System 1 to System 2: A Survey of Reasoning Large Language Models cites this paper.

From System 1 to System 2: A Survey of Reasoning Large Language Models Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 231

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T01:36:24.374302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T01:36:23.845366Z digest=sha256:b64304b377eaf90a175a4942593bb31b576a3edc0752ad4b52b20747a91aea07

Observation 36d78706-5955-4805-b2f0-56030b1edd87 · inbound

Search-Based Multi-Trajectory Refinement for Safe C-to-Rust Translation with Large Language Models cites this paper.

Search-Based Multi-Trajectory Refinement for Safe C-to-Rust Translation with Large Language Models Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-22T14:36:40.935994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T14:36:14.709972Z digest=sha256:135d11c75f43d6555c1ff1f7fa0fec39e700689668bf54863072478604acc0b4

Observation 088d5aa3-919b-412f-bf99-72f47975f052 · inbound

VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought cites this paper.

VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:10.998275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:10:10.998275Z digest=sha256:ffa0c4839440b6054c8d1ccc0d3f04ed87f83e3d08ea7ac2f04211eaabce1ea7

Observation 369bc846-c9c6-4fa4-ad5a-2aca8c78187a · inbound

Fast Quiet-STaR: Thinking Without Thought Tokens cites this paper.

Fast Quiet-STaR: Thinking Without Thought Tokens Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:22.061667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:46:22.061667Z digest=sha256:338edf479d76ca4039a5dd74afdde8cbfe91ac9bf1eb206d3d61f2cd5bf20f7c

Observation 260ed414-c322-4804-97af-d8b360aa700d · inbound

Self-Critique Guided Iterative Reasoning for Multi-hop Question Answering cites this paper.

Self-Critique Guided Iterative Reasoning for Multi-hop Question Answering Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:23:46.999033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:23:46.999033Z digest=sha256:db643daf22d53a1075e2fa62a3eb97ead80fd4c6545aeb8a5a80487627965648

Observation fa573d68-e3c0-47c5-adad-55b701f0db32 · inbound

MMATH: A Multilingual Benchmark for Mathematical Reasoning cites this paper.

MMATH: A Multilingual Benchmark for Mathematical Reasoning Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:23:08.712771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:23:08.712771Z digest=sha256:6f5b7d9329509cf52ffa4774db0cb15c95cd8b2a96284b5d072457422e6c64bb

Observation 59b4d008-5849-4dde-8bf4-48ebb1c29084 · inbound

Scalable, Symbiotic, AI and Non-AI Agent Based Parallel Discrete Event Simulations cites this paper.

Scalable, Symbiotic, AI and Non-AI Agent Based Parallel Discrete Event Simulations Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T13:07:50.276422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:07:50.276422Z digest=sha256:094fbf21ea040e564d6608f6558dd7249d1d637c1a09d4ed07147df0306e2d6a

Observation ae8e3374-ad35-4ae9-bc01-3de7c29f3cb3 · inbound

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence cites this paper.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:49.740854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:49.740854Z digest=sha256:834a7a79199d54008d516f5d278a372a33dfde938b4912cc3d452789a3e2d7bb

Observation 401ebe53-9236-41fc-aa5b-9ba4c81c56e5 · inbound

ProtInvTree: Deliberate Protein Inverse Folding with Reward-guided Tree Search cites this paper.

ProtInvTree: Deliberate Protein Inverse Folding with Reward-guided Tree Search Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:58:50.105199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:58:50.105199Z digest=sha256:01fef323982aec55871d54923ea8a97d199064b5c7b4e1567a951193e0c5fcbf

Observation 4cab0c0e-b358-4ad4-9987-312b364c2fb3 · inbound

LLM-First Search: Self-Guided Exploration of the Solution Space cites this paper.

LLM-First Search: Self-Guided Exploration of the Solution Space Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:43.114675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:43.114675Z digest=sha256:2241c82ad7cf227cfda5a00f697fba0426cffc04f9f41aaeecec81f1e079846d

Observation 8a1e9201-a555-4784-9b4d-594212cd067a · inbound

Token Signature: Predicting Chain-of-Thought Gains with Token Decoding Feature in Large Language Models cites this paper.

Token Signature: Predicting Chain-of-Thought Gains with Token Decoding Feature in Large Language Models Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T06:09:18.814824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:09:18.814824Z digest=sha256:9ee0f567f5471f784932fcfcfe3d920048a699086213e02a679809eed91678f0

Observation ae4bfb40-86e1-4e8f-ad17-e17e3fcb040e · inbound

Chain of Methodologies: Scaling Test Time Computation without Training cites this paper.

Chain of Methodologies: Scaling Test Time Computation without Training Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:50:36.463953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:50:36.463953Z digest=sha256:7d4a5aadf41d2875482fe23a9d358efecaa22257841e699270536636576e0a1c

Observation 015e5a0c-e089-4a1b-8558-d793b6c6cd01 · inbound

Boosting LLM's Molecular Structure Elucidation with Knowledge Enhanced Tree Search Reasoning cites this paper.

Boosting LLM's Molecular Structure Elucidation with Knowledge Enhanced Tree Search Reasoning Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T21:59:11.890969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:59:11.890969Z digest=sha256:b796cffc9c656bfda392585b47b785bc5ae345e2741280317d155dfe17b0460d

Observation 764e42ab-a1e1-4336-a2e9-c90d3b6edb80 · inbound

MiCoTA: Bridging the Learnability Gap with Intermediate CoT and Teacher Assistants cites this paper.

MiCoTA: Bridging the Learnability Gap with Intermediate CoT and Teacher Assistants Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T20:46:21.623327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:46:21.623327Z digest=sha256:daa11619f4b733de2e86fdff89be2d442ba583529d473f553de085bf220d2bb7

Observation 1ab23470-2b95-42f2-8ac8-d0fe97ba9b41 · inbound

Are They All Good? Evaluating the Quality of CoTs in LLM-based Code Generation cites this paper.

Are They All Good? Evaluating the Quality of CoTs in LLM-based Code Generation Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T18:57:13.092944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:57:13.092944Z digest=sha256:fb7c38006afcdfdc9def7621bd996035dc8782af8ac5dc93120eb4b5a34782f4

Observation 88d101ee-1d03-4d03-bf39-7a7376d74cf8 · inbound

Thinking Before You Speak: A Proactive Test-time Scaling Approach cites this paper.

Thinking Before You Speak: A Proactive Test-time Scaling Approach Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T16:23:35.889317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:23:35.889317Z digest=sha256:68c8e6143ea9743e6e021ccf7e8cb777d0c38fd5e9a7f093f8d5778aa3df480e

Observation ea3a972b-0e94-430b-babe-f68184fb8b26 · inbound

Mitigating Visual Context Degradation in Large Multimodal Models: A Training-Free Decoupled Agentic Framework cites this paper.

Mitigating Visual Context Degradation in Large Multimodal Models: A Training-Free Decoupled Agentic Framework Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T12:32:36.332335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-18T12:31:25.257879Z digest=sha256:debcda37cbaaa634f43d0ead4abc40c2318ce41122b05ba433d2eac7234753a4

Observation 0878ce24-1f4d-4c37-b545-8dfb1d9fc114 · inbound

DeepSearch: Overcome the Bottleneck of Reinforcement Learning with Verifiable Rewards via Monte Carlo Tree Search cites this paper.

DeepSearch: Overcome the Bottleneck of Reinforcement Learning with Verifiable Rewards via Monte Carlo Tree Search Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T12:12:35.898938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T12:12:25.437344Z digest=sha256:97f2a1211fb62025019b4a2520b343b78b426f0880d468e3bb56c119d20e7cdc

Observation e49369ee-2d24-4817-8bbe-d18c2ee685ec · inbound

Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs cites this paper.

Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T05:30:54.993450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T05:30:11.389756Z digest=sha256:72273d52b72e3add043190d70674439aeb7f9397b3cab89ef865c451e77f8aed

Observation 7ed99d0c-44d5-4e0c-acab-272436ffcb01 · inbound

OmniDrive-R1: Reinforcement-driven Interleaved Multi-modal Chain-of-Thought for Trustworthy Vision-Language Autonomous Driving cites this paper.

OmniDrive-R1: Reinforcement-driven Interleaved Multi-modal Chain-of-Thought for Trustworthy Vision-Language Autonomous Driving Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T22:38:37.782622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T22:34:00.895252Z digest=sha256:88ce942c30656786cd6d1f120e4b7e3cd889484a5bf37531f3e8888b32345fd1

Observation f2faae9e-da8e-46ff-80ca-098c053f0da2 · inbound

SVSR: A Self-Verification and Self-Rectification Paradigm for Multimodal Reasoning cites this paper.

SVSR: A Self-Verification and Self-Rectification Paradigm for Multimodal Reasoning Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:31:03.962349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T15:59:04.802241Z digest=sha256:f7b1e19688ae2bf2d171df237eaefb0c1dc9e65e9db7b47db4112d7622f64a5f

Observation 1f458e32-c60f-459e-9c33-9f8e34cf559f · inbound

SVSR: A Self-Verification and Self-Rectification Paradigm for Multimodal Reasoning cites this paper.

SVSR: A Self-Verification and Self-Rectification Paradigm for Multimodal Reasoning Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-12T22:47:33.917420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T22:47:33.917420Z digest=sha256:fe01fb9bdc74cfbeae5a1136329a3218c83b3304f4b26024a6283b77a8c0c258

Observation 87a31684-2ded-4a42-a2d3-f65adb39af86 · inbound

Free Energy-Driven Reinforcement Learning with Adaptive Advantage Shaping for Unsupervised Reasoning in LLMs cites this paper.

Free Energy-Driven Reinforcement Learning with Adaptive Advantage Shaping for Unsupervised Reasoning in LLMs Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 48

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T07:45:59.933722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T16:58:10.013475Z digest=sha256:4f9d3c555f90625d464a4b7f69e96cba9b9cd4d0fbb666f43e3108f62c2b99f8

Observation 1dbf3ab6-9e43-4a01-9582-2bdbf9a7a525 · inbound

Adapt to Thrive! Adaptive Power-Mean Policy Optimization for Improved LLM Reasoning cites this paper.

Adapt to Thrive! Adaptive Power-Mean Policy Optimization for Improved LLM Reasoning Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 48

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T08:00:59.947926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T16:51:19.555272Z digest=sha256:6c1250f3e7a325fed006511b6a78e3f2561547534a51cb43dcbab12c5772f3a1

Observation 0eb651f4-97e2-414c-af7a-43a9da5a0609 · inbound

APCD: Adaptive Path-Contrastive Decoding for Reliable Large Language Model Generation cites this paper.

APCD: Adaptive Path-Contrastive Decoding for Reliable Large Language Model Generation Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:26:25.223791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-12T05:22:25.475956Z digest=sha256:b6681e06225b9549079f486d89f3af93ef69fc5a8d6a46a9b4bc1d1aa208dce5

Observation 4d7eb8d8-5d58-41ad-9151-c89c73c44ac9 · inbound

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration cites this paper.

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 46

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T06:51:30.080800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T03:51:52.375703Z digest=sha256:cdde68726facfd96fbf607eb375873e73832b3a911ce8cfdcbb2b331e6018255

Observation 2c467a79-4875-4a67-8451-60d0dc32c500 · inbound

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration cites this paper.

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 46

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T05:15:03.462697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T05:11:32.053440Z digest=sha256:1d80c604b413d0f412fb48f1dcee4dbe1d3461ba9b246b4c9a616ccd2605234f

Observation d7f9824a-4fda-4dce-bd98-859ce6b90241 · inbound

Reducing Credit Assignment Variance via Counterfactual Reasoning Paths cites this paper.

Reducing Credit Assignment Variance via Counterfactual Reasoning Paths Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-21T00:49:18.778579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T00:45:41.280228Z digest=sha256:adfc07c4162f41ee83686622dbd98b47cbc92d534c6c9061097f2c77b58df684

Observation a1d82060-8488-42b8-b479-41fffd4f14e8 · inbound

DISA: Offline Importance Sampling for Distribution-Matching LLM-RL cites this paper.

DISA: Offline Importance Sampling for Distribution-Matching LLM-RL Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T15:13:24.922625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T15:11:43.235574Z digest=sha256:bcdb3207c1461766114f30987686bb16ddb2fa55f67cfcd373bb10bb9315d96b

Observation bae8a455-3f4a-4387-8fdc-b26eff2f985e · inbound

Tree of Thoughts as a Classical Heuristic Search Problem: Formal Foundations and Design Patterns cites this paper.

Tree of Thoughts as a Classical Heuristic Search Problem: Formal Foundations and Design Patterns Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-29T12:13:27.100782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T12:03:48.538010Z digest=sha256:110ffa7d4715cb5ca9f3266f4d86bc3c3ebde6da1447365bd68bd59a0d44d6a7

Observation 0f9d453b-3c1d-4ed2-9c56-18ef53b6f710 · inbound

Self-Improving Small Object Grounding in LVLMs cites this paper.

Self-Improving Small Object Grounding in LVLMs Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T22:06:16.284549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T15:48:01.833952Z digest=sha256:854c36b8776773eea5ec403541621b93e17a659ade087d9828e8dc1345186392

Observation 88315924-9e55-4c56-97d4-5f67d1c061f4 · inbound

RASFT: Rollout-Adaptive Supervised Fine-Tuning for Reasoning cites this paper.

RASFT: Rollout-Adaptive Supervised Fine-Tuning for Reasoning Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-06-27T22:31:21.143378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T22:25:52.959859Z digest=sha256:2725c77c04685bb911f438d641e7d31906523afadf9e79668e27a8f3418e31d8

Observation b6a7765b-aa45-40a4-bb57-373f437bff88 · inbound

DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination cites this paper.

DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 185

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:57:23.895049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T19:58:32.016341Z digest=sha256:ac1276327f15442fd21e716b30e4618c476ffb388dcce4567695420b0f015ed9

Observation a009eeb0-e452-4981-a04f-d6f2495dc28d · inbound

MetaPS: Adaptive Programmatic Strategy Selection for Market Agents cites this paper.

MetaPS: Adaptive Programmatic Strategy Selection for Market Agents Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T08:39:42.695889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-26T11:06:28.690956Z digest=sha256:f7f15c6da2ffbdfac0f50a9da8625d0409ce6c146aa190d574926e6bf013f9d3

Observation 643c1e6b-17c9-4546-a806-817b4d5f0176 · inbound

The Hitchhiker's Guide to Agentic AI: From Foundations to Systems cites this paper.

The Hitchhiker's Guide to Agentic AI: From Foundations to Systems Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 254

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T11:09:46.531751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T08:09:57.542558Z digest=sha256:7a489f099df87e1fe023a335554dd024a128e979044893e8d9bcff66768cbe7d

Observation b7e17d4d-bb20-4ae9-9be9-0f82d470ede6 · inbound

The Hitchhiker's Guide to Agentic AI: From Foundations to Systems cites this paper.

The Hitchhiker's Guide to Agentic AI: From Foundations to Systems Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 239

Resolution
unresolved
no resolver link, observed 2026-08-02T10:27:18.589000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T10:27:18.589000Z digest=sha256:bb9a1237cb11447fa3a2df7bc99f849f82a2f010de9cf71fe38cb38c0b98ac99

Observation 2d017af8-3cfc-4adb-8fb1-6a0c508bc27d · inbound

MILES: Modular Instruction Memory with Learnable Selection for Self-Improving LLM Reasoning cites this paper.

MILES: Modular Instruction Memory with Learnable Selection for Self-Improving LLM Reasoning Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-07-09T00:45:49.108733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-09T00:45:00.847714Z digest=sha256:63953d48f4d39a7486b25b6ae5842d6e1be7cf354049f0523f2c7c31669c8c60