Pith. sign in

Paper Citation Record · LEDGER

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models

As of 16 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 0 inbound Pith citation observations for arXiv:2508.12387.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.12387 v1

Coverage vector

measured 46 of 46 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T17:30:10.366860Z

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

46 of 46 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved45
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8432d6b2-4b6a-497c-bace-101932d726fe · outbound

This paper cites MCC-KD: Multi-CoT Consistent Knowledge Distillation.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models MCC-KD: Multi-CoT Consistent Knowledge Distillation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:09.997455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:09.997455Z digest=sha256:f9b84f061eeb9af5fcff37c20227c1080c248616a3e6591e2e7b828eccad3afc

Observation 8f640e9a-fbab-4c67-9d08-abd81a441dc8 · outbound

This paper cites an unresolved cited work.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.006417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.006417Z digest=sha256:9226e65b324e3def931f06b08b5a210537f79c1d560ced6571ebde535027b0f8

Observation 6fa91cd9-ee7b-4a85-9d1e-ed8569a11dc9 · outbound

This paper cites Universal Self-Consistency for Large Language Model Generation.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Universal Self-Consistency for Large Language Model Generation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.012367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.012367Z digest=sha256:46f75bdb646cafa4527105636c399cd261471a116de223cd2cb09b53dddd28d8

Observation 1e2aee67-d4b1-40b7-a6b9-5f8d6c987496 · outbound

This paper cites an unresolved cited work.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:11.523800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T17:30:10.018584Z digest=sha256:85850c89a9e293b16b3924ae326f558633b310846d72027d199851d46d9fec73

Observation 8ddc747b-4e1a-4548-b861-f99b7614675f · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.025196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.025196Z digest=sha256:030c3aba1dcaf244bfd7b098554e43ec6063155f6790064edde79786015abe9f

Observation 2adfa126-7575-4906-b9f8-4f93a3ee4105 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Training Verifiers to Solve Math Word Problems

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.032031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.032031Z digest=sha256:102280ae56a0a6c38020843d098f26d81ad5ca1ab5a80c0f4f8abec13c720b1c

Observation d60ea19f-685a-4960-9fe8-a93fd9afe398 · outbound

This paper cites Improve Student's Reasoning Generalizability through Cascading Decomposed CoTs Distillation.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Improve Student's Reasoning Generalizability through Cascading Decomposed CoTs Distillation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.039486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.039486Z digest=sha256:53977f178629a8c9a1ae4f01eba450c20f638804c970bab079a8915c00b07928

Observation f69e4fd5-e182-403a-a396-193db6ba8449 · outbound

This paper cites Investigating Symbolic Capabilities of Large Language Models.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Investigating Symbolic Capabilities of Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.046467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.046467Z digest=sha256:7e0752731b76ec6a8f2bec2a8224ac041d985c3b9821d844936bb014628ad0be

Observation 5bd4380b-62bf-44e2-b115-ea97a63a180b · outbound

This paper cites an unresolved cited work.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:11.498581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T17:30:10.051819Z digest=sha256:6778d778c656cb9a043333e8cfd6db06c16db5902cec4e170ec720ba64ef13a8

Observation 6b6cf652-22fa-4e10-be71-e57b220a7bb0 · outbound

This paper cites an unresolved cited work.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.057883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.057883Z digest=sha256:3095b312ea1ecd39042eff734f7ba163a2f2febb59f7c586358dfa4857650844

Observation e440bc80-48a1-438c-8670-1d35ddf5c36a · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.064832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.064832Z digest=sha256:ad10b77b5099d9098436d896fee995685e1e7fadf97f5406ab80b036e8c3a3ea

Observation c4f230cd-5547-4796-b405-5e7f9fc5bcbb · outbound

This paper cites an unresolved cited work.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:11.466446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T17:30:10.070606Z digest=sha256:25e1d06537ff51e2e3689c96280d3ebc28a5478f0f05cf7d78d1f340f2e64f3a

Observation 5d8c221f-b28f-48ed-8b47-f016461849c1 · outbound

This paper cites On the Impact of Knowledge Distillation for Model Interpretability.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models On the Impact of Knowledge Distillation for Model Interpretability

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.076003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.076003Z digest=sha256:c2bd392822b2b3b75e6c148db61790aceca0a4905f14d01be63ac7bba9eb7fe6

Observation a4860222-7179-4dd9-b878-8de876ce840d · outbound

This paper cites Measuring Massive Multitask Language Understanding.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Measuring Massive Multitask Language Understanding

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.081962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.081962Z digest=sha256:f6a749a7c75b50937cc3d5b8ac447af57ecb63211d3ffa17a11880c880a517dc

Observation eca87bf4-5dff-4c63-9cac-8f6c546b1b10 · outbound

This paper cites Distilling Step-by-Step! Outperforming Larger Language Models with Less Training Data and Smaller Model Sizes.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Distilling Step-by-Step! Outperforming Larger Language Models with Less Training Data and Smaller Model Sizes

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.087366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.087366Z digest=sha256:27b25c113973eb29ee176dd5224fa620381774cfd3fdb6e6e2bb30baa9c4e84e

Observation 5ce6bc4e-4f41-40c6-b904-87278c9c0d06 · outbound

This paper cites Qwen2.5-Coder Technical Report.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Qwen2.5-Coder Technical Report

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.093473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.093473Z digest=sha256:7c5259b5027b22c35a39e4307ac943af3780f5f22a36c47e065ff2b7779d94f9

Observation 37f53f8d-f51c-490d-bc38-47243f20455b · outbound

This paper cites Simple and Scalable Strategies to Continually Pre-train Large Language Models.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Simple and Scalable Strategies to Continually Pre-train Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.099822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.099822Z digest=sha256:bf6b37a586c0958961e07f5e9de0b0fed498ba669ca22ac777a566d5b6c849f6

Observation efd70028-0b3f-47f5-9d93-3e5bfc7e3beb · outbound

This paper cites OpenAI o1 System Card.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models OpenAI o1 System Card

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.108597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.108597Z digest=sha256:3f740df8472e6baa9465481b806d9a5a07ace187c8a8137f03eb5c64576e16c6

Observation ff61485d-33eb-4d24-8e62-4473686cfccb · outbound

This paper cites Scaling Laws for Neural Language Models.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Scaling Laws for Neural Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.114480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.114480Z digest=sha256:9d3d289d5e497d248ab8a219e8a16b723f23989ce444243974b19dc9f8a840a5

Observation a7236b47-a1d3-460d-8c57-eabcea247270 · outbound

This paper cites The CoT Collection: Improving Zero-shot and Few-shot Learning of Language Models via Chain-of-Thought Fine-Tuning.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models The CoT Collection: Improving Zero-shot and Few-shot Learning of Language Models via Chain-of-Thought Fine-Tuning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.119254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.119254Z digest=sha256:27f704d3ddad2573e47c2fa67e17c1d3c0178a240cac09444c62e0ab573d374f

Observation ce40e913-bef6-46b7-8b88-93710474b509 · outbound

This paper cites GSM-Plus: A Comprehensive Benchmark for Evaluating the Robustness of LLMs as Mathematical Problem Solvers.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models GSM-Plus: A Comprehensive Benchmark for Evaluating the Robustness of LLMs as Mathematical Problem Solvers

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.125206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.125206Z digest=sha256:48d5de66b72337e1714006defbe5734ee290266028ed165389ced31cf66f9ffb

Observation 72803b75-f54f-42e2-ae6b-679d7e66dfd5 · outbound

This paper cites Explanations from Large Language Models Make Small Reasoners Better.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Explanations from Large Language Models Make Small Reasoners Better

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.132783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.132783Z digest=sha256:6581f47ce58d69f2fbc8bf45c0b03ff588d1dfe3b9399521ed82baf79c1b13db

Observation d532f56f-1d36-4e14-a307-9d752f503eed · outbound

This paper cites Model Merging in Pre-training of Large Language Models.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Model Merging in Pre-training of Large Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.163726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.163726Z digest=sha256:0901c07151d5f0c5aefd16e0158b5df077f1b12ae65ed6213a2101e7f14d54b4

Observation 9cdbfe87-a25c-4528-81e5-64f0d3799ae5 · outbound

This paper cites $\textit{SKIntern}$: Internalizing Symbolic Knowledge for Distilling Better CoT Capabilities into Small Language Models.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models $\textit{SKIntern}$: Internalizing Symbolic Knowledge for Distilling Better CoT Capabilities into Small Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.202314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.202314Z digest=sha256:17ef3ee18381094c8c9293bf96ab72ac42aed07db8fe4ab7ecf5a16375d90fde

Observation f158f0e1-900f-44fb-82f0-f9a9af57a7da · outbound

This paper cites an unresolved cited work.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.208928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.208928Z digest=sha256:f71d6a535fa993fac32ddb917146282d7ff9f562a7fc128dcf2af6f40f1d3bfd

Observation dce8a3a3-818d-47c0-8276-c3770c2480a0 · outbound

This paper cites Mind's Mirror: Distilling Self-Evaluation Capability and Comprehensive Thinking from Large Language Models.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Mind's Mirror: Distilling Self-Evaluation Capability and Comprehensive Thinking from Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.216030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.216030Z digest=sha256:016d1aea503c7e10affa1c87d1cb921ab7c243fa1da292026b603f5f5776f101

Observation 2a29cc3b-2bae-4c45-b852-7083dd8091f0 · outbound

This paper cites an unresolved cited work.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.222024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.222024Z digest=sha256:7ec5fb30fc4a0d63de38f71f4cf8746df854002aecf563dc609d1e0c647d8f8b

Observation 80340091-be7f-4bd1-abaa-4ee8b83eb3ae · outbound

This paper cites SQuAD: 100,000+ Questions for Machine Comprehension of Text.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models SQuAD: 100,000+ Questions for Machine Comprehension of Text

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.228110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.228110Z digest=sha256:7c146233f9b64c37c7da194a85297494033214ea368ff3ec1ec02417e48dd9fd

Observation 739bd6f1-b775-42ab-be4d-6ee58d00d90e · outbound

This paper cites an unresolved cited work.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:30:11.411147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T17:30:10.234315Z digest=sha256:bdbdf8016d47b02bdaf94162221d3060182c3a49a5da1ce73a2175c19159a8db

Observation 878e4188-55bf-4e77-a1ef-a89b59af5b45 · outbound

This paper cites Active Learning for Convolutional Neural Networks: A Core-Set Approach.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Active Learning for Convolutional Neural Networks: A Core-Set Approach

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.243928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.243928Z digest=sha256:733c90ac54d922c7c3d5f74c224b74d3946a531d202b938d9e9162882cca442b

Observation f3776c99-b666-4ffb-acb3-8a308e1797a1 · outbound

This paper cites Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.251490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.251490Z digest=sha256:5b66aac61b4d31d42f18eb04620accff51e3f844e7f50155acc53be56d2466a1

Observation 143b60e9-9fad-474d-b27d-b569c6246724 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.261498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.261498Z digest=sha256:52db244a080be8d2b52be70f7c460b1f2faee4795feff6f6a986713c1aeeb897

Observation 24c9c5b6-544f-4ad0-88ce-436cde86f6a2 · outbound

This paper cites A.; Abid, A.; Fisch, A.; Brown, A.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models A.; Abid, A.; Fisch, A.; Brown, A

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.267698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.267698Z digest=sha256:8f88f40c64cb88f8f557c13092a2d1efbd6c7a0a544d590b086c2240c751dd0e

Observation 8e60cab9-2d9e-415a-9590-5a3a96fc5387 · outbound

This paper cites Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.276856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.276856Z digest=sha256:a5f9befc91a0931952ac529603e537efd91dd0aba4f1a52f14a0ce3523e46c39

Observation 39009085-8a5a-43f4-b847-496c1bef439b · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.285920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.285920Z digest=sha256:3ea293bf2bf56f23f192967948d4fc1da2ff2926c00d8de3ef1aeface0ebc338

Observation b2f2421b-99d1-4ea6-aaee-26f7de7dea2e · outbound

This paper cites Interpretable Preferences via Multi-Objective Reward Modeling and Mixture-of-Experts.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Interpretable Preferences via Multi-Objective Reward Modeling and Mixture-of-Experts

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.291665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.291665Z digest=sha256:00889f7b55d17080c89ea17d77b28cdf1295df9b4a3ffefc315bca244dfbb5d6

Observation a68d2bfc-8f75-42e3-8380-e6c34325ab38 · outbound

This paper cites Self-Consistency Improves Chain of Thought Reasoning in Language Models.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Self-Consistency Improves Chain of Thought Reasoning in Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.298033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.298033Z digest=sha256:6e71adcda049763f26f6760f4c7f58df1934d967dbd66b19b88c107c6de0fd41

Observation e32ed6e3-1c5b-42f6-a4a4-6e4ee7b7cdbd · outbound

This paper cites V.; Zhou, D.; et al.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models V.; Zhou, D.; et al

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.303249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.303249Z digest=sha256:58d019f887486ce719f1b2f3bc0737376ff7a94b656f8762d758c39f6122bd57

Observation 166a238d-5d8c-4071-8861-3695b7761c84 · outbound

This paper cites an unresolved cited work.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Unresolved cited work

Reference 39

Resolution
verified exact
raw_fallback, observed 2026-08-15T17:30:10.685464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T17:30:10.313356Z digest=sha256:43023d246ce1421f1bacdee69e3d8709906c36f519101c652eb81de231f7b914

Observation 0ca5fcd7-cc00-45bf-bc93-c9780812a81b · outbound

This paper cites LLMs-as-Instructors: Learning from Errors Toward Automating Model Improvement.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models LLMs-as-Instructors: Learning from Errors Toward Automating Model Improvement

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.319553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.319553Z digest=sha256:482426ab9c3d36f5bc6f5761849a415e841d7eba3136afc2f7e5fa39c3b78124

Observation fe457447-e8ca-4eae-a665-dc163fefc2e8 · outbound

This paper cites Scaling Relationship on Learning Mathematical Reasoning with Large Language Models.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Scaling Relationship on Learning Mathematical Reasoning with Large Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.331861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.331861Z digest=sha256:2543a3e0841e57050f0b236399351b11a534178fd297dd732b8d45edf16ee6f1

Observation f1f0bda7-5e68-4f86-a0c2-ad79bb686e11 · outbound

This paper cites CoT-based Synthesizer: Enhancing LLM Performance through Answer Synthesis.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models CoT-based Synthesizer: Enhancing LLM Performance through Answer Synthesis

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.340136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.340136Z digest=sha256:b5cb17f19dd2cda9945010058417a719cd12ac73b0b6d71eed9f8bcb9919ba10

Observation 5b09c08b-0b53-4b17-8b80-040962e910e2 · outbound

This paper cites Towards the Law of Capacity Gap in Distilling Language Models.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Towards the Law of Capacity Gap in Distilling Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.347134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.347134Z digest=sha256:522d15fc159b4acdc40137c8daaeb6f4c2d16ef468564d7d7a2810a5737c4ab2

Observation b782d167-d4c3-44e2-8bf8-c9b727fbe744 · outbound

This paper cites TinyLLaVA-Video-R1: Towards Smaller LMMs for Video Reasoning.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models TinyLLaVA-Video-R1: Towards Smaller LMMs for Video Reasoning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.352493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.352493Z digest=sha256:538dbc9407c29e0bf7ce38227730903e2f4e6d9ae1a485ea933f642e1256303b

Observation 6dd2f60a-b4a4-4803-8b55-5ed589bf5855 · outbound

This paper cites AGIEval: A Human-Centric Benchmark for Evaluating Foundation Models.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models AGIEval: A Human-Centric Benchmark for Evaluating Foundation Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.358136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.358136Z digest=sha256:37a5c5f3ef16efdb55e20793e01f8e1d92a970fe439302741d98598fa891bedd

Observation 69358d05-724e-4983-8aca-6792ff0d6fba · outbound

This paper cites Rethinking Soft Labels for Knowledge Distillation: A Bias-Variance Tradeoff Perspective.

ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models Rethinking Soft Labels for Knowledge Distillation: A Bias-Variance Tradeoff Perspective

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:10.366860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:30:10.366860Z digest=sha256:9e42ddeddfecec4aefc7bd4d4944e709402e932896d6399628c52a1a5d77c134

Pith citing papers

No inbound Pith citation observations are available.