Pith. sign in

Paper Citation Record · LEDGER

A Survey of Scaling in Large Language Model Reasoning

As of 7 August 2026, this Paper Citation Record lists 100 of 265 outbound references and 3 inbound Pith citation observations for arXiv:2504.02181.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.02181 v2

Coverage vector

measured 100 of 265 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-22T21:20:07.238992Z

measured 103 of 103 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-01T08:35:09.009897Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-01T08:35:33.342217Z

Reference resolution

100 of 265 outbound references displayed

  • verified exact57
  • verified fuzzy4
  • unresolved36
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3d1ecf94-a01e-4e52-80e4-1e436ff10bf3 · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.314456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:a2dff318de545fc0d5d64e135c21b5234ab6a25f46425e430cfb6326b94953de

Observation 21ff7433-45e3-458f-92f5-5240f640411a · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.302305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:33c2aee9fd5bbecb2e6ea3a7319a2be50ea261240f8c8fc361d50fccbfa12347

Observation c64f43c6-e0f2-4c65-b1c5-268342ce4e4e · outbound

This paper cites Back to Basics: Revisiting REINFORCE Style Optimization for Learning from Human Feedback in LLMs.

A Survey of Scaling in Large Language Model Reasoning Back to Basics: Revisiting REINFORCE Style Optimization for Learning from Human Feedback in LLMs

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-22T21:22:09.509264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:179d3cb5fc08975cb53303282ed9c39bc6d65f0c0b2006b91c246c4471f8e4d0

Observation 8d6dba17-85ae-4b32-920e-7d112aa8bfc7 · outbound

This paper cites Comparing Rationality Between Large Language Models and Humans: Insights and Open Questions.

A Survey of Scaling in Large Language Model Reasoning Comparing Rationality Between Large Language Models and Humans: Insights and Open Questions

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.129785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:538604bd606049783c4e802bf89ce87f69d71797573e30fa9209402955da1bfc

Observation 70eebbb4-6f26-418e-b8ed-a7d5af244f48 · outbound

This paper cites Concrete Problems in AI Safety.

A Survey of Scaling in Large Language Model Reasoning Concrete Problems in AI Safety

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-22T21:22:09.470182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:7db9101fcae948d1d3236290c1d990bbbcaecd5291e23b1ef701267a26bd64f7

Observation ee2e7284-731f-4b64-88f5-d71ee45b270c · outbound

This paper cites Revisiting In-Context Learning with Long Context Language Models.

A Survey of Scaling in Large Language Model Reasoning Revisiting In-Context Learning with Long Context Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:08.847660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:2d40b6b6ad6f31bfa2a1d810394bfa04a54d8fb61c19bd74444aac36aaf6db73

Observation 461b9cde-ab26-4bed-b39e-08c8f5bf26e6 · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.252006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:50bf52d6ebf4251779fc003e00d3f7e7dce275ad5dc3553d303406cfa2b33f77

Observation 5275f3a5-7e67-4f71-a5c5-1327ad322fb0 · outbound

This paper cites In-Context Learning with Long-Context Models: An In-Depth Exploration.

A Survey of Scaling in Large Language Model Reasoning In-Context Learning with Long-Context Models: An In-Depth Exploration

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.498310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:cd396443dbe62d076694f1da4d96edd8095704ac18b300ec0c1096be66586066

Observation ea69f26d-fba4-4e46-8669-07b0b146d682 · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.244634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:0985c7b2fe46965bb315457a180c37a73a080bab94b13e13a54372439c0e1d0c

Observation fb7899a0-4996-4c8c-bffd-3113b5eb1a17 · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.241285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:bcc9c3011d5f0600234bb71daa6bba5743b5ff22f0e8dbe459c55a1a2190d700

Observation 8806c47c-1293-408d-aeb7-56ed79f389bf · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

A Survey of Scaling in Large Language Model Reasoning Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-05-22T21:22:09.405178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:27c4304f709dcf4a7844f48acb5862ca946740dce22ea6ae1b77d9cc154d39b2

Observation c74935bf-61d3-4014-9a6b-8026f032c74b · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.230613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:4a1c29f39f234b72111bb542ebfca967692888d1aae24b8831fb2f27d58e8d84

Observation c0e29111-f78a-43c6-bd13-7b3da753a302 · outbound

This paper cites Scaling In-Context Demonstrations with Structured Attention.

A Survey of Scaling in Large Language Model Reasoning Scaling In-Context Demonstrations with Structured Attention

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.166055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:103cdefb431e65d796ab06abf1a8d4b050e94cff4db33e4e43743c6372e8d27e

Observation 090dc59d-8ec3-45a8-bfa8-d9f641d6097a · outbound

This paper cites Large Language Models as Tool Makers.

A Survey of Scaling in Large Language Model Reasoning Large Language Models as Tool Makers

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:08.776396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:b391e7b42369fbcc14520a6ef785f82777bbe112f6953d902a803349ad1d4ab0

Observation bb326f58-45ad-4dcf-950d-ec355d3ec111 · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.202504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:2229d21990aef4b27bc6f11ea20ee2cf19833382edf039ba060b004052d09814

Observation 675e1d1d-ac39-4bfc-80ff-a3863c0a37f4 · outbound

This paper cites Reading Wikipedia to Answer Open-Domain Questions.

A Survey of Scaling in Large Language Model Reasoning Reading Wikipedia to Answer Open-Domain Questions

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-22T21:22:09.106608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:44e3a77b80868973163a7225d9d46b73648a309d06f3c78a564946cc8ef3d643

Observation 0163a8d1-76d5-4036-812b-affbe3fdd77e · outbound

This paper cites HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs.

A Survey of Scaling in Large Language Model Reasoning HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-22T21:22:08.741033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:6322898be5793c0bfd2f22edb65f8187ea225647478f897136cb72f14f65bf15

Observation e4bb1187-2562-46e3-9dc3-335f6fe88efa · outbound

This paper cites Reinforcement Learning for Long-Horizon Interactive LLM Agents.

A Survey of Scaling in Large Language Model Reasoning Reinforcement Learning for Long-Horizon Interactive LLM Agents

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.374887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:d30ce0e258e4a8a874ebfc2bbe5efc3264b05059722f2dfa466a1217ecd1bec1

Observation 1339b975-6add-41a7-a0f5-5e03e1ac59c8 · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.180280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:d492d8f1f2062997e24e3c46816cf22f9ba2b8039663d2c726f8824fc93439f5

Observation 7395c582-b1bb-42db-8ccb-348b67f51284 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

A Survey of Scaling in Large Language Model Reasoning Evaluating Large Language Models Trained on Code

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-05-22T21:22:09.395451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:4ae33bfa97c6d64776e8107994958ac5dceda1fdf471537077974d2e891881d7

Observation bf62d06d-dc3d-44d2-90e3-1695ebee2ee2 · outbound

This paper cites LLM-empowered Chatbots for Psychiatrist and Patient Simulation: Application and Evaluation.

A Survey of Scaling in Large Language Model Reasoning LLM-empowered Chatbots for Psychiatrist and Patient Simulation: Application and Evaluation

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.246635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:43e6bb1bd4fc975377a83e866215fb10121910c7b36a17cc2fdb56c2bb91068d

Observation b520749e-ca3f-4cc1-b3c4-9a1c407015fe · outbound

This paper cites Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs.

A Survey of Scaling in Large Language Model Reasoning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-22T21:22:09.006698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:383d0e378602231bee9ff912d4399297bf9cff05a1741681b126fc20fd2c6a88

Observation 48fefaeb-12ac-4066-8bc2-c68ae83f8a61 · outbound

This paper cites Inner Thinking Transformer: Leveraging Dynamic Depth Scaling to Foster Adaptive Internal Thinking.

A Survey of Scaling in Large Language Model Reasoning Inner Thinking Transformer: Leveraging Dynamic Depth Scaling to Foster Adaptive Internal Thinking

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.171274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:b8fb4fa4418120aab3fb12d5678f77c21b06b290b752ff327cc7efee1531ba34

Observation d7a738cf-6525-4186-b1e9-d70d979fe702 · outbound

This paper cites SoulChat: Improving LLMs' Empathy, Listening, and Comfort Abilities through Fine-tuning with Multi-turn Empathy Conversations.

A Survey of Scaling in Large Language Model Reasoning SoulChat: Improving LLMs' Empathy, Listening, and Comfort Abilities through Fine-tuning with Multi-turn Empathy Conversations

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.224878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:c23fa12e0c38ad4a04207f3900cfc8c3a649188fd9dcfd5a14e5311e2fc894a7

Observation 76118bf6-9682-47c1-8501-8142d7d4721f · outbound

This paper cites MEDITRON-70B: Scaling Medical Pretraining for Large Language Models.

A Survey of Scaling in Large Language Model Reasoning MEDITRON-70B: Scaling Medical Pretraining for Large Language Models

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-22T21:22:08.711160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:73f3293ad58a1c11cc4c2bc589e1104e4283816ebd22dd778a69ba6e7c1b079f

Observation ceedceba-423e-480c-8ff8-2cade7c70811 · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.131111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:38ef0133e816a09f796b974093500c18f9fce86c650cd3de25cc9113b037bdc3

Observation d75a4784-6da8-46b7-bbfc-4744f1b49112 · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.188437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:bb0427a75239235759b988164b01c6bbc94001187e1e405d5ad6205fece9e07c

Observation a1f203df-f4ad-43ae-9ad1-8c54591dac1f · outbound

This paper cites TrojanRAG: Retrieval-Augmented Generation Can Be Backdoor Driver in Large Language Models.

A Survey of Scaling in Large Language Model Reasoning TrojanRAG: Retrieval-Augmented Generation Can Be Backdoor Driver in Large Language Models

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:08.830637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:3ccc50dbfd3721b5adcb317632777db1a7f6fe5fe16898115a2e3c272bd62f39

Observation 79087c8e-b017-49e2-a76d-a51d583a5245 · outbound

This paper cites TRACT: Regression-Aware Fine-tuning Meets Chain-of-Thought Reasoning for LLM-as-a-Judge.

A Survey of Scaling in Large Language Model Reasoning TRACT: Regression-Aware Fine-tuning Meets Chain-of-Thought Reasoning for LLM-as-a-Judge

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.241208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:ea0d0f4ff48dbbfe3b5547f190e2aa9df0648e911826e07e9eeec82642d10cad

Observation bbb48d40-aff3-44f4-8519-63e9d4d362f8 · outbound

This paper cites SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training.

A Survey of Scaling in Large Language Model Reasoning SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-05-22T21:22:08.728695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:adca92f361656776cd7cd230b22ca2febcdee1e72330b8d99f9e5961645cb037

Observation 74c993f4-a3e0-4595-9116-4dfb3794c4ea · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.113211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:5d3316a113c39522fe3a20ca3b9820533e6adf1035419b18b2c7b3b2bd15463d

Observation 41311be0-6faf-4d54-ac97-f62f77419bec · outbound

This paper cites Journal of Machine Learning Research 25, 70 (2024), 1–53.

A Survey of Scaling in Large Language Model Reasoning Journal of Machine Learning Research 25, 70 (2024), 1–53

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T21:22:10.109088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:1fd9005f76ebab6ff460a374115ae1a0601b120f0905d159567f2011de6f4d5c

Observation 9aa075b6-5358-4ce3-9a1e-5beadc8adf7b · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.177191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:79830d992b5a9ea01239f3f62cf93da2a0ffecedcb72e4f72b73b63dc1d982ec

Observation 7690a114-1d12-4590-a5bf-c6b6c503f257 · outbound

This paper cites Implicit Chain of Thought Reasoning via Knowledge Distillation.

A Survey of Scaling in Large Language Model Reasoning Implicit Chain of Thought Reasoning via Knowledge Distillation

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:08.770536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:d44fd946bfacddce8f0b46605d6f8288ad2adcd42bdf2a49d79407b4f46e39a5

Observation 5e8b1d34-d0ff-465d-923f-86503f312d25 · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.080614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:d1f03be05be1ce7405f9d699ad29373cc545376249c21e972aa9d9a9f7130e7a

Observation 1b41703e-1996-4582-91c1-8b60f6f7471a · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.137968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:1aaad589cd285387e5bdb6a2401ea37c0a9793c2a4e2be37387456a10460c6f4

Observation fe06efe5-d144-4c7f-8fa5-bffec20c7ff5 · outbound

This paper cites Agentic llm framework for adaptive decision discourse.

A Survey of Scaling in Large Language Model Reasoning Agentic llm framework for adaptive decision discourse

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.161187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:6289c4384ee845ca9b681611531f252255019bfaa47b11202b2b2af3ede999bc

Observation 9de56b95-e3d4-4a7f-bbcb-9d0e65a20b28 · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.067599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:997ad0f71b91b42ef8e68e32e7049418c293bd45d513abf16a9e5b329c4aeecd

Observation ce044c07-2bca-49c0-a9c5-b1e24cbea18c · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.064266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:4ef2c57d1ae383114cca55b028b48f07c4cc297d3ae5f3aea88c3cd830cdc759

Observation e15697cd-58a5-42d5-a51f-3098fe031b3d · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.077332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:3fb695a0e8dabc848d524db22710d1e83b6187c459c4ad151777530e025e1dad

Observation f4d1c5dc-a926-467d-a049-f96cd9737e36 · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.040194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:091e77a00cfc7a5845d306ed94f365b316e72f9c1632c21508c55dcad128031f

Observation 89d6027a-0986-455c-9000-fd762434b75c · outbound

This paper cites AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator.

A Survey of Scaling in Large Language Model Reasoning AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.176472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:0446f2d5aa4b1881399ee8144770d4b2cad6f2a1df1fae44c25bfddc121a1e06

Observation 0061fd63-7717-43b9-b1c3-ffaebff3e5ba · outbound

This paper cites A Multi-Agent Conversational Recommender System.

A Survey of Scaling in Large Language Model Reasoning A Multi-Agent Conversational Recommender System

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.192454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:402de4e26d67f60efc1d4bb7a84602d636b3a517c5c11644b018f35341e3d077

Observation d88f8c49-840f-4f07-81ae-4c0ec4f92594 · outbound

This paper cites Leveraging Large Language Models in Conversational Recommender Systems.

A Survey of Scaling in Large Language Model Reasoning Leveraging Large Language Models in Conversational Recommender Systems

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.083891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:7edf4a58dc261fb80908f5276c9b6c1ca707ca869058e16a8fc65e0b66cb9b35

Observation fa2d1f11-e831-4cc3-8e84-d63eab42fe7e · outbound

This paper cites Sliding Window Attention Training for Efficient Large Language Models.

A Survey of Scaling in Large Language Model Reasoning Sliding Window Attention Training for Efficient Large Language Models

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:08.801191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:c79a5cf98151bfba5a168dd27c9b8aedbbcf909c3a331c9ed20426194be18bf3

Observation e3ec0b30-336d-4794-9424-41fd728a6e8c · outbound

This paper cites U-NIAH: Unified RAG and LLM Evaluation for Long Context Needle-In-A-Haystack.

A Survey of Scaling in Large Language Model Reasoning U-NIAH: Unified RAG and LLM Evaluation for Long Context Needle-In-A-Haystack

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:08.853374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:363c869794d54d56408900783109dddacfba082623ef30f8390df1dc818dfcee

Observation 08049154-e517-4423-b60e-3a4235a971de · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.457228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:0fd12907d22af70f3921e738d7698a9dc31f39ac8371d374b9ae123cbd79e7b0

Observation bd15afa7-4e77-4fed-aec1-d9aee7874e70 · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.450036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:09d9bb56c5348d1ac559c1b3f2f36b64bf2811c5eb56e2dad0d745b27ccdf30b

Observation 65aa1c9b-04b2-4a42-8d03-26db744c506a · outbound

This paper cites Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach.

A Survey of Scaling in Large Language Model Reasoning Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach

Reference 49

Resolution
metadata mismatch
local_arxiv, observed 2026-05-22T21:22:08.835995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:d56f831964b785ed79c0119e24c90001eaf8518c506de1724a68f438547de7cf

Observation a19c3976-6e74-4402-aa44-4b6ff959c04b · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.437058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:356421871d7e50723ba5f5e7b1091b27144d76eb25a293c4f6d96f06f093a1dd

Observation 46d45b8a-78c4-40d1-bb43-cf67e5770f93 · outbound

This paper cites CRITIC: Large Language Models Can Self-Correct with Tool-Interactive Critiquing.

A Survey of Scaling in Large Language Model Reasoning CRITIC: Large Language Models Can Self-Correct with Tool-Interactive Critiquing

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-05-22T21:22:08.818983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:55870dff47c67f392ed9fa183f0d1c1ceaf8ed0d112ea3b60d2f1bda0e63b438

Observation bd97f8b2-89f7-4e8c-93fe-d37092050f0b · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

A Survey of Scaling in Large Language Model Reasoning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-05-22T21:22:09.055779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:1cfdcbe836d8fbd2e0da575c6f2e13c1741851eeba22399335a8552df5418262

Observation f916478f-19a4-4258-8c19-62a0fb7117dc · outbound

This paper cites Large Language Model based Multi-Agents: A Survey of Progress and Challenges.

A Survey of Scaling in Large Language Model Reasoning Large Language Model based Multi-Agents: A Survey of Progress and Challenges

Reference 54

Resolution
metadata mismatch
local_arxiv, observed 2026-05-22T21:22:09.328967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:2190c922e3db2514920e1dc236ccce80e3bd7485168f47aeee2a55c5b59eb84c

Observation f5755f86-2d3b-47e0-be50-f450c6b2a3ae · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.417158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:4b8f2089723e947c83545804b20341243aead6128f3f54c096a6797ebb92fe04

Observation 16237b5e-0a18-4d72-8942-aa327d699252 · outbound

This paper cites In International confer- ence on machine learning.

A Survey of Scaling in Large Language Model Reasoning In International confer- ence on machine learning

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T21:22:10.413400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:6b1b0bf82fb9ead1b21b04ab2554478a72b5dce3d037c86ba387d347b3497bb7

Observation ea15f9b6-59c1-4188-918b-536c27a73f19 · outbound

This paper cites Training Large Language Models to Reason in a Continuous Latent Space.

A Survey of Scaling in Large Language Model Reasoning Training Large Language Models to Reason in a Continuous Latent Space

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-05-22T21:22:09.344442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:7dd85146f41e37c0646c0e2e549766b32e07bda2a4692d7884e8f3e47b0b9416

Observation 28ce86d2-f281-4e86-88fb-8132e564bfc9 · outbound

This paper cites Structured Prompting: Scaling In-Context Learning to 1,000 Examples.

A Survey of Scaling in Large Language Model Reasoning Structured Prompting: Scaling In-Context Learning to 1,000 Examples

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:08.947253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:bd3a405d307d8d3c2582a0c7d982938f7d653e4172e65f67102704ebd0661b40

Observation b3d13ba6-fec2-4c4b-99e7-deb0ab509a2a · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.402498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:b1ff8949e399291799cfc4105dd5eea5d6126740b60f2ead498daa6745b2db47

Observation fd3f89ec-db34-4c16-8850-d4c00b0748e8 · outbound

This paper cites Large Language Models Are Reasoning Teachers.

A Survey of Scaling in Large Language Model Reasoning Large Language Models Are Reasoning Teachers

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.360185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:66839789e588e165d6643ccf3ea6bfe6689900dca73ac10d57bdd2518aa9b803

Observation 4cfaec74-32f0-4ff0-bc5c-47d1880cfa36 · outbound

This paper cites MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework.

A Survey of Scaling in Large Language Model Reasoning MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-05-22T21:22:08.952449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:5a6d3754023d8fd1282d48306433b66ef992eb95f10728051a1fa5d26f35e352

Observation 5edfe7af-59b4-4fac-a5d5-af2457690eb9 · outbound

This paper cites Does RLHF Scale? Exploring the Impacts From Data, Model, and Method.

A Survey of Scaling in Large Language Model Reasoning Does RLHF Scale? Exploring the Impacts From Data, Model, and Method

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.100904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:c416167a8bcb355f45617ee49ff8f883f1a32570a979c94acca1fc77dadd0c5f

Observation fc1017a8-151f-4dd1-91a1-23308a8b958e · outbound

This paper cites REINFORCE++: Stabilizing Critic-Free Policy Optimization with Global Advantage Normalization.

A Survey of Scaling in Large Language Model Reasoning REINFORCE++: Stabilizing Critic-Free Policy Optimization with Global Advantage Normalization

Reference 63

Resolution
verified exact
local_arxiv, observed 2026-05-22T21:22:09.380118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:689bff28405e2dc1cb91df456d7996050880ef3b0615a37bfa2e618b91c315f7

Observation 27cae4e0-e004-4edc-807b-eac5f8acefa3 · outbound

This paper cites O1 Replication Journey -- Part 3: Inference-time Scaling for Medical Reasoning.

A Survey of Scaling in Large Language Model Reasoning O1 Replication Journey -- Part 3: Inference-time Scaling for Medical Reasoning

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.454226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:2edc6d1f87893051decb749cc4334756644aa640255924b6ba1df73867d692d0

Observation a454aa83-2ff1-4918-8983-0d51f55992c7 · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.388961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:0580a1813930f4468ee454c9519beaeceacea5680fba9e19ce0726db2f10e39f

Observation 6a066174-9168-4cb0-b197-bdb1bb7b3283 · outbound

This paper cites OpenAI o1 System Card.

A Survey of Scaling in Large Language Model Reasoning OpenAI o1 System Card

Reference 66

Resolution
metadata mismatch
local_arxiv, observed 2026-05-22T21:22:08.995753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:9badd210854a4f33354651e9bb4fe862336dd3c866f0fdaa3025dd37a2811efc

Observation dcd0d35e-1e49-49ea-bc0a-c5f8d1ed843f · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.382092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:d6e591b2ae407ac0af3be3e1cf36e851ee832d2b4d6bbf3ecf80b5a428982e04

Observation ba73458c-46e0-449c-a172-0012277e0c68 · outbound

This paper cites Feedback-Guided Extraction of Knowledge Base from Retrieval-Augmented LLM Applications.

A Survey of Scaling in Large Language Model Reasoning Feedback-Guided Extraction of Knowledge Base from Retrieval-Augmented LLM Applications

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.400785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:85d4f693cbbf83f2a15d0336d5f24d2bb032cd1d2c326b707b2f7ac66840dd44

Observation e9b87ba3-cb91-401b-ae3e-c91b1dfe7eb9 · outbound

This paper cites SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities.

A Survey of Scaling in Large Language Model Reasoning SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:08.979508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:e3c126dd8169f5a7bb48182fd14fa294947753c267bd2ac2dcddb661d8c6c700

Observation 296f6d6b-2ffd-4b5a-9495-e1a3c5c74be0 · outbound

This paper cites A Survey on Large Language Models for Code Generation.

A Survey of Scaling in Large Language Model Reasoning A Survey on Large Language Models for Code Generation

Reference 70

Resolution
verified exact
local_arxiv, observed 2026-05-22T21:22:09.230153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:57834b39d6ad0890b1390b68578cb97a4a90cf162d6349101216f19150e84f6a

Observation 35ea2063-c611-40ee-8950-db416cbf987a · outbound

This paper cites LongRAG: Enhancing Retrieval-Augmented Generation with Long-context LLMs.

A Survey of Scaling in Large Language Model Reasoning LongRAG: Enhancing Retrieval-Augmented Generation with Long-context LLMs

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.207818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:a676da99fe63c6b840e7a991b6ebd55dd02d9328b741d8b3cbdd9787a3b98b47

Observation bdce46fe-c04c-4e6a-a53f-d14043478091 · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 73

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.359529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:ee01aa40ff8deb9d3338d6ca3c84cbc24740f1d28d4dc769fdf823a33eaf982e

Observation 8deac156-da47-45e0-b479-b6ede4ff291f · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 74

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.353449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:a2d267f18d398dcd4b42671ed277dc0099787c6d157182cc2658a5899d925cae

Observation 3d01b7e6-ea9d-4d9c-9583-79610d868e0d · outbound

This paper cites Scaling Laws for Neural Language Models.

A Survey of Scaling in Large Language Model Reasoning Scaling Laws for Neural Language Models

Reference 75

Resolution
verified exact
local_arxiv, observed 2026-05-22T21:22:09.259984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:df983a711cf7aafeb79aab44c451dadd9e7e4585011256ab6518f573ee2e2542

Observation 17b8b28a-5cad-4495-a643-0e9c7287528b · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 76

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.347352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:76d03a41369b2748f6839d6acc64e3d25d11c8b2ab50844528e4a3c8a08ef55f

Observation fe47977d-0797-4bdf-9880-d6f2f97de9b2 · outbound

This paper cites On scalable oversight with weak LLMs judging strong LLMs.

A Survey of Scaling in Large Language Model Reasoning On scalable oversight with weak LLMs judging strong LLMs

Reference 77

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.323611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:5f95dbfe341d371c0361f16a8025403f4813775c74e0c3ec868271ffe154bf51

Observation 670fcdd4-f431-4864-ade9-70c64403338d · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 78

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.341055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:bb1e4982939d01c9fccc38e299749f25a986dd64e34a2c076ab53cf382b41a11

Observation 1c9ff720-38ab-4e9e-ab89-d674d412b1de · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 79

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.220107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:931404b69e1a5aa29abd9507ad75275dc197bfa9f012d361db0847fc75176d7b

Observation e5c18a69-201e-4ce7-9004-323ddbd78f97 · outbound

This paper cites In Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing.

A Survey of Scaling in Large Language Model Reasoning In Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T21:22:10.333182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:feb9123891b0422b78b7c63a187466070e996ddc4efe963522d730519bc4cad9

Observation 6e38f73c-36e8-4f19-a155-76af83f1f4b4 · outbound

This paper cites Human Implicit Preference-Based Policy Fine-tuning for Multi-Agent Reinforcement Learning in USV Swarm.

A Survey of Scaling in Large Language Model Reasoning Human Implicit Preference-Based Policy Fine-tuning for Multi-Agent Reinforcement Learning in USV Swarm

Reference 81

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.218918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:1eb216bc02ce7defc405d52b4b186f8f44666988fdd3a44dd029e78284fe1659

Observation 47943f11-022f-4da2-a191-9f1339aadd19 · outbound

This paper cites Self-Generated In-Context Learning: Leveraging Auto-regressive Language Models as a Demonstration Generator.

A Survey of Scaling in Large Language Model Reasoning Self-Generated In-Context Learning: Leveraging Auto-regressive Language Models as a Demonstration Generator

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:08.925526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:5c960c70878d1630e1a689a2a9fa1045de748b788be4d803f08402ad9db41d94

Observation 39bc49d5-f654-46b3-bbde-1128ca0e7343 · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 83

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.318348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:01c6bc5b33d5187132f782ba056082029f087847d3a9880b0ea6f4f59eb3390a

Observation 4697bfd0-429c-469f-b23f-1594e78fb83a · outbound

This paper cites Advances in Neural Information Processing Systems 37 (2024), 79410–79452.

A Survey of Scaling in Large Language Model Reasoning Advances in Neural Information Processing Systems 37 (2024), 79410–79452

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T21:22:10.310480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:2a1bba57c1fffe6cb1e51ab4e24c73fb77a9d1d297f3b818f77c78582bda563f

Observation 6b425419-ad0a-48d8-9215-f5047af6627b · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 85

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.306331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:98474212747aa8879f17e4f70fd9fc2c44123914f2a1975770310bba7a3b01eb

Observation 4defa957-ee38-4bf7-88bf-7fe8e9ab6fce · outbound

This paper cites Understanding the Effects of Iterative Prompting on Truthfulness.

A Survey of Scaling in Large Language Model Reasoning Understanding the Effects of Iterative Prompting on Truthfulness

Reference 86

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:08.887136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:393d2ca0932180f526cf888d9fe5d5093dfdacd7e625e14699d55c8ac347bc94

Observation 47fe58b2-2d95-4b29-9aee-57fbd154bb4e · outbound

This paper cites Overthink: Slowdown attacks on reasoning llms.

A Survey of Scaling in Large Language Model Reasoning Overthink: Slowdown attacks on reasoning llms

Reference 87

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.140630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:07d5e05ce6cfe5dd42aedf3afb408efbcb235ad0a1a628ebaf51e12247e2106e

Observation 84a65939-9faa-4e51-b4b4-2bae29577fd1 · outbound

This paper cites LLM Post-Training: A Deep Dive into Reasoning Large Language Models.

A Survey of Scaling in Large Language Model Reasoning LLM Post-Training: A Deep Dive into Reasoning Large Language Models

Reference 88

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.504234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:a429ebf9e695283e3b46d6d66d645c99403491b14cff66554026ff94e9ab3fc2

Observation 73469dd1-e9df-4b99-84b8-a45efbbf4850 · outbound

This paper cites H-CoT: Hijacking the Chain-of-Thought Safety Reasoning Mechanism to Jailbreak Large Reasoning Models, Including OpenAI o1/o3, DeepSeek-R1, and Gemini 2.0 Flash Thinking.

A Survey of Scaling in Large Language Model Reasoning H-CoT: Hijacking the Chain-of-Thought Safety Reasoning Mechanism to Jailbreak Large Reasoning Models, Including OpenAI o1/o3, DeepSeek-R1, and Gemini 2.0 Flash Thinking

Reference 89

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.135337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:08670425c59f392095fd0f50d370a0ba62c8317ce733ac20e38fc35042b48a1f

Observation 93d8fe24-4734-45ac-95c3-ef3c3b90dc81 · outbound

This paper cites Multi-Agent Causal Discovery Using Large Language Models.

A Survey of Scaling in Large Language Model Reasoning Multi-Agent Causal Discovery Using Large Language Models

Reference 90

Resolution
verified exact
arxiv_id, observed 2026-05-27T02:05:00.588917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:7379f6c9ba7c2e5676f976f50cd869ad2ad1379720318cf1e7fe3e56b405da89

Observation a9c2f343-49d7-47b4-903c-5875d588a237 · outbound

This paper cites Inference Scaling for Bridging Retrieval and Augmented Generation.

A Survey of Scaling in Large Language Model Reasoning Inference Scaling for Bridging Retrieval and Augmented Generation

Reference 91

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.312768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:272a62b2d378563584df6f5b099960c076d3f83eaeec85b102073d7113a2cc6d

Observation cf5c1308-0075-4412-b1f4-7cc62ff40bef · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 92

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.273686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:6b2c63cef607cf4a957e5b136c8fb82ad7a1ff7d26cf43fb9bb67533e1b8c805

Observation 0f164dff-0bad-439c-9255-aa3e1a446d28 · outbound

This paper cites Harnessing Large Language Models for Disaster Management: A Survey.

A Survey of Scaling in Large Language Model Reasoning Harnessing Large Language Models for Disaster Management: A Survey

Reference 93

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:08.764590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:a65db5063772dbeea359104f709081fcc6f325b79800a9f9804510df5b157d4f

Observation 08c425f3-eb8f-43ae-8f35-43089d417208 · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 94

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.262870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:e7317b68b3b2b35d96dd56db11ce84c673cbe45a1863844ddb76e6c057b033a5

Observation 656fe646-9998-44a9-afbe-34bad594332e · outbound

This paper cites From generation to judg- ment: Opportunities and challenges of llm-as-a-judge.

A Survey of Scaling in Large Language Model Reasoning From generation to judg- ment: Opportunities and challenges of llm-as-a-judge

Reference 95

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:08.788678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:98b8b2f554cdc76caf48551a1802329fbc60105bff89fb4e475408e9616e0988

Observation fb7d6aac-fb2f-49b4-9359-6623d2a82992 · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 96

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.248267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:051637f2ee9c6ed3ae10eea64620c433a68a7360318e4150e0e6e9dabe76549c

Observation d4442445-3340-4eb0-b89c-21c0f3028ba4 · outbound

This paper cites INVESTORBENCH: A Benchmark for Financial Decision-Making Tasks with LLM-based Agent.

A Survey of Scaling in Large Language Model Reasoning INVESTORBENCH: A Benchmark for Financial Decision-Making Tasks with LLM-based Agent

Reference 97

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:08.963610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:9b2ecdc221012185c82f5a2731a2643876a726ee42cfa32615645cc2736aa25d

Observation 5e8a3a82-cc13-4dd8-b2d9-82c906fa1ca2 · outbound

This paper cites The Web Can Be Your Oyster for Improving Large Language Models.

A Survey of Scaling in Large Language Model Reasoning The Web Can Be Your Oyster for Improving Large Language Models

Reference 98

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.186902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:ca8f6a3cc5352ce83222c8e1868fa10c1ae8cc3348cf4879ba196fc15493b5ba

Observation 2123d773-c803-4469-8979-8c304932ed28 · outbound

This paper cites In-Context Learning with Many Demonstration Examples.

A Survey of Scaling in Large Language Model Reasoning In-Context Learning with Many Demonstration Examples

Reference 99

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.443530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:4c590b2ac4cd0c4d4a15e4686c6baa391166359795b181735d6671d827ee12f2

Observation 7bae226b-3fda-4278-babc-79873e33b3ca · outbound

This paper cites ParaICL: Towards Parallel In-Context Learning.

A Survey of Scaling in Large Language Model Reasoning ParaICL: Towards Parallel In-Context Learning

Reference 100

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.438889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:e16c822551d0478fe341b6c391ce78c1ca75a91f3338728a6e2f5c3e9b0233ec

Observation 59b9fc34-209f-4705-a7f5-c96c010a895c · outbound

This paper cites SCBench: A KV Cache-Centric Analysis of Long-Context Methods.

A Survey of Scaling in Large Language Model Reasoning SCBench: A KV Cache-Centric Analysis of Long-Context Methods

Reference 101

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.420903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:84282b35aaf14379b806faf6b7bd4d31ba70c9b7f7d27a71c9ebfe1fdac2776f

Observation 55806ae6-d51e-4433-9b1e-e7677b8cde95 · outbound

This paper cites an unresolved cited work.

A Survey of Scaling in Large Language Model Reasoning Unresolved cited work

Reference 102

Resolution
unresolved
raw_fallback, observed 2026-05-22T21:22:10.238273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:8704db4280969ab7f43133e6a5cb628e6e2758269139df3a7b2740d88fb9e198

Pith citing papers

Observation c1176bd8-16cb-4712-9bbd-06447b91b84e · inbound

Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models cites this paper.

Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models A Survey of Scaling in Large Language Model Reasoning

Reference 115

Resolution
verified exact
local_arxiv, observed 2026-05-12T08:40:41.940442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T08:40:40.910461Z digest=sha256:1bdb41346c2957249a664f14a7d8b084e401eb22912c36fd682e87c0f743846b

Observation c9e92fa3-d902-4cbd-b88f-4bca7635fb6b · inbound

A Survey of Context Engineering for Large Language Models cites this paper.

A Survey of Context Engineering for Large Language Models A Survey of Scaling in Large Language Model Reasoning

Reference 167

Resolution
verified exact
local_arxiv, observed 2026-05-13T20:58:45.466222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T20:58:45.060041Z digest=sha256:9e1c52f31e0e29d00a0b763db6ea0a29e56715a7384963db50b13bdc68af39b2

Observation 6b762dfc-c7d8-4da3-9a5d-b5464226ed6e · inbound

Quantifying Prior Dominance in RAG Systems cites this paper.

Quantifying Prior Dominance in RAG Systems A Survey of Scaling in Large Language Model Reasoning

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-07-01T08:35:33.343635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-01T08:35:09.009897Z digest=sha256:321823f9c0ed9f650552218736f1684981ab9156ced87bb3ce3871ef257e8dbc