Pith. sign in

Paper Citation Record · LEDGER

SLPO: Scaling Latent Reasoning via a Surrogate Policy

As of 14 August 2026, this Paper Citation Record lists 27 of 27 outbound references and 0 inbound Pith citation observations for arXiv:2607.19691.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.19691 v2

Coverage vector

measured 27 of 27 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T12:06:50.914516Z

measured 27 of 27 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

27 of 27 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved25
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 98f4168a-08d2-40ab-a803-9f9bc24e6bf7 · outbound

This paper cites Chi, Quoc V.

SLPO: Scaling Latent Reasoning via a Surrogate Policy Chi, Quoc V

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T12:06:50.361744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.361744Z digest=sha256:573a0f92cef2b100c0568470f7ea74822724bbe8c9b09807e5311ec4cbb7bb0d

Observation 4669b744-db50-4280-aa4f-4bbe32799149 · outbound

This paper cites Le, Ed H.

SLPO: Scaling Latent Reasoning via a Surrogate Policy Le, Ed H

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T12:06:50.476143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.476143Z digest=sha256:725ea3556f8878e63c79dad7c1d92985e0eb452eeefe5ebbf4e17c2c1fa26be5

Observation 467ec1b5-6a25-403e-9ad3-fc0fe407a96c · outbound

This paper cites s1: Simple test-time scaling.

SLPO: Scaling Latent Reasoning via a Surrogate Policy s1: Simple test-time scaling

Reference 3

Resolution
verified exact
doi, observed 2026-08-01T12:08:56.045800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-01T12:06:50.803213Z digest=sha256:01803e43db914fd83e1bb70763692fcc6d1c46fe06809e1ca4d47673ddf33205

Observation edf53b48-3f1e-4422-8ff2-7182b7f66e0c · outbound

This paper cites Scaling LLM test-time compute optimally can be more effective than scaling parameters for reasoning.

SLPO: Scaling Latent Reasoning via a Surrogate Policy Scaling LLM test-time compute optimally can be more effective than scaling parameters for reasoning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T12:06:50.823111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.823111Z digest=sha256:da3a4a43c27af9dddc0df6347cd440793ca9acc3472be3ecfa30e2341f0234f2

Observation a7fc9aec-8768-4a08-8677-9b45bf488850 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

SLPO: Scaling Latent Reasoning via a Surrogate Policy DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T12:06:50.827423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.827423Z digest=sha256:3d8074a75fee700120753d35aabb28d9e21744a4e91526b9dea759df08c8fec3

Observation e7eeed5c-9805-4e0c-ba32-16f4e5399aac · outbound

This paper cites Reinforcement Learning with Verifiable Rewards Implicitly Incentivizes Correct Reasoning in Base LLMs.

SLPO: Scaling Latent Reasoning via a Surrogate Policy Reinforcement Learning with Verifiable Rewards Implicitly Incentivizes Correct Reasoning in Base LLMs

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T12:06:50.831758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.831758Z digest=sha256:8e159f45bb8bf68d9e3e74eb34c460217fd08d21757ec0f5c1f70789c7a0d25f

Observation 4bcb7dc4-134a-4bc5-86fc-795749305c2f · outbound

This paper cites A Survey of Reinforcement Learning for Large Reasoning Models.

SLPO: Scaling Latent Reasoning via a Surrogate Policy A Survey of Reinforcement Learning for Large Reasoning Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T12:06:50.836674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.836674Z digest=sha256:3fc8e7c87689707a108504226d08b035678f342aa0b6fbc3d7a81df9cb843319

Observation 39689129-2aea-4c63-9fad-dc9391898c7a · outbound

This paper cites Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models.

SLPO: Scaling Latent Reasoning via a Surrogate Policy Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T12:06:50.840892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.840892Z digest=sha256:adc0bc26d39190f8a084d79a0995ccb1ff069f512d17bb23c1e850df70fea94b

Observation 5b173eeb-3447-47df-a150-207af4175eb8 · outbound

This paper cites Regular: Variational latent reasoning guided by rendered chain-of-thought, 2026.

SLPO: Scaling Latent Reasoning via a Surrogate Policy Regular: Variational latent reasoning guided by rendered chain-of-thought, 2026

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T12:06:50.845376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.845376Z digest=sha256:7225dd8b791ffaecbe5c23ac32226269b6354a0eb1ec7c967cd6e731b99d8dc9

Observation 3b060503-f3a0-491b-94d5-5139562e7341 · outbound

This paper cites Training large language models to reason in a continuous latent space.

SLPO: Scaling Latent Reasoning via a Surrogate Policy Training large language models to reason in a continuous latent space

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T12:06:50.849190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.849190Z digest=sha256:b36b9ff2affb66f0ed25e69b877dd52d562c49a529f7da022ae85a37069ed4d2

Observation f126d0d4-4ffc-4bda-a691-984cd45f8248 · outbound

This paper cites Reasoning beyond language: A comprehensive survey on latent chain-of-thought reasoning,.

SLPO: Scaling Latent Reasoning via a Surrogate Policy Reasoning beyond language: A comprehensive survey on latent chain-of-thought reasoning,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T12:06:50.853257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.853257Z digest=sha256:633e3310e5b3641dd06b6930bfcf8115ada71bc5f973a7b09d2a1c0368736726

Observation c477984f-b936-4c53-9ec0-ccd6bf7570b9 · outbound

This paper cites CODI:Compressingchain-of-thought into continuous space via self-distillation.

SLPO: Scaling Latent Reasoning via a Surrogate Policy CODI:Compressingchain-of-thought into continuous space via self-distillation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T12:06:50.861366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.861366Z digest=sha256:a10b31de0e31448dc97cf6bcf1f4ac4992ec6c727ce18890d64a2497d2be49f4

Observation 902f68a6-744e-41ed-ac7f-8260f49c8a5f · outbound

This paper cites Think silently, think fast: Dynamic latent compression of LLM reasoning chains.

SLPO: Scaling Latent Reasoning via a Surrogate Policy Think silently, think fast: Dynamic latent compression of LLM reasoning chains

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T12:06:50.865291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.865291Z digest=sha256:25b2657d7dcae6d9a4b6c6247588edff9395d37eb4134397438558ca0776f9af

Observation c99e7411-812c-4ccd-8641-49a27d1415f3 · outbound

This paper cites Sim-cot: Supervised implicit chain-of-thought, 2025.

SLPO: Scaling Latent Reasoning via a Surrogate Policy Sim-cot: Supervised implicit chain-of-thought, 2025

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T12:06:50.869443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.869443Z digest=sha256:da61439b260b18f3bfb78c834a6c7233359bdbdb5481ff95295bd1c3fcdea151

Observation 28f217ae-550c-4d4c-b0d8-b0640bc50b21 · outbound

This paper cites Render-of-Thought: Rendering Textual Chain-of-Thought as Images for Visual Latent Reasoning.

SLPO: Scaling Latent Reasoning via a Surrogate Policy Render-of-Thought: Rendering Textual Chain-of-Thought as Images for Visual Latent Reasoning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T12:06:50.873406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.873406Z digest=sha256:19ab54c40f2a311424b5972981ec43b75b80ffab983f167ac9fe36e94bd3deef

Observation c0354322-4a2e-4126-8a39-5bdf71165560 · outbound

This paper cites Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?.

SLPO: Scaling Latent Reasoning via a Surrogate Policy Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T12:06:50.877576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.877576Z digest=sha256:1b9e386817158e7acec8e4875d23644c179463401eaee9d3736b6319cb5b14dc

Observation 5d44ac5f-3876-4dc1-b36b-94cd2c1073d3 · outbound

This paper cites La, Duy M.

SLPO: Scaling Latent Reasoning via a Surrogate Policy La, Duy M

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T12:06:50.881769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.881769Z digest=sha256:417b886362b1deef97ad491012a95e8302c35ce08df357c4ed54de4a1025def3

Observation 5a9b0b3f-9d17-43c7-8244-7b4bc978dc7b · outbound

This paper cites DART:Distillingautoregressivereasoning to silent thought.

SLPO: Scaling Latent Reasoning via a Surrogate Policy DART:Distillingautoregressivereasoning to silent thought

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T12:06:50.885793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.885793Z digest=sha256:4ddb44417bf9cd0aa12181af36f875a6d1ce7128d801eb2943257817d422d11f

Observation faa32a5e-dd82-4e2d-b379-0d78c3276fdd · outbound

This paper cites LLM latent reasoning as chain of superposition, 2025.

SLPO: Scaling Latent Reasoning via a Surrogate Policy LLM latent reasoning as chain of superposition, 2025

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T12:06:50.889733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.889733Z digest=sha256:350ae93c2ef5ae9c55c992b017f3c4994e3c3703213c9725eddedaccc0837a93

Observation 16e548dc-f841-48d7-bf2a-7dd7995c63e4 · outbound

This paper cites Parallel test-time scaling for latent reasoning models.

SLPO: Scaling Latent Reasoning via a Surrogate Policy Parallel test-time scaling for latent reasoning models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T12:06:50.893604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.893604Z digest=sha256:c984f135a7680ea78bb6d9f5030123135e755c445f5516973be3679f41da3f07

Observation 08c6db03-c809-498e-a49a-3ea8b53bf8a3 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

SLPO: Scaling Latent Reasoning via a Surrogate Policy DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T12:06:50.897640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.897640Z digest=sha256:c9e94070ee8d59fdf65cca09ed62886e367282028597ca80e288161a9360b282

Observation c7e51b63-acdb-44de-a9fd-2741712cc433 · outbound

This paper cites Effective Reinforcement Learning for Reasoning in Language Models.

SLPO: Scaling Latent Reasoning via a Surrogate Policy Effective Reinforcement Learning for Reasoning in Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T12:06:50.901959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.901959Z digest=sha256:09b41a6f11d37f285d3b11baef3773911de7a5914ac38d0650c9a7e32461430b

Observation d6cfab06-5b62-478a-a928-aa918abdba56 · outbound

This paper cites LEPO: Latent Reasoning Policy Optimization for Large Language Models.

SLPO: Scaling Latent Reasoning via a Surrogate Policy LEPO: Latent Reasoning Policy Optimization for Large Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T12:06:50.905951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.905951Z digest=sha256:88f17ac3160d97673ae62cb81763993e2e07538c0ce65dd269005c40634b09ad

Observation 19ceb98c-c811-408f-bd69-3655c821862a · outbound

This paper cites Latent-GRPO: Group Relative Policy Optimization for Latent Reasoning.

SLPO: Scaling Latent Reasoning via a Surrogate Policy Latent-GRPO: Group Relative Policy Optimization for Latent Reasoning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T12:06:50.910188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.910188Z digest=sha256:0c41273bee8c998e78f31c8a81abbbd0b842d578c4a2fbdc3eb47f0df8459f4a

Observation 741454f6-b722-4f12-b3b7-fe3fae7eeff5 · outbound

This paper cites Soft Thinking: Unlocking the Reasoning Potential of LLMs in Continuous Concept Space.

SLPO: Scaling Latent Reasoning via a Surrogate Policy Soft Thinking: Unlocking the Reasoning Potential of LLMs in Continuous Concept Space

Reference 25

Resolution
malformed identifier
no resolver link, observed 2026-08-01T12:06:50.914516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.914516Z digest=sha256:0302baad406cd0fdd3c5147237f72a9f82c1b1d4bf77f1a59292f8aedcb267e3

Observation 61fd0556-9739-4f26-b70b-77070b783ae1 · outbound

This paper cites an unresolved cited work.

SLPO: Scaling Latent Reasoning via a Surrogate Policy Unresolved cited work

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-01T12:06:50.642661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.642661Z digest=sha256:8ad2081e4235a51200d142127996a82eb1311fb100096494ab5364dcf9b23150

Observation ff319080-1438-4872-8a44-b5cbabfddb53 · outbound

This paper cites an unresolved cited work.

SLPO: Scaling Latent Reasoning via a Surrogate Policy Unresolved cited work

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-01T12:06:50.857302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:06:50.857302Z digest=sha256:a7c8e9f60271fc089440ae24ecd0afcaaeb6bad7944dca15dacd0f163524dce5

Pith citing papers

No inbound Pith citation observations are available.