Pith. sign in

Paper Citation Record · LEDGER

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs

As of 11 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 0 inbound Pith citation observations for arXiv:2505.19075.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19075 v3

Coverage vector

measured 48 of 48 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-22T01:16:51.288077Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

48 of 48 outbound references displayed

  • verified exact24
  • verified fuzzy22
  • unresolved0
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7915efde-4014-49f2-9561-76044141a237 · outbound

This paper cites https://open-thoughts.ai.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs https://open-thoughts.ai

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T01:20:53.059308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:55b152ea77df6b979bccc848069defccdc0ef33dbb58628e87994b1697b23d78

Observation 2f6d41ee-b12f-4e8d-ac27-7c9ac5e7e1ed · outbound

This paper cites A distributional view on multi-objective policy optimization.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs A distributional view on multi-objective policy optimization

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T01:20:53.056170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:2ffd6a68bf475b462ae2cecbb1e6197950e11b11b6fed9b4eb0ba78869b9b32b

Observation 71ad09b3-2af8-4443-b8aa-0fb255219a3c · outbound

This paper cites Training language models to reason efficiently.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Training language models to reason efficiently

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-22T01:20:52.113085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:6e10d31691e3ef7717423921f29492c6c8f2234c132fe4b2cd02b59f94cc5816

Observation 7b682a7a-fdfe-477a-be31-469344dce1f5 · outbound

This paper cites Qwen2.5-VL Technical Report.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Qwen2.5-VL Technical Report

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-22T01:20:52.203476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:f117c6371f905d82042ae788a55f2a22c822dbbb5f003b0b8e78d401587404ac

Observation 47b5cb35-68e6-4067-98ac-baa9999a75aa · outbound

This paper cites Bespoke-stratos: The unreasonable effectiveness of reasoning distillation.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Bespoke-stratos: The unreasonable effectiveness of reasoning distillation

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T01:20:53.082399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:e76b75435ae5cfa6d1b28a2ddc784642f7d8d7e3861e84a54ca0ea1abe2b0548

Observation a6aaa778-a682-4be5-a3ec-56cec7314d6c · outbound

This paper cites Overview of the iwslt 2017 evaluation campaign.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Overview of the iwslt 2017 evaluation campaign

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T01:20:53.052749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:4fa9643cfa800776527c8e889194e74cbb87132d10b4bf7c5d7ef8268dab2114

Observation c4cef4f9-3fb2-4505-a16f-8cb7e023fb7e · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Training Verifiers to Solve Math Word Problems

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-22T01:20:52.198563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:93bcda734e28cdc0500fe421cf6202dd3a6f52f2fe0890e97b4bb6f696442e04

Observation aad8307f-e2ab-41f3-9028-b9f607f87145 · outbound

This paper cites Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T01:20:53.126831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:5af2c14308961fa4ff56164dbb76f820e7cffafb0f2821d7d36ba0bf68478a99

Observation 872b89c1-d2e2-4f20-b0b6-7124c13dd8c9 · outbound

This paper cites Controlled Text Generation via Language Model Arithmetic.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Controlled Text Generation via Language Model Arithmetic

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-22T01:20:52.182731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:1d24c919eb899720f08c1090e7f2d1388ef9fa2b1ef76fa33b40b8887e31ff76

Observation 3837ff0d-b0b5-45f2-91e7-47ed2836b30a · outbound

This paper cites Agent AI: Surveying the Horizons of Multimodal Interaction.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Agent AI: Surveying the Horizons of Multimodal Interaction

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-22T01:20:52.101452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:4bf57e4dbd17e1ce92ab17301664e4a3943b5a2de1775f484beebed1289ba845

Observation a11ac186-d2f9-44e8-a5ee-99cf35f7bd09 · outbound

This paper cites MT-R1-Zero: Advancing LLM-based Machine Translation via R1-Zero-like Reinforcement Learning.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs MT-R1-Zero: Advancing LLM-based Machine Translation via R1-Zero-like Reinforcement Learning

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-22T01:20:52.107060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:c8cb3435297e27f6420f8a4753fb6307209bbdf6cb0dc61c05e44c7c07a7f92b

Observation 9f5ebffa-9cee-4643-8d8b-ef82f2eb9ba6 · outbound

This paper cites Scaling laws for reward model overoptimization.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Scaling laws for reward model overoptimization

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T01:20:53.049385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:518fbba761d07e1897a5b9021ae595151a67da71c686a69053943a0f8a7c6c6f

Observation 056ea82c-06dc-487e-a52d-4e3767f22c0f · outbound

This paper cites The Llama 3 Herd of Models.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs The Llama 3 Herd of Models

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-05-22T01:20:52.079264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:613b2146faa1181f1782f590d346b0a1cc35995c9510364a1dc89cdf9f27df67

Observation 72068d0f-88ac-409c-be81-58046c28229c · outbound

This paper cites xcomet: Transparent machine translation evaluation through fine-grained error detection.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs xcomet: Transparent machine translation evaluation through fine-grained error detection

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T01:20:53.123230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:8bf083c431e4c5470bfe6e6abf335d1fd3560f60e95f5c5cd6330b9abd94c18f

Observation 1f00b3c2-cd6d-4976-8a55-85cc0fde0f2b · outbound

This paper cites Reinforcement learning with deep energy-based policies.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Reinforcement learning with deep energy-based policies

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T01:20:53.119671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:62f87f5ee12de110e200f9c04802330b2c9fe1792e7fbdd68f66ee6e66691a2a

Observation 28d7c55c-27d9-4503-ad9d-13a8f3f271d3 · outbound

This paper cites Value Augmented Sampling for Language Model Alignment and Personalization.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Value Augmented Sampling for Language Model Alignment and Personalization

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-22T01:20:52.085102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:d2c704769af792fb22cb218e0ebc3ee98fbeaf25ed506c5140375afedf8eb6a0

Observation aecab8b7-1f59-4848-b429-40e0533658f6 · outbound

This paper cites OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-22T01:20:52.193412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:17e8e4ab81e1d4a597084aba9c0c8732c06484e712479b39e9a6f1bd20140685

Observation 8a64ad45-09d0-477c-b8a3-cdc2851330af · outbound

This paper cites Lora: Low-rank adaptation of large language models.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Lora: Low-rank adaptation of large language models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T01:20:53.116330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:58e384334cdc0b5c8f798901e83326a83171f1fe327361f118c497a617683f8a

Observation 86da1fd4-3498-45fa-9415-d6c3aa076176 · outbound

This paper cites Open-reasoner-zero: An open source approach to scaling up reinforcement learning on the base model.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Open-reasoner-zero: An open source approach to scaling up reinforcement learning on the base model

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T01:20:53.112995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:0be1ba691435971b73348adc5fb9e487acf4e968fa87e7bd5ddb4c484a223150

Observation fdde4249-e2f8-4bf6-9e55-d2f4e23670b3 · outbound

This paper cites Solving quantitative reasoning problems with language models.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Solving quantitative reasoning problems with language models

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T01:20:53.109398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:2794252073e1591ff9a34512266b30e97d14f48653d6b3ba566c69f8965410b8

Observation 8220def0-689a-42ca-a85c-828cfd0329b4 · outbound

This paper cites RAIN: Your Language Models Can Align Themselves without Finetuning.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs RAIN: Your Language Models Can Align Themselves without Finetuning

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-22T01:20:52.188334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:6a070096fe0c5b22ff3acfdea01183edd65a846a516cb6830f345a518ef7c9a4

Observation 6ffd0ca5-2957-4209-ab91-054b8e481c64 · outbound

This paper cites Let’s verify step by step.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Let’s verify step by step

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T01:20:53.105883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:b9a905e1ecff07c61889a40ab587314d9a5f7eea358d688e5c49c64083103ae6

Observation 0645521b-7d67-4b62-a131-be0b61326922 · outbound

This paper cites Making ppo even better: Value-guided monte-carlo tree search decoding.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Making ppo even better: Value-guided monte-carlo tree search decoding

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T01:20:53.102722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:75ec55d61da2ef817c413dc26d4fae0f390c188a502026510c5dc022b0a54688

Observation ce7c1b11-dec3-4674-9a12-f52092031b24 · outbound

This paper cites Understanding R1-Zero-Like Training: A Critical Perspective.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Understanding R1-Zero-Like Training: A Critical Perspective

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-05-22T01:20:52.176802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:3a2b2014b3831dfdab0cda085ca6c490fbb2df34a42cdde092c9c6c137ca24bd

Observation 2aad626e-20b7-46ea-a334-ef43ac666b0c · outbound

This paper cites Inter-GPS: Interpretable Geometry Problem Solving with Formal Language and Symbolic Reasoning.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Inter-GPS: Interpretable Geometry Problem Solving with Formal Language and Symbolic Reasoning

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-22T01:20:52.171935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:58af03f5ecee29b0544c670f40b2c0415700bef42318b1ecdb7b6cdcae0dcdc8

Observation f976f99d-ee17-42b6-bdd3-06c5bc560059 · outbound

This paper cites Controlled Decoding from Language Models.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Controlled Decoding from Language Models

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-22T01:20:52.166735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:8b4f38ec80cda314b019b50b0917fca5dbefe4c1ebdf654757c6eef880e7123e

Observation 383dbb4b-e39d-4bde-99cb-c9145cedabdd · outbound

This paper cites s1: Simple test-time scaling.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs s1: Simple test-time scaling

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T01:20:53.099580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:e2038492ab3f06a5150463dd9e6915bc069b9814da8d9ee4ceb43cb721e75386

Observation 8d8d4176-d223-404f-a5c6-e5207e1d0711 · outbound

This paper cites Learning to reason with llms.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Learning to reason with llms

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T01:20:53.096141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:a80dd22d7bfadc22295143c76a84f3d227295a5528daa8b92da1e072b5e45829

Observation 7a2ccf24-1ca3-47b3-868e-26c56f619e68 · outbound

This paper cites BOLT: Bootstrap Long Chain-of-Thought in Language Models without Distillation.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs BOLT: Bootstrap Long Chain-of-Thought in Language Models without Distillation

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-22T01:20:52.209043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:051f613cf47cc4c3eda4613ec7b9e18c14124099a91a88ec3ae5f07583f74460

Observation 2c63b656-1f92-4a2e-8eee-6d9ebc65ac1d · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Bleu: a method for automatic evaluation of machine translation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T01:20:53.093173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:c5af9ef92444cfaab68cb661fabb07feb4ccc604dccbcd554132409fdbc86d55

Observation 65cfd4e5-9954-4cff-9412-6697d76c7c98 · outbound

This paper cites Direct preference optimization: Your language model is secretly a reward model.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Direct preference optimization: Your language model is secretly a reward model

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T01:20:53.089363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:4746a270f00e5f0d75519a9f727af1a6a0a2b3d9af55f30299da595d8f2c0974

Observation 222f1f60-7fab-46a8-bc84-085e4f79878c · outbound

This paper cites CometKiwi: IST-Unbabel 2022 Submission for the Quality Estimation Shared Task.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs CometKiwi: IST-Unbabel 2022 Submission for the Quality Estimation Shared Task

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-22T01:20:52.161220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:c1e2f01a8b8f9476831fb648d8c624f18acaa7548acad90af59dde6b42f0f5e3

Observation ec53a422-5e92-4b71-8b4b-c2660be71598 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Proximal Policy Optimization Algorithms

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-05-22T01:20:52.090514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:2fb2ba8fbf3f6dad9170fd4dad6d4871c5601c8bb5a00444c17aec6e985b1631

Observation 3fbcfc41-5b47-4203-ab4c-c532860b1a4a · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-05-22T01:20:52.096042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:e53003b1c7d28fef1e22beb69fa93891037f8d42b3de5b836612871a5643bf50

Observation 76f6d3ad-11b5-4b7a-8d9b-8852031dfa8c · outbound

This paper cites Language models are multilingual chain-of-thought reasoners.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Language models are multilingual chain-of-thought reasoners

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T01:20:53.086110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:d824e5b85543f01f01f1d3e2749cfc0d0a5ec14e7cde29369b2e8e10d4df0f2d

Observation 85e3a1ec-6c4e-408e-ad20-e2303f14f675 · outbound

This paper cites Offline RL for Natural Language Generation with Implicit Language Q Learning.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Offline RL for Natural Language Generation with Implicit Language Q Learning

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-22T01:20:52.155945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:08303de2306c1b9e77d843efc196ad1400c30caab6caf0bd2a41c3b8031e1cef

Observation a9dca9a3-4128-4376-acc7-12a18e9f0db2 · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-05-22T01:20:52.144708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:925f731fd17cbc5ecc288a75c0e878ff859680146163b647520016b155361a26

Observation 35c3cccf-e86d-4078-bcba-aa6592f5efe7 · outbound

This paper cites Inference-Time Alignment in Diffusion Models with Reward-Guided Generation: Tutorial and Review.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Inference-Time Alignment in Diffusion Models with Reward-Guided Generation: Tutorial and Review

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T01:20:52.134739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:84cb220a6cdfa5066bff5e5ceeab427a384be79bb557258e92ea908b17dc433b

Observation 04056c00-3775-4bae-8598-f20d88600f84 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Chain-of-thought prompting elicits reasoning in large language models

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T01:20:53.079265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:a8d8d20818808dfe519247358e7f6c12ef8217f8651afe2052b93b6b0f38ec66

Observation 513ae582-aa7c-4aba-a346-c1826eabce63 · outbound

This paper cites GenARM: Reward Guided Generation with Autoregressive Reward Model for Test-time Alignment.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs GenARM: Reward Guided Generation with Autoregressive Reward Model for Test-time Alignment

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-22T01:20:52.150446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:3eb4e9967114e8397cd1761fd58a3a0a53468b4318c5d83f0393e37e9f370625

Observation f55dfc8e-a788-4e6a-a5da-968495c6c492 · outbound

This paper cites Qwen2.5 Technical Report.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Qwen2.5 Technical Report

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-05-22T01:20:52.139872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:6749cee0750513103c3cacc65194a11664f2c3e966ed9952f4ff75ecd87b5f3c

Observation a9e5fb04-bd7b-46f4-a12c-9921810d1942 · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-05-22T01:20:52.129283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:f3d68662ee27b7194551c2da6b6778bbd9be95fde7fde589cecf7add1acfcb45

Observation 7e816517-ff3e-4e87-8683-8cda32623abd · outbound

This paper cites Preference-grounded token-level guidance for language model fine-tuning.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Preference-grounded token-level guidance for language model fine-tuning

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T01:20:53.074920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:31264c06d2ec19dda0310ddaf0acc5c6bbe66c5e9423edc5ab61c84974983973

Observation e627d3e4-b2a9-4f21-93e2-0d3947444551 · outbound

This paper cites Limo: Less is more for reasoning.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Limo: Less is more for reasoning

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T01:20:53.070822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:6cda16d829c05836fdd7f919c462793e1fa7fd912d6f4dcce85c08531c354318

Observation 69316f9c-050b-4ccf-8937-93ae51d826a2 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-05-22T01:20:52.124513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:1c680b60febaa0a363ef2477d4b5d922db54072e200faffadd5ae027d4cc844c

Observation d78bc88d-8e1a-4fcd-a8da-02a88cfa8749 · outbound

This paper cites LoRA-FA: Efficient and Effective Low Rank Representation Fine-tuning.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs LoRA-FA: Efficient and Effective Low Rank Representation Fine-tuning

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-05-22T01:20:52.118075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:06596685bf42bfcfc0f1c9e4ef2cbc330bfa67363e1d8dbfa3cc56a5420e7e86

Observation 74b1e685-a579-48e0-99de-47063553b50f · outbound

This paper cites Mathverse: Does your multi-modal llm truly see the diagrams in visual math problems? In European Conference on Computer Vision, pages 169–186.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Mathverse: Does your multi-modal llm truly see the diagrams in visual math problems? In European Conference on Computer Vision, pages 169–186

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T01:20:53.067029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:b88f0b35f5e8346ac39526c9dae4479baae264c40ba279f2242631d83b2e9fef

Observation 13641742-2f5d-4b0b-956c-2327087be5bf · outbound

This paper cites What about.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs What about

Reference 48

Resolution
malformed identifier
raw_fallback, observed 2026-05-22T01:20:53.063230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:7b752e7f0f839b12e4bcb90892e68616ac795ddf76f32a0b46f842df938ac6c2

Pith citing papers

No inbound Pith citation observations are available.