Pith. sign in

Paper Citation Record · LEDGER

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs

As of 6 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 2 inbound Pith citation observations for arXiv:2605.09063.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.09063 v3

Coverage vector

measured 33 of 33 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-20T22:24:04.625823Z

measured 35 of 35 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T07:51:11.899943Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-03T05:47:40.989192Z

Reference resolution

33 of 33 outbound references displayed

  • verified exact16
  • verified fuzzy7
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 38c89c7f-3209-4905-8b46-a641290d7539 · outbound

This paper cites First proof.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs First proof

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T22:24:07.808603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:dfceb026d571773422aeac4150c55d3400d5a1d4a4e7e37bb8668c544791b3ab

Observation cdae6753-e705-4059-8556-940d462018bc · outbound

This paper cites gpt-oss-120b & gpt-oss-20b Model Card.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs gpt-oss-120b & gpt-oss-20b Model Card

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T22:24:07.725125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:5fb5ae1950301ffcee2f82c9a2b371c0a77a5d19ac07178edec7bb193bfc9684

Observation 85c8bf02-d036-431b-ade6-140fe17ab51b · outbound

This paper cites Short proofs in combinatorics and number theory.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs Short proofs in combinatorics and number theory

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:24:07.813826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:efba83f690671c1947b9d725c94d227e73b55a99f1a4b01429ffc23eadb5a61b

Observation f89f8194-a745-4f4d-8f3f-04bf5530b99d · outbound

This paper cites Short proofs in combinatorics, probability and number theory II.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs Short proofs in combinatorics, probability and number theory II

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-20T22:24:07.738214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:f4e800ef8ec2904a025dcf06ed3d4fd63282d3decb2b14792a9988287a903c29

Observation c864122a-58ab-4fb5-a663-b762f98c04a2 · outbound

This paper cites arXiv preprint arXiv:2510.26768 , year=.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs arXiv preprint arXiv:2510.26768 , year=

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:24:07.763874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:8b63c7ff0840db4e8c38dc5d513571adc240f65c3259791f62412e5b04b911df

Observation 2aad720e-23ce-408a-a2fb-784ad312714c · outbound

This paper cites Introducing Claude Opus 4.5.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs Introducing Claude Opus 4.5

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T22:24:07.870333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:a9c64760b3f1923a0253f946145efc73a9a758026b919941c612583dee30cc58

Observation 6012d70c-aa78-43a1-beb8-7e0b003f744f · outbound

This paper cites American invitational mathematics examination (aime).

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs American invitational mathematics examination (aime)

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T22:24:07.854933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:5e05b4a6711ef03b636fd87d1270ce5eeb8b0293a2fc1d1fa7144a0f23617f5e

Observation f1b1ae87-2f79-4e0a-9a1e-d8c4e2ad425a · outbound

This paper cites an unresolved cited work.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-05-20T22:24:07.842239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:578672e3f85940e96a637a73999dc2405cfc200c6793f3349d32f3aeff610e94

Observation 1e19bb32-85f4-4727-8f1f-879be06ab68c · outbound

This paper cites an unresolved cited work.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-05-20T22:24:07.863939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:6b98ff71aaa81f928e2ffa00b19f4a4abee6fe602dc836db872112a012c1aae1

Observation 0d4cbd5c-3505-42aa-b55a-ddb77199a841 · outbound

This paper cites BeyondAIME: Advancing Math Reasoning Evaluation Beyond High School Olympiads.https://huggingface.co/datasets/ByteDance-Seed/BeyondAIME.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs BeyondAIME: Advancing Math Reasoning Evaluation Beyond High School Olympiads.https://huggingface.co/datasets/ByteDance-Seed/BeyondAIME

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T22:24:07.846216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:3ddf909a94543ea488026fd61f9b97b1733ad8190410e7b9d44e8a06989b587e

Observation 9a03e198-29f0-4da6-9689-5c790437e349 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs Training Verifiers to Solve Math Word Problems

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-05-20T22:24:07.758847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:ce8d562c394bb36de072f3c8804e25baf4d003a3143a49f8c0d7727f4951f003

Observation 2c511fb0-790b-4135-b25e-d015c5fd964f · outbound

This paper cites Semi-autonomous mathematics discovery with gemini: A case study on the erd\h{o}s problems.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs Semi-autonomous mathematics discovery with gemini: A case study on the erd\h{o}s problems

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:24:07.791742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:d554e85f9c404ca9d00cecf1175b465f22384ca3c82b491f34f66fed9e8ac225

Observation 0432c9a3-ce2d-4d1b-bd25-f8a8e39e01c2 · outbound

This paper cites an unresolved cited work.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-05-20T22:24:07.849288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:6331dd0a5b27f4272b5c251f71c599546cd0b5bf620a1e2abf4ea8067449e949

Observation ec04da43-7cfd-4db6-aef1-7502e4f8a40e · outbound

This paper cites Riemann-Bench: A Benchmark for Moonshot Mathematics.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs Riemann-Bench: A Benchmark for Moonshot Mathematics

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-20T22:24:07.752946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:17ff47fc13fc70980e2e6a7d4c3e51e5550dc04b9d182e2aea1d650d38c11ef9

Observation 5d266f89-6a95-46b4-a4c1-16f6e7516e9e · outbound

This paper cites FrontierMath: A Benchmark for Evaluating Advanced Mathematical Reasoning in AI.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs FrontierMath: A Benchmark for Evaluating Advanced Mathematical Reasoning in AI

Reference 15

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T22:24:07.822738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:a2472e6fd2be2031624ef6a2fa6fe7595a3ced1e62723f6c34465c5d88a3b1ec

Observation f77d596c-4f38-44c8-8634-43f8a56a958b · outbound

This paper cites Gemini 3.1 Pro.https://deepmind.google/models/gemini/ pro/.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs Gemini 3.1 Pro.https://deepmind.google/models/gemini/ pro/

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T22:24:07.876023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:1a01f58b434ee700f989a7a67fd5c764976d7f9f554ce5a2c8755d262c8466b7

Observation 7ad1f265-46d6-4070-9ff8-6a16a1c643f1 · outbound

This paper cites OpenThoughts: Data Recipes for Reasoning Models.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs OpenThoughts: Data Recipes for Reasoning Models

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-20T22:24:07.778193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:510b8a9a53d52eaf514f90967daf064761b481b2f61064c36d98de2dd7f393be

Observation bb8caf00-18e3-49f5-83f6-2b8da5d90cde · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-05-20T22:24:07.786902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:b7e9cc2265127090238726eddec4fea8f8948f77e811f9ba9a538e3195dd5501

Observation 9c1c6649-0ad1-4e81-9bf0-b632dc4c54a5 · outbound

This paper cites an unresolved cited work.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-05-20T22:24:07.867474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:eb34d482ac47b1e6331e6920f69fd563507cfa3de2dd2d442a8423204aad0d64

Observation ea3e4703-0df2-4af1-844c-1df4817614e9 · outbound

This paper cites HMMT.https://www.hmmt.org/.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs HMMT.https://www.hmmt.org/

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T22:24:07.873193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:8de43d7fc32b22d58de2afd9cbce7d427ae4fb61d7165ea57e2e5ecf95fc5ac5

Observation 63512ddc-8487-4c59-ac47-4d4877d75684 · outbound

This paper cites an unresolved cited work.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-05-20T22:24:07.839437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:4b5bfd47b69a25154cd215232d86597480429aa7e1c4d8c8b5e58aab014ee675

Observation 24da91b6-5249-4c59-97b9-e8ff5dde789c · outbound

This paper cites EternalMath: A Living Benchmark of Frontier Mathematics that Evolves with Human Discovery.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs EternalMath: A Living Benchmark of Frontier Mathematics that Evolves with Human Discovery

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-20T22:24:07.772977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:e1105dd5252ceddc7d0d8072ae659c0cc2994357f21a194a7f06ef725217ad32

Observation 210423f8-17a3-4993-baa9-a30ced3afdc9 · outbound

This paper cites proprietary ai foundation model.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs proprietary ai foundation model

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T22:24:07.852027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:53e737a4435ee7bf087bf365fea3b73e7ec8a184317ccd57397c8d59c3e41f5a

Observation 3e7aab8a-d5e1-4247-bed9-e661859d4a38 · outbound

This paper cites Humanity's Last Exam.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs Humanity's Last Exam

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-05-20T22:24:07.747968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:445cd9be24c9a1ab85c3d815bd776096ed0a2c4d60ba9f21f801a9fea32170a7

Observation f01528dd-f169-4bd6-a94c-d75730cf07d8 · outbound

This paper cites IMProofBench: Benchmarking AI on Research-Level Mathematical Proof Generation.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs IMProofBench: Benchmarking AI on Research-Level Mathematical Proof Generation

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-10T02:19:32.479333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:cc355e2961713b59e5528a358d9aa2d52701be44a6840645d5a04caa964f9b2e

Observation 0df6c967-6835-4223-965c-a391d08ef79c · outbound

This paper cites OpenAI GPT-5 System Card.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs OpenAI GPT-5 System Card

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-05-20T22:24:07.782392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:7d0d3b489c3ad4116982169ef8c0d1f44b27c4966b3e67922ecade8d4177c473

Observation 46a2f3c1-9845-48c5-b610-6870a0bdab0f · outbound

This paper cites an unresolved cited work.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-05-20T22:24:07.860967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:a19deed3a072da0b21b9bda7c2e15a29b4f43e75c83fdf4db6a98570972cfe65

Observation a6f4c571-b33e-4405-8564-4d8cf73e3a66 · outbound

This paper cites an unresolved cited work.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-05-20T22:24:07.836176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:ea7ae588c8124f65fe85d0b3dfd1d84158d0470bf8bfb0794731f940e77160ff

Observation 9d5056c3-4c56-43c9-8d32-b91ccca28b10 · outbound

This paper cites Kimi K2.5: Visual Agentic Intelligence.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs Kimi K2.5: Visual Agentic Intelligence

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-05-20T22:24:07.796186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:dbc32d77c3881de22fc8169de35df191e4340dbbe76ec3ad151974c9f5b05ffb

Observation 9c77a275-306a-459b-aa16-3ace68a761b8 · outbound

This paper cites Qwen3 Technical Report.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs Qwen3 Technical Report

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-05-20T22:24:07.800295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:28146cae7e20f8f6ac0ccdf3497ad533bb4d4c1369b77188ea02018ea394fabf

Observation c760bdf1-d822-444e-a9e8-b4da826a9d4e · outbound

This paper cites GLM-5.1: Towards Long-Horizon Tasks.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs GLM-5.1: Towards Long-Horizon Tasks

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T22:24:07.858204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:10baa02acfe752167c33c28dd91437092b427f90a076886187c22e34d22e5ddc

Observation 32a828db-cc2c-44b5-9c5c-88efc304143e · outbound

This paper cites Zhai et al.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs Zhai et al

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:24:07.828052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:d12c28a002edcfc5be214f83aa40ef37b7257944f6386ce6ed5e79c00b050104

Observation f32425d4-8502-4362-86f0-a681301cb3fe · outbound

This paper cites Sovereign AI Foundation Model.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs Sovereign AI Foundation Model

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:24:07.818261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T22:24:04.625823Z digest=sha256:ec9211f42c04e249dd330de387912914eff4468b0cbe84ba0cf98ad4f3aa5ef2

Pith citing papers

Observation 0bf0fcdc-f168-49ac-92a1-d9fc0155ca72 · inbound

The Grothendieck Constant is Less Than $\frac{\pi}{2 \log (1+ \sqrt{2})} - 10^{-5}$ cites this paper.

The Grothendieck Constant is Less Than $\frac{\pi}{2 \log (1+ \sqrt{2})} - 10^{-5}$ Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs

Reference 12

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T05:56:40.849823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T07:51:11.899943Z digest=sha256:e25648373248c4ab7b32171b834b445568d22bbc23ad13ee85b4cbc2f466240a

Observation 45112da2-8c11-43a0-94e7-905139f528e5 · inbound

KCSAT-ML: Probing Reasoning Models with Nationwide-Cohort Human Difficulty cites this paper.

KCSAT-ML: Probing Reasoning Models with Nationwide-Cohort Human Difficulty Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-07-03T05:47:40.990435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-27T13:07:59.509660Z digest=sha256:6bbe29557ad4ed81dc64359f40eae530ea86dd2e7f81e782819a7ed3065e615e