Pith. sign in

Paper Citation Record · LEDGER

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics

As of 19 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 12 inbound Pith citation observations for arXiv:2505.03171.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.03171 v1

Coverage vector

measured 33 of 33 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:04:38.456180Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:18:57.397021Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T18:17:33.661610Z

Reference resolution

33 of 33 outbound references displayed

  • verified exact1
  • verified fuzzy8
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 271822d3-054f-45d0-b4e2-499584185f04 · outbound

This paper cites Claude 3.7 sonnet, 2025.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics Claude 3.7 sonnet, 2025

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:04:39.155918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-16T00:04:38.299189Z digest=sha256:c9457ae053ee044644078a39b63aebb2de1e224e83cc176f2c3d114c2c37d2bf

Observation e36eddec-9d3c-4c99-a0a1-a35ae194a975 · outbound

This paper cites ProofNet: Autoformalizing and Formally Proving Undergraduate-Level Mathematics.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics ProofNet: Autoformalizing and Formally Proving Undergraduate-Level Mathematics

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T00:04:38.304499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:04:38.304499Z digest=sha256:f8a6e3f34123154022d8aba67885f2226dd8938730589a96f37a0fb02e4db7c7

Observation e87f5c6d-e14a-4a48-ac02-1d8295c98467 · outbound

This paper cites an unresolved cited work.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-16T00:04:39.140843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-16T00:04:38.309416Z digest=sha256:b822aa9f96c7f3444fbf90e77f6c3b82c919f0172a8625a0b603f78095339c80

Observation 5f571164-bf15-4283-90ef-c718c4c2d972 · outbound

This paper cites an unresolved cited work.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-16T00:04:39.126277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-16T00:04:38.314486Z digest=sha256:c3097855e595edbf77b962c678e53ddecd47912f5de7d05a281567f124b6e41b

Observation 38e7c6d4-4175-4071-89b1-b8a822688dbb · outbound

This paper cites an unresolved cited work.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-16T00:04:39.111430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-16T00:04:38.318943Z digest=sha256:04ba4f26f00a211aedc1508c21584eb74f3966a46462dbdbe3289c6ff599b47b

Observation bf58b874-5a58-40f9-be13-0863eb03d1da · outbound

This paper cites Chou, X.-S.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics Chou, X.-S

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:04:39.095644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-16T00:04:38.323862Z digest=sha256:778d36f16d149a8ed87830aff818eca1c6e53a16a27fc7902d5c50131361b9c7

Observation e10fc8cc-e407-424f-8124-b1a8c030f144 · outbound

This paper cites de Moura, S.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics de Moura, S

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:04:39.079183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-16T00:04:38.329100Z digest=sha256:55d75b3d24a1c42a4588a9518b89ec8e9dbc536cc6da0c4a0e98526a55f77c8c

Observation 34828bb5-7f77-4451-8c4e-a76f11f231ef · outbound

This paper cites Dong and T.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics Dong and T

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:04:39.063348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-16T00:04:38.333926Z digest=sha256:59163366162d17490e1e9150625cf21764f4e5b8e891bd0365c8a5e83ec0948a

Observation 1eb05b40-8015-44be-a945-3f934e1fea40 · outbound

This paper cites Gemini 2.5: Our most intelligent ai model, 2025.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics Gemini 2.5: Our most intelligent ai model, 2025

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:04:39.047931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-16T00:04:38.338423Z digest=sha256:2ec947bd4e342b49716e8c00515b857301452b4035815dd6c09dfea57dab429e

Observation 273ac6bc-9ae9-4cec-8e31-83d8d3a181ae · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T00:04:38.342897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:04:38.342897Z digest=sha256:2f354123072c92dfba45f2a007b1ac28cfdfb46d999dfcbc724dfba359eb5bfc

Observation d44e4573-b102-44b8-b335-6a6bab842091 · outbound

This paper cites OpenAI o1 System Card.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics OpenAI o1 System Card

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T00:04:38.347962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:04:38.347962Z digest=sha256:8fe83357360c2863e8bc52e545b0bd7e948586fbaa592d16ee0adb5d3c29af86

Observation 0be5b01d-6e81-4fb6-8e98-d384e2bc29f9 · outbound

This paper cites Draft, Sketch, and Prove: Guiding Formal Theorem Provers with Informal Proofs.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics Draft, Sketch, and Prove: Guiding Formal Theorem Provers with Informal Proofs

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T00:04:38.352880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:04:38.352880Z digest=sha256:488fa77302d0c271642b2c91eaa2577500cede6252d52b4d442f381d6ca6e29b

Observation 23b6e7a1-b02d-4512-9ce0-2d2d77826670 · outbound

This paper cites Goedel-Prover: A Frontier Model for Open-Source Automated Theorem Proving.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics Goedel-Prover: A Frontier Model for Open-Source Automated Theorem Proving

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T00:04:38.357808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:04:38.357808Z digest=sha256:b97905b5b3aff84e582b1a2967c0b34e274a2460150dc48cbb79eb6e37b2e730

Observation 33da42c2-99b2-4f58-92de-8c31a49504e6 · outbound

This paper cites an unresolved cited work.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-16T00:04:39.031934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-16T00:04:38.362606Z digest=sha256:78793f09a749761734526ceae570630f492e2a9f10d4b140a39a22fe36b9be27

Observation cf369646-4a8e-439b-a112-d5eb45a337e5 · outbound

This paper cites mathlib Community.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics mathlib Community

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T00:04:38.366963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:04:38.366963Z digest=sha256:debc4c9ccfb9abbda65ac7af4ceef2110fb019d6f1e2499fcbeeb3fadb90bc4e

Observation 796942e2-beee-4e06-938e-e8ab2921a59e · outbound

This paper cites an unresolved cited work.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-16T00:04:39.016230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-16T00:04:38.372185Z digest=sha256:820523c8f679a227c17df74827650cb4c159142ef11e54421610795649c87bdb

Observation fa95f311-9cac-435a-8adb-6dc0b9470d52 · outbound

This paper cites Openai o3-mini, 2025.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics Openai o3-mini, 2025

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T00:04:38.377268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:04:38.377268Z digest=sha256:87ae9611b180044b72d91bcd094c1df5a59b29cba2888ddb336db75de6845312

Observation 10f860fd-a62e-4db2-a761-3e114f5e85a7 · outbound

This paper cites Generative Language Modeling for Automated Theorem Proving.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics Generative Language Modeling for Automated Theorem Proving

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T00:04:38.382016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:04:38.382016Z digest=sha256:91e049da79fef47badd4cf87689bc321c7e10c592fcb4b733805b99590cbca36

Observation 7abc47ef-ac8a-49a3-bf9c-737cd25084ed · outbound

This paper cites A I M O P rize --- aimoprize.com.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics A I M O P rize --- aimoprize.com

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:04:38.990862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-16T00:04:38.386785Z digest=sha256:e1de953038cb4b675b059a68f7680b78f6a49488d39ab5b7e91df468bd4ecc49

Observation d49227e7-a579-4d05-b219-e13e64bdfdf7 · outbound

This paper cites an unresolved cited work.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-16T00:04:38.975932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-16T00:04:38.391434Z digest=sha256:ce69437de2b5c1fd7cecf44a59856bd27ad54718c58e90185077eb75e4e1290e

Observation 2ce0e71f-0f97-4312-831c-5c3ff5b79273 · outbound

This paper cites an unresolved cited work.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-16T00:04:38.960185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-16T00:04:38.395968Z digest=sha256:468250a1b852ae656f2e0757f08662af394bfbab5c543d7cfa06cb39ea4cf036

Observation a3367479-a069-42f7-b8d6-a67f3e4772fb · outbound

This paper cites The Coq Proof Assistant , Sept.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics The Coq Proof Assistant , Sept

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:04:38.942752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-16T00:04:38.400540Z digest=sha256:b391c932d0caa4ab075d1e379cf4fcaec624c9ac72fe41d0575b8b42fc0e9836

Observation 5ad56077-1b7b-4cfc-b005-e1e453a7c7ac · outbound

This paper cites PutnamBench: Evaluating Neural Theorem-Provers on the Putnam Mathematical Competition.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics PutnamBench: Evaluating Neural Theorem-Provers on the Putnam Mathematical Competition

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T00:04:38.404967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:04:38.404967Z digest=sha256:0b26f95fcb8a385d2eef84cc24ca445438515c1722d95ef35df94de970626c1d

Observation 7246e3c0-dcd5-436f-ae25-9b6a87e60d4a · outbound

This paper cites an unresolved cited work.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-16T00:04:38.927133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-16T00:04:38.409614Z digest=sha256:823f3ff8204e64ac37aecafabee14b7e4f28d3f87c1e39e7418ffc191b1cad61

Observation 4543389d-6284-4d0c-af4a-ae42adaf444a · outbound

This paper cites an unresolved cited work.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-16T00:04:38.911865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-16T00:04:38.414144Z digest=sha256:a513bc7d1012287c13212dc94253e48c84766e8f41bd4db4f636301365f2b150

Observation 6631b2b5-a931-4f4f-990e-ce329e9b6dc1 · outbound

This paper cites Kimina-Prover Preview: Towards Large Formal Reasoning Models with Reinforcement Learning.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics Kimina-Prover Preview: Towards Large Formal Reasoning Models with Reinforcement Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-16T00:04:38.418720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:04:38.418720Z digest=sha256:77ba1fede0d0a19d80915cca6df6a18816339a147877c21b4314df0b95f6eca6

Observation 8fc8521d-0820-4770-989d-7bc862e0d60b · outbound

This paper cites Wenzel, L.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics Wenzel, L

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:04:38.895485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-16T00:04:38.423749Z digest=sha256:28734b19d0f28b26f0b44d737f7997212b7fc5203a11cdaf15e47d73f2d226b9

Observation d823ffda-1738-442e-a47d-ede3af7289d8 · outbound

This paper cites an unresolved cited work.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-16T00:04:38.879361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-16T00:04:38.429150Z digest=sha256:1fb09388afb3d087f0c8503cdb74831c2ac8c5d072c0325d7a91d56c9821587c

Observation 284f4414-06e2-4d28-8895-f4b2f4677108 · outbound

This paper cites DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T00:04:38.434922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:04:38.434922Z digest=sha256:2bc90f2e183a13b556551d52cd992a8b8034c1f2b73239088bf064ceeacad631

Observation 9f1a46e6-a02e-4ff3-a25e-4f937c99b49c · outbound

This paper cites an unresolved cited work.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T00:04:38.439891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:04:38.439891Z digest=sha256:7631f77bbd2367550a3bb508ba2a70542ba068757ca90eb87ada0cd5fd470143

Observation 191eccc6-f994-406f-953c-c4214ba4e372 · outbound

This paper cites A Combinatorial Identities Benchmark for Theorem Proving via Automated Theorem Generation.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics A Combinatorial Identities Benchmark for Theorem Proving via Automated Theorem Generation

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-08-16T00:04:38.537787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-16T00:04:38.444500Z digest=sha256:4b09e7f0debba55268fcbb0e0f34a70584c8ca3d38baac2cf92d95fea9f9a9e7

Observation 1d08fa3b-94ae-490f-b486-935800308495 · outbound

This paper cites Leanabell-Prover: Posttraining Scaling in Formal Reasoning.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics Leanabell-Prover: Posttraining Scaling in Formal Reasoning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-16T00:04:38.449760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:04:38.449760Z digest=sha256:dcbac8eb53a41f4d08d68fa8c0c7ce6f18c9c24914b679298d99f1034fc7e09f

Observation 5a883dde-249b-4b4c-b293-4a07491a0194 · outbound

This paper cites MiniF2F: a cross-system benchmark for formal Olympiad-level mathematics.

CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics MiniF2F: a cross-system benchmark for formal Olympiad-level mathematics

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-16T00:04:38.456180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:04:38.456180Z digest=sha256:af8764b6cf477858af6b1e02cd721ec6056deefbef239730095df748ae0f76f7

Pith citing papers

Observation 12d6cf0d-b1a6-4584-8b73-eec6eaf0fd33 · inbound

Formally Solving Answer-Construction Problems in Lean cites this paper.

Formally Solving Answer-Construction Problems in Lean CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:34:08.568419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:34:08.568419Z digest=sha256:2c13e9a1231239fa61572acb0996ca61b29a0afb691d7e88b41980c30f24ad4e

Observation afa26eea-fd01-4f0f-b282-3c0315185717 · inbound

Using Reasoning Models to Generate Search Heuristics that Solve Open Instances of Combinatorial Design Problems cites this paper.

Using Reasoning Models to Generate Search Heuristics that Solve Open Instances of Combinatorial Design Problems CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:45.804899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:45.804899Z digest=sha256:c2275c9adf856907e416ff666d7bbbaf354bbfda522a0eb5cd3b6d7dabdf0304

Observation 7b6d46b7-1d1b-4ec0-a8ca-b228f0b511fa · inbound

Seed-Prover: Deep and Broad Reasoning for Automated Theorem Proving cites this paper.

Seed-Prover: Deep and Broad Reasoning for Automated Theorem Proving CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T10:30:44.632329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:30:44.632329Z digest=sha256:abbff68e2023130113497d14b604d5201e5acfc181ed0c44c18934631ecb8f19

Observation 411ff25b-e1c7-4732-bc91-1fa4d83bc43f · inbound

PuzzleClone: A DSL-Powered Framework for Synthesizing Verifiable Data cites this paper.

PuzzleClone: A DSL-Powered Framework for Synthesizing Verifiable Data CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T18:08:41.447037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T18:08:41.447037Z digest=sha256:80549b10331444ed24b1595a1e6d2486080a433084d232db7cb05dc645390b39

Observation ea9c2fc6-5462-481f-9c62-f7f119be5068 · inbound

CAM-Bench: A Benchmark for Computational and Applied Mathematics in Lean cites this paper.

CAM-Bench: A Benchmark for Computational and Applied Mathematics in Lean CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-20T13:38:19.375605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-20T13:35:04.729506Z digest=sha256:bf830616f01350029f1fd4f16b2348bac1b33040abec6cbf7570aee096459271

Observation 8ad20165-ffd4-4808-b141-ac450dcfb472 · inbound

GTBench: A Curriculum-Grounded Benchmark for Evaluating LLMs as Mathematical Research Assistants in Graph Theory cites this paper.

GTBench: A Curriculum-Grounded Benchmark for Evaluating LLMs as Mathematical Research Assistants in Graph Theory CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-02T03:06:30.307295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T10:19:03.670540Z digest=sha256:c72ca2bd65d523a7816f6cb8a9bd99eb0d59c85cc8e6a9128a31ccdda391bf48

Observation 8adee818-56e3-4c85-b038-e6a0001730c9 · inbound

TheoremBench: Evaluating LLMs on Theorem Proving in Formal Mathematics cites this paper.

TheoremBench: Evaluating LLMs on Theorem Proving in Formal Mathematics CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:47:31.569259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-27T16:19:11.123994Z digest=sha256:0b82f68c8750fca401317520c06ff1517dcc694191f02dc12a31c4540b91f9ee

Observation 359d3628-9664-4e30-96f0-1ea0999c09ff · inbound

Faults in Our Formal Benchmarking: Dataset Defects and Evaluation Failures in Lean Theorem Proving cites this paper.

Faults in Our Formal Benchmarking: Dataset Defects and Evaluation Failures in Lean Theorem Proving CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:04:21.947161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T06:56:32.253843Z digest=sha256:8859adf2f46ca3cd4daedc06e399b7e39ce6d8953dbf6ce77ac9d292f1cebd09

Observation 37f7c2a0-21c8-4a60-967c-08b179a944c0 · inbound

FormalRx: Rectify and eXamine Semantic Failures in Autoformalization cites this paper.

FormalRx: Rectify and eXamine Semantic Failures in Autoformalization CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics

Reference 111

Resolution
unresolved
no resolver link, observed 2026-07-11T15:42:50.296348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T15:42:50.296348Z digest=sha256:bd3861466aeff2c5229fc68e5193f99426e02721ff4a92d0a0c9d1eac92c8b76

Observation 79717f3b-2b55-47a3-8743-0a1cc9b819e3 · inbound

From Solvers to Research: Large Language Model-Driven Formal Mathematics at the Research Frontier cites this paper.

From Solvers to Research: Large Language Model-Driven Formal Mathematics at the Research Frontier CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics

Reference 150

Resolution
verified exact
local_arxiv, observed 2026-07-10T18:17:33.662864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-10T18:16:31.176239Z digest=sha256:6e5827f2c6552642b0738563b20a0dd9d3c1bba846a379012722e80973afd6d9

Observation aeafd0f9-a93b-410d-a445-b0bb8f5d0e21 · inbound

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability cites this paper.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T15:30:26.803162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:30:26.803162Z digest=sha256:a8bcc813344b50033ad41177032b546fc237395e160c95e75bec2b34e5b44431

Observation 45e50f55-3e41-4643-a318-260d8c62b23b · inbound

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability cites this paper.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:57.397021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:57.397021Z digest=sha256:547b99f832f6bc97c4cd75513cec3890b6005160c64b80496a2236f95e02bbdb