Pith. sign in

Paper Citation Record · LEDGER

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability

As of 11 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 0 inbound Pith citation observations for arXiv:2608.09538.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.09538 v1

Coverage vector

measured 46 of 46 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T15:30:26.930654Z

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

46 of 46 outbound references displayed

  • verified exact0
  • verified fuzzy24
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0fc0a445-f94f-4dbc-a555-e4b0e0a543ff · outbound

This paper cites The jacobian conjecture is false.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability The jacobian conjecture is false

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:30:28.020343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T15:30:26.708824Z digest=sha256:9facef50b158f9f283970bc1623091c5cea6133861601109fa619cbfb5f14fd5

Observation 9039ad13-3878-4a6d-92f4-2527c4f8805a · outbound

This paper cites GPT-4 Technical Report.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability GPT-4 Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T15:30:26.714450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:30:26.714450Z digest=sha256:6b41d60cd9c7da73f430158493f7ffbfb3575e9a17c73dbe9c807ab379cd2ca1

Observation 5d341a6e-484d-414f-acb8-7642602e1fe8 · outbound

This paper cites GEPA: Reflective Prompt Evolution Can Outperform Reinforcement Learning.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability GEPA: Reflective Prompt Evolution Can Outperform Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T15:30:26.719646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:30:26.719646Z digest=sha256:c74efef7bf049a4b0c2e7c98fc30a27b79c16514be376d3c39301926c4e26b01

Observation a06bd32b-7d27-4110-82b4-ab5c7458def5 · outbound

This paper cites AI achieves silver-medal standard solving international 178 mathematical olympiad problems.DeepMind blog, 179:45, 2024.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability AI achieves silver-medal standard solving international 178 mathematical olympiad problems.DeepMind blog, 179:45, 2024

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:30:28.005079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T15:30:26.725050Z digest=sha256:fdda51c00a7a8767a581291fbfa0e35cbaac6507779b769564eda4f0a8d4f7a7

Observation 67868338-95e4-4db8-9d18-2292a20b189f · outbound

This paper cites MathArena: Evaluating LLMs on Uncontaminated Math Competitions.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability MathArena: Evaluating LLMs on Uncontaminated Math Competitions

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T15:30:26.730065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:30:26.730065Z digest=sha256:9c02508f8b02d780c5e4dba4b7e9e079b804568eaeac811416a37244b3de2372

Observation 41ab451c-4651-4985-a058-7436dd6cf229 · outbound

This paper cites A Case Study on the Effectiveness of LLMs in Verification with Proof Assistants.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability A Case Study on the Effectiveness of LLMs in Verification with Proof Assistants

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T15:30:26.735092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:30:26.735092Z digest=sha256:69bc18ee721ac0fceb2171dcdb4f11fa21b2d52823ebc9f0846c240e8d8bc987

Observation 475b4496-aff0-44a2-9d7c-dd8fa67ec060 · outbound

This paper cites Autonomous chemical research with large language models.Nature, 624(7992):570–578, 2023.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability Autonomous chemical research with large language models.Nature, 624(7992):570–578, 2023

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T15:30:26.740935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:30:26.740935Z digest=sha256:481c8dc23805f9a4e1a3ec5cbeb26858ab3519de88371c3ecd5d0618e8daa297

Observation 24395958-9d8d-448b-9a96-c799c773ce71 · outbound

This paper cites Gold-medalist performance in solving olympiad geometry with alphageometry2.arXiv preprint arXiv:2502.03544, 2025.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability Gold-medalist performance in solving olympiad geometry with alphageometry2.arXiv preprint arXiv:2502.03544, 2025

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T15:30:26.745309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:30:26.745309Z digest=sha256:0d8fc7d04e2811edb15f4f7465cc9fbdce7bd563220b9b063b6519c96eacdfe9

Observation 3a71bb3d-ea2f-4903-aea7-c9212b8cd0ba · outbound

This paper cites Math- Construct: Challenging LLM reasoning with constructive proofs.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability Math- Construct: Challenging LLM reasoning with constructive proofs

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:30:27.978501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T15:30:26.749863Z digest=sha256:b6bb951a35675e7b349fe10b02c9d78f3e5d411a9098b210ffa4902e7b58cb98

Observation 934a182b-5df4-4381-980b-5e96e1a401cd · outbound

This paper cites The Open Proof Corpus: A Large-Scale Study of LLM-Generated Mathematical Proofs.arXiv preprint arXiv:2506.21621, 2025.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability The Open Proof Corpus: A Large-Scale Study of LLM-Generated Mathematical Proofs.arXiv preprint arXiv:2506.21621, 2025

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T15:30:26.754453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:30:26.754453Z digest=sha256:2d8b603a0bd70635661ee54778e115309cae21dd793285216179fadfdd1ba18d

Observation 3cdf85aa-b9ca-4ee2-a77b-2dbe9b6a69b7 · outbound

This paper cites Trinh, Garrett Bingham, Dawsen Hwang, Yuri Chervonyi, Junehyuk Jung, Joonkyung Lee, Carlo Pagano, Sang hyun Kim, Federico Pasqualotto, Sergei Gukov, Jonathan N.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability Trinh, Garrett Bingham, Dawsen Hwang, Yuri Chervonyi, Junehyuk Jung, Joonkyung Lee, Carlo Pagano, Sang hyun Kim, Federico Pasqualotto, Sergei Gukov, Jonathan N

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:30:27.963057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T15:30:26.758979Z digest=sha256:627f1b44c089ddd718ed07edffbda9f4cabdba991778611fc953975c8611819d

Observation 17d5f112-0e4e-4ed9-bd2f-b1b37c4ccf4a · outbound

This paper cites Towards an AI co-scientist.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability Towards an AI co-scientist

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T15:30:26.763898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:30:26.763898Z digest=sha256:4b76dab4100a5bf3ca09eca343e7caa47196d6a3c5a4e4b3587d829ba387896b

Observation 6f546cf6-afa1-4a4b-a5d4-218cda220ca7 · outbound

This paper cites CRISPR-GPT for Agentic Automation of Gene-editing Experiments.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability CRISPR-GPT for Agentic Automation of Gene-editing Experiments

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T15:30:26.768790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:30:26.768790Z digest=sha256:8db5bdb9325cfc68896ded690b6cc03465e1453ea9ca2e098aac22fcaf10b082

Observation 1bbaf41a-81b2-4af8-b290-5e9f26015078 · outbound

This paper cites Gemini 2.5 pro capable of winning gold at imo 2025.arXiv preprint arXiv:2507.15855, 2025.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability Gemini 2.5 pro capable of winning gold at imo 2025.arXiv preprint arXiv:2507.15855, 2025

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T15:30:26.773741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:30:26.773741Z digest=sha256:c0322bd1e671fc1657a2b83d35b50436b3d169e0e7816d0f248551719bb42ea5

Observation ff2865c9-0604-4eb7-9e88-fcb861007194 · outbound

This paper cites Chemformer: a pre- trained transformer for computational chemistry.Machine Learning: Science and Technology, 3(1):015022, 2022.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability Chemformer: a pre- trained transformer for computational chemistry.Machine Learning: Science and Technology, 3(1):015022, 2022

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:30:27.945930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T15:30:26.778488Z digest=sha256:0167ee35445f0157ad3e51877bdc4254531273795c2de450bd14921920dc9bb0

Observation 5bcfe8a1-a51e-4bf8-a54e-adec71a14f11 · outbound

This paper cites Gflownets for ai-driven scientific discovery.Digital Discovery, 2(3):557–577, 2023.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability Gflownets for ai-driven scientific discovery.Digital Discovery, 2(3):557–577, 2023

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:30:27.929328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T15:30:26.783006Z digest=sha256:c7eacc7c06cd9c96b8c0dbc24693e3ff5043e11ae717b0150e432298f88c9bf9

Observation 330a8451-9d93-4cf6-87a7-568d1abc2a2c · outbound

This paper cites Perfor- mance of chatgpt on usmle: potential for ai-assisted medical education using large language models.PLoS digital health, 2(2):e0000198, 2023.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability Perfor- mance of chatgpt on usmle: potential for ai-assisted medical education using large language models.PLoS digital health, 2(2):e0000198, 2023

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:30:27.912477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T15:30:26.787682Z digest=sha256:32fa680ae042a35d04892b1e734861a9a2c183ab9e132925b903695be22dd445

Observation 69292c8e-420b-4f96-88d5-568bcbe1026d · outbound

This paper cites Benchmarking automated theorem proving with large language models.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability Benchmarking automated theorem proving with large language models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:30:27.893830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T15:30:26.792588Z digest=sha256:88f7d67fb3cffb0ba47a58571205d0596c97cce6e66d3690ee1fa7e5abe89214

Observation 254e3d18-e4a5-4dab-a6c7-925e9320b7dc · outbound

This paper cites Proving Olympiad Inequalities by Synergizing LLMs and Symbolic Reasoning.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability Proving Olympiad Inequalities by Synergizing LLMs and Symbolic Reasoning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T15:30:26.797250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:30:26.797250Z digest=sha256:e224fbdc46adfa4e302f1036f34fff3016240b3101ef769b4440f7cbd31abc61

Observation aeafd0f9-a93b-410d-a445-b0bb8f5d0e21 · outbound

This paper cites CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T15:30:26.803162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:30:26.803162Z digest=sha256:7e7cdb77b4fe7dcb38037316d85fd0703df17e88275a6b771467a0c1c830ec61

Observation f8c269fe-e868-4948-bbb6-f42d051ab0e9 · outbound

This paper cites CHAMP: A Competition-level Dataset for Fine-Grained Analyses of LLMs' Mathematical Reasoning Capabilities.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability CHAMP: A Competition-level Dataset for Fine-Grained Analyses of LLMs' Mathematical Reasoning Capabilities

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T15:30:26.808039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:30:26.808039Z digest=sha256:d1adce1fe8be4bc3d17ee287d2ebb0f5c749be374cb1bf92ac54e19eb63c2c45

Observation fa73d2a6-c71e-43d0-9188-6be163742ab1 · outbound

This paper cites An openai model has disproved a central conjecture in discrete geometry.https: //openai.com/index/model-disproves-discrete-geometry-conjecture/ , 2026.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability An openai model has disproved a central conjecture in discrete geometry.https: //openai.com/index/model-disproves-discrete-geometry-conjecture/ , 2026

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:30:27.877263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T15:30:26.812993Z digest=sha256:9cac606cfb8a8ad8dbe9f19c861c88384ff99d208ffbe2431e84388648597d40

Observation 4b6cdbc3-d941-4ced-bab0-44a48facd5b0 · outbound

This paper cites Ten advances in mathematics and theoretical computer science.https://openai.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability Ten advances in mathematics and theoretical computer science.https://openai

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:30:27.861773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T15:30:26.817541Z digest=sha256:d5b244d2bd6cf61448cde06107dd64abf1d2b26069a1c6a996cdd78582f3553f

Observation a40e8709-6552-42dc-859d-f69ffdea258d · outbound

This paper cites Astroclip: a cross-modal foundation model for galaxies.Monthly Notices of the Royal Astronomical Society, 531(4):4990–5011, 2024.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability Astroclip: a cross-modal foundation model for galaxies.Monthly Notices of the Royal Astronomical Society, 531(4):4990–5011, 2024

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:30:27.845550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T15:30:26.821937Z digest=sha256:fd78d4d95898e5e08d662320ac80788c653f1cd2338e06d16c87ee957d3b8452

Observation 4d3eb6d6-bbe2-435d-bf9f-ef71c779ef51 · outbound

This paper cites BrokenMath: A Benchmark for Sycophancy in Theorem Proving with LLMs.arXiv preprint arXiv:2510.04721, 2025.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability BrokenMath: A Benchmark for Sycophancy in Theorem Proving with LLMs.arXiv preprint arXiv:2510.04721, 2025

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T15:30:26.826708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:30:26.826708Z digest=sha256:e2d96419cdd50ba40c475649d402398cb546e889634b192094d78004de89046e

Observation 09bfc06c-4615-4380-9c3b-14c7f4713fad · outbound

This paper cites LemmaBench: A Live, Research-Level Benchmark to Evaluate LLM Capabilities in Mathematics.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability LemmaBench: A Live, Research-Level Benchmark to Evaluate LLM Capabilities in Mathematics

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T15:30:26.831491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:30:26.831491Z digest=sha256:a8554f9df6b109268114b00d99365c7f8ecf21d600e3b2536dc7e51151459ec8

Observation 512e83f8-04ac-40f1-8adb-313778994f3d · outbound

This paper cites AI and the Everything in the Whole Wide World Benchmark.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability AI and the Everything in the Whole Wide World Benchmark

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T15:30:26.837015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:30:26.837015Z digest=sha256:4d2a8766578d21c7552992910b93fd857e7ca96f82e26cdd24b6db01d68e5d1a

Observation cd9312c2-8f68-4848-8bde-a72a4a09d5a5 · outbound

This paper cites Paperbench: Evaluating ai’s ability to replicate ai research.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability Paperbench: Evaluating ai’s ability to replicate ai research

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:30:27.829344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T15:30:26.842212Z digest=sha256:20baa8d29ca012e954d7899b7c70165f2d9a15ac8aa5eeb1d66dedd3a559e615

Observation 8b26023c-7e99-4541-86cc-37bf2d406949 · outbound

This paper cites Solving olympiad geometry without human demonstrations.Nature, 625(7995):476–482, 2024.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability Solving olympiad geometry without human demonstrations.Nature, 625(7995):476–482, 2024

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T15:30:26.847015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:30:26.847015Z digest=sha256:6d3dbfe2b12f3c082d1ebe38fb409f04ea44579f5d098a842d46b063cce7e2ac

Observation 7915923e-ab3b-49b5-a1ae-713da2253762 · outbound

This paper cites Gpt-4 can ace the bar, but it only has a decent chance of passing the cfa exams.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability Gpt-4 can ace the bar, but it only has a decent chance of passing the cfa exams

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:30:27.801204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T15:30:26.851799Z digest=sha256:23f32d912c9a9b10c71875eb8b3c0da17aa066036bed8902c93c4f7efda11df8

Observation 587175f0-91ce-4e20-a63c-00eee6202d6e · outbound

This paper cites The transformative potential of machine learning for experiments in fluid mechanics.Nature Reviews Physics, 5(9):536–545, 2023.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability The transformative potential of machine learning for experiments in fluid mechanics.Nature Reviews Physics, 5(9):536–545, 2023

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T15:30:26.856581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:30:26.856581Z digest=sha256:f094542c25a29f8b8365591939d30219cd2bfae8a1d2457bfef2440c5f0dd48a

Observation e5dbd1ee-950c-4477-8263-accc4f66a8ae · outbound

This paper cites Horizon- math: Measuring ai progress toward mathematical discovery with automatic verification.arXiv preprint arXiv:2603.15617, 2026.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability Horizon- math: Measuring ai progress toward mathematical discovery with automatic verification.arXiv preprint arXiv:2603.15617, 2026

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T15:30:26.861318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:30:26.861318Z digest=sha256:6d31a5e43f1a9a65f7cb1ab36d12ba6a5052ab977c75476efd6526200d550034

Observation 8b7281f1-01ec-499b-b326-bdfb9608eea5 · outbound

This paper cites Scientific discovery in the age of artificial intelligence.Nature, 620(7972):47–60, 2023.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability Scientific discovery in the age of artificial intelligence.Nature, 620(7972):47–60, 2023

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T15:30:26.866125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:30:26.866125Z digest=sha256:f3d780fc1cb5156d280a6ef0b8b20199355070aeb8d1170aad98b2a5f7dbf217

Observation 0a1c7bac-c3df-4270-9fe7-134f06e94718 · outbound

This paper cites Woodruff, Vincent Cohen-Addad, Lalit Jain, Jieming Mao, Song Zuo, Mohammad- Hossein Bateni, Simina Branzei, Michael P.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability Woodruff, Vincent Cohen-Addad, Lalit Jain, Jieming Mao, Song Zuo, Mohammad- Hossein Bateni, Simina Branzei, Michael P

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:30:27.764225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T15:30:26.870713Z digest=sha256:205995630a6f8f200e4fcaf22327b2c58be95f09637797c897519b0a2c840491

Observation f3edc235-b7ab-4673-a635-67a0a373aaca · outbound

This paper cites Exploring the role of large language models in the scientific method: from hypothesis to discovery.npj Artificial Intelligence, 1(1):14, 2025.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability Exploring the role of large language models in the scientific method: from hypothesis to discovery.npj Artificial Intelligence, 1(1):14, 2025

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:30:27.748566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T15:30:26.875225Z digest=sha256:491fd15e74add04506c8290386285308e3aec4a555d0f71b7f23e8a53ee8cdf2

Observation 96ae0ae8-d254-4055-9d66-5f0eb12fef80 · outbound

This paper cites AGIEval: A Human-Centric Benchmark for Evaluating Foundation Models.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability AGIEval: A Human-Centric Benchmark for Evaluating Foundation Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T15:30:26.879822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:30:26.879822Z digest=sha256:0a3257c3c013f7342de9d6519f28cc6049c4e446abd1a50667fa9376ab3750af

Observation 80584217-502a-43a6-a44e-aa3f94a1d21c · outbound

This paper cites safety margins.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability safety margins

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:30:27.731904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T15:30:26.884771Z digest=sha256:3008a0e620743a9c260eccdf13463c718d736501ee21e807c3f4bf0d60fa7760

Observation 0f7079c1-9be5-4a2f-a2ca-0538ecdee000 · outbound

This paper cites outer queries.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability outer queries

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:30:27.715477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T15:30:26.890015Z digest=sha256:e667ac7128933a6e2a060f3285692937249e402453fd7f1ed125327b27e9ebf6

Observation cdb8c14e-fc96-4064-a847-7d46b8b143e1 · outbound

This paper cites mapping limits matching identical parameters constraints mapping.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability mapping limits matching identical parameters constraints mapping

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:30:27.699735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T15:30:26.895088Z digest=sha256:473602ee4f1f89ced79deacd0b203263f7f18d277b2964afdfffcc3f65ebf7fe

Observation dd148b53-735f-464f-ae37-ec3f537ab054 · outbound

This paper cites - **Boundary and Floor/Ceiling Precision**: Bounds must be exactly supported.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability - **Boundary and Floor/Ceiling Precision**: Bounds must be exactly supported

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:30:27.682747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T15:30:26.900120Z digest=sha256:d3e8151b836cebfae4cd0f2657277b1bf3856e4d0412855a09c678ee948d585d

Observation de6ba752-a6d9-41aa-aa79-438990095047 · outbound

This paper cites an unresolved cited work.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:30:27.664987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T15:30:26.905361Z digest=sha256:43b8ccc300d8c38671bb1243584b71141f4e8f027c1a65a356b284513a7632e6

Observation e0c3ff3b-44e8-4945-b267-cd23d81b5fb9 · outbound

This paper cites This is for your own benefit, to maximize the chances that your proof is correct so 17 that you can pass the benchmark.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability This is for your own benefit, to maximize the chances that your proof is correct so 17 that you can pass the benchmark

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:30:27.647498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T15:30:26.910954Z digest=sha256:c0a0e43f891b028ca433aa588522d7109d8e2e600acde2648f6bf446f4f4e302

Observation bb131e81-4692-4f9e-b940-4b92418406eb · outbound

This paper cites 0" if the student introduces arbitrary numerical constraints or.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability 0" if the student introduces arbitrary numerical constraints or

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:30:27.632370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T15:30:26.916052Z digest=sha256:be6711e71c36e94af242458b1ea7c6602181ce09d119b41c154786a4d9bd6312

Observation d762a492-8d91-41ac-99d7-4a2765ae51f6 · outbound

This paper cites 0" if the student generalizes a property of a subset to a larger set (e.g., applying a u-uniformity property defined for.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability 0" if the student generalizes a property of a subset to a larger set (e.g., applying a u-uniformity property defined for

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:30:27.617762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T15:30:26.921125Z digest=sha256:849ba996de75dcc81f66727535739f58e9442686598cecaeee38507b187827d8

Observation 205b79c5-50e8-477d-a206-34b0a1108eee · outbound

This paper cites mapping limits matching identical parameters constraints mapping.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability mapping limits matching identical parameters constraints mapping

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:30:27.601914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T15:30:26.925904Z digest=sha256:bfa535400ec7de4f85a658e0c9a02bfad1c69d0475a8ba7c32f862ff76fd7f14

Observation 5bda6e9a-c82d-4eb2-9cfe-9651719923e2 · outbound

This paper cites 0". - Do not provide any explanation, feedback, or additional text. Your response must be only.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability 0". - Do not provide any explanation, feedback, or additional text. Your response must be only

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:30:27.585450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T15:30:26.930654Z digest=sha256:0b76fe9f7f297996742b5d7ced7ef58013eabe0fffa20718d090caf106474d66

Pith citing papers

No inbound Pith citation observations are available.