Pith. sign in

Paper Citation Record · LEDGER

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization

As of 7 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 0 inbound Pith citation observations for arXiv:2608.06301.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.06301 v1

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T06:02:25.819108Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

40 of 40 outbound references displayed

  • verified exact1
  • verified fuzzy4
  • unresolved34
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7c5f7cd5-2b14-40f1-a771-e60121272912 · outbound

This paper cites an unresolved cited work.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T06:02:26.840406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:02:25.616441Z digest=sha256:2015b7978f07ea7e1c1210e70bc34c6dee00062115fec7239c1786fe09f0bef6

Observation 84c6ef51-45d5-495f-9a85-66d6b116523e · outbound

This paper cites Program Synthesis with Large Language Models.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Program Synthesis with Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.621540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.621540Z digest=sha256:1b8fd176b96bd2637bd5164aace02ba68cfd265aafc0b3361bde2a433dbfc536

Observation 6b862b9b-a108-44ff-b7c2-bfc126a7f381 · outbound

This paper cites MLE-bench: Evaluating Machine Learning Agents on Machine Learning Engineering.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization MLE-bench: Evaluating Machine Learning Agents on Machine Learning Engineering

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.626627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.626627Z digest=sha256:f876872ad71e8d5fdea4276bdd6846760f6b1a92138f4742c75fec0df9bf5276

Observation 5fecd972-8608-4139-bedb-a94f5cec91d0 · outbound

This paper cites HarnessX: A Composable, Adaptive, and Evolvable Agent Harness Foundry.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization HarnessX: A Composable, Adaptive, and Evolvable Agent Harness Foundry

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-08-07T06:02:26.542719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:02:25.639914Z digest=sha256:2830efe7aa7fb8fe27498fd0e00cdd0ddb8d41af7f7bbef3d0f3169f1f36dfcf

Observation 06ae2d74-049d-4f54-a113-92f478ceaf97 · outbound

This paper cites an unresolved cited work.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T06:02:26.823414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:02:25.644768Z digest=sha256:3da0b8282a63f8030b8cb7420c29f77630e06136a3eaa34040e9d89150c72165

Observation 7e24c6bc-e978-4394-a654-b4a3c2cfc91a · outbound

This paper cites Trace is the Next AutoDiff: Generative Optimization with Rich Feedback, Execution Traces, and LLMs.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Trace is the Next AutoDiff: Generative Optimization with Rich Feedback, Execution Traces, and LLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.649833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.649833Z digest=sha256:c865606fcd15b377bebe54800116c98dd90885f5ed2676624992b375bd6f80c7

Observation 1ad6944b-e18e-41d3-b5e1-095ed755382e · outbound

This paper cites an unresolved cited work.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.655395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.655395Z digest=sha256:fe20cc5e39a47148ce1184bb6746afe33173c2dc8f018204b46f69c20171a4a8

Observation 4e70b33f-ef19-4b64-b8e2-73e9f5469edd · outbound

This paper cites an unresolved cited work.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T06:02:26.792632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:02:25.659935Z digest=sha256:118a2186ce71634e9e13663377df34eab06a4d7ac11639d8b9f8f1368511dde3

Observation 95aa86d1-88d4-415c-97c5-7a24d9379dfd · outbound

This paper cites DSPy: Compiling Declarative Language Model Calls into Self-Improving Pipelines.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization DSPy: Compiling Declarative Language Model Calls into Self-Improving Pipelines

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.668797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.668797Z digest=sha256:ec3c8c861537513eaaf6132528f87983971c70e23c2babc91849fad742c35374

Observation 72497daf-5131-42b4-be49-78a98649a566 · outbound

This paper cites ShinkaEvolve: Towards Open-Ended And Sample-Efficient Program Evolution.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization ShinkaEvolve: Towards Open-Ended And Sample-Efficient Program Evolution

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.673512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.673512Z digest=sha256:193ace9a8093480cb09cc72fdd681354d290d1725d031e2a7b0ecde124e4f46f

Observation b114c3e1-3999-451f-90ca-dd70996b620e · outbound

This paper cites Meta-Harness: End-to-End Optimization of Model Harnesses.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Meta-Harness: End-to-End Optimization of Model Harnesses

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.678912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.678912Z digest=sha256:13617c62eb12d63540903a4f7023b0385b81ee37d31f1e22f14884637f67455d

Observation 45050dee-36a5-4044-bb03-177129addcc5 · outbound

This paper cites Agentic Harness Engineering: Observability-Driven Automatic Evolution of Coding-Agent Harnesses.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Agentic Harness Engineering: Observability-Driven Automatic Evolution of Coding-Agent Harnesses

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.684162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.684162Z digest=sha256:2cc7fa3888221f37eba200b70766191c352e63b834c81d7dcb10ccf2b230ab6a

Observation 59ab5bed-bff7-4a79-bee9-45284fe28a2e · outbound

This paper cites RSIBench-Data: Benchmarking Data-Centric Research for Recursive Self-Improvement.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization RSIBench-Data: Benchmarking Data-Centric Research for Recursive Self-Improvement

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.689165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.689165Z digest=sha256:90252c8a5e8af3baca63800add0fe32df80ab0439061f16925d39e032de059d9

Observation 7a5ea708-3a4e-48e5-a781-7d8f087d9d65 · outbound

This paper cites Terminal-Bench: Benchmarking Agents on Hard, Realistic Tasks in Command Line Interfaces.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Terminal-Bench: Benchmarking Agents on Hard, Realistic Tasks in Command Line Interfaces

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.694406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.694406Z digest=sha256:5b986e787506b72c7e5a94186aed20d77cb17017d6a1143a4f2c5c6cdf23a38e

Observation 832e847c-45be-4983-ae0d-e267fdca825c · outbound

This paper cites GAIA: a benchmark for General AI Assistants.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization GAIA: a benchmark for General AI Assistants

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.699110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.699110Z digest=sha256:dc646f3fa5ed10d87dc57c433ae0e8102dcb2e4659f0fe5591f825ffc5d82368

Observation ac423a22-3cfe-41b3-918c-a10db151494c · outbound

This paper cites AlphaEvolve: A coding agent for scientific and algorithmic discovery.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization AlphaEvolve: A coding agent for scientific and algorithmic discovery

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.704618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.704618Z digest=sha256:51067372a414858805dc0f851368ceddcb67071e0f1040259170411aec36f431

Observation 6d6365fd-311e-4770-9698-675129b7400d · outbound

This paper cites Opsahl-Ong, A.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Opsahl-Ong, A

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.710081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.710081Z digest=sha256:c9d6251fa5e66cbadba2e8ab59c037ac4d1a7c6d25192627492f53a8f4052306

Observation ad972446-d52d-4349-a01d-73c312377a8b · outbound

This paper cites Ouyang, S.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Ouyang, S

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:02:26.773465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:02:25.714285Z digest=sha256:bd2bc15f512e6c19714aed0dc992ee5e0688b171e916b11cddccd58830f837ad

Observation 6a3d847d-bc83-419c-ba9a-46bfdafc02e9 · outbound

This paper cites an unresolved cited work.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.718442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.718442Z digest=sha256:69717e90e61879e973bbcdbf3781c7cec4bd8b136aa1cb152fb779da87f30bcd

Observation f8c1677d-df60-4c64-b1d0-0381581b932d · outbound

This paper cites Romera-Paredes, M.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Romera-Paredes, M

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.731989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.731989Z digest=sha256:f328ab05836e940731a862bbfa8d74af2ad81eb11999527482466547e177c404

Observation 72fcfa4e-52d1-4680-a932-1113a7b1aafa · outbound

This paper cites Ursekar, A.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Ursekar, A

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:02:26.755690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:02:25.736276Z digest=sha256:ca76405f78d5b223a58ae4672dc1377debb5927d0d76e7cc56f40fe22633010b

Observation 8c554198-4a8f-4996-add4-14b687b00309 · outbound

This paper cites an unresolved cited work.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T06:02:26.737554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:02:25.740882Z digest=sha256:78d16621d1f4902ab589f60d7b09e7e29f24991dbbb2462b4f3db6dcb542fb27

Observation 6b69b85f-28e1-4de3-95f4-1ebd8cd4499e · outbound

This paper cites an unresolved cited work.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.745383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.745383Z digest=sha256:5b71a646172188753ebd9b0d15ee96a65b2704151ce29e893cf9405464144934

Observation 5957b07f-23e7-492c-bfc7-8ae6bb7b80b7 · outbound

This paper cites an unresolved cited work.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T06:02:26.716924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:02:25.749972Z digest=sha256:e63e8fd565b05ff7305a49fc552f7a11169f90d2da6bcadf47bf76b430235a5a

Observation 5cadbf08-4f74-4106-b420-6f947be9a055 · outbound

This paper cites an unresolved cited work.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T06:02:26.692576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:02:25.754561Z digest=sha256:5375d11e7fa3d11957a5d4cacf45c53f062605377a1df907d2bccfbfcafe8560

Observation c3ee986a-611a-463a-941e-2e856cb77cad · outbound

This paper cites RE-Bench: Evaluating frontier AI R&D capabilities of language model agents against human experts.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization RE-Bench: Evaluating frontier AI R&D capabilities of language model agents against human experts

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.759327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.759327Z digest=sha256:5ccf54b70d29907b50d6fe8094b230763e6c287aea2ca341d005f71c6d42ad8a

Observation db2fe2c1-8763-4894-8e1e-9ca36870d67f · outbound

This paper cites an unresolved cited work.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T06:02:26.671001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:02:25.763834Z digest=sha256:bb3f538b0c7fe7d79ba03751fbd01d3d87740c5e0d3d9cc8dca0e663314360d7

Observation f0aa53e8-970e-4a68-b894-3261e7658649 · outbound

This paper cites Harness-Bench: Measuring Harness Effects across Models in Realistic Agent Workflows.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Harness-Bench: Measuring Harness Effects across Models in Realistic Agent Workflows

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.768245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.768245Z digest=sha256:aae7507d58272227fd3b420cdc92fb6ab4a89ed389fddf026043006418b0f8e0

Observation 07ccd406-2445-4169-8b89-6f761f2e1fb8 · outbound

This paper cites an unresolved cited work.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.772900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.772900Z digest=sha256:292df4448bc2ea3d350eed2188693f1865425c4d94ae9de6c6e3b8275d1d0032

Observation 5853d2ef-5886-4c79-908a-c3d9a07ba557 · outbound

This paper cites G\"odel Agent: A Self-Referential Agent Framework for Recursive Self-Improvement.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization G\"odel Agent: A Self-Referential Agent Framework for Recursive Self-Improvement

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.787406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.787406Z digest=sha256:c7ccc9ed11a3c51dff870fcbb5604c8de9fba75c694c2f347b1e66064e92784d

Observation fd38340b-54e5-4194-bb30-b1108b5bc95d · outbound

This paper cites TextGrad: Automatic "Differentiation" via Text.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization TextGrad: Automatic "Differentiation" via Text

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.793445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.793445Z digest=sha256:8b38acd24b56d62785094d4fc417d79a3387b368901a49f4cd183dc029166d9a

Observation 8b7b2ce4-092c-41fe-8d39-6e49194a1169 · outbound

This paper cites Zelikman, E.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Zelikman, E

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:02:26.647834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:02:25.799359Z digest=sha256:d73c21d002e39b4d06f4ae139d38de4f742421842f32b4d14bdc598e463ee352

Observation fc3bd319-50f2-4c22-b079-517dfd5d96b7 · outbound

This paper cites LLM-AutoDiff: Auto-Differentiate Any LLM Workflow.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization LLM-AutoDiff: Auto-Differentiate Any LLM Workflow

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.783053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.783053Z digest=sha256:e1519cfec19a4b712763f57664d59f38f4e883a11acd6c87abcf1c67a2be722d

Observation e86ce37a-0105-4da9-9a23-55caca814dba · outbound

This paper cites Zhang, J.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Zhang, J

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:02:26.629708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:02:25.809069Z digest=sha256:d4eeda7c9812765a8acbafc1a44a89c7755b6fd095bbe0539e371e99dfb61e84

Observation 0382457f-3707-41e7-b59a-3175ac6c7b7c · outbound

This paper cites Stop Comparing LLM Agents Without Disclosing the Harness.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Stop Comparing LLM Agents Without Disclosing the Harness

Reference 38

Resolution
malformed identifier
no resolver link, observed 2026-08-07T06:02:25.819108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.819108Z digest=sha256:0a246a84272a8b57a560888e5c1d2079c108a19fdd53716a292a26aabb29d5e3

Observation 73f9813b-71b4-42ab-adfc-fa76a52057a8 · outbound

This paper cites Darwin Godel Machine: Open-Ended Evolution of Self-Improving Agents.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Darwin Godel Machine: Open-Ended Evolution of Self-Improving Agents

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.804071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.804071Z digest=sha256:47170204f6a341cbe553f39659c75448ed7772ef17fb07ecc68dad24092854f0

Observation 9aa97db2-102f-483b-9dd4-c900c79cfe0e · outbound

This paper cites AFlow: Automating Agentic Workflow Generation.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization AFlow: Automating Agentic Workflow Generation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.813777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.813777Z digest=sha256:cf0696f0f4cdbfd1394566e2d6dd1864d87151115fe55fc4efa240255b3db5b9

Observation f1e92752-39a5-4a6e-98a4-f093a899e2ea · outbound

This paper cites Evaluating Large Language Models Trained on Code.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Evaluating Large Language Models Trained on Code

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.635628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.635628Z digest=sha256:139633fe07e12072947881334b53f97d43a1cc9489972788ac42e2390632aa41

Observation 234b9e0a-50f4-453b-8d80-4979d8b30ea7 · outbound

This paper cites SWE-bench: Can Language Models Resolve Real-World GitHub Issues?.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization SWE-bench: Can Language Models Resolve Real-World GitHub Issues?

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.664469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.664469Z digest=sha256:b291fc414bff6e214051721e86b20157a2980a59d755201de6b8cff302f9da9d

Observation 5ae73031-9670-42b4-b8d3-9d6859d5658e · outbound

This paper cites A Self-Improving Coding Agent.

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization A Self-Improving Coding Agent

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:25.727335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:25.727335Z digest=sha256:e8bab78e1719ce6d40477591111e6025f847af3f1ba9c2fef4477757fbd0a1e0

Pith citing papers

No inbound Pith citation observations are available.