Pith. sign in

Paper Citation Record · LEDGER

Agentic Auto-Research is Fuzz Testing

As of 19 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2608.09855.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.09855 v1

Coverage vector

measured 61 of 61 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T05:31:05.820448Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

61 of 61 outbound references displayed

  • verified exact0
  • verified fuzzy37
  • unresolved22
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 292ebfa1-7f72-4bc2-87f7-6e581460cfb1 · outbound

This paper cites AGI House community newslet- ter, July 23, 2026.

Agentic Auto-Research is Fuzz Testing AGI House community newslet- ter, July 23, 2026

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:07.300725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.474864Z digest=sha256:79fe738c2b9b31ef69a97d32c3f6200006381f12809fcaaed7fbba65f5e3c318

Observation 6ac6cdd4-2481-4035-8bad-7ff08b6c7441 · outbound

This paper cites X article, July 31, 2026.

Agentic Auto-Research is Fuzz Testing X article, July 31, 2026

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:07.276062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.482020Z digest=sha256:438dd7b6465864972aadde99d2331c9c21646487ed4936758b596917eede278d

Observation 2f73d93d-d090-49d4-9682-0d0eae95e7c3 · outbound

This paper cites The Oracle Problem in Software Testing: A Survey.

Agentic Auto-Research is Fuzz Testing The Oracle Problem in Software Testing: A Survey

Reference 3

Resolution
malformed identifier
no resolver link, observed 2026-08-11T05:31:05.486529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:31:05.486529Z digest=sha256:db536aca0bb258585e785c943dc706aeb28624ba10c3cfe0a1e753ab5ac55b1e

Observation d2a70ac0-3003-416b-a903-20801523b956 · outbound

This paper cites Directed Greybox Fuzzing.

Agentic Auto-Research is Fuzz Testing Directed Greybox Fuzzing

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T05:31:05.492817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:31:05.492817Z digest=sha256:653e8e1ac0bc3bd6faf5e30c302d506efcb68af1b50d3d3f67436053e5da7532

Observation 1f4fe1fe-efcd-4209-a3e8-3629bd64317d · outbound

This paper cites Coverage-based Greybox Fuzzing as Markov Chain.

Agentic Auto-Research is Fuzz Testing Coverage-based Greybox Fuzzing as Markov Chain

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T05:31:05.499251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:31:05.499251Z digest=sha256:abc1be4df9080e7d3593fed61afc381bfa03d0ea28f9a4a8fbd03ad866bee5c5

Observation a92efadc-d01a-4c47-85c6-4a2ae51b0c2f · outbound

This paper cites On the Re- liability of Coverage-Based Fuzzer Benchmarking.

Agentic Auto-Research is Fuzz Testing On the Re- liability of Coverage-Based Fuzzer Benchmarking

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T05:31:05.505312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:31:05.505312Z digest=sha256:689fc7978a06f0eeb9f33df78c930568b9946a6c49b66098627ae48364b0c780

Observation 197419a9-810d-452a-bbbd-e2cc95ca7a1d · outbound

This paper cites Autonomous chemical research with large language models.

Agentic Auto-Research is Fuzz Testing Autonomous chemical research with large language models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T05:31:05.513125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:31:05.513125Z digest=sha256:a162c482162c85f3feb02ed6827f1662dbaa0fefc12f6d0e57f22ce84a60dabd

Observation da915752-7dbf-4c1e-87e4-1b24e634bf34 · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

Agentic Auto-Research is Fuzz Testing Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T05:31:05.520379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:31:05.520379Z digest=sha256:df61abd8fa91b5c3feb58736a067cfe1721706184b0bde781d5154d8f9202098

Observation 848b6abf-4416-4809-b1f5-cbb32fc53c91 · outbound

This paper cites Exploration by Random Network Distillation.

Agentic Auto-Research is Fuzz Testing Exploration by Random Network Distillation

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:07.264602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.525497Z digest=sha256:c343b1a550b9323d29af2f9d6afe0c7296524a45babdc573437acfa22336cd85

Observation f2aed3c9-32e4-49ad-85e1-5b9a0dae752a · outbound

This paper cites MLR-Bench: Evaluating AI Agents on Open-Ended Machine Learning Research.

Agentic Auto-Research is Fuzz Testing MLR-Bench: Evaluating AI Agents on Open-Ended Machine Learning Research

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:07.252612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.530560Z digest=sha256:83b74ef4f2fdf73c0bd7fb158b76bbe7459e68b2ebc1c1aa32c800d2321187d7

Observation c902bab6-ca3a-4851-9532-99ce31a16992 · outbound

This paper cites Angora: Efficient Fuzzing by Principled Search.

Agentic Auto-Research is Fuzz Testing Angora: Efficient Fuzzing by Principled Search

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T05:31:05.540695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:31:05.540695Z digest=sha256:19d7a6ea8c78a4df5668fb080894aa71456da89517bbdbf851baeef7088f68ee

Observation 3cbf85b9-46ac-48e7-9df4-9d7406a94f51 · outbound

This paper cites An Unsolvable Problem of Elementary Number Theory.

Agentic Auto-Research is Fuzz Testing An Unsolvable Problem of Elementary Number Theory

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:07.239946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.545246Z digest=sha256:456ea00d20197d90b6436c20432b04790e9e5d5e3792fe29c48615b57d8c92e5

Observation 88b556f3-9e90-4b39-8704-bc037375b9b7 · outbound

This paper cites The reusable holdout: Preserv- ing validity in adaptive data analysis.

Agentic Auto-Research is Fuzz Testing The reusable holdout: Preserv- ing validity in adaptive data analysis

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:07.227165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.549624Z digest=sha256:a6628544a076dadd965a1a260e06eef62428088273480a96e362a44ece1a3921

Observation cede52f7-c5b1-4f8a-a549-c5468ecb9208 · outbound

This paper cites Transforming Science with Large Language Models: A Survey on AI-Assisted Scien- tific Discovery, Experimentation, Content Generation, and Evaluation.

Agentic Auto-Research is Fuzz Testing Transforming Science with Large Language Models: A Survey on AI-Assisted Scien- tific Discovery, Experimentation, Content Generation, and Evaluation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T05:31:05.557241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:31:05.557241Z digest=sha256:54aecbf4edb145311879fb56512f4f664059004bbda7c6994fae660d1ed29ed7

Observation ede82e73-2a80-4f52-a927-6c403a586144 · outbound

This paper cites AFL++: Combining Incremental Steps of Fuzzing Research.

Agentic Auto-Research is Fuzz Testing AFL++: Combining Incremental Steps of Fuzzing Research

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:07.214798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.562995Z digest=sha256:e95f94d2db4386d4c1b5fe02cbc95c829885e62f007d61b9a78ebc549c3f6a5c

Observation 6a2fcbb3-66ec-452c-9111-acaf35f683f9 · outbound

This paper cites A Learning Machine: Part I.

Agentic Auto-Research is Fuzz Testing A Learning Machine: Part I

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:07.201753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.568749Z digest=sha256:162e7dec1fcc2d587d8856203d987338805bad1b81367b1bbbd3efd656690435

Observation 8b7caed5-02f5-4f46-b52b-e337cfeaa2b6 · outbound

This paper cites Scaling Laws for Reward Model Overoptimization.

Agentic Auto-Research is Fuzz Testing Scaling Laws for Reward Model Overoptimization

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:07.189026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.576690Z digest=sha256:54c90081c37f389fe481938e0eb9319f8f0ea696bdfed73f8e90e8790fc4e7bb

Observation dad6d249-a49c-4167-ad5a-125044987bd5 · outbound

This paper cites Über formal unentscheidbare Sätze der Principia Mathematica und verwandter Systeme I.

Agentic Auto-Research is Fuzz Testing Über formal unentscheidbare Sätze der Principia Mathematica und verwandter Systeme I

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:07.177014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.583105Z digest=sha256:f98c77cc773a03032aeb7eedae20dff6355cd74ee0e63d12dbd4156d06904db1

Observation 3e3f8bcf-89a8-4b43-8905-4c9519949ef4 · outbound

This paper cites Speculations Concerning the First Ul- traintelligent Machine.

Agentic Auto-Research is Fuzz Testing Speculations Concerning the First Ul- traintelligent Machine

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:07.164986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.588273Z digest=sha256:3c95d9fd5974433a494979a007dfe6366bbdfb1cd1347c2e6e74219943500f59

Observation c6f72cb1-493e-40f9-8061-d23580ada192 · outbound

This paper cites Accelerating Scientific Discov- ery with Co-Scientist.

Agentic Auto-Research is Fuzz Testing Accelerating Scientific Discov- ery with Co-Scientist

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:07.152370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.596213Z digest=sha256:4f40a69d3aa5332f65dd13378a13f24ffc6c3f0ac3a4af6f1937e5e1971b64ad

Observation f7f6ac98-434c-4547-aa7a-19dfccbbf9a7 · outbound

This paper cites Content Fuzzing for Escaping Information Cocoons on Digital Social Me- dia.

Agentic Auto-Research is Fuzz Testing Content Fuzzing for Escaping Information Cocoons on Digital Social Me- dia

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:07.138796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.601942Z digest=sha256:7588c306e65a7de404853651c545011573b2e975cd19537d1d3202f1d144bff2

Observation e656f7f1-c5a5-4d42-9f9b-a98aa3d0b9be · outbound

This paper cites MLA- gentBench: Evaluating Language Agents on Ma- chine Learning Experimentation.

Agentic Auto-Research is Fuzz Testing MLA- gentBench: Evaluating Language Agents on Ma- chine Learning Experimentation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:07.124405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.607908Z digest=sha256:461ac932abf5342cc65033475e654aa069069572f9738e2a0675098931484384

Observation d1bdcd0c-ed0e-4271-8342-d7129988cbda · outbound

This paper cites ALE-Bench: A Benchmark for Long-Horizon Objective-Driven Algorithm Engineer- ing.

Agentic Auto-Research is Fuzz Testing ALE-Bench: A Benchmark for Long-Horizon Objective-Driven Algorithm Engineer- ing

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:07.109605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.614392Z digest=sha256:f3a922eb82f4d201668f534b2e8b306c0db0db591bdb4effdac361e3b66ac4b6

Observation 4f31befa-86ef-4230-803c-aa057a7b681d · outbound

This paper cites Evaluating Fuzz Testing.

Agentic Auto-Research is Fuzz Testing Evaluating Fuzz Testing

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T05:31:05.625385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:31:05.625385Z digest=sha256:3ee679a75a34f267462baa1c937fc8939bc1e89f30babfa8ec3b14b5cf3297ac

Observation dcf6a9a9-5092-4591-b094-b579a35317bb · outbound

This paper cites Abandoning Ob- jectives: Evolution through the Search for Novelty Alone.

Agentic Auto-Research is Fuzz Testing Abandoning Ob- jectives: Evolution through the Search for Novelty Alone

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T05:31:05.630357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:31:05.630357Z digest=sha256:fe937b21c04be0f13cebd9bcc9035850ab078d0703bf11af7a407dd03526dfc9

Observation 327d5189-2323-4076-8437-2fc726dd5868 · outbound

This paper cites On a Measure of the Information Provided by an Experiment.

Agentic Auto-Research is Fuzz Testing On a Measure of the Information Provided by an Experiment

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T05:31:05.636111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:31:05.636111Z digest=sha256:0d507955ada1fd54ac47566307fc0c3148a3a76e77ebba5aa52acdb6861c81be

Observation a2497f5c-202b-43aa-aa18-6c23d1e1bbe1 · outbound

This paper cites The Last Human-Written Paper: Agent-Native Research Artifacts.

Agentic Auto-Research is Fuzz Testing The Last Human-Written Paper: Agent-Native Research Artifacts

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T05:31:05.643545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:31:05.643545Z digest=sha256:f88e5b500e766f8a7387d58287bfacb796571e5b8e0ddacbaffbd498e16660e1

Observation 84105d83-9e65-4b30-9c70-4308e6a3555a · outbound

This paper cites AIGS: Generating Science from AI-Powered Automated Falsification.

Agentic Auto-Research is Fuzz Testing AIGS: Generating Science from AI-Powered Automated Falsification

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T05:31:05.649250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:31:05.649250Z digest=sha256:de284cfaa21be8bab0321a0fa3e9e173278bd9d821b6f5f5a09f6009f3d8c7fd

Observation 55cbf69a-8c55-4d52-8ceb-9c52824d2ebc · outbound

This paper cites The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery.

Agentic Auto-Research is Fuzz Testing The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T05:31:05.654896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:31:05.654896Z digest=sha256:dac0f3ceb5cb7b9fecbde9210950c800b7e62cb9ec9ae2a2999eefedc8f25f7f

Observation 69810099-83c5-4f73-be78-5b1523a4639b · outbound

This paper cites Prompt Fuzzing for Fuzz Driver Generation.

Agentic Auto-Research is Fuzz Testing Prompt Fuzzing for Fuzz Driver Generation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T05:31:05.662457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:31:05.662457Z digest=sha256:0dba78eceadbbe88a266ec56f10f6efee131428e27aa8b81e78e77d4adf92c19

Observation a878b3ff-5be8-4d0f-aad0-48bd62ddb05e · outbound

This paper cites The Art, Science, and Engineering of Fuzzing: A Survey.

Agentic Auto-Research is Fuzz Testing The Art, Science, and Engineering of Fuzzing: A Survey

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:07.095488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.667879Z digest=sha256:e9b007a619022a44ed78c9cecc865d984ee591a4a94c4c6466c0f91597ce59f7

Observation abe0c065-81c7-4daf-8def-71ea10c51c1f · outbound

This paper cites Illuminating search spaces by mapping elites.

Agentic Auto-Research is Fuzz Testing Illuminating search spaces by mapping elites

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T05:31:05.672812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:31:05.672812Z digest=sha256:de8b76973dfe9206e1119e07a4933c04cb9633426051846359f94958da04eacf

Observation edd5b415-64f5-40d3-a714-1412535c893f · outbound

This paper cites AlphaEvolve: A coding agent for scientific and algorithmic discovery.

Agentic Auto-Research is Fuzz Testing AlphaEvolve: A coding agent for scientific and algorithmic discovery

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T05:31:05.680580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:31:05.680580Z digest=sha256:eb3e9a42cce73ff68e6a0e3e67062b45051e5fd88248c81142e5763f5ea2c7d1

Observation 98d154a7-a2de-464c-bff0-1719fca737e8 · outbound

This paper cites The Effects of Reward Misspecification: Mapping and Mitigating Misaligned Models.

Agentic Auto-Research is Fuzz Testing The Effects of Reward Misspecification: Mapping and Mitigating Misaligned Models

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:07.081520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.685691Z digest=sha256:be9ac4ac41b1caf9dad1a0628668939b0baee78f3939d085166d9d8bee0d5e49

Observation 1b570125-e1be-4fa7-8aae-274736b48b0b · outbound

This paper cites LLM Evaluators Recognize and Favor Their Own Generations.

Agentic Auto-Research is Fuzz Testing LLM Evaluators Recognize and Favor Their Own Generations

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:07.068448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.689743Z digest=sha256:f9cc2ea8322f01ce7562bbb9c63234d5ff60f5aae55e4ed2e85573b7cdabc6a3

Observation dcbec012-f47d-4d52-8eca-1f097e44a7c4 · outbound

This paper cites Curiosity-driven Exploration by Self-supervised Prediction.

Agentic Auto-Research is Fuzz Testing Curiosity-driven Exploration by Self-supervised Prediction

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T05:31:05.693685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:31:05.693685Z digest=sha256:2fe62d7ab51fd754d8f0cea5e9e47342fa56d0ba921f693c6e6e33ef165b7fcb

Observation 41b61815-a489-4a45-9912-ca47b4969ff3 · outbound

This paper cites Modern Bayesian Experimental De- sign.

Agentic Auto-Research is Fuzz Testing Modern Bayesian Experimental De- sign

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T05:31:05.698039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:31:05.698039Z digest=sha256:89adb0775a4a7ad43afee4e14bad6ca49a67945389cdaf194960b1dbe0310730

Observation 4b692380-1411-42ad-a1ed-2a625398ee74 · outbound

This paper cites Towards Scientific Dis- covery with Generative AI: Progress, Opportunities, and Challenges.

Agentic Auto-Research is Fuzz Testing Towards Scientific Dis- covery with Generative AI: Progress, Opportunities, and Challenges

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:07.051576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.701986Z digest=sha256:8829fc47a53ff89ecaed3fb1065d1a58c80cf5202f9b14f861afd24ffca820c6

Observation aed8d400-f8a6-4d2f-9da4-51dfc71d1256 · outbound

This paper cites Classes of Recursively Enumerable Sets and Their Decision Problems.

Agentic Auto-Research is Fuzz Testing Classes of Recursively Enumerable Sets and Their Decision Problems

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:07.037142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.705473Z digest=sha256:dee16714550b4489cae2994ac14b65dd8c42fda62090c56d17f895f731441f0a

Observation 24555c14-c00f-4e40-bea1-b7a7a65dd82e · outbound

This paper cites Mathematical Discoveries from Program Search with Large Language Models.

Agentic Auto-Research is Fuzz Testing Mathematical Discoveries from Program Search with Large Language Models

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:07.023915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.709150Z digest=sha256:03820b901677daef91704ff363b4856c434c2a5041481fbd850a08682f63eb25

Observation 82729eb0-7210-409f-b00c-4915c235f44b · outbound

This paper cites Agent Laboratory: Using LLM Agents as Research Assistants.

Agentic Auto-Research is Fuzz Testing Agent Laboratory: Using LLM Agents as Research Assistants

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:07.010463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.712742Z digest=sha256:a8a050bee8b33767a5ceb84749060cc372a9c7f98defc4f4094c4d2129cae635

Observation 93eec5fd-7ca7-48ed-a0f9-7479716bd5a6 · outbound

This paper cites Ultimate Cognition à la Gödel.

Agentic Auto-Research is Fuzz Testing Ultimate Cognition à la Gödel

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:06.976105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.724853Z digest=sha256:4915b9d9fdfd4893d488f5262bf559441f40eab0910765c39edf083360f8594e

Observation b3ecce25-ea2f-470e-80f4-4ab4ecbd7341 · outbound

This paper cites Can LLMs Generate Novel Research Ideas? A Large-Scale Hu- man Study with 100+ NLP Researchers.

Agentic Auto-Research is Fuzz Testing Can LLMs Generate Novel Research Ideas? A Large-Scale Hu- man Study with 100+ NLP Researchers

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:06.958952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.728789Z digest=sha256:8d7128901cc28ab9fbd4bf8f556f958b218c33c550a8e9c71b7b66453a5979e5

Observation a0f3035e-3e98-43d7-84f1-38eb1e279e83 · outbound

This paper cites Defining and Characterizing Reward Gaming.

Agentic Auto-Research is Fuzz Testing Defining and Characterizing Reward Gaming

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:06.938619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.738529Z digest=sha256:3496eef589f68ee8b6930516bef86b39cf755112eeb079b9a3216296983aeb09

Observation 92bfde03-4666-4856-ac0e-9bc02d92e969 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally Can Be More Effec- tive than Scaling Parameters for Reasoning.

Agentic Auto-Research is Fuzz Testing Scaling LLM Test-Time Compute Optimally Can Be More Effec- tive than Scaling Parameters for Reasoning

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:06.924035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.745894Z digest=sha256:afcb73fbdad1974fc88b2f2e4ff5bad4aa2f14b4f2d43429ecfe22da32c84b93

Observation 91f095d7-c654-451b-8739-01a8a8443623 · outbound

This paper cites an unresolved cited work.

Agentic Auto-Research is Fuzz Testing Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-11T05:31:06.910001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.752407Z digest=sha256:c253cb92b0582a35da8b6ece26160c318d956a1778bb3325156ea149d05ecfda

Observation a7398345-f1b0-454c-b858-3690a35788fe · outbound

This paper cites The Virtual Lab of AI agents designs new SARS-CoV-2 nanobodies.

Agentic Auto-Research is Fuzz Testing The Virtual Lab of AI agents designs new SARS-CoV-2 nanobodies

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:06.898150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.757416Z digest=sha256:003090d2f6419644a4af8f0ff96c2964dee720725b5d2efbe1057753278bfd5f

Observation f488f85d-9881-4b70-a391-4c98a13d30d1 · outbound

This paper cites AI Research Agents for Ma- chine Learning: Search, Exploration, and General- ization in MLE-bench.

Agentic Auto-Research is Fuzz Testing AI Research Agents for Ma- chine Learning: Search, Exploration, and General- ization in MLE-bench

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:06.884106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.764745Z digest=sha256:805019d471fee21fb2dee6048590a9902bd0e72e0acc997e9d0328e91b3a0ef4

Observation 37909d0c-1004-46a1-8ff6-0629446e772b · outbound

This paper cites On Computable Numbers, with an Application to the Entscheidungsproblem.

Agentic Auto-Research is Fuzz Testing On Computable Numbers, with an Application to the Entscheidungsproblem

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:06.870103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.769303Z digest=sha256:4699af43d61e945c0c42e1a5a1c07a500a5038c303f96a98d4424f6feeb51433

Observation bcca53e0-4df5-45aa-b8be-f2bddb4583d8 · outbound

This paper cites Large Language Models are not Fair Evaluators.

Agentic Auto-Research is Fuzz Testing Large Language Models are not Fair Evaluators

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:06.855947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.777110Z digest=sha256:2750085dc75c313e64416d8c056c8ff38e9030d3a91dbef32825e95b6508a914

Observation 7619f226-e7de-45f8-accd-2b2b0483ce93 · outbound

This paper cites RE-Bench: Evaluating Frontier AI R&D Capabilities of Language Model Agents against Human Experts.

Agentic Auto-Research is Fuzz Testing RE-Bench: Evaluating Frontier AI R&D Capabilities of Language Model Agents against Human Experts

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:06.841600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.783973Z digest=sha256:e8aa5f9c106c2c6005b962e51cbffca922271ea7f652781a4c85abd1ae107d4d

Observation 9f82b0d5-4471-486d-a775-bfce9c322e00 · outbound

This paper cites The AI Scientist-v2: Workshop-Level Automated Scientific Discovery via Agentic Tree Search.

Agentic Auto-Research is Fuzz Testing The AI Scientist-v2: Workshop-Level Automated Scientific Discovery via Agentic Tree Search

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-11T05:31:05.788489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:31:05.788489Z digest=sha256:5a87b620fb211490131f2963c286c9c4d5ba92dac91736bc5f974c37250bda59

Observation 892aa5ec-4dbd-4629-ad3c-5b2d2302393d · outbound

This paper cites NAS Evaluation Is Frustratingly Hard.

Agentic Auto-Research is Fuzz Testing NAS Evaluation Is Frustratingly Hard

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:06.818921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.792859Z digest=sha256:414180bd01212f1bd08c9097de7d64c04f922f862b853de9a768bdbf01c2b71d

Observation a49e3060-957f-4a10-9ef6-213a7d1157ea · outbound

This paper cites Are We There Yet? Revealing the Risks of Utilizing Large Language Models in Scholarly Peer Review.

Agentic Auto-Research is Fuzz Testing Are We There Yet? Revealing the Risks of Utilizing Large Language Models in Scholarly Peer Review

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-11T05:31:05.797221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:31:05.797221Z digest=sha256:31af9081bb2e212e24aed0098b867f8c0b706a03fc668d1021ffbead68500ba5

Observation 10e3487c-48c1-42b2-8ae8-14f3fa9a4e41 · outbound

This paper cites Dolphin: Moving Towards Closed- loop Auto-research through Thinking, Practice, and Feedback.

Agentic Auto-Research is Fuzz Testing Dolphin: Moving Towards Closed- loop Auto-research through Thinking, Practice, and Feedback

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-11T05:31:05.801485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:31:05.801485Z digest=sha256:3469ed0867603219da76a1e61541ce7320afd5e9f62035da5063ea9b51760685

Observation 548ea2b2-01c4-40e8-8615-b60037ec01e3 · outbound

This paper cites Position: LLMs Can’t Jump.

Agentic Auto-Research is Fuzz Testing Position: LLMs Can’t Jump

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:06.804773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.805313Z digest=sha256:1e7b82f7c709d883f3873d37761726c512d7da37909ace5823e4a4e1f0f490a8

Observation ad627df0-b705-4295-ad88-1ee918b62c1c · outbound

This paper cites Zalewski.American Fuzzy Lop (AFL).

Agentic Auto-Research is Fuzz Testing Zalewski.American Fuzzy Lop (AFL)

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:06.789945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.808978Z digest=sha256:46c88d0673defc8a28362679420e65a2dca28724624fffd7a5a68587fc9178a0

Observation 885c986b-eac3-4395-b451-94c833bb0e6d · outbound

This paper cites LLAMA- FUZZ: Large Language Model Enhanced Greybox Fuzzing.

Agentic Auto-Research is Fuzz Testing LLAMA- FUZZ: Large Language Model Enhanced Greybox Fuzzing

Reference 58

Resolution
metadata mismatch
raw_fallback, observed 2026-08-11T05:31:06.005863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.812639Z digest=sha256:9ed874d4bf8a3f0a9383f9d6eaf5f3d9bf784aa41ded6ce2755f498fe9464612

Observation 9ce60152-0e53-42d5-9555-dbdde4bbee69 · outbound

This paper cites Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena.

Agentic Auto-Research is Fuzz Testing Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:06.777396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.816414Z digest=sha256:4a8abc98b9bb597f0b157379cce2315f4750952e6f129507acde3048465566e4

Observation 24e4d74f-04dd-4c9a-9b19-1fe4016cdedb · outbound

This paper cites Neural Architecture Search with Reinforcement Learning.

Agentic Auto-Research is Fuzz Testing Neural Architecture Search with Reinforcement Learning

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:06.763454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.820448Z digest=sha256:06b49e025bd09887e5b1003a8ae5bb630dbb8beefb9803c358504901f72ccf7a

Observation 5bbfe757-0dc0-45c7-84de-a895aec551b4 · outbound

This paper cites 5977–6043.DOI: 10.

Agentic Auto-Research is Fuzz Testing 5977–6043.DOI: 10

Reference 2025

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T05:31:06.993181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T05:31:05.720662Z digest=sha256:c4666166ab1ba8aebb684aa5cbdd97b853c20fe1ab9dc0ad3969746fdff68819

Pith citing papers

No inbound Pith citation observations are available.