Pith. sign in

Paper Citation Record · LEDGER

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels

As of 20 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 0 inbound Pith citation observations for arXiv:2607.24762.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.24762 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T12:30:32.026111Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

43 of 43 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved43
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1971731f-9819-49a5-9d68-e18cae51aae1 · outbound

This paper cites Efficient processing of deep neural networks: A tutorial and survey,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels Efficient processing of deep neural networks: A tutorial and survey,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:26.178857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:26.178857Z digest=sha256:73bb369ab9056d0f7fdfa83b2ea8aaf58734c8f1adb242de790ca675121988e9

Observation 25925247-bfd9-4f71-9fcc-0ae754836ca7 · outbound

This paper cites In-datacenter performance analysis of a tensor processing unit,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels In-datacenter performance analysis of a tensor processing unit,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:26.321349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:26.321349Z digest=sha256:98d503407380ec6427ba0c29610ecd2433290f2a044f5b4503003e6ba1f63cb5

Observation 8f6ff40f-c41e-42ba-b3db-daef252ba32a · outbound

This paper cites A systematic characterization of LLM inference on GPUs,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels A systematic characterization of LLM inference on GPUs,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:26.464343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:26.464343Z digest=sha256:13574abae93b27c59298efea429e88dc24bb699d08c169b112c3c26d6d63386e

Observation 1f88046d-bd44-410c-b9e8-366468ef7862 · outbound

This paper cites Triton: An intermediate language and compiler for tiled neural network computations,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels Triton: An intermediate language and compiler for tiled neural network computations,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:26.606355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:26.606355Z digest=sha256:1a1009eb1ebe85df0abd61331b7e4e01d3f8bfb295b0412ff568d6b87cbd1853

Observation a7dec49e-6593-424e-b528-9b37879f9239 · outbound

This paper cites FlashAttention: Fast and memory-efficient exact attention with IO-awareness,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels FlashAttention: Fast and memory-efficient exact attention with IO-awareness,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:26.776342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:26.776342Z digest=sha256:5558f7854306befe578696085086686680d559098b7ed1a3932e2045270d4647

Observation 15a5c079-575a-4971-9bb1-0ae777476a47 · outbound

This paper cites Autocomp: A powerful and portable code optimizer for tensor accelerators,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels Autocomp: A powerful and portable code optimizer for tensor accelerators,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:26.916959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:26.916959Z digest=sha256:23eda6b65768df5fe1ef2e0e6e65ac64ebf90d65f1029af92e0b4fc94c52a2af

Observation 6a424d0a-ac69-4f2d-9e89-e7c6f562f746 · outbound

This paper cites Geak: Introducing Triton Kernel AI Agent & Evaluation Benchmarks.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels Geak: Introducing Triton Kernel AI Agent & Evaluation Benchmarks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:27.102520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:27.102520Z digest=sha256:07c0af3d68a585342438819d8cfd1eeb2ec2965bde7b08b892c2008a19ce82e2

Observation 7acecd29-b2f7-4157-84bb-80508cd4c738 · outbound

This paper cites KernelBench: Can LLMs write efficient GPU kernels?.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels KernelBench: Can LLMs write efficient GPU kernels?

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:27.269678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:27.269678Z digest=sha256:da485e3f00be116721b77061eb74b43c75a3bb086c681df552b0c6bfca217b7e

Observation ebf936cc-483d-4676-a6c9-198a92e5bff9 · outbound

This paper cites KernelBenchX: A Comprehensive Benchmark for Evaluating LLM-Generated GPU Kernels.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels KernelBenchX: A Comprehensive Benchmark for Evaluating LLM-Generated GPU Kernels

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:27.349032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:27.349032Z digest=sha256:ab4685e00fd973e0073260c095e149fb9c009b789e3cf8690335e867e5248c79

Observation cf8a3bc0-b548-4ab6-a1ce-11250e5c3b1c · outbound

This paper cites TritonBench: Benchmarking large language model capabilities for generating Triton operators,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels TritonBench: Benchmarking large language model capabilities for generating Triton operators,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:27.467840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:27.467840Z digest=sha256:b652b2afce16f00329a1568caddf032d0caf8809446f1cac4b5c965d8eb2c850

Observation ee1d5d5f-ffef-448f-ba1b-f08fa11edc70 · outbound

This paper cites MultiKernelBench: A Multi-Platform Benchmark for Kernel Generation.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels MultiKernelBench: A Multi-Platform Benchmark for Kernel Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:27.620409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:27.620409Z digest=sha256:11b100cebf3e724f4356197f5c71ac25dc259ba2a4306c95914b3ba2f77b83ec

Observation 3e4ec3f7-dd9f-4493-8e70-bb36463823a5 · outbound

This paper cites CUDABench: Benchmarking LLMs for text-to-CUDA generation,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels CUDABench: Benchmarking LLMs for text-to-CUDA generation,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:27.759973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:27.759973Z digest=sha256:e54527a14ccce295ed57892566fb84334dc5d57585f392b90671acbb37fc0c2c

Observation 69e3f430-ff49-4f17-b5fa-381b2505f563 · outbound

This paper cites BackendBench: An evaluation suite for testing how well LLMs and humans can write PyTorch backends,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels BackendBench: An evaluation suite for testing how well LLMs and humans can write PyTorch backends,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:27.914125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:27.914125Z digest=sha256:8ed07b5531a7d219f71acd7af0a3c984f8d408bc52fcb2bd058d2a5fb1742c73

Observation ac666df2-100b-4f17-85ff-b6bada40bcdd · outbound

This paper cites FlashInfer- Bench: Building the virtuous cycle for AI-driven LLM systems,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels FlashInfer- Bench: Building the virtuous cycle for AI-driven LLM systems,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:28.074757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:28.074757Z digest=sha256:a69dd7dee355551a5ece2a2d0536de3c113c63c9044bd3c5df102dd1f806ddef

Observation 5a3d9507-9764-4c7b-8de6-db46453091de · outbound

This paper cites CudaForge: An agent framework with hardware feedback for CUDA kernel optimization,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels CudaForge: An agent framework with hardware feedback for CUDA kernel optimization,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:28.230400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:28.230400Z digest=sha256:14f3d7fe4745b1af08da742a1d539a983812ecce15bdbd1a2d65443922f9e4ef

Observation 4d0106e1-eee0-4b52-b862-7efe271f87e9 · outbound

This paper cites Astra: A multi-agent system for GPU kernel performance optimization,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels Astra: A multi-agent system for GPU kernel performance optimization,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:28.429142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:28.429142Z digest=sha256:9d8b20653832b596a6bfb100d8417207c53d52f3cf42cc610bf46c8053b7f9b8

Observation aac71d27-5e89-449f-8141-48457486147d · outbound

This paper cites Towards robust agentic CUDA kernel benchmarking, verification, and optimization,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels Towards robust agentic CUDA kernel benchmarking, verification, and optimization,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:28.664212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:28.664212Z digest=sha256:0c9d3b7771e6e36202568e9c0e856c0d63e2e4ff4c03e06d3da613081c95f2a9

Observation fe5c0d64-fea4-4f74-8e62-feac4f59f4eb · outbound

This paper cites KernelAgent: Multi-agent GPU kernel synthesis and optimization,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels KernelAgent: Multi-agent GPU kernel synthesis and optimization,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:28.844946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:28.844946Z digest=sha256:a9acdd10fc314814d6e99dbf99c814ad9c62e2a8c6a6174ba38e91c60f4376a7

Observation 0ab9e2cb-02dd-42ab-8e22-1d56bed21f32 · outbound

This paper cites Complete anytime beam search,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels Complete anytime beam search,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:29.054222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:29.054222Z digest=sha256:b779b2ed227b09d5e72a94f155a38d4ca068f3e7fdef0db80f217897f4852a2f

Observation 2048d913-47f0-4260-a38a-f78686939b10 · outbound

This paper cites Bandit based Monte-Carlo planning,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels Bandit based Monte-Carlo planning,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:29.198605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:29.198605Z digest=sha256:3d68865c3039aa8366889c027113563d8585501957a78ed9c8bd9634343af4af

Observation e2ad9d42-aa80-443c-96ed-3d5a88a17e51 · outbound

This paper cites Deep residual learning for image recognition,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels Deep residual learning for image recognition,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:29.351224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:29.351224Z digest=sha256:b8120ecf747d2345908732a1d3eb20a96c9c0dc50f07836056cb04c40f4f3dfa

Observation 011e335f-1c45-452f-8c61-007d0a3eb00e · outbound

This paper cites Stable diffusion 3.5 medium model card,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels Stable diffusion 3.5 medium model card,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:29.480014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:29.480014Z digest=sha256:602d99118ef9400e61982cbf6e95e90b14c1108cd9df58a11c3b3273b8f48e8e

Observation 4017e7e9-ff10-478a-9d07-2d1bca640a3d · outbound

This paper cites google/gemma-4-E2B-it model card,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels google/gemma-4-E2B-it model card,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:29.623340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:29.623340Z digest=sha256:28359ecf8df543d0619eca97b2dd925778d7f8d22a84e9789633342cb3e00c85

Observation 6d253fc1-25e5-470c-9124-7871f2c466a5 · outbound

This paper cites Qwen/Qwen3.5-35B-A3B model card,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels Qwen/Qwen3.5-35B-A3B model card,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:29.773982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:29.773982Z digest=sha256:87efeb1c8668a955911d54544423f8337d9021b934707916d447ab6303f0b1bd

Observation 609c88c6-3c38-4e72-9a5a-366e94dc4271 · outbound

This paper cites DGX Spark user guide: Hardware overview,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels DGX Spark user guide: Hardware overview,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:29.922278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:29.922278Z digest=sha256:fa073dc62e647ad92181c8ad8469339dff77e1a6724e130fe56dbb6750a18b4f

Observation 6a393e69-32a8-42b0-a8ce-719a6f99849d · outbound

This paper cites PyTorch: An imperative style, high-performance deep learning library,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels PyTorch: An imperative style, high-performance deep learning library,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:30.071296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:30.071296Z digest=sha256:a41eff274623ff2dad4ff536f164f76493256b3e7e5f1e2c59abd00d6c9c98f5

Observation 5f81cf21-94c3-4506-b1ea-468383c6dd30 · outbound

This paper cites CUDA programming guide,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels CUDA programming guide,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:30.262483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:30.262483Z digest=sha256:17006b61ef00dd236ae21a37c20e6e0e0d90b479c17f5c5ca3f3945637c62965

Observation d5a7e942-9b8c-4e34-b1b7-3bb10392d673 · outbound

This paper cites CUDA C++ best practices guide,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels CUDA C++ best practices guide,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:30.467461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:30.467461Z digest=sha256:1b5c6bb8ae518f318fad7b7d1a7f78fb414ac304d11b724b8e85a1ae518c1423

Observation 892d299e-27b6-4fa5-99b7-d48e9eae8511 · outbound

This paper cites Tvm: an automated end-to-end optimizing compiler for deep learning,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels Tvm: an automated end-to-end optimizing compiler for deep learning,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:30.689021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:30.689021Z digest=sha256:c25634c42a02bdbaa5370a267b795a1ac4725df9a49862692bb141609770269e

Observation 8ca0e47d-8324-4a05-80b1-edbf8bed4ac3 · outbound

This paper cites Learning to optimize tensor programs,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels Learning to optimize tensor programs,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:30.863019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:30.863019Z digest=sha256:02410c197132cc761725fe30162c75a1ced50accdc720206ef6986aadb16b2cb

Observation 789fd873-2679-4da6-bbdf-b6265dc57771 · outbound

This paper cites Ansor: Generating high-performance tensor programs for deep learning,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels Ansor: Generating high-performance tensor programs for deep learning,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:30.945666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:30.945666Z digest=sha256:810bca8b2d599da3ee7cec1f106928b193357652566c71118fcfe4abb7f009c5

Observation 4beaeafd-b4d0-40dd-b4e0-4a78a707faa8 · outbound

This paper cites torchvision.models.resnet50,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels torchvision.models.resnet50,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:31.004673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:31.004673Z digest=sha256:ec9c591617e2c8a0640a00bf8bd4943d39d8152311568d3d1442ceb173760469

Observation fe150bee-c4ab-46e3-ad8c-94d2388a9e38 · outbound

This paper cites Do ImageNet classifiers generalize to ImageNet?.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels Do ImageNet classifiers generalize to ImageNet?

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:31.089960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:31.089960Z digest=sha256:33fda9c039f252d1fe50359a90fcf873239ee69acaeacf74b15ae3c9c0a50da8

Observation 285896f3-d2c6-45cc-8af3-49dd7d60f4de · outbound

This paper cites Diffusers: State-of-the-art diffusion models,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels Diffusers: State-of-the-art diffusion models,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:31.198319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:31.198319Z digest=sha256:c44547b858bef8c6995ee175347ef6507421aab1ffec6415dcfbbd12b54afc86

Observation 27ebcfc4-0ac0-4268-af58-72fa2a32ac8a · outbound

This paper cites T2I-CompBench: A comprehensive benchmark for open-world compositional text-to-image generation,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels T2I-CompBench: A comprehensive benchmark for open-world compositional text-to-image generation,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:31.304018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:31.304018Z digest=sha256:27a47375d36a08a8b18b1ad7c8b750a13473f7a8b9a6e7de2b3d4c3852477665

Observation 7a74e52e-c2df-4419-9e66-851ad5d55655 · outbound

This paper cites Judging LLM-as-a-judge with MT-Bench and chatbot arena,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels Judging LLM-as-a-judge with MT-Bench and chatbot arena,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:31.421596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:31.421596Z digest=sha256:28df3af0d278992f8ff7b6aa8bafd2ed47f58ddca8b378a34c9c90d4e35e713b

Observation f84e7605-fc9a-495a-8aeb-063b34e1af03 · outbound

This paper cites ShareGPT Prompts Annotated,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels ShareGPT Prompts Annotated,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:31.503958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:31.503958Z digest=sha256:095f07b0e84d487575bdbab80b16e0e2249081b4325c76c6fee4a7624291d889

Observation 4060d277-5e4e-4587-ae63-2cc4018ec38f · outbound

This paper cites LongBench: A bilingual, multitask benchmark for long context understanding,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels LongBench: A bilingual, multitask benchmark for long context understanding,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:31.572369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:31.572369Z digest=sha256:a56e4cf4af6cfc44980823b7255af019aef6fa8a734fa1eedfb6c94064ff653c

Observation cf6d4265-ee8b-49e5-9bff-8447c34efbdd · outbound

This paper cites Introducing Claude Opus 4.7,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels Introducing Claude Opus 4.7,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:31.639521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:31.639521Z digest=sha256:6c36002768132021b7cf5813faeb2500aa2d0c295c1c4f0c4495c0b193dee670

Observation 3959a1e7-6f4b-46f4-a12e-07b16174161a · outbound

This paper cites ComputeEval: Evaluating large language models on CUDA,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels ComputeEval: Evaluating large language models on CUDA,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:31.722780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:31.722780Z digest=sha256:0d0875492652c67db306d8cb707339b15a2925a45e0038ffe191cb833b0a0743

Observation a9385890-c6d1-44c0-a485-d3b9b9803913 · outbound

This paper cites TritonGym: A benchmark for agentic LLM work- flows in Triton GPU code generation,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels TritonGym: A benchmark for agentic LLM work- flows in Triton GPU code generation,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:31.824296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:31.824296Z digest=sha256:346820775a4665f946dfd52443405e49def6be90106b29c87736544fc8f6faa6

Observation a25549b0-c343-4c98-94ed-a52768766abf · outbound

This paper cites KernelFalcon: Deep agent architecture for autonomous GPU kernel generation,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels KernelFalcon: Deep agent architecture for autonomous GPU kernel generation,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:31.939574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:31.939574Z digest=sha256:864ca19658299dc97a0d0bbd16d74fac3b971e03caf73e1efa2b4aa6b7cde0c0

Observation 87976858-c3f5-40f1-abef-9948308b0a7e · outbound

This paper cites torch.nn.functional.scaled_dot_product_attention,.

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels torch.nn.functional.scaled_dot_product_attention,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-02T12:30:32.026111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:30:32.026111Z digest=sha256:5780769019c13629e721a7e0009f5275d9a5cbb56ee1ca9a866a9a4dedd869df

Pith citing papers

No inbound Pith citation observations are available.