Pith. sign in

Paper Citation Record · LEDGER

Dissecting the NVIDIA Volta GPU Architecture via Microbenchmarking

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 22 inbound Pith citation observations for arXiv:1804.06826.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1804.06826 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 22 of 22 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:31:26.627631Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T06:15:00.866473Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 08440cbe-6057-494a-99de-2578b06cf6e8 · inbound

RegDem: Increasing GPU Performance via Shared Memory Register Spilling cites this paper.

RegDem: Increasing GPU Performance via Shared Memory Register Spilling Dissecting the NVIDIA Volta GPU Architecture via Microbenchmarking

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-25T02:00:12.232719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-25T01:57:46.860601Z digest=sha256:61cb8abf54b85fd4209370f5d384a643b80b0786100835007334581dffc1316d

Observation ec0647a6-d7d5-4af9-bd87-db8223e30404 · inbound

FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness cites this paper.

FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness Dissecting the NVIDIA Volta GPU Architecture via Microbenchmarking

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-12T16:22:09.024426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T16:22:08.801066Z digest=sha256:7b7c20fde3ec9f606fc2cd7dd2b2381fd1f5dbe24b676d53ff5e67b77e892f06

Observation aa9075db-eb78-4889-8848-ac14c9a5c83a · inbound

FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning cites this paper.

FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning Dissecting the NVIDIA Volta GPU Architecture via Microbenchmarking

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T02:39:44.843582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-11T02:39:44.770344Z digest=sha256:a31b3b557bd1a5d073895c2eeb41fea9ddd7af810ca88c507a86e7ebba1b037b

Observation 3c8887ad-961f-4d04-bde3-65fe63aeaf23 · inbound

Dissecting the NVIDIA Blackwell Architecture with Microbenchmarks cites this paper.

Dissecting the NVIDIA Blackwell Architecture with Microbenchmarks Dissecting the NVIDIA Volta GPU Architecture via Microbenchmarking

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T17:31:26.627631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:31:26.627631Z digest=sha256:bf8c498f81dd5cab748165f1c159c82fa291f3ee9bc95072260e47cf6aaddb46

Observation 347d6d97-2ac6-4159-8266-2bd4f0d3c79d · inbound

Energy Scaling Laws for Diffusion Models: Quantifying Compute in Image Generation cites this paper.

Energy Scaling Laws for Diffusion Models: Quantifying Compute in Image Generation Dissecting the NVIDIA Volta GPU Architecture via Microbenchmarking

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-17T21:10:16.424075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-17T21:08:23.500137Z digest=sha256:b8d2972d3b72b980de02b471088cca283ba9b249b0e4cc78efbe4dc5cad3d2f0

Observation 4893eb4a-9994-4704-b426-f98cd40fda77 · inbound

Ten-Four: An Open-Source Fused Dot Product Unit for Mixed-Precision GPGPU Tensor Cores cites this paper.

Ten-Four: An Open-Source Fused Dot Product Unit for Mixed-Precision GPGPU Tensor Cores Dissecting the NVIDIA Volta GPU Architecture via Microbenchmarking

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-17T20:40:15.016887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-17T20:36:28.314778Z digest=sha256:c4806db90d999fa684c16bc25ee33e0d8338b244787904083114c8e1e704a79f

Observation 07d3f984-48b6-4ee7-9173-ffb4d9b8a021 · inbound

Ten-Four: An Open-Source Fused Dot Product Unit for Mixed-Precision GPGPU Tensor Cores cites this paper.

Ten-Four: An Open-Source Fused Dot Product Unit for Mixed-Precision GPGPU Tensor Cores Dissecting the NVIDIA Volta GPU Architecture via Microbenchmarking

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T21:26:16.082986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:26:16.082986Z digest=sha256:574537508017b9e5c3e0f75e9cb3f78ccb0dc2d6c52fb9afe3fdb0a203521d80

Observation feeb1c8e-1287-4588-b2c7-2cd1e043acf2 · inbound

TileLoom: Automatic Dataflow Planning for Tile-Based Languages on Spatial Dataflow Accelerators cites this paper.

TileLoom: Automatic Dataflow Planning for Tile-Based Languages on Spatial Dataflow Accelerators Dissecting the NVIDIA Volta GPU Architecture via Microbenchmarking

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-16T21:58:35.792860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T21:55:59.146697Z digest=sha256:0b0bc98898acd9ffeb0ce3cf15821b23feffd7d992e433cf3086663ecafeb22c

Observation 8c52cd65-df55-4e72-bfcc-89b4451eaf4a · inbound

Analyzing Reverse Address Translation Overheads in Multi-GPU Scale-Up Pods cites this paper.

Analyzing Reverse Address Translation Overheads in Multi-GPU Scale-Up Pods Dissecting the NVIDIA Volta GPU Architecture via Microbenchmarking

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:38:14.391040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T20:37:49.558450Z digest=sha256:fa65d6c5e675e4d552ad8f7bd3d6187d880fabe91a73693a5becbbf083862432

Observation e4b1609c-55b0-4128-88aa-55981b48059e · inbound

AdaSplash-2: Faster Differentiable Sparse Attention cites this paper.

AdaSplash-2: Faster Differentiable Sparse Attention Dissecting the NVIDIA Volta GPU Architecture via Microbenchmarking

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T11:20:10.059580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:19:13.934756Z digest=sha256:812c4deff687d7dc5d543dc8c6112ce19167a9e5337ec90dd7b3c19d0b84e7e2

Observation a43d099d-e40a-48ea-9626-59cd57442f6c · inbound

M100: An Orchestrated Dataflow Architecture Powering General AI Computing cites this paper.

M100: An Orchestrated Dataflow Architecture Powering General AI Computing Dissecting the NVIDIA Volta GPU Architecture via Microbenchmarking

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-10T05:51:09.774369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:50:06.986471Z digest=sha256:bc3ed445cee45fe9a1fa77f6e6d36919fee1135787a371f76d57b15a761bca02

Observation e7f2e3c6-71fb-48bb-a415-a1798bcfc562 · inbound

Ocean: Fast Estimation-Based Sparse General Matrix-Matrix Multiplication on GPU cites this paper.

Ocean: Fast Estimation-Based Sparse General Matrix-Matrix Multiplication on GPU Dissecting the NVIDIA Volta GPU Architecture via Microbenchmarking

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-10T02:38:17.318545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T02:38:08.805896Z digest=sha256:5d68a862b5962842fd6ccc0d7913c4a5f690697ad89528025cd1a720324f3536

Observation 6ccbac46-e780-4ab8-8589-cd809d50ee03 · inbound

CUDA Kernel Optimization and Counter-Free Performance Analysis for Depthwise Convolution in Cloud Environments cites this paper.

CUDA Kernel Optimization and Counter-Free Performance Analysis for Depthwise Convolution in Cloud Environments Dissecting the NVIDIA Volta GPU Architecture via Microbenchmarking

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-12T00:26:13.763035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T15:26:41.301014Z digest=sha256:6e82402d8bafd4cf4d688df0287ecf9fb0a58826083d52fe5075ecc141ccac66

Observation 96b08911-3b34-4380-a176-cc0eaaeb7d10 · inbound

AME-PIM: Can Memory be Your Next Tensor Accelerator? cites this paper.

AME-PIM: Can Memory be Your Next Tensor Accelerator? Dissecting the NVIDIA Volta GPU Architecture via Microbenchmarking

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:21:28.290656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T06:37:17.349312Z digest=sha256:58e4a99dcbdab9db831e0e41547d641fcd15fbf74ddeac4feb6393fc47fc4f89

Observation fd7ff87e-225c-4b4c-8be6-d0bcbb6bf7dd · inbound

Microbenchmark-Driven Analytical Performance Modeling Across Modern GPU Architectures cites this paper.

Microbenchmark-Driven Analytical Performance Modeling Across Modern GPU Architectures Dissecting the NVIDIA Volta GPU Architecture via Microbenchmarking

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:30:42.803738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T18:23:44.730845Z digest=sha256:c94cb017d78e5183a01aca7ff05900c0c3f4907c3e66d51ff180dfb4c8152dc7

Observation c0e3aaa3-52b3-471f-8e6d-4845caee57ab · inbound

Rigel: Reverse-Engineering the Metal 4.1 Tensor Compute Path on the Apple M4 Max GPU cites this paper.

Rigel: Reverse-Engineering the Metal 4.1 Tensor Compute Path on the Apple M4 Max GPU Dissecting the NVIDIA Volta GPU Architecture via Microbenchmarking

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-07-03T14:08:21.725116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T07:14:52.640891Z digest=sha256:f85b07d636d30dc0fc00159abc68b02bb2981b47ba3b07bb8f2346c78c18de06

Observation ac4c459f-45b7-4f63-99b6-39470febe5da · inbound

Non-Uniform L2 Cache Latency Across the Streaming Multiprocessors of an NVIDIA L40 cites this paper.

Non-Uniform L2 Cache Latency Across the Streaming Multiprocessors of an NVIDIA L40 Dissecting the NVIDIA Volta GPU Architecture via Microbenchmarking

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-04T09:49:43.792881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T09:35:42.284568Z digest=sha256:c79027ef24b79ff41a2773c2470fe6607bcdbb13fa60e18a556dec2b669e9773

Observation f3568636-2d93-4e58-922a-85d19dd62d37 · inbound

Unprivileged Topology Certificates for Cloud GPU Attestation cites this paper.

Unprivileged Topology Certificates for Cloud GPU Attestation Dissecting the NVIDIA Volta GPU Architecture via Microbenchmarking

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-07-04T11:19:50.742344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T08:01:11.527450Z digest=sha256:2ad903382845c3ee95f2ae1f209a6473486d78118c54c9c9bf916406ceacc122

Observation 0d732189-3db8-4618-ab1e-1c4cf82f11a1 · inbound

Unprivileged Topology Certificates for Cloud GPU Attestation cites this paper.

Unprivileged Topology Certificates for Cloud GPU Attestation Dissecting the NVIDIA Volta GPU Architecture via Microbenchmarking

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-06-26T08:49:15.939674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T08:01:11.527450Z digest=sha256:1c08299671ca1c393885e4422b9a0666018aac31f4a82f93b7ac93fc3e883ee9

Observation 50cd40d0-3d4c-4514-b840-8f47131953ec · inbound

KernelSight-LM: A Kernel-Level LLM Inference Simulator cites this paper.

KernelSight-LM: A Kernel-Level LLM Inference Simulator Dissecting the NVIDIA Volta GPU Architecture via Microbenchmarking

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-06-30T00:54:06.291293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-30T00:48:19.207465Z digest=sha256:6e1b3e7aca49dd36a36eef86900674b528b7b74cb7791a898519fb24937f33d3

Observation e94042d1-9e3d-411b-a0a5-d2b58262e0bb · inbound

KernelSight-LM: A Kernel-Level LLM Inference Simulator cites this paper.

KernelSight-LM: A Kernel-Level LLM Inference Simulator Dissecting the NVIDIA Volta GPU Architecture via Microbenchmarking

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-07-03T23:19:02.223677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-03T23:09:38.092583Z digest=sha256:63ce0109b622035f3b69c43434b58dbe4919477e96b44061a3863f7c437ac43c

Observation 07539891-7b71-4c8d-86ec-12878f1dd380 · inbound

DGNA: Dissecting GPU NUMA Architecture through Microbenchmarking and Data Analysis cites this paper.

DGNA: Dissecting GPU NUMA Architecture through Microbenchmarking and Data Analysis Dissecting the NVIDIA Volta GPU Architecture via Microbenchmarking

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:36.766856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:36.766856Z digest=sha256:4dc3826af5a9c1eaf748d1500d73145f1f1f4c978f295ebb2dcd2d5848fef8d1