Pith. sign in

Paper Citation Record · LEDGER

Accelerating Matrix Multiplication: A Performance Comparison Between Multi-Core CPU and GPU

As of 12 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 1 inbound Pith citation observation for arXiv:2507.19723.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.19723 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T14:09:27.354350Z

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T16:52:36.806908Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

14 of 14 outbound references displayed

  • verified exact1
  • verified fuzzy12
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation b186840b-c4d3-4ed1-84df-5224c0e0e7a4 · outbound

This paper cites Eijkhout, Introduction to high performance scientific computing.

Accelerating Matrix Multiplication: A Performance Comparison Between Multi-Core CPU and GPU Eijkhout, Introduction to high performance scientific computing

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:09:29.799849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T14:09:25.786059Z digest=sha256:7176143c408bc7d42802dfa6139a728f42e898ec01dbd1c3063a646f607c5b66

Observation c62bda44-39ae-4d50-8e4b-33cad290ebbc · outbound

This paper cites GPU computing performance analysis on matrix multiplication,.

Accelerating Matrix Multiplication: A Performance Comparison Between Multi-Core CPU and GPU GPU computing performance analysis on matrix multiplication,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:09:29.634953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T14:09:25.880965Z digest=sha256:f66aa775f2d2de67355a23605ad95c457eeab8197b6d5eaf06a4c67683ade37f

Observation 58a25c64-85da-4046-8ddd-4ff6fb20edf1 · outbound

This paper cites Neuromorphic computing and beyond,.

Accelerating Matrix Multiplication: A Performance Comparison Between Multi-Core CPU and GPU Neuromorphic computing and beyond,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:09:29.490214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T14:09:25.984102Z digest=sha256:be38358e71e9d6bba2ec70ffd58128c9c63821f995a45345441d7cde2edff9c7

Observation a28ffbfc-d77d-4d30-aea7-cf6695974b5f · outbound

This paper cites The landscape of parallel computing research: A view from berkeley,.

Accelerating Matrix Multiplication: A Performance Comparison Between Multi-Core CPU and GPU The landscape of parallel computing research: A view from berkeley,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:09:29.288904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T14:09:26.074868Z digest=sha256:c4e3a6c828dc8746b4022a9dcf53159491732f3b25ef92d4ee16eeced62a428f

Observation cb4e3de1-44ba-4067-b3e0-d20b7c086123 · outbound

This paper cites Optimization principles and application performance evaluation of a multithreaded gpu using cuda,.

Accelerating Matrix Multiplication: A Performance Comparison Between Multi-Core CPU and GPU Optimization principles and application performance evaluation of a multithreaded gpu using cuda,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:09:29.123921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T14:09:26.177094Z digest=sha256:36f49edae03e041b016ab243e0e84daec5ad261704e38d0be0f8398a582d9892

Observation ffe6ae4f-2514-47a9-9879-b371f4e99cb3 · outbound

This paper cites A survey of cpu-gpu heterogeneous computing techniques,.

Accelerating Matrix Multiplication: A Performance Comparison Between Multi-Core CPU and GPU A survey of cpu-gpu heterogeneous computing techniques,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:09:28.940485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T14:09:26.331535Z digest=sha256:6fbaff559bbe0e6ea2b3a8fae6d48bbfab4aaa112d4dafd6b7189fb7eed217e5

Observation 4c1deeb3-6516-4951-86d4-0fb5cc23e80d · outbound

This paper cites Openmp: an industry standard api for shared-memory programming,.

Accelerating Matrix Multiplication: A Performance Comparison Between Multi-Core CPU and GPU Openmp: an industry standard api for shared-memory programming,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:09:28.774242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T14:09:26.505848Z digest=sha256:646c648d43aa47d3677313b6be8b7065708262bfc88aae12cd7c777878675fbd

Observation 4e794139-93f9-4a33-a1f5-ec607c8e1ea8 · outbound

This paper cites CUDA C Programming Guide,.

Accelerating Matrix Multiplication: A Performance Comparison Between Multi-Core CPU and GPU CUDA C Programming Guide,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:09:28.583904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T14:09:26.664275Z digest=sha256:6e2a68eaea7c28c324944958de97b01389169c16cef741e3307b0f38883924d1

Observation ddc9dab9-0f7f-4921-a7b4-13948109d629 · outbound

This paper cites Debunking the 100x gpu vs. cpu myth: an evaluation of throughput computing on cpu and gpu,.

Accelerating Matrix Multiplication: A Performance Comparison Between Multi-Core CPU and GPU Debunking the 100x gpu vs. cpu myth: an evaluation of throughput computing on cpu and gpu,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:09:28.397092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T14:09:26.837127Z digest=sha256:467b5453d51edf8e14fc7ed887d82874f51bc4afdd97ffdcf1dca7b2e70651dd

Observation a454a7c5-9d0d-4635-92a5-738f68e1bd2e · outbound

This paper cites Performance Analysis and Efficient Execution on Systems with multi-core CPUs, GPUs and MICs.

Accelerating Matrix Multiplication: A Performance Comparison Between Multi-Core CPU and GPU Performance Analysis and Efficient Execution on Systems with multi-core CPUs, GPUs and MICs

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-06T14:09:27.616253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T14:09:26.978488Z digest=sha256:d7206b75395c8e2d4376f0e329bfbedbb17e3bebf2f7ac386e75e054bb55adad

Observation 8001c516-e126-450f-9335-af2a2d64af60 · outbound

This paper cites Understanding the efficiency of gpu algorithms for matrix- matrix multiplication,.

Accelerating Matrix Multiplication: A Performance Comparison Between Multi-Core CPU and GPU Understanding the efficiency of gpu algorithms for matrix- matrix multiplication,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:09:28.218649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T14:09:27.049132Z digest=sha256:61d1abfd00306ed3c784ac25a87a40357a859ecafe7dccdfa565b69ac51499b2

Observation a42bee58-a988-466b-87dc-eede18ec976e · outbound

This paper cites Tsm2: optimizing tall-and-skinny matrix-matrix multiplication on gpus,.

Accelerating Matrix Multiplication: A Performance Comparison Between Multi-Core CPU and GPU Tsm2: optimizing tall-and-skinny matrix-matrix multiplication on gpus,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:09:28.009369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T14:09:27.162777Z digest=sha256:a627c0639f491bdcf692b1da3be048b9f9d6da10c8642a3bdd67bac150095873

Observation 28ff72b4-b9df-499c-99c1-ba5af5e1b68c · outbound

This paper cites Automatically tuning sparse matrix-vector multiplication for gpu architectures,.

Accelerating Matrix Multiplication: A Performance Comparison Between Multi-Core CPU and GPU Automatically tuning sparse matrix-vector multiplication for gpu architectures,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:09:27.782089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T14:09:27.243953Z digest=sha256:7baa3c1fb4bfe4f204d1fb664cade36f048893aedd550b98fbe8c447aeb81799

Observation 4c99f52f-2478-46cf-a51c-a93f85ca9e6e · outbound

This paper cites A Performance Comparison of CUDA and OpenCL.

Accelerating Matrix Multiplication: A Performance Comparison Between Multi-Core CPU and GPU A Performance Comparison of CUDA and OpenCL

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T14:09:27.354350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:09:27.354350Z digest=sha256:e0192497b73dbde1dac1cfc97a8133f03dd02c8ac42a97f8b64a72555704b328

Pith citing papers

Observation 6a158e03-4930-4dd1-b322-e13d3b1b395c · inbound

MAC Performance and Algorithmic Optimization in Matrix Multiplication Workloads cites this paper.

MAC Performance and Algorithmic Optimization in Matrix Multiplication Workloads Accelerating Matrix Multiplication: A Performance Comparison Between Multi-Core CPU and GPU

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-28T17:02:23.816997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-28T16:52:36.806908Z digest=sha256:2134fbc3e2f800c4bd696f4f68f02c48fb71db187527d178d8eed0ef1f3ac58a