Pith. sign in

Paper Citation Record · LEDGER

On Design Principles for Efficient Heterogeneous DRAM-PIM-GPU Systems

As of 21 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 0 inbound Pith citation observations for arXiv:2608.04169.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.04169 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T00:26:16.011775Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

14 of 14 outbound references displayed

  • verified exact0
  • verified fuzzy12
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6166bf76-160a-4325-8ab0-646f50260aed · outbound

This paper cites The Landscape of Compute-near-memory and Compute-in-memory: A Research and Commercial Overview.

On Design Principles for Efficient Heterogeneous DRAM-PIM-GPU Systems The Landscape of Compute-near-memory and Compute-in-memory: A Research and Commercial Overview

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T00:26:15.966271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:26:15.966271Z digest=sha256:c2bb4c108788907794971ba6ccc256ec0fed7f638528621e6ec1381e853e9655

Observation bda73d1d-345a-4c78-90fd-d495b38d17d0 · outbound

This paper cites PIM Is All You Need: A CXL-Enabled GPU-Free System for Large Language Model Inference,.

On Design Principles for Efficient Heterogeneous DRAM-PIM-GPU Systems PIM Is All You Need: A CXL-Enabled GPU-Free System for Large Language Model Inference,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:26:16.147955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-08T00:26:15.971521Z digest=sha256:5dff0cc5319b2ec3c1a94e8fc153cfcc58301d119990b1c3efcf22d3c8aa9e74

Observation 57cd3f78-3d75-49e6-b86e-7d4f9b5edc49 · outbound

This paper cites Pimba: A Processing-in-Memory Acceleration for Post- Transformer Large Language Model Serving,.

On Design Principles for Efficient Heterogeneous DRAM-PIM-GPU Systems Pimba: A Processing-in-Memory Acceleration for Post- Transformer Large Language Model Serving,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:26:16.139667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-08T00:26:15.975687Z digest=sha256:f0445a8310f6223ac47f4124babdba3cbc78c52af64c3334be0fd376b7699853

Observation 4cd3f8b6-bd02-4a9c-b0b4-077edefcc3e2 · outbound

This paper cites AttAcc! Unleashing the Power of PIM for Batched Transformer-based Generative Model Inference,.

On Design Principles for Efficient Heterogeneous DRAM-PIM-GPU Systems AttAcc! Unleashing the Power of PIM for Batched Transformer-based Generative Model Inference,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:26:16.130620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-08T00:26:15.979815Z digest=sha256:9db8e96c26e0dfe1be6aaad39f6c2c87d4763a732087de0b7c7c6e1f7c064c2b

Observation 465a2555-817f-4476-ad20-54291f3ff734 · outbound

This paper cites PAPI: Exploiting Dynamic Parallelism in Large Language Model Decoding with a Processing-In-Memory-Enabled Computing System,.

On Design Principles for Efficient Heterogeneous DRAM-PIM-GPU Systems PAPI: Exploiting Dynamic Parallelism in Large Language Model Decoding with a Processing-In-Memory-Enabled Computing System,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:26:16.120825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-08T00:26:15.983041Z digest=sha256:42f0cced3bb2c36a14ae82eb36f945a2fea919da7ba5dc9030fd8d5b5cf9e787

Observation 35c25566-cfa3-4a75-8221-6a58068d5fee · outbound

This paper cites NicePIM: Design Space Exploration for Processing-In- Memory DNN Accelerators With 3-D Stacked-DRAM,.

On Design Principles for Efficient Heterogeneous DRAM-PIM-GPU Systems NicePIM: Design Space Exploration for Processing-In- Memory DNN Accelerators With 3-D Stacked-DRAM,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:26:16.112394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-08T00:26:15.986624Z digest=sha256:9a10045ac4e2f18b5e3b4b7028b9299fc831f4f7bb3dac3c72b84703171944ee

Observation 777429a4-2d65-459f-b911-4f6695a794b0 · outbound

This paper cites What Your DRAM Power Models Are Not Telling You: Lessons from a Detailed Experimental Study,.

On Design Principles for Efficient Heterogeneous DRAM-PIM-GPU Systems What Your DRAM Power Models Are Not Telling You: Lessons from a Detailed Experimental Study,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:26:16.104107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-08T00:26:15.990559Z digest=sha256:844b58de1ccb620a7f70106e15dc76f17a833520283aa453e0063b982c05a4c1

Observation ef54de1c-ee8b-4278-accb-5e7eac0c9c06 · outbound

This paper cites DRAMPower 5: An Open-Source Power Simulator for Current Generation DRAM Standards,.

On Design Principles for Efficient Heterogeneous DRAM-PIM-GPU Systems DRAMPower 5: An Open-Source Power Simulator for Current Generation DRAM Standards,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:26:16.094264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-08T00:26:15.993327Z digest=sha256:abedc843a948a80a1d473e311212119c271222fb993a67806477fac7208956a0

Observation 42eecebe-0f31-4388-8fc2-d49cdc92d6e5 · outbound

This paper cites Ramulator 2.0: A Modern, Modular, and Extensible DRAM Simulator,.

On Design Principles for Efficient Heterogeneous DRAM-PIM-GPU Systems Ramulator 2.0: A Modern, Modular, and Extensible DRAM Simulator,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:26:16.084462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-08T00:26:15.996245Z digest=sha256:3d6b0b1b5ae7525df0435cd337ea7299665327bf9428987a7b785b54117fd64f

Observation a4f2aee8-b3a5-43bd-b3f2-e7dd14481cee · outbound

This paper cites NVIDIA A100 Tensor Core GPU: Performance and Innovation,.

On Design Principles for Efficient Heterogeneous DRAM-PIM-GPU Systems NVIDIA A100 Tensor Core GPU: Performance and Innovation,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:26:16.076081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-08T00:26:15.999088Z digest=sha256:2e6f2d78b796779910f5a6a03eee8539705b47a6bc4f9375a33031b644988469

Observation 902d8e3d-2bd6-4ab3-8d15-b2fb0254e289 · outbound

This paper cites With Shared Microexponents, A Little Shifting Goes a Long Way,.

On Design Principles for Efficient Heterogeneous DRAM-PIM-GPU Systems With Shared Microexponents, A Little Shifting Goes a Long Way,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:26:16.068313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-08T00:26:16.002290Z digest=sha256:04a6d7202a4a393471c73cab93ba16e70d60f4b60cd42e535a1a95d95652925f

Observation 940b1038-6ba5-4566-8c88-bdc22e7d89a0 · outbound

This paper cites OPT: Open Pre-trained Transformer Language Models,.

On Design Principles for Efficient Heterogeneous DRAM-PIM-GPU Systems OPT: Open Pre-trained Transformer Language Models,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:26:16.059020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-08T00:26:16.005154Z digest=sha256:b52a8209cd0e8129d0beb3eafe890409a14b50a38886fec5a757758a1045e0ec

Observation 74fd6d98-2bec-4bf0-8aa4-96dc39f56856 · outbound

This paper cites Transformers are SSMs: generalized models and effi- cient algorithms through structured state space duality,.

On Design Principles for Efficient Heterogeneous DRAM-PIM-GPU Systems Transformers are SSMs: generalized models and effi- cient algorithms through structured state space duality,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:26:16.050222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-08T00:26:16.011775Z digest=sha256:0a9934fbc1168a0463ea7ed2b736dac55ee45a54318f596ea190d300daee0733

Observation 10eff353-e98b-47e3-8785-50d4e1809aa1 · outbound

This paper cites OPT: Open Pre-trained Transformer Language Models.

On Design Principles for Efficient Heterogeneous DRAM-PIM-GPU Systems OPT: Open Pre-trained Transformer Language Models

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-08T00:26:16.008092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:26:16.008092Z digest=sha256:8c8d3ff22cb83923265c3463e1d8aa858f707385b5f8c922f4db9aab506be2a2

Pith citing papers

No inbound Pith citation observations are available.