Pith. sign in

Paper Citation Record · LEDGER

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production

As of 11 August 2026, this Paper Citation Record lists 90 of 90 outbound references and 0 inbound Pith citation observations for arXiv:2605.11733.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.11733 v1

Coverage vector

measured 90 of 90 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-13T05:17:24.147248Z

measured 90 of 90 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

90 of 90 outbound references displayed

  • verified exact41
  • verified fuzzy34
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7a73db7a-d303-4502-91dd-d36a7156eae8 · outbound

This paper cites Energy and ai.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Energy and ai

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.256594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:6327635339355c966cdf538b953cf09046185691e9027f2f60ccb3ad8fa25d17

Observation e7c8a8cd-0057-49ea-8fdf-4633bd4d463c · outbound

This paper cites Analyzing artificial intelligence and data center energy consumption.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Analyzing artificial intelligence and data center energy consumption

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:27:19.474450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:b1c181759257900ad5be5b239d107cb69dbb2a04179509faa6ee2693c7d53ac1

Observation 59e6994f-93d6-49ac-8b94-c227d1ab5fa5 · outbound

This paper cites Ai factories.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Ai factories

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.266037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:b2201c32c233668aa5eaab303071488bc23ca957828f57ddf658188ab25366c4

Observation 65031900-53f4-404e-91be-95d04483228e · outbound

This paper cites Ai inference.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Ai inference

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.262793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:87c717a413d384297ae1706755bd30d08be64431836f2c1438dbd561abaa3bb4

Observation c8413658-dab2-4096-abbe-31da21207ead · outbound

This paper cites Api pricing.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Api pricing

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.261107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:6985087b2bd54d3e1a041c407c7eb1f48a5579c5a6b36a5c1d49fbc9f4a03358

Observation 3eb02857-4a97-4d6f-b3d9-222ed5b70694 · outbound

This paper cites Models overview and api pricing.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Models overview and api pricing

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.267668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:33fe79bb410a6a80069b0bbb851c830a8da8e79e6d400e3845d86a8f9e9564ba

Observation abc144d8-6bae-48b8-af45-0cba65a6be1a · outbound

This paper cites Models and pricing.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Models and pricing

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.264442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:2d2a049c558ca2982095ac5a94d3685ab5c665e15cbfdeeed8ebcf3ee5fc34fc

Observation 1ed037e6-6394-4445-8bbc-bd86833ecf85 · outbound

This paper cites Smith, and Oren Etzioni.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Smith, and Oren Etzioni

Reference 8

Resolution
verified exact
doi, observed 2026-05-13T05:22:18.958105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:380f685ee9f400f08d5eb6e9d11cce075c8a60c33bbe25f7f9f57eba764c5ae1

Observation 3100e300-11dd-44fc-808b-370644d00ca9 · outbound

This paper cites Patterson, J.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Patterson, J

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.258422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:9856d1a2068c67a728dc4824be6028dca6c885c65c456c1304467086be0f1e8f

Observation f3c2c794-1f25-4ec4-be2c-2d9664c4ee4b · outbound

This paper cites Carbon Emissions and Large Neural Network Training.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Carbon Emissions and Large Neural Network Training

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-13T05:27:19.471326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:99ddeabcbe67f8ccb95fb655bba1e8701d43c978ee00f9930fa1769d2a29eb61

Observation aed7967d-58e9-4689-a0f3-615913343e5e · outbound

This paper cites Patterson, Joseph Gonzalez, Urs Hölzle, Quoc V.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Patterson, Joseph Gonzalez, Urs Hölzle, Quoc V

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:22:18.956217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:785a2a261bcfa8a97fc3bcb0ca3135b47ff5c0d6580ba194257e2bb29118ef40

Observation afbd22f0-89ec-4217-bd6a-a3e670da904b · outbound

This paper cites Sustainable AI: Environmental Implications, Challenges and Opportunities.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Sustainable AI: Environmental Implications, Challenges and Opportunities

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:27:19.377263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:09e920d90236e009e8e64949214078940812fb7e16d55c8c7da7034a98062168

Observation a9951de1-d5e9-4da9-b36f-b287dd22bab3 · outbound

This paper cites Quantifying the Carbon Emissions of Machine Learning.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Quantifying the Carbon Emissions of Machine Learning

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:52:44.534862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:ac858acd2f91752f41cf833c53cdf88e2fea3c1524a589ef6b5a4b379df1cea8

Observation a4439e91-1fb0-447f-8976-a97262c5ca2e · outbound

This paper cites Energy and Policy Considerations for Deep Learning in NLP.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Energy and Policy Considerations for Deep Learning in NLP

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:27:19.423169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:37ceb69221c4d71050e4694c4f17147fb073d153f23c8ab8d328556d967b4d42

Observation a23c13ff-2777-4a44-844e-764538deafcb · outbound

This paper cites Beyond Individual Accountability: (Re-)Asserting Democratic Control of AI.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Beyond Individual Accountability: (Re-)Asserting Democratic Control of AI

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T05:22:18.951522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:a34f933685a2055f1c3732df203a9a0b2ca4c66e7c3d3879ecccaaa309aab0c9

Observation 737134fc-2ddb-4aca-acdc-32f860e038c4 · outbound

This paper cites Mlperf inference v4.1 power results.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Mlperf inference v4.1 power results

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.316895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:d77fb2d628ad76ad969b91777bdc0fe8f1dde6f75c592a542a7767303941e07b

Observation 6ec9305f-b802-4c9d-97eb-1c6e88df3d98 · outbound

This paper cites an unresolved cited work.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-05-13T10:42:38.318728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:806489ff2fd638bcbb4a6a7865c4fe4b2129768091b4230afa065d95caeba8d0

Observation 5c62286c-2475-48ab-ac45-87dcacd0e872 · outbound

This paper cites an unresolved cited work.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Unresolved cited work

Reference 18

Resolution
verified exact
doi, observed 2026-05-13T05:22:18.953304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:df3a62db3476bf957b8b6b1fe87b51db38860be1c221c3d0aad86aa36056a08c

Observation 151263ce-1cc2-450f-9d2f-47bfc6479c14 · outbound

This paper cites 2024 global data center survey results.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production 2024 global data center survey results

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.335016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:db6f1c1e888e62f47e5e93a671a43477d535ceefbf35576f6c2e5d8291a67ad9

Observation 6a91dd2b-00a9-4b45-ae3a-5b2db1210cb9 · outbound

This paper cites an unresolved cited work.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-05-13T10:42:38.315192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:f17844555b3785227fcc870bc7d5a925cc6564d1a5c8536fb200d5761dfce383

Observation 4aa516c3-7f5c-4946-a1e1-611e2fae73ad · outbound

This paper cites From Words to Watts: Benchmarking the Energy Costs of Large Language Model Inference.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production From Words to Watts: Benchmarking the Energy Costs of Large Language Model Inference

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:27:19.429463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:99433c1d8b95d551f3e4f27570122b6928ecdd03d58ae6902489789a8b628573

Observation 43f42230-32e5-4121-bbfc-fa008067d4b9 · outbound

This paper cites Tokenpowerbench: Benchmarking the power consumption of llm inference.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Tokenpowerbench: Benchmarking the power consumption of llm inference

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:27:19.370752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:7d8339a87c75abec9aff23f16b31b1fc08238f4cab9f7ec8d865f29126405289

Observation e63ca9ad-0a31-4608-a110-f97301767b85 · outbound

This paper cites year =.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production year =

Reference 23

Resolution
verified exact
doi, observed 2026-05-13T05:22:18.939138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:3e803c6ef5636f899da093882906b478b369dae61dae2b2a8384b37b87e654a2

Observation 2abd72ed-86f1-4039-884d-cfbbd5ce43fa · outbound

This paper cites Roofline: An insightful visual performance model for multicore architectures.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Roofline: An insightful visual performance model for multicore architectures

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T05:22:18.941810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:1e2965b389b95b69c30ab9cc13d27f77f378a96402206be1b1c5aa7d80c2c3c4

Observation 1d04ef3d-1e8d-4efb-b0f7-d6360198bc46 · outbound

This paper cites Sevilla and E.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Sevilla and E

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.308423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:c7b3e62d0a9473d271a39b6bf514450f1a1c257d39a7faab1c8ae6ec176717fb

Observation 09a7ce7e-3961-42b1-a95a-0deb6611faa9 · outbound

This paper cites Where do the joules go? diagnosing inference energy consumption.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Where do the joules go? diagnosing inference energy consumption

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:27:19.405195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:fa37adf888b8d27aca8a3afb06404bb91ddb2d7f7a6b2c1d576ee83420a81e3e

Observation 3fc8d1af-dbc3-4c9a-aa0f-1ed3f87e9f68 · outbound

This paper cites Delavande, R.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Delavande, R

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:27:19.385538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:edf818c08673ab5911a8674bd7c219b746aa5669118f34702efdab1d84d9cd2a

Observation c2341ae7-b72c-4f56-8704-ec472cc048a8 · outbound

This paper cites an unresolved cited work.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-05-13T10:42:38.306607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:5f987749a7e1882d9813437684acc6e1e128f2bae4a73cb8584b3bc81b9c457a

Observation 764826af-88f7-4541-9c57-1109151e9968 · outbound

This paper cites SweetSpot: An Analytical Model for Predicting Energy Efficiency of LLM Inference.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production SweetSpot: An Analytical Model for Predicting Energy Efficiency of LLM Inference

Reference 29

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T05:27:19.426549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:161c55c319d596711bf83f2ec3122bfad5b681c4373881729114eaf5edbf3afc

Observation b9f2525e-f157-4293-952e-b01d9941666b · outbound

This paper cites Nvidia h100 tensor core gpu: Product specifications.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Nvidia h100 tensor core gpu: Product specifications

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.326162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:99cde47fa7a0b6146fed540fd97ead7bc297c6640b19d3bbc35c0a96889acc17

Observation 80194f8c-23f1-4b99-a49d-ff4472ae8a6c · outbound

This paper cites Nvidia hgx platform specifications (hgx h100 4/8-gpu).

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Nvidia hgx platform specifications (hgx h100 4/8-gpu)

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.310045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:c3344a4969d146a0e5131dcdb68f6dcf7489e7f351a9a5b0105aa6d2384bac40

Observation e9814d3a-fe80-422c-9e88-92fbb81094ad · outbound

This paper cites Ai is set to drive surging electricity demand from data centres while offering the potential to transform how the energy sector works.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Ai is set to drive surging electricity demand from data centres while offering the potential to transform how the energy sector works

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.311882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:9ada1457451eeb0404c5d417004cdde5b6e4d3306cffc18ba504ff030559bf2a

Observation e393c692-bbe8-4c5c-8255-b43f386acac6 · outbound

This paper cites DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-05-13T05:27:19.459368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:26ffef6ea0e0cfff556a5fc66d6a87aa78320cb13c969c44b5ba1c8a5892d478

Observation 887a3b89-ba09-42f8-ba38-80b15edebb73 · outbound

This paper cites Scaling Laws for Neural Language Models.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Scaling Laws for Neural Language Models

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-05-13T05:27:19.374193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:7ffbce01983ae566932af9cf6a023fa1da31abdf8dfa9f0ce404a1d85ad4c9f7

Observation fac672f3-b55f-459f-978d-a4836d507c14 · outbound

This paper cites Training Compute-Optimal Large Language Models.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Training Compute-Optimal Large Language Models

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-05-13T05:27:19.379963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:5f5325ac6a63433aa33b4e5835fb40967e0db480936a06dbf37c3c3f8cd2fa5a

Observation 5f58ee32-16b0-476e-982b-abfdc488ba72 · outbound

This paper cites Scaling Language Models: Methods, Analysis & Insights from Training Gopher.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Scaling Language Models: Methods, Analysis & Insights from Training Gopher

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-05-13T05:27:19.462039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:7f488c037e9d041ba29170babb270cc933f5f0de3ef350c1cb2f68551223afc1

Observation 8f64c5a3-b694-43df-bffc-e9a8aa97715a · outbound

This paper cites Compute Trends Across Three Eras of Machine Learning.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Compute Trends Across Three Eras of Machine Learning

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:27:19.440978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:90201d84ee33dc9444fbfe2d643053eee0819ff9f9ffad24416bbbaf7a74d767

Observation 0b81757e-f4b6-4e93-aaac-d9715733017e · outbound

This paper cites FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-05-13T05:27:19.443530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:68c143c7ad2b75f958fc969272c95e776f72d8bc114c9ae4e430f734e97c6c9f

Observation 2f701fc2-3453-4ce1-866f-64a16063b7a8 · outbound

This paper cites Efficient Memory Management for Large Language Model Serving with PagedAttention.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Efficient Memory Management for Large Language Model Serving with PagedAttention

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-05-13T05:27:19.432534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:5081acaf97a7af75f70529552a969b9944dfae12f775134923208d5921bee27b

Observation a3e9881c-8adc-485a-bdd0-b99f88ce3ad2 · outbound

This paper cites GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-05-13T05:27:19.435378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:98b2a8dc501b3828c02a365924b801c8a1da7e2b8880fdcd7bb1bd58f3b0b924

Observation a23779db-e224-46f5-85ce-f55102be1db0 · outbound

This paper cites AWQ: Activation-aware Weight Quantization for LLM Compression and Acceleration.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production AWQ: Activation-aware Weight Quantization for LLM Compression and Acceleration

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-05-13T05:27:19.438188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:9deb1c4a41d46a17c1d0c069fa915712ba2e3f01f32f3a02f2d56a5e423d66be

Observation 3e686972-1574-4ea8-b1f7-8c2c52cb941c · outbound

This paper cites Dissecting the Runtime Performance of the Training, Fine-tuning, and Inference of Large Language Models.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Dissecting the Runtime Performance of the Training, Fine-tuning, and Inference of Large Language Models

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:27:19.391042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:a6cc09451f49c3d4907cd1e8950d2fb71d5a2e31b33e50ac4ae4ae4209092f6f

Observation e9a4404b-f545-4a68-931a-16742e93f134 · outbound

This paper cites Knowledge-Centric Hallucination Detection.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Knowledge-Centric Hallucination Detection

Reference 44

Resolution
verified exact
doi, observed 2026-05-13T05:22:18.947071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:d63eb3002a06d600d987ea3c9bbe1a3899b11f90030dfce0b6dab6ad02769e5e

Observation 2571e0da-c795-4886-88f8-afc45f61a3fd · outbound

This paper cites Department of Energy.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Department of Energy

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.320296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:cf632fe9da171b8132b4975e7e3e6391006563148f13e8db16eda9adc07af644

Observation 30530823-64d3-434e-8c6f-bc0a82584b13 · outbound

This paper cites Juniewicz.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Juniewicz

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.301279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:67ff83586f20a70c6296d82658942a5997d7a04caed0356e26bec3e98700b257

Observation 33fcbdba-9968-4a3b-9ca4-83579d06d6b6 · outbound

This paper cites Combined Alphabet, Amazon, Meta, Microsoft, and Oracle capex extracted from SEC EDGAR 10-Q/10-K filings.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Combined Alphabet, Amazon, Meta, Microsoft, and Oracle capex extracted from SEC EDGAR 10-Q/10-K filings

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.304987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:3215c1a007828ff5b18f3747852c5b882c1cab5c96f60039f5fddf65c21569f5

Observation 84ffc216-3f1d-4350-880c-c91fa5fab685 · outbound

This paper cites an unresolved cited work.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-05-13T10:42:38.291342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:8dd5bbcd6fb965db3ee5952eb7c4b33755674c1c853b2b26e1dff9c56905c9d2

Observation 549c7c84-6409-4fcc-94b8-e72a251695aa · outbound

This paper cites Doubao surpasses 120 trillion daily tokens as usage doubles in three months.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Doubao surpasses 120 trillion daily tokens as usage doubles in three months

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.293780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:00f946c598966203714059025e718b6fd002a399d8de9eaba7ba0bb5cdda78a1

Observation bcc33e17-4d56-4a2e-b0a2-a3b3c07011bd · outbound

This paper cites Wulf and Sally A.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Wulf and Sally A

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:22:18.944786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:3d977f761054de93010011d0131d4e1c6665a9d166e2a5d9abbb9d1ace1979d5

Observation c3c6545e-a9d3-46d8-8c8f-78df0d022980 · outbound

This paper cites Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-16T23:46:30.294205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:25d66b13e8523685ac247941b7d33cfbf045313f613afc5edd90eb3464265200

Observation 71187678-6832-414d-baf3-55c3763a22ad · outbound

This paper cites Deepseek-v4: Towards highly efficient million-token context intelligence.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Deepseek-v4: Towards highly efficient million-token context intelligence

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.289661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:a36d64df3798117b8c7793d0a9e43fdce31c0318d6340df73b3b3f33aea6f313

Observation eae5b671-f73b-47f5-a380-eab60f7249a2 · outbound

This paper cites Chunkkv: Semantic-preserving KV cache compression for efficient long-context LLM inference.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Chunkkv: Semantic-preserving KV cache compression for efficient long-context LLM inference

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:27:19.417404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:7c72f8b10acbffc451554c188d7153339d29c1a5edbec97fb77ee5174054f590

Observation 8da5e788-22d0-48ef-a47e-460901d27162 · outbound

This paper cites H$_2$O: Heavy-Hitter Oracle for Efficient Generative Inference of Large Language Models.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production H$_2$O: Heavy-Hitter Oracle for Efficient Generative Inference of Large Language Models

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-17T18:00:50.699870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:6b5145fca4797291c7ca92682cdeffd7af8e78c79758284f3b09106954060748

Observation 82cc6781-1167-4e40-b11c-4477ac3a65d7 · outbound

This paper cites FlowKV: Enhancing multi-turn conversational coherence in LLMs via isolated key-value cache management.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production FlowKV: Enhancing multi-turn conversational coherence in LLMs via isolated key-value cache management

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:27:19.402283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:171ee0d3d1a38a2e7b2cd854a7fef19183123ec3bae64ba265ef4203bfc64b68

Observation 144f0371-1b6e-452a-9c2e-22af451bcb4b · outbound

This paper cites Semantic integrity matters: Benchmarking and preserving high-density reasoning in KV cache compression.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Semantic integrity matters: Benchmarking and preserving high-density reasoning in KV cache compression

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.327931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:86b0203dd471d3b9861ac9a50d5ff3d497a2d4f3924bcc5221ad0ecf544a26ac

Observation e96c7aaf-e00f-42ae-aac4-f8733be98775 · outbound

This paper cites Semantic Integrity Matters: Benchmarking and Preserving High-Density Reasoning in KV Cache Compression.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Semantic Integrity Matters: Benchmarking and Preserving High-Density Reasoning in KV Cache Compression

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-05-13T05:27:19.414579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:69d14a4197a1681d04b5b572b0b2de45a7343fd5d29d33d1022f69d59ad04eb4

Observation b4a0c4d9-93e4-47e7-bd0b-efabd9260bc0 · outbound

This paper cites SONIC: Segmented optimized nexus for information compression in key-value caching.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production SONIC: Segmented optimized nexus for information compression in key-value caching

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:27:19.411568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:8c359b1d1ab6ad312a0d8e5ffe461cf6180c08b91d8fbd7c8cf0612aa02d02fc

Observation d4b9d3f2-7d38-4e1d-a234-191f236f2afb · outbound

This paper cites AnTKV: Anchor token-aware sub-bit vector quantization for KV cache in large language models.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production AnTKV: Anchor token-aware sub-bit vector quantization for KV cache in large language models

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:27:19.382778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:611514763c29a9a104be3ee630cef2ad0b1bf9773dad3b2e1d49090a638c4c79

Observation b4966c3c-fa63-459c-920f-018492c9da66 · outbound

This paper cites Ora- cleKV: Oracle guidance for question-independent KV cache eviction.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Ora- cleKV: Oracle guidance for question-independent KV cache eviction

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.286144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:b86046ee4e7694962d74f7146ead2a5a95e5354b2c65f1d20df2f566c3b5cd12

Observation 7da1e05c-8151-459e-b767-4162b988f6b2 · outbound

This paper cites FlexGen: High-Throughput Generative Inference of Large Language Models with a Single GPU.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production FlexGen: High-Throughput Generative Inference of Large Language Models with a Single GPU

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:27:19.393998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:a73641d246e2737d64e735da6cb5a1fc4b53a0df925ba81911eac9a1f7f76428

Observation f588b8c9-2806-4213-be57-769b4169f552 · outbound

This paper cites MiniMax-01: Scaling Foundation Models with Lightning Attention.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production MiniMax-01: Scaling Foundation Models with Lightning Attention

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-16T06:26:38.921226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:515d8539013618498eccd419f10f2ecfe359923f24dd19e7bca7136fc257a6b3

Observation dda5657e-3b70-4bed-80c1-a5d67aae9bfb · outbound

This paper cites DiffAdapt: Difficulty-Adaptive Reasoning for Token-Efficient LLM Inference.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production DiffAdapt: Difficulty-Adaptive Reasoning for Token-Efficient LLM Inference

Reference 63

Resolution
verified exact
local_arxiv, observed 2026-05-13T05:27:19.396706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:79d99b8cd83a83a5d26a4377bb02dd340eb262c3588d037250f4619b13155469

Observation 1a22711f-598c-4884-8ed0-da478d025ca1 · outbound

This paper cites Reasoning language model inference serving unveiled: An empirical study.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Reasoning language model inference serving unveiled: An empirical study

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:27:19.452420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:417de790ff99cb5320129277f6b8fa73e99734b7c769176aec279cee24019d79

Observation b1569d95-b7d3-46cd-a1f4-4c604567c0b5 · outbound

This paper cites Can Compressed LLMs Truly Act? An Empirical Evaluation of Agentic Capabilities in LLM Compression.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Can Compressed LLMs Truly Act? An Empirical Evaluation of Agentic Capabilities in LLM Compression

Reference 65

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T05:27:19.388471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:e9d682ba1e0cb9a83c1969146df30bbeddef46fdeb2f9d4ce64ee9dc8a8ca548

Observation f675fd1b-1e3a-492e-b43d-f020615ca2fd · outbound

This paper cites The Lottery LLM Hypothesis, Rethinking What Abilities Should LLM Compression Preserve?.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production The Lottery LLM Hypothesis, Rethinking What Abilities Should LLM Compression Preserve?

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:27:19.449599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:1b656e95539e10a232421fa3144e7f114b9f879511335f77ea13e6168b2becfe

Observation 9ea6e8d6-eb6d-4ad0-aed5-d423ab23e1a9 · outbound

This paper cites National energy administration releases 2024 national electric power industry statistics.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production National energy administration releases 2024 national electric power industry statistics

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:27:19.446585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:83c27a7ab6bebe4219800351e5d239474d6519261886dcd8917fa2a4e41c3592

Observation 825fa77b-7f25-41f0-9330-dcdc7a6a920e · outbound

This paper cites Interpretation of the work plan for stabilizing growth in the power equipment industry (2025–2026).

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Interpretation of the work plan for stabilizing growth in the power equipment industry (2025–2026)

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.287945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:f7962da1106309db61385d8750c04932572d4d68d57fb58b66f76cf1aeea6ee8

Observation 81faf1a9-83b0-47b2-8e8b-48b59fb2958b · outbound

This paper cites Launch of second data centre – call for application.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Launch of second data centre – call for application

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.295654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:fc9bd833f50ee88ba1f6bac864c32bf3c12c7d8724d6cc88419e08724b82da21

Observation 7e899b80-7d62-45b7-8e4c-b2c713540033 · outbound

This paper cites State of ai: token-usage rankings, q1 2026.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production State of ai: token-usage rankings, q1 2026

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.297414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:f9eb4b119d0381c21675de9d539655cd41ac0b6a70707d92ac7f7986ed767dc1

Observation 036b269c-7875-4501-8d4a-595be0548bff · outbound

This paper cites 2025 ai index report: Ai model performance gaps narrowing, compute costs plummeting.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production 2025 ai index report: Ai model performance gaps narrowing, compute costs plummeting

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.303034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:54b487164e48973fb8d9798e8443ea0e7a027bba88706b34217925ba6ae1b91f

Observation 2441701d-2c7f-4d98-aa6a-46c32420ff9f · outbound

This paper cites The Computational Limits of Deep Learning.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production The Computational Limits of Deep Learning

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:27:19.408590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:755db8da698613f1aacf886d3061b442bc30962e9aa08a147471ee52d5cc240b

Observation c92b5f69-69d9-45d5-a68b-90612a8f6dc0 · outbound

This paper cites an unresolved cited work.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Unresolved cited work

Reference 73

Resolution
unresolved
raw_fallback, observed 2026-05-13T10:42:38.281654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:785ecf3985a6960d84a1f31cf4c221a6400e5638e53c412b5393b212a3599539

Observation f1574f00-b03b-4bea-83b9-52e8b12f229f · outbound

This paper cites an unresolved cited work.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Unresolved cited work

Reference 74

Resolution
verified exact
doi, observed 2026-05-13T05:22:18.948887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:2c13046721214e6bda940ee884a430d9ee75e28c92e3f2efb8666baff7d45f24

Observation 45d5564e-0640-4913-88ed-e0f3fe9f68d1 · outbound

This paper cites Appenzeller.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Appenzeller

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.279994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:9873ca15f4c55834fa9912c58aa4a6ed47b35322b0b5fda9b33df6b48f8cccba

Observation 9b62728c-52c2-4fb1-bab5-b8800a554e29 · outbound

This paper cites Demirer, A.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Demirer, A

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.274692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:276710b0556d7aba60a36b1b9f13e3741452a80ea3b3b65f33fe8e9f28ad1116

Observation a91901e4-55e7-4dd5-b0ea-fc2878c00f4d · outbound

This paper cites NBER Working Paper No.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production NBER Working Paper No

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.278276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:887376ad1ea44194a963df1d095e335297f5b81f042eb8348bdec8cf2c0d829d

Observation b9b728c0-b2d7-4a30-a405-19788b75a303 · outbound

This paper cites High cost of energy: industrial electricity prices in the eu vs the us and china.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production High cost of energy: industrial electricity prices in the eu vs the us and china

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.271272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:5c05b2e61d8666ebd117b35f9d5bcc8494a17c3689fcec56085b16206c66f235

Observation 0b25e718-7e18-489b-8736-f06b0272f5ca · outbound

This paper cites Shapiro and H.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Shapiro and H

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.269493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:50b9ac64943d73b20ee7d2c489b398702380eee19472ac9f5f4beacfe08f04ea

Observation 2e92250a-11c9-4675-af0c-0d2b6e083be8 · outbound

This paper cites Energy Information Administration.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Energy Information Administration

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.329812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:d296978c11aac4113dbec4301b8d904684546f2f15b516d644c113522643d231

Observation 7af78cb9-78c5-4bc4-9f5e-3e1fc3bce3b1 · outbound

This paper cites an unresolved cited work.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Unresolved cited work

Reference 81

Resolution
unresolved
raw_fallback, observed 2026-05-13T10:42:38.272931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:4c2d61e5f3b5d0dedd190631aa129e27409285362f8db71b546ac53d1f9b3c83

Observation d814c45a-47fa-4772-96d6-058b6c4d1100 · outbound

This paper cites Will we run out of data? Limits of LLM scaling based on human-generated data.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Will we run out of data? Limits of LLM scaling based on human-generated data

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:27:19.456297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:2a19dd6fb497ad5bff0f63a6bf26d28c8140405ed0ca9cfc3dcf4ce151e87550

Observation 05cba738-48f0-475c-a26a-2ce8ea5c1a3f · outbound

This paper cites Xiaomi mimo api open platform.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Xiaomi mimo api open platform

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.276397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:337a17a6a0b9063d554093caf059a6c72511e49946f08ef4d268fe89db374613

Observation 3c766683-cab9-4fb3-8f3e-daa4f1a2e932 · outbound

This paper cites Z.ai developer documentation: pricing overview.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Z.ai developer documentation: pricing overview

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.284278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:21b9ba40ac0b6649726fd956f9923f80583999d1d6bf7c998139314416b4d273

Observation 06226aa9-3c97-43a9-ad4b-257d88340173 · outbound

This paper cites Kimi api platform: model inference pricing.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Kimi api platform: model inference pricing

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.324370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:751d429e4c7324a15b32c041fb2e68862e25ef2a3ff347ab0a874d6fbbab9cbf

Observation 3ef06c40-6fd2-4cc1-a3f1-cb4dcbbf4e2d · outbound

This paper cites how cheaply can we generate tokens?.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production how cheaply can we generate tokens?

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T10:42:38.313488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:d8d7e11887b74ae6630a30dea3f1494907af6118855e0da1b6a720b91f0b493d

Observation 3b2a94ca-a9fa-48fb-a1b4-4ae46ae886c2 · outbound

This paper cites an unresolved cited work.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Unresolved cited work

Reference 87

Resolution
unresolved
raw_fallback, observed 2026-05-13T10:42:38.299124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:7c4c63a2da1004da2791f95e96dec513aec6d964987f2cfc6c4f40c8990c8832

Observation 5a04bb42-8e1e-4c0b-a289-17aa1b982c89 · outbound

This paper cites an unresolved cited work.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Unresolved cited work

Reference 88

Resolution
unresolved
raw_fallback, observed 2026-05-13T10:42:38.322063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:165037287f5e9baecf860b84ae0aea7bd8ef1dbb3f177cb39c58dbb229e44a8e

Observation 9d290c00-853b-4560-98a3-75d85a14933f · outbound

This paper cites an unresolved cited work.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Unresolved cited work

Reference 89

Resolution
unresolved
raw_fallback, observed 2026-05-13T10:42:38.336780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:6327441c98c42aa8708177799f30de6423b92d622fffc59669fa271b90f4dd16

Observation 7c383df8-8411-4b3a-bb17-48d9154860b4 · outbound

This paper cites an unresolved cited work.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Unresolved cited work

Reference 90

Resolution
unresolved
raw_fallback, observed 2026-05-13T10:42:38.333218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:2585cf48ee58783448c97081fce94d8ef5650882d02130100d0a410031cb9955

Observation f1958883-a74c-4df5-960d-938f8af83ef7 · outbound

This paper cites an unresolved cited work.

Position: LLM Inference Should Be Evaluated as Energy-to-Token Production Unresolved cited work

Reference 91

Resolution
unresolved
raw_fallback, observed 2026-05-13T10:42:38.331595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:17:24.147248Z digest=sha256:6e107a11528ca61fe37ce786705e82c6a456bf224dfbc562ae367492c40a2f6f

Pith citing papers

No inbound Pith citation observations are available.