Pith. sign in

Paper Citation Record · LEDGER

8-bit Optimizers via Block-wise Quantization

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 65 inbound Pith citation observations for arXiv:2110.02861.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2110.02861 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 65 of 65 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:30:56.666182Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

4
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ddd28ee3-33c2-46e9-9363-080b227bd50a · inbound

Rethinking the Role of Demonstrations: What Makes In-Context Learning Work? cites this paper.

Rethinking the Role of Demonstrations: What Makes In-Context Learning Work? 8-bit Optimizers via Block-wise Quantization

Reference 167

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T09:51:46.844616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-15T09:51:46.701149Z digest=sha256:ef93dc9e5b1d1c60dea10041f7193f737df27115887cb72878eac51f0209a0ee

Observation 5d175240-32ff-4a1c-b08a-03049c74372f · inbound

BitMoD: Bit-serial Mixture-of-Datatype LLM Acceleration cites this paper.

BitMoD: Bit-serial Mixture-of-Datatype LLM Acceleration 8-bit Optimizers via Block-wise Quantization

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T18:18:05.352436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:18:05.352436Z digest=sha256:cd0448cd444feabd8d07e0048a92603e056a124807675c1b4e1e44c6816ff4da

Observation 30954bc3-704c-464c-a27a-5f1c0e75e42a · inbound

COAP: Memory-Efficient Training with Correlation-Aware Gradient Projection cites this paper.

COAP: Memory-Efficient Training with Correlation-Aware Gradient Projection 8-bit Optimizers via Block-wise Quantization

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T12:41:58.306850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:41:58.306850Z digest=sha256:9cf14eb769ad6577be7df80d3278e1bbfdb5a5edba5af2ebfd4f879bca796f32

Observation bf204469-eab7-4e1c-b3af-16f1d15f1dfe · inbound

AdaScale: Dynamic Context-aware DNN Scaling via Automated Adaptation Loop on Mobile Devices cites this paper.

AdaScale: Dynamic Context-aware DNN Scaling via Automated Adaptation Loop on Mobile Devices 8-bit Optimizers via Block-wise Quantization

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T05:09:32.477647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:09:32.477647Z digest=sha256:1409056188b112cb0d43ad7be134a9b9b4c275ba6893287cdd1aa62189b917ab

Observation 3e3f046d-c067-4e0b-b608-6ce0b35fb1f9 · inbound

ControlFace: Harnessing Facial Parametric Control for Face Rigging cites this paper.

ControlFace: Harnessing Facial Parametric Control for Face Rigging 8-bit Optimizers via Block-wise Quantization

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T04:43:10.381217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:43:10.381217Z digest=sha256:458a2604dc8a2dbb4a7b892eff65c9e267c75a3b476f30690abe6ed8c9bd8f30

Observation 6c2eba9e-4b83-4c80-a650-9a46e8451219 · inbound

VibrantVS: A high-resolution multi-task transformer for forest canopy height estimation cites this paper.

VibrantVS: A high-resolution multi-task transformer for forest canopy height estimation 8-bit Optimizers via Block-wise Quantization

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T16:10:26.477257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:10:26.477257Z digest=sha256:d6126dea0ab6de9a5ad164e8569a9ebfbc886175b111f4faf5c49132c1d29d67

Observation c3a1db67-6fec-43e1-bcb1-1f138d582f8d · inbound

Memory-Efficient 4-bit Preconditioned Stochastic Optimization cites this paper.

Memory-Efficient 4-bit Preconditioned Stochastic Optimization 8-bit Optimizers via Block-wise Quantization

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T15:50:24.116174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:50:24.116174Z digest=sha256:c75da2a6f064068399c9cc4014b74e97fe2f6c40b4b2c693207183c609286dfc

Observation 4dd42e93-584e-4f38-a86c-57b19a161d3b · inbound

No More Adam: Learning Rate Scaling at Initialization is All You Need cites this paper.

No More Adam: Learning Rate Scaling at Initialization is All You Need 8-bit Optimizers via Block-wise Quantization

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T14:41:00.210604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:41:00.210604Z digest=sha256:3947b85e640f7d3a48c051b5740ccd899d70590204fab1f367a3563476ec2547

Observation d87533f9-a5dd-4855-a7ff-dbd7526217bb · inbound

Fine-tuning Whisper on Low-Resource Languages for Real-World Applications cites this paper.

Fine-tuning Whisper on Low-Resource Languages for Real-World Applications 8-bit Optimizers via Block-wise Quantization

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T11:12:55.515201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:12:55.515201Z digest=sha256:f9984c96c801542cdfa9950c2331a099d6236d77466ead34ce66e56db4dac5b9

Observation 7640b6f2-e1ce-428e-9341-5263c77915d1 · inbound

Gradient Weight-normalized Low-rank Projection for Efficient LLM Training cites this paper.

Gradient Weight-normalized Low-rank Projection for Efficient LLM Training 8-bit Optimizers via Block-wise Quantization

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T00:15:00.234867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:15:00.234867Z digest=sha256:9d0c6a745de978fff872f73492e3ca100f4d307f2a2e58de43357a18b54f1245

Observation a4623ac3-e50c-431e-8457-f0b7523e542d · inbound

Scaling Large Language Model Training on Frontier with Low-Bandwidth Partitioning cites this paper.

Scaling Large Language Model Training on Frontier with Low-Bandwidth Partitioning 8-bit Optimizers via Block-wise Quantization

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T21:42:38.187634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:42:38.187634Z digest=sha256:6a5fdf3f55b0f248ec3cd51fce567f216bf0bdae71b2f924d343337ca82401e1

Observation 563f671c-e81c-4388-8c1c-1abfaf7093da · inbound

Affordably Fine-tuned LLMs Provide Better Answers to Course-specific MCQs cites this paper.

Affordably Fine-tuned LLMs Provide Better Answers to Course-specific MCQs 8-bit Optimizers via Block-wise Quantization

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T21:10:10.889603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:10:10.889603Z digest=sha256:a0e39fdadfeab4882ec72f692e85031bcadbcdc8857c2db42f0bc91056b3bf17

Observation e30bc0c9-edc9-471f-b566-125d8d1c27c8 · inbound

SPAM: Spike-Aware Adam with Momentum Reset for Stable LLM Training cites this paper.

SPAM: Spike-Aware Adam with Momentum Reset for Stable LLM Training 8-bit Optimizers via Block-wise Quantization

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-10T20:56:25.828161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:56:25.828161Z digest=sha256:d62bc5e140d37a66a4e392c07a4afeca42bfbaf0878b13da331aaacec4b76c10

Observation 80ea1a75-f208-4640-a79c-4e6dd5496c74 · inbound

GWT: Scalable Optimizer State Compression for Large Language Model Training cites this paper.

GWT: Scalable Optimizer State Compression for Large Language Model Training 8-bit Optimizers via Block-wise Quantization

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-23T05:57:36.672293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-23T05:57:09.276224Z digest=sha256:9541f0331e1bcb14cfa762c0661a97862e5bd2914a26e0bb5d216435a2f15536

Observation e8dc9f2c-94ec-4702-896d-390e59653968 · inbound

Atleus: Accelerating Transformers on the Edge Enabled by 3D Heterogeneous Manycore Architectures cites this paper.

Atleus: Accelerating Transformers on the Edge Enabled by 3D Heterogeneous Manycore Architectures 8-bit Optimizers via Block-wise Quantization

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T19:57:25.671736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:57:25.671736Z digest=sha256:6e79c33950347c374e546188eedbc2b44f5f0fbab1b0c2b9464e785c3e716535

Observation c80adc21-bcf1-4a73-b946-aa352a7a2fa0 · inbound

Irrational Complex Rotations Empower Low-bit Optimizers cites this paper.

Irrational Complex Rotations Empower Low-bit Optimizers 8-bit Optimizers via Block-wise Quantization

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T16:49:47.067643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T16:49:47.067643Z digest=sha256:cb59308268cfcdd98e60796cec76613c39e672384972ca2534f258cfef9956ca

Observation 1894d2d6-2b8b-47ac-a1fb-823bbad37bbc · inbound

Fine-Tuned Language Models as Space Systems Controllers cites this paper.

Fine-Tuned Language Models as Space Systems Controllers 8-bit Optimizers via Block-wise Quantization

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T12:10:35.679586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T12:10:35.679586Z digest=sha256:adf5a0bf09769d332505aec35d9be5c61e05353bdf2a6e2d64042716521924bc

Observation 10c0fdfc-4e48-4586-bc29-41a1ab6ee0f0 · inbound

Query-Aware Learnable Graph Pooling Tokens as Prompt for Large Language Models cites this paper.

Query-Aware Learnable Graph Pooling Tokens as Prompt for Large Language Models 8-bit Optimizers via Block-wise Quantization

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-23T04:37:32.089763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-23T04:36:06.247071Z digest=sha256:392678562188f81c80d655019ebc2c9402df861c793cf5904557712cbd5f44d7

Observation 74641c09-8f2c-4b8e-9cdc-e80004964567 · inbound

SANA 1.5: Efficient Scaling of Training-Time and Inference-Time Compute in Linear Diffusion Transformer cites this paper.

SANA 1.5: Efficient Scaling of Training-Time and Inference-Time Compute in Linear Diffusion Transformer 8-bit Optimizers via Block-wise Quantization

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-09T23:37:12.518282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T23:37:12.518282Z digest=sha256:3daa9e044b1d5fefa8aea2ea75a2b185bc85522193bfd6e9b8ebd1fe3ec4cdda

Observation 68d31d00-a39f-45f7-9514-84cd3ca516b4 · inbound

Optimization Strategies for Enhancing Resource Efficiency in Transformers & Large Language Models cites this paper.

Optimization Strategies for Enhancing Resource Efficiency in Transformers & Large Language Models 8-bit Optimizers via Block-wise Quantization

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T20:10:20.323263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:10:20.323263Z digest=sha256:423629f772ca92f074ba82c5697d6db0314fc346619220f67a15afffa732b77f

Observation 06501bb7-f57a-4160-87d8-217e03c47c00 · inbound

Resource-Efficient & Effective Code Summarization cites this paper.

Resource-Efficient & Effective Code Summarization 8-bit Optimizers via Block-wise Quantization

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-09T04:23:34.861574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:23:34.861574Z digest=sha256:573c5ea064b7c3efdb8efbaaee1c1d73461573b2cdab330b1ab0967bc17443c3

Observation b6f0a2ef-2891-483c-8653-7110ecacb8e6 · inbound

SSH: Sparse Spectrum Adaptation via Discrete Hartley Transformation cites this paper.

SSH: Sparse Spectrum Adaptation via Discrete Hartley Transformation 8-bit Optimizers via Block-wise Quantization

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T19:01:09.214424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T19:01:09.214424Z digest=sha256:b02e0e6f92b14a10d058459f44a6873295dcc26e104566adc054e30818bd24f3

Observation 056f036f-95e6-449f-a3f5-d5081f4ff537 · inbound

Low-Precision Training of Large Language Models: Methods, Challenges, and Opportunities cites this paper.

Low-Precision Training of Large Language Models: Methods, Challenges, and Opportunities 8-bit Optimizers via Block-wise Quantization

Reference 168

Resolution
unresolved
no resolver link, observed 2026-08-16T04:30:56.666182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:30:56.666182Z digest=sha256:66bfb7901d648bd425bee21d400e54aae9354ab63cbdf1119686c84b35f6d957

Observation ce373dd9-84dc-4846-b661-1976e0f98472 · inbound

Symbol-based entity marker highlighting for enhanced text mining in materials science with generative AI cites this paper.

Symbol-based entity marker highlighting for enhanced text mining in materials science with generative AI 8-bit Optimizers via Block-wise Quantization

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-15T22:57:02.114377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:57:02.114377Z digest=sha256:db0b8211b456e6bbaf73865bde742eda52223c2365bd8c661fd17639d55dc3b9

Observation e64fa1bc-0016-42d3-8344-8e797838f702 · inbound

Quaff: Quantized Parameter-Efficient Fine-Tuning under Outlier Spatial Stability Hypothesis cites this paper.

Quaff: Quantized Parameter-Efficient Fine-Tuning under Outlier Spatial Stability Hypothesis 8-bit Optimizers via Block-wise Quantization

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:37.149873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:37.149873Z digest=sha256:63edeba2b8c2ea602b3701cbd8dc39a0a4a41db5e39c4b5da4680bc7eea7f761

Observation 1d85fe88-6af5-4a02-a9ad-a0323c9912d8 · inbound

Comparative Evaluation of Prompting and Fine-Tuning for Applying Large Language Models to Grid-Structured Geospatial Data cites this paper.

Comparative Evaluation of Prompting and Fine-Tuning for Applying Large Language Models to Grid-Structured Geospatial Data 8-bit Optimizers via Block-wise Quantization

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:00.585372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:16:00.585372Z digest=sha256:173568af67d76f19a33c2fc1fb5c586653168f517fd28013be33f84ddce8292c

Observation f320d63e-c91c-46e4-9a45-535ee198c85f · inbound

Subspace Networks: Scaling Decentralized Training with Communication-Efficient Model Parallelism cites this paper.

Subspace Networks: Scaling Decentralized Training with Communication-Efficient Model Parallelism 8-bit Optimizers via Block-wise Quantization

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:55:13.250462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:55:13.250462Z digest=sha256:feff81e31db48f958083adf95d365c8d78ca6fcf9b24523d08be8a3f9015f6e8

Observation 46ab0f85-5288-406a-9555-ca4a273f1bd3 · inbound

A MISMATCHED Benchmark for Scientific Natural Language Inference cites this paper.

A MISMATCHED Benchmark for Scientific Natural Language Inference 8-bit Optimizers via Block-wise Quantization

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T10:42:23.400254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:42:23.400254Z digest=sha256:7abd3b2a362b1acabe923bbf2a1849389aa9a9290a9e8afb51a58ef04335ad05

Observation f082671f-8362-49d4-9d86-5122c2bd84f8 · inbound

Slimming Down LLMs Without Losing Their Minds cites this paper.

Slimming Down LLMs Without Losing Their Minds 8-bit Optimizers via Block-wise Quantization

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:41.968247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:19:41.968247Z digest=sha256:38d9d65a400a3cf780fefafeae59442df13dfba3373a5158878a2b75e7c9fedb

Observation 6762342b-de94-4203-8733-6fd2d808e4e6 · inbound

SlimMoE: Structured Compression of Large MoE Models via Expert Slimming and Distillation cites this paper.

SlimMoE: Structured Compression of Large MoE Models via Expert Slimming and Distillation 8-bit Optimizers via Block-wise Quantization

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T18:58:33.787850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:58:33.787850Z digest=sha256:af1b66cfa35b18ad1be6050ec2004ecc5c387a8c42cc98c3be469c4215885a03

Observation 37add5ea-1f3f-4a0b-8e0b-d16bbe160b7b · inbound

Low-rank Momentum Factorization for Memory Efficient Training cites this paper.

Low-rank Momentum Factorization for Memory Efficient Training 8-bit Optimizers via Block-wise Quantization

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T18:35:40.111493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:35:40.111493Z digest=sha256:a1905abba6a5b9a87c5fcb868eb4ef8836371173a9132572a3ca07e65610db33

Observation 2c8cc62a-b52d-42d7-8929-54e27d22a9fa · inbound

Is Quantization a Deal-breaker? Empirical Insights from Large Code Models cites this paper.

Is Quantization a Deal-breaker? Empirical Insights from Large Code Models 8-bit Optimizers via Block-wise Quantization

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T17:56:21.261233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:56:21.261233Z digest=sha256:2ac5c63a909b700fa786f3458ba023d5a985c870d053ca20afbc5d8fb3aa1bb8

Observation 19effd8d-d214-47d5-95bf-bc6f0dd197a3 · inbound

Droplet3D: Commonsense Priors from Videos Facilitate 3D Generation cites this paper.

Droplet3D: Commonsense Priors from Videos Facilitate 3D Generation 8-bit Optimizers via Block-wise Quantization

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T15:10:34.440764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:10:34.440764Z digest=sha256:a563e2487255bc3865bbe85fa9eca41ceb3cdc65e7ef0686c1da79842405bd6b

Observation ebfa13f1-ea5b-4b2e-b086-13096f9b453b · inbound

Quantized Large Language Models in Biomedical Natural Language Processing: Evaluation and Recommendation cites this paper.

Quantized Large Language Models in Biomedical Natural Language Processing: Evaluation and Recommendation 8-bit Optimizers via Block-wise Quantization

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-05T10:38:35.001738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:38:35.001738Z digest=sha256:3f38f36bb47119e1ccda92855a1747ea04cf807f0eb6d161c2265c2e4bce885c

Observation b1810150-1745-434f-8810-5bcf97bc4d58 · inbound

NeuronMLP: Efficient LLM Inference via Singular Value Decomposition Compression and Tiling on AWS Trainium cites this paper.

NeuronMLP: Efficient LLM Inference via Singular Value Decomposition Compression and Tiling on AWS Trainium 8-bit Optimizers via Block-wise Quantization

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T02:52:21.861968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T02:51:19.275111Z digest=sha256:a7e82a807a04a6fd4f721e4a3d493c7ea41269efb10234ca53200997c3e80b0d

Observation 7ed4500e-3886-4ab4-98d1-7d5c1c067285 · inbound

Fine-Tuning Causal LLMs for Text Classification: Embedding-Based vs. Instruction-Based Approaches cites this paper.

Fine-Tuning Causal LLMs for Text Classification: Embedding-Based vs. Instruction-Based Approaches 8-bit Optimizers via Block-wise Quantization

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-25T07:25:29.063199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-25T07:23:53.703286Z digest=sha256:b7f4eee2c5f0b6ec314efb62b6780bdb012801f4fc73ce2ee4f5862bcd90124f

Observation d63b4287-f6aa-4180-bd21-fae2a57d8e46 · inbound

Fine-Tuning Causal LLMs for Text Classification: Embedding-Based vs. Instruction-Based Approaches cites this paper.

Fine-Tuning Causal LLMs for Text Classification: Embedding-Based vs. Instruction-Based Approaches 8-bit Optimizers via Block-wise Quantization

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T16:38:13.326091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:38:13.326091Z digest=sha256:094d835d9a2599259224365f5e25c8c92e329f23c6da4be0a1540b514365bf81

Observation f61e057b-ad46-4ba0-92f0-230c6f43287b · inbound

When Less is More: 8-bit Quantization Improves Continual Learning in Large Language Models cites this paper.

When Less is More: 8-bit Quantization Improves Continual Learning in Large Language Models 8-bit Optimizers via Block-wise Quantization

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T14:52:19.865798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T14:52:19.865798Z digest=sha256:6bc53f0c945178f5534b52a8aca330818dadae60f4d6f8ef2dbbaa33afe82fa0

Observation 6566ab8f-907b-452e-8a8d-5848c03cdb55 · inbound

AdaFRUGAL: Adaptive Memory-Efficient Training with Dynamic Control cites this paper.

AdaFRUGAL: Adaptive Memory-Efficient Training with Dynamic Control 8-bit Optimizers via Block-wise Quantization

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T19:18:19.272518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-16T19:13:30.177633Z digest=sha256:f95c588037014b6ec77f4a97919a0f097a0b36396a8a5f0acbec1c354ebbbba8

Observation 45df5dfe-cff6-47ea-8b13-d22114a12d93 · inbound

HeteroCache: A Dynamic Retrieval Approach to Heterogeneous KV Cache Compression for Long-Context LLM Inference cites this paper.

HeteroCache: A Dynamic Retrieval Approach to Heterogeneous KV Cache Compression for Long-Context LLM Inference 8-bit Optimizers via Block-wise Quantization

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T13:17:54.857068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-16T13:16:31.568604Z digest=sha256:b0bd37c5e1d03be220d2dfd4aa2c61e726aa4a425c605642b32961819fadd98c

Observation bcd8a09d-2f6c-48b1-9b12-d1c0a94e7295 · inbound

Sustainability Is Not Linear: Quantifying Performance, Energy, and Privacy Trade-offs in On-Device Intelligence cites this paper.

Sustainability Is Not Linear: Quantifying Performance, Energy, and Privacy Trade-offs in On-Device Intelligence 8-bit Optimizers via Block-wise Quantization

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T09:29:56.963343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-21T09:27:34.204972Z digest=sha256:612254607f3b168df93badaa48d3f0cee7ba98d617e79f8e765c2d5acc14f961

Observation b35a8172-04e3-4064-94fa-f70c1bbedace · inbound

STQuant: Spatio-Temporal Adaptive Framework for Optimizer Quantization in Large Multimodal Model Training cites this paper.

STQuant: Spatio-Temporal Adaptive Framework for Optimizer Quantization in Large Multimodal Model Training 8-bit Optimizers via Block-wise Quantization

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:55:49.370629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T18:49:40.758234Z digest=sha256:0cb6292a6dc0ca659ffc1a7125f5e1c68bb68cfce7f31c5ac0c3b435cba155e2

Observation d071ecff-fe92-4659-b28a-fbb1d5822426 · inbound

GSQ: Highly-Accurate Low-Precision Scalar Quantization for LLMs via Gumbel-Softmax Sampling cites this paper.

GSQ: Highly-Accurate Low-Precision Scalar Quantization for LLMs via Gumbel-Softmax Sampling 8-bit Optimizers via Block-wise Quantization

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T05:36:02.321113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T05:29:51.182114Z digest=sha256:01b9ec0857353eb93777af6242e7a2c4726a67d6bce6f02487ef82a9c5334f30

Observation 6faadd91-7601-44f0-92b8-aad68f6d1ee2 · inbound

GSQ: Highly-Accurate Low-Precision Scalar Quantization for LLMs via Gumbel-Softmax Sampling cites this paper.

GSQ: Highly-Accurate Low-Precision Scalar Quantization for LLMs via Gumbel-Softmax Sampling 8-bit Optimizers via Block-wise Quantization

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T18:02:42.154497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-19T18:01:08.514022Z digest=sha256:47b34efc065ec0ff49c85c35cac024b5ca94e22d8ec5b0c6a031642876644a0c

Observation 2c8c744a-9be6-4545-8a47-d364c764bd8c · inbound

Hybrid JIT-CUDA Graph Optimization for Low-Latency Large Language Model Inference cites this paper.

Hybrid JIT-CUDA Graph Optimization for Low-Latency Large Language Model Inference 8-bit Optimizers via Block-wise Quantization

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T20:41:14.307289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-08T08:13:51.176435Z digest=sha256:8d3834610da2b68b2e0a1be7d9d8388520731ca5bd21034a899c6d43eb08e4e2

Observation f4cdca61-a59f-4ba4-b976-b4889318014b · inbound

Cloud to Edge: Benchmarking LLM Inference On Hardware-Accelerated Single-Board Computers cites this paper.

Cloud to Edge: Benchmarking LLM Inference On Hardware-Accelerated Single-Board Computers 8-bit Optimizers via Block-wise Quantization

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T20:21:12.025263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-08T09:22:05.851456Z digest=sha256:650370498b6b1301b042629f70127c494cfd7ae7cfa5b0802c80f48ab4d161ce

Observation bf3171ca-d31b-4492-843c-437b112ec8b7 · inbound

Cloud to Edge: Benchmarking LLM Inference On Hardware-Accelerated Single-Board Computers cites this paper.

Cloud to Edge: Benchmarking LLM Inference On Hardware-Accelerated Single-Board Computers 8-bit Optimizers via Block-wise Quantization

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T05:34:06.504733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:34:06.504733Z digest=sha256:bb830ba3f4eb4cdcc21a3fc6ad72d7e888a867184a6d4327e92dc5827a5a2799

Observation ae91f5d5-1d90-4616-9f9c-86e39e8b2b4f · inbound

Salca: A Sparsity-Aware Hardware Accelerator for Efficient Long-Context Attention Decoding cites this paper.

Salca: A Sparsity-Aware Hardware Accelerator for Efficient Long-Context Attention Decoding 8-bit Optimizers via Block-wise Quantization

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:11:18.879146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-07T17:56:39.124969Z digest=sha256:051dfef1fc71a2a5081ce139d676da4c50b89704df1c0962042ee3975e6e4f29

Observation 009929a1-0701-4642-b55f-4e2b9d833f13 · inbound

Maistros: A Greek Large Language Model Adapted Through Knowledge Distillation From Large Reasoning Models cites this paper.

Maistros: A Greek Large Language Model Adapted Through Knowledge Distillation From Large Reasoning Models 8-bit Optimizers via Block-wise Quantization

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:00:37.376768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-08T19:08:09.259620Z digest=sha256:27d4c505b33dc38115d5b7b9ddcfa0974b9d9048a6ed3a61a3c650e7f10d5bd7

Observation 6eccfed7-6512-4a67-81f8-5f484d422187 · inbound

Revealing Modular Gradient Noise Imbalance in LLMs: Calibrating Adam via Signal-to-Noise Ratio cites this paper.

Revealing Modular Gradient Noise Imbalance in LLMs: Calibrating Adam via Signal-to-Noise Ratio 8-bit Optimizers via Block-wise Quantization

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T16:41:05.506387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-09T15:39:51.611115Z digest=sha256:f83f3fed4fcfde54ed8ee3daeb5f8618077f9a7181fd27626f79540c3a2cec62

Observation 45819efa-d7c1-4c4e-b210-3528816a4eb4 · inbound

Q-LocalAdam: Memory-Efficient Client-Side Adaptive Optimization for Edge Federated Learning cites this paper.

Q-LocalAdam: Memory-Efficient Client-Side Adaptive Optimization for Edge Federated Learning 8-bit Optimizers via Block-wise Quantization

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:13:21.230707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-20T14:11:53.371521Z digest=sha256:48de37d8cfd757a26dc7345b9242570b375c9825c4b72ec5c61bc6fb975b58da

Observation 4fcba71d-9a04-47dc-b5a5-b08b92e64a1c · inbound

Bounded-Compute Multimodal Regression for Product-Rating Prediction cites this paper.

Bounded-Compute Multimodal Regression for Product-Rating Prediction 8-bit Optimizers via Block-wise Quantization

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T18:03:48.009935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T17:58:58.353336Z digest=sha256:780e2e274a592152feef13eea9c4137c10db87ba0e001915b24342e74fb728e5

Observation 7efc0b72-f121-45fc-96f7-7b3f49829d79 · inbound

GNMR: Runtime Stability Control for Low-Precision Large Language Model Training cites this paper.

GNMR: Runtime Stability Control for Low-Precision Large Language Model Training 8-bit Optimizers via Block-wise Quantization

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T19:42:36.064792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T18:53:47.187437Z digest=sha256:e131f97950eaa16870153e18615aaa6cd361195f46e72b5bef2bc580340123b7

Observation 73323cac-f7b3-4759-85c9-0fa8bc9595ce · inbound

TadA-Bench: A Million-Variant Benchmark for Future-Round Discovery Toward Agentic Protein Engineering cites this paper.

TadA-Bench: A Million-Variant Benchmark for Future-Round Discovery Toward Agentic Protein Engineering 8-bit Optimizers via Block-wise Quantization

Reference 109

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T20:12:37.861872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-28T20:01:10.638647Z digest=sha256:2ff519240e98844116d0e2686bd6e77bcbfafa40a92b7f9a05b0f81b9514380b

Observation 33ab8464-441b-4381-a4b9-634d871b16c5 · inbound

Alignment Collapse Under KV Cache Quantization: Diagnosis and Mitigation cites this paper.

Alignment Collapse Under KV Cache Quantization: Diagnosis and Mitigation 8-bit Optimizers via Block-wise Quantization

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T22:16:16.038824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-28T15:37:34.129339Z digest=sha256:09c4035c94dfcb1c95d8436942466f453ff6df2a19ea0110fb097e1737ed9cc0

Observation 85af1874-7e65-4b84-a222-377e963a6480 · inbound

Gefen: Optimized Stochastic Optimizer cites this paper.

Gefen: Optimized Stochastic Optimizer 8-bit Optimizers via Block-wise Quantization

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-14T18:00:57.430970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T18:00:57.430970Z digest=sha256:06ab26069e0441c4e2ed9fec71c2e344223c9755327e3dfd5c2b808e933b0ba4

Observation c95ffa89-222d-4498-8419-fc0785e05f02 · inbound

Zero-order Parameter-free Optimization for LMO-based Methods: Novel Approach for Efficient Fine-tuning cites this paper.

Zero-order Parameter-free Optimization for LMO-based Methods: Novel Approach for Efficient Fine-tuning 8-bit Optimizers via Block-wise Quantization

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T17:08:43.652628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-27T04:33:10.554853Z digest=sha256:f4f272118f0e49f82220c3a03a5ff2105e0abae28b2c6ef713041251aa345766

Observation a8f6dc79-f9db-49f2-b0a0-efad0d2b543e · inbound

MaineCoon: Pursuing A Real-Time Audio-Visual Social World Model cites this paper.

MaineCoon: Pursuing A Real-Time Audio-Visual Social World Model 8-bit Optimizers via Block-wise Quantization

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T20:48:55.723997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-27T01:09:02.214590Z digest=sha256:6230409d3770bac2d300440ffdd8ec6854e4e48d19581a2107e3bd5c0d594aa7

Observation 1329faa3-a8a5-4a9c-8cb6-774a33d77948 · inbound

P-MTP: Efficient Document Parsing via Multi-Token Prediction with Progressive Depth Scaling cites this paper.

P-MTP: Efficient Document Parsing via Multi-Token Prediction with Progressive Depth Scaling 8-bit Optimizers via Block-wise Quantization

Reference 27

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T16:39:58.543364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-26T00:18:07.278889Z digest=sha256:56f827dd05177029cb53c1d74d1e7454a3d6d64bf48033587a6b963e5b40654a

Observation c8eb76c7-31e2-4905-8db5-7b2165347bce · inbound

OmniOpt: Taxonomy, Geometry, and Benchmarking of Modern Optimizers cites this paper.

OmniOpt: Taxonomy, Geometry, and Benchmarking of Modern Optimizers 8-bit Optimizers via Block-wise Quantization

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-11T22:10:49.683444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T22:10:49.683444Z digest=sha256:db4cc52b20c39aa4804e3349c16b6da170f8a34c7e1d46b23a4f27e5f412d36b

Observation 0814f2e2-e831-447a-8c90-bd2e633f1f36 · inbound

Full-Stack FP4: Stable LLM Pretraining with Quantized Projections, Optimizers, and Attention cites this paper.

Full-Stack FP4: Stable LLM Pretraining with Quantized Projections, Optimizers, and Attention 8-bit Optimizers via Block-wise Quantization

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-11T19:17:59.044982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:17:59.044982Z digest=sha256:e3c43c799b71bf945162b30e86c53c1225c68ed5f59353549f9dfb08de9dd3f6

Observation aec9d49b-5138-48d3-979d-83f9af323cdb · inbound

RAGAL: A Frugal, Fully Local Retrieval-Augmented Assistant for Technical Support at a Government Agency cites this paper.

RAGAL: A Frugal, Fully Local Retrieval-Augmented Assistant for Technical Support at a Government Agency 8-bit Optimizers via Block-wise Quantization

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T14:30:35.482308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:30:35.482308Z digest=sha256:edf9edc7292436d14d50ac17d24a7554fdc5b2eaae271536c6cf4107fc7faaa2

Observation 7dafb524-0f23-4066-8a28-05a8baa2d963 · inbound

Reasoning Before Translation: Enhancing Legal Machine Translation with Structured Reasoning cites this paper.

Reasoning Before Translation: Enhancing Legal Machine Translation with Structured Reasoning 8-bit Optimizers via Block-wise Quantization

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T13:15:34.905327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:15:34.905327Z digest=sha256:24ca81b824610820e0bb112ebf8abd038799398be171ae0b9ff91e1881d6a35d

Observation de6d4400-f87c-4cdc-ac9f-1dfc53ca0cea · inbound

FBLayout: Optimizing Memory Layout for Efficient LLM Finetuning on Mobile GPUs cites this paper.

FBLayout: Optimizing Memory Layout for Efficient LLM Finetuning on Mobile GPUs 8-bit Optimizers via Block-wise Quantization

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T08:23:07.462482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:23:07.462482Z digest=sha256:352d4223794aa6dd108f07c2eb1702eb2bca4a4ecd807674ab522497ac171620

Observation 3c39e27b-609a-4a3b-8538-d8b0576acda7 · inbound

Efficient Multilingual Neural Machine Translation via Corpus-Driven Vocabulary Pruning: An English-Arabic Case Study cites this paper.

Efficient Multilingual Neural Machine Translation via Corpus-Driven Vocabulary Pruning: An English-Arabic Case Study 8-bit Optimizers via Block-wise Quantization

Reference 2012

Resolution
unresolved
no resolver link, observed 2026-08-05T18:20:39.185565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:20:39.185565Z digest=sha256:d9386480e9b90acb6600ddb36404dcad4c5c49bd9d44560e96ee5b3a943d9fba