Pith. sign in

Paper Citation Record · LEDGER

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts

As of 7 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 3 inbound Pith citation observations for arXiv:2506.07533.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.07533 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:37:52.831949Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T07:06:32.220604Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T07:06:44.760367Z

Reference resolution

41 of 41 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved41
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 93ef760f-9fe2-455e-bfb2-7cc69503ccb0 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.101344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.726359Z digest=sha256:38c464b17dbe70ff40c0190a391bf0ab16a009f674a4841e20d260c75fe4584a

Observation 00c1b072-33df-4c31-8784-c606e5bad556 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.093931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.729848Z digest=sha256:5d3b2a7bb5ad5253e2da343a2b3d2e9b56769cf775389c64a40abb9b9ae6f245

Observation cad4eba4-2e70-483a-86af-74d99a9302d8 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.086525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.732621Z digest=sha256:bb78c1cce2e47dccd7d39e28d1be4beda10f9781e1571bf5f7d15c5f3db5023f

Observation ff7b7b0d-b588-4130-bf45-8099c29fedc2 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.079284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.736276Z digest=sha256:c56d8c83de981866ffc7d7120dd2d331c2affd651ff779d451dafd79cba96372

Observation d0cc57ba-1a11-44bc-81c1-d6d6084505d2 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.739085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.739085Z digest=sha256:aa01148849e8879baf59a6441bedb2e79472bc42e400559140b1baf3ba0f5127

Observation 9d270926-113e-4682-8a45-974699604732 · outbound

This paper cites The Llama 3 Herd of Models.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts The Llama 3 Herd of Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.741766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.741766Z digest=sha256:db4f297c6d83e8d4858b05fe23bc9f440d41927ad49e0e3642f087421ccf2354

Observation 11cae4a9-7209-4517-a833-125d20dfee12 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.744828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.744828Z digest=sha256:bd2164868212e422866ae518c5f13aa2484058a6e9fcdb095fb5b00f594d49c6

Observation 58e4bff1-2dd4-414c-9faa-12c37e7358e2 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.064517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.747319Z digest=sha256:36a1ab45359b273bc11c82af4c8eca33ed257263faddd6d341b2851c0d1f8733

Observation 1a19405e-cd93-4e6e-a6cb-9cb971805fcc · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.057502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.749750Z digest=sha256:a8b263f2f59d8423614a19b8061721368a699bc9de211f16a034128b5d5daa56

Observation 62b3619a-8255-4aad-a1c8-50cf07bc25a0 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.049905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.752293Z digest=sha256:0c9a4fa38a5db8e582c376faff561ca1b8399658373f8b31f851d8d101029f0a

Observation 1b096e79-904e-4e7a-9c35-dab5e3ecdf9f · outbound

This paper cites FastDecode: High-Throughput GPU-Efficient LLM Serving using Heterogeneous Pipelines.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts FastDecode: High-Throughput GPU-Efficient LLM Serving using Heterogeneous Pipelines

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.754758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.754758Z digest=sha256:2c63bb9f80975ba9970a149804154012caf781a8d5af993fb58d9bb0b11729c5

Observation 8eaa42d7-0799-46b7-b4f1-a8016b6bfc54 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.757583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.757583Z digest=sha256:089804a8f88104cea9e0c58ff6429d0ecc4aedc5eea4b3a8ddc5f0dc12a8649f

Observation 70738aa0-3e22-43c7-b94d-a7b9c39670f9 · outbound

This paper cites Mistral 7B.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Mistral 7B

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.759987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.759987Z digest=sha256:d364654a21b30e1f2d7a350890952cf28258818d9e500d50136dff02abc5e7a3

Observation 2fa381d1-f621-471f-bd56-2dd8f5858f38 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.038373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.762469Z digest=sha256:7bbae815ca995b0a620ae57e9fbb8058f334cf0d76781ff86dfb5095f5eede83

Observation 33383077-b0d1-46c5-a5dc-d23f0f20442a · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.031098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.764935Z digest=sha256:3bc311a5f58efd334f68ccb389eb7ba5f569175d5114596d07de5afb25a58774

Observation 24978516-5df3-4c95-b05e-6100d9279c55 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.767321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.767321Z digest=sha256:d1feca6fe0309a9dbe840b126d20443806f0e3ec140846ee5c74c2340e5d1ca9

Observation 471a6197-9f0c-4740-b8f7-e7daa127c0f8 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.020173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.769657Z digest=sha256:5752020d69e57203e0e6e902b53224d5b28b7c6e4c9ee0d7ec358ff55c409c88

Observation 391055aa-87be-4d0e-a8ab-a3c80408998e · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.011936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.772099Z digest=sha256:8a3e2e3a15eb3ee98ac6db8fb3dab893d0c8e47601aa74d5d03359d8718485bd

Observation 69fc9aba-5171-4328-8028-6c167f724b79 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.004184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.774596Z digest=sha256:7a19191f363605a27fa77f6eed832f5d8957817f4f266073b01c040a1be099db

Observation fbcfab01-79d6-4a7c-b889-80c989f83034 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.996040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.777238Z digest=sha256:fa4bb9867c4983bfcf20c9c6e930382dcd0833f1509b92cabbc5258039cfc5c4

Observation be14830c-359a-4c1a-89ec-ff7a6c519f25 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.988534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.779700Z digest=sha256:54a76238b45e5735c7d17bcd89acbf2879b64a797deaff6cf6e6f9890ff213ed

Observation 1ea7fddc-0f24-4453-93ed-d2c6a9d2e76f · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.980782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.782013Z digest=sha256:c78dda385419f8b9c3afcf021fbd649e898a8bc026e8f52f0b15fe04b8c262b5

Observation 3c4f9fe3-4537-4d32-af70-8a79c5127e2f · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.973484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.784541Z digest=sha256:c6f1c1a380bee5802a428817438fb8364abfeea356677a86b00dff19d07f9faa

Observation bbfdb19f-daaf-43af-b8c1-5dbaab7ca91b · outbound

This paper cites Faster Causal Attention Over Large Sequences Through Sparse Flash Attention.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Faster Causal Attention Over Large Sequences Through Sparse Flash Attention

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.787126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.787126Z digest=sha256:40774a8dfcfc89cb5d64ba2cb0e8bed591d6696b41d1e008e3b27ed9645ba53e

Observation f653f2a2-fb66-4252-9dde-b75c79ef059a · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.965936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.789767Z digest=sha256:13bf25e7f9f3629ae99c967f4ffefacb111853760b25f5af741fb2851620e390

Observation 193f9b5b-60dc-402d-9cac-c2b29da981d9 · outbound

This paper cites MADLLM: Multivariate Anomaly Detection via Pre-trained LLMs.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts MADLLM: Multivariate Anomaly Detection via Pre-trained LLMs

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.792294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.792294Z digest=sha256:9008101005d305e21387e2912007cb3ea168845b12441e470f6b09ffcd64c287

Observation 35f25e79-6c45-4749-a952-90cca57ef983 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.958128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.795055Z digest=sha256:d38fcc87cd4d8d2e620e6b8a8284c0b6bc5830a0373ad735a84ae8671e1fb9f6

Observation 300f7af0-503a-407a-8bd8-a0b576c0ad48 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.797726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.797726Z digest=sha256:450a45b0953ddafafa994ff0d5292afe73b97526225e1927326639c19bca436a

Observation f99d2bfa-44fc-4718-9bcd-42802985e054 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts LLaMA: Open and Efficient Foundation Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.800505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.800505Z digest=sha256:16c1cb619c4ba3dfffae8ba256e2f1b0d8d7c48bfde08cff41633ddd5e5fc3ce

Observation 80714b5a-a09a-423f-977f-ce829752f20b · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.803173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.803173Z digest=sha256:9d2b3a99bc8cc89ffdd682d0f237a47ec7948d3a69daa91ee9bd1fe9a45459f2

Observation 8ec93e94-a880-4197-bd4c-a5ffe10393f7 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.805537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.805537Z digest=sha256:ecf964333ef5626c95bac2a6a97248086bfb06fe0c3be83a84679c5489de0bde

Observation feb5f522-870d-4ed0-a219-fdf23f4847e2 · outbound

This paper cites PowerInfer-2: Fast Large Language Model Inference on a Smartphone.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts PowerInfer-2: Fast Large Language Model Inference on a Smartphone

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.808116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.808116Z digest=sha256:660962bbe38d9f5d1b9fb09c61d9be781e1550479b1d503eb8cb1bd32669197d

Observation edbfd869-51d3-4373-a409-88549bb739b6 · outbound

This paper cites No Token Left Behind: Reliable KV Cache Compression via Importance-Aware Mixed Precision Quantization.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts No Token Left Behind: Reliable KV Cache Compression via Importance-Aware Mixed Precision Quantization

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.810890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.810890Z digest=sha256:056eaaef62ade06e4f25cf255fbffe24ed29113b0b530dac798334bb70a925c5

Observation dabcf591-c317-44d8-805c-5be471d6d578 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.813698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.813698Z digest=sha256:230efa9222001bfb5d1da2db2eb4f2301a116372750e6e328c8ecc7f05a0da0f

Observation fb1188af-ed16-4d39-b644-994bcfb1a400 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.942824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.816113Z digest=sha256:aeee734258cdd0e4ad947aa6c47b65e60b3472139876f78a26da5522b2213908

Observation 2da58a12-dbd8-46a8-8a91-4fca6f72382b · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.935733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.818713Z digest=sha256:1b98c23930b044071d3f791790b248c2cd0f674c970ae95463c7cac23a62ca67

Observation d6ba924a-3ae2-4abe-8cfd-a5028b7e80e3 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.927800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.821094Z digest=sha256:a809de38900c26a8689dbe8bcb7c6f17839ac5be5f024d2d83a1ec400d86c9b0

Observation ec05a50f-2d00-48a0-9a0a-d4c08bd40119 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.824237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.824237Z digest=sha256:d0a55f526c99a48e45c28610815921085b4abcce578a30e6dcfd7fcceb7cbf22

Observation 5e5aa577-2ea5-4556-9901-5379ceb2738a · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.826625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.826625Z digest=sha256:dfd7207b0764e75c156d3d7e801af954f1c1245701a10b2a4525129e697f7958

Observation 87fac50a-0b1e-4fb8-a603-688c0dc16a9b · outbound

This paper cites online" 'onlinestring :=.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts online" 'onlinestring :=

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.829127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.829127Z digest=sha256:323e1d16e0dd7e2819988a6f7a722705751a5e9bf7e4f60127cec0e34dc345e3

Observation 874692fc-bb50-47bc-9e4e-424041ea12d6 · outbound

This paper cites write newline.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts write newline

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.831949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.831949Z digest=sha256:4e249b83de8816f1f7f12ff977522fe3d04e2a2642bb84f8450dc4f4051535f2

Pith citing papers

Observation f73211dc-0b3b-4b67-9f8b-08170c1576cf · inbound

LayerScope: Predictive Cross-Layer Scheduling for Efficient Multi-Batch MoE Inference on Legacy Servers cites this paper.

LayerScope: Predictive Cross-Layer Scheduling for Efficient Multi-Batch MoE Inference on Legacy Servers MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:41:22.811448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T12:38:31.783807Z digest=sha256:a3463b19f3d6517cdbffaaba0ae310d388fa721944846257a813607c93fe6402

Observation f42a5624-500e-456e-945c-895ae92cf14b · inbound

LBLLM: Lightweight Binarization of Large Language Models via Three-Stage Distillation cites this paper.

LBLLM: Lightweight Binarization of Large Language Models via Three-Stage Distillation MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts

Reference 95

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:46:04.826877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-10T03:04:14.900791Z digest=sha256:1d53fb1226ca8b3778a18b88bd370b3213dba626a09df32b81df1554856ab988

Observation 3319422f-3338-450e-8a37-083fda5a5ae0 · inbound

AlphaQ: Calibration-Free Bit Allocation for Mixture-of-Experts Quantization cites this paper.

AlphaQ: Calibration-Free Bit Allocation for Mixture-of-Experts Quantization MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T07:06:44.762157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T07:06:32.220604Z digest=sha256:925df35d46fb8f5c9f107a2352a9423ab6e8738f83a52b8f2bb39faa3cafe133