Pith. sign in

Paper Citation Record · LEDGER

SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval

As of 19 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 1 inbound Pith citation observation for arXiv:2412.12009.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.12009 v2

Coverage vector

measured 24 of 24 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T14:27:27.353817Z

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-30T14:35:02.484377Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T14:44:45.645532Z

Reference resolution

24 of 24 outbound references displayed

  • verified exact0
  • verified fuzzy16
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6cfb7fff-0c53-459d-ad20-d62b262953a2 · outbound

This paper cites A survey on speech large language models,.

SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval A survey on speech large language models,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T14:27:27.207916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:27:27.207916Z digest=sha256:3b4f71cec0b3da4ced390407e0de9e336e15edfe43eb9c3352b2058e552a8c49

Observation 615b9647-943e-49d8-96ce-4193edc48367 · outbound

This paper cites Qwen2-Audio Technical Report.

SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval Qwen2-Audio Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T14:27:27.214993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:27:27.214993Z digest=sha256:5dbab785981f3da2282bbed7b3551c1684d24305619a436e00400a6ea1e6be76

Observation 91cc5029-1500-4b98-b8c1-4dd268680a26 · outbound

This paper cites Distilling an End-to-End Voice Assistant Without Instruction Training Data.

SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval Distilling an End-to-End Voice Assistant Without Instruction Training Data

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T14:27:27.223349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:27:27.223349Z digest=sha256:153a209a8589973c0cb9484a406196a9451b5bacbf3ec350f7f0b6fff39ba5b4

Observation d991f7e9-421f-42e6-b482-c0a6254c6d7d · outbound

This paper cites SpeechGPT: Empowering large language models with intrinsic cross-modal conversational abilities,.

SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval SpeechGPT: Empowering large language models with intrinsic cross-modal conversational abilities,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:27:27.949943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T14:27:27.229904Z digest=sha256:7e8fefb086bd7bb8f146529f2fcec54467af76cc54b0643b03e03fc2c90a0cec

Observation 8682b50a-fe3b-4c59-b3e9-6671b498d8e7 · outbound

This paper cites AnyGPT: Unified multimodal LLM with discrete sequence modeling,.

SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval AnyGPT: Unified multimodal LLM with discrete sequence modeling,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:27:27.932649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T14:27:27.237231Z digest=sha256:12474a264cce66e61c327c996653d54d1beae87cb6b2d74d4750d65764c32418

Observation 37b91436-91f8-4d47-b7f1-c354d36ba71e · outbound

This paper cites Dynamic-SUPERB phase-2: A collaboratively expanding benchmark for measuring the capabilities of spoken language models with 180 tasks,.

SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval Dynamic-SUPERB phase-2: A collaboratively expanding benchmark for measuring the capabilities of spoken language models with 180 tasks,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:27:27.914804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T14:27:27.242211Z digest=sha256:6159834a4cb383c83ac34ab0d465358b4de977af24b9cafaaf56915de89e38e9

Observation 6334cba9-d4d5-421f-ba93-fbbb90c882ac · outbound

This paper cites MMAU: A massive multi-task audio understanding and reasoning benchmark,.

SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval MMAU: A massive multi-task audio understanding and reasoning benchmark,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:27:27.898104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T14:27:27.248819Z digest=sha256:45232fff0b37a9ff7a5bbe0b91f9848bb26fb3513a023ac45096d569be172694

Observation 6b5f935f-5e54-4c99-b575-ea8ef3f15a60 · outbound

This paper cites AudioBench: A Universal Benchmark for Audio Large Language Models.

SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval AudioBench: A Universal Benchmark for Audio Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T14:27:27.254326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:27:27.254326Z digest=sha256:226fc9bbd3d9401554ec5dd2bd848026f8c7af4aa6dbdb1e83f0353287746850

Observation 16451f2e-2264-436d-ac06-90a331ddb238 · outbound

This paper cites On the com- putational complexity of self-attention,.

SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval On the com- putational complexity of self-attention,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:27:27.880027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T14:27:27.259908Z digest=sha256:43425b124645ce9c974d6437481123f2fb2a4891c8937d6379ac5ff2f2a1828a

Observation 2c3b74b5-599f-4d8e-9ca8-18781f522004 · outbound

This paper cites Robust speech recognition via large-scale weak supervision,.

SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval Robust speech recognition via large-scale weak supervision,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T14:27:27.265395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:27:27.265395Z digest=sha256:bc36ff48b433243e2bd24351a45f9f7a5aeb5c6effd9e9cfb4c495a09e38a70e

Observation 5c4387b6-9496-468d-a057-55d8aa83cdfa · outbound

This paper cites Llava-prumerge: Adaptive token reduction for efficient large multimodal models,.

SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval Llava-prumerge: Adaptive token reduction for efficient large multimodal models,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T14:27:27.271000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:27:27.271000Z digest=sha256:c64726e19b0dde6802be41e48fdf14c96939554912ea8c420647be691cf3c4dd

Observation 6eaf1794-7992-4709-9582-56a9c0f65cae · outbound

This paper cites Styletts 2: Towards human-level text-to-speech through style diffusion and adversarial training with large speech language models,.

SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval Styletts 2: Towards human-level text-to-speech through style diffusion and adversarial training with large speech language models,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:27:27.852739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T14:27:27.278405Z digest=sha256:26ad53c8d5c4426bd56ca3f7d33d7ea7bd7e6705173547e2b817e23550317e92

Observation eb8cfd2f-580e-4966-b0fb-99ea0702f51a · outbound

This paper cites LibriTTS: A corpus derived from librispeech for text- to-speech,.

SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval LibriTTS: A corpus derived from librispeech for text- to-speech,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:27:27.835898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T14:27:27.284337Z digest=sha256:ce30ba87e1536fd25bec3ea8ca4644e4af9dfc7c367730fea62e6ffac68efcd5

Observation 88fff4b9-643f-40ef-a897-f37563c920ba · outbound

This paper cites Utmos: Utokyo-sarulab system for voicemos challenge 2022,.

SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval Utmos: Utokyo-sarulab system for voicemos challenge 2022,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:27:27.819049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T14:27:27.289656Z digest=sha256:16ff65843870156ec1a48ea93de92ac2459dcc81883b81c76446847f0576510b

Observation fec33ec5-9763-4267-9c15-a26c18fe8eab · outbound

This paper cites Listen, think, and understand,.

SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval Listen, think, and understand,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:27:27.802854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T14:27:27.295112Z digest=sha256:e72cdbc5fd8612dc5fbf029956bb2b66dbbe642608cb2a95667d2a46fbe1bcd4

Observation b96a12fb-37fc-4f2e-b0ec-67f5581f292c · outbound

This paper cites SALMONN: Towards generic hearing abilities for large language models,.

SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval SALMONN: Towards generic hearing abilities for large language models,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:27:27.786330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T14:27:27.301158Z digest=sha256:9de9f5633f52f2e729f16921194db83c79bf3ee2508e25ad4da42fb32daf46a7

Observation 07045fb6-57ee-4793-9e2e-aee6c25e3365 · outbound

This paper cites SpeechVerse: A Large-scale Generalizable Audio Language Model.

SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval SpeechVerse: A Large-scale Generalizable Audio Language Model

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T14:27:27.306355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:27:27.306355Z digest=sha256:78df83f3f0bdc1d62de7bb0b343b6944ef8f1a9edfe776ec153b4d3c654714b7

Observation 29a50415-014b-4365-94d0-4dc07d53b2a1 · outbound

This paper cites Attention is all you need,.

SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval Attention is all you need,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:27:27.768934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T14:27:27.317192Z digest=sha256:2cb92d3b73118157578c26481aa33f925a88ec6767e6d3601c108de419c462c9

Observation cf0e7d12-b0ff-49eb-80a7-bc588827afe6 · outbound

This paper cites Transformer quality in linear time,.

SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval Transformer quality in linear time,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:27:27.750112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T14:27:27.322853Z digest=sha256:0c86474cc6aacef7d0eea184e22681b53e1ec5b8ed5b8b4f017d0b5c7d0c7695

Observation 0bdeeabc-0705-4dda-a662-d1d1e81ff128 · outbound

This paper cites Learned token pruning for transformers,.

SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval Learned token pruning for transformers,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:27:27.734347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T14:27:27.327836Z digest=sha256:4398d1b3f06a59f29daec6cf7d76a76d73a9f6237652684e256bfca23140b79d

Observation b1377ffc-43f5-44e1-8593-d1475917ee1c · outbound

This paper cites Cortical oscillations and speech processing: Emerging computational principles and operations,.

SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval Cortical oscillations and speech processing: Emerging computational principles and operations,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:27:27.715513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T14:27:27.334489Z digest=sha256:71844b19797341c05f88be0c715b984c9e74868ffc16eef60442ed1d9fde2793

Observation 3eddaf34-10e1-4853-bcea-ca125b391400 · outbound

This paper cites LLM Inference Unveiled: Survey and Roofline Model Insights.

SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval LLM Inference Unveiled: Survey and Roofline Model Insights

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T14:27:27.340179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:27:27.340179Z digest=sha256:f651088f00d7c46fd24b2923b3f2f220898d3114747523baf62030df25c31683

Observation 3c719319-0ede-44e5-ba2a-bb6eecdc3606 · outbound

This paper cites Dream: A challenge data set and models for dialogue- based reading comprehension,.

SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval Dream: A challenge data set and models for dialogue- based reading comprehension,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:27:27.697902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T14:27:27.348476Z digest=sha256:7206eb3b6ef7809124e901df20ffe4c572e47665d4c91ef887a6102093e51c3d

Observation c834cfa0-87fe-4051-add2-51de44ff5887 · outbound

This paper cites WavLLM: Towards robust and adaptive speech large language model,.

SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval WavLLM: Towards robust and adaptive speech large language model,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:27:27.677871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T14:27:27.353817Z digest=sha256:ccf017838b3870ca8df7c5db22c1f3024dbd49592d6e013ed8907af04d1b4cb2

Pith citing papers

Observation 27cbed04-321f-4d1b-b206-b87602c53b06 · inbound

EVA: Accelerating LLM Decoding via an Efficient Vector Quantization Architecture cites this paper.

EVA: Accelerating LLM Decoding via an Efficient Vector Quantization Architecture SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-06-30T14:44:45.647323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T14:35:02.484377Z digest=sha256:de82d3195560af0ac09ab79ed77dc427415b781efe3fc3fe26e5864f371c4d74