Pith. sign in

Paper Citation Record · LEDGER

Locality Matters for Training-Free Audio Token Compression in Audio-Language Models

As of 20 August 2026, this Paper Citation Record lists 5 of 5 outbound references and 0 inbound Pith citation observations for arXiv:2605.25179.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.25179 v1

Coverage vector

measured 5 of 5 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-30T11:39:41.560725Z

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

5 of 5 outbound references displayed

  • verified exact1
  • verified fuzzy2
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0af7bde2-e635-42e2-a408-3b0668f72d04 · outbound

This paper cites InInter- national Conference on Learning Representations.

Locality Matters for Training-Free Audio Token Compression in Audio-Language Models InInter- national Conference on Learning Representations

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T09:16:07.165404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T11:39:41.560725Z digest=sha256:9daf1cd0f7909fe58714bd261074f5652fd8e64cbe93e29df82a4ae2fa3cd915

Observation b71f69bb-56ca-4d89-abb3-60ae8d8a68b6 · outbound

This paper cites Qwen2-Audio Technical Report.

Locality Matters for Training-Free Audio Token Compression in Audio-Language Models Qwen2-Audio Technical Report

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T11:44:38.353807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T11:39:41.560725Z digest=sha256:be6ccf3f3ba4b5ea22270eb69a5d3ca5296ca7d223172c63338e86f5a9b8830f

Observation 29a2bb70-bba0-48cf-b300-454338361a75 · outbound

This paper cites Segmentwise pruning in audio-language models.

Locality Matters for Training-Free Audio Token Compression in Audio-Language Models Segmentwise pruning in audio-language models

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:44:38.348947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T11:39:41.560725Z digest=sha256:498d2fde00904ea5342645e049553dd1d0cbe5edce6e564537625d4365b22ae5

Observation 37f6b381-e56a-4e33-a38a-e603e6bee3fa · outbound

This paper cites Beyond Text-Visual Attention: Exploiting Visual Cues for Effective Token Pruning in VLMs.

Locality Matters for Training-Free Audio Token Compression in Audio-Language Models Beyond Text-Visual Attention: Exploiting Visual Cues for Effective Token Pruning in VLMs

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T11:44:38.351501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T11:39:41.560725Z digest=sha256:2eccd1cc867da2285b96183565f8ae89337e0c972f8312016c2e970b78d24f6e

Observation 7b9bd658-b5ee-4b59-9da9-c5a47ea43512 · outbound

This paper cites InProceed- ings of the IEEE/CVF International Conference on Computer Vision.

Locality Matters for Training-Free Audio Token Compression in Audio-Language Models InProceed- ings of the IEEE/CVF International Conference on Computer Vision

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T09:16:07.163463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T11:39:41.560725Z digest=sha256:d6f0de354bfb4402ddde1c16c7f970ea0600a0e8a47b7d89993c0f12ef92cd35

Pith citing papers

No inbound Pith citation observations are available.