Pith. sign in

Paper Citation Record · LEDGER

Sparks of Large Audio Models: A Survey and Outlook

As of 23 July 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2308.12792.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.12792 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-07-22T06:31:00.163083+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-22T20:44:57.476464Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T20:45:08.153364Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 348d5f4f-9348-4e04-bbce-ceabc736624b · inbound

On The Landscape of Spoken Language Models: A Comprehensive Survey cites this paper.

On The Landscape of Spoken Language Models: A Comprehensive Survey Sparks of Large Audio Models: A Survey and Outlook

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-22T20:45:08.156543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-22T20:44:57.476464Z digest=sha256:169c59eda424a5f6c90eee085a1527b04d95a32a98a88bdaf1eca173a263b086

Observation 1e380dfc-05f5-4624-8b86-ba2db684e2f1 · inbound

Game-Time: Evaluating Temporal Dynamics in Spoken Language Models cites this paper.

Game-Time: Evaluating Temporal Dynamics in Spoken Language Models Sparks of Large Audio Models: A Survey and Outlook

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T11:52:35.681744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-18T11:51:43.561210Z digest=sha256:895cc7e860a2139ba4ad9f1e00fc684e34d71fb242cc0e1b4bc1a42e07e4b1e2

Observation 1b5d64a3-e72e-4c99-85af-ec6faa9844d4 · inbound

Hearing to Translate: The Effectiveness of Speech Modality Integration into LLMs cites this paper.

Hearing to Translate: The Effectiveness of Speech Modality Integration into LLMs Sparks of Large Audio Models: A Survey and Outlook

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-16T21:51:17.776305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-16T21:49:21.785096Z digest=sha256:36ccfe92ff7d01a56d31820753c2ec00349f55365c444f53c646c875acfc9848

Observation 4a84b826-a364-475a-8450-b2583259cb58 · inbound

Generative AI in Signal Processing Education: An Audio Foundation Model Based Approach cites this paper.

Generative AI in Signal Processing Education: An Audio Foundation Model Based Approach Sparks of Large Audio Models: A Survey and Outlook

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T08:47:37.375859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-16T08:44:02.494895Z digest=sha256:23a94f7e778910072dec1a2665b725e12c9bd9d6d4f805070ae6db28447d4633

Observation 1c77c32c-4c42-4db6-93c0-c645629de473 · inbound

Heterogeneity-Aware Dataset Scheduling for Efficient Audio Large Language Model Training cites this paper.

Heterogeneity-Aware Dataset Scheduling for Efficient Audio Large Language Model Training Sparks of Large Audio Models: A Survey and Outlook

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T07:18:07.311169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-20T07:13:29.041056Z digest=sha256:4aa15e9396bf3ee7518d6f285cf5da76ccc2ec834993eeddc2408d0142a7d76c

Observation 6c6d05f5-03a7-48d3-b83c-8c7ed597307e · inbound

A Survey of Large Audio Language Models: Generalization, Trustworthiness, and Outlook cites this paper.

A Survey of Large Audio Language Models: Generalization, Trustworthiness, and Outlook Sparks of Large Audio Models: A Survey and Outlook

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T07:39:48.960571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-21T07:38:23.099479Z digest=sha256:48ade78b551a553f8a604e5efdb7d4d9c1b5e9626fb462291ca871410b09f2a1

Observation ca5800a4-c99a-4c78-94eb-dd7fa1bf053a · inbound

A Survey of Audio Reasoning in Multimodal Foundation Models cites this paper.

A Survey of Audio Reasoning in Multimodal Foundation Models Sparks of Large Audio Models: A Survey and Outlook

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T02:09:24.461469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-21T02:08:06.976461Z digest=sha256:19c51cdf0cbd457d4f2afa75bc7aada23bd56e36d19deeb7a2b742a593362a4d