Pith. sign in

Paper Citation Record · LEDGER

Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2409.10999.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2409.10999 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T05:22:08.784841Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T17:56:58.130779Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6c44bf4d-c152-429c-8290-4f1a117dabf7 · inbound

AV-Odyssey Bench: Can Your Multimodal LLMs Really Understand Audio-Visual Information? cites this paper.

AV-Odyssey Bench: Can Your Multimodal LLMs Really Understand Audio-Visual Information? Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-11T23:19:11.594038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:19:11.594038Z digest=sha256:7887c6852e87f28a42850fb02e4b6d941f8992a09dfa5506fc104655a1de3d34

Observation afb63e51-5556-42cb-8129-10ff2b0f666a · inbound

Typhoon 2: A Family of Open Text and Multimodal Thai Large Language Models cites this paper.

Typhoon 2: A Family of Open Text and Multimodal Thai Large Language Models Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-11T12:56:13.308933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T12:56:13.308933Z digest=sha256:e2b5d100a9311a1edff76deabe79c4ea797a74418ae0e8bc494fbff1c88e2d80

Observation 3057d1ba-c2ca-4ac1-8e8c-e78614a1a48e · inbound

Contrastive Learning for Task-Independent SpeechLLM-Pretraining cites this paper.

Contrastive Learning for Task-Independent SpeechLLM-Pretraining Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T11:13:48.504255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:13:48.504255Z digest=sha256:ec97e778f3bff03feefa9753a8c2aa95e20b90490e83e40cb98fe57640f55d01

Observation 9c338a2d-068b-4f27-b55c-50c714a857b4 · inbound

Enhancing Non-Core Language Instruction-Following in Speech LLMs via Semi-Implicit Cross-Lingual CoT Reasoning cites this paper.

Enhancing Non-Core Language Instruction-Following in Speech LLMs via Semi-Implicit Cross-Lingual CoT Reasoning Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T05:22:08.784841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:22:08.784841Z digest=sha256:787397f6e5e226cae8dceeafb4892c78b93274086a989bdb8f7f8b69a3f17f13

Observation f546cba8-f8c1-4c1a-b2b0-7d4610d039f5 · inbound

Speechless: Speech Instruction Training Without Speech for Low Resource Languages cites this paper.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:16.797400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:16.797400Z digest=sha256:cdd232ca2b48da406b968e60260e92b78625c1f3ec351d8a5c7e2f506fe13006

Observation f84b21eb-086a-4199-95f3-c3341fda2126 · inbound

Weakly Supervised Data Refinement and Flexible Sequence Compression for Efficient Thai LLM-based ASR cites this paper.

Weakly Supervised Data Refinement and Flexible Sequence Compression for Efficient Thai LLM-based ASR Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T13:21:02.170670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:21:02.170670Z digest=sha256:72e7a228c96a0e1fb3c3af82c60b20a4aec320e3d16b3c502d388e84bd6649c7

Observation 809a6685-b873-4b14-9281-a9ece5f55ba6 · inbound

Breaking the Barriers of Text-Hungry and Audio-Deficient AI cites this paper.

Breaking the Barriers of Text-Hungry and Audio-Deficient AI Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T11:28:53.960789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:28:53.960789Z digest=sha256:f07a3507cbf2f712e525ca4fc542ac24910153370d933e56f225685b7b31f2c5

Observation 250c51c1-5224-424b-9189-07f2f9f51ff9 · inbound

AudioJudge: Understanding What Works in Large Audio Model Based Speech Evaluation cites this paper.

AudioJudge: Understanding What Works in Large Audio Model Based Speech Evaluation Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T16:47:56.560926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:47:56.560926Z digest=sha256:69c857f7740294f2a0bdf3d9e938cb03249182d40c7ddc07adc6f8b28477a953

Observation c1022617-f09f-4b4e-a144-18a27dace9ff · inbound

LLaSO: A Foundational Framework for Reproducible Research in Large Language and Speech Model cites this paper.

LLaSO: A Foundational Framework for Reproducible Research in Large Language and Speech Model Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models

Reference 55

Resolution
verified exact
local_arxiv, observed 2026-08-05T17:56:58.331121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T17:56:53.680396Z digest=sha256:4d00504e5959480824712c7ee17d2f26c21f5b63b33b1013590c5c88d09db8d9