Pith. sign in

Paper Citation Record · LEDGER

Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2409.10999.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2409.10999 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T05:22:08.784841Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T17:56:58.130779Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6c44bf4d-c152-429c-8290-4f1a117dabf7 · inbound

AV-Odyssey Bench: Can Your Multimodal LLMs Really Understand Audio-Visual Information? cites this paper.

AV-Odyssey Bench: Can Your Multimodal LLMs Really Understand Audio-Visual Information? Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-11T23:19:11.594038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:19:11.594038Z digest=sha256:2fe0f555f63f0f207e895ffbfa0baac42a4b9b7dbfebc68dcf5ea7fa72fb21e2

Observation afb63e51-5556-42cb-8129-10ff2b0f666a · inbound

Typhoon 2: A Family of Open Text and Multimodal Thai Large Language Models cites this paper.

Typhoon 2: A Family of Open Text and Multimodal Thai Large Language Models Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-11T12:56:13.308933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T12:56:13.308933Z digest=sha256:59900afd0730120182eb6d21528650f5e2267e32535baa320ec1d2c62c28f61e

Observation 3057d1ba-c2ca-4ac1-8e8c-e78614a1a48e · inbound

Contrastive Learning for Task-Independent SpeechLLM-Pretraining cites this paper.

Contrastive Learning for Task-Independent SpeechLLM-Pretraining Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T11:13:48.504255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:13:48.504255Z digest=sha256:c7fc5e60fab250781920190b1ff1351e2e2ddc581c9c9a750dc9f8ef40f032e9

Observation 9c338a2d-068b-4f27-b55c-50c714a857b4 · inbound

Enhancing Non-Core Language Instruction-Following in Speech LLMs via Semi-Implicit Cross-Lingual CoT Reasoning cites this paper.

Enhancing Non-Core Language Instruction-Following in Speech LLMs via Semi-Implicit Cross-Lingual CoT Reasoning Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T05:22:08.784841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:22:08.784841Z digest=sha256:44e2360f03962f092ba352cb1e453d2cdc4bc1fd3ef1261e12ff4319cf083bca

Observation f546cba8-f8c1-4c1a-b2b0-7d4610d039f5 · inbound

Speechless: Speech Instruction Training Without Speech for Low Resource Languages cites this paper.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:16.797400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:16.797400Z digest=sha256:341bad96b258a8447b91160f60d7b12835967b21b5f2d4add810c6e397165f3c

Observation f84b21eb-086a-4199-95f3-c3341fda2126 · inbound

Weakly Supervised Data Refinement and Flexible Sequence Compression for Efficient Thai LLM-based ASR cites this paper.

Weakly Supervised Data Refinement and Flexible Sequence Compression for Efficient Thai LLM-based ASR Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T13:21:02.170670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:21:02.170670Z digest=sha256:629c63a89709818cb757fe8a7f3769ce10346f015418a0a1edac458814fe9f91

Observation 809a6685-b873-4b14-9281-a9ece5f55ba6 · inbound

Breaking the Barriers of Text-Hungry and Audio-Deficient AI cites this paper.

Breaking the Barriers of Text-Hungry and Audio-Deficient AI Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T11:28:53.960789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:28:53.960789Z digest=sha256:731df4272d02a80bdb1a4fe87aa54caba0c8acbf40df0ee574349788ece6da02

Observation 250c51c1-5224-424b-9189-07f2f9f51ff9 · inbound

AudioJudge: Understanding What Works in Large Audio Model Based Speech Evaluation cites this paper.

AudioJudge: Understanding What Works in Large Audio Model Based Speech Evaluation Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T16:47:56.560926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:47:56.560926Z digest=sha256:e9fffec8bead2d3809502de4fa0e4c9590a4a0258cac64676f67944991c89cbb

Observation c1022617-f09f-4b4e-a144-18a27dace9ff · inbound

LLaSO: A Foundational Framework for Reproducible Research in Large Language and Speech Model cites this paper.

LLaSO: A Foundational Framework for Reproducible Research in Large Language and Speech Model Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models

Reference 55

Resolution
verified exact
local_arxiv, observed 2026-08-05T17:56:58.331121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:56:53.680396Z digest=sha256:56a3ab4abab110284b50e3e86f614eb40043b61929ee7b4d563fa4ffa8ad358f