Pith. sign in

Paper Citation Record · LEDGER

Speechless: Speech Instruction Training Without Speech for Low Resource Languages

As of 19 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 2 inbound Pith citation observations for arXiv:2505.17417.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.17417 v1

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:51:19.432830Z

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:51:15.305445Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T11:58:44.561636Z

Reference resolution

45 of 45 outbound references displayed

  • verified exact1
  • verified fuzzy22
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 11d56fbd-b22c-4a54-b72e-5a8c8ecf1c8e · outbound

This paper cites Speechless: Speech Instruction Training Without Speech for Low Resource Languages.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Speechless: Speech Instruction Training Without Speech for Low Resource Languages

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.305445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:15.305445Z digest=sha256:6f30ae1328087fa6f07da8bdd078e570afc7134c8f1f4dd70cd9cc107cd36d0f

Observation 900b0269-1f87-4455-87c5-55626e2fffb6 · outbound

This paper cites First, we train a residual vector quantizer (RVQ) to en- code speech into discrete semantic tokens that align with Whis- per’s encoder representations.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages First, we train a residual vector quantizer (RVQ) to en- code speech into discrete semantic tokens that align with Whis- per’s encoder representations

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:51:23.699872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:51:15.341648Z digest=sha256:47d9fb4115f51eaf8198f36dc7efef5b2e620e5b1a4f504105bef37ce833161b

Observation 51b3eb90-d3d4-4fbb-ae2d-e431f0cab151 · outbound

This paper cites Datasets For Stage 1, we utilized two automatic speech recognition (ASR) datasets: viV oice (Vietnamese) and LibriTTS-R[22] (English).

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Datasets For Stage 1, we utilized two automatic speech recognition (ASR) datasets: viV oice (Vietnamese) and LibriTTS-R[22] (English)

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:51:23.550068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:51:15.413140Z digest=sha256:204326034ff26bc3757156825d3783d29a143da558b61f4a020adaf22df0520c

Observation 1702fe69-67d2-46e0-8a75-5b85d782328b · outbound

This paper cites ASR and Speechless Comparisons To evaluate the performance of the Speechless model alone, we make use of ASR test sets.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages ASR and Speechless Comparisons To evaluate the performance of the Speechless model alone, we make use of ASR test sets

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:51:23.379474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:51:15.506491Z digest=sha256:bab475bab4bb7f9493a70f96fbf5cbde7e90954302bb2dbbbc35bad9f9993643

Observation dacab840-508a-49a8-a599-669601336807 · outbound

This paper cites By lever- aging a quantized Whisper encoder, Speechless generates se- mantic speech tokens, effectively addressing challenges in low- resource languages.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages By lever- aging a quantized Whisper encoder, Speechless generates se- mantic speech tokens, effectively addressing challenges in low- resource languages

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:51:22.925666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:51:15.761479Z digest=sha256:0110a9d41f219f8f8e849bf70cb3d423bd7278e745646226c4314432182d9a62

Observation 2c1e8d43-313a-4cff-8756-24ca52554909 · outbound

This paper cites This is also clear when see that with added noise (VBD noisy), the Whisper encoder starts to generate tokens that show poorer WER in comparison.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages This is also clear when see that with added noise (VBD noisy), the Whisper encoder starts to generate tokens that show poorer WER in comparison

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:51:23.064274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:51:15.667264Z digest=sha256:fe6a86b18804c2b4527c13297b2729e36cb47087d5eef2e3c8df07baa0f0568a

Observation eaf8e15b-238e-44e0-a5c9-eadd4d06e17a · outbound

This paper cites SALMONN: Towards generic hearing abilities for large language models,.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages SALMONN: Towards generic hearing abilities for large language models,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:16.343843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:16.343843Z digest=sha256:8f7a6db7b1250a148db901c696ac5571717d5a760ffe879abb0f6dbfc8a91119

Observation 2345f4fd-92d2-4833-8314-f1a8d9e0a7c2 · outbound

This paper cites Ichigo: Mixed-Modal Early-Fusion Realtime Voice Assistant.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Ichigo: Mixed-Modal Early-Fusion Realtime Voice Assistant

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:51:19.788559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:51:15.830544Z digest=sha256:5556fd104e8624c1abee38faccded2b8c5835891c17869b0d71b6995620a6966

Observation 24297cbb-faa3-4c2a-a25a-9cf84fe8d9b4 · outbound

This paper cites WavChat: A Survey of Spoken Dialogue Models.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages WavChat: A Survey of Spoken Dialogue Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.917688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:15.917688Z digest=sha256:52ecade508fea6395b27c8df91c19b90c43ce62ad03ec858602d970f770c8bae

Observation 8b62c3c8-bdc0-46e4-bea0-2e93f641eeb5 · outbound

This paper cites Recent Advances in Speech Language Models: A Survey.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Recent Advances in Speech Language Models: A Survey

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.984529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:15.984529Z digest=sha256:1611c94d9334a59bb27443894d516ac0f2d611b1d318e1ec1761e73a1aa30c61

Observation 0d13692c-21c8-42a9-bc12-cfe3f818b5c3 · outbound

This paper cites LLaMA-Omni: Seamless Speech Interaction with Large Language Models.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages LLaMA-Omni: Seamless Speech Interaction with Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:16.069890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:16.069890Z digest=sha256:266347eda2b1a9d03be6ac7e96ec03e634589e8b7cc869a1a446c5e6cec6abcd

Observation a563852e-78bc-42c8-8b2d-111ec6d27c1c · outbound

This paper cites Instruction data generation and unsupervised adap- tation for speech language models,.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Instruction data generation and unsupervised adap- tation for speech language models,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:51:22.786323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:51:16.167458Z digest=sha256:6a4987ccdb3fc157391a9655bd00a0514c6d199121b313023c9e2e87d2174053

Observation 4f491142-bab0-4c43-95c5-2aa0aaed3b49 · outbound

This paper cites Tango 2: Aligning diffusion-based text-to-audio generations through direct preference optimization,.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Tango 2: Aligning diffusion-based text-to-audio generations through direct preference optimization,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:51:22.581169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:51:16.269017Z digest=sha256:3deef64ee59423c5a7360b78737720b31b4bf4d0904b4c789abd97a42aeb2528

Observation caa0ee6b-f3a4-45e4-98b1-de8f6327de53 · outbound

This paper cites Distilling an End-to-End Voice Assistant Without Instruction Training Data.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Distilling an End-to-End Voice Assistant Without Instruction Training Data

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:17.014767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:17.014767Z digest=sha256:c637f419afea53bad183a4d293e31f5bafb9999008503bfe77ca762556ca0d65

Observation 607d190d-46da-4014-af69-c27e2ebce769 · outbound

This paper cites COSMIC: Data Efficient Instruction-tuning For Speech In-Context Learning.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages COSMIC: Data Efficient Instruction-tuning For Speech In-Context Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:16.436586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:16.436586Z digest=sha256:7c41405092045fc979333fe333e2189f6e0178c27af22df535853b18d50bbaa5

Observation a6c1edec-a87e-4e7f-a463-d1e943a859ce · outbound

This paper cites LibriSQA: A Novel Dataset and Framework for Spoken Question Answering with Large Language Models.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages LibriSQA: A Novel Dataset and Framework for Spoken Question Answering with Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:16.525195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:16.525195Z digest=sha256:6b063276e46c95ed1a752f6a3bf0f0d67049db59b1ccb0175a0c28c3a0bf7ded

Observation 2c331902-1aca-4b33-9d3c-e062d51eaabd · outbound

This paper cites An efficient and high fidelity vietnamese streaming end-to-end speech synthesis,.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages An efficient and high fidelity vietnamese streaming end-to-end speech synthesis,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:51:22.428889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:51:16.609766Z digest=sha256:ad8be29e30463049362386a935aa434feef6a5e4a475164e226d6455eb4fc64a

Observation 53ecc50b-f8f7-484c-ba70-7a61eb396779 · outbound

This paper cites Low-resource multilingual and zero-shot multispeaker TTS,.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Low-resource multilingual and zero-shot multispeaker TTS,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:51:22.262103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:51:16.720043Z digest=sha256:3466bceac809285e7aafcde6d722cb075d3ff7e03eb807ee16bb952ed141f1eb

Observation f546cba8-f8c1-4c1a-b2b0-7d4610d039f5 · outbound

This paper cites Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:16.797400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:16.797400Z digest=sha256:cdd232ca2b48da406b968e60260e92b78625c1f3ec351d8a5c7e2f506fe13006

Observation c790d769-9dcd-4498-b404-bb479547d23d · outbound

This paper cites Unsupervised cross-modal alignment of speech and text embedding spaces,.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Unsupervised cross-modal alignment of speech and text embedding spaces,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:51:22.078323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:51:16.900732Z digest=sha256:6df8e03db5daef1aa315b1effd62d92b01f928e48b1bee146227fcc3ea72aea5

Observation a31e3807-8636-42a3-b356-ff6020e822fb · outbound

This paper cites Alpaca: A strong, replicable instruction-following model,.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Alpaca: A strong, replicable instruction-following model,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:51:21.053822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:51:17.763096Z digest=sha256:9956230195b911ee416f7e163e385fa9352a43574a8f8ac1ade92d8d43c8d567

Observation 7c2dd311-c4b4-4b2c-8cdf-0d8e87ddef20 · outbound

This paper cites An analysis of semantically-aligned speech-text embeddings,.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages An analysis of semantically-aligned speech-text embeddings,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:51:21.907697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:51:17.124483Z digest=sha256:ab8c8f3717fce86e962531770944fbe44a213669db8103b30918bfb44372cbdd

Observation c1591c75-ab13-4a3b-b6fc-2972fd1bf797 · outbound

This paper cites Astra: Aligning speech and text representa- tions for asr without sampling,.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Astra: Aligning speech and text representa- tions for asr without sampling,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:51:21.694139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:51:17.245173Z digest=sha256:690f1fa9960ede66c6af080520d9ad091a167aae8bed613b4d60d7102329ea19

Observation abf51dc9-7f4b-4e10-8e0f-db41f1142271 · outbound

This paper cites Robust speech recognition via large-scale weak supervision,.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Robust speech recognition via large-scale weak supervision,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:17.321338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:17.321338Z digest=sha256:9ebc0c1f87283cd42929d2507cfee7ab5bb80554a8b6e2b49e2ec324068a1347

Observation 9347ce0c-ce5b-4d54-aaac-79a2f5c4f7cb · outbound

This paper cites Speecht5: Unified-modal encoder-decoder pre-training for spoken language processing,.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Speecht5: Unified-modal encoder-decoder pre-training for spoken language processing,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:51:21.495300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:51:17.427170Z digest=sha256:4abefbce88ab2ef3d432b8167c3a44135ddea902f43c2009cda79e0d53c57651

Observation d1757e7c-322e-43a2-914b-f4543644f43c · outbound

This paper cites Sailor 2 dataset,.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Sailor 2 dataset,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:51:21.318498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:51:17.534692Z digest=sha256:8241774a342cb8c87aaf469fc9a74bd90e78db94683ea4998ee530a32035e704

Observation d3fad88d-c6b6-4f61-954b-221df6fbe0e2 · outbound

This paper cites Vtsnlp instruct general dataset,.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Vtsnlp instruct general dataset,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:51:21.209910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:51:17.650905Z digest=sha256:9c39fb661c0c55211939fc691a802f71c71c0a2b69f8d67e2e78941d8ec79bed

Observation a89f7a66-63ba-40b4-9f0a-949c3ae289c1 · outbound

This paper cites VoiceBench: Benchmarking LLM-Based Voice Assistants.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages VoiceBench: Benchmarking LLM-Based Voice Assistants

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:18.448509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:18.448509Z digest=sha256:2866275fef647ce3eed1e9e99333c269ff80cea01b07fff635a0165cea578a06

Observation 9cfdabc0-e0dc-4eb3-8c5c-aa231bd844b1 · outbound

This paper cites LibriTTS: A Corpus Derived from LibriSpeech for Text-to-Speech.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages LibriTTS: A Corpus Derived from LibriSpeech for Text-to-Speech

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:17.873469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:17.873469Z digest=sha256:04a50c14f124f3b5a5114323ae14a1ed69615fafbfe90153d794ec46f65c1a15

Observation acede9d8-bcef-4531-b951-8ca6819b8fb1 · outbound

This paper cites vivoice: Enabling vietnamese multi-speaker speech synthesis,.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages vivoice: Enabling vietnamese multi-speaker speech synthesis,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:51:20.884120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:51:17.955220Z digest=sha256:a884a6e350d9f5c6bee55604f352bc359d3db9ef353acb9592da4926bfafe029

Observation 389b06a8-4a30-4540-af90-b95b5f135e3a · outbound

This paper cites an unresolved cited work.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:23.175121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:51:15.602667Z digest=sha256:3e360a213fac4a940f04506722879472616cf91cddff0df744491e3d99771858

Observation 0cc6ac39-bfac-41f9-91ac-54f907faddbf · outbound

This paper cites Libritts-p: A corpus with speaking style and speaker identity prompts for text-to-speech and style captioning,.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Libritts-p: A corpus with speaking style and speaker identity prompts for text-to-speech and style captioning,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:51:20.728142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:51:18.061792Z digest=sha256:2c12a6f364ab83a76d767401073bafa756c1ea1e9050884b7b38457f10262cd1

Observation 50bd9f11-e106-401b-9dec-ce817a39b5dd · outbound

This paper cites MLS: A Large-Scale Multilingual Dataset for Speech Research.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages MLS: A Large-Scale Multilingual Dataset for Speech Research

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:18.155649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:18.155649Z digest=sha256:2afb0e3fd4c938dc6d4546e2538642a2243b91615e090fc343c0d061aa9a9f98

Observation 0e5160ce-8d1b-4afb-82b0-c0bd78df15ff · outbound

This paper cites Efficient Memory Management for Large Language Model Serving with PagedAttention.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Efficient Memory Management for Large Language Model Serving with PagedAttention

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:18.265223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:18.265223Z digest=sha256:1fe15b1cb2390d05ec5b9c866b7ec64f916d30640475d6d8b0e092efead3ab64

Observation 2b6be35c-d36b-4d4c-a7a6-61f9e11e547c · outbound

This paper cites Ray: A Distributed Framework for Emerging AI Applications.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Ray: A Distributed Framework for Emerging AI Applications

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:18.378750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:18.378750Z digest=sha256:23e9540cbb8d36602f1de4e9780abb5951f91b01a517d65abd468808e69d4169

Observation ee710545-1be9-43f2-8cd7-d8b722303ef4 · outbound

This paper cites Lib- rispeech: an asr corpus based on public domain audio books,.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Lib- rispeech: an asr corpus based on public domain audio books,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:18.542975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:18.542975Z digest=sha256:ef5e250810eb6da383e2109ee770125f2afd1d96f3298411d9a310d3e377fadb

Observation abf34e67-8708-42fb-bb91-58d90c51c0d2 · outbound

This paper cites Speech enhancement for a noise-robust text-to-speech synthe- sis system using deep recurrent neural networks,.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Speech enhancement for a noise-robust text-to-speech synthe- sis system using deep recurrent neural networks,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:51:20.561260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:51:18.623468Z digest=sha256:d21f5dca794336aa0c019e90a3c01161cd6c0a5be9f6cdf0ffacfac7c5839670

Observation 18225890-ce84-465f-95d6-18d6f01ce73a · outbound

This paper cites The diverse environments multi-channel acoustic noise database (demand): A database of multichannel environmental noise recordings,.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages The diverse environments multi-channel acoustic noise database (demand): A database of multichannel environmental noise recordings,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:18.760225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:18.760225Z digest=sha256:6cc72cddbffa6cf49e3045a55dc28ecf35af12940601dbe49e5e67a8a4035c7a

Observation be6b2239-8b98-4800-b2aa-c7d0d53ef519 · outbound

This paper cites Com- mon voice: A massively-multilingual speech corpus,.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Com- mon voice: A massively-multilingual speech corpus,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:51:20.395648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:51:18.825142Z digest=sha256:8beddb36b984b55d0f65ea68493380674cae51d0f57d5df5ac9b2dfa03ae9623

Observation 520c3d00-e442-4328-a5f0-0dd46f1321b7 · outbound

This paper cites Length-Controlled AlpacaEval: A Simple Way to Debias Automatic Evaluators.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Length-Controlled AlpacaEval: A Simple Way to Debias Automatic Evaluators

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:18.912898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:18.912898Z digest=sha256:a0e7beb1fca30082b3ce8ecaaba58d0ae7f87ff3f4e5c69f7f70158401bd2597

Observation ccb497d4-3c47-4829-a70a-ba964cd1bfd5 · outbound

This paper cites SD-QA: Spoken dialectal question answering for the real world,.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages SD-QA: Spoken dialectal question answering for the real world,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:51:20.223654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:51:19.032030Z digest=sha256:cde52446ff6acdd78bdc765268867172f86fd4e1207a1d7f8439c2a29a806530

Observation 41d8c8c1-34f7-4e8a-ba81-3a2af63b33d9 · outbound

This paper cites Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:19.139471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:19.139471Z digest=sha256:05fc4ca7dec484e1ba6acfc50eddfcd4d46f66dc887a7d054a637ee0dbd16173

Observation 9ccdc6d9-0ca4-42ca-8e96-0d8ac76004b9 · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:19.219916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:19.219916Z digest=sha256:7b52ded80f876a6e14d111585c241b7a87aedf02e204759f89be8b9da6e56553

Observation bb07a2e8-90bc-4f87-a40a-8a4236a11a76 · outbound

This paper cites Moshi: a speech-text foundation model for real-time dialogue.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Moshi: a speech-text foundation model for real-time dialogue

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:19.350579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:19.350579Z digest=sha256:cd436f12bb6f5e56c4ee0223aa24a96800af82dfb0c81729246e77e75e8389ef

Observation 21615463-daac-4d2e-8f38-928b36f96ce0 · outbound

This paper cites BLSP: Bootstrapping language-speech pre-training via behavior alignment of continuation writing,.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages BLSP: Bootstrapping language-speech pre-training via behavior alignment of continuation writing,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:51:20.026378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:51:19.432830Z digest=sha256:fbc53adda45d18f1bce53f81789acb3651bf2a2c963ca7f0f449780130ffa9a4

Pith citing papers

Observation 11d56fbd-b22c-4a54-b72e-5a8c8ecf1c8e · inbound

Speechless: Speech Instruction Training Without Speech for Low Resource Languages cites this paper.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Speechless: Speech Instruction Training Without Speech for Low Resource Languages

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.305445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:15.305445Z digest=sha256:6f30ae1328087fa6f07da8bdd078e570afc7134c8f1f4dd70cd9cc107cd36d0f

Observation 6df42d5f-d314-4609-a221-5560619ae578 · inbound

TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment cites this paper.

TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment Speechless: Speech Instruction Training Without Speech for Low Resource Languages

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:58:44.627131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:58:41.531921Z digest=sha256:ce6b0749362daad05a4c51ade4a7e49a78141a501551d15ec8e65ca1599cf5a5