Pith. sign in

Paper Citation Record · LEDGER

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs

As of 22 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 3 inbound Pith citation observations for arXiv:2508.17863.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.17863 v1

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T16:46:49.282451Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-29T00:56:43.991936Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T07:27:44.920929Z

Reference resolution

57 of 57 outbound references displayed

  • verified exact5
  • verified fuzzy0
  • unresolved52
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a9b52bf9-b6e4-40c0-b305-81f589a83a6f · outbound

This paper cites On The Landscape of Spoken Language Models: A Comprehensive Survey.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs On The Landscape of Spoken Language Models: A Comprehensive Survey

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:47.064133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:47.064133Z digest=sha256:afc3706904326811481fe42095269f131877e5e62c9267c195374af44c553efa

Observation 3aa58d23-2c8b-4667-9f07-2639666c394f · outbound

This paper cites Qwen Technical Report.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Qwen Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:47.124721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:47.124721Z digest=sha256:64a99cdd09116daf821652693cd51da99e3a81a881fa15bebea8a396a2dd179d

Observation 555c79fd-572b-4e21-916f-edbf18b4c574 · outbound

This paper cites SLURP: A Spoken Language Understanding Resource Package.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs SLURP: A Spoken Language Understanding Resource Package

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:47.223367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:47.223367Z digest=sha256:4e4b60ec9b1337e5a8b7ebb42bd0af6be92d3e12e7405930f944bc230ffd7bc2

Observation ef17f421-9602-4c3a-9911-b47eb4b02ad4 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.957548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-05T16:46:47.328362Z digest=sha256:5797eb4c60879d182cbb35d455f626749aeb5b6efc401086a1b7ce284d225695

Observation 21e7f332-5c39-4404-8cd5-2d104bc89f32 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.949765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-05T16:46:47.390216Z digest=sha256:7c838edb1183ea7b5c9ac1054cacf4fa8ba5318b925d9dae4a429e6a385c41c3

Observation aeff3fc1-8532-4454-a968-25da03344a4f · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:47.487403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:47.487403Z digest=sha256:d0529fbfecdd51754d16aa7804a030a311e280474f60a86f5c00b1eeb241f3a5

Observation 63cde880-590b-4f8a-a196-9b77100c7de0 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.941858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-05T16:46:47.555373Z digest=sha256:4fc5024d11396f30d6da1d93e0160ed33c38718c89d3c7bf261269d347aa80f7

Observation 59d58ead-f7e1-4f16-a889-aa6095d78d70 · outbound

This paper cites GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:47.691984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:47.691984Z digest=sha256:732b1f49eefc0d507c3b6f34d9d72a4b54707e349e6c42d406842a8cb9be31d1

Observation 8c6a3ded-de9b-4d0d-8255-efc9e2389adb · outbound

This paper cites LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:47.765382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:47.765382Z digest=sha256:32979555a2e304a9cd6465e3604016799251d53a440c4e771b85167219ed5280

Observation 52db9ce8-41ec-454d-8e0a-076ed64207bc · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.934533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-05T16:46:47.860324Z digest=sha256:bbfa5dfe466116eaee123d500fa13106e1c168cd2d9bf9f3848f5fcad3fcf30c

Observation 273e77fb-f566-4c2a-bd53-c60e976523b4 · outbound

This paper cites Qwen2-Audio Technical Report.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Qwen2-Audio Technical Report

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:47.924933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:47.924933Z digest=sha256:b611d98c165026a334ca2cdac79751d56a2be2b697c8b22d09a7cb7a55d28caf

Observation 251c40b2-4f2c-478d-b659-c45cef6a6c6c · outbound

This paper cites Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:48.021893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:48.021893Z digest=sha256:d8713b3f424f9a365783468902f1e1fe52d0ac500032dd024f07b7710d11e33f

Observation b90e7ae5-32b0-4904-aa3f-5e5b23cb3264 · outbound

This paper cites Recent Advances in Speech Language Models: A Survey.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Recent Advances in Speech Language Models: A Survey

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:48.070568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:48.070568Z digest=sha256:d3f73c1dfdd880878c628ca777b3dcb9757395daa1fc7e8ae88af98885295ed2

Observation 25d5a620-79c0-4275-94f5-f82068149967 · outbound

This paper cites High Fidelity Neural Audio Compression.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs High Fidelity Neural Audio Compression

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:48.093307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:48.093307Z digest=sha256:80c55863987b2db96b8bfb826ff31a5c74aec60f744ad78d413d0f5c72c03b96

Observation 53e548ca-4d3d-48ad-8491-64055f7569f8 · outbound

This paper cites Exploring the Benefits of Tokenization of Discrete Acoustic Units.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Exploring the Benefits of Tokenization of Discrete Acoustic Units

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-05T16:46:49.652664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-05T16:46:48.202777Z digest=sha256:fc2649a689ddd01f7a028f1e372a057f98186a17a56f134bb083b4259f78b926

Observation f513f72a-db5f-4106-bc4c-05b4d09d7087 · outbound

This paper cites The Llama 3 Herd of Models.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs The Llama 3 Herd of Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:48.356620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:48.356620Z digest=sha256:7baa78fd90ea2446a649c03112e63e2e8de37b2597b663f7dbe1a50cf8233047

Observation 2bba73a3-5426-4bf6-bedb-0330b9655176 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.926577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-05T16:46:48.486899Z digest=sha256:0e177d6ae62be041fb6df44f4ea166c0a7d5e3ee4a778ce55f7b9a5d97afaeae

Observation 60e763a5-eb26-4df5-868b-81e21fab4b52 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.918359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-05T16:46:48.635372Z digest=sha256:10aaefda3648b7d47b1e2d88f43125b147eb33c18c49ec130537fd3e83db9456

Observation 003dd5a5-a870-433e-8397-f892f8ba0cb1 · outbound

This paper cites Listen, Think, and Understand.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Listen, Think, and Understand

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:48.796047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:48.796047Z digest=sha256:ff67aa4ee601ec990c69d8595ab7fbde1bcec3a1ed3f413672b09c13f3a10207

Observation 9ea639e1-8de5-4d71-86c4-f09baeca2449 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.909209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-05T16:46:48.963009Z digest=sha256:d84d1af1c91ae466701eb38488c200ae291a88ab0c9562d507a2cfa02f153136

Observation 8e34093a-3e8c-481c-93d0-767033c9d995 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs LoRA: Low-Rank Adaptation of Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.182080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.182080Z digest=sha256:b29cc7aae445ea7877f18e048eb12f1a37d102a3007754ec17739330046925b0

Observation 5288e3c4-2220-4f74-bd8a-ccf2e787a69d · outbound

This paper cites WavChat: A Survey of Spoken Dialogue Models.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs WavChat: A Survey of Spoken Dialogue Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.184979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.184979Z digest=sha256:e94e1382483c66f5a488f98c9952c311272075c95e01909f4ca790bfb653e550

Observation 127383d8-3888-406f-bf51-3a37cc4d4d52 · outbound

This paper cites WavTokenizer: an Efficient Acoustic Discrete Codec Tokenizer for Audio Language Modeling.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs WavTokenizer: an Efficient Acoustic Discrete Codec Tokenizer for Audio Language Modeling

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.187588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.187588Z digest=sha256:6747a6365a8a323e074f2bfec1af7d43464b6785adc1e4a5fa90b9d3c7b7f90d

Observation 950e6ffd-d212-4326-9577-37c7267beefe · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.190729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.190729Z digest=sha256:a778272084f6eeafd44d41b2990fc229c7f1614055555d7ae92208cdb0baa0ea

Observation 8a3261b7-ad91-47e5-9be0-886189330964 · outbound

This paper cites Continuous Speech Tokenizer in Text To Speech.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Continuous Speech Tokenizer in Text To Speech

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-05T16:46:49.602331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.193445Z digest=sha256:3b08377e60a95b5e94dc1418834911768ae1004c72fc2a0aa57b277a27b81438

Observation ca7456a1-cfec-42dc-8ed9-fd7c83588c02 · outbound

This paper cites An Embarrassingly Simple Approach for LLM with Strong ASR Capacity.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs An Embarrassingly Simple Approach for LLM with Strong ASR Capacity

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.196106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.196106Z digest=sha256:2cd4ba491a61b36a681ede72806416cf073fce4bc2d6400f2cd0cac38f511021

Observation 288f89b5-1b4f-4b77-a6be-f36784bd7c75 · outbound

This paper cites PSLM: Parallel Generation of Text and Speech with LLMs for Low-Latency Spoken Dialogue Systems.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs PSLM: Parallel Generation of Text and Speech with LLMs for Low-Latency Spoken Dialogue Systems

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-08-05T16:46:49.581631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.199048Z digest=sha256:e1d84eb6477bfcc3159691a1bd22ed3a3dba088ac887eacb575a3abdc142e728

Observation 7dcb80e2-5104-49ea-bc20-9dc7a316d8d6 · outbound

This paper cites How Should We Extract Discrete Audio Tokens from Self-Supervised Models?.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs How Should We Extract Discrete Audio Tokens from Self-Supervised Models?

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.201803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.201803Z digest=sha256:e1564fcc6ea9a330591f13c9d110c2c7a3437e37bcbc538b29c4f8dda2de3645

Observation 36df5c38-9ff3-478f-b9d6-01f7a273e5bc · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.900264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.204517Z digest=sha256:69186d03e64e42e5521d183394f8ce4457feefb0d4dc447971f9197f395adda2

Observation bd0e828b-c896-4ba8-9bdb-1f81344a04c7 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.892093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.206910Z digest=sha256:bfc0f2b1973e7595ceb126f254ee87dfa0885013950e6511fe8a34d0ff9bdfd5

Observation 6296b8dd-7b47-4047-8c23-5c862a9ecd22 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.883742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.209410Z digest=sha256:a936fa02c84ed1a92eeb2c25801b343024562696d59568192ee608502877c0f7

Observation 49a1d482-f3a4-417d-8dde-03d944f1e11b · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.211813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.211813Z digest=sha256:a34a3d23057287fdecde94610dc610c894b4eefee73b9993524e9d7cf7217eeb

Observation e55f726f-0bc6-463b-b818-abeb078d964f · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.875522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.214337Z digest=sha256:996f31afa19028ef7c8125a185e24467bfd10eb905bac9ecc84883f56b256505

Observation 6b3b7ae4-e59d-4345-a001-4097b9b36a86 · outbound

This paper cites AudioPaLM: A Large Language Model That Can Speak and Listen.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs AudioPaLM: A Large Language Model That Can Speak and Listen

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.216718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.216718Z digest=sha256:76221b524094eb4c49b75685bfa4c6432117cc64147846b746cc8e3fbe8b2035

Observation 871b63f0-a24a-474b-8214-bd978d1f0d83 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.867337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.220533Z digest=sha256:9e87d046f4bc029e5d2ddb114e70f875ab5399f12e38e4482d84c40162f64b6f

Observation c582b0e4-44dc-4def-87c9-84d8c2620ae5 · outbound

This paper cites DiscreteSLU: A Large Language Model with Self-Supervised Discrete Speech Units for Spoken Language Understanding.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs DiscreteSLU: A Large Language Model with Self-Supervised Discrete Speech Units for Spoken Language Understanding

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-05T16:46:49.441586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.223732Z digest=sha256:9c1bb8fbe338f12880a87d25f1fae5a7523ed46de43e29aab854064b865cc136

Observation 4073ece4-78b8-40ac-baa2-e52fee4a7071 · outbound

This paper cites SALMONN: Towards Generic Hearing Abilities for Large Language Models.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs SALMONN: Towards Generic Hearing Abilities for Large Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.226321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.226321Z digest=sha256:e82469ece22bffb2cd51e7134e2ce83bdab8969f07a4e3d5bd5ddefed28acc76

Observation 7a2eb8ed-93bf-4b58-9b78-14df767b8381 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.858172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.229228Z digest=sha256:0597ccdf087fbad55672bc6f03ecde8932dbc653976eaf50fda0af2942c67196

Observation 58daee93-3456-4c1e-a19a-d8a94f8dfe07 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.849882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.231950Z digest=sha256:daaf4ee914845b0ca3cf1a8bbaf22e835ac7ba3e5cc3eded2eae604bd167a7c3

Observation c278616e-00d3-432b-aeb3-9c587e7b02bb · outbound

This paper cites BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.234700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.234700Z digest=sha256:04071f8aeb9322dabb5dcdbcb0795d542d0ac50d141a6438b1e69e972e932c7e

Observation b3502b97-07a2-4492-a9f7-d06e1864cb13 · outbound

This paper cites InSerter: Speech Instruction Following with Unsupervised Interleaved Pre-training.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs InSerter: Speech Instruction Following with Unsupervised Interleaved Pre-training

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-08-05T16:46:49.412149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.237800Z digest=sha256:a83233f82da6e87da57f3c45b14ecbef06397faeca6994140f1cce227d8c20d1

Observation 82e7c989-1115-41fd-9d0b-10bca00a2b5c · outbound

This paper cites VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.240587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.240587Z digest=sha256:887123bb2991b4194597fb40b7a5b083aa8eb5eba80be39d8dd56be293351cde

Observation 31db02af-61fd-466c-b863-13297a7e3d11 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.841549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.243540Z digest=sha256:d2ee0e8cde1b97a021491ce075d55bd2bea758b44a6139e6fe7232e21564da72

Observation 1e47ca38-e878-4336-b0eb-ce153a8f1336 · outbound

This paper cites Speech Commands: A Dataset for Limited-Vocabulary Speech Recognition.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Speech Commands: A Dataset for Limited-Vocabulary Speech Recognition

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.246073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.246073Z digest=sha256:1f34a2f6375af4bc438f339dcbab13dd7fabb0a29bd3bb8f491419ac7597a56b

Observation 794ca939-5f0f-460b-8cb1-be4930fd03a9 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.832557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.248919Z digest=sha256:372ee8bfb62e250bd9f99153c642ad831192e8730535db1fbf8e7069c5444b77

Observation 60f263ac-2bba-43eb-8519-23d2f46f4e95 · outbound

This paper cites Mini-Omni2: Towards Open-source GPT-4o with Vision, Speech and Duplex Capabilities.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Mini-Omni2: Towards Open-source GPT-4o with Vision, Speech and Duplex Capabilities

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.251449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.251449Z digest=sha256:d82cd73c57351cad910384c23b75546ce449970836160ca28e9c59001fc468b9

Observation 5b315595-faf4-4481-9962-1b2c55c9451b · outbound

This paper cites Qwen2.5-Omni Technical Report.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Qwen2.5-Omni Technical Report

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.254645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.254645Z digest=sha256:e7712c810506433087f2e26ea3a48eada05fea0e5f46251ba3d302404af15f97

Observation b1a419f0-0c28-45fa-8f38-e6035255882b · outbound

This paper cites Comparing Discrete and Continuous Space LLMs for Speech Recognition.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Comparing Discrete and Continuous Space LLMs for Speech Recognition

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.257437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.257437Z digest=sha256:4673b80434d27b32c7555719f563d9a8a0777f0f8865af48a10ac7d107f5052e

Observation edfe98b9-dc87-4e50-9f77-c65c579c2dbf · outbound

This paper cites ALMTokenizer: A Low-bitrate and Semantic-rich Audio Codec Tokenizer for Audio Language Modeling.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs ALMTokenizer: A Low-bitrate and Semantic-rich Audio Codec Tokenizer for Audio Language Modeling

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.260250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.260250Z digest=sha256:2ea0c6feb2f045a40361154ee7189b5ea14b0f5cab108080fd39e58d971fbbf5

Observation 144c10e5-3bbb-4014-b9ec-f8b7e4a9d11e · outbound

This paper cites GigaST: A 10,000-hour Pseudo Speech Translation Corpus.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs GigaST: A 10,000-hour Pseudo Speech Translation Corpus

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.262707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.262707Z digest=sha256:5c45401c140b8816f9c5d8c53ffa1ca0ad91f94938ed9794cc662084493b0377

Observation a7fb8689-7e29-4934-9057-0f440c292082 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.265779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.265779Z digest=sha256:d91e40a0b3860198506c9b92d2a372154e4f501c142922db55ca7083419684e5

Observation e0daf9c9-2cad-4e82-9a27-23774e80b77b · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.823915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.268492Z digest=sha256:9ba3f2188427855306e724d339933a7d640389f80893585c06852dd2754534e7

Observation 494656bc-0d91-41ad-b14b-d49cb69a3022 · outbound

This paper cites GLM-4-Voice: Towards Intelligent and Human-Like End-to-End Spoken Chatbot.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs GLM-4-Voice: Towards Intelligent and Human-Like End-to-End Spoken Chatbot

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.270919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.270919Z digest=sha256:fe27c96573df996aa8a424298d92a680732d6f5d4e2fbf65ce65008d45f07756

Observation fcec29af-ec7c-4da5-a927-e41fc3fa48d8 · outbound

This paper cites SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.273789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.273789Z digest=sha256:aa4eaafeb0d310ac56f684183494052a0b6018e0114a362037913e73bd70a384

Observation e89c1e01-398a-4a9b-9a5f-14c3a259b39d · outbound

This paper cites SpeechTokenizer: Unified Speech Tokenizer for Speech Large Language Models.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs SpeechTokenizer: Unified Speech Tokenizer for Speech Large Language Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.276513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.276513Z digest=sha256:af05d0af5c4e0c4d8a112b2288154a53906a50931ecc53f53903f89d081e6895

Observation 5d45cb9e-8915-4788-a5b7-956688efa298 · outbound

This paper cites online" 'onlinestring :=.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs online" 'onlinestring :=

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.279342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.279342Z digest=sha256:d447d9c829f0ee36929e863d8bb29aee6eab0ccbe717bc645f81c87d85efcdc7

Observation 446a4bcb-cf18-4417-b33f-cc99e74a81ab · outbound

This paper cites write newline.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs write newline

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.282451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.282451Z digest=sha256:d5622f08d28eaaf4f9410e5848590be01f9dcf0b98f43f28ec78373557d4478f

Pith citing papers

Observation 5e0eb325-70d6-4597-a064-93b3bef5b1ae · inbound

Towards Building Speech Large Language Models for Multitask Understanding in Low-Resource Languages cites this paper.

Towards Building Speech Large Language Models for Multitask Understanding in Low-Resource Languages Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-18T16:31:37.222863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T16:27:37.596817Z digest=sha256:c7382da5b7dac8d76932127bcf81d0fe3f77234de93c881b8a0026844dd51d5c

Observation 3667a07c-94e2-4be2-8936-8432bcc774f9 · inbound

Speech Meets ELF: Audio Conditional Continuous-Target Diffusion for Speech Recognition and Translation cites this paper.

Speech Meets ELF: Audio Conditional Continuous-Target Diffusion for Speech Recognition and Translation Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-03T07:27:44.922554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-27T12:04:50.483329Z digest=sha256:e8c5a22d7ce6d07bd68594dc0332c52fadbdd4ba57b8431cca65fcc591bed845

Observation 89a5f7bf-cfe4-4d3e-be39-55c2f5ca66a6 · inbound

HybridCodec: Modeling Discrete and Continuous Representations for Efficient Speech Language Models cites this paper.

HybridCodec: Modeling Discrete and Continuous Representations for Efficient Speech Language Models Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-06-29T01:02:56.248185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T00:56:43.991936Z digest=sha256:5ed6aaf439d8b4ec6de6048113fd5e549064c053b0062c04ea0e39645ea69477