Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-07T14:53:07.512543Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 0 inbound Pith citation observations for arXiv:2607.05365.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-07T14:53:07.512543Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
55 of 55 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation dc7450df-764d-4b75-a05a-f4c73542b24b · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Large language models: a survey of their development, capabilities, and applications
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c0214adf-85b2-4627-8455-f5efab4214b6 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models On The Landscape of Spoken Language Models: A Comprehensive Survey
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 427e98d8-9ab9-4078-a8d8-bb828dbff0cf · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models A full-duplex speech dialogue scheme based on large language model,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 3c52f983-cf20-40c2-8b83-86ae60aaa1d8 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Human conversational behavior,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f059b1cd-71b6-457b-98ae-4e4e97b2114f · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models 6 conversation analytic approaches to the relevance and uses of relationship categories in interaction,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 433d7776-909d-433f-87ed-3cbac7bbcdd1 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models τ-voice: Benchmarking full-duplex voice agents on real-world domains
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6e032512-eb4d-4e5c-b6e0-4afe4a7c4b3c · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Wildspeech-bench: Benchmarking end-to-end speechllms in the wild
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 53d58aa6-b453-4969-a737-34720ed263dd · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Full-duplex-bench v1. 5: Evaluating overlap handling for full-duplex speech models,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 773c2f82-c687-45cb-bfac-10ecfa0f2a43 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Seamless Interaction: Dyadic Audiovisual Motion Modeling and Large-Scale Dataset
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 958289ae-5a6d-414d-a0a9-a99bbd810356 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models TRACE: Temporal Relationship-Aware Conversational Entrainment Detection in Dyadic Speech
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 20d93fba-99d0-44ec-882d-0df67967d16a · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models The embedded deformation problem for monomial ideals
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6145d9e5-6b24-4931-af5b-6192f990cacd · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models TurnNat: Automatic Evaluation of Turn-Taking Naturalness in Dyadic Spoken Dialogue
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 246f4cc5-1e1a-46d1-9a95-d50bd62fcb29 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Vox-Profile: A Speech Foundation Model Benchmark for Characterizing Diverse Speaker and Speech Traits
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 14813de2-cbdc-4dbe-bd0d-51dbad3be517 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models V oxlect: A speech foundation model benchmark for modeling dialects and regional languages around the globe,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation bacf8d78-19bc-4613-8e4d-05e873f6dd5f · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Reference-based prosody and rhythm evaluation for spoken dialogue systems,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 2a3ba648-7f3d-461d-8de8-5a2e3c0b2b57 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Gpt-4 technical report,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 37c4a746-c714-4e72-b644-0e91bdc7d1e9 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Qwen3-Omni Technical Report
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1f68a723-5ecc-486a-8b52-c87389211192 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Gemini: A Family of Highly Capable Multimodal Models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e6451e05-4307-47ec-9477-f9e7730f820b · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models SpeechTokenizer: Unified Speech Tokenizer for Speech Large Language Models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 68ee057e-8f93-4c68-8492-a4431557e16e · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Whisper-AT: Noise-Robust Automatic Speech Recognizers are Also Strong General Audio Event Taggers
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1b25f448-0516-4ce6-9eb5-5b22cf2f06ad · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Qwen3-ASR Technical Report
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e4a0d144-2bc2-425e-a32b-fcc7328db780 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Hubert: Self-supervised speech representation learning by masked prediction of hidden units
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation ebc32acf-4467-4bf3-9e49-bd6764b6407b · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6523c40a-596a-4e69-acc8-8482fa6de632 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Wavlm: Large-scale self-supervised pre- training for full stack speech processing
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6e4c7289-d80d-4a0b-906f-af201dded00d · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models WavLLM: Towards Robust and Adaptive Speech Large Language Model
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 5cee6b7c-d768-438a-b8c4-e136211ac1bd · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 2af54508-abda-4e58-9ccc-cea4c3721ea5 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models gpt-audio-1.5,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f8cfa9fe-64d6-45b9-b037-949d80bc8d77 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Qwen2.5-Omni Technical Report
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6875b2f9-7546-44c0-b64c-2ce42d4095d1 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models gpt-realtime-2,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 45f5a163-01fb-4a37-9d1c-a3b040895e44 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 2ef3b95b-e22e-4b52-b992-09178b36372c · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models gemini-3.1-flash-live-preview,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 672dc707-08c5-4bf4-879a-e306953c57e0 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Available: https://ai.google.dev/gemini-api/docs/models/ gemini-3.1-flash-live-preview
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 833cd489-af27-4686-a29a-c5550f69ea82 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Mini-Omni: Language Models Can Hear, Talk While Thinking in Streaming
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 3f0d91e3-f456-4abc-988f-0eb5097149a9 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Couper-Kuhlen and M
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7de617d4-6af4-465b-9871-cb130e888db0 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Phonetics and prosody in conversation,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 3ef9f11a-b666-45b0-8a15-289f9e09dd15 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Classifying conversational entrainment of speech behavior: An expanded framework and review,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6c539f27-3f84-48e2-b3b9-902682b1394e · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models A cross-linguistic analysis of the temporal dynamics of turn-taking cues using machine learning as a descriptive tool,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b178f409-8385-43b5-bf6b-4cb8bd75f504 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Real-time changes to social dynamics in human-robot turn-taking,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 2a5d849a-7bfb-46ab-a75d-972749ffa3a9 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Talking Turns: Benchmarking Audio Foundation Models on Turn-Taking Dynamics
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 58a2debd-d7d6-4ee8-a37d-569b059d42b7 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models UTMOS: UTokyo-SaruLab System for VoiceMOS Challenge 2022
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7963645b-9cd7-4068-b5d1-df5bcd427f8c · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models SUPERB: Speech processing Universal PERformance Benchmark
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 72874aac-fc2e-4d92-a913-36525bedd6ec · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Wildbench: Benchmarking llms with chal- lenging tasks from real users in the wild,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation df14cba5-6600-43aa-8d0a-03da7266391e · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Mt-bench-101: A fine-grained benchmark for evaluat- ing large language models in multi-turn dialogues,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 3fd36e6a-1423-48c6-be2f-d35fc0a42f88 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Air-bench: Benchmarking large audio-language models via generative comprehension
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation eb11fc14-eb85-4ad7-9e72-e1cd047fd8f4 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models GlobeAudio: A Multilingual Multicultural Benchmark for Naturalistic Evaluation of Large Audio-Language Models
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 95d45456-7a28-49fb-8a71-4bc749c5846a · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models A systematic survey and critical review on evaluating large language models: Challenges, limitations, and recommendations,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 19bc96f1-8ff8-4314-8945-1e6bd1df7318 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Flexi: Benchmarking full-duplex human-llm speech interaction
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 22847e6f-472d-47e5-ba87-f613595d12f8 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Speech to Speech AI Model & Provider Leaderboard,
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c171b547-75ff-440c-9627-1f245b46f0ea · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Speech Arena Leaderboard,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 4896ad97-aefb-460b-95fc-6f5c3b7af505 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Methodology — V oice Arena TTS Leaderboard,
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a5bc260d-d5cf-4da2-bf96-6f608301c6a2 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Robust speech recognition via large-scale weak supervi- sion
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e4d3f6f8-6931-484c-9920-270b1c54a3b3 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Silero vad: pre-trained enterprise-grade voice activity detec- tor (vad), number detector and language classifier
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7c7bcfee-af00-4e22-89d1-41686b8569ba · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Scaling speech technology to 1,000+ languages,
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 190783a9-186b-4d73-866f-6a86eab5e424 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Fleurs: Few-shot learning evaluation of universal representations of speech
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e4a23d02-3103-45fa-b218-a91b57672ee8 · outbound
SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
No inbound Pith citation observations are available.