Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:58:44.194819Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 1 inbound Pith citation observation for arXiv:2506.06343.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:58:44.194819Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T05:50:33.363981Z
A source-named dated measurement, never combined with another source.
Source: cited_works
29 of 29 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c806618a-d343-424a-993a-e6ff54626dfd · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment Moshi: a speech-text foundation model for real-time dialogue
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb7aac0f-dd7d-4a7b-b1c4-2e3a9c3bc669 · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a70dcad-3b11-496f-98e0-1fda8883fab2 · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment Qwen2-Audio Technical Report
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 631d0733-5633-44b9-8519-626565e9a63e · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment Baichuan-omni-1.5 technical report,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 076e0b54-13c8-48fa-aa92-1a3b5b45342a · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment LLaMA-Omni: Seamless Speech Interaction with Large Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e806ba7-755f-4f05-b107-ee54be2e91b8 · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment Freeze-Omni: A Smart and Low Latency Speech-to-speech Dialogue Model with Frozen LLM
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 123184ec-ddfa-45d5-854a-097c2d554f96 · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bed4a7ab-c49e-4df2-8453-db75df57dc7e · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment InSerter: Speech Instruction Following with Unsupervised Interleaved Pre-training
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fec8c50-3c76-4ff4-837a-7c90a0fa5985 · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment Distilling an End-to-End Voice Assistant Without Instruction Training Data
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6df42d5f-d314-4609-a221-5560619ae578 · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment Speechless: Speech Instruction Training Without Speech for Low Resource Languages
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4e0f11e2-22a8-4db6-94d5-26769db66329 · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment SpeechT5: Unified-Modal Encoder-Decoder Pre-Training for Spoken Language Processing
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6c26aea-4ace-4858-abee-3cebbce44cf4 · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment SLAM: A Unified Encoder for Speech and Language Modeling via Speech-Text Joint Pre-Training
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 073b2990-7cce-476c-be3a-76ec126d729b · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment mSLAM: Massively multilingual joint pre-training for speech and text
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cc83ce2-6990-4d94-9e94-e5b4ad163fc3 · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment SeamlessM4T: Massively Multilingual & Multimodal Machine Translation
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1f7fc13-e024-4ffa-882b-2ba265f37872 · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment No Language Left Behind: Scaling Human-Centered Machine Translation
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e7d927f-2eff-4969-982b-308d2f6ec336 · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment W2v-bert: Combining contrastive learning and masked language mod- eling for self-supervised speech pre-training,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3c87180-6074-434f-bc2e-2c9d23dd2acd · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment VoiceBench: Benchmarking LLM-Based Voice Assistants
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ad31277-d93e-4068-a1cd-f9b30221ff29 · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment GLM-4-Voice: Towards Intelligent and Human-Like End-to-End Spoken Chatbot
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87aa7824-7b99-471b-86b6-de5acded7abb · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment Paralinguistics-Aware Speech-Empowered Large Language Models for Natural Conversation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8818343-d73f-4f41-966a-38e8a1e32526 · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment Spirit-lm: Interleaved spoken and written language model,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cb7bff6c-c7a8-4840-8086-89d7b676e829 · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment Dissecting learning and forgetting in language model finetuning,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 12c6ad8f-6c7b-484d-8d59-4e2841806b24 · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment The Llama 3 Herd of Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3aa62cf-f587-4f1a-a40b-15fce72d20b0 · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment Openwebtext corpus,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 09a79cd4-a638-4dfc-a433-35c79d77fbd4 · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment Enhancing chat language models by scaling high-quality instructional conversations,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 46b02533-0f52-45e5-9613-d44a4c88cc04 · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment Openhermes 2.5: An open dataset of synthetic data for generalist llm assistants,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b10ec70f-cf80-4788-8a44-f29e9e14ff3e · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment Can a suit of armor conduct electricity? a new dataset for open book question answering,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfadb9da-0797-43a0-b1b7-f9913edbc643 · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment CommonsenseQA: A question answering challenge targeting commonsense knowledge,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bb23e8f6-66ff-44a6-8976-98ff6d1c6755 · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment Librispeech: an asr corpus based on public domain audio books,
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a8231c3-cc6a-4139-afe9-79d01fc0f81a · outbound
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment CoVoST 2 and Massively Multilingual Speech-to-Text Translation
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1302492f-f91d-49b9-82bd-e80a90afee7b · inbound
MEUSLI: a Multilingual Projector for LLM-based ASR and Beyond TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.