Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:28:54.976320Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 100 of 165 outbound references and 3 inbound Pith citation observations for arXiv:2506.02443.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:28:54.976320Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-26T09:46:15.701501Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T09:39:46.872455Z
100 of 165 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c6c2304a-23ac-4914-ba28-4f55fc9e71c3 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Simultron: On-device simultaneous speech to speech translation
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 873835c3-1a3e-4f91-b08d-eda15d57db18 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Qwen 2.5: A comprehensive review of the leading resource-efficient llm with potentioal to surpass all competitors
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32a0ea4b-8888-4f91-83a8-5d35e1858e85 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Analysis of layer- wise training in direct speech to speech translation using bi-lstm
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a273406-4bde-489e-bdd9-71d54fcd955a · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Precipitation nowcasting with generative diffusion models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b8fe074-38c4-4a16-aaab-be8e14324918 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Mean-Field-Type Game Theory: Applica- tions, volume 2
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50a12541-1011-418a-890b-61774aaf52b1 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Unresolved cited work
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7990542-4be9-4a87-90b4-55da5b2b7ea4 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Listen and Translate: A Proof of Concept for End-to-End Speech-to-Text Translation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c201fa9-12a9-4fbd-933a-bcdce127e9ac · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Unresolved cited work
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c72b7324-91de-4b01-9e1e-f8fae1dd1ebd · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Large language models are strong audio-visual speech recognition learners
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82501d3c-2e97-4271-9597-0295d15f6637 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Low frame-rate speech codec: a codec designed for fast high-quality speech llm training and inference
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8355f08a-cb85-47f3-8dae-a183f17c899a · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI A speech-to-speech translation based interface for tourism
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 257d59fa-52f2-480f-942e-00eab9818c01 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Exploring in-context learning of textless speech language model for speech classification tasks
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffd1783b-d96b-45b3-875b-cd32e9e3f7cc · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Audio Large Language Models Can Be Descriptive Speech Quality Evaluators
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50a4e804-f797-4395-8160-ed86c96a8184 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Multi-modal generative ai: Multi-modal llm, diffusion and beyond
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfb96779-2c01-4c02-8f85-b926d2500d09 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI BLASER: A Text-Free Speech-to-Speech Translation Evaluation Metric
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e0a91960-f798-408d-9327-e419336192c6 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Opportunities and challenges of diffusion models for generative ai
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d123d686-eb6c-44ea-b2e1-1c3d8cda3ec5 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Speech-to-Speech Translation For A Real-world Unwritten Language
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 45848d7e-2c1f-42fb-ad7b-5181aebd560e · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Beyond Single-Audio: Advancing Multi-Audio Processing in Audio Large Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f85aa4a1-dd53-4ecd-855f-9464586307fb · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI MAVFlow: Preserving Paralinguistic Elements with Conditional Flow Matching for Zero-Shot AV2AV Multilingual Translation
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4cd6b87b-9d0a-4756-914e-c97a4f230078 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI V2sflow: Video-to-speech generation with speech decomposition and rectified flow
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed150bb4-fd36-4975-ba85-8a9f662e12e9 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Qwen2-Audio Technical Report
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70e5c870-ac9f-41f7-bfa6-0dd9bddde940 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4b26d0b-0a3d-4474-ab62-54688b122d44 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Diffusion models in vision: A survey
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 499a2de5-a6d6-49b1-9bbe-b68e66bb5ecb · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Exploring the Benefits of Tokenization of Discrete Acoustic Units
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aceb2984-3db4-4261-9616-02507bf01c93 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI ADIFF: Explaining audio difference using natural language
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab55b59d-fcab-4bd1-88ca-d70844f0a013 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI French- fulfulde textless and cascading speech translation: Towards a dual architecture
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1013065-7116-412b-8efe-45c6e9a37755 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Textless Speech-to-Speech Translation With Limited Parallel Data
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a17fd686-12f2-4ec2-b5a1-979ad80e7a4b · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI PolyVoice: Language Models for Speech to Speech Translation
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ed51a5e-b3e9-4b28-9310-5c372fc528a5 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4826969a-6b2d-429a-8547-47185ea11b6d · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Analyzing Speech Unit Selection for Textless Speech-to-Speech Translation
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0eac4a58-04d9-427d-bad9-8e49c446503e · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Enhancing expressivity transfer in textless speech-to-speech translation
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb15fd2f-4b1b-47a7-8885-81dd702209da · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Towards massive parallel corpus creation for hausa-to-english machine translation
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43061353-529b-4bd1-b9e2-4e95189b84e7 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Auditory-visual perception of speech
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91bd5a83-bffa-449d-ab01-199c8b972f30 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Cascade or direct speech translation? a case study
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5b44b4b-c15b-4aa9-b989-426f7342c4ab · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI CTC-based Non-autoregressive Textless Speech-to-Speech Translation
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5a56457-cccd-405c-a2f8-3d656195576b · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Can We Achieve High-quality Direct Speech-to-Speech Translation without Parallel Speech Data?
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aef21d98-ff4e-4d1a-b6bc-c81f4bd7ba12 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Generative learning of the solution of parametric partial differential equations using guided diffusion models and virtual observations
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2623a34d-a06c-4916-b297-75ecf9cc5921 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Unsupervised speech technology for low-resource languages
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0da42731-d665-45ed-a794-9f9946f11ffa · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Speech-to-speech translation
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e31a667-aa47-4940-9947-2d4e8bbd8fba · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Audio Dialogues: Dialogues dataset for audio and music understanding
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60b14585-6067-444f-a72e-8091a3e33557 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Multilingual Speech-to-Speech Translation into Multiple Target Languages
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8cd0254-62a1-47a9-9fda-2f42e927c5f1 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Joint audio and speech understanding
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39fa5909-c21d-46eb-a46f-2ee384003423 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Tibetan–chinese speech-to-speech translation based on dis- crete units
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95e3ae8a-f5a0-4b8e-b1b8-f29255c55b25 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Recent advances in discrete speech tokens: A review
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 666aa822-1a2a-4348-bae2-6e5fe90247a1 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Direct Speech-to-Speech Neural Machine Translation: A Survey
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2022b7aa-ae5c-427c-a578-058301bac9b2 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Onellm: One framework to align all modalities with language
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd866ec3-b451-408d-b0a0-8748559e606d · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Physics-inspired approaches in generative diffusion models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f355817d-0e17-4169-9cbd-8afe0eb53b49 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Exploring In-Context Learning of Textless Speech Language Model for Speech Classification Tasks
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fbc096ac-803e-475f-9254-687f59be384f · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Chain-of-thought prompting for speech translation
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a16b2aec-5c45-419e-bdd6-f59e89cb8817 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI TranSpeech: Speech-to-Speech Translation With Bilateral Perturbation
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 645479fc-6505-4d81-8dc5-c3e9f1e7c109 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Textless Acoustic Model with Self-Supervised Distillation for Noise-Robust Expressive Speech-to-Speech Translation
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4f7aa9c0-3d17-439d-b26a-4b34f3d7c969 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Massively multi- lingual forced aligner leveraging self-supervised discrete units
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 411f20dd-1332-4924-ac1f-3640152ea1a7 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI LibriS2S: A German-English Speech-to-Speech Translation Corpus
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b3c742b7-aa8e-4ea6-b906-76cf35b79278 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI WavChat: A Survey of Spoken Dialogue Models
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73b92d89-70e8-41c1-936a-f94529374e76 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Can Generative Geospatial Diffusion Models Excel as Discriminative Geospatial Foundation Models?
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b4fde17-57c6-4b7f-b51f-693b03ae020e · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Listra automatic speech translation: English to lingala case study
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8164b2a9-f011-4b13-8074-1a8b4d7bb00d · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Gdplan: Generative network planning via graph diffusion model
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffa65f22-3545-4468-a606-d1383c7257be · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Direct Punjabi to English speech translation using discrete units
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 47a8994c-a3da-440a-93fb-7ce5beed79f2 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Textless unit-to-unit training for many- to-many multilingual speech-to-speech translation
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1df9749-84f9-43fa-a302-e888b5d5135b · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Phi dm-dialog: an experimental speech-to-speech dialog translation system
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1bcde44-36e0-4eca-ae94-34ec2c0fdf4f · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI High-Fidelity Simultaneous Speech-To-Speech Translation
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee857ae4-18f4-4465-91a5-e7e86832f1ea · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Janus-iii: Speech-to-speech translation in multiple languages
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a544d151-2432-4379-afe7-49bd57a39dac · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Textless Speech-to-Speech Translation on Real Data
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bd8c795-7df6-4383-9fdd-0cfa0c5444aa · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Video diffusion models are strong video inpainter
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f212478-dc0f-4540-8470-5ec1cb2360fd · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Speech proportion and accuracy in simultaneous interpretation from english into korean
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5462798-f80a-43c7-a9e3-edd0a298c7b6 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Diffusion models for audio restoration: A review [special issue on model-based and data-driven audio signal processing]
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2bfdd863-1ecb-41df-8902-eb01fd97d21a · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Conditional diffusion model for missing value imputation
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69f88707-5b02-41bd-b2e2-f5ae64cbdcff · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI BrainECHO: Semantic Brain Signal Decoding through Vector-Quantized Spectrogram Reconstruction for Whisper-Enhanced Text Generation
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a861450a-14e7-4ae6-b631-4ebdec81cd59 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7976c58-03ca-45cc-b81f-0ef8ea4a3323 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Textless direct speech-to-speech translation with discrete speech representation
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 396661f3-0e18-4ef1-a595-31a025022a83 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Beyond Words: AuralLLM and SignMST-C for Sign Language Production and Bidirectional Accessibility
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ccef7ffc-1061-4497-8b00-b6dc666de485 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Align-SLM: Textless Spoken Language Models with Reinforcement Learning from AI Feedback
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70b4be9b-ed03-4861-b1d3-1846e98fd97e · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Unresolved cited work
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79af8f6c-b061-4ac3-88c1-3a856ac09d70 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Handdiffuse: generative controllers for two-hand interactions via diffusion models
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0222dac1-65b2-4a42-b3f4-ba5a3b496eda · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI A Preliminary Exploration with GPT-4o Voice Mode
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 550e152a-e5dc-42bc-8b1f-74d8a62b5a70 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Recent highlights in multilingual and multimodal speech translation
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a738617c-b7fd-458e-825c-4d7de2be9be0 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Speech-to-speech low-resource translation
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b92ec96-557f-43f8-a985-c861c919bd5f · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Listening and seeing again: Generative error correction for audio-visual speech recognition
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 497f7294-6426-4c88-8ee4-a722d4b124f6 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI SLIDE: Integrating Speech Language Model with LLM for Spontaneous Spoken Dialogue Generation
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51da33f3-43af-4f58-b49d-4360c75d864c · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Llamapartialspoof: An llm-driven fake speech dataset simulating disinformation generation
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8a008de-eb3f-46d5-a025-6441456e2a81 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Build llm-based zero-shot streaming tts system with cosyvoice
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7da2b580-0764-4db3-813a-999550cfc9a4 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Auto-avsr: Audio-visual speech recognition with automatic labels
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7fa3d4a9-2460-499f-a34e-4f104d4c58c9 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Real-Time Textless Dialogue Generation
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3cb01864-40b8-4d19-ae41-18a246831544 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Slamming: Training a Speech Language Model on One GPU in a Day
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 809a6685-b873-4b14-9281-a9ece5f55ba6 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f53f924c-240e-4287-b891-99000f3eec12 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Make some noise: Towards llm audio reasoning and generation using sound tokens
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e180aeb-c0ca-4e72-9646-087233593a67 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Deep networks as denoising algorithms: Sample-efficient learning of diffusion models in high-dimensional graphical models
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4e61795-1735-49f3-b201-3cbf5270002a · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Amharic speech recognition for speech translation
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85a5e050-475a-4d8f-a090-e283bb7612b0 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Parrot: Autoregressive spoken dialogue language modeling with decoder-only transform- ers
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation beb0db37-1d76-4902-a215-7357bd88f8eb · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Towards to a direct speech to speech for endangered languages in africa
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a87d303b-cef5-4093-a2d6-25f6d4b7597b · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI A Unit-based System and Dataset for Expressive Direct Speech-to-Speech Translation
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5d30e757-ba22-422a-b20a-df145a1b26f0 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Spoken Question Answering and Speech Continuation Using Spectrogram-Powered LLM
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 404457b9-4d9e-466b-89c6-ef2fde59496c · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Towards real-time multilingual multimodal speech-to-speech translation
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29d563db-e77d-4365-8701-2432871951b3 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI One Model, Many Languages: Meta-learning for Multilingual Text-to-Speech
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6f82a242-5629-4085-b9f6-6be3c3184587 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Spoken Language Modeling from Raw Audio
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a296a4c6-ff87-4be2-ac85-65a5381ba322 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Visually Grounded Speech Models for Low-resource Languages and Cognitive Modelling
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 04064344-4acb-46fc-8596-1a702b74407a · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Verbmobil: The use of prosody in the linguistic components of a speech understanding system
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34d099fb-c96f-4cec-85e7-4e0ec50d7b47 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Phonology-Guided Speech-to-Speech Translation for African Languages
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 899aff8c-87f0-48a0-b348-c1531fdc2af2 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Let's Go Real Talk: Spoken Dialogue Model for Face-to-Face Conversation
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69f5a34b-30cf-4b75-8c5b-9544d342ba84 · outbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Long-Form Speech Generation with Spoken Language Models
Reference 100
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b97f05c3-f9ac-40d9-8c1e-c8c2a4cb0a72 · inbound
Achieving Generational Peace in Mali through Intergenerational Mean-Field-Type Game-based Incentives Breaking the Barriers of Text-Hungry and Audio-Deficient AI
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e83bdcdb-5e9f-4667-9220-e8c24470eaa5 · inbound
Mitigating Polycentric Conflict-Trap Risk in Mali via Intergenerational Volterra Mean-Field-Type Games Breaking the Barriers of Text-Hungry and Audio-Deficient AI
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f7b6c78e-a369-4e67-867a-2a4f67fa9319 · inbound
Risk-Aware Information Theory Breaking the Barriers of Text-Hungry and Audio-Deficient AI
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.