Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T20:08:58.582356Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:2508.11187.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T20:08:58.582356Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
44 of 44 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2987e7b9-c65c-420f-b679-ceba3ebe440a · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Retrieval and browsing of spoken content,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation bf2f4c16-bcd1-4ed3-8686-cffeca2a4677 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style V oice-based information retrieval—how far are we from the text-based information retrieval?
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b7769f79-5289-4bb2-a1d8-24ac5c0951d7 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Spoken content retrieval: A survey of techniques and technologies,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0460dde0-cf7d-4deb-b65a-92845e325e4e · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Spoken content retrieval—beyond cascading speech recognition with text retrieval,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3a8108d8-5251-4262-924a-7044d05beca3 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style SpeechDPR: End-to-end spoken passage retrieval for open-domain spoken question answering,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3447d916-4f3c-4c70-8c79-e20aedddd086 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Retrieval augmented end-to-end spoken dialog models,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 88292e66-930c-407c-b267-94e6aee151af · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style WavRAG: Audio-Integrated Retrieval Augmented Generation for Spoken Dialogue Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbbe176b-4fc8-4de3-acf0-0507ebdf67f6 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Speech retrieval-augmented generation without automatic speech recognition,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 048add9f-8558-4874-90a5-8277fe7b6ca7 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Prompting audios using acoustic properties for emotion representation,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fd5d1968-4cc0-4b3e-b386-f081012e6276 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Large-scale contrastive language-audio pretraining with feature fusion and keyword-to-caption augmentation,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9409ad26-5b5a-419b-a9d6-2d18f66845ab · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style IEMOCAP: Interactive emotional dyadic motion capture database,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2b00b25b-503a-462f-b924-f2987c8e4824 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Emotional voice conversion: Theory, databases and ESD,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a631dcd0-1d34-4f75-9465-1b2af2119817 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Expresso: A benchmark and analysis of discrete expressive speech resynthesis,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 383da48d-113c-467a-9679-9cf71a7c0c63 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style PromptTTS: Controllable text-to-speech with text descriptions,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4f6c3a39-8f78-4e9d-8f30-8295154507e9 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style PromptStyle: Controllable style transfer for text-to-speech with natural language descriptions,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cefdac49-bc5f-4d28-b2f0-5f0bbdc2d1c4 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style PromptTTS 2: Describing and generating voices with text prompt,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4ca6dce2-e074-4916-aa4f-9695a3d41165 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style DreamV oice: Text-guided voice conversion,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a5f17ddf-3b7c-43e4-be54-ad62714dc880 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style StyleCap: Automatic speaking-style captioning from speech based on speech and language self-supervised learning models,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d516af7a-56b6-4212-a7c5-7033bc9a7978 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style LibriTTS-P: A corpus with speaking style and speaker identity prompts for text-to-speech and style captioning,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 96f0f9f0-92d8-460f-97fd-6d22f5e94e2b · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style V ocabulary independent spoken term detection,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7c8bf58c-9105-4f0e-84ba-6ff34ea07034 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Statistical lattice-based spoken document retrieval,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e7af4a0a-085a-48da-b017-70bc70675d86 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Improved semantic retrieval of spoken content by document/query expansion with random walk over acoustic similarity graphs,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 35a12892-4b63-4ab0-a6fd-87b1bbd05521 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Evaluating ASR output for information retrieval,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 89d11070-717c-47cf-8bf1-3df3ad898e04 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Investigating the global semantic impact of speech recognition error on spoken content collections,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 166dafa7-2cb9-4188-9026-d506fbb459bc · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Speech-centric information processing: An optimization-oriented approach,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e1876a5c-0e8c-4fef-b17a-322ba56511c9 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Unsupervised spoken-term detection with spoken queries using segment-based dynamic time warping
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a1b05aaa-b43b-41c9-b51d-705d123fa730 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Memory efficient subsequence dtw for query-by-example spoken term detection,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3cddcd3e-154d-49d2-bdde-aaa3697cb39e · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style The spoken web search task at MediaEval 2012,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 29f2ff46-011e-4f9a-8824-b0d76f1b53e0 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style RECAP: Retrieval-augmented audio captioning,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 710f025a-c68c-43a0-a5c5-7e05f52846b8 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Beyond speaker identity: Text guided target speech extraction,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4ce8b136-65d5-40bd-b6db-cb39af5900cd · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style WavLM: Large-scale self-supervised pre- training for full stack speech processing,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02b9ac51-317e-4829-b155-f89905685ddb · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style emo- tion2vec: Self-supervised pre-training for speech emotion representation,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 448eae5e-275a-4316-9d96-bf266f4d046b · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style BERT: Pre- training of deep bidirectional transformers for language understanding,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 79accd8c-d82c-4429-9aa6-6ce3e18e6b2b · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style RoBERTa: A Robustly Optimized BERT Pretraining Approach
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d53a25fb-05a3-405c-8373-612ac569b576 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Exploring the limits of transfer learning with a unified text-to-text transformer,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f681b8d2-59e9-4d9b-9092-8f3d21cf3520 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style The flan collection: Designing data and methods for effective instruction tuning,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9b112b2f-79a3-4cd5-a356-4d8360b9f105 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Sentence-T5: Scalable sentence encoders from pre-trained text-to-text models,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 83195699-1641-47d1-af58-b3e32cbbb861 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Domain-adversarial training of neural networks,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 24279dca-a0f0-444d-b756-51de2a8d3365 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Unsupervised domain adaptation by backpropagation,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b802835a-261b-4e40-bf7e-a3b5a91fb24e · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style GPT-4o System Card
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96bdecf8-9edd-4b16-a788-e8cf996cd90e · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Decoupled weight decay regularization,
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5870a2b-bf04-496f-b0cd-2d8a5fdf736e · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style PromptTTS++: Controlling speaker identity in prompt-based text-to-speech using natural language descriptions,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c3b5de4c-bd7f-4b68-b677-f8ce08f3a5e3 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style PromotiCon: Prompt- based emotion controllable text-to-speech via prompt generation and matching,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f6bc0f0e-80fd-42db-9cf4-d6fef5e76115 · outbound
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Visualizing data using t-SNE,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
No inbound Pith citation observations are available.