Pith. sign in

Paper Citation Record · LEDGER

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style

As of 22 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:2508.11187.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.11187 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:08:58.582356Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

44 of 44 outbound references displayed

  • verified exact0
  • verified fuzzy39
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2987e7b9-c65c-420f-b679-ceba3ebe440a · outbound

This paper cites Retrieval and browsing of spoken content,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Retrieval and browsing of spoken content,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:08.591689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:53.581070Z digest=sha256:678f9c44865055cf96024c89b44bae776f2b70ac9521c359ea2e83618066cceb

Observation bf2f4c16-bcd1-4ed3-8686-cffeca2a4677 · outbound

This paper cites V oice-based information retrieval—how far are we from the text-based information retrieval?.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style V oice-based information retrieval—how far are we from the text-based information retrieval?

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:08.410347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:53.693568Z digest=sha256:808f9d39a157ed53deb9d3f25c6bb178234344a351e05d032a7a340e72363c0f

Observation b7769f79-5289-4bb2-a1d8-24ac5c0951d7 · outbound

This paper cites Spoken content retrieval: A survey of techniques and technologies,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Spoken content retrieval: A survey of techniques and technologies,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:08.272085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:53.819901Z digest=sha256:7e3668f128e86e4f744fc6c26695e0be22973477e863b20f9b2ba58e1857af4c

Observation 0460dde0-cf7d-4deb-b65a-92845e325e4e · outbound

This paper cites Spoken content retrieval—beyond cascading speech recognition with text retrieval,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Spoken content retrieval—beyond cascading speech recognition with text retrieval,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:08.061091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:53.970416Z digest=sha256:79b7f6ae4dea5d9ae19385c765e03bbcc484eb54f315bb93a8c1114cd20a65d1

Observation 3a8108d8-5251-4262-924a-7044d05beca3 · outbound

This paper cites SpeechDPR: End-to-end spoken passage retrieval for open-domain spoken question answering,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style SpeechDPR: End-to-end spoken passage retrieval for open-domain spoken question answering,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:07.834944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:54.089386Z digest=sha256:6e91654bb686fb886e13ac1612d253a537eae17ba1ba101380c73e53c7ecd01c

Observation 3447d916-4f3c-4c70-8c79-e20aedddd086 · outbound

This paper cites Retrieval augmented end-to-end spoken dialog models,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Retrieval augmented end-to-end spoken dialog models,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:07.599886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:54.211609Z digest=sha256:8d574429e41cd739e2a36bd21c9c17b1bf656e760e915cf779f744a6b5dc952c

Observation 88292e66-930c-407c-b267-94e6aee151af · outbound

This paper cites WavRAG: Audio-Integrated Retrieval Augmented Generation for Spoken Dialogue Models.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style WavRAG: Audio-Integrated Retrieval Augmented Generation for Spoken Dialogue Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T20:08:54.327856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:08:54.327856Z digest=sha256:26deaa7df62dfb11086cdb16501935611789401c45c456fd95bf41a3f480a8f7

Observation bbbe176b-4fc8-4de3-acf0-0507ebdf67f6 · outbound

This paper cites Speech retrieval-augmented generation without automatic speech recognition,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Speech retrieval-augmented generation without automatic speech recognition,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:07.394761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:54.447022Z digest=sha256:e21271c99cc5939ad9741f04a5c8030e48fbe2248c25186d94388b02f9593002

Observation 048add9f-8558-4874-90a5-8277fe7b6ca7 · outbound

This paper cites Prompting audios using acoustic properties for emotion representation,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Prompting audios using acoustic properties for emotion representation,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:07.222536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:54.615420Z digest=sha256:08a4a30c92f110abbb5a8e691b5aff7cf06c72383f177cedfabba2952a6fad4a

Observation fd5d1968-4cc0-4b3e-b386-f081012e6276 · outbound

This paper cites Large-scale contrastive language-audio pretraining with feature fusion and keyword-to-caption augmentation,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Large-scale contrastive language-audio pretraining with feature fusion and keyword-to-caption augmentation,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:07.029219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:54.731774Z digest=sha256:a95e9cf5f9d2b9091c95ce8f1811476713b13eb417fcdec7e18a0da18c2ce024

Observation 9409ad26-5b5a-419b-a9d6-2d18f66845ab · outbound

This paper cites IEMOCAP: Interactive emotional dyadic motion capture database,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style IEMOCAP: Interactive emotional dyadic motion capture database,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:06.804722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:54.894348Z digest=sha256:c347eb5ea0dd1fcf78b259bde19d8247e5787522c619446aa295004c8aff707d

Observation 2b00b25b-503a-462f-b924-f2987c8e4824 · outbound

This paper cites Emotional voice conversion: Theory, databases and ESD,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Emotional voice conversion: Theory, databases and ESD,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:06.618778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:55.011380Z digest=sha256:b90ff8a5c76bd487c0b1ab51099b71dbc7241094977e4b48517876a331725f7c

Observation a631dcd0-1d34-4f75-9465-1b2af2119817 · outbound

This paper cites Expresso: A benchmark and analysis of discrete expressive speech resynthesis,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Expresso: A benchmark and analysis of discrete expressive speech resynthesis,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:06.446662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:55.123871Z digest=sha256:3d3686e146421075e697dcfbf6dfbb22b2215610287c151cd8d39ab36bb6783c

Observation 383da48d-113c-467a-9679-9cf71a7c0c63 · outbound

This paper cites PromptTTS: Controllable text-to-speech with text descriptions,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style PromptTTS: Controllable text-to-speech with text descriptions,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:06.256855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:55.239997Z digest=sha256:0aab70ac7109ebf88fb43fe129e29b1f4e0e938884e58faa4afb16e094e78804

Observation 4f6c3a39-8f78-4e9d-8f30-8295154507e9 · outbound

This paper cites PromptStyle: Controllable style transfer for text-to-speech with natural language descriptions,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style PromptStyle: Controllable style transfer for text-to-speech with natural language descriptions,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:06.074349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:55.336467Z digest=sha256:8bb5ae830c36aacfce0e07981247007e3e1cdbee02690f470fec0fec28e4ee3e

Observation cefdac49-bc5f-4d28-b2f0-5f0bbdc2d1c4 · outbound

This paper cites PromptTTS 2: Describing and generating voices with text prompt,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style PromptTTS 2: Describing and generating voices with text prompt,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:05.843066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:55.462957Z digest=sha256:285f3c0601771f2bb4607e09216eb86dec3aaacb6633234db6fd6e1535c2b548

Observation 4ca6dce2-e074-4916-aa4f-9695a3d41165 · outbound

This paper cites DreamV oice: Text-guided voice conversion,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style DreamV oice: Text-guided voice conversion,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:05.624428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:55.587642Z digest=sha256:da0246c1c95ebbbc9b01fe95760243b5bc1a27e429d83dbd278ed2243d961509

Observation a5f17ddf-3b7c-43e4-be54-ad62714dc880 · outbound

This paper cites StyleCap: Automatic speaking-style captioning from speech based on speech and language self-supervised learning models,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style StyleCap: Automatic speaking-style captioning from speech based on speech and language self-supervised learning models,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:05.438828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:55.706895Z digest=sha256:0ddbb49c0a8a2fa6727b60fb2de938cbec7fea58c5913eacfd78ca31b208d244

Observation d516af7a-56b6-4212-a7c5-7033bc9a7978 · outbound

This paper cites LibriTTS-P: A corpus with speaking style and speaker identity prompts for text-to-speech and style captioning,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style LibriTTS-P: A corpus with speaking style and speaker identity prompts for text-to-speech and style captioning,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:05.193490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:55.823325Z digest=sha256:2a0662419120dd1931729c3b30280fcaa656db0f46af0829fd71832261cfffe2

Observation 96f0f9f0-92d8-460f-97fd-6d22f5e94e2b · outbound

This paper cites V ocabulary independent spoken term detection,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style V ocabulary independent spoken term detection,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:04.909259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:56.015717Z digest=sha256:e548152e4c514d971c37e4cc05bb06aeb92e6a91939a99969b872ba6ef426406

Observation 7c8bf58c-9105-4f0e-84ba-6ff34ea07034 · outbound

This paper cites Statistical lattice-based spoken document retrieval,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Statistical lattice-based spoken document retrieval,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:04.619531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:56.143258Z digest=sha256:58bd42bad9b431d00c5c1869ec5a770c40cfcc0f9ad5013ed6dde8564b0a4ec9

Observation e7af4a0a-085a-48da-b017-70bc70675d86 · outbound

This paper cites Improved semantic retrieval of spoken content by document/query expansion with random walk over acoustic similarity graphs,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Improved semantic retrieval of spoken content by document/query expansion with random walk over acoustic similarity graphs,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:04.334449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:56.260715Z digest=sha256:851126a0ae29d54e8ad7b5461230bf7c8c085e46179ecaa0714f1b973120d68c

Observation 35a12892-4b63-4ab0-a6fd-87b1bbd05521 · outbound

This paper cites Evaluating ASR output for information retrieval,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Evaluating ASR output for information retrieval,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:04.030441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:56.414580Z digest=sha256:15067877d09ce18dd06dafca5ce25da29e7d9fee9587ccaf29e324119b4ffb7c

Observation 89d11070-717c-47cf-8bf1-3df3ad898e04 · outbound

This paper cites Investigating the global semantic impact of speech recognition error on spoken content collections,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Investigating the global semantic impact of speech recognition error on spoken content collections,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:03.622881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:56.573972Z digest=sha256:08c64afb7470d239dfc75928547d1f921231581c9e2d734977c149793902fcc2

Observation 166dafa7-2cb9-4188-9026-d506fbb459bc · outbound

This paper cites Speech-centric information processing: An optimization-oriented approach,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Speech-centric information processing: An optimization-oriented approach,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:03.133742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:56.654325Z digest=sha256:883d6aa0f420ca96fbd14efa33ccbc677ed34e1d32dacd4e1cd141cf6119b101

Observation e1876a5c-0e8c-4fef-b17a-322ba56511c9 · outbound

This paper cites Unsupervised spoken-term detection with spoken queries using segment-based dynamic time warping.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Unsupervised spoken-term detection with spoken queries using segment-based dynamic time warping

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:02.778218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:56.754424Z digest=sha256:c2f92c288add88fafa0a8891ccf7a52834afb77612124831901119f0777373c3

Observation a1b05aaa-b43b-41c9-b51d-705d123fa730 · outbound

This paper cites Memory efficient subsequence dtw for query-by-example spoken term detection,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Memory efficient subsequence dtw for query-by-example spoken term detection,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:02.467412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:56.873814Z digest=sha256:6e8f2d179ab55dbf9e5c7ebbf12515bcd237eadba662a4aaa4433f8eddf62719

Observation 3cddcd3e-154d-49d2-bdde-aaa3697cb39e · outbound

This paper cites The spoken web search task at MediaEval 2012,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style The spoken web search task at MediaEval 2012,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:02.140897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:57.012996Z digest=sha256:45a81d1d74d5c76484bc9edd79d08384e5a9827d09f62f4a267e1bff849f1418

Observation 29f2ff46-011e-4f9a-8824-b0d76f1b53e0 · outbound

This paper cites RECAP: Retrieval-augmented audio captioning,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style RECAP: Retrieval-augmented audio captioning,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:01.880751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:57.134441Z digest=sha256:a800d14687adbdb3449022780ba611b96ee1becd661aeca59b36e0c2f760df48

Observation 710f025a-c68c-43a0-a5c5-7e05f52846b8 · outbound

This paper cites Beyond speaker identity: Text guided target speech extraction,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Beyond speaker identity: Text guided target speech extraction,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:01.586330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:57.182972Z digest=sha256:14397214a3dfd791d544f1d04bec0a6f96f55ffe505dbd36680e3bc0b2e9b2c0

Observation 4ce8b136-65d5-40bd-b6db-cb39af5900cd · outbound

This paper cites WavLM: Large-scale self-supervised pre- training for full stack speech processing,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style WavLM: Large-scale self-supervised pre- training for full stack speech processing,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T20:08:57.318770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:08:57.318770Z digest=sha256:0aeb3f50857cdf439a222453015d72081896a5354907d3ef059faa12ce807865

Observation 02b9ac51-317e-4829-b155-f89905685ddb · outbound

This paper cites emo- tion2vec: Self-supervised pre-training for speech emotion representation,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style emo- tion2vec: Self-supervised pre-training for speech emotion representation,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:01.170207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:57.429737Z digest=sha256:92275190bebb2f05ca28cafe81789ce8f60cad0a471a20cdaba7b5d1f48b8def

Observation 448eae5e-275a-4316-9d96-bf266f4d046b · outbound

This paper cites BERT: Pre- training of deep bidirectional transformers for language understanding,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style BERT: Pre- training of deep bidirectional transformers for language understanding,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:00.963076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:57.528474Z digest=sha256:600c1d8a61d7da2b7caed259995fe2c3fefc795e9cea8f6de834ec4c2dabd2c5

Observation 79accd8c-d82c-4429-9aa6-6ce3e18e6b2b · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T20:08:57.647214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:08:57.647214Z digest=sha256:3dc20f27ce61588dce6989db6cda0f583a54077b5112ee44539c69758c911f11

Observation d53a25fb-05a3-405c-8373-612ac569b576 · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Exploring the limits of transfer learning with a unified text-to-text transformer,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:00.773425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:57.795428Z digest=sha256:a5bac94ba34c151dc2012aea605bcf06066214816dc07757b0baa2ba3bd74474

Observation f681b8d2-59e9-4d9b-9092-8f3d21cf3520 · outbound

This paper cites The flan collection: Designing data and methods for effective instruction tuning,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style The flan collection: Designing data and methods for effective instruction tuning,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:00.530926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:57.923525Z digest=sha256:64559058e8d0beaf94416ce2d251242a6a43b26b22122efa2c93d010f9c183fd

Observation 9b112b2f-79a3-4cd5-a356-4d8360b9f105 · outbound

This paper cites Sentence-T5: Scalable sentence encoders from pre-trained text-to-text models,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Sentence-T5: Scalable sentence encoders from pre-trained text-to-text models,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:09:00.271532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:58.045198Z digest=sha256:603fe0994deb97b18425908d9f7b8dd541a7d7cabc6888955edb0b39f4500975

Observation 83195699-1641-47d1-af58-b3e32cbbb861 · outbound

This paper cites Domain-adversarial training of neural networks,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Domain-adversarial training of neural networks,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:08:59.969229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:58.052707Z digest=sha256:bf7d3fb3ef3092d4f2d46627d06989f2269d43581399a7839ccdf95878c47f17

Observation 24279dca-a0f0-444d-b756-51de2a8d3365 · outbound

This paper cites Unsupervised domain adaptation by backpropagation,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Unsupervised domain adaptation by backpropagation,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:08:59.689390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:58.056120Z digest=sha256:0421b2fa8a03c46f194a86cf0d6a8895118a64a9c402d2f1002ade4215710c5b

Observation b802835a-261b-4e40-bf7e-a3b5a91fb24e · outbound

This paper cites GPT-4o System Card.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style GPT-4o System Card

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T20:08:58.063279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:08:58.063279Z digest=sha256:678fa40d0cb383a53291506797656b1fa2c1c0158ea093debf907f499fcc0850

Observation 96bdecf8-9edd-4b16-a788-e8cf996cd90e · outbound

This paper cites Decoupled weight decay regularization,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Decoupled weight decay regularization,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-05T20:08:58.165179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:08:58.165179Z digest=sha256:ceef1b6c0641a19472f8869f11a807a2afd48b08b838acdb3190b0b5a8f13c22

Observation f5870a2b-bf04-496f-b0cd-2d8a5fdf736e · outbound

This paper cites PromptTTS++: Controlling speaker identity in prompt-based text-to-speech using natural language descriptions,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style PromptTTS++: Controlling speaker identity in prompt-based text-to-speech using natural language descriptions,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:08:59.461612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:58.281102Z digest=sha256:ce88c0cf076cb62adf325a42bf8ad9ca50550d5ad6031f5f6395a53a983baa51

Observation c3b5de4c-bd7f-4b68-b677-f8ce08f3a5e3 · outbound

This paper cites PromotiCon: Prompt- based emotion controllable text-to-speech via prompt generation and matching,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style PromotiCon: Prompt- based emotion controllable text-to-speech via prompt generation and matching,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:08:59.165146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:58.359548Z digest=sha256:4521291e285c167305a5a5009cca03a8e10fed6e21efb877a9d357ba6cdbe284

Observation f6bc0f0e-80fd-42db-9cf4-d6fef5e76115 · outbound

This paper cites Visualizing data using t-SNE,.

Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style Visualizing data using t-SNE,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:08:58.841982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T20:08:58.582356Z digest=sha256:f6b9b662b4fc06a0eb4c3a188e391e2f2809e9ab5b9bf5bd706add76bf8132b3

Pith citing papers

No inbound Pith citation observations are available.