Pith. sign in

Paper Citation Record · LEDGER

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems

As of 6 August 2026, this Paper Citation Record lists 100 of 100 outbound references and 0 inbound Pith citation observations for arXiv:2508.10456.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.10456 v1

Coverage vector

measured 100 of 100 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:31:48.143276Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 100 outbound references displayed

  • verified exact11
  • verified fuzzy57
  • unresolved32
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 11db9d7c-8e3b-4198-a6bc-1b8e5fffbb68 · outbound

This paper cites Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.040946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.040946Z digest=sha256:55298c50d0270ba0918f00580167b12e595ee11c215610421cf4118f5e5234bc

Observation f7f1ea8e-1476-4043-9d28-2b0aa8db8904 · outbound

This paper cites Hybrid ctc/attention architecture for end-to-end speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Hybrid ctc/attention architecture for end-to-end speech recognition,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.145089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.145089Z digest=sha256:123a3f44d7c9989b1c38dabe2a84cd72f97b9379baea09ae81929bd8780d19d6

Observation 2095d3b1-e8d6-49e8-afab-564009151f8f · outbound

This paper cites Listen, attend and spell: A neural network for large vocabulary conversational speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Listen, attend and spell: A neural network for large vocabulary conversational speech recognition,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.228233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.228233Z digest=sha256:17da5c729e670a9336667d451f608c11c2f7c3d8c3d45cb8f04f7a9a307c3d8c

Observation b4c69068-b269-4da4-9bc8-7ac0012145a2 · outbound

This paper cites Attention is all you need,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Attention is all you need,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.340737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.340737Z digest=sha256:1435b300dea131ae19e8d317a65298ce087044d6e47fb82ffa384e61244e31f3

Observation ab75db89-276c-470b-9458-8c626c82fc0f · outbound

This paper cites Speech-transformer: a no-recurrence sequence-to- sequence model for speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Speech-transformer: a no-recurrence sequence-to- sequence model for speech recognition,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.447669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.447669Z digest=sha256:46fa93706e4684df16f7d8bc7e6cc828be2b81904a04d0fd54c94da8027f88d2

Observation e1fd1b12-f44a-4d9d-bafc-2a70d90e6f47 · outbound

This paper cites A comparative study on transformer vs rnn in speech applications,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems A comparative study on transformer vs rnn in speech applications,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.537168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.537168Z digest=sha256:3e38a745d96706fab832580587254d2c76b3eacadaee8784aeabf1b5d2d06995

Observation 4bee5d7f-62eb-48c2-82d2-5097c1aac52e · outbound

This paper cites Conformer: Convolution-augmented Transformer for Speech Recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Conformer: Convolution-augmented Transformer for Speech Recognition,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.633846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.633846Z digest=sha256:0a933ae80c9231ba2f53347972b4f09c8e85c2e8eb770e227cb74860a05399fe

Observation d2390ca1-8927-4b0a-acd4-f2140f2dfb33 · outbound

This paper cites Recent developments on espnet toolkit boosted by conformer,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Recent developments on espnet toolkit boosted by conformer,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.712245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.712245Z digest=sha256:d1106d5a59906bd6723ee3044695c675ff73b45cd85e08ef0875df846a52ab25

Observation e7c11225-302e-4ce8-bc5e-05b7baf6788f · outbound

This paper cites Sequence Transduction with Recurrent Neural Networks.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Sequence Transduction with Recurrent Neural Networks

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.796983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.796983Z digest=sha256:7d9f9f34088a98052e094c9f947cf9df92d002dedfaf824c6903f16bd618006c

Observation 047e0586-806e-42cd-81aa-7574d9fa98cd · outbound

This paper cites Recurrent neural networks,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Recurrent neural networks,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.914416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.914416Z digest=sha256:1b816ed3f71de82a4d526ae38f29a860b434a5061839cde5eda94f2084a5bde3

Observation 3c41ce41-2d7b-4270-a270-7a0ebda4a6d8 · outbound

This paper cites Long Short-Term Memory Based Recurrent Neural Network Architectures for Large Vocabulary Speech Recognition.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Long Short-Term Memory Based Recurrent Neural Network Architectures for Large Vocabulary Speech Recognition

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.032849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.032849Z digest=sha256:6cc47a9f4988c81a02a82030bbbc67f3d6ff2c55e3ee7463dc3932174e5b41fe

Observation 6d4faef4-48f6-4892-a2c7-dc86704e821e · outbound

This paper cites Efficient training of neural transducer for speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Efficient training of neural transducer for speech recognition,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.112682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.112682Z digest=sha256:2f5441864ceed0ed28fa243ce2c855eadd26deaf575dc7c02ab4b17a1fb451ea

Observation 995b2255-cb3e-47ef-98d0-352f5904b3aa · outbound

This paper cites On the limit of english conversational speech recog- nition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems On the limit of english conversational speech recog- nition,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.213383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.213383Z digest=sha256:8f733eac3bb4f9365c377278cefbb5d2015f99a052c0310cf5c7e5812ed34155

Observation 7146709b-3bbc-4ffe-be3a-99477a6906f5 · outbound

This paper cites Improving the training recipe for a robust conformer-based hybrid model,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Improving the training recipe for a robust conformer-based hybrid model,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.327114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.327114Z digest=sha256:37c85ba21448dcff15b3163aada03bb92cefbe6b48046ce48bd182d2e5954367

Observation 9cd39ca8-bd36-4e63-9213-3cfcdf89c110 · outbound

This paper cites Confidence score based conformer speaker adaptation for speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Confidence score based conformer speaker adaptation for speech recognition,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.415984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.415984Z digest=sha256:c14859b49479013437fd068919dbfa6a225697e61da21ced2cfecb0db8cc0313

Observation 00bb68b6-2a8c-4174-b0bf-d35ecab8dbb3 · outbound

This paper cites Diagonal State Space Augmented Transformers for Speech Recognition.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Diagonal State Space Augmented Transformers for Speech Recognition

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:51.243108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:40.515370Z digest=sha256:f00da51c22dd5b9ee1daf9c9ce279d70d2f73ddc7145ff5a14489d9932451cca

Observation 0617d538-d4e0-42ce-b3bd-101d1af0fe27 · outbound

This paper cites Recent advances in end-to-end automatic speech recog- nition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Recent advances in end-to-end automatic speech recog- nition,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.601112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.601112Z digest=sha256:b500140dfe4c444b1b9d0e910297bce84c36f333ec281f83682e95af2b045271

Observation 0bcb61f5-291a-410e-b19c-38ee77418444 · outbound

This paper cites Improving Transformer-based Conversational ASR by Inter-Sentential Attention Mechanism.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Improving Transformer-based Conversational ASR by Inter-Sentential Attention Mechanism

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:51.028105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:40.689037Z digest=sha256:e7c0b824118c9de92d2ef63af10f01ea68406dc33f0715ac3d30fcd5ccfc44e6

Observation ec1cfdea-1210-4095-8423-246400ab3e25 · outbound

This paper cites Improving asr contextual biasing with guided attention,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Improving asr contextual biasing with guided attention,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.773045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.773045Z digest=sha256:61558c347d8b987b43c667069de7eacb92a1c0e445c6c4d2d3d59f0e55cf0ef8

Observation 0d3c780a-92b5-4b2b-8207-abaff055228d · outbound

This paper cites Optimizing byte-level representation for end-to-end asr,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Optimizing byte-level representation for end-to-end asr,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.851093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.851093Z digest=sha256:49ea49ccdab3187b387b7d746bc1ab8f8386ee4e1f90006ab5e0cd20692a4fd2

Observation d9907ec3-f172-4699-9634-6a4165bf1ed9 · outbound

This paper cites Contextual modeling for document-level asr error correction,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Contextual modeling for document-level asr error correction,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.979654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.979654Z digest=sha256:22efe681be36332011d48c9b633f853cf6c59ec7801286268832debef2d8203c

Observation 9c199e58-a953-4b7a-9239-740734c6c521 · outbound

This paper cites Promptasr for contextualized asr with controllable style,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Promptasr for contextualized asr with controllable style,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:41.065159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:41.065159Z digest=sha256:b11df8994d1556897013c3706c9119c81bde1b98b43236bce7643d0910e8f0f9

Observation 6f881a35-e3a6-4ae2-80d2-cd350f18f212 · outbound

This paper cites Conversational speech recognition by learning audio- textual cross-modal contextual representation,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Conversational speech recognition by learning audio- textual cross-modal contextual representation,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:41.143119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:41.143119Z digest=sha256:ab61d76b6ae9e189d0248a8b32bd77b76706f1de430219e1c45f8f851507a6fc

Observation 5efe2905-0318-4095-96cc-99e8f0f248b1 · outbound

This paper cites Dual-mode nam: Effective top-k context injection for end-to-end asr,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Dual-mode nam: Effective top-k context injection for end-to-end asr,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:41.229606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:41.229606Z digest=sha256:ad3f416e0f5d455172a21bb864d0810611a0afa40425aa50833d8f5306397086

Observation 3e8e1dbc-1089-4d79-ab86-e7db061ad785 · outbound

This paper cites CPPF: A contextual and post-processing-free model for automatic speech recognition.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems CPPF: A contextual and post-processing-free model for automatic speech recognition

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:50.766549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:41.333464Z digest=sha256:092f49ca97218da7442495f66a7fae86457a7ad4aef425bc9d1d798c436c2f1c

Observation 71a14c27-013a-4ac2-9140-e92a55b71557 · outbound

This paper cites CASA-ASR: Context-Aware Speaker-Attributed ASR.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems CASA-ASR: Context-Aware Speaker-Attributed ASR

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:50.531977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:41.404760Z digest=sha256:d76b616930f681070cab49b06a7aa10bae16bb87dd14638628c7614647cab22b

Observation 593b9995-e5c3-43dc-8322-202f0bcb9e72 · outbound

This paper cites Using External Off-Policy Speech-To-Text Mappings in Contextual End-To-End Automated Speech Recognition.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Using External Off-Policy Speech-To-Text Mappings in Contextual End-To-End Automated Speech Recognition

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:41.493267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:41.493267Z digest=sha256:ae5639e7d31c191dcda3f404d1f370bbdbb560f565e284ca737ca4b1cfe4d2de

Observation 461d666f-6e32-4906-83d4-166bd3dec097 · outbound

This paper cites Bring dialogue-context into rnn-t for streaming asr,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Bring dialogue-context into rnn-t for streaming asr,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:01.257605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:41.598588Z digest=sha256:ad27feb724bf8836a801a80887b29da6b9706a56af5ad14467162e6f42a901f3

Observation 74b8dca5-c583-4fd1-acbd-7016962c5497 · outbound

This paper cites CopyNE: Better Contextual ASR by Copying Named Entities.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems CopyNE: Better Contextual ASR by Copying Named Entities

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:50.264086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:41.707826Z digest=sha256:f5c0ecb0c89f6c75fb7d4e7d5bac94a8b32f237fd4faaa598be172984989f3b0

Observation 08faf3b5-bde3-4fae-a1fe-1c067bc2b645 · outbound

This paper cites Improving Neural Biasing for Contextual Speech Recognition by Early Context Injection and Text Perturbation.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Improving Neural Biasing for Contextual Speech Recognition by Early Context Injection and Text Perturbation

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:50.050215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:41.805565Z digest=sha256:22109abb0b08d7161619ddf6ed9fb47096cccb844a217794c5306436c6cfae33

Observation 6976842c-c120-407c-8157-c58085e3a12f · outbound

This paper cites Training language models for long-span cross-sentence evaluation,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Training language models for long-span cross-sentence evaluation,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:01.036595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:41.884033Z digest=sha256:a26b9e9fa784244fc65069fb31e56e7cad21883491fcd2b72b2eaeff8f9f7eed

Observation d274690e-fcbd-4c06-aefb-1596b24930c2 · outbound

This paper cites LSTM-LM with Long-Term History for First-Pass Decoding in Conversational Speech Recognition.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems LSTM-LM with Long-Term History for First-Pass Decoding in Conversational Speech Recognition

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:49.827274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:41.984032Z digest=sha256:06445aa0e0a52ea8ba94e33ebc794c60e143e22463e66b710f6ddc45fa9c1fa5

Observation a44c5369-b7b4-4c4f-938f-b250fe31b8f0 · outbound

This paper cites Session-level language modeling for conversational speech,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Session-level language modeling for conversational speech,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:00.851580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:42.099735Z digest=sha256:ac264df04bcef24bbb029d3d53bb2ab70653c51cff6e9fc3f85dad8803c2d168

Observation 1a018f5b-d065-4fb8-9e11-3bc6c34ad837 · outbound

This paper cites Transformer-xl: attentive language models beyond a fixed-length context,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Transformer-xl: attentive language models beyond a fixed-length context,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:00.636395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:42.175833Z digest=sha256:757d760802e7d323ef71c60d8ed6ee30708458ab0630f91d94f5a095a6ae988c

Observation de503af8-0a5f-4b5b-8e64-72b83282e881 · outbound

This paper cites End-to-end speech recognition on conversations,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems End-to-end speech recognition on conversations,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:00.225580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:42.289439Z digest=sha256:86af840c45d8527ad843fb7fdd2607e24fbb9dc345fba582546e9d74c6dfbeca

Observation 20aad9f9-3f85-4409-9fd3-f433b7213289 · outbound

This paper cites Contextualizing ASR lattice rescoring with hybrid pointer network language model,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Contextualizing ASR lattice rescoring with hybrid pointer network language model,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:59.827745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:42.363244Z digest=sha256:be77727dfbc1a33acedd9856b0a6d76283ceed00c955f3c070c1c724d3a86c49

Observation 893e0851-d763-4751-9ca6-b08b2789e449 · outbound

This paper cites Use of contexts in language model interpolation and adaptation,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Use of contexts in language model interpolation and adaptation,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:59.689635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:42.455433Z digest=sha256:2319d35ff183bcafe3f7cccfee4f9f1d3ce7d27816bdc599b173485d28009970

Observation 4a711db4-3544-43ce-89b3-b878233fdb68 · outbound

This paper cites Longformer: The Long-Document Transformer.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Longformer: The Long-Document Transformer

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:42.516968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:42.516968Z digest=sha256:fac408e778f8e40ab753b34cdf5db84b100e6462c1682cd09f092cfc996f0797

Observation 29dfea9b-e96d-4063-8f0c-75fe30248ea4 · outbound

This paper cites Transformer language models with lstm-based cross- utterance information representation,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Transformer language models with lstm-based cross- utterance information representation,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:59.523332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:42.603756Z digest=sha256:2e22fa50528bd403f05ed179ff99eef512a2d5c6edd1653fb9be5ce3b44d3669

Observation 9d5d7fc7-5001-4462-9ccb-d1aa6801088e · outbound

This paper cites Ctc-assisted llm-based contextual asr,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Ctc-assisted llm-based contextual asr,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:59.365324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:42.670064Z digest=sha256:d8b67f72e5510c30e07b3c9894aa5098ca594df641c2524ac33564d25765bab7

Observation 49e04d82-0591-4a2b-bcfa-165cee1248a3 · outbound

This paper cites Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:42.749315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:42.749315Z digest=sha256:f50207be4dfc91308889fa0706fbd99c804578a9c155ae70ae85a9482677ac73

Observation 40e7f0d3-ce42-4085-9876-cddf36846caf · outbound

This paper cites Contextual asr error handling with llms augmentation for goal-oriented conversational ai,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Contextual asr error handling with llms augmentation for goal-oriented conversational ai,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:59.195690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:42.859181Z digest=sha256:b17caf01f46de6d474e6e114f57e79d9be2691d7826862b7c2a442fa2d02a2c4

Observation 091432e8-62b9-48a7-86de-4e49e62bfb26 · outbound

This paper cites Towards asr robust spoken language understanding through in-context learning with word confusion networks,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Towards asr robust spoken language understanding through in-context learning with word confusion networks,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:59.005515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:42.920140Z digest=sha256:5b5b334dc1564c617e110ba0ccb935fe8fb9df76310ec891eff6dd3c44fe785e

Observation 14d87e6b-ea9c-43bd-a846-96e9386cd530 · outbound

This paper cites Contextualization of asr with llm using phonetic retrieval- based augmentation,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Contextualization of asr with llm using phonetic retrieval- based augmentation,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:58.856857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:43.000136Z digest=sha256:f45821a8bd4a054f960f288baca2afff8e6fbad2185e30aeb72779ecf4045b25

Observation fd7b20e2-e5d2-4c48-8491-71acd464f171 · outbound

This paper cites End-to-end speech recognition contextualization with large language models,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems End-to-end speech recognition contextualization with large language models,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:58.746831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:43.104408Z digest=sha256:2c27b10b662d8a6d0773f2ca04ce28edb4fb853edbb4c8a6f7b2b0e23263ffd2

Observation 91402f0c-236e-4798-b02d-1be9e1a24467 · outbound

This paper cites Improving domain-specific asr with llm-generated contextual descriptions,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Improving domain-specific asr with llm-generated contextual descriptions,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:58.599667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:43.219288Z digest=sha256:81b0c28f5c6540eba971471850147354cd4de3538ca82a4449bf5452be428b75

Observation e5436655-70a5-4320-af73-a6360e8e52d1 · outbound

This paper cites MaLa-ASR: Multimedia-Assisted LLM-Based ASR.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems MaLa-ASR: Multimedia-Assisted LLM-Based ASR

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:49.498400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:43.295375Z digest=sha256:0aab997746ffa0840fc03314500c2ef9636b6dc7aea747fc87149a07c01ad09b

Observation fde25a80-3665-471e-971e-2ca0bdc51efe · outbound

This paper cites Contextualized speech recognition: rethinking second- pass rescoring with generative large language models,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Contextualized speech recognition: rethinking second- pass rescoring with generative large language models,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:58.403844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:43.373827Z digest=sha256:5af0bf5e2116452ac15d39a82900bf37a18fde72eb38b2e05a302a781348a516

Observation 4eecbe69-a9bd-41d0-9f06-79167b1a6412 · outbound

This paper cites Improving speech recognition with prompt-based contextualized asr and llm-based re-predictor,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Improving speech recognition with prompt-based contextualized asr and llm-based re-predictor,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:58.237641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:43.456893Z digest=sha256:8c37e77a7dbfcb04be4737511c37e415ce06fd9113e19cb1c7ae3a1a4cf556e9

Observation af0de628-c71a-4cc2-94db-f74b2e7f04d3 · outbound

This paper cites Using Large Language Model for End-to-End Chinese ASR and NER.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Using Large Language Model for End-to-End Chinese ASR and NER

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:43.555162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:43.555162Z digest=sha256:1fb98391ddd3669db7f9b4bd98c7aa64b0aa91aa1812b166304c21df2249b3d9

Observation cdcaf6aa-a5b6-4586-b7a8-9ab090614695 · outbound

This paper cites Compressive Transformers for Long-Range Sequence Modelling.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Compressive Transformers for Long-Range Sequence Modelling

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:43.616971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:43.616971Z digest=sha256:1a089b0f780139bbf8b348c8f532dee5f27a0e6b7f324c823d01f368fb310532

Observation 4d49d666-82cd-4600-a11c-daaf11c69055 · outbound

This paper cites Speaker-aware Speech-transformer,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Speaker-aware Speech-transformer,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:58.090493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:43.710776Z digest=sha256:f577091a3339fccd2be193218bf93793713913fc375dcda515cb6631c4342ab1

Observation a151e8a4-7ff0-434a-9d7a-4e0a24c7906e · outbound

This paper cites Transformer ASR with Contextual Block Process- ing,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Transformer ASR with Contextual Block Process- ing,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:57.946512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:43.789641Z digest=sha256:ff4872d12b8ca9e9284d1ad942f3005fe7576787039608000fe6d39cca4ca696

Observation c9307482-10d9-4ae7-bc12-dbcab7aa9904 · outbound

This paper cites Transformer-based long-context end-to-end speech recognition.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Transformer-based long-context end-to-end speech recognition

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:57.806972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:43.856177Z digest=sha256:d6f6aa8ce04f4bbe46b7983782fe89d68bb9c9dcca22d1b671cd42bd18be4e60

Observation 24bdc12b-946d-40df-a2f9-8cf21cc5a672 · outbound

This paper cites Advanced long-context end-to-end speech recognition using context-expanded transformers,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Advanced long-context end-to-end speech recognition using context-expanded transformers,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:57.684237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:43.939035Z digest=sha256:8be9494ac8ebbf7689820d9c30784a03d049d5de51394cecc6cdef83988a55eb

Observation 89b091be-dc99-4f00-a415-0b62fc89a353 · outbound

This paper cites Context-aware end-to-end asr using self-attentive embedding and tensor fusion,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Context-aware end-to-end asr using self-attentive embedding and tensor fusion,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:57.527917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:44.010039Z digest=sha256:10175c386e9942ff76bdcdc079953604eb4d3eee023692f1d556f195775a0ee6

Observation 97cf2b87-6708-4372-9ff2-d8c13f6f46ce · outbound

This paper cites Longfnt: Long-form speech recognition with factorized neural transducer,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Longfnt: Long-form speech recognition with factorized neural transducer,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:57.374355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:44.107285Z digest=sha256:c5b10788f34bf7e24662d2241caa2da9e519a60973316c2fe6870148f520dc63

Observation 9bf31ccf-200c-406b-bb7e-aaf0ef13288c · outbound

This paper cites Advanced long-content speech recognition with factorized neural transducer,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Advanced long-content speech recognition with factorized neural transducer,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:57.213167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:44.172636Z digest=sha256:0c17d8d25d729a9c490c66712568eff19046df225e5a953c105c7be60d7c10dc

Observation 288a580c-37a1-4974-87da-c5a9a22e0703 · outbound

This paper cites Gated Embeddings in End-to-End Speech Recognition for Conversational-Context Fusion.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Gated Embeddings in End-to-End Speech Recognition for Conversational-Context Fusion

Reference 59

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:49.180696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:44.256170Z digest=sha256:6011f576eecd9e910d82916d0be6948221e874d2b21ad06fa95139020cb50d03

Observation dace3c57-6b23-4716-b1f7-579469cae9fa · outbound

This paper cites Dialog-context aware end-to-end speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Dialog-context aware end-to-end speech recognition,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:57.039618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:44.339765Z digest=sha256:38900fd7d6faa63557217919274689598c2eab813845bf943d7217f527aad9ea

Observation 500088fc-abbc-4dc7-970a-27a59f22f618 · outbound

This paper cites Improving RNN-T ASR Accuracy Using Context Audio.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Improving RNN-T ASR Accuracy Using Context Audio

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:48.854090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:44.437262Z digest=sha256:bfcc6ddf4234048e3d480d164d41ecf443b2f314c4f9ad911222380b1e3ad8af

Observation 58ce565f-11bf-4f36-af09-e80c3f164fc3 · outbound

This paper cites Large-context automatic speech recognition based on rnn transducer,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Large-context automatic speech recognition based on rnn transducer,

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:56.894247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:44.553967Z digest=sha256:bc3d3fc24d1de20e7347c11bda9a6ad5f00a6f2a88cf60fa06ad8bdc31748629

Observation 15b7a962-8d94-426f-92b8-22bd2eff6411 · outbound

This paper cites Phoneme-aware encoding for prefix-tree-based contextual asr,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Phoneme-aware encoding for prefix-tree-based contextual asr,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:56.762269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:44.635612Z digest=sha256:1f21b3659b3a13ffd7b51bda442c7fa747945d9d1b8e1de9b57abdf8be9b5114

Observation 35ac3e39-c498-464b-8d23-00baa8686d08 · outbound

This paper cites Deferred NAM: Low-latency Top-K Context Injection via Deferred Context Encoding for Non-Streaming ASR.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Deferred NAM: Low-latency Top-K Context Injection via Deferred Context Encoding for Non-Streaming ASR

Reference 64

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:48.551140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:44.729117Z digest=sha256:cfdcabc6017ef012dd9acdc5d6d9e7e7a8f00303ff14fca18e8a014d1ca86a59

Observation 93f9e98f-fa5e-4f4c-a78a-0cc1a17b62a3 · outbound

This paper cites Contextualized end-to-end speech recognition with contextual phrase prediction network,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Contextualized end-to-end speech recognition with contextual phrase prediction network,

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:56.611543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:44.809319Z digest=sha256:e7dde43ab6ede05d0d6c5735e002ff342be698dad3fd3a9e52acc37e7d8b7dd4

Observation 90f4525e-4987-4c89-9d6d-2f4f2e527839 · outbound

This paper cites Semi-autoregressive streaming asr with label context,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Semi-autoregressive streaming asr with label context,

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:56.459164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:44.893812Z digest=sha256:94ad2bad7a8c92a6d037e400d953b588c44928ee8d1cdead4de6e28711763a95

Observation 8e0f3d92-8326-49b5-b0de-56fd29b0fb72 · outbound

This paper cites Context-aware transformer transducer for speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Context-aware transformer transducer for speech recognition,

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:56.309374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:44.983848Z digest=sha256:899d573744925d73c82ca5fb600be039ac7283a4253930c4686be5cfd76e688a

Observation 10454637-34e0-4bc5-a807-76b2982a8845 · outbound

This paper cites Towards effective and compact contextual representation for conformer transducer speech recognition systems,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Towards effective and compact contextual representation for conformer transducer speech recognition systems,

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:56.180065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:45.047116Z digest=sha256:098b1bc6ae5ca7080c3f9ab953c8a683a68a7d132742868aee8a53d563205820

Observation 3c8d4152-b711-4728-aacd-1b96368900e8 · outbound

This paper cites Wav2vec 2.0: a framework for self-supervised learning of speech representations,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Wav2vec 2.0: a framework for self-supervised learning of speech representations,

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:56.044753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:45.119425Z digest=sha256:dc01c8c210e7aced8cfe4c0fecc5fb3b5159fd56b95969488050a2ce3d777e2a

Observation 5701ae8b-ddff-49f9-80a5-e7d78d2fd852 · outbound

This paper cites Wavlm: large-scale self-supervised pre-training for full stack speech processing,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Wavlm: large-scale self-supervised pre-training for full stack speech processing,

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:55.898704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:45.187611Z digest=sha256:ce269588fe92ebda66360b39dbeb801b97498663ff912c88dbad641b0281818c

Observation 3816819a-e1c4-41b2-91fc-8eda76b5cfc6 · outbound

This paper cites Hubert: Self-supervised speech representation learn- ing by masked prediction of hidden units,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Hubert: Self-supervised speech representation learn- ing by masked prediction of hidden units,

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:55.727367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:45.269084Z digest=sha256:cdfc3f445d2562ca31d432e1582f871983aa5e7818182a79cf33af01c3a9a659

Observation d25de274-8d78-40ef-8a6c-2c55cd28f302 · outbound

This paper cites data2vec: A General Framework for Self-supervised Learning in Speech, Vision and Language,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems data2vec: A General Framework for Self-supervised Learning in Speech, Vision and Language,

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:55.559832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:45.364942Z digest=sha256:5d253ca0fd13dd5d02399913a8238ef16898652771af1d82bd113c25e69aa526

Observation 1b8b7bd8-9bc8-49d5-8bcf-a14c1f119d77 · outbound

This paper cites XLS-R: Self-supervised Cross-lingual Speech Repre- sentation Learning at Scale,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems XLS-R: Self-supervised Cross-lingual Speech Repre- sentation Learning at Scale,

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:55.366817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:45.431578Z digest=sha256:9d8f077157bd075f70991919d657a35fe96c3c326bebdd25b0b6c27a7e996d79

Observation 6d1bce3b-7bed-4b95-a086-62e88057544f · outbound

This paper cites Data2vec: A general framework for self-supervised learning in speech, vision and language,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Data2vec: A general framework for self-supervised learning in speech, vision and language,

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:55.203752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:45.508204Z digest=sha256:cd030dd91f8d554cb42208b9089460e8dd40f05df65be0e7dc266fa6d2f92ca5

Observation 2f732a3c-5b53-4682-ad90-78c29eb96669 · outbound

This paper cites Robust speech recognition via large-scale weak supervision,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Robust speech recognition via large-scale weak supervision,

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:55.018209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:45.576141Z digest=sha256:8baf696522055f95ac81b78c4348c765f3df7d6a364b6583dc51192ed352e30a

Observation d1bf3b75-9894-4e2f-9d00-f8bb23dc60eb · outbound

This paper cites fairseq: A Fast, Extensible Toolkit for Sequence Modeling.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems fairseq: A Fast, Extensible Toolkit for Sequence Modeling

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:45.668112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:45.668112Z digest=sha256:826602e783e3e7a1ba0a174a4f47fc213d92ef2a154b71afb01dba40af997c31

Observation e171fbec-8b85-4df6-ac32-35cbfffa24d1 · outbound

This paper cites Zipformer: A faster and better encoder for automatic speech recognition.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Zipformer: A faster and better encoder for automatic speech recognition

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:45.751156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:45.751156Z digest=sha256:c6efe3886de28c5e4367bddaaf69f3f488a9324870f3f8acea074ab6c97c9f09

Observation d7e200e4-be8f-4994-bed1-132c5b6756fa · outbound

This paper cites ESPnet: End-to-End Speech Processing Toolkit,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems ESPnet: End-to-End Speech Processing Toolkit,

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:54.878240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:45.863510Z digest=sha256:f97c9e55bc0aa4bf8167aae655ea7a155daca3a7e806ce5a65ae14703e5fbc64

Observation 0a05ab93-c5b3-4b3c-9891-049835cbe534 · outbound

This paper cites GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:45.946423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:45.946423Z digest=sha256:9ad64c9bfe20e4f56576583950d826fded3cbc4352c52ff587fcd43096f2021e

Observation bf08cde8-13d4-4769-93f2-da483918513b · outbound

This paper cites Wenetspeech: A 10000+ hours multi-domain man- darin corpus for speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Wenetspeech: A 10000+ hours multi-domain man- darin corpus for speech recognition,

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:54.678706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:46.004011Z digest=sha256:d2082f1376f767fa37f84959c52dae8e58c8f3b5c6bd63a09bb1227b989ebf21

Observation 41321727-56bf-411a-a94d-3d2ea0de4481 · outbound

This paper cites The natural history of alzheimer’s disease: description of study cohort and accuracy of diagnosis,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems The natural history of alzheimer’s disease: description of study cohort and accuracy of diagnosis,

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:54.500000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:46.098424Z digest=sha256:b788532638e3159a6168e6da8c61440741a64ab89deba183474b1bf0c251ab71

Observation 877244d9-6847-4c06-97dc-e304f44a1a21 · outbound

This paper cites Speaker turn aware similarity scoring for diarization of speech-based cognitive assessments,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Speaker turn aware similarity scoring for diarization of speech-based cognitive assessments,

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:54.337853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:46.188228Z digest=sha256:9cd88b74dbda2457996fd2479c55e37c9fff235e05989f2d617e952305dd1d1a

Observation f96dc9c6-d6f3-4266-81f9-4c88cbab1a89 · outbound

This paper cites Self-supervised asr models and features for dysarthric and elderly speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Self-supervised asr models and features for dysarthric and elderly speech recognition,

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:54.172237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:46.265388Z digest=sha256:04b97436315d26282a17bef5b8f9e9f33a0bb094f0508073c37b4a25d492d388

Observation fffb07a8-d62b-4a14-b776-a01fae785d9e · outbound

This paper cites Developing real-time streaming transformer transducer for speech recognition on large-scale dataset,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Developing real-time streaming transformer transducer for speech recognition on large-scale dataset,

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:53.979937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:46.350732Z digest=sha256:ee2d6e1d9871834e8784fd5e1d97b4df6e183b86208531c6e4209ed81a5f6537

Observation 6a90f7b1-9a5d-412f-bc2f-ffe669f11b70 · outbound

This paper cites Transformer transducer: a streamable speech recog- nition model with transformer encoders and RNN-T loss,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Transformer transducer: a streamable speech recog- nition model with transformer encoders and RNN-T loss,

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:53.827935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:46.434543Z digest=sha256:456f038032f74bc989c575fc45cdbf4ef90cdae7cf28711522a83f07bbfe5eaa

Observation 9a4deac2-3a99-4335-b266-c56ae1695911 · outbound

This paper cites Transformer-Transducer: End-to-End Speech Recognition with Self-Attention.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Transformer-Transducer: End-to-End Speech Recognition with Self-Attention

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:46.546745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:46.546745Z digest=sha256:5c3a7504bff07af91a5bc764943d578fd2a7a08619bc0f1bc731f567a5428567

Observation 9cd3241e-12b5-4540-9b0d-224aa63924af · outbound

This paper cites Language modeling with gated convolutional networks,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Language modeling with gated convolutional networks,

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:53.598121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:46.629844Z digest=sha256:0a2f119ab65177dd120c9c46b85c0ffa01ef4bc48336b0d2dce36e1b107e723d

Observation ead4386f-fbc7-462c-87c3-4eb7b118fbc1 · outbound

This paper cites Attentive Pooling Networks.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Attentive Pooling Networks

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:46.700394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:46.700394Z digest=sha256:3b2f42eb8c6feadfb333796b414b7b7003562e914e57836c4a038bdd8f57c22a

Observation 0ae1fd88-4d05-4373-bcbd-1ce6a34afdb4 · outbound

This paper cites Alzheimer’s Dementia Recognition Through Sponta- neous Speech: The ADReSS Challenge,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Alzheimer’s Dementia Recognition Through Sponta- neous Speech: The ADReSS Challenge,

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:53.437465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:46.810510Z digest=sha256:d6491eb8c79f61cb50444d40329e3adf472d2b889db1499cbddf590d3856defd

Observation bf71d3d1-869a-41fa-8a5f-4ea6c3ede40d · outbound

This paper cites Development of the cuhk elderly speech recognition system for neurocognitive disorder detection using the dementiabank corpus,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Development of the cuhk elderly speech recognition system for neurocognitive disorder detection using the dementiabank corpus,

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:53.276825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:46.871585Z digest=sha256:6c2855a2593248f7675d454be76c71f0c72cac4b995cca9594a33b1c32bf11bf

Observation 0d376c8b-bcb5-424a-b8d6-1e2a86371335 · outbound

This paper cites Speaker Turn Aware Similarity Scoring for Diarization of Speech-Based Cognitive Assessments,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Speaker Turn Aware Similarity Scoring for Diarization of Speech-Based Cognitive Assessments,

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:53.119919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:47.037997Z digest=sha256:8afe3447ab3265265e50a7c577891f7acd6606cb340612cca169f8057a601718

Observation d2f7a809-7897-40ff-8f05-8ffa777f06f6 · outbound

This paper cites Speaker adaptation using spectro-temporal deep features for dysarthric and elderly speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Speaker adaptation using spectro-temporal deep features for dysarthric and elderly speech recognition,

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:52.961044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:47.136299Z digest=sha256:4096bb86c598490182799bc7e8f6dffac6ea9bafaa514a0e40c9cdf39dbfe1e9

Observation 6b58fa7c-73a7-4d23-bdf8-1f82f5f72748 · outbound

This paper cites Some statistical issues in the comparison of speech recognition algorithms,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Some statistical issues in the comparison of speech recognition algorithms,

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:52.783057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:47.141110Z digest=sha256:7bb359a818ae51a42f2e96492c945fa43a565174dbc250b77c36e4bff4caaaf7

Observation b9a8482f-8ffd-4bea-ac15-1162bb3c1154 · outbound

This paper cites SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition,

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:52.623835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:47.221732Z digest=sha256:fb3c470127f55433bfe7de8a7f982f8fe8831fe7defa040d924a7e692414f024

Observation 4c969da5-5255-4ec0-93e0-c7c517cc5419 · outbound

This paper cites Tools for the analysis of benchmark speech recognition tests,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Tools for the analysis of benchmark speech recognition tests,

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:52.403485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:47.380300Z digest=sha256:6af7e99e4527bb221382a312edd865bcd2db67fc3974058cc3d9146d4db39f0a

Observation 44c0c64e-9953-4c8c-8a7e-eb64a6067c21 · outbound

This paper cites Exploiting Cross-domain And Cross-Lingual Ultrasound Tongue Imaging Features For Elderly And Dysarthric Speech Recog- nition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Exploiting Cross-domain And Cross-Lingual Ultrasound Tongue Imaging Features For Elderly And Dysarthric Speech Recog- nition,

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:52.219031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:47.521344Z digest=sha256:a3d75e314c52886c50c6cf8353f522e65d0cad36d60a41515f4074428423e0ad

Observation ff45c5f5-5b0f-4e81-8a1a-ce0dfad35a06 · outbound

This paper cites Homogeneous speaker features for on-the-fly dysarthric and elderly speaker adaptation and speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Homogeneous speaker features for on-the-fly dysarthric and elderly speaker adaptation and speech recognition,

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:52.029574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:47.653179Z digest=sha256:926552a3c8d7ea97354bfbc8d80c75e512b3e557426b4221a31d2551aea30243

Observation bc9f7db8-6cba-4681-9c3c-de2bb60715e9 · outbound

This paper cites Bayesian Parametric and Architectural Domain Adap- tation of LF-MMI Trained TDNNs for Elderly and Dysarthric Speech Recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Bayesian Parametric and Architectural Domain Adap- tation of LF-MMI Trained TDNNs for Elderly and Dysarthric Speech Recognition,

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:51.813760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:47.808713Z digest=sha256:8e4256af30f042e3a6b9837902bb664d25b0c2fc2d7a50e413bd60f32194ba0d

Observation 65f23f34-7a71-4544-b5e4-7a1fe7f6ee55 · outbound

This paper cites Conformer Based Elderly Speech Recognition System for Alzheimer’s Disease Detection,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Conformer Based Elderly Speech Recognition System for Alzheimer’s Disease Detection,

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:51.616251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:48.005043Z digest=sha256:affa369f488a25e7569f5a65f67cb086840621d01ba49059e723796348a868bb

Observation fc64bde5-cc7b-458a-adf0-b7ea595b99f4 · outbound

This paper cites Hyper-parameter Adaptation of Conformer ASR Sys- tems for Elderly and Dysarthric Speech Recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Hyper-parameter Adaptation of Conformer ASR Sys- tems for Elderly and Dysarthric Speech Recognition,

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:51.418413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T20:31:48.143276Z digest=sha256:4186f57e0760aabff396cba7e0f8586e0e1b323c853e14a5cc0e9022d0e18db9

Pith citing papers

No inbound Pith citation observations are available.