Pith. sign in

Paper Citation Record · LEDGER

Joint Speech Recognition and Speaker Diarization via Sequence Transduction

As of 8 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 4 inbound Pith citation observations for arXiv:1907.05337.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1907.05337 v1

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-25T00:59:00.525481Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:32:27.969261Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-05-25T01:00:09.574967Z

Reference resolution

40 of 40 outbound references displayed

  • verified exact3
  • verified fuzzy31
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation df2148e4-efbe-47bf-9f6a-169c1c4e18ba · outbound

This paper cites Joint Speech Recognition and Speaker Diarization via Sequence Transduction.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Joint Speech Recognition and Speaker Diarization via Sequence Transduction

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-25T01:00:09.578155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:ba73700b183a0a9cd84d34fb76ccd4f0a99ccd51e568429e9edcf16cbaeb0991

Observation dd1d165d-6cab-4e0b-89f6-5acb1a9381b2 · outbound

This paper cites Problem Formulation and Proposed Solution Many machine learning tasks can be expressed as mapping an input sequence into an output sequence.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Problem Formulation and Proposed Solution Many machine learning tasks can be expressed as mapping an input sequence into an output sequence

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.300720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:30cbb22e49c4ff29d9c425a31f8de38f9e402132abd7ef284631109372cfe773

Observation d88a58d1-b5af-475b-8a69-feb4e749a31c · outbound

This paper cites an unresolved cited work.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-05-25T01:00:10.362086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:60588e679ede0e805820798e10f5d827ec2c94bfd10b5062a3f76659a557b62f

Observation 7993a0b8-0a36-4f24-a645-d9ff7c842276 · outbound

This paper cites an unresolved cited work.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-05-25T01:00:10.240446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:e441c460e331ce46b6b02be603a193eecf379de9752859376926512e98fc4afc

Observation 29d16264-0740-4a61-bd05-4d6862a3a615 · outbound

This paper cites an unresolved cited work.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-05-25T01:00:10.381972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:4c695c4ed0a358efc7bc60bd448c285474e7c4498d6586acb4f95d94b9ba8a3c

Observation 97e3ca92-1e81-4f84-bc05-55482e9c14ab · outbound

This paper cites an unresolved cited work.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-05-25T01:00:10.228658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:9353b076c768ed458fa1bcd4be34b5a94621aaa95e2a53ac8011dd1bfae4620c

Observation 98cd67ce-d600-4ad2-b90a-5ac506107060 · outbound

This paper cites an unresolved cited work.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-05-25T01:00:10.236456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:63304dece6d9329ae17ec2f7b8d99a8e7ad49f817bea6c84a274b2316405abed

Observation d3f6adbc-08e1-4b4d-8dc7-68c887a5b5a9 · outbound

This paper cites We demonstrated the performance of our approach by evaluating it on a large corpus of clinical conversa- tions between physicians and patients.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction We demonstrated the performance of our approach by evaluating it on a large corpus of clinical conversa- tions between physicians and patients

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.292305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:51a03d1e5f4aedd8a802095d17f1cf5a640dbad7861958257e9eefb31d9ed7ce

Observation 07b88adf-565a-46cb-a228-2d068797be7f · outbound

This paper cites an unresolved cited work.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-05-25T01:00:10.366683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:20738642a78281b3639aae0e011eea4066f3df75001b24bee97c42c025ba12ab

Observation 175393fa-3556-4888-ad78-2b18e7e2507a · outbound

This paper cites An overview of auto- matic speaker diarization systems.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction An overview of auto- matic speaker diarization systems

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.375332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:d7d5ef8cf144a061be6c850c7bb7316482f0087db6d89d5f2495b1ff585f70f8

Observation 5e9e5dc7-7e68-4d98-a67c-e79a4aedee57 · outbound

This paper cites Speaker diarization: A review of recent research.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Speaker diarization: A review of recent research

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.378649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:c5a133120362962f8501881acea1a10b2c15eefe6f2aebbd133902179bec26f0

Observation e92e4095-589b-462b-a11f-c7b73fd89946 · outbound

This paper cites A robust speaker clustering algo- rithm.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction A robust speaker clustering algo- rithm

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.350458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:7cd8868965cf035a579f7787bc7381d6ca92b0de9b861cde93ec6bfaa23fa917

Observation 23ca0396-8c5c-4579-b7e9-b1e22ac3d0ae · outbound

This paper cites Multistage speaker diarization of broadcast news.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Multistage speaker diarization of broadcast news

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.334842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:b26512f605b1f44c16d6f9704db9667dac5ceeb16a547571b638c318bf889285

Observation 8fd2b52d-a994-4177-ba77-294dc10d82d6 · outbound

This paper cites Speaker diarization with PLDA i-vector scoring and unsupervised calibration.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Speaker diarization with PLDA i-vector scoring and unsupervised calibration

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.287755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:e7f4dfe1cb05445b32e45c93b0040b576f938fdf1c7cf1b449865b1f50dde59c

Observation e487747d-1160-474c-b7ad-dbc58b25a482 · outbound

This paper cites Speaker diarization using deep neural network embeddings.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Speaker diarization using deep neural network embeddings

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.393043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:0c7eea832995be122efa7203e00d1720dcb2c5cdf8664bb3bcedffeb194409a8

Observation bea5c21e-eceb-46a1-9788-309e58dfb63e · outbound

This paper cites Speaker diarization with LSTM.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Speaker diarization with LSTM

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.338336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:1a3c7bb4f62260f627074a75287961f0e36c87b17a5c97698ed17c10ce19d459

Observation fc8682a6-ce90-476f-9257-60ca9c24f8b0 · outbound

This paper cites X-vectors: Robust DNN embeddings for speaker recogni- tion.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction X-vectors: Robust DNN embeddings for speaker recogni- tion

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.358453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:158b87a570b53df7697a8ba501ebbe4fd9574bd7b7a8ab2287637caee9ba78f7

Observation 8f3d676e-8def-4d48-82c4-8756d3ccabc1 · outbound

This paper cites Diarization is hard: Some experiences and lessons learned for the JHU team in the inaugural DIHARD chal- lenge.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Diarization is hard: Some experiences and lessons learned for the JHU team in the inaugural DIHARD chal- lenge

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.371970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:a2613d3dbd51b7cfe7d4246e69289889990061aad2e1b93411c6dac1be33ac64

Observation 8805d2e1-9e96-4d0e-b8a6-5d50f4562c44 · outbound

This paper cites Tristounet: Triplet loss for speaker turn embedding.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Tristounet: Triplet loss for speaker turn embedding

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.389529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:3d9852a2c782e69ff654d4814114388580f8a82e53fdc5bc0e9d2b964d52cba8

Observation 7d2a9063-ea67-4c88-84bb-a4d65a32540a · outbound

This paper cites Fully Supervised Speaker Diarization.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Fully Supervised Speaker Diarization

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-05-25T01:00:09.565790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:661bb5bbf54cfbf2bcd89d9b23dce23fe64af27ab6860d77c6d69c3510ee32d7

Observation cc1057fd-5d2d-46d6-97cc-be24bc01a83a · outbound

This paper cites Speaker di- arization from speech transcripts.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Speaker di- arization from speech transcripts

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.385332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:c24608d7bec29608fa098f8d77986677e555bcb47c8080ab57a2ba6fb39e6db4

Observation 7ec0a56b-c577-4a66-be1c-ac97361682bc · outbound

This paper cites Multimodal speaker segmentation and diarization using lexical and acoustic cues via sequence to se- quence neural networks.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Multimodal speaker segmentation and diarization using lexical and acoustic cues via sequence to se- quence neural networks

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.354194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:12f2145a01386106f418a62488430f2b314cc738176262e31edc8307a55249d2

Observation 42913212-c0ef-4597-b76e-e529b5aecee3 · outbound

This paper cites The use of recurrent neural networks in continuous speech recognition.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction The use of recurrent neural networks in continuous speech recognition

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.273558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:2370ee722cbba3f4950da48452785679f1697e2406b8cdeb7eb26edf1c8d9aa0

Observation f178c1d3-fdcf-4ecd-b75e-b7e18570297d · outbound

This paper cites Connectionist temporal classification: labelling unsegmented se- quence data with recurrent neural networks.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Connectionist temporal classification: labelling unsegmented se- quence data with recurrent neural networks

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.277767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:35bb551c96c76d55e5d24b8b3cda33d4a28677cd18ac2fdaef30cf1d109c8bf3

Observation bf21111b-9427-4ff4-b75b-b0132871c77a · outbound

This paper cites Sequence transduction with recurrent neural net- works.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Sequence transduction with recurrent neural net- works

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.282214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:f6f8fe833ab97a1f6c22cea54d212aa47cbee00a293efd17c462ab100f9af752

Observation 197d9a67-66ea-442a-80cc-916a7cfb3806 · outbound

This paper cites Speech recognition with deep recurrent neural networks.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Speech recognition with deep recurrent neural networks

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.232528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:0541c35346d7bc6c9da9d7fdbebd3d60cb1cbb45f03c521381755a6517007585

Observation 6eb82b7c-d5fa-4d52-a9cc-cd65ea5ee047 · outbound

This paper cites Streaming End-to-end Speech Recognition For Mobile Devices.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Streaming End-to-end Speech Recognition For Mobile Devices

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-05-25T01:00:09.571824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:bf105429e282d29ab49dd6433848b7772a057849b4d0458034f5c7035dd7c805

Observation eacc9491-ce87-4ba4-a3cd-4c130cd3a670 · outbound

This paper cites In- datacenter performance analysis of a tensor processing unit.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction In- datacenter performance analysis of a tensor processing unit

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.342565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:403757c126a680aeff4f05a45a4ceed0080225519c859973d78633f8d8efd374

Observation c5f8e694-5f06-469f-a95c-4dae748b3ed5 · outbound

This paper cites Improving the efficiency of forward-backward algorithm using batched computation in tensorflow.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Improving the efficiency of forward-backward algorithm using batched computation in tensorflow

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.327322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:3b55e7fe9292664c6379cfc3f1cc00eddb1546476fe71f809b348d297e441b8d

Observation 04feb0f1-01c7-4c3c-b3ac-b1108f51d59b · outbound

This paper cites Efficient implementation of recurrent neu- ral network transducer in tensorflow.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Efficient implementation of recurrent neu- ral network transducer in tensorflow

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.305342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:f1e545de7d2a7c2c3f6bbafd6fa7128f3c454a56522aef7969e9e98c17fdf6f5

Observation f513f50c-df4a-451d-af65-857af8e9acf2 · outbound

This paper cites Neural speech recognizer: Acoustic-to-word LSTM model for large vocabulary speech recognition.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Neural speech recognizer: Acoustic-to-word LSTM model for large vocabulary speech recognition

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.396781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:5a3e193cf0c72d119e3c2130d34caf08210721e43046d6bfe6dd5133172a45a6

Observation 9ec9e717-29f6-49a8-94b9-7bdfed362b15 · outbound

This paper cites Morfessor 2.0: Python implementation and extensions for morfessor base- line.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Morfessor 2.0: Python implementation and extensions for morfessor base- line

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.323290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:1a843677afa07e8f4db5bb191708c91bd19be12aa76864b05aad2f18dfa31dbb

Observation ddc453d7-d7fe-4be0-8c52-ea030f679765 · outbound

This paper cites Phoneme recognition using time-delay neural networks.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Phoneme recognition using time-delay neural networks

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.296298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:ad00d6a901c475d07e9a5c04a53a74091f75e35b81114908016733a139919734

Observation 77f9fa79-f727-49da-9c24-555afded2daa · outbound

This paper cites Reducing the computational complexity for whole word models.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Reducing the computational complexity for whole word models

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.331126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:c1286a762d76c5740daa857148ddcc68306bda2a4acc95a193d37cd8ffca28c8

Observation b0b87f38-9eb6-403b-9f69-ff41a168d1d7 · outbound

This paper cites Long short-term memory.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Long short-term memory

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.346710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:cd0107eeccda0a121f08dd38c07ccf8c6cbe9e6c93ef15106c3bd689df9cd9de

Observation 86aae55a-6dbc-4151-b93a-d76934d4abf6 · outbound

This paper cites Adam: A method for stochastic opti- mization.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Adam: A method for stochastic opti- mization

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.268808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:d1ce0847bbb82b0664af7d00e77acc518f3b53f29194b98e7dffe8c8e342d094

Observation 9cbfc75a-8ef7-479f-9943-34efdec55b15 · outbound

This paper cites The Rich Transcription Fall 2003 (RT-03F) Evalu- ation Plan.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction The Rich Transcription Fall 2003 (RT-03F) Evalu- ation Plan

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.244402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:5feac47a0111f19d1096a2a1afb3d5f41714e9ac06c01322fbd559ab33d47384

Observation 345aaecc-4234-4cd8-b7af-20877cdb534e · outbound

This paper cites Feature learn- ing with raw-waveform cldnns for voice activity detection.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Feature learn- ing with raw-waveform cldnns for voice activity detection

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.310789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:d2d89ba81ad0106bcc4f84ea494a1e1994351dea05a4f1a4725b9dc7c37f3612

Observation 2d39334c-f8b3-4a0e-befd-0d94e98cd608 · outbound

This paper cites End-to- end text-dependent speaker verification.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction End-to- end text-dependent speaker verification

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.314975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:3ab085396db6af95808a550e802367427e1f53c8e592b07c3fda774ad89bec50

Observation 7f4a7fb6-5a2c-4945-8ac6-00b760f5b051 · outbound

This paper cites V oxceleb2: Deep speaker recognition.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction V oxceleb2: Deep speaker recognition

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:00:10.318901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:5d70f4c016471be5ef30b0759451233fee024f8235396bff1a5c6d51bef92e31

Pith citing papers

Observation df2148e4-efbe-47bf-9f6a-169c1c4e18ba · inbound

Joint Speech Recognition and Speaker Diarization via Sequence Transduction cites this paper.

Joint Speech Recognition and Speaker Diarization via Sequence Transduction Joint Speech Recognition and Speaker Diarization via Sequence Transduction

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-25T01:00:09.578155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:59:00.525481Z digest=sha256:ba73700b183a0a9cd84d34fb76ccd4f0a99ccd51e568429e9edcf16cbaeb0991

Observation d9b4c47a-ad4d-405a-b808-72a72c24843f · inbound

Joint ASR and Speaker Role Tagging with Serialized Output Training cites this paper.

Joint ASR and Speaker Role Tagging with Serialized Output Training Joint Speech Recognition and Speaker Diarization via Sequence Transduction

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T04:32:27.969261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:32:27.969261Z digest=sha256:0f34cfc6ecab47f53f3cd2505ad8a5a2e89523eb19b1b1c53ed4d8bd29771196

Observation 80d5c565-14c0-4f84-9a26-578fc0102d5e · inbound

Do We Still Need Audio? Rethinking Speaker Diarization with a Text-Based Approach Using Multiple Prediction Models cites this paper.

Do We Still Need Audio? Rethinking Speaker Diarization with a Text-Based Approach Using Multiple Prediction Models Joint Speech Recognition and Speaker Diarization via Sequence Transduction

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T04:15:56.818718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:15:56.818718Z digest=sha256:6a1439e2c16c3943834c7f7c041a1ba8dcb435335a7b46d531cfb026ab32e27e

Observation 612daea9-4d17-4202-b167-cc2aa4c16239 · inbound

TokenVerse++: Towards Flexible Multitask Learning with Dynamic Task Activation cites this paper.

TokenVerse++: Towards Flexible Multitask Learning with Dynamic Task Activation Joint Speech Recognition and Speaker Diarization via Sequence Transduction

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T15:27:29.255358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:27:29.255358Z digest=sha256:83a3609bd08aaa267dc09fc55eaebc0c861d82de51622111b7ccd1c76f873f34