Pith. sign in

Paper Citation Record · LEDGER

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation

As of 22 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 2 inbound Pith citation observations for arXiv:2605.29430.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.29430 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-29T07:30:11.718647Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T00:41:58.099350Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T00:42:03.927134Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact18
  • verified fuzzy0
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a80f878b-efd0-4d4a-b948-e1de831c42f7 · outbound

This paper cites Automatic recognition of spoken digits,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Automatic recognition of spoken digits,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:2b5b3d30fe6b37b7da1eab7fc5f51b1b758a85d4bf8c06588e8818f274bb5681

Observation 83e885d3-fbf9-4d0c-b77e-7f5d02a9926c · outbound

This paper cites Slm: Bridge the thin gap between speech and text foundation models,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Slm: Bridge the thin gap between speech and text foundation models,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:c8885300748a2ba961aa9cd856436adb8b5b0c6466107a7ae54512d5d67feb82

Observation 7985b1d0-131a-47f5-a507-dddbbf1fab99 · outbound

This paper cites Qwen3-ASR Technical Report.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Qwen3-ASR Technical Report

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-06-29T07:33:13.676721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:bce6c97c380a6a683c25328475045a4b25fdedeeb1e6b70ea7a0c1ee955ffbda

Observation 3b47f032-012f-417b-95d9-2714668eb182 · outbound

This paper cites Listen, attend and spell: A neural network for large vocabulary conversational speech recognition,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Listen, attend and spell: A neural network for large vocabulary conversational speech recognition,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:0b3a71504e0bf17b5caa26974601ddb6d6a99b8950188a65ffd62c2a4eba72b4

Observation efff2c18-217a-4043-8404-f18286dc0108 · outbound

This paper cites Robust speech recognition via large-scale weak supervi- sion,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Robust speech recognition via large-scale weak supervi- sion,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:a01ea0a9890760a6d37237cd42e1325a59421eac0e7158c4f9c0255239c66dea

Observation 6a42cb78-7de2-4228-aeba-f70578effd85 · outbound

This paper cites Grounding in communication,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Grounding in communication,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:61b059e4add974187d5f95a858199495f62266eff8642c7bc13351f732be8f3e

Observation 2a7dc854-c50a-4145-9fb9-fae13561ac13 · outbound

This paper cites The preference for self- correction in the organization of repair in conversation,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation The preference for self- correction in the organization of repair in conversation,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:95ac7c5e408eee622aea8029db57ca21fafdf3974a95880285a9a02b5c3704eb

Observation 4f6ba823-a7f5-4e09-b115-d1f10ae8a27a · outbound

This paper cites How to evaluate asr output for named entity recognition?.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation How to evaluate asr output for named entity recognition?

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:e564c0f54af4bfefbd049077a34a7644dad6585ae7fc8993f01359992a9c6920

Observation cbb5d508-2716-4c7a-a06e-1da9bf0f58b2 · outbound

This paper cites Is word error rate a good indicator for spoken language understanding accuracy,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Is word error rate a good indicator for spoken language understanding accuracy,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:db67c2e97b05afb5c79d63c65ef6722ac28f1c1d6aeb70b59c93bfc0f49caa03

Observation f1a163cb-d02f-4540-be76-23936a45b2f0 · outbound

This paper cites Jelinek,Statistical methods for speech recognition.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Jelinek,Statistical methods for speech recognition

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:4b67fb4e9b5092885b2d1b46c0ff933e56b3eaacb59ce7bf4b375396f8f555fb

Observation 31ac5f47-ce76-4cba-9606-476ac477e46b · outbound

This paper cites Semantic-WER: A Unified Metric for the Evaluation of ASR Transcript for End Usability.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Semantic-WER: A Unified Metric for the Evaluation of ASR Transcript for End Usability

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T07:33:13.679947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:e077c7620b1136e3c03f6b77bdff08b5648ec709aa11bc1e43c345dbba82bfa2

Observation 831c2531-bb1d-4a96-b665-84ee7a77f1f7 · outbound

This paper cites Semantic Distance: A New Metric for ASR Performance Analysis Towards Spoken Language Understanding.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Semantic Distance: A New Metric for ASR Performance Analysis Towards Spoken Language Understanding

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.677790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:5d024ff9f1e70ae06021f817f6a97c7f71353b2584b7f93bbc67a380b55f3d2c

Observation 1e37cb97-b24d-4227-af4d-db6dd46b0f85 · outbound

This paper cites Automatic estimation of word significance oriented for speech-based information retrieval,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Automatic estimation of word significance oriented for speech-based information retrieval,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:ac3c10c678c1741980a15fead97ff058d0c3ca9a90d78cf895af9d5a1eeb2a8f

Observation 8bf49f3a-f214-46d6-bcf8-2b09127398ba · outbound

This paper cites Heval: A new hybrid evaluation metric for automatic speech recognition tasks,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Heval: A new hybrid evaluation metric for automatic speech recognition tasks,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:55742de1a0eeb897dbf4c13dfb95a052a94abfdcaf41a26fc8b10d2cf7ca88b3

Observation 4def9281-c34b-46dd-9e4b-b38b1ddf9029 · outbound

This paper cites BERTScore: Evaluating Text Generation with BERT.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation BERTScore: Evaluating Text Generation with BERT

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-06-29T07:33:13.715084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T10:38:08.909666+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:9dc455cb6fd4d18b73364259cae683e3adcd6944f54676794a0ced4a8617235e

Observation ab7e752c-117c-4e31-b06a-cf0ba1abd488 · outbound

This paper cites Laser: An llm-based asr scoring and evaluation rubric,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Laser: An llm-based asr scoring and evaluation rubric,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:c8b4afa4a7d39b717c0cdd538ffc032f655f255235b866a8579f72d46a4939aa

Observation 6e548d9e-e4f8-479b-9e86-328dd4422431 · outbound

This paper cites An approach to measuring the performance of Automatic Speech Recognition (ASR) models in the context of Large Language Model (LLM) powered applications.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation An approach to measuring the performance of Automatic Speech Recognition (ASR) models in the context of Large Language Model (LLM) powered applications

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.710757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:9fbcd8ba90bc7911c826305f0b3e431c9d7a54ea2e56aec4617757d0a2d78446

Observation 0283a8bb-b2c8-4375-b9ba-15798062f03a · outbound

This paper cites Multimodal error correction for speech user interfaces,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Multimodal error correction for speech user interfaces,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:1e99f0e5bd4aa1ce703eb47a62dc356ac6d23e1abb1e56a792071ceadf2868a2

Observation b32d745b-b43b-48dd-9f2d-c123806e4826 · outbound

This paper cites V oice typing: a new speech interaction model for dictation on touchscreen devices,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation V oice typing: a new speech interaction model for dictation on touchscreen devices,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:7e0509abaa7933b8ae24752c3f99ab048dc4af15334ec4133a41b9e5d6168b14

Observation 3d5033ef-cdf1-4e62-8a42-d56c13c8d909 · outbound

This paper cites Ef- ficient speech transcription through respeaking.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Ef- ficient speech transcription through respeaking

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:3f4784b36ee803697447ceeba50f93594316a46be860c58f84c86b61c774d119

Observation c2a0ae98-0c42-420d-aff0-01ebc77aa388 · outbound

This paper cites The Gift of Feedback: Improving ASR Model Quality by Learning from User Corrections through Federated Learning.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation The Gift of Feedback: Improving ASR Model Quality by Learning from User Corrections through Federated Learning

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.719822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:157c09bd51047cee659563c2d5f4c5f46e476c3ad004378a20d0d41bd67c0bcc

Observation 1aabef63-0bc8-43aa-80a5-fc37eb48afdc · outbound

This paper cites React: Synergizing reasoning and acting in language models,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation React: Synergizing reasoning and acting in language models,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:335dd55ef4c3f0810302c5d28549a38fc32ed0515c90345065e07719a90d5143

Observation cdc9bad7-0e20-4c1d-943f-f29a3aa5fc53 · outbound

This paper cites Large Language Models Are State-of-the-Art Evaluators of Translation Quality.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Large Language Models Are State-of-the-Art Evaluators of Translation Quality

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T07:33:13.692874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:14f63dd174a5d1219a5f5412c3c91efbb2043664d59f7c35d16dbc7f86ae6ba3

Observation bef01f40-086c-4ada-b463-97055944eb4a · outbound

This paper cites Judging llm-as-a-judge with mt-bench and chatbot arena,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Judging llm-as-a-judge with mt-bench and chatbot arena,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:1409b695fed83c6cabbd4637ab0c824482e36757b3c806f47aa520fe4ae3ca7e

Observation 956bd722-6beb-4302-81f1-fc84f25c0c35 · outbound

This paper cites G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-06-29T07:33:13.704387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:0c9c774504b184b1a7ee48f37319415c179f3c812c61b38c6770e23c33835e71

Observation eb18251f-3739-4434-8640-d8307a67368b · outbound

This paper cites Evaluating speech recognition perfor- mance towards large language model based voice assistants,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Evaluating speech recognition perfor- mance towards large language model based voice assistants,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:7ebdece9e3f01da60c82a971a9192e5c942051c78c3e0b6cbac975d81482286f

Observation f62db5cf-c6a7-4f9a-8686-111d41d06fb0 · outbound

This paper cites Large language models as a proxy for human evaluation in assessing the comprehensibility of disordered speech transcription,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Large language models as a proxy for human evaluation in assessing the comprehensibility of disordered speech transcription,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:c5b985df67cd6ca1e7dfa73a3c93a4d7f243fb93e0cdd5a9e5e6ec8f3fd7b043

Observation 86ca9340-c4a6-432a-bc4e-402cbfb1262a · outbound

This paper cites Evaluating Large Language Models at Evaluating Instruction Following.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Evaluating Large Language Models at Evaluating Instruction Following

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.707801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:5d90fe048cd651751122251950e5a8866453d4c546305b9a5b6594a9a391b06b

Observation ccfacce3-3f69-428f-b470-e3e7de33a320 · outbound

This paper cites Judging the judges: A systematic study of position bias in llm-as-a-judge.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Judging the judges: A systematic study of position bias in llm-as-a-judge

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.702000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:e54b6d8b9013fdd4820e4e138b2aef1be323a09ce891602feffb2974062e5a3d

Observation 41ee52e8-735f-4a8e-be94-78f10ecc1b31 · outbound

This paper cites AISHELL-NER: Named Entity Recognition from Chinese Speech.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation AISHELL-NER: Named Entity Recognition from Chinese Speech

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.716653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:8adf7dc4c6586122af5693ac7a6f2b2d9acc72db77711b8803d7c7c48cba137d

Observation 47bdaa9f-0171-403b-9500-97ff09ce30b7 · outbound

This paper cites Code-Switching in End-to-End Automatic Speech Recognition: A Systematic Literature Review.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Code-Switching in End-to-End Automatic Speech Recognition: A Systematic Literature Review

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.726044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:0be662730c7a6dbc23daf82d220834efc1e69c45e958d34625243637e860ca1c

Observation 797e1605-a7f9-4a1d-94bf-01485999b4b1 · outbound

This paper cites Qwen3 Technical Report.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Qwen3 Technical Report

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-06-29T07:33:13.670925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:8024051f558c51ce95fe80484e22c75a490c51f487423a16ac6d6e89dcdc2c47

Observation c8a2368c-09d1-4a4a-85c6-16870403ee90 · outbound

This paper cites IndexTTS: An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation IndexTTS: An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.683613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:8683d400cfa1afcab8ab723de321bfc80b57202aa84b14d5ad074facb343a8b4

Observation 7b8c78ff-3fb0-44c9-8187-4d9844e2d9c2 · outbound

This paper cites GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.728960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:b2a75cf4135040d5394538207e6f2de066ad3537d4acdf119d8beb244d30d51b

Observation 368cb05c-aa02-41d9-b07b-6349648f31c3 · outbound

This paper cites Wenetspeech: A 10000+ hours multi-domain mandarin corpus for speech recognition,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Wenetspeech: A 10000+ hours multi-domain mandarin corpus for speech recognition,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:9a5ac37438dfcccdfe7080940c1abdfbefd5afb0066e02a73006c26eb18d54e6

Observation 25f914eb-55d1-42e0-a004-c416044ac839 · outbound

This paper cites AISHELL-1: An Open-Source Mandarin Speech Corpus and A Speech Recognition Baseline.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation AISHELL-1: An Open-Source Mandarin Speech Corpus and A Speech Recognition Baseline

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-06-29T07:33:13.722668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:337efbd45d58bcff047a7f3d83e1e1a9670bfe3f70f9a94bcb721e4606f63ae2

Observation f46561d3-6ed3-4122-8673-09576d05063d · outbound

This paper cites The ASRU 2019 Mandarin-English Code-Switching Speech Recognition Challenge: Open Datasets, Tracks, Methods and Results.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation The ASRU 2019 Mandarin-English Code-Switching Speech Recognition Challenge: Open Datasets, Tracks, Methods and Results

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.709713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:31a094d3c45ed325f8ea3ad2c898b76866ff9d2f1fda599e0afd2717ca23adb7

Observation b36ce334-f5b4-43e3-98c2-b2df15955548 · outbound

This paper cites CS-Dialogue: A 104-Hour Dataset of Spontaneous Mandarin-English Code-Switching Dialogues for Speech Recognition.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation CS-Dialogue: A 104-Hour Dataset of Spontaneous Mandarin-English Code-Switching Dialogues for Speech Recognition

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.707090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:e30098f0cb6e35b0e6695db764d356b4ee8eedc1882dcf663888237c3999af7d

Observation b2d20a84-171f-4c9b-8511-d67383474cee · outbound

This paper cites Zechner and K.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Zechner and K

Reference 39

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:cd190a523b3d5847bbc2d46490035e4e05ac09d01feae76668319834b55f52c4

Observation 18563bf5-3c9d-4dda-a34c-169602f1da5c · outbound

This paper cites Automated Speech Scoring System Under The Lens: Evaluating and interpreting the linguistic cues for language proficiency.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Automated Speech Scoring System Under The Lens: Evaluating and interpreting the linguistic cues for language proficiency

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.695666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:32eb51611bec44612d205567c8c69efcdee080d1e9f8824486c449a2eb8ba87d

Observation b8565199-fdc6-43b9-9789-73c13c95fca6 · outbound

This paper cites Vii. note on regression and inheritance in the case of two parents,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Vii. note on regression and inheritance in the case of two parents,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:ca1e7dd145845884905a0ed3a4e29fd3fea60edeec4f255ed8de49d0ff4e101e

Observation 409ae099-8b61-45de-8e73-82819403d8ea · outbound

This paper cites FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.699366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:274f021baff474b98ff9b73b7bbcfe76d9540d4ca970f59e10a27c934f592839

Pith citing papers

Observation 721d140f-dfa4-418e-b732-5a8e9b0531cf · inbound

AgenticASR: Refining Speech Recognition in Real-World Scenarios via an Agentic Approach cites this paper.

AgenticASR: Refining Speech Recognition in Real-World Scenarios via an Agentic Approach Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-31T15:38:32.655980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T15:38:32.655980Z digest=sha256:b0bac0d2dc7e1d0c9585a610a0ca4a2b4c82465cdade3c73dc7e7b3e14c4350d

Observation 34c3857b-52c2-4752-b0cd-c54d73bae652 · inbound

CallScreenBench: Benchmarking On-Device Models as Phone Secretaries cites this paper.

CallScreenBench: Benchmarking On-Device Models as Phone Secretaries Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation

Reference 20

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T00:42:04.022861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-06T00:41:58.099350Z digest=sha256:fee89612ab85f9cbbd13eaad1720b8a0d6106bbaa48d4be174e62cae916ba049