Pith. sign in

Paper Citation Record · LEDGER

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation

As of 6 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 2 inbound Pith citation observations for arXiv:2605.29430.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.29430 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-29T07:30:11.718647Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T00:41:58.099350Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T00:42:03.927134Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact18
  • verified fuzzy0
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a80f878b-efd0-4d4a-b948-e1de831c42f7 · outbound

This paper cites Automatic recognition of spoken digits,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Automatic recognition of spoken digits,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:ea953e125d0b8192b4f9891145ad988c0638c9d2e4fc9bf909493273f647adfe

Observation 83e885d3-fbf9-4d0c-b77e-7f5d02a9926c · outbound

This paper cites Slm: Bridge the thin gap between speech and text foundation models,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Slm: Bridge the thin gap between speech and text foundation models,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:eabd6d773b54586388133ca497a16d01ca2927335dfd6897f71d0a14e1d50467

Observation 7985b1d0-131a-47f5-a507-dddbbf1fab99 · outbound

This paper cites Qwen3-ASR Technical Report.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Qwen3-ASR Technical Report

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-06-29T07:33:13.676721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:7ca142a7acd667a5724bbc2db8015b9969724029a7a70cef0c86891d9a97ff5d

Observation 3b47f032-012f-417b-95d9-2714668eb182 · outbound

This paper cites Listen, attend and spell: A neural network for large vocabulary conversational speech recognition,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Listen, attend and spell: A neural network for large vocabulary conversational speech recognition,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:ec2104c5da272784072ca8ef9c211a3aa87002ec918a4d114d6c350b4e44b458

Observation efff2c18-217a-4043-8404-f18286dc0108 · outbound

This paper cites Robust speech recognition via large-scale weak supervi- sion,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Robust speech recognition via large-scale weak supervi- sion,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:ba5d80919e8fd8fa7a39dfbefa19759bc74ec407c3249f7291b92f473d405770

Observation 6a42cb78-7de2-4228-aeba-f70578effd85 · outbound

This paper cites Grounding in communication,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Grounding in communication,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:539a449cbf00f2747c367961acc3116b5764656fb25c99d0775a1c1f81c7bc3b

Observation 2a7dc854-c50a-4145-9fb9-fae13561ac13 · outbound

This paper cites The preference for self- correction in the organization of repair in conversation,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation The preference for self- correction in the organization of repair in conversation,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:510dfd53e0595badd9fbe7b31f8319e5b3c70bc8e62f66103b29b72b4b3e3baa

Observation 4f6ba823-a7f5-4e09-b115-d1f10ae8a27a · outbound

This paper cites How to evaluate asr output for named entity recognition?.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation How to evaluate asr output for named entity recognition?

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:29eed8f7dcb3e4806de26aa9e2952b8a759b731407e004af28824b692a8ce7e4

Observation cbb5d508-2716-4c7a-a06e-1da9bf0f58b2 · outbound

This paper cites Is word error rate a good indicator for spoken language understanding accuracy,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Is word error rate a good indicator for spoken language understanding accuracy,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:25b50b556952932c7044b8ee2125790b4ca7b4688b6e0551c00e4e922dd0e4f0

Observation f1a163cb-d02f-4540-be76-23936a45b2f0 · outbound

This paper cites Jelinek,Statistical methods for speech recognition.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Jelinek,Statistical methods for speech recognition

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:ec9fafc8e2f11dba27bf40bec25a05ad560463a8a95d9807283bde6e29b05a7d

Observation 31ac5f47-ce76-4cba-9606-476ac477e46b · outbound

This paper cites Semantic-WER: A Unified Metric for the Evaluation of ASR Transcript for End Usability.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Semantic-WER: A Unified Metric for the Evaluation of ASR Transcript for End Usability

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T07:33:13.679947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:3514de3eba9bcaf2b7d7724c2a44096fb0e0fec35d1640c354439bd6c25ddb9a

Observation 831c2531-bb1d-4a96-b665-84ee7a77f1f7 · outbound

This paper cites Semantic Distance: A New Metric for ASR Performance Analysis Towards Spoken Language Understanding.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Semantic Distance: A New Metric for ASR Performance Analysis Towards Spoken Language Understanding

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.677790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:8b4ca73b3c181e20db10060422d284d2570505e068e32158ed693a969d8ed4cc

Observation 1e37cb97-b24d-4227-af4d-db6dd46b0f85 · outbound

This paper cites Automatic estimation of word significance oriented for speech-based information retrieval,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Automatic estimation of word significance oriented for speech-based information retrieval,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:a5d113a96cc5c485a17cc4334f3e1fcfed842e7903c25b9cd4b8c5584ffe0e9a

Observation 8bf49f3a-f214-46d6-bcf8-2b09127398ba · outbound

This paper cites Heval: A new hybrid evaluation metric for automatic speech recognition tasks,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Heval: A new hybrid evaluation metric for automatic speech recognition tasks,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:7d0078c4a8b7116c8969187d4911ff6f4c9ff71e296237666a538df976a28c06

Observation 4def9281-c34b-46dd-9e4b-b38b1ddf9029 · outbound

This paper cites BERTScore: Evaluating Text Generation with BERT.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation BERTScore: Evaluating Text Generation with BERT

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-06-29T07:33:13.715084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:a4cc1ae66624277c287c9a598ed2e3b02f9813373ff5fe9c36b1ba12333ef272

Observation ab7e752c-117c-4e31-b06a-cf0ba1abd488 · outbound

This paper cites Laser: An llm-based asr scoring and evaluation rubric,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Laser: An llm-based asr scoring and evaluation rubric,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:3a2d3ab5d360493560ce9068a722a0ccf2e70ef1a4e02b727bf169981d95577d

Observation 6e548d9e-e4f8-479b-9e86-328dd4422431 · outbound

This paper cites An approach to measuring the performance of Automatic Speech Recognition (ASR) models in the context of Large Language Model (LLM) powered applications.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation An approach to measuring the performance of Automatic Speech Recognition (ASR) models in the context of Large Language Model (LLM) powered applications

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.710757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:c107098e955731910e4aff5d7c6f1a413166d62417349d32f57b9682296492ab

Observation 0283a8bb-b2c8-4375-b9ba-15798062f03a · outbound

This paper cites Multimodal error correction for speech user interfaces,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Multimodal error correction for speech user interfaces,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:d09c0d2bd906d30fb567bbc19c80fcdaa1aa0647d7dbab6f5e29bc19cca708b5

Observation b32d745b-b43b-48dd-9f2d-c123806e4826 · outbound

This paper cites V oice typing: a new speech interaction model for dictation on touchscreen devices,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation V oice typing: a new speech interaction model for dictation on touchscreen devices,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:f697f153326ba8577e4b8350d4518ffb8e2b44bb457cc2dfc62cd5cd7649d629

Observation 3d5033ef-cdf1-4e62-8a42-d56c13c8d909 · outbound

This paper cites Ef- ficient speech transcription through respeaking.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Ef- ficient speech transcription through respeaking

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:42e28d472154070888306a1e0d6a1a684b58b0554a538b4af2219fa9f1de2143

Observation c2a0ae98-0c42-420d-aff0-01ebc77aa388 · outbound

This paper cites The Gift of Feedback: Improving ASR Model Quality by Learning from User Corrections through Federated Learning.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation The Gift of Feedback: Improving ASR Model Quality by Learning from User Corrections through Federated Learning

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.719822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:9e741ff60f62f83079b580ff1311b07cb5b5d07ab87da31f6febc62f8f352e4f

Observation 1aabef63-0bc8-43aa-80a5-fc37eb48afdc · outbound

This paper cites React: Synergizing reasoning and acting in language models,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation React: Synergizing reasoning and acting in language models,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:be4029c877eba90ed1f6b535c33830fa8fc371f434f56217e9e1e98e76d44ac9

Observation cdc9bad7-0e20-4c1d-943f-f29a3aa5fc53 · outbound

This paper cites Large Language Models Are State-of-the-Art Evaluators of Translation Quality.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Large Language Models Are State-of-the-Art Evaluators of Translation Quality

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T07:33:13.692874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:a8b7f3659db77119ce5addfecf3784189bb9856d20780f069680f89892ebd952

Observation bef01f40-086c-4ada-b463-97055944eb4a · outbound

This paper cites Judging llm-as-a-judge with mt-bench and chatbot arena,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Judging llm-as-a-judge with mt-bench and chatbot arena,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:64b5fa8bd4bbf74e24b9deb162767d795ae26839d39d18b92631541d5f3be0cf

Observation 956bd722-6beb-4302-81f1-fc84f25c0c35 · outbound

This paper cites G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-06-29T07:33:13.704387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:ace0c40d66787da44e0da8c471a36c82a8cafff24b067c9051105f7a5116c59e

Observation eb18251f-3739-4434-8640-d8307a67368b · outbound

This paper cites Evaluating speech recognition perfor- mance towards large language model based voice assistants,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Evaluating speech recognition perfor- mance towards large language model based voice assistants,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:6cc0fabc0187685b41b6e3ab231b2aab3932c74630a5990a126f1a6b4ff5657e

Observation f62db5cf-c6a7-4f9a-8686-111d41d06fb0 · outbound

This paper cites Large language models as a proxy for human evaluation in assessing the comprehensibility of disordered speech transcription,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Large language models as a proxy for human evaluation in assessing the comprehensibility of disordered speech transcription,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:593535d153dfdd86004554c082d7ff277268107cb27de96e5bcd6b056affa23b

Observation 86ca9340-c4a6-432a-bc4e-402cbfb1262a · outbound

This paper cites Evaluating Large Language Models at Evaluating Instruction Following.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Evaluating Large Language Models at Evaluating Instruction Following

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.707801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:2eaae31e5924c5d3f573515b07de1f37f3157a24ce272576ef0bbf74209dcc9f

Observation ccfacce3-3f69-428f-b470-e3e7de33a320 · outbound

This paper cites Judging the judges: A systematic study of position bias in llm-as-a-judge.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Judging the judges: A systematic study of position bias in llm-as-a-judge

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.702000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:bc4b03371ce34f3c05ba7b81ad9c1cd4150d87545a9a47b7ee1bd2387511588b

Observation 41ee52e8-735f-4a8e-be94-78f10ecc1b31 · outbound

This paper cites AISHELL-NER: Named Entity Recognition from Chinese Speech.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation AISHELL-NER: Named Entity Recognition from Chinese Speech

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.716653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:62b93d431473b9aa56aa51f5ce6bbd9744ba4815171b38907db81de73d655687

Observation 47bdaa9f-0171-403b-9500-97ff09ce30b7 · outbound

This paper cites Code-Switching in End-to-End Automatic Speech Recognition: A Systematic Literature Review.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Code-Switching in End-to-End Automatic Speech Recognition: A Systematic Literature Review

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.726044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:2604aa8e364d9accbab1b1c9f98e97c540c2d3b9ab0c396cc606416133053585

Observation 797e1605-a7f9-4a1d-94bf-01485999b4b1 · outbound

This paper cites Qwen3 Technical Report.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Qwen3 Technical Report

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-06-29T07:33:13.670925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:22f9e97151c4fa2edde8171ce0a0f06d647f39e7731610d253d98eca9b84ebbd

Observation c8a2368c-09d1-4a4a-85c6-16870403ee90 · outbound

This paper cites IndexTTS: An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation IndexTTS: An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.683613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:254e7c06e4362815afc4870e29d0ef6e38cd30c9ab857c0692017ebb3976650b

Observation 7b8c78ff-3fb0-44c9-8187-4d9844e2d9c2 · outbound

This paper cites GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.728960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:33d67c66f222cc3519615fe822fb43c4eabc580967c079a619fbbc3a30b34a82

Observation 368cb05c-aa02-41d9-b07b-6349648f31c3 · outbound

This paper cites Wenetspeech: A 10000+ hours multi-domain mandarin corpus for speech recognition,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Wenetspeech: A 10000+ hours multi-domain mandarin corpus for speech recognition,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:5202ffe14100c56425e89ff3021cd5255b983d97b0266cab80576d2e676ec29f

Observation 25f914eb-55d1-42e0-a004-c416044ac839 · outbound

This paper cites AISHELL-1: An Open-Source Mandarin Speech Corpus and A Speech Recognition Baseline.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation AISHELL-1: An Open-Source Mandarin Speech Corpus and A Speech Recognition Baseline

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-06-29T07:33:13.722668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:50a115272bd22fd4a6a20b0cc6b47ed879708774bfdbb1bd381f3db928442afb

Observation f46561d3-6ed3-4122-8673-09576d05063d · outbound

This paper cites The ASRU 2019 Mandarin-English Code-Switching Speech Recognition Challenge: Open Datasets, Tracks, Methods and Results.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation The ASRU 2019 Mandarin-English Code-Switching Speech Recognition Challenge: Open Datasets, Tracks, Methods and Results

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.709713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:a38b880b6ed3ebf81e7d0cb4ca34eea14905ca2e8b4a14e68d8f447b9b237677

Observation b36ce334-f5b4-43e3-98c2-b2df15955548 · outbound

This paper cites CS-Dialogue: A 104-Hour Dataset of Spontaneous Mandarin-English Code-Switching Dialogues for Speech Recognition.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation CS-Dialogue: A 104-Hour Dataset of Spontaneous Mandarin-English Code-Switching Dialogues for Speech Recognition

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.707090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:a122aa3d5ffa30738155499d4fc75670a20edfac8e24c30ae9380a78c5c8180b

Observation b2d20a84-171f-4c9b-8511-d67383474cee · outbound

This paper cites Zechner and K.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Zechner and K

Reference 39

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:8384572f60221c6240639ad2b88ba6130f373f7f325d849005e68f5965bec273

Observation 18563bf5-3c9d-4dda-a34c-169602f1da5c · outbound

This paper cites Automated Speech Scoring System Under The Lens: Evaluating and interpreting the linguistic cues for language proficiency.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Automated Speech Scoring System Under The Lens: Evaluating and interpreting the linguistic cues for language proficiency

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.695666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:f652ddd66321ddf349f052048319acf35db9dd5e5f4aee8563414a31096f4c90

Observation b8565199-fdc6-43b9-9789-73c13c95fca6 · outbound

This paper cites Vii. note on regression and inheritance in the case of two parents,.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation Vii. note on regression and inheritance in the case of two parents,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-06-29T07:30:11.718647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:dfcf3ef02478d9e3a88afc63762e238ca0c66dd9c2ba25906ce32c934168b0bb

Observation 409ae099-8b61-45de-8e73-82819403d8ea · outbound

This paper cites FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.699366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:3be388743a8848af5dcbf672a05cf46ab7e90042abab5c2cdbca9c0144c2f5ff

Pith citing papers

Observation 721d140f-dfa4-418e-b732-5a8e9b0531cf · inbound

AgenticASR: Refining Speech Recognition in Real-World Scenarios via an Agentic Approach cites this paper.

AgenticASR: Refining Speech Recognition in Real-World Scenarios via an Agentic Approach Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-31T15:38:32.655980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T15:38:32.655980Z digest=sha256:14e195db87842e410621eccae46690cf0c15caf54bae6e2dbf8d1356cca9c348

Observation 34c3857b-52c2-4752-b0cd-c54d73bae652 · inbound

CallScreenBench: Benchmarking On-Device Models as Phone Secretaries cites this paper.

CallScreenBench: Benchmarking On-Device Models as Phone Secretaries Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation

Reference 20

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T00:42:04.022861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T00:41:58.099350Z digest=sha256:29fdd8f8795ff01964d40e06c3ee7b99ae83cf15b49c54eb460d46d5655625c9