Pith. sign in

Paper Citation Record · LEDGER

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks

As of 7 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 2 inbound Pith citation observations for arXiv:2506.23049.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.23049 v1

Coverage vector

measured 34 of 34 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:55:55.302080Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-21T02:08:06.976461Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T02:09:24.310163Z

Reference resolution

34 of 34 outbound references displayed

  • verified exact1
  • verified fuzzy10
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6a050def-1168-4fe9-8295-ecffce12c92a · outbound

This paper cites Espnet-sds: Unified toolkit and demo for spoken dialogue systems,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Espnet-sds: Unified toolkit and demo for spoken dialogue systems,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:55:57.345167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:55:53.060791Z digest=sha256:268a1780ead9d1250f08fd5c69543986454e9474db60e926de5b5f00939352a2

Observation 0dc7d8ef-a7bb-4d91-b572-23ebe7d1ba76 · outbound

This paper cites SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:53.134838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:53.134838Z digest=sha256:e937dc3867be56597112115659b5a4a7b18f2ff3d3d4d20501db9403cfaac106

Observation be26c935-a7ac-46cc-8640-094e2ff6b65b · outbound

This paper cites Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:53.227261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:53.227261Z digest=sha256:5b49c2a67c8d9dba6608c554a41d62e7904d9526c9f4534f328e28f84f35ba5d

Observation 83cfe486-8963-4984-b578-6021b9fedb3c · outbound

This paper cites Mini-Omni: Language Models Can Hear, Talk While Thinking in Streaming.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Mini-Omni: Language Models Can Hear, Talk While Thinking in Streaming

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:53.426594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:53.426594Z digest=sha256:5fc3c10e970c09f977dbdd56a3b26f89ee72100c28826d601f6cf1f247f20fb2

Observation 6c4f318d-5e12-4868-9bfc-5745befd587c · outbound

This paper cites Openomni: Advancing open-source omnimodal large language models with progressive multimodal alignment and real-time self-aware emotional speech synthesis,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Openomni: Advancing open-source omnimodal large language models with progressive multimodal alignment and real-time self-aware emotional speech synthesis,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:53.535944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:53.535944Z digest=sha256:32d2b0139248ee2b6593766885d207874065af71cadb74cecd7f40fa609a0b5b

Observation 29136298-6744-4ed0-ad02-6f8d95d881cf · outbound

This paper cites VoiceBench: Benchmarking LLM-Based Voice Assistants.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks VoiceBench: Benchmarking LLM-Based Voice Assistants

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:53.632465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:53.632465Z digest=sha256:9f87955d254ea043de2b5a6022776e4608815d006a6e99b734b0fb18bc56f574

Observation 0a0bb2e1-e7cd-43f1-8433-9cbfe2ada353 · outbound

This paper cites MultiWOZ -- A Large-Scale Multi-Domain Wizard-of-Oz Dataset for Task-Oriented Dialogue Modelling.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks MultiWOZ -- A Large-Scale Multi-Domain Wizard-of-Oz Dataset for Task-Oriented Dialogue Modelling

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:53.745815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:53.745815Z digest=sha256:dbd85e932a61b2098b15547358d16adc7b77e82275242af78510839545a91057

Observation 1729b4ea-7e13-4fd2-8e35-6e6537c543bc · outbound

This paper cites SpokenWOZ: A Large-Scale Speech-Text Benchmark for Spoken Task-Oriented Dialogue Agents.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks SpokenWOZ: A Large-Scale Speech-Text Benchmark for Spoken Task-Oriented Dialogue Agents

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:53.858082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:53.858082Z digest=sha256:48e1532a51053c9f1bf9b9f18b6d7ef3ddf067f62a9783315587f1ee1aa30b52

Observation 1b3c17e9-60ec-4a67-93ca-76db04cd35ad · outbound

This paper cites Training language models to follow instructions with human feedback.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Training language models to follow instructions with human feedback

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:53.956373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:53.956373Z digest=sha256:afd9f07718d10e74347889ad44942cc1aabfb7fce124d5a52655bfef24e018ca

Observation 556260fb-0aa7-484a-98ae-0209d064c99e · outbound

This paper cites HuggingGPT: Solving AI Tasks with ChatGPT and its Friends in Hugging Face.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks HuggingGPT: Solving AI Tasks with ChatGPT and its Friends in Hugging Face

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:54.050660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:54.050660Z digest=sha256:86ff52ff1caadf5448a8fd2337b217c7f2a5879805354c82ce3d7a42cca60ef7

Observation 5e8bd2b4-b095-4eaa-96f2-ef35ab2f289f · outbound

This paper cites API-Bank: A Comprehensive Benchmark for Tool-Augmented LLMs.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks API-Bank: A Comprehensive Benchmark for Tool-Augmented LLMs

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:54.130776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:54.130776Z digest=sha256:fab72a70382374bcff430c0fd6ba6d434e39978ddf1bc87e6f90b8d9867fc1c7

Observation 44b5b849-70de-40c6-b1f0-aaa0cafb8fa0 · outbound

This paper cites ToolDial: Multi-turn Dialogue Generation Method for Tool-Augmented Language Models.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks ToolDial: Multi-turn Dialogue Generation Method for Tool-Augmented Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:54.171642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:54.171642Z digest=sha256:23a10d4df4b2be85f69af943aa7c41b7c71b06a5b30a6aa157938e514cb4f45f

Observation 6fb7d218-242a-4f0f-af21-6204b876b14c · outbound

This paper cites Rethinking task-oriented dialogue systems: From complex modularity to zero-shot autonomous agent,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Rethinking task-oriented dialogue systems: From complex modularity to zero-shot autonomous agent,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:55:57.237848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:55:54.215277Z digest=sha256:e62bc7f849cf13476626fece2aaa8d5ac44d5a1c4c31fbb22cba3e6180a08ddf

Observation a9230911-5ef5-4918-bd37-7749709b1904 · outbound

This paper cites React: Synergizing reasoning and acting in language models,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks React: Synergizing reasoning and acting in language models,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:54.285113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:54.285113Z digest=sha256:cce90e8925211ea5c5603dd0551e2fad15e726750913b60e0845e3d9dd75d0a3

Observation 899075e1-9ab3-44f9-bedd-970ffb07a5f2 · outbound

This paper cites Audio-cot: Exploring chain-of-thought reasoning in large audio language model,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Audio-cot: Exploring chain-of-thought reasoning in large audio language model,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:55:57.139353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:55:54.342505Z digest=sha256:dafc6874ad46305afa5d5ba71e3a0d7125f9f13fb2771c559f6a48129263af8d

Observation 75024118-a0c6-4ac4-b6bf-d45f86f1d41e · outbound

This paper cites Audio-reasoner: Improving reasoning capability in large audio language models,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Audio-reasoner: Improving reasoning capability in large audio language models,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:54.436315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:54.436315Z digest=sha256:2fdfc3a43659d2b92e6e28c81ae4042c004d18c07bb70df9fd5eb61dc95ee20c

Observation 13a23776-4513-41db-ab68-e23694a70cd5 · outbound

This paper cites Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:54.386404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:54.386404Z digest=sha256:cef3ff7550be85b176129683a4edf0abc2ed97a2c33e907cc7c580770fcad75d

Observation 5f6776fd-53f7-48e7-a592-eb9359226dfc · outbound

This paper cites Can a suit of armor conduct electricity? a new dataset for open book question answering,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Can a suit of armor conduct electricity? a new dataset for open book question answering,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:54.539532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:54.539532Z digest=sha256:371ad5a2e27a0d01567605450c8c4aefccbff39ed079c14428ef2dbde391e265

Observation e08f5fdd-a684-4b24-9a06-34a544179e27 · outbound

This paper cites ReSpAct: Harmonizing Reasoning, Speaking, and Acting Towards Building Large Language Model-Based Conversational AI Agents.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks ReSpAct: Harmonizing Reasoning, Speaking, and Acting Towards Building Large Language Model-Based Conversational AI Agents

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-06T21:55:55.474892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:55:54.473272Z digest=sha256:7be52dbc44cb134bb202af3c5e82f18d7c1a5ff66ade11a93bccd4e7a9878525

Observation 11404fd2-5124-4b62-a298-96fbd8fde36c · outbound

This paper cites Gpt-4o: Openai’s new multimodal flagship model,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Gpt-4o: Openai’s new multimodal flagship model,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:55:57.037215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:55:54.625379Z digest=sha256:f5bf435186345bb822b725fac2a235716654a016228bd81717e89dcd90bc25af

Observation 13a632cc-e96a-46d8-820f-e45d952320fc · outbound

This paper cites Robust Speech Recognition via Large-Scale Weak Supervision.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Robust Speech Recognition via Large-Scale Weak Supervision

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:54.575759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:54.575759Z digest=sha256:945c3c6183198d45ec2d6c0ca3ab6162a3df61b1bc29ec16911668f655f4df05

Observation 3705390a-447d-4e6f-b7d1-6b65e8ad02ce · outbound

This paper cites Espnet-TTS: Unified, reproducible, and integratable open source end-to-end text-to-speech toolkit,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Espnet-TTS: Unified, reproducible, and integratable open source end-to-end text-to-speech toolkit,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:55:56.795127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:55:54.714602Z digest=sha256:f66a029d59c2417329381eacc6166aa1b3eef87fd9c183b59884ddeb4d9c750a

Observation 37471f67-73b0-461a-8515-e273a18636a3 · outbound

This paper cites Owsm v3.1: Better and faster open whisper-style speech models based on e-branchformer,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Owsm v3.1: Better and faster open whisper-style speech models based on e-branchformer,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:55:56.937871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:55:54.650824Z digest=sha256:86e832ec293d780b2b873971dcee3718ceb55321d9892c5290222a40be9336ef

Observation 66ba6451-65ec-41fd-9e2c-1ef0c70223ec · outbound

This paper cites Efficient Memory Management for Large Language Model Serving with PagedAttention.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Efficient Memory Management for Large Language Model Serving with PagedAttention

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:54.808526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:54.808526Z digest=sha256:1a756c3e8f4ae20ce421c1970b6db0664edd794d5f5781069e20de2ecb4e285b

Observation 4d4e4b70-79e5-4f2e-bca8-d5e1e66c4115 · outbound

This paper cites The Llama 3 Herd of Models.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks The Llama 3 Herd of Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:54.758002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:54.758002Z digest=sha256:bee46a1ca4eb81e41564e12ba8e1821d9d521ee03447a46df3c61fb6c1fbe960

Observation c741c3e5-60be-4f0c-a95b-f7eb3d2c3196 · outbound

This paper cites Gpt-4o: Openai’s new multimodal flagship model,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Gpt-4o: Openai’s new multimodal flagship model,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:55:56.399442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:55:54.920074Z digest=sha256:d5c5758eb88290e5b2ad3ff3bfd3da0c49c15b9392479f71448e18333e984cd8

Observation 64c3b58c-b427-4d79-946b-af719df12293 · outbound

This paper cites Alpacaeval: An automatic evaluator of instruction-following models,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Alpacaeval: An automatic evaluator of instruction-following models,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:55:56.544396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:55:54.850695Z digest=sha256:eed3ad1e7f9c7f7cc60f70251ab64054eb65b5de32e6d6fdee73a35f32f86c6e

Observation 4c43f8dc-f9de-4f62-8a4b-20a10ca0f237 · outbound

This paper cites Moshi: a speech-text foundation model for real-time dialogue.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Moshi: a speech-text foundation model for real-time dialogue

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:55.026534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:55.026534Z digest=sha256:557b9f0b7cc1df42ba8bad62c770d6a010cea6b42af5a916222665760a0f3b6f

Observation 32a86f2f-3495-438e-ad5a-6029d8f42bd6 · outbound

This paper cites Gpt-4o: Openai’s new multimodal flagship model,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Gpt-4o: Openai’s new multimodal flagship model,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:55:56.224799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:55:54.953118Z digest=sha256:a953179689da7ea16cae1c301c075f168ee37a9677a10924aa9592f080ef0f5b

Observation f6d58d67-3f07-4053-a581-dfd13108f015 · outbound

This paper cites Kimi-Audio Technical Report.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Kimi-Audio Technical Report

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:55.166303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:55.166303Z digest=sha256:f6499ec8ab8be0b5353e45cc9a553e5106ed6ab6b02ab41453fa82759f9865c3

Observation 63d253aa-d644-45da-a7b8-a64c00d06d94 · outbound

This paper cites Mini-Omni2: Towards Open-source GPT-4o with Vision, Speech and Duplex Capabilities.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Mini-Omni2: Towards Open-source GPT-4o with Vision, Speech and Duplex Capabilities

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:55.082723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:55.082723Z digest=sha256:9be6fe9a0e1096b9a11d6abbbbebea273de3a69f8f938fd3c44a96321221d0e0

Observation 342d84ad-921d-4ceb-a81f-bd574f5321f8 · outbound

This paper cites Parakeet-tdt-0.6b-v2,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Parakeet-tdt-0.6b-v2,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:55:56.017766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:55:55.302080Z digest=sha256:999526de59f5fc30bf7488485cbb6fded042fab77fba8068f03d8242864fc135

Observation 7640f52d-8697-41b2-b649-200ececb7083 · outbound

This paper cites Qwen3 Technical Report.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Qwen3 Technical Report

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:55.221863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:55.221863Z digest=sha256:6cf462bbdd627a29a95dd1c8a1d27df5ab1e618b0b33b1c0374cb762bf7f3bb0

Observation 853e5859-4f46-4486-a17e-4cc553b47172 · outbound

This paper cites ESPnet-SDS: Unified Toolkit and Demo for Spoken Dialogue Systems.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks ESPnet-SDS: Unified Toolkit and Demo for Spoken Dialogue Systems

Reference 2025

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T21:55:55.822601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:55:53.100157Z digest=sha256:5b17a23f8eb40efff6adc4bbdbde42d45e837d5fda4cc6bc016fc88011d7234a

Pith citing papers

Observation 9231afea-d3b6-4e49-9b70-35d3f6ae2021 · inbound

From Reactive to Proactive: Assessing the Proactivity of Voice Agents via ProVoice-Bench cites this paper.

From Reactive to Proactive: Assessing the Proactivity of Voice Agents via ProVoice-Bench AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.002818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T11:44:12.373082Z digest=sha256:b71e58cbacc959ec319416f1ca24f118e527693605f557905827b7d0bf273044

Observation 878f8158-8b47-4503-8238-062795f7805e · inbound

A Survey of Audio Reasoning in Multimodal Foundation Models cites this paper.

A Survey of Audio Reasoning in Multimodal Foundation Models AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks

Reference 109

Resolution
verified exact
arxiv_id, observed 2026-05-21T02:09:24.313261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T02:08:06.976461Z digest=sha256:2ef69854127f8c08f423ab5f03aa68d7412b50172470b0a3dc4cd7049ce4bb98