Pith. sign in

Paper Citation Record · LEDGER

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks

As of 10 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 2 inbound Pith citation observations for arXiv:2506.23049.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.23049 v1

Coverage vector

measured 34 of 34 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:55:55.302080Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-21T02:08:06.976461Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T02:09:24.310163Z

Reference resolution

34 of 34 outbound references displayed

  • verified exact1
  • verified fuzzy10
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6a050def-1168-4fe9-8295-ecffce12c92a · outbound

This paper cites Espnet-sds: Unified toolkit and demo for spoken dialogue systems,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Espnet-sds: Unified toolkit and demo for spoken dialogue systems,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:55:57.345167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T21:55:53.060791Z digest=sha256:a76a60c9576fe5061e8c61811ebab2a9fe0b2fe59a72ce3704ad2a273a54d704

Observation 0dc7d8ef-a7bb-4d91-b572-23ebe7d1ba76 · outbound

This paper cites SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:53.134838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:53.134838Z digest=sha256:2eca426dae14bb7b46e4b42259328dc552671ddb4806fc23a51ad8fe53bcae53

Observation be26c935-a7ac-46cc-8640-094e2ff6b65b · outbound

This paper cites Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:53.227261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:53.227261Z digest=sha256:059c0080795c2aa039511d0959d0cf60b7769cf722c32730e68617a8a5c0c69b

Observation 83cfe486-8963-4984-b578-6021b9fedb3c · outbound

This paper cites Mini-Omni: Language Models Can Hear, Talk While Thinking in Streaming.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Mini-Omni: Language Models Can Hear, Talk While Thinking in Streaming

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:53.426594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:53.426594Z digest=sha256:e00b61441a72427b6c06365d0ea5528001e6e84b82ebdc544536ee8055540403

Observation 6c4f318d-5e12-4868-9bfc-5745befd587c · outbound

This paper cites Openomni: Advancing open-source omnimodal large language models with progressive multimodal alignment and real-time self-aware emotional speech synthesis,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Openomni: Advancing open-source omnimodal large language models with progressive multimodal alignment and real-time self-aware emotional speech synthesis,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:53.535944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:53.535944Z digest=sha256:5d77732424afc06bf8d0fbc48b37cd1cd24b57816131906e11209ec289900a74

Observation 29136298-6744-4ed0-ad02-6f8d95d881cf · outbound

This paper cites VoiceBench: Benchmarking LLM-Based Voice Assistants.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks VoiceBench: Benchmarking LLM-Based Voice Assistants

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:53.632465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:53.632465Z digest=sha256:cff2603c4155e551ec02bf3a546a1052b4d03634aed31fb1540e24abe9514095

Observation 0a0bb2e1-e7cd-43f1-8433-9cbfe2ada353 · outbound

This paper cites MultiWOZ -- A Large-Scale Multi-Domain Wizard-of-Oz Dataset for Task-Oriented Dialogue Modelling.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks MultiWOZ -- A Large-Scale Multi-Domain Wizard-of-Oz Dataset for Task-Oriented Dialogue Modelling

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:53.745815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:53.745815Z digest=sha256:d29c1c207fe9826fdb09b135d86f3663832072927a7a2ffcd037cbfae3a10ab8

Observation 1729b4ea-7e13-4fd2-8e35-6e6537c543bc · outbound

This paper cites SpokenWOZ: A Large-Scale Speech-Text Benchmark for Spoken Task-Oriented Dialogue Agents.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks SpokenWOZ: A Large-Scale Speech-Text Benchmark for Spoken Task-Oriented Dialogue Agents

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:53.858082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:53.858082Z digest=sha256:5a83b6b20e094e8cab1bd79750fa1eadf9c592d0dc83ea1715515c29044e7856

Observation 1b3c17e9-60ec-4a67-93ca-76db04cd35ad · outbound

This paper cites Training language models to follow instructions with human feedback.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Training language models to follow instructions with human feedback

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:53.956373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:53.956373Z digest=sha256:0703ca70da2f7a12061d88618f83d7c5889b1be06f4d5693f41661805e6988ab

Observation 556260fb-0aa7-484a-98ae-0209d064c99e · outbound

This paper cites HuggingGPT: Solving AI Tasks with ChatGPT and its Friends in Hugging Face.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks HuggingGPT: Solving AI Tasks with ChatGPT and its Friends in Hugging Face

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:54.050660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:54.050660Z digest=sha256:cf4723972c93c4382d421b55d2a6e947809eaffe848ff8774a3acdf9318b2500

Observation 5e8bd2b4-b095-4eaa-96f2-ef35ab2f289f · outbound

This paper cites API-Bank: A Comprehensive Benchmark for Tool-Augmented LLMs.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks API-Bank: A Comprehensive Benchmark for Tool-Augmented LLMs

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:54.130776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:54.130776Z digest=sha256:7345e4f64eaa215bb6cd6b6674cc77aab705865c7548276da29d86f8a41290d0

Observation 44b5b849-70de-40c6-b1f0-aaa0cafb8fa0 · outbound

This paper cites ToolDial: Multi-turn Dialogue Generation Method for Tool-Augmented Language Models.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks ToolDial: Multi-turn Dialogue Generation Method for Tool-Augmented Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:54.171642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:54.171642Z digest=sha256:cc50d5a1bbda3dc3296bbd7c6478d1db6991270cbc1508ffe481b24fa0446df1

Observation 6fb7d218-242a-4f0f-af21-6204b876b14c · outbound

This paper cites Rethinking task-oriented dialogue systems: From complex modularity to zero-shot autonomous agent,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Rethinking task-oriented dialogue systems: From complex modularity to zero-shot autonomous agent,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:55:57.237848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T21:55:54.215277Z digest=sha256:2972ab7f4f45c0d435ade923be3ca9ceec798a7e0a3c6a08cb5a8eef167940a3

Observation a9230911-5ef5-4918-bd37-7749709b1904 · outbound

This paper cites React: Synergizing reasoning and acting in language models,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks React: Synergizing reasoning and acting in language models,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:54.285113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:54.285113Z digest=sha256:b4b187b7a07d20eb7460aecd3558b8632545e7c70b3160840ebe7dcbe96cb4ea

Observation 899075e1-9ab3-44f9-bedd-970ffb07a5f2 · outbound

This paper cites Audio-cot: Exploring chain-of-thought reasoning in large audio language model,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Audio-cot: Exploring chain-of-thought reasoning in large audio language model,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:55:57.139353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T21:55:54.342505Z digest=sha256:8387830441306048773ab8a653e5e1d49b2d537dd4636657f03839e1824f68f1

Observation 75024118-a0c6-4ac4-b6bf-d45f86f1d41e · outbound

This paper cites Audio-reasoner: Improving reasoning capability in large audio language models,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Audio-reasoner: Improving reasoning capability in large audio language models,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:54.436315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:54.436315Z digest=sha256:bfec1f97fb238c8d9108bf35243242ac22f0cd693807f13e66c4b7482d198977

Observation 13a23776-4513-41db-ab68-e23694a70cd5 · outbound

This paper cites Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:54.386404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:54.386404Z digest=sha256:ab2503899a1127ad9b90e763d1b8cc9601568575052634bd02051210b33102d7

Observation 5f6776fd-53f7-48e7-a592-eb9359226dfc · outbound

This paper cites Can a suit of armor conduct electricity? a new dataset for open book question answering,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Can a suit of armor conduct electricity? a new dataset for open book question answering,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:54.539532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:54.539532Z digest=sha256:d922fc3653261c8db71883bd6f8495d9cfacac33af79222dd6363a4890b7158b

Observation e08f5fdd-a684-4b24-9a06-34a544179e27 · outbound

This paper cites ReSpAct: Harmonizing Reasoning, Speaking, and Acting Towards Building Large Language Model-Based Conversational AI Agents.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks ReSpAct: Harmonizing Reasoning, Speaking, and Acting Towards Building Large Language Model-Based Conversational AI Agents

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-06T21:55:55.474892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T21:55:54.473272Z digest=sha256:5e2d6c17c3bafb52e858685b0a518c4b6405412879ad82b51710649dabd9e4b7

Observation 11404fd2-5124-4b62-a298-96fbd8fde36c · outbound

This paper cites Gpt-4o: Openai’s new multimodal flagship model,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Gpt-4o: Openai’s new multimodal flagship model,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:55:57.037215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T21:55:54.625379Z digest=sha256:89d1f68b43f15a46d3c9f12b99d6677af5201f8a3e8add4fdb975609b62d47db

Observation 13a632cc-e96a-46d8-820f-e45d952320fc · outbound

This paper cites Robust Speech Recognition via Large-Scale Weak Supervision.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Robust Speech Recognition via Large-Scale Weak Supervision

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:54.575759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:54.575759Z digest=sha256:452eac42bf4f3e3c05b1d67f74848c9af1570a8fe2d9b58b52e2fba2975d881c

Observation 3705390a-447d-4e6f-b7d1-6b65e8ad02ce · outbound

This paper cites Espnet-TTS: Unified, reproducible, and integratable open source end-to-end text-to-speech toolkit,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Espnet-TTS: Unified, reproducible, and integratable open source end-to-end text-to-speech toolkit,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:55:56.795127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T21:55:54.714602Z digest=sha256:2ae9579c5f0a425e6297787d9dd1b8ebbdd8d1c4010151fe2b1f7fdd1cdbf6da

Observation 37471f67-73b0-461a-8515-e273a18636a3 · outbound

This paper cites Owsm v3.1: Better and faster open whisper-style speech models based on e-branchformer,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Owsm v3.1: Better and faster open whisper-style speech models based on e-branchformer,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:55:56.937871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T21:55:54.650824Z digest=sha256:3a4221adce5657e0889f3d4fa2819304294de7efd0ab8318983d0b4e7579c89e

Observation 66ba6451-65ec-41fd-9e2c-1ef0c70223ec · outbound

This paper cites Efficient Memory Management for Large Language Model Serving with PagedAttention.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Efficient Memory Management for Large Language Model Serving with PagedAttention

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:54.808526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:54.808526Z digest=sha256:ae383367e743103fcf92de29ca71a1f67ac9f411fa1418b499beb0f2879e1f38

Observation 4d4e4b70-79e5-4f2e-bca8-d5e1e66c4115 · outbound

This paper cites The Llama 3 Herd of Models.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks The Llama 3 Herd of Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:54.758002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:54.758002Z digest=sha256:ecc4b2b17d251b471014238e5a5bb19b00270d92680bd4358ab0378bb0fcf3b2

Observation c741c3e5-60be-4f0c-a95b-f7eb3d2c3196 · outbound

This paper cites Gpt-4o: Openai’s new multimodal flagship model,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Gpt-4o: Openai’s new multimodal flagship model,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:55:56.399442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T21:55:54.920074Z digest=sha256:3e6661178f9384f2efb70701695b98eb3abc0d5bf61781e0499853f9f6b394f3

Observation 64c3b58c-b427-4d79-946b-af719df12293 · outbound

This paper cites Alpacaeval: An automatic evaluator of instruction-following models,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Alpacaeval: An automatic evaluator of instruction-following models,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:55:56.544396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T21:55:54.850695Z digest=sha256:536e5cd8d00929cd20e789ea3ce00733b0ec77970ce97582786078560bc914e4

Observation 4c43f8dc-f9de-4f62-8a4b-20a10ca0f237 · outbound

This paper cites Moshi: a speech-text foundation model for real-time dialogue.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Moshi: a speech-text foundation model for real-time dialogue

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:55.026534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:55.026534Z digest=sha256:b73d8ea8cc3fd0eb88fc7cd1b9f33edc67c88f5d90e34120923dec3f446c1d38

Observation 32a86f2f-3495-438e-ad5a-6029d8f42bd6 · outbound

This paper cites Gpt-4o: Openai’s new multimodal flagship model,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Gpt-4o: Openai’s new multimodal flagship model,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:55:56.224799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T21:55:54.953118Z digest=sha256:e3b87fec6e708d1550046fd5eb47e0835455cae47611ab4989653d489de9ade7

Observation f6d58d67-3f07-4053-a581-dfd13108f015 · outbound

This paper cites Kimi-Audio Technical Report.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Kimi-Audio Technical Report

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:55.166303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:55.166303Z digest=sha256:55002057fa7b4b08a75bb4a4e5e4456bda6b618b96eeecae448823df294f7ff8

Observation 63d253aa-d644-45da-a7b8-a64c00d06d94 · outbound

This paper cites Mini-Omni2: Towards Open-source GPT-4o with Vision, Speech and Duplex Capabilities.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Mini-Omni2: Towards Open-source GPT-4o with Vision, Speech and Duplex Capabilities

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:55.082723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:55.082723Z digest=sha256:6f9cb30b231ec1f2757f07bc570b4fffaf2f6962dfbddad3bf3335109a0416a1

Observation 342d84ad-921d-4ceb-a81f-bd574f5321f8 · outbound

This paper cites Parakeet-tdt-0.6b-v2,.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Parakeet-tdt-0.6b-v2,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:55:56.017766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T21:55:55.302080Z digest=sha256:73ae7e88dee3752a94090aa71cfb7eb2c2becd59840a29b78745a0f6c6002512

Observation 7640f52d-8697-41b2-b649-200ececb7083 · outbound

This paper cites Qwen3 Technical Report.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Qwen3 Technical Report

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:55.221863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:55.221863Z digest=sha256:c5f9801f15f912c65efa11d0a9ac3212b5099af2f39c8ac6e8cd5c86879b0d0d

Observation 853e5859-4f46-4486-a17e-4cc553b47172 · outbound

This paper cites ESPnet-SDS: Unified Toolkit and Demo for Spoken Dialogue Systems.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks ESPnet-SDS: Unified Toolkit and Demo for Spoken Dialogue Systems

Reference 2025

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T21:55:55.822601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T21:55:53.100157Z digest=sha256:97360f55a42710c5a4b574e5d43b03666d7960b46602836c6d3fd7f42c1d8dec

Pith citing papers

Observation 9231afea-d3b6-4e49-9b70-35d3f6ae2021 · inbound

From Reactive to Proactive: Assessing the Proactivity of Voice Agents via ProVoice-Bench cites this paper.

From Reactive to Proactive: Assessing the Proactivity of Voice Agents via ProVoice-Bench AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.002818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T11:44:12.373082Z digest=sha256:137e79f19a668a2ce08a0215006cd4685f1da1c30b28b3d02ff2ef3d7676684b

Observation 878f8158-8b47-4503-8238-062795f7805e · inbound

A Survey of Audio Reasoning in Multimodal Foundation Models cites this paper.

A Survey of Audio Reasoning in Multimodal Foundation Models AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks

Reference 109

Resolution
verified exact
arxiv_id, observed 2026-05-21T02:09:24.313261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T02:08:06.976461Z digest=sha256:5d279ba21832abfaf866e53aca6fef02b51ae37d1c1e8000713c2afd836d1547