Pith. sign in

Paper Citation Record · LEDGER

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems

As of 8 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 4 inbound Pith citation observations for arXiv:2505.15000.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.15000 v1

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:29:11.225870Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T11:20:18.032456Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T02:09:24.101764Z

Reference resolution

53 of 53 outbound references displayed

  • verified exact2
  • verified fuzzy1
  • unresolved50
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 85c6914a-5e75-49ce-8570-921a040a7767 · outbound

This paper cites online" 'onlinestring :=.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:05.383707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:05.383707Z digest=sha256:1aac7857eba487be71391e7715f1f18819ddb91387da6a9320835fe76c607157

Observation 75dc0901-1399-48a2-81e0-7de589607136 · outbound

This paper cites write newline.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:05.492840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:05.492840Z digest=sha256:5fe4c93f8393ac1f9c05c64793b1a440fc82dcee10549919933be897d9daf97f

Observation 266e389a-678d-4477-8ce9-a6f32b368a9a · outbound

This paper cites Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:05.603798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:05.603798Z digest=sha256:ffb8fa711942fa9c0598ee1e9d31079ede45aaa85e3fa2dde593e8b59ba13190

Observation a291c044-2852-4cb3-b259-d15d1708d90b · outbound

This paper cites FunAudioLLM: Voice Understanding and Generation Foundation Models for Natural Interaction Between Humans and LLMs.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems FunAudioLLM: Voice Understanding and Generation Foundation Models for Natural Interaction Between Humans and LLMs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:05.746864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:05.746864Z digest=sha256:7e84f8619f326ba536ea6ebcc1c94f8f1a79af45a4925b14641044f0cef5e722

Observation e6c5dcdc-367a-46f6-a876-48e8e644b503 · outbound

This paper cites Qwen2-Audio Technical Report.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Qwen2-Audio Technical Report

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:05.891772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:05.891772Z digest=sha256:e2dc7305a0db62aad0095f73f4192362ae51c640f6655ae56ca7611e767f7185

Observation c528cbb1-192e-42ed-9ea7-6f26663486f2 · outbound

This paper cites Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:05.982526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:05.982526Z digest=sha256:a8e8111c2a0fa99fbe820293982958f336cf301e14e5f9fe0e5218bba72c06aa

Observation c9620cbb-0b11-4707-b5cd-91bcd36fbc23 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Training Verifiers to Solve Math Word Problems

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:06.046164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:06.046164Z digest=sha256:6964120bdb4457b49b5f9583cc229321694aacc27617cf623bef48c9982c2f86

Observation 3664a2fe-b193-4ef5-afd8-b995ec8f5b92 · outbound

This paper cites VoxEval: Benchmarking the Knowledge Understanding Capabilities of End-to-End Spoken Language Models.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems VoxEval: Benchmarking the Knowledge Understanding Capabilities of End-to-End Spoken Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:06.156425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:06.156425Z digest=sha256:166fce852fed3d9c2566ccb4e5a96746162dab6023d46654beb41bf6225a15c1

Observation 776258d4-e59f-4fe8-a55e-cb813b7dfa3a · outbound

This paper cites Recent Advances in Speech Language Models: A Survey.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Recent Advances in Speech Language Models: A Survey

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:06.241559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:06.241559Z digest=sha256:9f50b28c96bb9c45c5585dcb56fe46494250979af3e3af3865af8a8b6a2829e6

Observation be7bd3bd-d8fd-4713-8ca7-15a84fd921ec · outbound

This paper cites GAMA: A Large Audio-Language Model with Advanced Audio Understanding and Complex Reasoning Abilities.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems GAMA: A Large Audio-Language Model with Advanced Audio Understanding and Complex Reasoning Abilities

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:06.347770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:06.347770Z digest=sha256:bdf3466874fac048050b654818960c8ec0a537437d234c72f1831568e9494e47

Observation b6743a7c-d15e-47ab-850a-7e2686a526c5 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:14.642682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T15:29:06.447238Z digest=sha256:e73e5e7bb034a9287053cd070460ba53e11829ac0cd83aed34f9a443a628777c

Observation 410a78f5-e00f-4e21-aee0-e7cd25e3da7c · outbound

This paper cites The Llama 3 Herd of Models.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems The Llama 3 Herd of Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:06.506922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:06.506922Z digest=sha256:3890db3bd9eaaf6db591f39709adba36072b650d0ad3dec2956f5bbb0602ea27

Observation 042367db-bc70-4ca2-9105-b29867ac3b75 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:06.661963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:06.661963Z digest=sha256:d6277e682d0dfd3ba592a23c00eb624c075240035d43686da5ec6d348c69fa43

Observation e211f9c4-279c-4b75-8982-982754be555f · outbound

This paper cites InfiMM-Eval: Complex Open-Ended Reasoning Evaluation For Multi-Modal Large Language Models.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems InfiMM-Eval: Complex Open-Ended Reasoning Evaluation For Multi-Modal Large Language Models

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:29:12.179165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T15:29:06.752573Z digest=sha256:b4515dae61d56bd24a3bccd7d853c6b03b353f3691a08cd64982b6b54a2f1ae3

Observation ca4d68d7-8604-484e-b562-19be0e03b0cf · outbound

This paper cites Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:06.836442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:06.836442Z digest=sha256:1e626f02991102728ce2c884870fcc77bc5157bb4bb3719df6956157c9e09f94

Observation 0aa76744-67c5-4a22-ae0c-6d7edda09bd1 · outbound

This paper cites MERaLiON-AudioLLM: Bridging Audio and Language with Large Language Models.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems MERaLiON-AudioLLM: Bridging Audio and Language with Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:06.918057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:06.918057Z digest=sha256:79e961995e26a674bece3a3e2ffba3bcc085e33640c3915591ac37dd431794ff

Observation afbdd6d7-55f5-4469-a78a-8f7fd1c431aa · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:06.990799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:06.990799Z digest=sha256:a6f45029016a52ac8db0c883ad7d7dc6827ef44335496915f7d9a0fd931edf42

Observation e6641d6b-4106-4342-8a7a-3cfd18529825 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:14.372863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T15:29:07.060061Z digest=sha256:d23acfd83bf7386f9ec4bb82ffef5a913e7924b53249b3ccfcaa88e50a104077

Observation c7bc4686-5ada-4b63-9ff4-27152db6d3f0 · outbound

This paper cites Step-Audio: Unified Understanding and Generation in Intelligent Speech Interaction.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Step-Audio: Unified Understanding and Generation in Intelligent Speech Interaction

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:07.130989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:07.130989Z digest=sha256:d1128e8aa82d8ebd2a67462d0474058fe587999f6d28907fa501e721ffefc076

Observation 9cf64359-b0d1-4418-8645-ae6f4fb80d1d · outbound

This paper cites GPT-4o System Card.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems GPT-4o System Card

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:07.276254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:07.276254Z digest=sha256:d0477d968918fb29cea8344b2796dab6191d3c74c3be8b6a2d19b9a9bab682ca

Observation 9e65ca95-cf97-4614-a103-166d69d5f444 · outbound

This paper cites Identifying and mitigating vulnerabilities in llm-integrated applications.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Identifying and mitigating vulnerabilities in llm-integrated applications

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:14.146759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T15:29:07.386967Z digest=sha256:aa03ac3d696ef9694b54a3196f0f64e9e9152ef781b795cf899749ffa1222826

Observation c86d34b2-23f7-473a-b8d1-bc4262b5ef0d · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:07.575107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:07.575107Z digest=sha256:47b5b2d2374087e530be88fec2b5281b55ecd9d4bf4c2b1abde81390cfad57a4

Observation f5fd6abf-ca37-4e7b-a9ef-1b505acd5909 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:07.705019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:07.705019Z digest=sha256:b174af09cfb2c6aa3c0581cad5cd5e33b3126393ebb7f8dcc1760ca6965a6b34

Observation 8a518f1d-38d9-4f42-a14c-d4df6cadc9e7 · outbound

This paper cites MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:07.815220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:07.815220Z digest=sha256:970ae736b19bfba96d2ab3840187fc8c95ee1ea88d4766778b7f3f0f0179510e

Observation 45ab8b93-6c95-4f35-900e-c6b760657cf9 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:13.850609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T15:29:07.970827Z digest=sha256:d4f84d7143521b3c62b74994fef5af113520e5cfb57b769b44bcab6c37bc248a

Observation f57a2608-05d9-447f-a12f-d5f2eae2d3d4 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:13.607467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T15:29:08.140174Z digest=sha256:57038749abe1b9a8f71f090ea5868348803d65dff27d25d9ec90d7fea295ba88

Observation 5b4d033c-4740-4584-9ebc-6f69d021e827 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:13.299629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T15:29:08.269687Z digest=sha256:93f4642d12add7530cdb0ecedcc214d8001b976ed4eab04b0765bb1457aaf4de

Observation edf34150-7c3f-4609-aa30-344adb293b33 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:08.439478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:08.439478Z digest=sha256:6246888f006486699c13b7c313009a59d1568f49078867ae0bd8168aabf526e4

Observation 7b93e13b-f795-48cc-87b6-d2ceb4957420 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:08.670367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:08.670367Z digest=sha256:c6492ef64cdb598c132e828d64cae771ff4233d1bad30d180cf05e5e881d3797

Observation 3eff4b52-9fde-4a5d-8613-7e369c6fd26b · outbound

This paper cites We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:08.761996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:08.761996Z digest=sha256:6d15648f6768c3b0b41d2f08058d00b726da1f7d89e535ada23eb92ce33d7039

Observation 7096508b-4efd-4f84-924d-279cddb03ad3 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:08.811385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:08.811385Z digest=sha256:a25bb72f5b9e26f16fb346dc7c47adcc40e84124d6fae952cf98e0fbe0943663

Observation e4fc5f9c-95cc-4e26-884d-586c8bd12475 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:13.025456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T15:29:08.866273Z digest=sha256:1f173ed7e5ded5304ab3477ae3980a43331ea6effe7f3eaddcfc0b90b105bd7e

Observation 11b05881-b78f-4b2b-9a64-cbcade8cdf47 · outbound

This paper cites MMAU: A Massive Multi-Task Audio Understanding and Reasoning Benchmark.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems MMAU: A Massive Multi-Task Audio Understanding and Reasoning Benchmark

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:08.907699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:08.907699Z digest=sha256:bc3c809bf0196f10dbd4c2b0aea3c7421c343d08c727607e39ac291d97de5cbe

Observation d217b324-4312-4423-b330-978c7bc4eca3 · outbound

This paper cites ARB: Advanced Reasoning Benchmark for Large Language Models.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems ARB: Advanced Reasoning Benchmark for Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:08.957582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:08.957582Z digest=sha256:281fad1d2abe66e097c91cfcebc693a23df7d1eb2d604ad852f0a5b07e3e73ed

Observation e4b22b5a-e1c1-4fbd-823c-6265b5be0127 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:09.009917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:09.009917Z digest=sha256:93c14964d0a33c17ed26b27d5270e7e8cb0d8b177ba1faa7520b04745b1ecaa4

Observation 62377a18-2e9d-4fc0-9855-1509e8db42d4 · outbound

This paper cites Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:09.060214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:09.060214Z digest=sha256:838d2a8ae4a587d22623acec28345d71339febf79df780452bca7821fb135301

Observation ae38544a-e1a5-4962-a735-e79322499815 · outbound

This paper cites SALMONN: Towards Generic Hearing Abilities for Large Language Models.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems SALMONN: Towards Generic Hearing Abilities for Large Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:09.112978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:09.112978Z digest=sha256:b4cd8097d42757bbf9446ba15533346b22843f87753197eda0378767465e0fae

Observation 334d6059-03cd-458a-8931-d1168dcc0057 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:09.165431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:09.165431Z digest=sha256:df6595c136fb086859e3944a2fafe3fa41ff9cf25b7657795970fe729a570eeb

Observation bf90be1b-a9c8-48d5-ba24-832a9fcdf08e · outbound

This paper cites Gemma 2: Improving Open Language Models at a Practical Size.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Gemma 2: Improving Open Language Models at a Practical Size

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:09.231365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:09.231365Z digest=sha256:abfbf0b886aba01d475051eecb8b1d7870588d267ca97d5febb9a485420875aa

Observation 957facb3-f800-4cfd-a49b-e014d32d746e · outbound

This paper cites LlamaV-o1: Rethinking Step-by-step Visual Reasoning in LLMs.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems LlamaV-o1: Rethinking Step-by-step Visual Reasoning in LLMs

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:09.303581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:09.303581Z digest=sha256:9fe71fdec16166266370bae4fd0d9b89aae756cb552ae8eb4e4dfb7a683cce3f

Observation 1dfd7160-01e1-4910-accc-69f467d5916d · outbound

This paper cites OpenMathInstruct-1: A 1.8 Million Math Instruction Tuning Dataset.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems OpenMathInstruct-1: A 1.8 Million Math Instruction Tuning Dataset

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:09.451527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:09.451527Z digest=sha256:7a1928ba632a1274839cf8f49f7f7c0d1144a01e290a071f772b625375007844

Observation 6cb4693f-c41a-4568-9d0d-7b5c6e890d98 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:09.630568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:09.630568Z digest=sha256:eb68f4d3711d658cf8c9fb96a2bf9e60c984e5be72eb5a6898cc2086741e5974

Observation 033c6da8-70f6-4ba8-87ce-e643a5b99b3a · outbound

This paper cites CoinMath: Harnessing the Power of Coding Instruction for Math LLMs.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems CoinMath: Harnessing the Power of Coding Instruction for Math LLMs

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:29:11.653822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T15:29:09.802739Z digest=sha256:68a684f9840dbfc35ae550df4046912541bc3ae0290746181f64ffb7558bf0d0

Observation bf8e362f-a722-42c1-b9f2-df5b4b3d1c19 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:09.918646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:09.918646Z digest=sha256:c4f98466a74c894f6cb6fd84fb4542d8ce4a59c0cd79db19a12ea44002ec43e9

Observation 9f7fcb83-621b-48f1-8aa4-ea55d87c5f76 · outbound

This paper cites LLaVA-CoT: Let Vision Language Models Reason Step-by-Step.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems LLaVA-CoT: Let Vision Language Models Reason Step-by-Step

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:10.069369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:10.069369Z digest=sha256:a1c0597fb3b2960cbe4d49f6c0ff0c8ddb42cc237461bd7704ddf28229e5084f

Observation 2b724981-3f44-4de4-b84d-c906395ff3f3 · outbound

This paper cites VisuLogic: A Benchmark for Evaluating Visual Reasoning in Multi-modal Large Language Models.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems VisuLogic: A Benchmark for Evaluating Visual Reasoning in Multi-modal Large Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:10.165713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:10.165713Z digest=sha256:73e3b8b99d6f6470619ae29fa3e3a40e49966388b358112be95374b80b3e3574

Observation 05a65c43-edc5-4eb8-9042-f3dfef4010e6 · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:10.500394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:10.500394Z digest=sha256:94d179dab3a5237819b618445d9eef6a340535f753e29d25b1e56c745b1cedeb

Observation 2455d858-69a6-4ed7-b71c-3f6252576548 · outbound

This paper cites AIR-Bench: Benchmarking Large Audio-Language Models via Generative Comprehension.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems AIR-Bench: Benchmarking Large Audio-Language Models via Generative Comprehension

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:10.617952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:10.617952Z digest=sha256:69cee8c6b3c14b18a0d0a89abb26cf9ad137279f04f1b07b3c22e545b145403f

Observation 3484289d-0fc8-4e26-9ec8-b14aca858f4e · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:10.696136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:10.696136Z digest=sha256:107df4f466044d6d6274b207a29f1ceb9a8bc31b3f634141ee564d749ef7042a

Observation 98cbffa3-3cc2-4f6b-8840-fc938a69a4b5 · outbound

This paper cites MR-Ben: A Meta-Reasoning Benchmark for Evaluating System-2 Thinking in LLMs.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems MR-Ben: A Meta-Reasoning Benchmark for Evaluating System-2 Thinking in LLMs

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:10.816937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:10.816937Z digest=sha256:1a29871efdd908c7d79f3a60a95c2152ee8ab348600d2c52211dc75547af8f30

Observation cca22b89-d645-42ee-b32b-5a6c2779d378 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:12.749419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T15:29:10.988219Z digest=sha256:90696001a8f60db8e7060fea164af9abd0fe26621b41e531f0e33fb2b8228478

Observation ba3d7663-9a9a-4344-af18-9833fe446361 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:12.493642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T15:29:11.108727Z digest=sha256:143114f0728a1fb87c33566fd56326d5ee09a812089825bb5856d77639ddc7a2

Observation ea63c9ff-5161-465b-841f-43870f4dff29 · outbound

This paper cites What Makes Large Language Models Reason in (Multi-Turn) Code Generation?.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems What Makes Large Language Models Reason in (Multi-Turn) Code Generation?

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:11.225870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:11.225870Z digest=sha256:ed9e19ab32dcc4307bfd285e2c695fdbac0efebdd7b520efe402a851ebe164f1

Pith citing papers

Observation 5a02317f-104d-4e39-a2a5-766b072b0843 · inbound

Mind-Paced Speaking: A Dual-Brain Approach to Real-Time Reasoning in Spoken Language Models cites this paper.

Mind-Paced Speaking: A Dual-Brain Approach to Real-Time Reasoning in Spoken Language Models Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-18T07:46:03.523212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T07:43:23.913399Z digest=sha256:b7bd979eff8aed989eabc564af9b56d63da500f82ad2a982647c3a0fbe12f076

Observation 1597cf4d-188a-4f55-802e-283e695a2adc · inbound

Step-Audio-R1.5 Technical Report cites this paper.

Step-Audio-R1.5 Technical Report Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-12T00:46:13.579952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T14:12:56.278256Z digest=sha256:422b437c0122f0ec541a231fba4adf9b8682d5dea2b79f40c8c7f4473c00cf75

Observation e831c21e-5c89-43e9-a2b8-e51abbb0ae8b · inbound

A Survey of Audio Reasoning in Multimodal Foundation Models cites this paper.

A Survey of Audio Reasoning in Multimodal Foundation Models Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems

Reference 127

Resolution
verified exact
arxiv_id, observed 2026-05-21T02:09:24.104663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T02:08:06.976461Z digest=sha256:801e03f3636ce89337fe471d97022a5fd1716bf62e3ae1a8144401ce47c1ba38

Observation 57d2b1d6-7214-482b-80b6-266c8d138531 · inbound

Efficient Chain-of-Modality Reasoning via Progressive Compression for Spoken Language Models cites this paper.

Efficient Chain-of-Modality Reasoning via Progressive Compression for Spoken Language Models Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-01T11:20:18.032456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:20:18.032456Z digest=sha256:3295710a62ff655c6385b14056cc6d36cb32aaae6959613bfbd1ac0e0e788094