Pith. sign in

Paper Citation Record · LEDGER

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems

As of 21 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 4 inbound Pith citation observations for arXiv:2505.15000.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.15000 v1

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:29:11.225870Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T11:20:18.032456Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T02:09:24.101764Z

Reference resolution

53 of 53 outbound references displayed

  • verified exact2
  • verified fuzzy1
  • unresolved50
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 85c6914a-5e75-49ce-8570-921a040a7767 · outbound

This paper cites online" 'onlinestring :=.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:05.383707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:05.383707Z digest=sha256:59b868a7306bfff540a6cf5e15e3547d05fed50c3ce787731ba3e6d3b656b06a

Observation 75dc0901-1399-48a2-81e0-7de589607136 · outbound

This paper cites write newline.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:05.492840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:05.492840Z digest=sha256:7a83d52abd60997d35f0e8fd0260aeeb004fb1e233b4b1817f1b3ed95fb7c37f

Observation 266e389a-678d-4477-8ce9-a6f32b368a9a · outbound

This paper cites Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:05.603798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:05.603798Z digest=sha256:905c20bd6055865e48f329c1c17baafb2cc3b52b2673f7b0837618b47e7b03c7

Observation a291c044-2852-4cb3-b259-d15d1708d90b · outbound

This paper cites FunAudioLLM: Voice Understanding and Generation Foundation Models for Natural Interaction Between Humans and LLMs.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems FunAudioLLM: Voice Understanding and Generation Foundation Models for Natural Interaction Between Humans and LLMs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:05.746864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:05.746864Z digest=sha256:2b4b351847a956d4aff20b146f333928077babf589dc86126d634f2da08822a0

Observation e6c5dcdc-367a-46f6-a876-48e8e644b503 · outbound

This paper cites Qwen2-Audio Technical Report.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Qwen2-Audio Technical Report

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:05.891772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:05.891772Z digest=sha256:94c5c365fd368d6764ec0a9eeae59d37b9aa31dde4efeb53e230aef407c891b0

Observation c528cbb1-192e-42ed-9ea7-6f26663486f2 · outbound

This paper cites Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:05.982526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:05.982526Z digest=sha256:8fd6597429bced311fd31bf9aec0096b5b9765309e404ba5806f223ee0bda4db

Observation c9620cbb-0b11-4707-b5cd-91bcd36fbc23 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Training Verifiers to Solve Math Word Problems

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:06.046164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:06.046164Z digest=sha256:fb979d5d91c36519900d0aea7f8eb3994f8607173e332dc49665a378fcf14e9d

Observation 3664a2fe-b193-4ef5-afd8-b995ec8f5b92 · outbound

This paper cites VoxEval: Benchmarking the Knowledge Understanding Capabilities of End-to-End Spoken Language Models.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems VoxEval: Benchmarking the Knowledge Understanding Capabilities of End-to-End Spoken Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:06.156425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:06.156425Z digest=sha256:62e81e10847cb37f805cfdcae0d33b8562fae37f72ef8d595114bce3d5d3b211

Observation 776258d4-e59f-4fe8-a55e-cb813b7dfa3a · outbound

This paper cites Recent Advances in Speech Language Models: A Survey.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Recent Advances in Speech Language Models: A Survey

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:06.241559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:06.241559Z digest=sha256:6bb5ea61d09a3dc238a46bcb98a8ead6c9ebc79476692beab2b8e8b7ca1bb709

Observation be7bd3bd-d8fd-4713-8ca7-15a84fd921ec · outbound

This paper cites GAMA: A Large Audio-Language Model with Advanced Audio Understanding and Complex Reasoning Abilities.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems GAMA: A Large Audio-Language Model with Advanced Audio Understanding and Complex Reasoning Abilities

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:06.347770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:06.347770Z digest=sha256:f1051f0d9b29d389c6b7b03ebc093aad6e30381db90ea6a2245cfadc43482b71

Observation b6743a7c-d15e-47ab-850a-7e2686a526c5 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:14.642682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T15:29:06.447238Z digest=sha256:748840a064c082cd0487a677ff11cd84b90bc3c77e2c3a95d54e1541c82be090

Observation 410a78f5-e00f-4e21-aee0-e7cd25e3da7c · outbound

This paper cites The Llama 3 Herd of Models.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems The Llama 3 Herd of Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:06.506922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:06.506922Z digest=sha256:667be59b7526b2329810c2dde97a7c680953c5956941eb73dcba38a0a9cab1e3

Observation 042367db-bc70-4ca2-9105-b29867ac3b75 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:06.661963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:06.661963Z digest=sha256:ca65c6280ee346e7cce2331c75239c78da0f5c12ed3b29724d61e0bf37ad7cd8

Observation e211f9c4-279c-4b75-8982-982754be555f · outbound

This paper cites InfiMM-Eval: Complex Open-Ended Reasoning Evaluation For Multi-Modal Large Language Models.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems InfiMM-Eval: Complex Open-Ended Reasoning Evaluation For Multi-Modal Large Language Models

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:29:12.179165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T15:29:06.752573Z digest=sha256:abfe43d52a2b320437ecbbed9a97400666881f59541362ef0cf1b6c9a9991304

Observation ca4d68d7-8604-484e-b562-19be0e03b0cf · outbound

This paper cites Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:06.836442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:06.836442Z digest=sha256:94626597c9b456c17ea683d56ac15d3488448b18ed92793a52d4f102f19fddcd

Observation 0aa76744-67c5-4a22-ae0c-6d7edda09bd1 · outbound

This paper cites MERaLiON-AudioLLM: Bridging Audio and Language with Large Language Models.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems MERaLiON-AudioLLM: Bridging Audio and Language with Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:06.918057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:06.918057Z digest=sha256:9cd7fc2fa8a0b2c70f5f4f8acf7fc7dfe585a0f085b138676c691179928654e6

Observation afbdd6d7-55f5-4469-a78a-8f7fd1c431aa · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:06.990799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:06.990799Z digest=sha256:0ddcc3f173eb9b3b1e62c456b8e4c0bb88b8dfb42e2369dfaafe98a86aa7bbbc

Observation e6641d6b-4106-4342-8a7a-3cfd18529825 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:14.372863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T15:29:07.060061Z digest=sha256:5f0d057e9855189539130b1d89bfe8a1363aa77fd861d811cd23bbc596908f48

Observation c7bc4686-5ada-4b63-9ff4-27152db6d3f0 · outbound

This paper cites Step-Audio: Unified Understanding and Generation in Intelligent Speech Interaction.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Step-Audio: Unified Understanding and Generation in Intelligent Speech Interaction

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:07.130989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:07.130989Z digest=sha256:043e26953b631287db72c4198d9544f27b10f85fefd3cc32633c978b61436ea4

Observation 9cf64359-b0d1-4418-8645-ae6f4fb80d1d · outbound

This paper cites GPT-4o System Card.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems GPT-4o System Card

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:07.276254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:07.276254Z digest=sha256:d7a6d39216c6474ca01b4e7ecd1385a93d6be32ee2074f73d27d716f38e531b5

Observation 9e65ca95-cf97-4614-a103-166d69d5f444 · outbound

This paper cites Identifying and mitigating vulnerabilities in llm-integrated applications.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Identifying and mitigating vulnerabilities in llm-integrated applications

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:14.146759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T15:29:07.386967Z digest=sha256:13117553670a1be396f8a65aeafef84f34f2bd758afef790d47bc4093e90c0dc

Observation c86d34b2-23f7-473a-b8d1-bc4262b5ef0d · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:07.575107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:07.575107Z digest=sha256:267640f5c8b445ce1a059482850d8ef7e8178b7b378ff5f45d7b174cd4acc63d

Observation f5fd6abf-ca37-4e7b-a9ef-1b505acd5909 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:07.705019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:07.705019Z digest=sha256:9ad3dbc17b5169dd366f856260ced2d08c79c6af2075e397d6e133f6c57b1619

Observation 8a518f1d-38d9-4f42-a14c-d4df6cadc9e7 · outbound

This paper cites MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:07.815220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:07.815220Z digest=sha256:1cb7f2be3230efd9dbc3ac54ba27dd44b9efbb0623d27c9569f377efac3d9133

Observation 45ab8b93-6c95-4f35-900e-c6b760657cf9 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:13.850609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T15:29:07.970827Z digest=sha256:693a61317aa3fdd8eb934ec250d3a23376ac2793fbe813f7d19960fc1c07f85e

Observation f57a2608-05d9-447f-a12f-d5f2eae2d3d4 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:13.607467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T15:29:08.140174Z digest=sha256:03f5ff44c2cb94a91628a6499d62a0d90a4c195352aa0d637ef0d6ea3f9b07f1

Observation 5b4d033c-4740-4584-9ebc-6f69d021e827 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:13.299629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T15:29:08.269687Z digest=sha256:157a1d2855feff8745a748a26822fb5e60fc141d17d28b0811327a4d798a71ee

Observation edf34150-7c3f-4609-aa30-344adb293b33 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:08.439478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:08.439478Z digest=sha256:482814efb74983cefcb49a176d9fbc375ef923e2695476789ac323b2ebfedb68

Observation 7b93e13b-f795-48cc-87b6-d2ceb4957420 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:08.670367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:08.670367Z digest=sha256:26b2062c4fb8d298552657d9b4fc65a0e6a8be29a3e1d6aff1873aac050443b8

Observation 3eff4b52-9fde-4a5d-8613-7e369c6fd26b · outbound

This paper cites We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:08.761996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:08.761996Z digest=sha256:de7ade02a1a733e02f5cfd1dfe2e2181f67c8788427b9e7d30cf2e6d91300ea5

Observation 7096508b-4efd-4f84-924d-279cddb03ad3 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:08.811385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:08.811385Z digest=sha256:c922ef784ae94f32108099e87b1567f593e2cac5f4b502a5e6351caa3bd41435

Observation e4fc5f9c-95cc-4e26-884d-586c8bd12475 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:13.025456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T15:29:08.866273Z digest=sha256:38cb44b737af503946bce311b1811208dd0ddfbeb542e42eed05eb5d1b418fad

Observation 11b05881-b78f-4b2b-9a64-cbcade8cdf47 · outbound

This paper cites MMAU: A Massive Multi-Task Audio Understanding and Reasoning Benchmark.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems MMAU: A Massive Multi-Task Audio Understanding and Reasoning Benchmark

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:08.907699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:08.907699Z digest=sha256:4615be5b275acf291a72718a414ff7c6132c6858b4a8011b8be8e99c71f85df5

Observation d217b324-4312-4423-b330-978c7bc4eca3 · outbound

This paper cites ARB: Advanced Reasoning Benchmark for Large Language Models.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems ARB: Advanced Reasoning Benchmark for Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:08.957582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:08.957582Z digest=sha256:47cd5238c75763ba683ba20dbda4d7cdc86609725dd39d4355eeedd3d719d77f

Observation e4b22b5a-e1c1-4fbd-823c-6265b5be0127 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:09.009917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:09.009917Z digest=sha256:2d82a296ce7b615a002c9b614649a186d5ad9af153acd9c0a50fa56cb58fa629

Observation 62377a18-2e9d-4fc0-9855-1509e8db42d4 · outbound

This paper cites Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:09.060214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:09.060214Z digest=sha256:a3d166c6212f4ee5a2ed60e16ca62348ffbbbd87163cde9bc52fa82284eb413b

Observation ae38544a-e1a5-4962-a735-e79322499815 · outbound

This paper cites SALMONN: Towards Generic Hearing Abilities for Large Language Models.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems SALMONN: Towards Generic Hearing Abilities for Large Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:09.112978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:09.112978Z digest=sha256:f4dd3d87060c6f361ea9d3e7bb3431a0a59bcd86cd50692414cf11e0e02eaa78

Observation 334d6059-03cd-458a-8931-d1168dcc0057 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:09.165431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:09.165431Z digest=sha256:4a517e8d02e288d8b048b5d0ca3547ad076f4550d243366d5053bd07d71420fa

Observation bf90be1b-a9c8-48d5-ba24-832a9fcdf08e · outbound

This paper cites Gemma 2: Improving Open Language Models at a Practical Size.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Gemma 2: Improving Open Language Models at a Practical Size

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:09.231365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:09.231365Z digest=sha256:a22b3e9f7e53c978f5d6e86537c0bb1e8128e8e113718f27962ca569c49e67a6

Observation 957facb3-f800-4cfd-a49b-e014d32d746e · outbound

This paper cites LlamaV-o1: Rethinking Step-by-step Visual Reasoning in LLMs.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems LlamaV-o1: Rethinking Step-by-step Visual Reasoning in LLMs

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:09.303581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:09.303581Z digest=sha256:dafa11cd5231777a6352b9443ea57ef4f0bf42f87e6a115d6b14e26195db7b84

Observation 1dfd7160-01e1-4910-accc-69f467d5916d · outbound

This paper cites OpenMathInstruct-1: A 1.8 Million Math Instruction Tuning Dataset.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems OpenMathInstruct-1: A 1.8 Million Math Instruction Tuning Dataset

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:09.451527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:09.451527Z digest=sha256:aa284b7229fa357f02c9bac0798903a77b7365ae6316e08c7d6fe51ffa4f493c

Observation 6cb4693f-c41a-4568-9d0d-7b5c6e890d98 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:09.630568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:09.630568Z digest=sha256:92d903882f0bca4b24d1999bd4508cc67a819458b19468719a4b5f363934b2db

Observation 033c6da8-70f6-4ba8-87ce-e643a5b99b3a · outbound

This paper cites CoinMath: Harnessing the Power of Coding Instruction for Math LLMs.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems CoinMath: Harnessing the Power of Coding Instruction for Math LLMs

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:29:11.653822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T15:29:09.802739Z digest=sha256:a35b0870264090cdbaafc5bc499eb0a53a19c2399031f04ab2d98b06aee24c62

Observation bf8e362f-a722-42c1-b9f2-df5b4b3d1c19 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:09.918646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:09.918646Z digest=sha256:eb32492afbcfed7f70ca1f04f3d9bc8854e388cb07efe090778ee2613350940b

Observation 9f7fcb83-621b-48f1-8aa4-ea55d87c5f76 · outbound

This paper cites LLaVA-CoT: Let Vision Language Models Reason Step-by-Step.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems LLaVA-CoT: Let Vision Language Models Reason Step-by-Step

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:10.069369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:10.069369Z digest=sha256:3ccca0d38f31846bf9e6d7765520b4a0e02010cb687bfba26150e3aa74477a7d

Observation 2b724981-3f44-4de4-b84d-c906395ff3f3 · outbound

This paper cites VisuLogic: A Benchmark for Evaluating Visual Reasoning in Multi-modal Large Language Models.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems VisuLogic: A Benchmark for Evaluating Visual Reasoning in Multi-modal Large Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:10.165713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:10.165713Z digest=sha256:a2b6feffe9e8dae794f26cc458c280b74f3f659c39fff5a0c7ebc37702fe3460

Observation 05a65c43-edc5-4eb8-9042-f3dfef4010e6 · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:10.500394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:10.500394Z digest=sha256:2e0dfcda9523b0a21232dcd7be1498bcbfb8aca77177e8ebd82f0b0dd1733306

Observation 2455d858-69a6-4ed7-b71c-3f6252576548 · outbound

This paper cites AIR-Bench: Benchmarking Large Audio-Language Models via Generative Comprehension.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems AIR-Bench: Benchmarking Large Audio-Language Models via Generative Comprehension

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:10.617952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:10.617952Z digest=sha256:c18008c1dbb0daf7cffeeda268294bae7ad3203d84e59dd4b26fd969b7a98eb2

Observation 3484289d-0fc8-4e26-9ec8-b14aca858f4e · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:10.696136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:10.696136Z digest=sha256:dbb16650a6381e8e19d3c22901142510d6d2fc0de4069bc342ca124eec01dba8

Observation 98cbffa3-3cc2-4f6b-8840-fc938a69a4b5 · outbound

This paper cites MR-Ben: A Meta-Reasoning Benchmark for Evaluating System-2 Thinking in LLMs.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems MR-Ben: A Meta-Reasoning Benchmark for Evaluating System-2 Thinking in LLMs

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:10.816937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:10.816937Z digest=sha256:d7cdad9b6b354e9844bcc3dc086a6439b082ec58871d98ef46529b6408fee174

Observation cca22b89-d645-42ee-b32b-5a6c2779d378 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:12.749419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T15:29:10.988219Z digest=sha256:89a1f5d22791cd2b1ca7b2cb73bd1e998842ea8dec47aac6bc4a5de5c81e0b5a

Observation ba3d7663-9a9a-4344-af18-9833fe446361 · outbound

This paper cites an unresolved cited work.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:12.493642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T15:29:11.108727Z digest=sha256:0c3d11348d6bb34a9f8990a95fb591d8a6f918f68f50ddd681ae4b268f311e93

Observation ea63c9ff-5161-465b-841f-43870f4dff29 · outbound

This paper cites What Makes Large Language Models Reason in (Multi-Turn) Code Generation?.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems What Makes Large Language Models Reason in (Multi-Turn) Code Generation?

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:11.225870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:11.225870Z digest=sha256:5ccdf3bd9ef01798665f04836079c434775e22c7c8c025d3fa564df1a335b4ee

Pith citing papers

Observation 5a02317f-104d-4e39-a2a5-766b072b0843 · inbound

Mind-Paced Speaking: A Dual-Brain Approach to Real-Time Reasoning in Spoken Language Models cites this paper.

Mind-Paced Speaking: A Dual-Brain Approach to Real-Time Reasoning in Spoken Language Models Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-18T07:46:03.523212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T07:43:23.913399Z digest=sha256:859594738ebc0be0d3c02afccb7be95548d0230176a9bce0fa2b51f526a2edb6

Observation 1597cf4d-188a-4f55-802e-283e695a2adc · inbound

Step-Audio-R1.5 Technical Report cites this paper.

Step-Audio-R1.5 Technical Report Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-12T00:46:13.579952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-07T14:12:56.278256Z digest=sha256:50eb415a1f05c641236168d2742c95c1d4abc2b1efad8c080611c6871f22172d

Observation e831c21e-5c89-43e9-a2b8-e51abbb0ae8b · inbound

A Survey of Audio Reasoning in Multimodal Foundation Models cites this paper.

A Survey of Audio Reasoning in Multimodal Foundation Models Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems

Reference 127

Resolution
verified exact
arxiv_id, observed 2026-05-21T02:09:24.104663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-21T02:08:06.976461Z digest=sha256:dfb07132e8799362b22e05f86322cfaa4407e7731a12e98228ef1830a440a42f

Observation 57d2b1d6-7214-482b-80b6-266c8d138531 · inbound

Efficient Chain-of-Modality Reasoning via Progressive Compression for Spoken Language Models cites this paper.

Efficient Chain-of-Modality Reasoning via Progressive Compression for Spoken Language Models Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-01T11:20:18.032456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:20:18.032456Z digest=sha256:47811d533c0886a92854a017990c9c107c351aac40772ca009b9960c04a9db15