Pith. sign in

Paper Citation Record · LEDGER

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation

As of 8 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 0 inbound Pith citation observations for arXiv:2607.13477.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.13477 v1

Coverage vector

measured 34 of 34 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T05:08:08.044744Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

34 of 34 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved34
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 66b1f27c-9864-45e7-85c5-0934150ceb2e · outbound

This paper cites Qwen2-Audio Technical Report.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation Qwen2-Audio Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:03.178486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:03.178486Z digest=sha256:f033b00e312c897945e4e0211e15dcab548c39cdbb6e6dd1ea598be6aead6dac

Observation 73b93a94-6749-491b-bc4d-abcf614fe0c5 · outbound

This paper cites Qwen2.5-Omni Technical Report.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation Qwen2.5-Omni Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:03.283757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:03.283757Z digest=sha256:60b7e4ab5871f7f7cfd9e326f062a94527a601b8593b30b3724203d96e6535e1

Observation b4faa0e3-9973-462e-b624-5587a85b5e29 · outbound

This paper cites Qwen3-Omni Technical Report.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation Qwen3-Omni Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:03.339448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:03.339448Z digest=sha256:b0803a6e0a05886623d68d36f551800fc6d33c09da22b6e7609a9fe79d6fdde0

Observation 57fb78aa-6083-4f9f-a8f1-8e0407d37986 · outbound

This paper cites Audio flamingo 3: Advancing audio intelligence with fully open large audio language models,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation Audio flamingo 3: Advancing audio intelligence with fully open large audio language models,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:03.541744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:03.541744Z digest=sha256:8dfaaa228a369224bec0322f30129a71640d709f0bcbe9fa52af96fb285f8538

Observation eee421ae-6f81-4633-954c-6d145337df2e · outbound

This paper cites Voxtral.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation Voxtral

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:03.719567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:03.719567Z digest=sha256:899294ae37c6b7f8e6c4a0d0ba096b7514057ab0c143221756d7e8a8a3fd4a14

Observation cabaf573-e2dc-41ab-8756-f24375fb29bb · outbound

This paper cites Judging LLM-as-a-judge with MT-Bench and chatbot arena,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation Judging LLM-as-a-judge with MT-Bench and chatbot arena,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:03.830382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:03.830382Z digest=sha256:fcf3efc2b991d86d332ead4b4056d02b01f5f3fb138a5a7c0c31cb99583ecf5b

Observation a9d79a93-61f5-4eb5-a683-c196d5e33bf2 · outbound

This paper cites CLAIR: Evaluating image captions with large language models,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation CLAIR: Evaluating image captions with large language models,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:03.951106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:03.951106Z digest=sha256:d7c89e8ec7ac5577326b70d8560e2effeed390c722082876cd335dffb9516730

Observation 99821d0c-df41-4936-9659-0ecf328c43e4 · outbound

This paper cites CLAIR-A: Lever- aging large language models to judge audio captions,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation CLAIR-A: Lever- aging large language models to judge audio captions,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:04.141943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:04.141943Z digest=sha256:43955029980b36318990e01dbdb3386a7b26e64c0d0a5f0bcfaf5914b05bd49b

Observation db7e028b-a52e-4c8f-8d17-676d3054059c · outbound

This paper cites Audio large language models can be descriptive speech quality evaluators,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation Audio large language models can be descriptive speech quality evaluators,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:04.274586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:04.274586Z digest=sha256:55579012e1ddde5466212b42531ed04f8f17e4d1e9ba12a8125ba97d22ea99d6

Observation 5568fe46-e569-4d0e-a5be-fa2833a76138 · outbound

This paper cites EmergentTTS- Eval: Evaluating TTS models on complex prosodic, expressiveness, and linguistic challenges using model-as-a-judge,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation EmergentTTS- Eval: Evaluating TTS models on complex prosodic, expressiveness, and linguistic challenges using model-as-a-judge,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:04.385402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:04.385402Z digest=sha256:42f950ba0b555df6a1ee54836e7642d2733b8127064a526323656a8f99ecb7a1

Observation ce6e5120-c663-4fa3-805a-fd0067f1c597 · outbound

This paper cites InstructTTSEval: Benchmarking Complex Natural-Language Instruction Following in Text-to-Speech Systems.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation InstructTTSEval: Benchmarking Complex Natural-Language Instruction Following in Text-to-Speech Systems

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:04.544761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:04.544761Z digest=sha256:72e03b32d2a30a5e037976377fb15f4c4e7b4a14bb67d04c0ab168bbca1cfeeb

Observation 0fe25140-e892-4b65-bf95-0830f5199c99 · outbound

This paper cites SpeechLLM-as-judges: Towards general and inter- pretable speech quality evaluation,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation SpeechLLM-as-judges: Towards general and inter- pretable speech quality evaluation,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:04.704822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:04.704822Z digest=sha256:731b4abbaba0217e763c7eb6a77da1dad69e14393e2e57b260e8cde1108e5f46

Observation abfb1ffd-54bd-40e1-a1f8-80b90e0f420f · outbound

This paper cites AudioJudge: Understanding What Works in Large Audio Model Based Speech Evaluation.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation AudioJudge: Understanding What Works in Large Audio Model Based Speech Evaluation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:04.890814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:04.890814Z digest=sha256:35524085a024280fafaf71283c8b4b8ba774832da8b3542a7ff01d394a54d50d

Observation 9bfb8815-c8b4-47e9-829c-5b24c269b73d · outbound

This paper cites Hearing between the lines: Unlocking the reasoning power of LLMs for speech evaluation,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation Hearing between the lines: Unlocking the reasoning power of LLMs for speech evaluation,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:05.039877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:05.039877Z digest=sha256:5dad8c4f100b3c2acf25edf45e4726457726c0a24280d5f6d0e6409f142c008a

Observation 1430c498-e0b0-49b9-9ab1-9ef3d9a7ae11 · outbound

This paper cites All That Glitters Is Not Audio: Rethinking Text Priors and Audio Reliance in Audio-Language Evaluation.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation All That Glitters Is Not Audio: Rethinking Text Priors and Audio Reliance in Audio-Language Evaluation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:05.235383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:05.235383Z digest=sha256:6d2c93928dddfe288a6edaed853f78ab65943cf7a47ccf198bf949cf867031db

Observation b37c2199-a75a-46c7-92f6-9a89ef8297ea · outbound

This paper cites Do audio LLMs really LISTEN, or just transcribe? measuring lexical vs. acoustic emotion cues reliance,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation Do audio LLMs really LISTEN, or just transcribe? measuring lexical vs. acoustic emotion cues reliance,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:05.422904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:05.422904Z digest=sha256:8d25a66b6b697a6bdff3a26a076129565933f40e7433537acbb597af942ce22e

Observation 552c5de6-83c0-4459-9a0e-9efe7e31b24d · outbound

This paper cites When audio-LLMs don’t listen: A cross-linguistic study of modality arbitration,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation When audio-LLMs don’t listen: A cross-linguistic study of modality arbitration,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:05.582882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:05.582882Z digest=sha256:6582581518bc4e896752b918561d0d477f8a62931e40dabb088cb7f5bda7dcdf

Observation cb618e57-5fce-451e-b284-8b18b0649791 · outbound

This paper cites Do audio LLMs listen or read? analyzing and mitigating paralinguistic failures with V oxParadox,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation Do audio LLMs listen or read? analyzing and mitigating paralinguistic failures with V oxParadox,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:05.702454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:05.702454Z digest=sha256:a43288efaf47bd072073a6acbacca6c995100f8f366c5fda975694a28a5df042

Observation 99656150-51ed-41df-a093-a938cf9b832c · outbound

This paper cites LALM-as-a-Judge: Benchmarking large audio-language models for safety evaluation in multi-turn spoken di- alogues,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation LALM-as-a-Judge: Benchmarking large audio-language models for safety evaluation in multi-turn spoken di- alogues,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:05.880445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:05.880445Z digest=sha256:a37c0badca62b3e7434d0ec0ae7ecbba34c7ec7ed69ad734eb124f3d8a112f78

Observation 52e711d1-46c8-4589-ad52-a8cf90dae114 · outbound

This paper cites Large language models are not fair evaluators,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation Large language models are not fair evaluators,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:06.052440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:06.052440Z digest=sha256:db1d5a8fa9b7591d5e56537caebbe05d33fa3597f2c64a76c2e82414af0299af

Observation eb1b5ce6-eeb0-4903-9310-a72dba581d3a · outbound

This paper cites Justice or prejudice? quantifying biases in LLM-as-a-judge,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation Justice or prejudice? quantifying biases in LLM-as-a-judge,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:06.177892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:06.177892Z digest=sha256:506b044abb69e8a6c05da98c5929a002537dfe13bb78a2453f178c5072a8c424

Observation b9421d3c-a333-428d-94c2-c6c9cb8f68da · outbound

This paper cites Towards understanding sycophancy in language models,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation Towards understanding sycophancy in language models,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:06.323955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:06.323955Z digest=sha256:a6943d75797db8de935570e197406910dab09795eb281e31fadeb86097c1a64d

Observation d96c7da1-a475-4ce9-9f1f-12e87f54d0e4 · outbound

This paper cites SycEval: Evaluating LLM sycophancy,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation SycEval: Evaluating LLM sycophancy,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:06.446080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:06.446080Z digest=sha256:d8fa1338c548e1551fdebf523621245b115a8cf7e919078f1b89d310f795cf4c

Observation 8c741d28-874c-4861-80b2-950acba9fa9f · outbound

This paper cites Shortcut learning in deep neural networks,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation Shortcut learning in deep neural networks,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:06.546354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:06.546354Z digest=sha256:9df809c5c7c977be3d21347cf4f2a663f1a56274428af084bad65bcc844c90f6

Observation 467f0370-3211-4d19-91c2-3864eaf79c16 · outbound

This paper cites The pitfalls of simplicity bias in neural networks,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation The pitfalls of simplicity bias in neural networks,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:06.676234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:06.676234Z digest=sha256:cd4af3a8c37060a6f6197c8824bb2719420b7cd3e81644a2cd41c532519e9724

Observation 7149e545-69d2-41c0-8536-f05c89b5fb22 · outbound

This paper cites Large language models can be lazy learners: Analyze shortcuts in in-context learning,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation Large language models can be lazy learners: Analyze shortcuts in in-context learning,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:06.804051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:06.804051Z digest=sha256:d46a36f6be9949a233ff933a9c6bd24bd965e385f3a4bfed9dde69db38b761d4

Observation 75ece925-619b-4109-b033-21b4acfbb9fc · outbound

This paper cites The Geneva minimalistic acoustic parameter set (GeMAPS) for voice research and affective computing,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation The Geneva minimalistic acoustic parameter set (GeMAPS) for voice research and affective computing,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:07.041096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:07.041096Z digest=sha256:08552c56e4857b1a0a6de21974a9472c29da5fcc415880d4898d30e4121ec48b

Observation edca2696-62a2-4645-82f3-660795770a74 · outbound

This paper cites The Ryerson Audio-Visual Database of Emotional Speech and Song (RA VDESS): A dynamic, multimodal set of facial and vocal expressions in North American English,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation The Ryerson Audio-Visual Database of Emotional Speech and Song (RA VDESS): A dynamic, multimodal set of facial and vocal expressions in North American English,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:07.218577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:07.218577Z digest=sha256:4f4eac924343efa2cc9daffc882084a86c61169131011b2f32abc6cd6b4e09cf

Observation 42618ece-aca4-4a7a-aa5d-ef143bae8e17 · outbound

This paper cites The V oiceMOS Challenge 2022,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation The V oiceMOS Challenge 2022,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:07.381205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:07.381205Z digest=sha256:ca144971e22462e6e7647f4529de550afa0023612375e9378a6123c039dc3237

Observation 2b7cfe02-09c9-4039-9a83-10a6ce90bbed · outbound

This paper cites V oxCeleb: A large-scale speaker identification dataset,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation V oxCeleb: A large-scale speaker identification dataset,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:07.501372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:07.501372Z digest=sha256:2c38e8c744bb388c6a2f515088eb2754d6ce5521728b7d0c7e6ac0531f4a123d

Observation 6ef20922-36e1-42a5-996f-95f93a549e0c · outbound

This paper cites emotion2vec: Self-supervised pre-training for speech emotion rep- resentation,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation emotion2vec: Self-supervised pre-training for speech emotion rep- resentation,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:07.610514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:07.610514Z digest=sha256:db946319a11269805e505ff034a5c6b941f6be340127b4b0b192e2cde256a8e4

Observation 943f673c-b636-4277-9aae-4564bb91cb29 · outbound

This paper cites Robust speech recognition via large-scale weak super- vision,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation Robust speech recognition via large-scale weak super- vision,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:07.753369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:07.753369Z digest=sha256:78c755743007b003e650e62e8d0b4254551becd74e35744245c2142937602313

Observation e21f231a-651a-42ae-b072-7f838562f484 · outbound

This paper cites ECAPA-TDNN: Emphasized channel attention, propagation and aggregation in TDNN based speaker verification,.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation ECAPA-TDNN: Emphasized channel attention, propagation and aggregation in TDNN based speaker verification,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:08.044744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:08.044744Z digest=sha256:5657518ef5cc3a8f0561bfc89528a21e58b56733f7e9b817be8f551df38f42f7

Observation 65858456-936b-4c05-b5f7-9ee5ad7be467 · outbound

This paper cites 28 492–28 518.

Auditing Protocol-Level Shortcuts in Large Audio Language Model Judges for Speech Evaluation 28 492–28 518

Reference 202

Resolution
unresolved
no resolver link, observed 2026-08-02T05:08:07.869223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:08:07.869223Z digest=sha256:83edda929f88d855e81ff67d7b048fa365c53bc64ea0f16ac2777db06e35d785

Pith citing papers

No inbound Pith citation observations are available.