Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:48:15.208113Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 1 inbound Pith citation observation for arXiv:2505.23509.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:48:15.208113Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:48:12.576291Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T12:48:15.322444Z
31 of 31 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 8544dcc9-e87f-4db7-9eb8-e673504d47a4 · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds A critical step in developing powerful audio deep neu- ral networks (DNNs) is converting audio signals into meaning- ful acoustic feature representations
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0aadf0c6-816f-41b8-8f4b-384280ddac7f · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f2750332-72bc-4a06-bf00-a49fa2093d16 · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 24d82cb6-04c9-4e60-9ef4-d5c3262ff414 · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8675b5d3-a907-408a-acc6-b2e0b0a43027 · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fe22b275-bcc3-4752-a3a7-eac1d6bcf3e4 · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds A survey of audio classification using deep learning,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3c3ad048-3e4c-406b-9f9f-526af2a1075f · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds AST: Audio spectrogram transformer,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4d6b88ac-86f1-4ff6-b5ae-eed6d1bcaed5 · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds Audio set: An ontology and human-labeled dataset for audio events,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1a8f05df-e420-4517-8438-6cb96e23429b · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds CNN architectures for large-scale audio classification,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5e6ade39-ae94-47c4-a668-e71ca072b41d · outbound
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 40e6dc28-1cd9-4d32-9a18-9f5a3ab4abbb · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds Robust DOA esti- mation from deep acoustic imaging,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ea9ece3b-248c-422b-b87c-8eede74c9d04 · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds A review of differentiable digital signal processing for music and speech synthesis,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7f2f58c9-283b-4aca-960a-a70352002351 · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds The modulation spectrogram: In pursuit of an invariant representation of speech,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 877c4e90-4e92-45e6-b052-e052aa694581 · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds Speech intel- ligibility prediction using spectro-temporal modulation analysis,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dcebb7a8-0ab8-4381-9f6a-33e0796c1f89 · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds Com- paring different flavors of spectro-temporal features for ASR
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8e4bf80a-5f0b-4b2b-9399-418f04723c79 · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds Automatic mu- sic genre classification based on modulation spectral analysis of spectral and cepstral features,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 356e24d9-5f44-49f9-8696-51733ebf700c · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds Discrimination of speech from nonspeech based on multiscale spectro-temporal modulations,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 69ea92cb-6151-4cd2-8565-9a4b2a858655 · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds Distinct sensitivity to spectrotemporal modulation supports brain asym- metry for speech and melody,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b5adde92-8f18-4045-ad0d-a933daca95a4 · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds Spectrotemporal modulation provides a unifying framework for auditory cortical asymmetries,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 835fff49-c290-46c4-a377-b5dc8068f740 · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds Spectro-temporal acoustical markers differenti- ate speech from song across cultures,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 21b5f124-a30b-4780-a68e-532498bba7ad · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds The human auditory system uses amplitude modulation to distinguish music from speech,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4242752c-6a66-456a-a2c9-c71f7ea44354 · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds Distinct cortical pathways for music and speech revealed by hypothesis-free voxel decomposition,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 90697415-7289-4fb8-8c05-010c665b4efd · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds Hybrid Transformers for Music Source Separation,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 76195d72-4c58-4a45-9e20-519b6659d383 · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds SONYC Urban Sound Tagging (SONYC-UST): a multilabel dataset from an urban acoustic sen- sor network,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2a17fc99-58f8-456f-831c-f80d8b0c6432 · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds eBird: A citizen-based bird observation network in the biological sciences,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1245d9cd-6b81-4dc6-87fa-a517e5e32531 · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds The cortical organization of speech processing,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0f0ad36d-613d-47db-8681-092911200ac5 · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds Facing Imbalanced Data Recommendations for the Use of Performance Metrics,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 86b0af17-e0d9-46a4-a594-1bca5ffc8f7c · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds Survey on deep learning with class imbalance,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9801d646-ddce-47f4-9124-7696b9d7ee52 · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds Many but not all deep neural network audio models capture brain responses and exhibit correspondence between model stages and brain regions,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 10cc63e4-d7eb-4b2f-b089-2da056a4b2ed · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds Reconstructing the spectrotem- poral modulations of real-life sounds from fMRI response pat- terns,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5d5eec24-f548-4559-977b-dfe74a136534 · outbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds Decoding spectrotemporal features of overt and covert speech from the hu- man cortex,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0aadf0c6-816f-41b8-8f4b-384280ddac7f · inbound
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.