Pith. sign in

Paper Citation Record · LEDGER

Auditory Intelligence: Understanding the World Through Sound

As of 12 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 0 inbound Pith citation observations for arXiv:2508.07829.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.07829 v1

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T21:55:00.437900Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

51 of 51 outbound references displayed

  • verified exact9
  • verified fuzzy38
  • unresolved4
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d799a26c-be9e-4a90-bdc0-004f4d199ce9 · outbound

This paper cites GPT-4 Technical Report.

Auditory Intelligence: Understanding the World Through Sound GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T21:54:55.018861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:54:55.018861Z digest=sha256:7aded76d7c267b300308d1bc7d682f4f644e704efa759f520566abc08974eb90

Observation 4e1cdf58-475e-45f1-b332-78498679b8c4 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Auditory Intelligence: Understanding the World Through Sound Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T21:54:55.093314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:54:55.093314Z digest=sha256:12cd264576ee8984009db792f7a526fcc349cd1ea65a7aaacb280b670a26f0b0

Observation 24a23040-184e-4408-8fb7-4eb7dd8f6171 · outbound

This paper cites DeepSeek-V3 Technical Report.

Auditory Intelligence: Understanding the World Through Sound DeepSeek-V3 Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T21:54:55.207940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:54:55.207940Z digest=sha256:b6dec2da8e809b98addb48efac8bb1595280c059609efd25ff6da440b97a808f

Observation 706dea90-ca0a-4a26-8799-cf7c8e25b31a · outbound

This paper cites Virtanen, M.

Auditory Intelligence: Understanding the World Through Sound Virtanen, M

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:09.763230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:55.290539Z digest=sha256:a1dfc11ab0292e184b8ede3c3ffa1ea68203382dfe8887d14985d622ff726677

Observation fef540cf-73ea-4132-a74f-d3f7504bc6b3 · outbound

This paper cites Sound event detection in domestic environments with weakly labeled data and soundscape synthesis,.

Auditory Intelligence: Understanding the World Through Sound Sound event detection in domestic environments with weakly labeled data and soundscape synthesis,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:09.449372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:55.368526Z digest=sha256:0136bc69408ef24a9a3833e5dc8e9ebe6d329514c19e36e541dd64ba3cd23feb

Observation 5bd1999d-f935-467f-9f97-9e972676f921 · outbound

This paper cites Convolutional recurrent neural networks for polyphonic sound event detection,.

Auditory Intelligence: Understanding the World Through Sound Convolutional recurrent neural networks for polyphonic sound event detection,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:09.138402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:55.459643Z digest=sha256:bcd63f9a9dc3710626e231fc7bf46e343bbae30d1cb0a07879bb665c02758bb4

Observation eb209ae2-e210-4867-93c9-415fbddc20b0 · outbound

This paper cites Metrics for polyphonic sound event detection,.

Auditory Intelligence: Understanding the World Through Sound Metrics for polyphonic sound event detection,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:08.824581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:55.545487Z digest=sha256:516748eb7982fccad23d14a2f6d393b7ab822fb2b79af016a0da250dd20a07d2

Observation 230cba28-708a-4c03-b05e-054c3d08235c · outbound

This paper cites A framework for the robust evaluation of sound event detection,.

Auditory Intelligence: Understanding the World Through Sound A framework for the robust evaluation of sound event detection,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:08.447221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:55.616813Z digest=sha256:2155e35be90ffc3a7124c042a40d68faa295560887fb479a3f7ea63b9ee1a0c8

Observation 9aa04b05-3827-4cbe-a996-4e467a4ee220 · outbound

This paper cites Towards Understanding of Frequency Dependence on Sound Event Detection.

Auditory Intelligence: Understanding the World Through Sound Towards Understanding of Frequency Dependence on Sound Event Detection

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:55:02.185813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:55.618843Z digest=sha256:342f0f775ee46a0faa088fb7345ecee84f98ed91df189f3e2d61544ce72e2605

Observation b63d5103-2042-45f4-a322-ae8232c20e11 · outbound

This paper cites JiTTER: Jigsaw Temporal Transformer for Event Reconstruction for Self-Supervised Sound Event Detection.

Auditory Intelligence: Understanding the World Through Sound JiTTER: Jigsaw Temporal Transformer for Event Reconstruction for Self-Supervised Sound Event Detection

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:55:02.001185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:55.621227Z digest=sha256:02ba7bd1500e22a795e8bd9b487b6a20ae3215ffece8a36736c42d1bf6747f77

Observation 5af3869f-a693-4e89-8801-49d2dc06c548 · outbound

This paper cites SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition,.

Auditory Intelligence: Understanding the World Through Sound SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:08.124515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:55.623704Z digest=sha256:0ef1187b19464300923458f67fd37216fe3dac1eed224021cf7140eb810e2b26

Observation cd7fd502-ae51-4374-af28-ea291568350d · outbound

This paper cites Conformer: Convolution- augmented Transformer for Speech Recognition,.

Auditory Intelligence: Understanding the World Through Sound Conformer: Convolution- augmented Transformer for Speech Recognition,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:07.854378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:55.696331Z digest=sha256:992013039fd6d0416123a2145c2b99b32482257cb5fc42131ce555a12cbd2cd8

Observation e6ce8a0c-c5a9-4e18-8add-d94a0a1e06a2 · outbound

This paper cites Coherence-based phonemic analysis on the ef- fect of reverberation to practical automatic speech recognition,.

Auditory Intelligence: Understanding the World Through Sound Coherence-based phonemic analysis on the ef- fect of reverberation to practical automatic speech recognition,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:07.597433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:55.792917Z digest=sha256:9f05b0f041d17af35c0442f38b34c3c902735a62a85cb492d6a6228c96a33152

Observation 9e209df9-6509-4644-808f-04218b52ba55 · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech representations,.

Auditory Intelligence: Understanding the World Through Sound wav2vec 2.0: A framework for self-supervised learning of speech representations,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:07.258853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:55.856904Z digest=sha256:5237b390b2525ede306cc5f663742fcf0722518cd5d8144f31f7780e0c091e57

Observation 8f743f9e-7593-4972-969e-e9173af92efa · outbound

This paper cites Hubert: Self-supervised speech representation learning by masked prediction of hidden units,.

Auditory Intelligence: Understanding the World Through Sound Hubert: Self-supervised speech representation learning by masked prediction of hidden units,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:07.011330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:55.890183Z digest=sha256:9a4c41c20d94a894314f56d80b76bc30392ed5a64c2234e085f1083008f4cfea

Observation 60b67069-b919-4e57-8af0-b7ef5baa3df6 · outbound

This paper cites Attentive statistics pooling for deep speaker embedding,.

Auditory Intelligence: Understanding the World Through Sound Attentive statistics pooling for deep speaker embedding,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T21:54:56.043026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:54:56.043026Z digest=sha256:289e54bc2d4d63f63726fc8d782cfbedc4f32dae7f91a09157a67153b291ac73

Observation 68704b73-a938-4901-9684-5e008bac89d6 · outbound

This paper cites Exploring the encoding layer and loss function in end-to-end speaker and language recognition system,.

Auditory Intelligence: Understanding the World Through Sound Exploring the encoding layer and loss function in end-to-end speaker and language recognition system,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:06.780807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:56.245420Z digest=sha256:7ca9489ea687e76f7543d84c3a9ee39503aee49f11f7aebf1993481b59bb67bd

Observation 188f8834-4afc-4ca9-bfd0-63c267134a61 · outbound

This paper cites Analysis-based optimization of temporal dynamic convolutional neural network for text-independent speaker verification,.

Auditory Intelligence: Understanding the World Through Sound Analysis-based optimization of temporal dynamic convolutional neural network for text-independent speaker verification,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:06.713473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:56.417343Z digest=sha256:f3adfbf04b2b81260df22d36a60361684c343b7b332955889b232c9a0fa5211e

Observation e88911b6-715e-4cb2-abf6-f55cbece2c99 · outbound

This paper cites Integrating fre- quency translational invariance in tdnns and frequency positional in- formation in 2d resnets to enhance speaker verification,.

Auditory Intelligence: Understanding the World Through Sound Integrating fre- quency translational invariance in tdnns and frequency positional in- formation in 2d resnets to enhance speaker verification,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:06.589708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:56.581826Z digest=sha256:caa7d92bb78ac9a609639453f4b22ebda64a19e02fcb5c41e7bf140d38702b3b

Observation f6828fbc-3eea-4ec9-858f-b650999147e0 · outbound

This paper cites Convolution-based channel-frequency atten- tion for text-independent speaker verification,.

Auditory Intelligence: Understanding the World Through Sound Convolution-based channel-frequency atten- tion for text-independent speaker verification,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:06.462490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:56.697963Z digest=sha256:e535f3f27f7b77fdae78c82327307622e67f744ba1f4305ded22a6f79628db16

Observation b0868592-5570-465e-a0d1-f6f08c9e6e85 · outbound

This paper cites Panns: Large-scale pretrained audio neural networks for audio pattern recognition,.

Auditory Intelligence: Understanding the World Through Sound Panns: Large-scale pretrained audio neural networks for audio pattern recognition,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:06.313111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:56.862622Z digest=sha256:bd7ae42df0813a6b26ac72080c169ac61ed389582649e7093eaab9dc579d7b57

Observation 2de542e1-234c-4a89-9208-af40a6aa3584 · outbound

This paper cites Deep learning based cough detection camera using enhanced features,.

Auditory Intelligence: Understanding the World Through Sound Deep learning based cough detection camera using enhanced features,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:06.135508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:57.030581Z digest=sha256:b0f791c6a09bf7a6c1e55862c380672d5883dc3ff38384ab795eec8e276ae981

Observation d64d0547-8f1c-40a6-8018-c7acb8450811 · outbound

This paper cites Real-time sound recognition system for human care robot considering custom sound events,.

Auditory Intelligence: Understanding the World Through Sound Real-time sound recognition system for human care robot considering custom sound events,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:05.927205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:57.193849Z digest=sha256:21b43b56930abe3bd60f3b5755afdf87e98238b2cbeef5bdd158a83ea02a0c37

Observation f3480612-f4b5-44cb-815d-fbe705c16e93 · outbound

This paper cites Ast: Audio spectrogram trans- former,.

Auditory Intelligence: Understanding the World Through Sound Ast: Audio spectrogram trans- former,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:05.744457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:57.356267Z digest=sha256:3a668fc198465afa39c3d8bc5ea0c2bf6ac11f5e83cfb3241688adf1981af0ff

Observation 049bd74c-97f3-445e-8e5c-83cf7d744dfe · outbound

This paper cites Beats: Audio pre-training with acoustic tokenizers,.

Auditory Intelligence: Understanding the World Through Sound Beats: Audio pre-training with acoustic tokenizers,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:05.563083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:57.480436Z digest=sha256:602c2e1a5705caff49bd1a9fd946307be3b366008dea91b876199d21b1b94fe7

Observation a325f9c7-30a8-497f-9c02-cd21b79416e7 · outbound

This paper cites Overview and evaluation of sound event localization and detection in dcase 2019,.

Auditory Intelligence: Understanding the World Through Sound Overview and evaluation of sound event localization and detection in dcase 2019,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:05.397189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:57.627088Z digest=sha256:e4e85f4970e52884a6d9d5a58177f6b189bc5ba9660847eee146bd97cb9c50a0

Observation 02947e24-3dad-417e-89fe-c00cdcf9eba2 · outbound

This paper cites STARSS22: A dataset of spatial recordings of real scenes with spa- tiotemporal annotations of sound events,.

Auditory Intelligence: Understanding the World Through Sound STARSS22: A dataset of spatial recordings of real scenes with spa- tiotemporal annotations of sound events,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:05.229253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:57.738938Z digest=sha256:208a788d4915fa76a09df53d62fba0b2c08daa5ede00ff0d7db4c556c102ccda

Observation 9e0b3c59-c1e2-4daa-bfd4-278b4d3cd8c5 · outbound

This paper cites Data augmentation and squeeze-and-excitation network on multiple dimension for sound event localization and detection in real scenes,.

Auditory Intelligence: Understanding the World Through Sound Data augmentation and squeeze-and-excitation network on multiple dimension for sound event localization and detection in real scenes,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:05.050564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:57.899374Z digest=sha256:fbc8f28fc250e74c97a0a12a6298d88c4f24b6dc1583c3da6b675655fd97c15f

Observation 30625b5f-9cdb-4af7-b716-9cdc483e3830 · outbound

This paper cites Binaural Sound Event Localization and Detection based on HRTF Cues for Humanoid Robots.

Auditory Intelligence: Understanding the World Through Sound Binaural Sound Event Localization and Detection based on HRTF Cues for Humanoid Robots

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:55:01.847274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:58.024131Z digest=sha256:090c589fa14ae9713da92469ec7e983363ad2330947030e5bd0a0cb471eb852c

Observation 1e23078e-4c6b-4e27-8ea9-b43acbe34508 · outbound

This paper cites Automated audio captioning with recurrent neural networks,.

Auditory Intelligence: Understanding the World Through Sound Automated audio captioning with recurrent neural networks,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:04.881036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:58.170275Z digest=sha256:14d1c41f7116ede6117cb227b8b24aa4d78308161edeeeacf55b80c243ae86e3

Observation 1a0665ed-0268-45ac-b880-3125acd53a1d · outbound

This paper cites Clotho: an audio captioning dataset,.

Auditory Intelligence: Understanding the World Through Sound Clotho: an audio captioning dataset,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:04.676532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:58.286059Z digest=sha256:74fdc757b8d9d47100ad185bbbbe19d899a4210340ee8dbcf613c267074e4529

Observation 77818c72-35bb-4d60-bb00-17b96f46a42f · outbound

This paper cites Chatgpt caption paraphrasing and fense-based caption filtering for automated audio captioning,.

Auditory Intelligence: Understanding the World Through Sound Chatgpt caption paraphrasing and fense-based caption filtering for automated audio captioning,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:04.531602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:58.404359Z digest=sha256:a4102c9cd42ab6006c480330f2ca70b565a03b5fca9f88e6e8072f42eec54f64

Observation 556bbac1-d0bc-4baf-a758-c59642c58474 · outbound

This paper cites Mind the Domain Gap: a Systematic Analysis on Bioacoustic Sound Event Detection.

Auditory Intelligence: Understanding the World Through Sound Mind the Domain Gap: a Systematic Analysis on Bioacoustic Sound Event Detection

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:55:01.629069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:58.488390Z digest=sha256:c7f3e55b824aea60fe3ea275f6cec55dbf2f5c913fe670914be44a6c6e3461ba

Observation 9633a1d8-64e0-48fd-bcfc-fabe39d98b4f · outbound

This paper cites Few-shot bioacoustic event detection utilizing spectro-temporal receptive field,.

Auditory Intelligence: Understanding the World Through Sound Few-shot bioacoustic event detection utilizing spectro-temporal receptive field,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:04.395425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:58.572096Z digest=sha256:72b109eda145c4eec2907b35c2b97588867c7f98026909d67966b33a4108e0eb

Observation a613f9fe-77bb-4d9b-b2a5-ca500286eb07 · outbound

This paper cites Prtfnet: Hrtf individual- ization for accurate spectral cues using a compact prtf,.

Auditory Intelligence: Understanding the World Through Sound Prtfnet: Hrtf individual- ization for accurate spectral cues using a compact prtf,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:04.245019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:58.674078Z digest=sha256:8b600949511385baaeda58afd0a4cece8699aab1bdd4cdda3723e62be9c3b83e

Observation d6f13d25-fadb-4cf4-90c1-cdf14ff29460 · outbound

This paper cites Filteraugment: An acoustic environmental data augmentation method,.

Auditory Intelligence: Understanding the World Through Sound Filteraugment: An acoustic environmental data augmentation method,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:04.073974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:58.787521Z digest=sha256:33334c0b66984af7acde57c4c3ecace1c328f5ec93835f039a281eae0a5fbae1

Observation 369327be-28db-4ef5-baf5-9b2f03d07e45 · outbound

This paper cites DNN based HRIRs Identification with a Continuously Rotating Speaker Array.

Auditory Intelligence: Understanding the World Through Sound DNN based HRIRs Identification with a Continuously Rotating Speaker Array

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:55:01.438903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:58.901523Z digest=sha256:3ce7e4fdc1b5f75b047709dbe0f64c77062007fb4bc4403fb1e53b2f08b7ba5b

Observation 792542b4-877c-4b7c-922c-55834e3b1dff · outbound

This paper cites AudioLDM: Text-to-audio generation with latent diffusion models,.

Auditory Intelligence: Understanding the World Through Sound AudioLDM: Text-to-audio generation with latent diffusion models,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:03.901509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:58.988369Z digest=sha256:1495f1e43f59a6a8930e9600cd0bf409e392d371fddaa4a1883c929d7a353df6

Observation ada6fa31-0d43-4223-a3dc-defab0890554 · outbound

This paper cites Audiogen: Textually guided audio generation,.

Auditory Intelligence: Understanding the World Through Sound Audiogen: Textually guided audio generation,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:03.753031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:59.102365Z digest=sha256:c870fa6391506a2350ac8f4854c11dc7c54b2b0a8ce9144620d9977ee17f1e20

Observation d7c109d0-d67f-4729-bc6d-1c51d4b0c51b · outbound

This paper cites Vifs: An end-to-end variational inference for foley sound synthesis,.

Auditory Intelligence: Understanding the World Through Sound Vifs: An end-to-end variational inference for foley sound synthesis,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:03.554341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:59.215508Z digest=sha256:01499d1804742cb4aef1e7578d7b1937979dbf9dca928e2beb7d9313cc8921b0

Observation 7bcbe46f-95b9-49f2-91b4-e1d3c2899ae0 · outbound

This paper cites Heavily augmented sound event detection utilizing weak predictions,.

Auditory Intelligence: Understanding the World Through Sound Heavily augmented sound event detection utilizing weak predictions,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:03.411377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:59.349962Z digest=sha256:7ce9a07f792289d56c858c2e69b16d048277f33a4e459aa399add065d2ecbf9e

Observation 21a5f27c-e399-4348-a384-1d74fc22a414 · outbound

This paper cites Filteraugment: An acoustic environmental data augmentation method,.

Auditory Intelligence: Understanding the World Through Sound Filteraugment: An acoustic environmental data augmentation method,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:03.264445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:59.472133Z digest=sha256:9b9561733c44085ac33ce547687e528864b9749fcc9c25e957f0dba9f99b90d3

Observation 1973ca4d-8cb7-4cb9-a7b5-7759246abec0 · outbound

This paper cites Frequency Dynamic Convolution: Frequency-Adaptive Pattern Recognition for Sound Event Detection,.

Auditory Intelligence: Understanding the World Through Sound Frequency Dynamic Convolution: Frequency-Adaptive Pattern Recognition for Sound Event Detection,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:03.075348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:59.555478Z digest=sha256:cc7242bd6f2ad654cbdcf5df0c6ba5e8c64dfe90a9350d46d5222b3c37ec4dfc

Observation ded12c34-a415-43b6-9e1d-3e0d9102ebec · outbound

This paper cites Self training and ensembling frequency dependent networks with coarse prediction pooling and sound event bounding boxes,.

Auditory Intelligence: Understanding the World Through Sound Self training and ensembling frequency dependent networks with coarse prediction pooling and sound event bounding boxes,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:02.909992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:59.654198Z digest=sha256:10808143b7c91ba4575bd99db4ba7edf65a1128602d2ac47f30ee72a8d673426

Observation 0ae9c7db-d015-4fed-a50c-4e1f603a90e2 · outbound

This paper cites Diversifying and expanding frequency-adaptive convolution kernels for sound event detection,.

Auditory Intelligence: Understanding the World Through Sound Diversifying and expanding frequency-adaptive convolution kernels for sound event detection,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:02.747332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:59.795404Z digest=sha256:bed583a33b7b8fcf489f205731b429e4f8c44a022802ffba962801f67333df64

Observation 05d450ad-e4a7-4d4c-820a-035c26b9ba0f · outbound

This paper cites Pushing the Limit of Sound Event Detection with Multi-Dilated Frequency Dynamic Convolution.

Auditory Intelligence: Understanding the World Through Sound Pushing the Limit of Sound Event Detection with Multi-Dilated Frequency Dynamic Convolution

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:55:01.245826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:59.894205Z digest=sha256:4165484ee31e40dd6c520a9dc7711326ae627a5cdf773f7ef168faeb1a3df709

Observation 72f54b94-765d-44ca-ac9e-ddc6677757af · outbound

This paper cites Temporal Attention Pooling for Frequency Dynamic Convolution in Sound Event Detection.

Auditory Intelligence: Understanding the World Through Sound Temporal Attention Pooling for Frequency Dynamic Convolution in Sound Event Detection

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:55:01.067562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:54:59.989934Z digest=sha256:e00757846e1617a1090b3cdc1d0c0b099c8d702e9bb37ea40931218b5218795e

Observation 773b6fc6-25d1-44e5-86c6-7f4ef1b8a7d8 · outbound

This paper cites Frequency Dynamic Convolutions for Sound Event Detection.

Auditory Intelligence: Understanding the World Through Sound Frequency Dynamic Convolutions for Sound Event Detection

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:55:00.851246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:55:00.075447Z digest=sha256:060793b8260fc79eba0bd71a489a0b37c8ac825c96b9d1fc50cbdf2278044cf4

Observation f3f7fd21-81b5-464d-82dc-1d6b272b7750 · outbound

This paper cites FLAM: Frame-Wise Language-Audio Modeling.

Auditory Intelligence: Understanding the World Through Sound FLAM: Frame-Wise Language-Audio Modeling

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:55:00.654739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:55:00.164281Z digest=sha256:0283eb989370302f9ae0ae7b8de60adf0cc54039e24b31ca8fc93e0af2e06397

Observation 2d676abf-5417-479d-85b2-bcc09366cf9a · outbound

This paper cites AudioCaps: Generating captions for audios in the wild,.

Auditory Intelligence: Understanding the World Through Sound AudioCaps: Generating captions for audios in the wild,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:02.500837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:55:00.299272Z digest=sha256:3d8ffab20243c243093dc5b4f2d6953934a40568e440acd8346e208cd2051cc6

Observation 476396e1-5fa4-4778-a4b8-db33cee5beeb · outbound

This paper cites Wavcaps: A chatgpt-assisted weakly-labelled audio captioning dataset for audio-language multimodal research,.

Auditory Intelligence: Understanding the World Through Sound Wavcaps: A chatgpt-assisted weakly-labelled audio captioning dataset for audio-language multimodal research,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:02.344444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:55:00.437900Z digest=sha256:8e9fbfcd20c0fb491a5893514b1c330ade5142da51ebe03afb27f457f20fc93d

Pith citing papers

No inbound Pith citation observations are available.