Pith. sign in

Paper Citation Record · LEDGER

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance

As of 22 July 2026, this Paper Citation Record lists 79 of 79 outbound references and 0 inbound Pith citation observations for arXiv:2509.23727.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.23727 v2

Coverage vector

measured 79 of 79 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-18T13:10:18.700497Z

measured 79 of 79 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-07-22T06:31:00.163083+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

79 of 79 outbound references displayed

  • verified exact24
  • verified fuzzy53
  • unresolved0
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a4e84992-2612-4d42-81c5-50ea7e8e67ed · outbound

This paper cites MusicLM: Generating Music From Text.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance MusicLM: Generating Music From Text

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:11:23.738082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:be50b4082e35fecf456215d9d2ea7a33c6c3fd232111c445f829a953f3571fc4

Observation 81484efc-df85-4baa-b34b-a0fda6b5402f · outbound

This paper cites The million song dataset.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance The million song dataset

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.764585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:252ea75ac914c63fe771c7c669c43010997fe0a279d989dab9e4597d00c272db

Observation 2c3f4c62-6941-4e2a-ab30-29c8e2662c8e · outbound

This paper cites Classifier-Free Guidance is a Predictor-Corrector.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Classifier-Free Guidance is a Predictor-Corrector

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:11:23.744095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:421370f88c0e8b6382ce16c5757131974d6a480c366162504e49e0493947d9d1

Observation 0f3e69ae-0b13-477f-b634-8a018cab9032 · outbound

This paper cites Vggsound: A large-scale audio-visual dataset.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Vggsound: A large-scale audio-visual dataset

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.822172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:dd384e67bfb7ca2781402d2051f4b2fad6802b698d7ec9130fa977afae1ebe97

Observation b1472ab8-8d6f-4cc3-8f51-7cea02260ac6 · outbound

This paper cites Mmaudio: Taming multimodal joint training for high-quality video-to-audio synthesis.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Mmaudio: Taming multimodal joint training for high-quality video-to-audio synthesis

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.933944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:47f97406fc38b6cb673b2995e850156322b35b88461d869df216d0e63d5917eb

Observation 4d520348-3d92-41fe-b1e3-8f0477b6d88e · outbound

This paper cites What does guidance do? A fine-grained analysis in a simple setting.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance What does guidance do? A fine-grained analysis in a simple setting

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:11:23.784504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:919ae54ff88afa1fcd4d7f5ff91096afcda6221351ebdc61a01cddcc04c69421

Observation d4c794e9-e4b4-4ccf-8781-d8592472710d · outbound

This paper cites Scaling instruction-finetuned language models.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Scaling instruction-finetuned language models

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.953254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:67f6e3853031996121db45f54bfa5eee76dd852581874d8ca06e285e12389347

Observation 66853980-4fd6-49d8-aac3-a13df09cb649 · outbound

This paper cites Syncfusion: Multimodal onset-synchronized video-to-audio foley synthesis.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Syncfusion: Multimodal onset-synchronized video-to-audio foley synthesis

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.957152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:ee42414d3d4319b55b9e1a0fc9afdd36fb0ce4b4b12fda997b9f0e5422e4c4b8

Observation 544377e3-aea6-4183-86d0-1028397fd91f · outbound

This paper cites Simple and controllable music generation.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Simple and controllable music generation

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.928064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:1511d4daae10a2fe5ece878d684690f3080106b1d6084a7e5836105c48115539

Observation 8205c891-3462-4416-a8a8-d7edb75a6565 · outbound

This paper cites Text-to-Audio Generation using Instruction-Tuned LLM and Latent Diffusion Model.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Text-to-Audio Generation using Instruction-Tuned LLM and Latent Diffusion Model

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:11:23.841838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:8e4b965d715309d5eaabddcd1d35a52c90a95a9a27f5f57de74b2ce181e8ad97

Observation e95830f5-79d3-4cc2-845b-f90fd8f009c5 · outbound

This paper cites FMA: A Dataset For Music Analysis.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance FMA: A Dataset For Music Analysis

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:11:23.846476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:e3f6b44972e94df621e50672f7b94042550eafdb8bb6e5793946cf86d3394e15

Observation 6f80ec5b-8cea-40e4-8b43-cce94cf62352 · outbound

This paper cites Clotho: An audio captioning dataset.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Clotho: An audio captioning dataset

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.804437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:6ac1f4c5fe08b9d91149bd404406761af0502b5824e0da91cf663866f9c0ca4c

Observation a5756389-abee-4b20-b709-ce4a3c10883a · outbound

This paper cites Conditional generation of audio from video via foley analogies.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Conditional generation of audio from video via foley analogies

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.814056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:847d766871b9d16e4f6efec8e5a7792984dee2e9d472ea6564767687853095da

Observation 13c96fa3-2e5e-4572-92ac-06d517cf92ce · outbound

This paper cites Fast timing-conditioned latent audio diffusion.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Fast timing-conditioned latent audio diffusion

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.808052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:e49c8b841d4a3b74041c01c8a7fc1564c48ead32c749903f5a7bed140975c8f9

Observation a58daf82-ee2b-406d-9ddf-ec89349fe941 · outbound

This paper cites Stable audio open.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Stable audio open

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.810842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:f4b279fc08ffaf0adf45fa632435a149d0a5538cdb80d37e8b5720ff936ba4b9

Observation 34559a5c-e750-4746-96e2-924bb9460fae · outbound

This paper cites Fsd50k: an open dataset of human-labeled sound events.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Fsd50k: an open dataset of human-labeled sound events

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.819767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:8a5d869bf0e4ddbf0fbbeeff4e1e16038751987246c5b4bc5218cdada664562b

Observation 1f2994e9-1a7d-4491-a6d4-e4ae2f8fa8ff · outbound

This paper cites Riffusion-stable diffusion for real-time music generation.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Riffusion-stable diffusion for real-time music generation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.816945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:410d13ab0fa463bf6a0b0f701c7aa0bda0b6b4f4e07d3a0f34f0fa3234bd7658

Observation 42a741e9-9276-49ee-be08-5ed0d5388ed7 · outbound

This paper cites Audio set: An ontology and human-labeled dataset for audio events.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Audio set: An ontology and human-labeled dataset for audio events

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.939782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:25f6c2e76003c0ff673c977b81dca00d15a317e323b4e788209c811ed7c7d285

Observation 7d59908c-228c-4580-bf01-64494771ae7b · outbound

This paper cites Imagebind: One embedding space to bind them all.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Imagebind: One embedding space to bind them all

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.943303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:6d50d76d354e165ceb33a902c396ca1bc6eb6c6f33caa987bc53b08c6e7be9b6

Observation c675b593-b3bc-4f92-b533-5414b37be5a0 · outbound

This paper cites Gemmeke, Aren Jansen, R.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Gemmeke, Aren Jansen, R

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.930661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:d4b8496ceee3f39bd0115ff4091020ea59751c678f5fc8216dc9b9ea7f728a46

Observation cdbbde3d-e43e-445e-9946-4049ad6cf896 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilibrium.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Gans trained by a two time-scale update rule converge to a local nash equilibrium

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.947581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:2e36facd22a19c2bcb910b8a1a88fe51ba773b2c3a1e956e523a1feb198f5702

Observation 48f138c7-2879-438a-902e-b9fd37e28d01 · outbound

This paper cites Classifier-Free Diffusion Guidance.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Classifier-Free Diffusion Guidance

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:11:23.829946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:552305a3ab26970f4dfa168f76ceedff13729038d14ea1ea283ff8549a1544ab

Observation 641e5fc4-0c35-4e47-8749-562bbf929605 · outbound

This paper cites Improving sample quality of diffusion models using self-attention guidance.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Improving sample quality of diffusion models using self-attention guidance

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.921840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:57ad1d55b4a578d73bbb7278c6ea7b11a6aff7c8b1fb82c80498022700e7ebef

Observation 9eb675b3-b552-4e73-b57e-3ff154ace35e · outbound

This paper cites Make-An-Audio 2: Temporal-Enhanced Text-to-Audio Generation.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Make-An-Audio 2: Temporal-Enhanced Text-to-Audio Generation

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:11:23.759064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:3b7e5bb2e01c3b014d2782c1db566a6ce6f39882c11d8dfb8124a4d12f01afe4

Observation 388ebe9c-653c-4407-84b4-383b9f66e6bf · outbound

This paper cites Masked autoencoders that listen.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Masked autoencoders that listen

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.893822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:80c490e7dba1ee4d8d8cd6fd087ecd7f701aecc09826fb2c65d992e1025289a6

Observation fec7c95a-4838-4c99-8ad5-6ecb0fcd2a79 · outbound

This paper cites Make-an-audio: Text-to-audio generation with prompt-enhanced diffusion models.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Make-an-audio: Text-to-audio generation with prompt-enhanced diffusion models

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.899533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:f4de2ac161c5cb8f1dee195fc3be26660c36243bee2f5e0d9c9e8eb088a83128

Observation 7f2626d6-579d-4430-8d7b-a17467c8e3bc · outbound

This paper cites TangoFlux: Super Fast and Faithful Text to Audio Generation with Flow Matching and Clap-Ranked Preference Optimization.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance TangoFlux: Super Fast and Faithful Text to Audio Generation with Flow Matching and Clap-Ranked Preference Optimization

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:11:23.820034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:135074e020eadd64db3aff475f04fb68090bdd52f19bfe5257f23f7fe36abb5f

Observation f1fac55a-4d85-4ea7-bd2b-505deafb6d03 · outbound

This paper cites Spatiotemporal skip guidance for enhanced video diffusion sampling.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Spatiotemporal skip guidance for enhanced video diffusion sampling

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.918829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:4cbdf3e688b95d3baf9043a1bbf04096670999ae8c24d2b0298ff2e298b38b20

Observation c12da6aa-2a08-40f7-8181-06c7a4a4edc0 · outbound

This paper cites SPG: Improving Motion Diffusion by Smooth Perturbation Guidance.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance SPG: Improving Motion Diffusion by Smooth Perturbation Guidance

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:11:23.868072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:b3110e15ef6bb3b34e25a70f7067f28ec9626c72f1d543e031e5d92b860a1fae

Observation 689c51ec-fe7f-4d5f-a2af-89429960d0d5 · outbound

This paper cites Read, Watch and Scream! Sound Generation from Text and Video.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Read, Watch and Scream! Sound Generation from Text and Video

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:11:23.800632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:15033a6f41b619d77afe35ce9d8a778cf88716b660acac6c174c0476106994d5

Observation 5d6d9a29-b6fe-4f4d-9348-f845bf5f90bf · outbound

This paper cites Freeaudio: Training-free timing planning for controllable long-form text-to-audio generation.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Freeaudio: Training-free timing planning for controllable long-form text-to-audio generation

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:11:23.809871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:37a28f02adb218b09f3ce238b2a5f791500229e7e75be39c60b33af011aa645d

Observation 9389098b-a6bd-424a-9bd9-c82ab88b568a · outbound

This paper cites Guiding a diffusion model with a bad version of itself.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Guiding a diffusion model with a bad version of itself

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.876574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:cf45beab7bfe0ae573d8b363ff6bc15ca5fcd71641abeacf79d238f33fbf42f6

Observation 165b5df4-5b60-414f-9a12-2b49d91f9ee8 · outbound

This paper cites Analyzing and improving the training dynamics of diffusion models.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Analyzing and improving the training dynamics of diffusion models

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.879086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:ce435a50717594cb9aa8b51276c8c296148d7cbcf93f09ecfcb06b612effb0ef

Observation c321ed7c-a3cb-4807-96c3-da2922f48349 · outbound

This paper cites AutoLoRA: AutoGuidance Meets Low-Rank Adaptation for Diffusion Models.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance AutoLoRA: AutoGuidance Meets Low-Rank Adaptation for Diffusion Models

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:11:23.773291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:a239cec3227726aa6cd74e4a7e210d3bdbdac5d8f5e45d571908dd40547034e2

Observation 4fc3de16-b3c3-4833-93cc-a274a0116f44 · outbound

This paper cites Sim, and Paris Smaragdis.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Sim, and Paris Smaragdis

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.873131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:e9e0555a2cfd228930615763ee09c0cb961ce84fbc1bd45d442e92fdd189f034

Observation c9cd8c45-1194-4bd4-ab90-a119b60e021a · outbound

This paper cites Audiocaps: Generating captions for audios in the wild.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Audiocaps: Generating captions for audios in the wild

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.867618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:2f317e44accfd9f69b85c325a1debedab29890e0b25fddd5259d0e7f217bf4e3

Observation 42814239-0d86-41a5-8f33-f9e2db72a901 · outbound

This paper cites Auto-encoding variational bayes.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Auto-encoding variational bayes

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.870375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:4ec6873c92e68bd38e308f6f934cd762ec1520aef49465e8aa31d872556104e0

Observation e5115ff1-b127-43cd-985a-cf170f7fc5d6 · outbound

This paper cites Plumbley.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Plumbley

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.884411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:d56d629a109ac9d7d8e44be431847ea0d5ae5037daeed27dd2c30ce370d0546d

Observation 98d0d72c-814d-448f-96be-2c18c456fb85 · outbound

This paper cites Improving Text-To-Audio Models with Synthetic Captions.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Improving Text-To-Audio Models with Synthetic Captions

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:11:23.857545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:80110ec7d2597344202c52f0b6d42727756ddea3c96cc2007a5b7b0fdd44e98f

Observation cdc799a6-0e8a-482a-ad42-19bc116d3940 · outbound

This paper cites AudioGen: Textually Guided Audio Generation.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance AudioGen: Textually Guided Audio Generation

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T13:11:23.852477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:54e194ac649f774e26042cec8f3d7d0c8118e1032656b5c4a656eb7717d26f5d

Observation c7fb5466-eea3-47fe-bd06-73f895a2e2c5 · outbound

This paper cites an unresolved cited work.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Unresolved cited work

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.833667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:eedd82ebd2b159324b0421257801fa82684fd3d004a83fb02d397bd1b48e280a

Observation afff307d-8ec0-4e3a-a93d-a4c3991fd5cc · outbound

This paper cites Efficient neural music generation.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Efficient neural music generation

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.836493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:7f817ce5b39a090d479bf3bd593e0bc13f561d1f3ede30fa754fb3f1368df1ce

Observation 32c9a11d-0304-4293-a5df-3b3699793829 · outbound

This paper cites Evaluation of algorithms using games: The case of music tagging.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Evaluation of algorithms using games: The case of music tagging

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.912495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:1206adcead197b3c9e4f25e2cc830a853a455e314a0d596b42098593d283259a

Observation 0c2acfa3-591b-450c-8898-eaa313769d98 · outbound

This paper cites Etta: Elucidating the design space of text-to-audio models.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Etta: Elucidating the design space of text-to-audio models

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.830888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:8e708fd2bdf73019dc1103bf95c2c6c1fe618a9f250df23658e4b98289aa55d4

Observation ef7fabf5-ffa2-4348-9f6d-a5e7a3ac2eeb · outbound

This paper cites Quality-aware masked diffusion transformer for enhanced music generation.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Quality-aware masked diffusion transformer for enhanced music generation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.839476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:6d56dae14bb874b4f98fcc3c3c71ae0e8f4117051e6411a7890b220c27ac8866

Observation d6065d82-12c2-4860-bce6-300d78370f07 · outbound

This paper cites Jen-1: Text-guided universal music generation with omnidirectional diffusion models.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Jen-1: Text-guided universal music generation with omnidirectional diffusion models

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.842535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:00a89e4e8accddd9b3e5694756c3fc41b5586eee5d030f3617260a1b02b10c12

Observation e1f7960b-6d7f-4d38-b773-8abb4cd53724 · outbound

This paper cites Self-guidance: Boosting flow and diffusion generation on their own.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Self-guidance: Boosting flow and diffusion generation on their own

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.861376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:e9f0ca44f93b3b739ac914fe5706b478a2cf8458c7d643fefc145b7b66382a3f

Observation 29112a42-28a7-48ca-a089-aee4158f9105 · outbound

This paper cites AudioLDM: Text-to-Audio Generation with Latent Diffusion Models.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance AudioLDM: Text-to-Audio Generation with Latent Diffusion Models

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:11:23.790250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:f610d583942c694379d993a425d8548c7bdfc44e4517e3d98050e83128fc3b05

Observation eb5bd2c9-fd6f-4764-b71a-472e088c38aa · outbound

This paper cites Audioldm 2: Learning holistic audio generation with self-supervised pretraining.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Audioldm 2: Learning holistic audio generation with self-supervised pretraining

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.828013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:fd97ba0d8b283ff02dd441ddebaaf7570eafcccf3dbbac19999dcb15fc31baaf

Observation 79d0da04-f8d5-457c-9de3-3e08e889cf9d · outbound

This paper cites Tell what you hear from what you see - video to audio generation through text.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Tell what you hear from what you see - video to audio generation through text

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.799381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:ed65705920dff6eadd0019502cdd4e28cd11126bebad1600a5f3a20329b4239b

Observation 2c75ec9b-ddde-468f-bddf-75e013780ef9 · outbound

This paper cites Decoupled Weight Decay Regularization.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Decoupled Weight Decay Regularization

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:11:23.824722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:44cbbf6f125e375f36a35ee3b1fb7fd409da0281a334e853b3d963c4137549f9

Observation 4b6782ee-2a2c-4400-a7a1-6b8eaef5a30f · outbound

This paper cites DPM-Solver++: Fast Solver for Guided Sampling of Diffusion Probabilistic Models.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance DPM-Solver++: Fast Solver for Guided Sampling of Diffusion Probabilistic Models

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:11:23.749361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:e164db83dc0b5e8e1666aa0175454c2b09ae278f45d62b70f919466c6035b9ce

Observation 80cfe31a-d91d-4c9c-85e7-6f391036d7a6 · outbound

This paper cites Diff-foley: Synchronized video-to-audio synthesis with latent diffusion models.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Diff-foley: Synchronized video-to-audio synthesis with latent diffusion models

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.825273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:14c396c270bc904f719419817e74c80862a5572e075a2059b5accfa374e91090

Observation f7a8b0fd-ccf7-466e-979c-0f2307b2c0fe · outbound

This paper cites Tango 2: Aligning diffusion-based text-to-audio generations through direct preference optimization.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Tango 2: Aligning diffusion-based text-to-audio generations through direct preference optimization

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.776986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:512493b723df65d66cd7a2c73687e1d23c6a7e1e563e32fe581a75e23b565a34

Observation 99ee3880-62f1-4c87-bf69-31b3aeddf128 · outbound

This paper cites Foleygen: Visually-guided audio generation.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Foleygen: Visually-guided audio generation

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.779995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:c5a920ed5b77da5b674ec0fb5825e5abe229541147a1a7455ad0d00741556be8

Observation 97fe2c6e-5ef7-4fd4-85f8-f5ac0c1072cf · outbound

This paper cites an unresolved cited work.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Unresolved cited work

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.767446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:419150f87ebaedc2e5fff3bcaa6f3f5b4a18444e9234cd4065457089696ec13a

Observation 07ede07c-a0a1-4093-8ac7-a76968a76c19 · outbound

This paper cites Scalable diffusion models with transformers.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Scalable diffusion models with transformers

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.773427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:6ccf00affe190e7788f05684b1c713198af31def9d0f29c377ce327d81e3b0b3

Observation de644d8b-8b19-4349-a6fd-53bce01365f1 · outbound

This paper cites Unconditional priors matter! improving conditional generation of fine-tuned diffusion models.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Unconditional priors matter! improving conditional generation of fine-tuned diffusion models

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:11:23.732969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:8d66edac059f84e94d415f235c8e023be7758fcd33006c8efe6f41c9cef679f5

Observation 7a8de36d-cf3c-47c7-b551-0198db56fbfc · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Learning Transferable Visual Models From Natural Language Supervision

Reference 59

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:11:23.778353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:5bbf1ef3e079e95a027d57d335c0025fc00210b32dbc626cd25368622e569eeb

Observation 4caecfe5-df23-49ec-bc4d-399797a43750 · outbound

This paper cites Direct preference optimization: Your language model is secretly a reward model.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Direct preference optimization: Your language model is secretly a reward model

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.784347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:21d9a7b7cbefa8fb168b8594fc41bc58e8905bbf49896553e9638d0967877f70

Observation 51120807-e73e-4f8f-9dfe-619b50750a01 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance High-resolution image synthesis with latent diffusion models

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.793073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:bffc5afb73b048f5b17feec14cf618561715c9e1608bfb93c8c1f788091080aa

Observation 375ff664-876c-490f-95de-5c9b54e7f61d · outbound

This paper cites No Training, No Problem: Rethinking Classifier-Free Guidance for Diffusion Models.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance No Training, No Problem: Rethinking Classifier-Free Guidance for Diffusion Models

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:11:23.795440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:e9aaafc370f103cbba5cb1dab650764fdb8b19d6c6ea607e8db29d0049ddb077

Observation 8bca6522-2b12-4412-9e0e-6b84a7a41b9a · outbound

This paper cites Improved techniques for training gans.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Improved techniques for training gans

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.795967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:e405b7c397e35f9271c765dc9ddcb8a5ba683928ec060d2c50c0092e2e67a079

Observation dcd21e07-2051-4675-a944-4ea174227dff · outbound

This paper cites Mo\^usai: Text-to-Music Generation with Long-Context Latent Diffusion.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Mo\^usai: Text-to-Music Generation with Long-Context Latent Diffusion

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:11:23.753882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:21ff82c0fb140e27dbb17eef8a12e3d31b5123b933d4788bfd440c65a390bdb6

Observation 25f447be-2fba-4347-9118-afd64b22f80b · outbound

This paper cites I hear your true colors: Image guided audio generation.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance I hear your true colors: Image guided audio generation

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.962301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:f2828ad971e00b58d4535ec5f686213681f30d2466d08fdc5b2a8326208471c9

Observation 76428a5b-8200-43c9-a5fe-e62957cc6005 · outbound

This paper cites From Vision to Audio and Beyond: A Unified Model for Audio-Visual Representation and Generation.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance From Vision to Audio and Beyond: A Unified Model for Audio-Visual Representation and Generation

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:11:23.763813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:dccb030159315d86c9e00cac9182b7555ba87eea69a8eaa6db8156aa454b4d8a

Observation 6cab8fd4-9322-4fa6-8099-079e6a803d95 · outbound

This paper cites Neural discrete representation learning.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Neural discrete representation learning

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.906105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:a591031b2420edabb58395bb586f3e2cf305bdc9c2fe1ecf98d81921a14d5e5e

Observation a0be768d-85df-4bac-90e2-b26242fb4f8a · outbound

This paper cites V2a-mapper: A lightweight solution for vision-to-audio generation by connecting foundation models.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance V2a-mapper: A lightweight solution for vision-to-audio generation by connecting foundation models

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.890931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:926b9c5e2d4cf93e321412d62ec47bb54faea149a568db1d1fa5ca8d11172282

Observation fc09729f-35f4-4826-85ea-25047df365e9 · outbound

This paper cites Tiva: Time-aligned video-to-audio generation.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Tiva: Time-aligned video-to-audio generation

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.896478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:99899c20fd49438f63ae87c4fb2e41e302b56f51b7d6af7ddc15b35e5bf9959c

Observation 1fdeee9a-60fa-45d6-83b7-3b8ef0d0480a · outbound

This paper cites Large-scale contrastive language-audio pretraining with feature fusion and keyword-to-caption augmentation.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Large-scale contrastive language-audio pretraining with feature fusion and keyword-to-caption augmentation

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.909406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:578d69e3d166a3e70160e657828492259afd7d3f9cccad30d0b3dcb1c9e7f5dc

Observation 8c2c9650-4d58-4ca5-86e0-e4f4cb8ed2a3 · outbound

This paper cites Sonicvisionlm: Playing sound with vision language models.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Sonicvisionlm: Playing sound with vision language models

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.915578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:8354093fbccb7ec9e08696172e19245ab7180ddbde32a7b0b184acaa48ff1d5e

Observation 1c89f49b-6dd4-4868-8384-4f148663444a · outbound

This paper cites Video-to-Audio Generation with Hidden Alignment.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Video-to-Audio Generation with Hidden Alignment

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:11:23.814761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:7946b6b959ef4596db0f382cd75052e631d032bdb803f939b6b490bfd9a3fe74

Observation 2cdeb5f2-6be2-4c31-932c-87a77070e292 · outbound

This paper cites Diffsound: Discrete diffusion model for text-to-sound generation.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Diffsound: Discrete diffusion model for text-to-sound generation

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.925156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:914bc255fd6957a943a785ed7f6b1e3a43f164b585752f591faaa6d20c288f57

Observation bb145c55-3032-4ca0-b7ad-a7d9cd1ec4c1 · outbound

This paper cites FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:11:23.863177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:422e7bd6b2976d370a37b257cf11eb4361dc4878eab05160ff2c6b9c3be8d6d2

Observation df6ef58a-0497-428d-a197-094f0ee6c0d6 · outbound

This paper cites Domain Guidance: A Simple Transfer Approach for a Pre-trained Diffusion Model.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Domain Guidance: A Simple Transfer Approach for a Pre-trained Diffusion Model

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:11:23.769116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:3951cce7ac69677112637a62786fa86e60eaa1811042fc0bf01c2fd7ed744695

Observation 0f622c84-33b0-4e14-9ae5-208e681723c9 · outbound

This paper cites write newline.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance write newline

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.881780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:b702ca7de1c91a2d5cdea9aee89fb11727c5396b99bee5c457b8273e13102f5c

Observation 668b3409-581c-40dd-b202-4c57705a595d · outbound

This paper cites @esa (Ref.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance @esa (Ref

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.887932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:9d56f6032e0468fc74e8b9d837a03b49633d4aaf42384e634081edf73cc30884

Observation a2b97fa1-460d-4d41-852b-60ee1f55d1ec · outbound

This paper cites an unresolved cited work.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Unresolved cited work

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:11:24.864289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:096361d128e2eb7d9e9ac3b39c7ab78da56c322330c2fe41aa53288844f74115

Observation c45d8bd8-8b73-4722-807b-d4b05e967cdb · outbound

This paper cites Wڵk䤦455e ­ZJVիu-_h `Y(.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Wڵk䤦455e ­ZJVիu-_h `Y(

Reference 79

Resolution
malformed identifier
raw_fallback, observed 2026-05-18T13:11:24.858426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:706e463c3a9890cfe83f8f76cee2326549b6a82347a38469d5bb7995a0dd714e

Pith citing papers

No inbound Pith citation observations are available.