Pith. sign in

Paper Citation Record · LEDGER

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast

As of 12 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2606.07356.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.07356 v1

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-27T20:52:26.812505Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

36 of 36 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved32
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3ab35fd4-7d5d-475e-99b8-f5b0f24b8420 · outbound

This paper cites Zero-Shot Unsupervised and Text-Based Audio Editing Using.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast Zero-Shot Unsupervised and Text-Based Audio Editing Using

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:1eba358a7e7df41ebeed78b1665134c148e8599972e64aa95527490c4bebe473

Observation 7f5379f5-db75-4284-a191-6b6971f90e3c · outbound

This paper cites an unresolved cited work.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:b429c5d6f90655f41513c81a95eb32bab492fa7c353a90a5cba1d009a9a8a2eb

Observation 9aa1dcf4-0b9f-43e4-b0b5-3a12f293dac1 · outbound

This paper cites Separate anything you describe , volume =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast Separate anything you describe , volume =

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:f87877743c099afda3ebf6de664bca28ee41009ba7d36138085bce921b8944a5

Observation e9b0ed47-9d02-4887-8996-1953bccf8e12 · outbound

This paper cites Audioldm 2: Learning holistic audio generation with self-supervised pretraining , volume =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast Audioldm 2: Learning holistic audio generation with self-supervised pretraining , volume =

Reference 4

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:bcf28f79ec4e702807e0108a60bda9945470bdef9d5de665798d87cb05eebc48

Observation 72d8e8df-57ad-4cac-a2fa-a882d03bd652 · outbound

This paper cites Mandic and Wenwu Wang and Mark D.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast Mandic and Wenwu Wang and Mark D

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:ebdd3bc15cc384fcd68ad1e3cd4a08877cfc8f2e864d8d17a1fa249ead2a8f33

Observation 9a90a4ea-b23b-46da-93ad-1dde462ae46d · outbound

This paper cites an unresolved cited work.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:27309a259f0cee89ade58543e740398be6b68942a2ec0c03caf8d4b4b3ecbaa2

Observation c4b541b2-c7c9-41a3-978e-da68a9dbcead · outbound

This paper cites Audioeditor: A training-free diffusion-based audio editing framework , year =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast Audioeditor: A training-free diffusion-based audio editing framework , year =

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:56a7902295e481a8d8c6a92af3416c26a1d4ca4be20d69b4f7f2c4a9c5262090

Observation f637a2be-c8e5-4efd-a57d-bb60d619297b · outbound

This paper cites Tango 2: Aligning diffusion-based text-to-audio generations through direct preference optimization , year =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast Tango 2: Aligning diffusion-based text-to-audio generations through direct preference optimization , year =

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:554ca547c7cfc648fc09222f61fb51bf22a47d396e2b3a4c10f3853ab5583e92

Observation 4541ea1e-636c-404b-a178-925fd3552415 · outbound

This paper cites Auffusion: Leveraging the power of diffusion and large language models for text-to-audio generation , year =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast Auffusion: Leveraging the power of diffusion and large language models for text-to-audio generation , year =

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:225fe1d4f2b52388f425a4b40a5b1ef92d33bc716722d44f23c547ebb17a9ef4

Observation ad320ebc-3094-4c88-bc84-69a2e79a96e0 · outbound

This paper cites Flowedit: Inversion-free text-based editing using pre-trained flow models , year =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast Flowedit: Inversion-free text-based editing using pre-trained flow models , year =

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:4870ec0492e712ef623103609c62e3cdd8a5c8ac6c22d48717b45c718a4e1965

Observation 9a8ac4cd-8e13-49e7-b25d-943c7d5f001f · outbound

This paper cites A udio C aps: Generating captions for audios in the wild.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast A udio C aps: Generating captions for audios in the wild

Reference 11

Resolution
metadata mismatch
doi, observed 2026-06-27T22:51:22.371826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:e0caef545f0fdb40c79a4494b118506536a3b90b31d601578d74e347ff9f96c6

Observation ec7bc1ba-0f72-44d3-998c-fd4031357486 · outbound

This paper cites In: ICASSP 2023 - 2023 IEEE Inter- national Conference on Acoustics, Speech and Signal Processing (ICASSP), pp.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast In: ICASSP 2023 - 2023 IEEE Inter- national Conference on Acoustics, Speech and Signal Processing (ICASSP), pp

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-06-27T22:51:22.374364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:45d67eacf9ca9eeb829810487761ea92390ba6cc2059fdbbea74245d74bad494

Observation 7d7e6c6e-2d03-4947-96ff-8f23fd59f501 · outbound

This paper cites Make-An-Audio: Text-To-Audio Generation with Prompt-Enhanced Diffusion Models , url =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast Make-An-Audio: Text-To-Audio Generation with Prompt-Enhanced Diffusion Models , url =

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:4f9d6293bdbc42a816a37f3b9638708b14f4b64f2fd70a58b4e19facc6168a05

Observation 1efcab7c-9aa1-4d1f-b26c-da9dbd176ebc · outbound

This paper cites an unresolved cited work.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:01d25388155cb9872559080a3f6d344f26a5c441d4eacbf1be6f65281486d2a5

Observation 39a41f62-d53d-460e-8753-5b6eef917e41 · outbound

This paper cites InstructME: An Instruction Guided Music Edit Framework with Latent Diffusion Models , url =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast InstructME: An Instruction Guided Music Edit Framework with Latent Diffusion Models , url =

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:9927f0e55931ff8861652344eb25c1f12d5197a46c1cce83752af115e6e482cb

Observation f07a119d-7444-49c1-97cd-afbaaf15005a · outbound

This paper cites InstructSpeech: Following Speech Editing Instructions via Large Language Models , url =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast InstructSpeech: Following Speech Editing Instructions via Large Language Models , url =

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:5ffeb85555df408fb6a0c4ed7a50a5c2af5e6ba702d16aadb482278d983fcde2

Observation 717df371-86a6-4911-a189-a1ae7cc73b3f · outbound

This paper cites Steermusic: Enhanced musical consistency for zero-shot text-guided and personalized music editing , volume =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast Steermusic: Enhanced musical consistency for zero-shot text-guided and personalized music editing , volume =

Reference 17

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:966c4d3095f6831c28892916fdff007b29119741c35310bef376f174b1321c9f

Observation 152fee70-6e59-4ff0-ac0f-b0c7e53530c1 · outbound

This paper cites Scaling Rectified Flow Transformers for High-Resolution Image Synthesis , url =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast Scaling Rectified Flow Transformers for High-Resolution Image Synthesis , url =

Reference 18

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:f390971d414336fdaf36b3a72062ccc19dd15d70fbac33a50f66652145d01c42

Observation 5a89c7c3-ad0b-467e-bc50-c966f09bdbc9 · outbound

This paper cites Classifier-Free Diffusion Guidance , url =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast Classifier-Free Diffusion Guidance , url =

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:0a91cd36fd716369832d58477541471b5fef1745863891c60b7fcc71806af674

Observation 37615aeb-08af-46ce-bfee-d626136f9928 · outbound

This paper cites an unresolved cited work.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:79d94918caa7ff91de04d1e4114606f7b299bc56fac426e9ccad8c87d79d3e83

Observation e9966058-2c1a-4af0-aca4-03e6418fdda5 · outbound

This paper cites Goodfellow and Wojciech Zaremba and Vicki Cheung and Alec Radford and Xi Chen , bibsource =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast Goodfellow and Wojciech Zaremba and Vicki Cheung and Alec Radford and Xi Chen , bibsource =

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:d77bd8b7607e51196ac9356f3ebc34ed8c570c20b2fa7b21828a2190b871fb74

Observation 9b27da6e-16cb-49b5-9c5b-e82c2c1a5bd2 · outbound

This paper cites Image quality assessment: from error visibility to structural similarity , volume =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast Image quality assessment: from error visibility to structural similarity , volume =

Reference 22

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:d720446d487d2e883d25ec92eaa20f0aab20ff45de3bbea20e3780c430bd2892

Observation 0a784be8-943d-4144-9749-de86d3fe5e71 · outbound

This paper cites Comparing individual means in the analysis of variance , year =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast Comparing individual means in the analysis of variance , year =

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:92516d208346b1b51e31be8a73dc856360e2fd3de7d63df41fa1a87e16bb1f22

Observation 7b981bf0-9aba-442f-915c-24fabbd9c2d0 · outbound

This paper cites DeepSeek-V3 Technical Report , url =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast DeepSeek-V3 Technical Report , url =

Reference 24

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:cb79f8cd61464c6d317c648e567e03c56a4abf9b288b8dc43245a5b49d30f1b5

Observation 1a7fa51c-5978-4f8a-92d0-cd6d23458c75 · outbound

This paper cites AudioMorphix: Training-free audio editing with diffusion probabilistic models , url =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast AudioMorphix: Training-free audio editing with diffusion probabilistic models , url =

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:1fbe14d57f5686a55fe3672601640de16bcbde6f882f90106924f72b4a8b9796

Observation c8a24ac6-af8a-4bb7-99dc-36c635fec41e · outbound

This paper cites SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations , url =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations , url =

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:d4cb67814ac872efb1d3582cd6ec07c738ea53e63a8fdf63812d6b2bfc4d86c3

Observation d7dbb0c2-1d8c-4f9e-a302-f2fa7ac3cd74 · outbound

This paper cites Emogen: Emotional image content generation with text-to-image diffusion models.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast Emogen: Emotional image content generation with text-to-image diffusion models

Reference 27

Resolution
metadata mismatch
arxiv_id, observed 2026-06-27T22:51:22.379612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:8b668c4d4209440d581f6afb19fea41f5c9bb438b180421300f2fa38fe924b5d

Observation af8edf29-5b0e-4570-9735-16d81fcebe9f · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-06-27T22:51:22.376959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:0e0bbf319162aacf4b010f30067f3336d5c3dafd260f9c0b18e23e2f7adc8b9a

Observation 3364943a-27c5-4402-b243-035296bca026 · outbound

This paper cites SteerFlow: Steering Rectified Flows for Faithful Inversion-Based Image Editing , url =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast SteerFlow: Steering Rectified Flows for Faithful Inversion-Based Image Editing , url =

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:991eb23c0386b63c6022548fafa61847dfd2d8107e5ff882ba54dbe0c7118daf

Observation 861fa5ea-707b-4c01-b98b-e9b2ac67077d · outbound

This paper cites Flowalign: Trajectory-regularized, inversion-free flow-based image editing , url =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast Flowalign: Trajectory-regularized, inversion-free flow-based image editing , url =

Reference 30

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:d1e252e56d7845ee939f644cad5f14b36f6ae33ccb7f3cc1287a0078d72a3a8a

Observation a1235c27-4731-4500-a9ae-33d02197bce0 · outbound

This paper cites Advances in text-guided 3D editing: a survey , volume =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast Advances in text-guided 3D editing: a survey , volume =

Reference 31

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:5846ceae3fdeedade52e06368650a0f000bd7ce443911de0ad4aa76a41232c46

Observation c05f1be6-d320-4ea4-8433-30167c8914af · outbound

This paper cites A survey of multimodal-guided image editing with text-to-image diffusion models , url =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast A survey of multimodal-guided image editing with text-to-image diffusion models , url =

Reference 32

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:ef8e0d2587ee1222e8d67d6de4ebb02715928dac006c7385951839498f2f4b12

Observation ff9b7ce7-1785-4c0e-9e77-ada0019440c3 · outbound

This paper cites Diffusion model-based image editing: A survey , year =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast Diffusion model-based image editing: A survey , year =

Reference 33

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:b9e3c0c3f8148bb2189766dde590ad1ac35b6ce6d8b5a1772183d1a176584422

Observation b7e5de5f-ab52-4191-a4d0-6c7964866d21 · outbound

This paper cites Guiding audio editing with audio language model , url =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast Guiding audio editing with audio language model , url =

Reference 34

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:f369656c9ef43b98e8b75140c32d7f3dc4d469e08b4831773e93769dbe172c38

Observation 2e59e7de-49ef-42b6-b8b5-116f3f755bb6 · outbound

This paper cites WavCraft: Audio editing and generation with large language models , url =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast WavCraft: Audio editing and generation with large language models , url =

Reference 35

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:e1cbf23a798760fb51f1cdb220777d97e802a46a202798ec32b2c2a696494c3f

Observation 8dcb46b4-6d0a-4b3e-b1b1-2fd361392ac7 · outbound

This paper cites Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow , url =.

DirectAudioEdit: Inversion-Free Text-Guided Audio Editing via Diffusion Prediction Contrast Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow , url =

Reference 36

Resolution
unresolved
no resolver link, observed 2026-06-27T20:52:26.812505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T20:52:26.812505Z digest=sha256:4ba12d244dad4721a32d6fcec5d6e44d704ce7af7d9e9aa67991fe6296a41b02

Pith citing papers

No inbound Pith citation observations are available.