Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-27T21:02:04.021977Z
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 1 inbound Pith citation observation for arXiv:2606.07229.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-27T21:02:04.021977Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-31T15:38:32.786994Z
A source-named dated measurement, never combined with another source.
Source: cited_works
31 of 31 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 24629961-3d15-4f16-9e62-a79a951a5d87 · outbound
MMAE: A Massive Multitask Audio Editing Benchmark Nano Banana 2: Google’s latest AI image generation model., 2026
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4944dca6-00f5-4d3b-b708-a98ec6736ecb · outbound
MMAE: A Massive Multitask Audio Editing Benchmark Gemini Omni: Native multimodal generation and video model., 2026
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 908bdb12-5b54-4c8f-b202-c6134f2858ef · outbound
MMAE: A Massive Multitask Audio Editing Benchmark Guiding audio editing with audio language model
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1448b5ec-527f-4ecb-9861-eb8fbeb13edf · outbound
MMAE: A Massive Multitask Audio Editing Benchmark Mmedit: A unified framework for multi-type audio editing via audio language model,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 511ea7c9-f983-4344-8af4-035b527c6965 · outbound
MMAE: A Massive Multitask Audio Editing Benchmark Ming-uniaudio: Speech llm for joint understanding, generation and editing with unified representation
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7c26e1b0-4d6a-45da-9476-e7f08e184aad · outbound
MMAE: A Massive Multitask Audio Editing Benchmark Step-audio-editx technical report
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2736e59b-65a2-4b86-ba67-73e3eb2c443f · outbound
MMAE: A Massive Multitask Audio Editing Benchmark Audiochat: Unified audio sto- rytelling, editing, and understanding with transfusion forcing,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 65371372-0a0b-4c25-9c73-8416bba30717 · outbound
MMAE: A Massive Multitask Audio Editing Benchmark Audio-Omni: Extending Multi-modal Understanding to Versatile Audio Generation and Editing
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46f2db77-7b18-4051-ac78-0da88f7d4c6f · outbound
MMAE: A Massive Multitask Audio Editing Benchmark Voicecraft: Zero-shot speech editing and text-to-speech in the wild
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a53f7f4-c286-4d21-b417-d03c4f5eaede · outbound
MMAE: A Massive Multitask Audio Editing Benchmark Rubrics as Rewards: Reinforcement Learning Beyond Verifiable Domains
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b6e32a0a-eff3-41b7-b4f3-d7f8e6631ce6 · outbound
MMAE: A Massive Multitask Audio Editing Benchmark The interspeech 2026 audio reasoning chal- lenge: Evaluating reasoning process quality for audio reasoning models and agents
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e5ceee80-3aab-4981-815e-e39da6559633 · outbound
MMAE: A Massive Multitask Audio Editing Benchmark Edit-Compass & EditReward-Compass: A Unified Benchmark for Image Editing and Reward Modeling
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5c5ab581-eeb1-4024-a142-e0af8a4f9780 · outbound
MMAE: A Massive Multitask Audio Editing Benchmark Fluentspeech: Stutter- oriented automatic speech editing with context-aware diffusion models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 173652ec-d260-4c50-bdd1-f6c21dbb0334 · outbound
MMAE: A Massive Multitask Audio Editing Benchmark Ssr-speech: T owards stable, safe and robust zero-shot text-based speech editing and synthesis
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e49812c5-c163-4796-881b-ab1cb714c36c · outbound
MMAE: A Massive Multitask Audio Editing Benchmark Recomposer: Event-roll-guided generative audio editing
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f94d8411-ddf2-48ed-9a7c-0d494c25f213 · outbound
MMAE: A Massive Multitask Audio Editing Benchmark Prompt-guided precise audio editing with diffusion models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71cb0130-9758-44af-b716-9be88c10ede9 · outbound
MMAE: A Massive Multitask Audio Editing Benchmark Zero-shot unsupervised and text-based audio editing using DDPM inversion
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a896a11-11aa-4042-abf7-e1de48272e99 · outbound
MMAE: A Massive Multitask Audio Editing Benchmark Instructspeech: Following speech editing instructions via large language models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c41a948-3bf0-4c50-abe6-82abf5d5b579 · outbound
MMAE: A Massive Multitask Audio Editing Benchmark WavCraft: Audio Editing and Generation with Large Language Models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 242ed538-66b8-4ad5-a0ed-379822119b49 · outbound
MMAE: A Massive Multitask Audio Editing Benchmark Audit: Audio editing by following instructions with latent diffusion models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c2be435-a2a6-4130-a2de-37583994c917 · outbound
MMAE: A Massive Multitask Audio Editing Benchmark Audioeditor: A training-free diffusion-based audio editing framework
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e3d45a2-4f8b-4179-b483-6f334cfb80f6 · outbound
MMAE: A Massive Multitask Audio Editing Benchmark AudioMorphix: Training-free audio editing with diffusion probabilistic models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 004d4398-a5b1-46a4-b97a-a9d5c99bd38a · outbound
MMAE: A Massive Multitask Audio Editing Benchmark VoiceCraft-X: Unifying Multilingual, Voice-Cloning Speech Synthesis and Speech Editing
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d311cd7-530d-45b2-8829-db578420ddd0 · outbound
MMAE: A Massive Multitask Audio Editing Benchmark CosyEdit: Unlocking End-to-End Speech Editing Capability from Zero-Shot Text-to-Speech Models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0e8c4fdf-8e7f-4f00-9572-c79cf55e14c6 · outbound
MMAE: A Massive Multitask Audio Editing Benchmark InstructAV2AV: Instruction-Guided Audio-Video Joint Editing
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f3b64323-eb9d-4c46-a72a-2bf434cef0be · outbound
MMAE: A Massive Multitask Audio Editing Benchmark SpongeBob: Sync-Aware Harmonious Audio-Visual Generative Editing
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 45b72c30-3e4f-46ea-90ae-8345a5d86c71 · outbound
MMAE: A Massive Multitask Audio Editing Benchmark Flam: Frame-wise language-audio modeling
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c143df92-39d7-4394-85f2-565de928ec65 · outbound
MMAE: A Massive Multitask Audio Editing Benchmark Omni-captioner: Data pipeline, models, and benchmark for omni detailed perception
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69a7bd31-4b73-49a8-8da1-e88c97e3bc53 · outbound
MMAE: A Massive Multitask Audio Editing Benchmark Qwen3-Omni Technical Report
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e2085551-2634-4c86-a5c1-f91bad2e3bd9 · outbound
MMAE: A Massive Multitask Audio Editing Benchmark Qwen3.5-Omni Technical Report
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 880df646-a006-4570-8207-b0da9a2ff514 · outbound
MMAE: A Massive Multitask Audio Editing Benchmark id": "69e898163a050f39ac567501
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 554ce836-3db2-4da5-8e18-af27c14b8e05 · inbound
AgenticASR: Refining Speech Recognition in Real-World Scenarios via an Agentic Approach MMAE: A Massive Multitask Audio Editing Benchmark
Reference 2026
Source-reported events for the cited work
Unavailable: canonical work link unavailable.