Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2306.15687.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T11:43:40.731660Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T15:47:06.014261Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation f456c7f4-541e-47cd-856c-1b8ad3962f2c · inbound
Movie Gen: A Cast of Media Foundation Models Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation bd427a1c-693f-42af-b43f-14463a3e3cd5 · inbound
Speech Watermarking with Discrete Intermediate Representations Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b360166c-1f97-4adf-93ca-1ce8e1519132 · inbound
TangoFlux: Super Fast and Faithful Text to Audio Generation with Flow Matching and Clap-Ranked Preference Optimization Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 996e2d5b-fdd1-48d5-88b8-1a4ee0220a00 · inbound
Continuous Autoregressive Modeling with Stochastic Monotonic Alignment for Speech Synthesis Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 352967e5-f9f1-4b2a-8471-9f1a25b21d67 · inbound
OmniAudio: Generating Spatial Audio from 360-Degree Video Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c436648-c066-4e7d-bf43-27b29dadef12 · inbound
Improving Trajectory Stitching with Flow Models Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 228b822e-6e8d-4a88-ad92-ce19c9c4def3 · inbound
EZ-VC: Easy Zero-shot Any-to-Any Voice Conversion Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c44eb73f-b23e-479f-822f-8d99bf5d7fe1 · inbound
VoiceStar: Robust Zero-Shot Autoregressive TTS with Duration Control and Extrapolation Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f142280b-7379-477a-bfee-909d5bfdb87c · inbound
Unlocking Speech Instruction Data Potential with Query Rewriting Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ead94f66-2b8d-4b78-928d-1d85dd3b77a5 · inbound
Technical report: Impact of Duration Prediction on Speaker-specific TTS for Indian Languages Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a26f963b-c51b-4469-a6b0-76b0961ebf9d · inbound
Generative Model Unlearning: A Survey through Target Events, Unlearning Operators, and Evaluation Protocols Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale
Reference 119
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90acce7b-57d5-49f2-9b22-d93e9238a3c7 · inbound
Making Separation-First Multi-Stream Audio Watermarking Feasible via Joint Training Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9d4e9c5-ff4c-451c-acc1-e5a81b6d86e5 · inbound
F3-Tokenizer: Taming Audio Autoencoder Latents for Understanding and Generation Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 02bcab2f-d7b1-462a-b062-50ba2e73d25c · inbound
Optimal Self-Distillation for Rectified Flow via Linear Probing Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b386eac8-8089-4c0d-89f8-1bac72d4b507 · inbound
X-Translator: A Real-Time Multilingual Speaker-Aware Speech-to-Speech Translation System Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbde2eef-b8a9-41e2-9698-44f2f76c5e36 · inbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0556a7c5-451b-4c06-94bb-3e5db3357557 · inbound
Confucius4-TTS: Transcript-Free Cross-Lingual Zero-Shot TTS with a Learnable Speaker Encoder Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.