Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T12:32:40.579862Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 68 of 68 outbound references and 0 inbound Pith citation observations for arXiv:1908.07094.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T12:32:40.579862Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
68 of 68 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 35ad6b23-c40c-48e7-a66f-bda27a2c4e9d · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Deep speech 2: End-to-end speech recognition in english and mandarin
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation cbaadc98-e1c7-4c7d-928c-a793e9a2c4de · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Bottom-up and top-down attention for image captioning and visual question answering
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1509ed93-aed0-40c6-9850-3f3386e0f534 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Deep voice: Real-time neural text-to-speech
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 9ea4c8e1-891d-4262-988b-de226798bd6d · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck One-sided unsupervised do- main mapping
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 15406ef7-2d88-4131-9483-ecf74b271bdb · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Microsoft COCO Captions: Data Collection and Evaluation Server
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0918578d-903d-408a-9083-37f831003fcd · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Weiss, Kanishka Rao, Katya Gonina, Navdeep Jaitly, Bo Li, Jan Chorowski, and Michiel Bacchiani
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6d75fb6c-4c44-49e7-91e8-b631717e85e8 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Learning phrase representations using rnn encoder-decoder for statistical machine translation
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0dba30e2-25b4-461e-88c4-5c7022836145 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck StarGAN: Unified gener- ative adversarial networks for multi-domain image-to-image translation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0fbf77f8-2d26-4c74-84bb-23eca167d1c3 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ab4d28ca-1ed6-4844-88a4-38219b608731 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Unsupervised domain adaptation by backpropagation
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3f08e268-2703-4f73-8cf7-420252166d2a · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Auditory-visual inte- gration during multimodal object recognition in humans: a behavioral and electrophysiological study
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f545e8c4-795e-48d1-bde4-d3a0b34a6ae5 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Deep voice 2: Multi-speaker neural text-to-speech
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a463a38d-d7be-409b-9e47-46c3bc12fa0b · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Generative adversarial nets
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 55867ffc-d869-4696-b7f0-c23500615042 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Griffin and Jae Lim
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a9f517cf-adea-4265-83b5-cf604f0182e0 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Unresolved cited work
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a901f9a3-bdd6-4f8a-9474-2796c9971ffa · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Deep neural networks for acoustic modeling in speech recognition
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ec1e21c5-7697-4666-a19b-21c451dc9a46 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Huang, Z
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 4f98a6bf-714b-4324-b9c3-e4446dd196aa · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 33c9fec7-9ce1-4584-99ab-329778a5c0b0 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Recurrent fusion network for image captioning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6522e939-ca0e-4114-b621-468754b8b2c3 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Learning to discover cross-domain relations with generative adversarial networks
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3dbaa0dc-3b94-4030-8570-69154d0a83a1 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Adam: A method for stochastic optimization
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39a0947e-dbaf-4c35-85a7-b63450f74f29 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Auto-encoding varia- tional bayes
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d394991e-db9c-4945-9809-0f3aeb1ff483 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Shamma, Michael S
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0df4bacb-40aa-4a6b-841f-f526c8ad1899 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Letter-Based Speech Recognition with Gated ConvNets
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97f37b93-b26e-48fc-9836-b7d6f98cb6f2 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck DA-GAN: instance-level image
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a0dc435e-e9af-4f45-a0f4-04a77d8939ee · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Neural TTS styl- ization with adversarial and collaborative games
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6c2909de-9943-465a-b9a9-84e0aa189c2b · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Semstyle: Learning to generate stylised image captions using unaligned text
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e963026b-5cd3-49a3-9fd2-e22378dd0b3b · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Deep multi-scale video prediction beyond mean square error
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 246b088b-86e9-476a-83d3-30e418bd7a41 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Fitting new speakers based on a short untranscribed sample
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 183b6c45-042c-45ee-8f6d-f2a6a96a0ba2 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Librispeech: An ASR corpus based on public domain audio books
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 9cf086df-c52c-485a-9e60-b0c2df823191 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Attend to you: Personalized image captioning with context sequence memory networks
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f73743cb-664a-4c60-b154-e764c072554d · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Beyond sensory im- ages: Object-based representation in the human ventral path- way
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 2ebac211-0c85-4b2a-a2db-75e0951216f2 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Arik, Ajay Kannan, Sharan Narang, Jonathan Raiman, and John Miller
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 81148e27-f988-4720-b030-3d5de3180dc6 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Generative ad- versarial text to image synthesis
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 1e437235-6436-4bd5-921f-496d3aa71b31 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Faster r-cnn: Towards real-time object detection with region proposal networks
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation cc08e4ad-b405-4ca7-a066-d8804c3554de · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Natural TTS Synthesis by Conditioning WaveNet on Mel Spectrogram Predictions
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39f2086d-8a96-4a69-b83e-0abc87399f1f · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Char2wav: End-to-end speech synthesis
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 15590278-f90c-4c7a-b371-9e1ad76a8ddb · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Szegedy, Wei Liu, Yangqing Jia, P
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 7494df59-bec1-44c6-9b46-d2080ae75d06 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Unsupervised cross-domain image generation
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b60e1574-6217-40e3-ac05-4b017bf2481b · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck V oiceloop: V oice fitting and synthesis via a phonolog- ical loop
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d91550cd-0fe8-416f-953a-1b62f9056112 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck The information bottleneck method
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b5488199-086e-4b76-b792-c0391b49f077 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck WaveNet: A Generative Model for Raw Audio
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation beaf1d4a-2160-4e1e-99a8-cc27cecad1aa · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Attention is all you need
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c148cdd4-3077-4afe-9cab-f11e86832d9f · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Gomez, Lukasz Kaiser, and Illia Polosukhin
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 27f4cf11-4028-4c02-9d5b-7aadf127df87 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Unresolved cited work
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 2d9e1973-0cce-46d7-b994-8c91366a3021 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Show and tell: A neural image caption gen- erator
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0689af5f-ddea-4cec-bea7-52183f9196c5 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Tacotron: Towards end- to-end speech synthesis
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 7b0e0399-1fdc-4ee2-9582-09cbbd2aeea6 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Unresolved cited work
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f35ff7ab-c1c2-438b-8ad2-39c01ecbca5b · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Memory networks
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 68744106-0ce7-4a5a-9bdf-fece81af9708 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Courville, Ruslan Salakhutdinov, Richard S
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aeedcc49-4e5b-4512-ade5-3389ee9adb75 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Attngan: Fine- grained text to image generation with attentional generative adversarial networks
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3e2215f9-f6e5-4012-870b-d75b57324378 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Image captioning with semantic attention.CVPR, 2017
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a0e63788-9d66-41a1-a2ca-70f492b48c75 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Stackgan: Text to photo-realistic image synthesis with stacked genera- tive adversarial networks
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 60d31258-403f-42cf-9b6f-f52dcfab64e9 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Visual to sound: Generating natural sound for videos in the wild
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da4a2775-032b-41c5-93ea-9445644bcae2 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Im- proving end-to-end speech recognition with policy learning
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a3c5a992-6bd7-492b-b704-36bc6a53f641 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Unresolved cited work
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ec905cbd-4e7b-4139-94d4-c13339b689f3 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Efros, Oliver Wang, and Eli Shechtman
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 1a4ac184-6260-4cac-baf4-fb788c5abeac · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck We encour- age the readers to refer to Figure 3 and Figure 4 of our main paper when reading this section
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 5399a283-5229-4460-afb4-13fe1a6708d2 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Unresolved cited work
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 1cb707da-a64a-4932-b58d-57eddc55bec4 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Unresolved cited work
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 43abc4ef-1f57-4c11-bafb-174ade716b24 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Unresolved cited work
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 224ff9f3-0ebb-4843-8926-383cad96ea42 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Unresolved cited work
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation c04a12fe-06cb-46ad-b544-7e871c475c59 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Unresolved cited work
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation c79de29b-7e42-4cd9-a4e7-5369cc87ce5a · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Unresolved cited work
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e3f706e1-9dc3-4c1a-a3c8-f80baa5dbd8a · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Unresolved cited work
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e61b71e8-6b17-4bbf-81e9-5fab85ad1eb1 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck dog” and “zebra
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 33f8c0bb-e942-47d4-b291-eabd291d05b6 · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck bottleneck
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 7c2b0122-2aba-4a76-a0b5-7d5f60ae7a6a · outbound
Unpaired Image-to-Speech Synthesis with Multimodal Information Bottleneck Unresolved cited work
Reference 256
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
No inbound Pith citation observations are available.