Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T15:48:33.963779Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 1 inbound Pith citation observation for arXiv:2507.14935.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T15:48:33.963779Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T21:53:28.021948Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T21:53:34.084878Z
65 of 65 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9a653eda-7ed8-461e-9c05-254125e640ed · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Robust cross-modal representation learning with progressive self- distillation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation da722cd7-14fe-4fba-9270-de4615f6ea92 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation On the effectiveness of image rotation for open set domain adaptation
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f27956e9-7a49-4e87-ac8a-f86e3d175bdf · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Domain generalization by solving jigsaw puzzles
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation dee88b25-f9c6-4d64-b7b1-a9ce145e3d7c · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Collecting highly paral- lel data for paraphrase evaluation
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 396e61a4-d780-417d-a576-1fbeb402b0c1 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Vggsound: A large-scale audio-visual dataset
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 40a90f4a-be71-4554-a7b7-24b51ec6dc28 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Hts-at: A hierarchi- cal token-semantic audio transformer for sound classifica- tion and detection
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9eacc5aa-1288-437f-a5f3-2e665af6c317 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation VALOR: Vision-Audio-Language Omni-Perception Pretraining Model and Dataset
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4528c55e-2a93-4781-ae74-7c72c085ea0e · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Uniter: Universal image-text representation learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7406c049-2539-466b-aa97-8122938036f1 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Sinkd: Sinkhorn distance minimization for knowledge distillation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d694a12a-1115-452f-949a-aca3c1e76ad3 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Sinkhorn distance minimization for knowledge distilla- tion
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1b229b04-2c2d-4e19-a83e-327b027cdebf · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Optical: Leveraging optimal trans- port for contribution allocation in dataset distillation
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2f994dd1-b7c2-4e6d-8938-a2e94fdcde70 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Layoutenc: Leveraging enhanced layout rep- resentations for transformer-based complex scene synthesis
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a3cf246e-ee9e-41e9-9ee8-07a0e685898f · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Streetsurfgs: Scalable ur- ban street surface reconstruction with planar-based gaussian splatting
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation af806a17-5526-4171-8036-6b1e99c5917e · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Towards Multimodal Open-Set Domain Generalization and Adaptation through Self-supervision
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 33e87a74-1eb6-464c-b70a-86a9749d9f10 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Simmmdg: A simple and effective framework for multi-modal domain generalization
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 507023a4-0327-44b5-b8ea-a620ee805697 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Clotho: An audio captioning dataset
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 09215f9d-9f71-494e-b01f-299865375760 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Multi-modal align- ment using representation codebook
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 51531b0d-6a4b-410d-be22-2e8426a6893c · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Ace: A generative cross-modal retrieval framework with coarse-to-fine semantic modeling
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2f814264-7211-4be5-a0d5-027e32162906 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Slowfast networks for video recognition
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd407d55-5386-4d64-9c20-ccaf26e2e10f · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Domain-adversarial training of neural networks
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 4c64e29f-8989-4401-991e-dc5f1e2c3740 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Imagebind: One embedding space to bind them all
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5693eb52-ffb9-4a0e-a0b6-c3fd604f14b2 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Enhancing Multimodal Unified Representations for Cross Modal Generalization
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 302450ac-17aa-42a6-8c26-e1ed2f82c0bd · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Semantic residual for multimodal unified discrete representation
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation bdc993b6-96dd-4fdc-b61a-fd7b27d75e34 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0f93d54-2536-44a3-91c0-881c016cfe4f · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation WavTokenizer: an Efficient Acoustic Discrete Codec Tokenizer for Audio Language Modeling
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 807fa901-b252-4444-8f09-06bbc1d8feb9 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Learning to generalize: Meta-learning for do- main generalization
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e4e07c83-b58c-4ac9-9336-480b1ebb98c7 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Domain generalization for med- ical imaging classification with linear-dependency regular- ization
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 882ef81a-0a97-43bc-8971-19b5b257d701 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Adjustment and alignment for unbiased open set domain adaptation
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b9e429f1-b3ba-46b4-82a2-594c73c93eff · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Cross-Modal Discrete Representation Learning
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c7a639d-2e41-4ca0-b4a7-ed46348417e7 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Feddg: Federated domain generalization on medical image segmentation via episodic learning in continuous fre- quency space
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a53e4ee0-723a-424f-9877-952fddd56894 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Swin transformer: Hierarchical vision transformer using shifted windows
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d48944d-23b5-4973-ace7-88db099cecdf · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Unified-IO: A Unified Model for Vision, Language, and Multi-Modal Tasks
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 604260d1-9db9-4e96-9617-9735f183c9ad · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Unsupervised learning of visual representations by solving jigsaw puzzles
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a19c92f6-8dc0-4a89-b56e-f1530b4f90c6 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Two at once: Enhancing learning and generalization capacities via ibn-net
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0a5add89-0bdf-4104-b171-e11c1a236067 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Estimating Visual Information From Audio Through Manifold Learning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b771fb0b-8b4f-4812-8c64-7d035d63b104 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Audio-visual speech recognition with a hybrid ctc/attention architecture
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0a61c64f-76dc-42b1-930c-c750e10ea19f · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Domain generalization through audio- visual relative norm alignment in first person action recog- nition
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation eb751c58-b667-4e10-8fff-866a3022576c · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Learning transferable visual models from natural language supervi- sion
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f742929-52d9-43e4-a3c9-5c0dc224b5e2 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Mask2anomaly: Mask transformer for uni- versal open-set segmentation
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 62bdf01a-5e88-4d31-b262-8bf4980429fa · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation XKD: Cross-modal Knowledge Distillation with Domain Alignment for Video Representation Learning
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 24c727c0-79b1-4434-95d5-26697d9abdf6 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Open domain generalization with domain- augmented meta-learning
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7281a904-b633-496b-bf6b-43836baae6e1 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef32786e-a737-49b6-8428-03a2dcfd6dff · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Audio-visual event localization in unconstrained videos
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9651d3e8-bc5d-4b0a-bc27-b1d713ae757f · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Unified mul- tisensory perception: Weakly-supervised audio-visual video parsing
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5de95d9f-85c0-426e-9575-b89836e6821f · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Domain randomization for transferring deep neural networks from simulation to the real world
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ec0d5cd5-d07d-4e68-a099-f64740d5f903 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Deep Domain Confusion: Maximizing for Domain Invariance
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f23e8e0-ffe8-4fd5-b1ec-ea7f4cc4fc3e · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation IRBridge: Solving Image Restoration Bridge with Pre-trained Generative Diffusion Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02e59edd-c139-4d7d-a293-980bf430387a · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Towards Transformer-Based Aligned Generation with Self-Coherence Guidance
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 449e3041-9ccf-43fd-b742-3b06a1d220cb · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Vlmixer: Unpaired vision-language pre-training via cross-modal cutmix
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 4d17012a-086b-478e-98a8-7edd453a268c · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation General- izable decision boundaries: Dualistic meta-learning for open set domain generalization
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 820ec6f5-c22a-461a-9ce0-6d02045a981c · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Achiev- ing cross modal generalization with multimodal unified rep- resentation
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 19e3608f-e6cf-420e-aaaa-a654c73fda0a · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Hyper- spectral image classification based on unsupervised hetero- geneous domain adaptation cyclegan
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5d2da7bd-d3ac-45e0-a4d3-cf220f061a64 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Class semantics modulation for open-set in- stance segmentation
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0c7c5007-6929-411b-8fac-ce29d6d3443e · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Learn- ing domain-invariant and discriminative features for homo- geneous unsupervised domain adaptation
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c55857b4-5572-411e-8a7d-b081986d96b4 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation A du- ality based approach for realtime tv-l 1 optical flow
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7f10b910-2872-4c0b-8b06-b8d5b1dc0448 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation mixup: Beyond empirical risk management
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 78b6b8c4-a9aa-4cd1-91a4-8751551b2670 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Extending multi-modal contrastive rep- resentations
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 92ae4c94-4bfa-47f6-af31-dc2573259f2d · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Towards effective multi-modal interchanges in zero-resource sounding object localization
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 55de53c6-1ece-419d-9a92-e46d4b6bf367 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Positive sample propagation along the audio- visual event line
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 97ddbd13-ffc0-4fa6-b81a-5493df3479fa · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Contrastive pos- itive sample propagation along the audio-visual event line
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9e964db3-0843-4c9d-8e06-6d5cebaacee9 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation As shown in Figure 4, applying the same mask to paired multimodal samples helps improve model perfor- mance
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ea7d8129-f526-46f0-a843-fcf27af389a9 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation As shown in Figure 5, we experimented with five different settings: 256, 400, 512, 800, and 1024
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation de07682c-87f8-4353-a6c1-e34fea5fdacb · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation Lcoarse serves as the foundation of the model, while Lf ine and Lcujp further refine the unified representation space and enhance the model’s open-domain detection capabilities
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0ed9da58-6fc8-4401-ab6b-a4ca137ccec8 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation CUJP8, despite having more split block reorder- ing, optimizes memory usage and reduces training time compared to MMJP6 [14]
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 4d271cca-30a9-434f-9f9a-a9835f3a91a5 · outbound
Open-set Cross Modal Generalization via Multimodal Unified Representation The visualization maps audio-video-text triplets from the Valor32K dataset [7] into the unified rep- resentation space (codebook)
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation aa53c0d5-f2ba-4c21-9783-00fd73b7fa6f · inbound
TAP: Parameter-efficient Task-Aware Prompting for Adverse Weather Removal Open-set Cross Modal Generalization via Multimodal Unified Representation
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.