Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:04:38.641568Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 0 inbound Pith citation observations for arXiv:2506.08887.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:04:38.641568Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
63 of 63 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 69266413-019c-46c6-9327-ef3c60df9862 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Localizing mo- ments in video with natural language
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8819666a-0375-46b3-b42a-42f9bb34c443 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Vqa: Visual question answering
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4628a9ca-b838-4b54-b7ce-56bd9964f1c3 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Frozen in time: A joint video and image encoder for end-to-end retrieval
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0b308772-8625-4b7a-b330-03ca9c3b1748 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Cross modal retrieval with querybank normalisation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dc198174-2b83-4f14-a585-9fe2ece5a579 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval RAP: Efficient text-video retrieval with sparse-and- correlated adapter
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b5ee4d76-0774-4ee6-9b80-77e60927dd13 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Adaptformer: Adapt- ing vision transformers for scalable visual recognition.Ad- vances in Neural Information Processing Systems, 2022
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 23c14efb-88db-425d-b8a0-770b3f7f91d7 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Vast: A vision-audio-subtitle-text omni-modality foundation model and dataset.Advances in Neural Information Processing Sys- tems, 2023
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 32d8a49a-cd82-4d9a-ac3e-fd9236325f46 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Improving Video-Text Retrieval by Multi-Stream Corpus Alignment and Dual Softmax Loss
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f48e58cb-3cbe-4924-ba15-f4b2fe443d8c · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Prompt switch: Efficient clip adaptation for text-video re- trieval
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 97dfb5ab-823f-42b0-a475-1ea517437e7a · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Acnet: Strengthening the kernel skeletons for powerful cnn via asymmetric convolution blocks
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 41a91c59-7797-47da-890b-a138a5bbbdcf · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Repvgg: Making vgg-style convnets great again
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1a91173e-b2c3-4f50-a64c-20c3fbbf3af1 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Improving clip training with language rewrites.Advances in Neural Information Processing Sys- tems, 2024
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 32c04091-3f0f-4b2c-8cd5-ef0c393c6c98 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Multi-modal transformer for video retrieval
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9f8a736e-eac1-4dc4-acbb-27e3224e3cfe · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval X-pool: Cross-modal language-video attention for text- video retrieval
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 093d98e5-9f5f-4bc0-af90-1ad33cb684ec · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Accurate, Large Minibatch SGD: Training ImageNet in 1 Hour
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55a5bf61-fbbc-45a9-8b9b-848a786f3dde · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Framewise phoneme classification with bidirectional lstm and other neural net- work architectures.Neural Networks, 2005
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 57571e0b-2ffa-47b9-996f-bb3704489bd4 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Towards a unified view of parameter-efficient transfer learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63af92d2-9708-40f8-bd12-91613cb213f3 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Secret: Self-consistent pseudo label refinement for unsupervised domain adaptive person re-identification
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3e4554c0-31bb-48be-b182-ab87fffcf3f1 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Gaussian Error Linear Units (GELUs)
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82e4b745-76e2-4cf8-b463-6c323e630f29 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Parameter-efficient transfer learning for nlp
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6625c00c-b459-4507-a833-94e918b277d1 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval LoRA: Low-Rank Adaptation of Large Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0329a248-4f29-4b20-8274-a3a2a8f55253 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval V op: Text-video co- operative prompt tuning for cross-modal retrieval
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1c1c649c-d85c-4583-90dc-8678cbbb7fbd · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Vi- sual prompt tuning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e48bacc2-7389-4021-ac23-3c6a801563a9 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Video- text as game players: Hierarchical banzhaf interaction for cross-modal representation learning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c79fb09d-18d4-46c3-8db5-a85b3247de15 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Mv-adapter: Multimodal video transfer learning for video text retrieval
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 903f48d1-5eee-4799-b8b3-5b30427eb340 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Deep visual-semantic align- ments for generating image descriptions
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 66598c33-e1f8-43e7-8b1e-b96415f926ba · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Maple: Multi-modal prompt learning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e1b60c54-ab64-4a65-b8c8-ddbdc45a349a · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Self-regulating prompts: Foundational model adaptation without forgetting
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation be6e81ff-b0a4-4c84-96dd-753df8724479 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Dense-captioning events in videos
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 37d14fd7-c852-4144-b417-c8a86a2965dd · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Courier Corporation, 1997
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a4cfc7b-c815-4033-bf1b-934abdf1bb8f · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Less is more: Clipbert for video-and-language learning via sparse sampling
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 92789e07-17c7-4199-bb2a-94da68badea5 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Align before fuse: Vision and language representation learn- ing with momentum distillation.Advances in Neural Infor- mation Processing Systems, 2021
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6669b5e9-2724-497d-9548-25bcd1d9e577 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 472df5ca-6f3a-4ae8-8f94-535037418fd3 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Unmasked teacher: Towards training-efficient video foundation models
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2470c477-442b-4551-8d82-6811e1f69e24 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Llava-next: Im- proved reasoning, ocr, and world knowledge, 2024
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c4330bd7-a4c8-4b56-b714-198727c40189 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Sgdr: Stochastic gradient descent with warm restarts
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e6db3627-e858-42d0-b83d-3907c87c4145 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Clip4clip: An empirical study of clip for end to end video clip retrieval and captioning.Neu- rocomputing, 2022
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d1a8c765-90e9-45e0-8f5a-0a79c9e6f4c4 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Ea-vtr: Event-aware video-text retrieval
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 675359d4-1a39-4b9f-b23d-8ee148482d73 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Howto100m: Learning a text-video embedding by watching hundred million narrated video clips
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cb69a46c-7eb6-4721-ab43-ac139614fd12 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Representation Learning with Contrastive Predictive Coding
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4603b17-f9ba-4280-bf00-2e3f29eb956a · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Language models are unsu- pervised multitask learners.OpenAI blog, 2019
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5a8f1e4-476f-4dbb-aadf-fa421504ecd9 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Learn- ing transferable visual models from natural language super- vision
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3ddc66e8-4b1c-40f1-a540-51cc0dec58eb · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval The long-short story of movie description
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 20943e42-6839-4b62-ad48-600cbe32f003 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval TempMe: Video Temporal Token Merging for Efficient Text-Video Retrieval
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5050387a-a682-48d0-87e4-b2319976e873 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval X-reid: Cross-instance transformer for identity-level person re- identification
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a487edaf-4706-4cd6-8dc1-127138457d33 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Zerocap: Zero-shot image-to-text generation for visual- semantic arithmetic
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f644fda7-d10d-4b96-8176-32ee03ac7e91 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Yolov10: Real-time end-to-end object de- tection.Advances in Neural Information Processing Systems,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aaef0ede-80b5-4deb-9093-06115b7bdf9d · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Text is mass: Modeling as stochastic embedding for text-video retrieval
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 318d79f3-c194-4bb6-a2ea-277328ae9909 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Disentangled Representation Learning for Text-Video Retrieval
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 300d6d22-2ed4-4a14-89bc-5e3f9fe69607 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval InternVideo: General Video Foundation Models via Generative and Discriminative Learning
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13d9f64f-a526-4fb3-8666-890cc69a6917 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Cap4video: What can auxiliary captions do for text-video retrieval? InProceedings of the IEEE Confer- ence on Computer Vision and Pattern Recognition, 2023
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d3bb257c-fec5-4d6a-b455-d1824b8df63d · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Demystifying clip data
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 62a76fcd-6b96-400d-a602-ea4d4ecd08fa · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Msr-vtt: A large video description dataset for bridging video and language
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a2d4fc0f-88d0-4624-877e-0ea3b83a0e31 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Show, attend and tell: Neural image caption gen- eration with visual attention
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ac8193da-961f-4068-b1fb-a10e92076d7c · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Clip-vip: Adapting pre-trained image-text model to video-language alignment
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4429bc67-e6b9-413d-8cfc-0eeb2161aa64 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval LLMI3D: MLLM-based 3D Perception from a Single 2D Image
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c7252e5e-663e-4109-b549-6fbcafb3c3dd · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval HEIE: MLLM-Based Hierarchical Explainable AIGC Image Implausibility Evaluator
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbdea981-e643-4f0b-9a62-f98e4d62c816 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Dgl: Dynamic global-local prompt tuning for text-video re- trieval.Proceedings of the AAAI Conference on Artificial Intelligence, 2024
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e338126f-3b20-4a2c-813f-89a4f1771003 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Cross-modal and hierarchical modeling of video and text
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d77fac82-44a4-4399-8190-8b237b154d11 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Neural Prompt Search
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e438381-88ec-420d-8564-76c6862d6121 · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Conditional prompt learning for vision-language mod- els
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 60c6db4a-d10a-4ccb-b7c0-3814677297bb · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval Learning to prompt for vision-language models.Inter- national Journal of Computer Vision, 2022
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 207f0633-d482-4615-9278-a64b4d1f201a · outbound
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval 14,αandβare set to0.3and1.0, respectively
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.