Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-16T10:40:01.902254Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 81 of 81 outbound references and 0 inbound Pith citation observations for arXiv:2601.20597.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-16T10:40:01.902254Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
81 of 81 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 1116804f-5659-4cc0-9b45-1d9667496262 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4f82f5e7-7c65-4c5d-9e40-dee810632c0d · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Ashok, K
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6e0b33d6-b5fd-43b8-860c-b96508f0c9e3 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Frozen in time: A joint video and image encoder for end-to-end retrieval
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 37f15509-ad8c-4d42-87e0-b2800cb4290c · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Non-autoregressive cross- modal coherence modelling
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1982bc2e-ebc8-4b69-9931-0a75c947a475 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Activitynet: A large-scale video benchmark for human activity understanding
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation fdc776cf-03f6-418a-a7fd-e847250d8444 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Online fast adaptation and knowledge accumulation (osaka): A new approach to continual learning.Advances in Neural Information Processing Systems, 33:16532–16545
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation dcab8dd2-25f9-4a20-9dff-5d406f0fad9e · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Clumo: Cluster- based modality fusion prompt for continual learning in visual question answering.Journal of Artificial Intelligence Research, 83
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7681fd1e-64b4-4899-8472-6a0cfa999907 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Castro, Manuel J
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3b2f0bbb-149a-4ba9-8c8a-369115acf023 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Fine-grained video-text retrieval with hierarchical graph reasoning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation da9974f5-19a1-4cfb-9e61-c1b7a53ee94f · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Vision- sensor attention based continual multimodal egocentric ac- tivity recognition
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 169e921f-6c01-41e7-abb1-4d21d640f6a5 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Teachtext: Cross-modal generalized distillation for text-video retrieval
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8c686c0e-6338-41a2-9249-55373b401d8d · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Don't Stop Learning: Towards Continual Learning for the CLIP Model
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 227e82ff-e973-409c-b3cb-f3a67b75ef71 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Podnet: Pooled outputs distillation for small-tasks incremental learning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d136fbec-b734-4c12-aa50-905633e73a58 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval A feature- space multimodal data augmentation technique for text- video retrieval
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation cd9416df-fac6-48d9-91d0-cdce30f56e10 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Uatvr: Uncertainty-adaptive text-video retrieval
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c99db1a2-c20f-4648-a484-33184fd37811 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Transferring image-clip to video-text retrieval via temporal relations.IEEE Transactions on Multimedia, 25:7772–7785
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c0b5fd60-b1d3-4dee-b8fa-af287fae4eef · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Multi-modal transformer for video retrieval
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 52c14299-a180-40fa-93af-00ca30203a9a · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval X-pool: Cross-modal language-video attention for text- video retrieval
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ea1e9ff3-71d0-499b-b885-fedd0577eae9 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Dyson: Dynamic feature space self- organization for online task-free class incremental learning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 235bf544-8942-41bd-9e04-cc4ce4c8aa7b · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Learning a unified classifier incrementally via rebalancing
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 23e57aa6-d95f-4629-a46d-21e504b01840 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Curiosity-driven class- incremental learning via adaptive sample selection.IEEE Transactions on Circuits and Systems for Video Technology, 32(12):8660–8673
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e604ae7e-8b2e-4f53-a92d-f059bf6d3837 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Distilling causal effect of data in class-incremental learning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8480298d-d87b-4d0b-b8ea-1a0e5e4af4bf · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Neural collapse inspired federated learning with non-iid data
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4dfa2c9d-afc7-46a9-8162-c85451acfb4f · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8b9fcda3-507a-41ba-92c2-0993e30f8931 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Unresolved cited work
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ddb1b485-f960-4d80-b661-c6fd3649bfd3 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Hybrid-tower: Fine-grained pseudo-query interaction and generation for text-to-video retrieval
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ab85f800-dd7a-4da5-a188-4f29c50e368d · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Bakker, Nicu Sebe, and Michael S
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation afe372b3-fb8c-47ef-b4ee-d415d6312775 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Dynamic integration of task-specific adapters for class incremental learning
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a2dd7117-829c-4ad6-ba6f-a74700987012 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Multi-modal inductive framework for text-video retrieval
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3bcd1dc6-93b9-493d-8c62-a0ebf4ce32be · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Learning without forgetting
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation edebad7e-2da5-4a8d-9e83-d68f5976bc11 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Anchor assisted experience replay for online class- incremental learning.IEEE Transactions on Circuits and Systems for Video Technology, 33(5):2217–2232
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b3b04c33-f98c-4d7e-b06a-1c2d3b4ab5ff · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Pre-train, prompt, and predict: A systematic survey of prompting methods in natural language processing.ACM Computing Surveys, 55(9):1–35
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6f9179a4-001e-4107-967b-5669b4a120f1 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Use what you have: Video retrieval using representations from collaborative experts
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation eb99e289-a5c8-44d6-bc01-1f01e42901f9 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Adaptive aggregation networks for class-incremental learning
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation dd4eb987-5ff5-4cc8-8321-8b6609bd0d7d · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Ts2-net: Token shift and selection transformer for text-video retrieval
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a06aaf53-a0db-4a1f-8699-a4c9bf11e4e1 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Clip4clip: An empirical study of clip for end to end video clip retrieval and captioning
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b9c1695b-e4ae-4135-9907-1c55d730415a · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval X-clip: End-to-end multi-grained contrastive learning for video-text retrieval
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation cc713bd8-6722-4803-9dcb-d562db5fcfc9 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Packnet: Adding multiple tasks to a single network by iterative pruning
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 69ef110c-35c8-4964-b321-bec612227c23 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Prevalence of neural collapse during the terminal phase of deep learning training.Proceedings of the National Academy of Sciences, 117(40):24652–24663
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ca335b31-e5ac-4caf-8464-f3fcb3b3b882 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Prabhu, P
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 396540dc-c96c-4a1c-9ffd-0ff774ac9832 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Learning transferable visual models from natural language supervision
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c2124e43-fd82-4aff-86a1-65bc36a5367d · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval De Melo, Benjamin Van Durme, and Rama Chellappa
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation db425f6b-aa98-4063-b043-ce955592cef6 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Unresolved cited work
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 01acc3f3-6af0-4b52-971d-ed6d2bd58277 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Relation triplet construction for cross-modal text-to-video retrieval
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1e61f864-8c9c-446a-8fd5-297b26d80f5d · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Spatial-temporal graphs for cross-modal text2video retrieval
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1921058a-4e86-402b-8514-441de59d9813 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Learning endogenous attention for 12 incremental object detection
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a58a7f6e-660d-4dac-b416-4111b8a21b09 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Multimodal continual learning using online dictionary updating.IEEE Transactions on Cognitive and Developmental Systems, 13(1):171–178
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3fdd9fa0-b3d5-4eb6-b708-802f7c1f237c · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Unresolved cited work
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8f1985a2-8ea5-4538-ae4f-cee317aeadb9 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Topology-preserving class-incremental learning
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b7428cfb-c325-49d1-b8a4-7d36a9a4d9cf · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval new” while consolidating “known
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e983ac38-969c-421f-ace1-8e98a8a855b8 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Holistic features are almost sufficient for text-to- video retrieval
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c3ecd1c6-91d7-406b-82a4-28246964eb74 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Dualcp: Rehearsal- free domain-incremental learning via dual-level concept prototype
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2ef25a4f-1dcc-43b9-b5cd-c8d759a4a279 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Semantic knowledge guided class-incremental learning.IEEE Transactions on Circuits and Systems for Video Technology, 33(10):5921–5931
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5be729ff-f5be-4411-86c5-2238436c0c81 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Non-exemplar class-incremental learning via adaptive old class reconstruction
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b24ac00e-0264-4a54-9f7b-0aca9746794b · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval T2vlad: Global-local sequence alignment for text-video retrieval
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2a1936ad-5e4f-43f7-9ece-dedbc41a08e5 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Unresolved cited work
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 37d9e1c2-d717-4481-b2c1-fdb53fc4fe02 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Unified coarse-to-fine alignment for video-text retrieval
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 72e1b749-ee1a-482f-8266-452eb054661d · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Unresolved cited work
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f599089f-2aec-42db-b05b-dcf40d1976cc · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Unresolved cited work
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8575d487-c113-46d4-b50c-953a9fb3053d · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Striking a balance between stability and plasticity for class-incremental learning
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation bf5cbcf8-e89d-44a7-b88a-955fb23350a6 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Large scale incremental learning
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 76bf1e52-9479-4f74-ad57-51d64cfcb98e · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Msr-vtt: A large video description dataset for bridging video and language
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 323b25a9-ab67-4122-84d9-172ef450bf81 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Clip-vip: Adapting pre- trained image-text model to video-language alignment
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3a3e8bb0-5807-46a3-bcf9-5b05ae7d85a1 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Der: Dynamically expandable representation for class incremental learning
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 293d60e1-c208-42ac-b50e-fd430be27fa1 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Low-rank prompt interaction for continual vision-language retrieval
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e20f0040-64d1-4218-bc02-81aef12cc4ae · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Dynamic support network for few-shot class incremental learning.IEEE Transactions on Pattern Analysis and Machine Intelligence
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e9a3d12a-91c7-4c8e-ab7f-6b50dee6a5ed · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Taco: Token-aware cascade contrastive learning for video-text alignment
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e30ce547-9178-4dd1-8db6-b74f4d7e132e · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Recent advances of multimodal contin- ual learning: A comprehensive survey
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 254ee880-1289-427b-92b8-a25f65be6876 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Boosting continual learning of vision-language models via mixture-of-experts adapters
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 08d621ff-ad31-4363-b78a-363b35155a17 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval A joint sequence fusion model for video question answering and retrieval
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 30180881-0e62-4084-8fc1-e23f12b15cf6 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Quantifying and narrowing the unknown: Interactive text-to-video retrieval via uncertainty minimization
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b71c631a-5f11-4577-89f4-5fa686a4dc94 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Mpt: Multi-grained prompt tuning for 13 text-video retrieval
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 01de61a7-4be2-4976-a52d-d3be55779950 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Vqacl: A novel visual question answering continual learning setting
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ecf872ac-51ad-464d-8f45-838e029c0fbf · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Mgsvf: Multi-grained slow versus fast framework for few-shot class-incremental learning.IEEE Transactions on Pattern Analysis and Machine Intelligence, 46(3):1576–1588
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c018a695-a164-4cdd-a9e8-4969c9698d9b · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Centerclip: Token clustering for efficient text-video retrieval
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5282bee4-e656-420a-bff7-aebc91fd4343 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Continual text-to-video retrieval with frame fusion and task-aware routing
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 30e8075a-d419-4b5c-8dd1-9e0ea72e66f0 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Preventing zero-shot transfer degradation in continual learning of vision-language models
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0e8f7948-cea6-4260-ad51-2af72f224ec2 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Understanding imbalanced semantic segmentation through neural collapse
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation fc92ad12-59cd-4713-ab6a-5ddf29ecbce1 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Unresolved cited work
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f406e808-57cf-4780-88ec-ec31b8bf628e · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Unresolved cited work
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 68c220ba-d509-46e9-b739-cd48f2e79ba2 · outbound
StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval Complementarity- aware space learning for video-text retrieval.IEEE Transactions on Circuits and Systems for Video Technology, 33(8):4362–4374
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
No inbound Pith citation observations are available.