Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T12:55:41.858385Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2507.21353.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T12:55:41.858385Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
41 of 41 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9e6b3f22-96c8-4fdf-bd98-99d6283d1359 · outbound
Group Relative Augmentation for Data Efficient Action Detection Frozen in time: A joint video and image encoder for end-to-end retrieval
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 349e228c-2515-4f54-bef2-3dc0a1259e28 · outbound
Group Relative Augmentation for Data Efficient Action Detection Exploiting vlm localizability and semantics for open vocabulary action detection
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a4e19a9e-a116-40f2-8a43-dda1673bbca1 · outbound
Group Relative Augmentation for Data Efficient Action Detection Frozen feature augmentation for few-shot image classification
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1315f334-b352-4329-9662-6fdb5718d5f4 · outbound
Group Relative Augmentation for Data Efficient Action Detection Visualgpt: Data- efficient adaptation of pretrained language models for image captioning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fddb37e3-7c7c-4e32-9b25-f5f1054d910e · outbound
Group Relative Augmentation for Data Efficient Action Detection Adversarial Feature Augmentation and Normalization for Visual Recognition
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 254383df-acec-4f27-a082-f36dbfe6d5b6 · outbound
Group Relative Augmentation for Data Efficient Action Detection PerceptionLM: Open-Access Data and Models for Detailed Visual Understanding
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10f9b7f0-37fe-406c-ae7c-6055b09f2a18 · outbound
Group Relative Augmentation for Data Efficient Action Detection Autoaugment: Learning augmentation strategies from data
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a1ef18b4-e09c-4fbb-b061-84da64288753 · outbound
Group Relative Augmentation for Data Efficient Action Detection Clip-adapter: Better vision-language models with fea- ture adapters
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 94005041-74c4-4cf2-bfba-70c183cdab04 · outbound
Group Relative Augmentation for Data Efficient Action Detection Ava: A video dataset of spatio-temporally localized atomic visual actions
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f85fcdf5-329b-4cfb-a473-79972efaf6f6 · outbound
Group Relative Augmentation for Data Efficient Action Detection Lora: Low-rank adaptation of large language models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db1f0c87-1b3e-4238-9dd2-754a670352ac · outbound
Group Relative Augmentation for Data Efficient Action Detection Interaction-aware prompting for zero-shot spatio-temporal action detection
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4fd9715a-7dda-45ce-8a63-e2764a9c28a3 · outbound
Group Relative Augmentation for Data Efficient Action Detection Interaction-aware prompting for zero-shot spatio-temporal action detection
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6a36ac45-cfb3-456a-ac98-2a8c50b20fbd · outbound
Group Relative Augmentation for Data Efficient Action Detection Spatio-temporal context prompting for zero-shot action detection
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0b32c2b8-056c-47cf-ac57-d1f5024f7549 · outbound
Group Relative Augmentation for Data Efficient Action Detection Scaling up visual and vision-language representation learning with noisy text supervision
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bbc9bb6-f8f0-4f41-bc00-37147d0d2d25 · outbound
Group Relative Augmentation for Data Efficient Action Detection Region-aware pretraining for open- vocabulary object detection with vision transformers
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 69262b86-ce9e-44f5-8156-17915e2b7ea9 · outbound
Group Relative Augmentation for Data Efficient Action Detection On feature normalization and data augmentation
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 27a371b8-107f-4dbc-b820-05748e87c9ce · outbound
Group Relative Augmentation for Data Efficient Action Detection Blip: Bootstrapping language- image pre-training for unified vision-language understanding and generation
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 96be9944-1086-4730-a0c1-ae2ec5354e4f · outbound
Group Relative Augmentation for Data Efficient Action Detection Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f55c61e-cd19-4f48-a533-c1f51e3fd8ab · outbound
Group Relative Augmentation for Data Efficient Action Detection Learning Object-Language Alignments for Open-Vocabulary Object Detection
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af6a9cf9-bbff-42cd-8e17-e1b110e90cce · outbound
Group Relative Augmentation for Data Efficient Action Detection UniVL: A Unified Video and Language Pre-Training Model for Multimodal Understanding and Generation
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c71580c9-2a89-4a71-bd0b-32a618059631 · outbound
Group Relative Augmentation for Data Efficient Action Detection Moma: Multi-object multi-actor activity parsing
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f39fccbd-82bb-4898-96f7-bd52d3dc0bbf · outbound
Group Relative Augmentation for Data Efficient Action Detection Film: Visual reasoning with a general conditioning layer
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81659a12-2f1c-4cb7-98ec-2f0e983214af · outbound
Group Relative Augmentation for Data Efficient Action Detection Learning transferable visual models from natural language supervision
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 228b6ebe-8c05-44ea-9c66-b84aab72939c · outbound
Group Relative Augmentation for Data Efficient Action Detection Videomae: Masked autoen- coders are data-efficient learners for self-supervised video pre-training
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c5aa8eb1-7c74-400e-815e-eb4541b7dc87 · outbound
Group Relative Augmentation for Data Efficient Action Detection Internvideo2: Scaling foundation models for multimodal video understanding
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 742852c3-85d0-42b1-a66e-8bb309ec9bde · outbound
Group Relative Augmentation for Data Efficient Action Detection Feature adaptation with clip for few-shot classification
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f6f3f7f4-705c-4f09-99ec-ec3291c07080 · outbound
Group Relative Augmentation for Data Efficient Action Detection Cora: Adapting clip for open- vocabulary detection with region prompting and anchor pre-matching
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 90e63cc7-bf95-4ea5-aa75-fa1289e5b836 · outbound
Group Relative Augmentation for Data Efficient Action Detection VideoCLIP: Contrastive Pre-training for Zero-shot Video-Text Understanding
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be5f889f-9d9e-4c31-bc0c-d8ab0347161f · outbound
Group Relative Augmentation for Data Efficient Action Detection VideoCoCa: Video-Text Modeling with Zero-Shot Transfer from Contrastive Captioners
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a435b824-26a2-43b8-ab32-aa11018d5207 · outbound
Group Relative Augmentation for Data Efficient Action Detection Vid2seq: Large-scale pretraining of a visual language model for dense video captioning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c0d2707b-60cc-43ad-ac3a-837849da7257 · outbound
Group Relative Augmentation for Data Efficient Action Detection Image Data Augmentation for Deep Learning: A Survey
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1dfee674-368b-4c65-a70e-9e3d971d1b73 · outbound
Group Relative Augmentation for Data Efficient Action Detection Textmania: Enriching visual feature by text-driven manifold augmentation
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 029b2d6d-5e9d-45b5-aa3e-3c75799d9622 · outbound
Group Relative Augmentation for Data Efficient Action Detection CoCa: Contrastive Captioners are Image-Text Foundation Models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9b956d0-0445-4635-96b6-4c7a73d489f2 · outbound
Group Relative Augmentation for Data Efficient Action Detection Cutmix: Regularization strategy to train strong classifiers with local- izable features
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 67fffa54-4971-474c-8cf5-517ee3aec1ce · outbound
Group Relative Augmentation for Data Efficient Action Detection Fasa: Feature augmentation and sampling adaptation for long-tailed instance segmentation
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3778ec90-35a5-4105-af86-eea23753776b · outbound
Group Relative Augmentation for Data Efficient Action Detection Open- vocabulary object detection using captions
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f06c7f30-fcb4-4a29-84aa-7b381c0b3191 · outbound
Group Relative Augmentation for Data Efficient Action Detection Sigmoid loss for language image pre-training
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 894229cf-580e-40d6-bde8-c47c77ae6cb1 · outbound
Group Relative Augmentation for Data Efficient Action Detection Tip-Adapter: Training-free CLIP-Adapter for Better Vision-Language Modeling
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fead464-266c-48d7-bad9-279b38e5a680 · outbound
Group Relative Augmentation for Data Efficient Action Detection Don't Judge by the Look: Towards Motion Coherent Video Representation
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ee2b1069-a638-46dd-8642-57333de3b69e · outbound
Group Relative Augmentation for Data Efficient Action Detection Regionclip: Region- based language-image pretraining
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 15a2ed48-d5a7-425b-9651-0d15ca654883 · outbound
Group Relative Augmentation for Data Efficient Action Detection Not all features matter: Enhancing few-shot clip with adaptive prior refine- ment
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.