Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T15:53:31.889301Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 2 inbound Pith citation observations for arXiv:2411.13836.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T15:53:31.889301Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T20:04:44.106636Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-13T22:21:15.797283Z
51 of 51 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e8d4222c-c7a9-42dd-8f17-90c0bb26a74d · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Weakly su- pervised learning of instance segmentation with inter-pixel relations
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6c6c5a6a-84a4-4045-aa37-4967a6e6d61d · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Zero-shot semantic segmentation
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d18a48e3-a586-4cd5-b360-d17f0e54e0a9 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Coco- stuff: Thing and stuff classes in context
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 197c4d1a-34e0-4973-abff-865a2fdd9132 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Emerg- ing properties in self-supervised vision transformers
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 2fb29648-8d04-4c9b-b931-1e7a56e978bd · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Learn- ing to generate text-grounded mask for open-world semantic segmentation from only image-text pairs
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation dfd3e05b-4488-4f30-9402-a96375d05028 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Masked-attention mask transformer for universal image segmentation
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation bfa636b9-7eb7-4292-983e-64c112e2cc83 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Reproducible scaling laws for contrastive language-image learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 5272cc8f-b55b-45cd-92b1-ba3b62e3fe37 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation CAT-Seg: Cost Aggregation for Open-Vocabulary Semantic Segmentation
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c3e7584-fa90-43a4-8185-0a4fe387b6ed · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Open-Vocabulary Universal Image Segmentation with MaskCLIP
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d277024b-19b1-4fc2-b1f8-077edea43f61 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78f240fe-ec8b-4796-aadf-3b4b65160abe · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation The pascal visual object classes (voc) challenge
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0881ea7d-4ff2-4390-8134-9d8081342a30 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Scaling up visual and vision-language representa- tion learning with noisy text supervision
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 2274c32d-a5ad-4847-8272-79a291cfca52 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Diffusion models for zero-shot open-vocabulary segmentation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6b827282-485e-49ef-88eb-515e84b63026 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Segment anything
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 423ec143-ef68-46a0-8dd1-e6d917960360 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation ProxyCLIP: Proxy attention improves clip for open-vocabulary segmentation
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6dc9b584-b0b6-4919-87f8-cb622f46ef15 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation ClearCLIP: Decom- posing clip representations for dense vision-language infer- ence
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0e733d73-93f9-4fc4-b743-8420134f2b27 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Anti- adversarially manipulated attributions for weakly and semi- supervised semantic segmentation
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9fa5cd2c-22c7-4835-9522-776468f09def · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation A Closer Look at the Explainability of Contrastive Language-Image Pre-training
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29a748cd-ff89-4e14-8ada-6c5fa9bea170 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Open-vocabulary semantic segmentation with mask-adapted clip
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation cafa38ba-9d28-4b8a-a449-f375d9a8c235 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation CLIP is also an efficient segmenter: A text-driven approach for weakly su- pervised semantic segmentation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 449e98ac-1963-48a5-babc-1bd3c2baeffb · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation TagCLIP: A local-to-global framework to enhance open-vocabulary multi-label classification of clip without training
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9744087a-2167-4d00-b9a8-cbc741bbe050 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Fully convolutional networks for semantic segmentation
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 66449b3c-c91c-4a61-8ca8-18a167b6a519 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation SegCLIP: Patch aggregation with learnable centers for open-vocabulary semantic segmentation
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 4cbf94d9-3b2a-4eb4-9e34-5ee0d946aa24 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation The role of context for object detection and semantic segmentation in the wild
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 82a288b3-01c7-41c6-9ee1-579ecf5ed70b · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Learning transferable visual models from natural language supervision
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d3175b13-b7f6-4075-a543-9417b6f3b717 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Per- ceptual grouping in contrastive vision-language models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e7ea2c76-8fe9-4204-97ac-b02c22479857 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation ViewCo: Discovering text-supervised segmentation masks via multi-view semantic consistency
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation dc4bb987-e0bd-48da-8858-9c13f054bca6 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation High-resolution image syn- thesis with latent diffusion models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 8db690b3-cad4-42db-9946-f1bf9512de3a · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation To- ken contrast for weakly-supervised semantic segmentation
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ff6fa4f0-1142-43e1-ac19-a7faab042712 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Laion-5B: An open large-scale dataset for training next generation image-text models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 3ef15bd6-675a-46e3-ae5b-d66e8b41b6db · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation ReCo: Re- trieve and co-segment for zero-shot transfer
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 76af3956-dcb7-43e3-82b4-dade04414c75 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation iSeg: An Iterative Refinement-based Framework for Training-free Segmentation
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ee87535e-4d4b-41ad-8387-f8c20bdb7c73 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation CLIP as RNN: Segment countless visual concepts with- out training endeavor
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0e62b1ef-f7ee-410a-9bee-51decc852a1f · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Sclip: Rethinking self-attention for dense vision-language inference
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 537c9fe7-9180-47fc-beb5-44f643700f30 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Sam-clip: Merging vision foundation models towards semantic and spatial understanding
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f1e187ce-2dfd-4150-8125-3768a669df13 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Diffusion Model is Secretly a Training-free Open Vocabulary Semantic Segmenter
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e021f2c3-832d-4647-886f-e9733f0fac57 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Clip-dinoiser: Teaching clip a few dino tricks for open- vocabulary semantic segmentation
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 5268430a-a04e-4686-bde8-0acc42f79d99 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation From Text to Mask: Localizing Entities Using the Attention of Text-to-Image Diffusion Models
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0fd812d0-79ab-49e5-8126-9f64443fd613 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Sed: A simple encoder-decoder for open- vocabulary semantic segmentation
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6818b0d7-7fbb-4083-b73b-4807879449f2 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Alvarez, and Ping Luo
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 82e486ea-48b1-475c-bb02-2c9994ba1305 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Clims: Cross language image matching for weakly supervised se- mantic segmentation
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 179b678d-4304-41da-b23a-b38123aa2586 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Rewrite caption semantics: Bridging seman- tic gaps for language-supervised semantic segmentation
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation aa91ddfa-5545-4106-83cb-1dbe4aaf9647 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Groupvit: Semantic segmentation emerges from text supervision
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 34810fe9-a51d-4920-83db-c970a2641733 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Learning open-vocabulary seman- tic segmentation models from natural language supervision
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 2f00c189-f6a3-4241-a794-bda1d061773d · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Open-vocabulary panoptic segmentation with text-to-image diffusion models
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 16223171-313d-4031-a611-a09d431217d6 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Multi-class token transformer for weakly supervised se- mantic segmentation
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0732b187-27e6-4bd9-9ecd-79a5528df144 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Side adapter network for open-vocabulary semantic segmentation
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation dedbb98a-65d2-424e-b811-32470ae6b714 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Convolutions Die Hard: Open-Vocabulary Segmentation with Single Frozen Convolutional CLIP
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91ee6cf4-fc84-47f8-bc5f-272acd54846b · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Open vocabulary scene parsing
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c5655f20-8b1e-4db0-8adb-12423fbf6252 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Semantic under- standing of scenes through the ade20k dataset
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3241dc5e-106a-4fec-82df-4def85422460 · outbound
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation Extract free dense labels from clip
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a7d5046b-d83a-4772-9b7f-346ca2037fe4 · inbound
CorrCLIP: Reconstructing Patch Correlations in CLIP for Open-Vocabulary Semantic Segmentation CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fee426a-3576-4ef6-b5f8-473170d116dc · inbound
Perception Encoder: The best visual embeddings are not at the output of the network CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation
Reference 128
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.