Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:03:26.439718Z
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 64 of 64 outbound references and 2 inbound Pith citation observations for arXiv:2504.14032.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:03:26.439718Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T16:44:21.060148Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-15T16:44:21.321080Z
64 of 64 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d0ffb65c-58df-4d85-bbc0-3b1021f6e9d5 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Coco- stuff: Thing and stuff classes in context
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 01f35811-148b-41cb-b54e-8ffb14b462e7 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Learning continuous image representation with local implicit image function
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e9c0e46-3301-4ed2-857b-ecc69915721d · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Schwing, Alexan- der Kirillov, and Rohit Girdhar
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e56aec69-b1e2-4620-94ae-0e2809f1937a · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models The cityscapes dataset for semantic urban scene understanding
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 022968e6-3d6d-46b1-82ef-2922d1d0438a · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Learning affinity- aware upsampling for deep image matting
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation ca019ac8-593c-4049-9663-0a2d7ce38b8d · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models An image is worth 16x16 words: Transformers for image recognition at scale
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 505ff348-5e1b-4845-8cb2-46dabc56123a · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Lanczos filtering in one and two dimen- sions
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 5e150a97-d669-448f-b0c6-319e472ba37b · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models A guide to convolution arithmetic for deep learning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa18085d-94d2-407b-a4c5-1aae4ea1dd1b · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Depth map prediction from a single image using a multi-scale deep net- work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation db8a2fdb-888a-4d9f-8712-ba01dc5645e7 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Prob- ing the 3d awareness of visual foundation models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation bd70ed5a-d8d0-4be5-9f14-d16827fad094 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Single image 3d without a single 3d image
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 4a049e1b-2bdb-42f3-b660-b3014ab6a82e · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Brandt, Axel Feld- mann, Zhoutong Zhang, and William T
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c35f94f4-6859-442a-a8bc-92cd0f0b0cbc · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Unsupervised Semantic Segmentation by Distilling Feature Correspondences
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7ed419e-94a4-446c-96c1-f176c7880ba1 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Semantic contours from inverse detectors
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c6a9d133-f79a-4a4b-a8b6-46e09aa0bf5a · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Guided image fil- tering
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 6d06c215-6ed2-42f5-b217-da2748f066e8 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Renovating names in open-vocabulary segmenta- tion benchmarks
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 0de4ac1e-2248-4778-a4c4-f17a2c4e68e2 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Space-time correspondence as a contrastive random walk
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 6c6f07e1-f108-4a21-b5d8-50cd5674bb90 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models NA VI: Category- agnostic image collections with high-quality 3d shape and pose annotations
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 8af58bbf-caba-4d98-b1d2-1ae5393671d7 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Adam: A Method for Stochastic Optimization
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2bc5a16-7675-47c3-a076-7c3e2114eb26 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Pointrend: Image segmentation as rendering
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2cd2bed3-ead7-4009-aac3-6bbc9c52687e · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Segment any- thing
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 40d9a4f9-1c4b-4564-a45a-5086f753bfe8 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Joint bilateral upsampling
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 232f6dbb-738e-454b-80a8-9364eab6cd5f · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Proxyclip: Proxy at- tention improves clip for open-vocabulary segmentation
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 196ec7ff-9ef8-4a2e-add6-0660b46dd435 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Exploring plain vision transformer backbones for object de- tection
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71fd4f56-99d8-411f-b3c8-8a27c20f7527 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Vision transformer for nerf-based view synthesis from a single input image
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a01da6c6-ab06-47af-abcf-75b8e72ea4c5 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Microsoft coco: Common objects in context
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 3267e676-8e67-45c9-98d6-243ca5b93306 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Simpleclick: Interactive image segmentation with sim- ple vision transformers
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d3345656-a2ff-4fde-9e68-c1b78363431a · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Decoupled Weight Decay Regularization
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99c14532-b431-4d91-bd21-787b37665704 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Index networks
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation f8cb0c0e-4ef1-472e-8e5e-ed860884cc3b · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Fade: Fusing the assets of decoder and encoder for task-agnostic upsampling
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 263aa2cb-f4c1-46c8-b9b5-89d27f3e8e83 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Sapa: Similarity-aware point affiliation for feature upsampling
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 354d878f-b330-4d04-8869-2a286ed67b26 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models A database of human segmented natural images and its application to evaluating segmentation algorithms and measuring ecological statistics
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 30bd273d-5e16-416b-8ce6-3edf10bb59b9 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Cubic spline interpola- tion
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c27c24ed-ea4d-4ca1-9a85-d6e405ece2e6 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Nerf: Representing scenes as neural radiance fields for view syn- thesis
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 55bd3c3b-4c8c-49ef-9e15-d6d694543b9a · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Learning deconvolution network for semantic segmentation
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2b03ef9e-9b34-47aa-b5ea-01314e784260 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models De- convolution and checkerboard artifacts
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d727c67b-eb1c-4c69-af69-946e810b4e50 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 13e9a62c-9180-43e9-9f91-3b27ff8cc12e · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models A benchmark dataset and evaluation methodology for video object segmentation
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 02277b65-c1f7-48de-b693-a3e2db9baac9 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models The 2017 DAVIS Challenge on Video Object Segmentation
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ada388d-e73a-436a-875d-ba7be166a87c · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Three pillars improving vision foundation model distillation for lidar
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d262d93-7d02-412a-a2f7-2a67fe7921ae · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Learn- ing transferable visual models from natural language super- vision
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 725bb191-03e2-42a4-adf0-3019693d319e · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Vi- sion transformers for dense prediction
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation ac334253-fc50-46a8-b987-182ee83d0360 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Am-radio: Agglomerative vision foundation model reduce all domains into one
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 48d976be-5493-42cc-b6c4-1f3500cc6201 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Glamm: Pixel grounding large multimodal model
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 961460bf-1f9f-430f-baae-4abbf2619409 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models SAM 2: Segment Anything in Images and Videos
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37be4766-e079-4087-b36f-db1d112c4781 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Grounded sam: Assembling open-world models for diverse visual tasks,
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8dbe0da4-e5fc-47b1-9f8d-6040100805a8 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models U- net: Convolutional networks for biomedical image segmen- tation
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1da3f8cb-c7f6-41ef-b19a-68b61616c32c · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models ” grabcut” interactive foreground extraction using iterated graph cuts
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 4b2eb3bf-ded7-4ce3-b308-8b776fcdd063 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Is the deconvolution layer the same as a convolutional layer?
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e4b9436-0eaa-46e6-b6ed-8f48ef5a589e · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Adaptis: Adaptive instance selection network
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 86083929-f0fa-4ffa-8891-0d78ba74b985 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Re- viving iterative training with mask guidance for interactive segmentation
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation cdc25dd5-e865-4a1b-b8e6-d003968ebab9 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Lift: A surprisingly simple lightweight feature transform for dense vit descriptors
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13d99416-c639-4449-8a2c-a7b10b47ae51 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Splatter image: Ultra-fast single-view 3d recon- struction
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c3cc74e-5ec7-4408-93ed-93ec6a2af364 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 295f7f4a-ef11-49ae-bc6c-3ea906ce0f26 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Dino-tracker: Taming dino for self-supervised point tracking in a single video
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 298b4c20-b600-4c03-bf9d-81cabb0287c3 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Carafe: Content-aware reassembly of fea- tures
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 05ad44e8-54c2-43d7-a1c5-6ca5e87281bf · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Self-supervised trans- formers for unsupervised object discovery using normalized cut
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 95033c19-ddf3-4674-b23a-ac402da74efd · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Clip-dinoiser: Teaching clip a few dino tricks for open- vocabulary semantic segmentation
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7812cdc5-345c-4907-8ec8-5606f99aa4d7 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Featuren- erf: Learning generalizable nerfs by distilling foundation models
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 158102a0-61ae-42f9-b3db-6794dd99c9e0 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models pixelnerf: Neural radiance fields from one or few images
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1261df93-e68b-472f-8fc0-b38887bbceff · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Convolutions die hard: Open-vocabulary seg- mentation with single frozen convolutional clip
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8041424f-202d-4b2b-94e3-88a92b90848c · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models Sigmoid loss for language image pre-training
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 0d108903-f6f4-4f17-baff-916114e47eaf · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models LoftUp: earning a Coordinate-Based Feature Upsampler for Vi- sion Foundation Models
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e9bbc884-d312-4ede-b569-6e8492229de2 · outbound
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models 1, 2, 3, 4, 6, 7, 12
Reference 128
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2a3c4294-09b6-4365-b009-02afd379bed9 · inbound
Maybe you don't need a U-Net: convolutional feature upsampling for materials micrograph segmentation LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7fad0f2a-d470-45f4-aa9d-30d5c9227b70 · inbound
UPLiFT: Efficient Pixel-Dense Feature Upsampling with Local Attenders LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.