Pith. sign in

Paper Citation Record · LEDGER

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation

As of 7 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 0 inbound Pith citation observations for arXiv:2506.11131.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.11131 v1

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:01:23.882872Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

58 of 58 outbound references displayed

  • verified exact0
  • verified fuzzy31
  • unresolved27
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1e509b39-0717-4dce-a7b4-5ced0adbec71 · outbound

This paper cites Zerowaste dataset: To- wards deformable object segmentation in cluttered scenes.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Zerowaste dataset: To- wards deformable object segmentation in cluttered scenes

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.746199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.694098Z digest=sha256:c8d5b8a4eedb499b0b4a2358b73430ad8c191cfdc3ae4f8b34f4cc55725adda9

Observation 28594f12-b209-49ab-a85a-cfe51ce3f4f5 · outbound

This paper cites Token Merging: Your ViT But Faster.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Token Merging: Your ViT But Faster

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.698370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.698370Z digest=sha256:8381113896929a999afaf58b39954a02e1efcaddef430f88c5afd045105ee2a4

Observation c86c28c2-8ee7-40b4-9a06-65874c14c5fd · outbound

This paper cites The eccentricity effect: Target eccentricity af- fects performance on conjunction searches.Perception & psychophysics, 57:1241–1261, 1995.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation The eccentricity effect: Target eccentricity af- fects performance on conjunction searches.Perception & psychophysics, 57:1241–1261, 1995

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.694667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.702224Z digest=sha256:021aa8194bd45eb0b846493da5f09de8a56e849ea6cd7fc946b50209c6b1c5d4

Observation e2183c98-8fcb-4a6c-bb8d-b4cd06db1fb2 · outbound

This paper cites Pelk: Parameter-efficient large kernel con- vnets with peripheral convolution.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Pelk: Parameter-efficient large kernel con- vnets with peripheral convolution

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.642470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.705985Z digest=sha256:e5031e82dd18041b3c4333d22c8548817c81f470f483276b58f409669934deac

Observation 55063883-09bd-41a8-89dd-fafbd792858a · outbound

This paper cites Diffrate: Differentiable compression rate for efficient vision transformers.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Diffrate: Differentiable compression rate for efficient vision transformers

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.627947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.709324Z digest=sha256:46b9538d7864151464301cab9b379048a3f9fbcf53ccda54b12067f71020f013

Observation 9b378b04-b665-45ab-af7e-f1ae93f46f5b · outbound

This paper cites The cityscapes dataset for semantic urban scene understanding.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation The cityscapes dataset for semantic urban scene understanding

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.712548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.712548Z digest=sha256:e7558ceae1a53d74b0b2140dd8055bcaf7c5a19155e1b366989f4674304c8672

Observation 8327ace9-0161-4fb0-914d-af88c483025c · outbound

This paper cites Rescaling egocentric vision: Collection, pipeline and chal- lenges for epic-kitchens-100.International Journal of Com- puter Vision, pages 1–23, 2022.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Rescaling egocentric vision: Collection, pipeline and chal- lenges for epic-kitchens-100.International Journal of Com- puter Vision, pages 1–23, 2022

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.607128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.716082Z digest=sha256:8ede99cfcd1449759dfb22e2048a0042aeb529c023851597ac9bf6822f6dad69

Observation 55363477-3f56-4009-bd58-67cd24293673 · outbound

This paper cites Vision Transformers Need Registers.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Vision Transformers Need Registers

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.719371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.719371Z digest=sha256:566b752a7263405a63cf1e6ae04ba91c798d8cddb491bb113fe4ca95ada991f4

Observation eaf99e91-6ff2-42c3-a664-fb632d5f4745 · outbound

This paper cites Epic-kitchens visor benchmark: Video segmenta- tions and object relations.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Epic-kitchens visor benchmark: Video segmenta- tions and object relations

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.593588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.722720Z digest=sha256:627e791a95c2263be95bbb8a72d319d568b225a21122feda002842b701a7f190

Observation 0f531688-f6f3-4537-9755-c3accd759b1f · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Imagenet: A large-scale hierarchical image database

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.580040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.725866Z digest=sha256:96e2d8c8f461849b7f48fdd57a84c938877375812b65c974e586411361479797

Observation 5bd0c6c2-48f7-406f-8e90-cb51c855634a · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.728704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.728704Z digest=sha256:7290171f3f5d6560356c78a4505977c1eb6ebc91e5ffc6f355f079c132a831df

Observation 72ea821b-b086-4d6d-b65a-2f06aad9f9ab · outbound

This paper cites Project Aria: A New Tool for Egocentric Multi-Modal AI Research.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Project Aria: A New Tool for Egocentric Multi-Modal AI Research

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.732463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.732463Z digest=sha256:84048140454ec919fd617d1af2e043c37ca322a0515cf885783fd863ef2a19c4

Observation 2b6c4ab8-f20c-40e8-8e11-08ac62ec1bf0 · outbound

This paper cites Adaptive token sampling for efficient vision transformers.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Adaptive token sampling for efficient vision transformers

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.554245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.735627Z digest=sha256:6523ee915641b9e259c58eda92bda9f56f4e39861c441b25ff9761b5d8652320

Observation 21c265b8-1918-4a76-8f95-917f683ff375 · outbound

This paper cites Instance segmen- tation for autonomous log grasping in forestry operations.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Instance segmen- tation for autonomous log grasping in forestry operations

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.515908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.738525Z digest=sha256:c819e81106d6823cc7ca210701deb62deaa3aa12ad835858ec6ad59a986854b5

Observation 99a98a1c-017e-41cd-a25a-d4f760767f8c · outbound

This paper cites Deep residual learning for image recognition.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Deep residual learning for image recognition

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.463612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.741588Z digest=sha256:049ea901624b5d548aa6cd236a4ce25f89e4b7d7219239cf7210caa3328d0abc

Observation 18b36cfc-202c-4303-b830-5c8676a3cd26 · outbound

This paper cites Masked autoencoders are scalable vision learners.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Masked autoencoders are scalable vision learners

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.445343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.744538Z digest=sha256:b9b711e41d897e30bf3c6b9b324cbda97a8145c8147b65abdf08c0b07f676188

Observation 73d796ab-7a26-4cb7-a8bc-8af8e40e5868 · outbound

This paper cites Bytes Are All You Need: Transformers Operating Directly On File Bytes.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Bytes Are All You Need: Transformers Operating Directly On File Bytes

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.747349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.747349Z digest=sha256:1f53173109c511e93c801e0d285330bfac97f1939b9ce87669c3a4d4b478cf85

Observation 40a1255e-c691-4dd0-81b2-330d8e2c0bec · outbound

This paper cites FoveaTer: Foveated Transformer for Image Classification.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation FoveaTer: Foveated Transformer for Image Classification

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.751062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.751062Z digest=sha256:4dd3c0d8ca406746edb4891f70e1363dd77038537d5cf2ec40b64cffebd1769c

Observation 0fa90366-b06b-4c18-999a-852cc05fc66c · outbound

This paper cites Scaling Laws for Neural Language Models.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Scaling Laws for Neural Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.754450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.754450Z digest=sha256:1b4467bd21fb2510d75bb4325a3b224d9e341b3bd63b1682dcd76c502753ef89

Observation 8e62f6ee-dc64-4cb4-96a4-a8b2d340878b · outbound

This paper cites Segment any- thing.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Segment any- thing

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.433749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.758293Z digest=sha256:316818f56f268c9e69eab074c020fe032da7f55d7d9eb1401f93a8c76d17fd0a

Observation c7db607a-ee3c-4385-8c74-e1c25ba8a50b · outbound

This paper cites Spvit: Enabling faster vision transformers via latency-aware soft token pruning.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Spvit: Enabling faster vision transformers via latency-aware soft token pruning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.422770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.761078Z digest=sha256:3b3e00d54a03019232dcd3b33df2627ff48be73f075c7265a3c9af7f740b9de8

Observation 9ca957d8-7dc4-43bd-9854-dd88108d4785 · outbound

This paper cites GazeGPT: Augmenting Human Capabilities using Gaze-contingent Contextual AI for Smart Eyewear.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation GazeGPT: Augmenting Human Capabilities using Gaze-contingent Contextual AI for Smart Eyewear

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.764359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.764359Z digest=sha256:9decb4d1c8b2ac01682adaeae6a99630123d919cc80e31a485bac707262d438b

Observation 590a6293-dc25-4f1d-9958-f665e0569fb0 · outbound

This paper cites Imagenet classification with deep convolutional neural net- works.NeurIPS, 25, 2012.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Imagenet classification with deep convolutional neural net- works.NeurIPS, 25, 2012

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.411540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.767641Z digest=sha256:dfe04e5872854cf986e4d61c2c4c6ff7c74ba94f34576c5ba7c89654f34a9d57

Observation 63c6b7bf-c21b-49ac-a983-781244081ff1 · outbound

This paper cites Microsoft coco: Common objects in context.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Microsoft coco: Common objects in context

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.771069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.771069Z digest=sha256:22dec36b093060e30b4591fab27eeda4a6f0f854af89fe678caecdcdef10652f

Observation 36157df9-59b5-4627-a570-b74ea1f749c3 · outbound

This paper cites Efficientvit: Memory efficient vision transformer with cascaded group attention.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Efficientvit: Memory efficient vision transformer with cascaded group attention

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.390481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.774471Z digest=sha256:2e09a7769f9f6871c54dd66766208bd1049cfb955bcf455a5a94cfbee5af95c6

Observation 084b4e53-0296-46e9-af71-b67bc7baa978 · outbound

This paper cites Nymeria: A Massive Collection of Multimodal Egocentric Daily Motion in the Wild.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Nymeria: A Massive Collection of Multimodal Egocentric Daily Motion in the Wild

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.777693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.777693Z digest=sha256:6e430a24f5bf9c3a86fb33480d2280bc1d5abf2835356ca398a7a5477756e6ff

Observation 788a1717-40dd-4ee4-8aa0-f37e6cdbb4a8 · outbound

This paper cites Token Pooling in Vision Transformers.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Token Pooling in Vision Transformers

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.781218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.781218Z digest=sha256:0eb9a66a1a35b018f2e6f30960362f57cd2ae0e149a93cf7eae905eedfa64da2

Observation a7e9de63-7a8c-483d-a0e1-f877df98dc81 · outbound

This paper cites Adavit: Adaptive vision transformers for efficient image recognition.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Adavit: Adaptive vision transformers for efficient image recognition

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.377985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.785126Z digest=sha256:4840cbf5465d65e6ed7767fc281de6f61f087391ca6184c4cff93c1d5340e165

Observation 4c17ce08-a961-4e2a-b3ca-24f2b5bde1e1 · outbound

This paper cites Peripheral vision transformer.NeurIPS, 35:32097–32111,.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Peripheral vision transformer.NeurIPS, 35:32097–32111,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.365746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.788274Z digest=sha256:6961d75fff93791d672e304f576657cc9727931ed6a4031b47e2749f90f34bd0

Observation f56ad5c0-676c-4893-aac9-98ab94bf32b6 · outbound

This paper cites Finely-grained annotated datasets for image-based plant phenotyping.Pattern recognition letters, 81:80–89, 2016.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Finely-grained annotated datasets for image-based plant phenotyping.Pattern recognition letters, 81:80–89, 2016

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.354346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.791261Z digest=sha256:e870a3c2377fec8aab5d4c6ef8a292308ef25c070963a635aa21b044b46c6047

Observation 1235dc9c-c215-4930-968f-9f7b0e3a0247 · outbound

This paper cites Rgb no more: Minimally- decoded jpeg vision transformers.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Rgb no more: Minimally- decoded jpeg vision transformers

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.341396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.794455Z digest=sha256:a0169dbca6e5e75d167e11ad6b2f558de9f5931c10480f798d4a000f227a60c0

Observation cbf95ed2-4d45-4c90-a0bd-b8014a99c4cd · outbound

This paper cites Dynamicvit: Efficient vision transformers with dynamic token sparsification.NeurIPS, 34:13937–13949, 2021.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Dynamicvit: Efficient vision transformers with dynamic token sparsification.NeurIPS, 34:13937–13949, 2021

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.329601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.797553Z digest=sha256:239fa03a032c077bb2e8c7f6ec413eaadebc67a73f6fd80da7fe515d58b89f04

Observation 69d40b15-99d7-4f63-8ae8-d2561eba0191 · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation SAM 2: Segment Anything in Images and Videos

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.800874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.800874Z digest=sha256:beef052663dbde8a920809ea069980973c1f66db1d51cb410ff5f645ef070cac

Observation c65e594d-c154-445b-a67a-def22ddd5241 · outbound

This paper cites Learning to Merge Tokens in Vision Transformers.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Learning to Merge Tokens in Vision Transformers

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.804525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.804525Z digest=sha256:350025c210621ad8c39c9358f43e147b62f7cb73e8935bddd32481a1e35be027

Observation 7ccae9cd-45a8-41a1-a6a8-b6cb26b7a985 · outbound

This paper cites CP-ViT: Cascade Vision Transformer Pruning via Progressive Sparsity Prediction.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation CP-ViT: Cascade Vision Transformer Pruning via Progressive Sparsity Prediction

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.807850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.807850Z digest=sha256:2fda0d890f86148a361a5e612033482fb769fcf33d9abf272b59875b2058428e

Observation e8f10b38-f447-44eb-968a-3d3eaf864169 · outbound

This paper cites On Efficient Variants of Segment Anything Model: A Survey.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation On Efficient Variants of Segment Anything Model: A Survey

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.811323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.811323Z digest=sha256:346eb446cedc7dbf26d88096b7bff07ba4c8bdd34a07134ada88989641e34432

Observation d23f2a49-6e4d-41b8-84f9-aba653abed04 · outbound

This paper cites NDD20: A large-scale few-shot dolphin dataset for coarse and fine-grained categorisation.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation NDD20: A large-scale few-shot dolphin dataset for coarse and fine-grained categorisation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.814417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.814417Z digest=sha256:b59978dd71bd077d6e3b3fbba1e43b12f112005ee2693fd3b2ddae1ea0729c4d

Observation 43bd1eec-cba5-4ced-a584-1a7f911a7b8d · outbound

This paper cites Neural discrete representation learning.NeurIPS, 30, 2017.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Neural discrete representation learning.NeurIPS, 30, 2017

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.818098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.818098Z digest=sha256:974452431d29a2bc02e5e0c811773c2fbbba36847b9fc0cfb4085fc42a6e7e03

Observation b4ca801e-802b-431c-a17c-b8c84bc54fda · outbound

This paper cites SqueezeSAM: User friendly mobile interactive segmentation.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation SqueezeSAM: User friendly mobile interactive segmentation

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.821808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.821808Z digest=sha256:b8c25f5aec94f6b1817576989c48030c359f774802fe83611b28d8fab771171b

Observation 39280563-a8b6-411a-928b-0b8736568912 · outbound

This paper cites Efficientsam: Leveraged masked image pretraining for efficient segment anything.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Efficientsam: Leveraged masked image pretraining for efficient segment anything

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.309971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.825145Z digest=sha256:c87707ac274fffbb9bd77d1b51e014e008e1e7c2dcb6c63362f7802ac4064fc4

Observation 4fb11bf9-fcfe-488a-bb6c-46e94179d6f8 · outbound

This paper cites ElasticTok: Adaptive Tokenization for Image and Video.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation ElasticTok: Adaptive Tokenization for Image and Video

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.828275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.828275Z digest=sha256:9ce91defdca5167a44b3093951117e127b2e452f4d1424245661df1d8f620bd6

Observation d5c30b82-a863-4e20-8dfa-befadddf50fa · outbound

This paper cites A-vit: Adaptive tokens for efficient vision transformer.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation A-vit: Adaptive tokens for efficient vision transformer

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.299277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.831259Z digest=sha256:35c6cce82acf053eae98fbe4897e7d00c7b03ab100941bdb6d97b7487a9f5de7

Observation 0e845d06-82ed-4ee2-8822-c956a82a55ec · outbound

This paper cites Woodscape: A multi-task, multi-camera fisheye dataset for autonomous driving.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Woodscape: A multi-task, multi-camera fisheye dataset for autonomous driving

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.288980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.834300Z digest=sha256:3d7a3f3b9290040ccc21a225c2447ed344c1343706270b4bae466b53f579068c

Observation 51258c3e-8c81-49a3-8ae8-791671ac755e · outbound

This paper cites Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.837602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.837602Z digest=sha256:68858b8e854937fee3a79a112b66d40a601ab030050255d9c7039306b63196aa

Observation 1406d1ef-0d27-4c18-a42b-e0873baf7cbf · outbound

This paper cites An Image is Worth 32 Tokens for Reconstruction and Generation.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation An Image is Worth 32 Tokens for Reconstruction and Generation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.840987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.840987Z digest=sha256:c6ba16fd3fc722e7ab2f84d9e5a040624259427bc5e475b69b17dbda48831dc5

Observation c971f6c8-29c5-4407-8401-1fe63677d38d · outbound

This paper cites Faster Segment Anything: Towards Lightweight SAM for Mobile Applications.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Faster Segment Anything: Towards Lightweight SAM for Mobile Applications

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.844273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.844273Z digest=sha256:67e6bd4d176b930c02bd74ddee5e74cdcd4265ccb7ac5ede6ce3569f83b4e55c

Observation 09b06f40-1aee-42b8-9c5a-6f73400cc95c · outbound

This paper cites Fine-grained egocentric hand-object segmentation: Dataset, model, and applications.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Fine-grained egocentric hand-object segmentation: Dataset, model, and applications

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.278276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.847206Z digest=sha256:a16480b559673e4425dc3b78232cc1e77aff06d064208c489948ea78ac50aaeb

Observation 13e08c05-b92e-4eb8-a142-e51cfd207e2d · outbound

This paper cites Efficientvit-sam: Accelerated segment anything model without performance loss.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Efficientvit-sam: Accelerated segment anything model without performance loss

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.267838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.850081Z digest=sha256:21e23420a64b1749ad2e7e657b9c2d5f7a38d15511bbb168188f5e2a1c46c77b

Observation e57b58ee-d1ee-4a3b-bb66-c6f4e81d541b · outbound

This paper cites Fast Segment Anything.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Fast Segment Anything

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.853152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.853152Z digest=sha256:6ce9251a290a1b2e77275c9f9e2ef6147e474d6b3974805e74131db96ff4e866

Observation 7ce1e0d3-1524-42d2-bd27-260cfd55224e · outbound

This paper cites Semantic under- standing of scenes through the ade20k dataset.IJCV, 127: 302–321, 2019.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Semantic under- standing of scenes through the ade20k dataset.IJCV, 127: 302–321, 2019

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.256436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.856235Z digest=sha256:9927ea5fce24041d8a3e934b868cec43b14498c9afbf49621f31c5928cd70c48

Observation 4f101be8-ea3f-41fc-980f-1c44a8ac2330 · outbound

This paper cites EdgeSAM: Prompt-In-the-Loop Distillation for SAM.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation EdgeSAM: Prompt-In-the-Loop Distillation for SAM

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.859111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.859111Z digest=sha256:2e4214baedcba96606afe0a61118f5528ffa488f0b46bd746cf89d6114de8efb

Observation 9a6763cf-b739-4160-aff5-205b06f3fbe4 · outbound

This paper cites MAE Pre-training We pre-trained our foveated image encoders using MAE pre-training for 500K iterations.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation MAE Pre-training We pre-trained our foveated image encoders using MAE pre-training for 500K iterations

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.246207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.862624Z digest=sha256:921a37df0adbfef11251d4c8e2bfec74728a1de8d79df826937d985024103a16

Observation ed16bc5d-091d-4f38-bd2f-fd6573629003 · outbound

This paper cites Here we give a more formal definition of the parameterization of such a pattern.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Here we give a more formal definition of the parameterization of such a pattern

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.235663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.866031Z digest=sha256:cb2a06a032918310eac0d261fc4fecd9b686a0e743084478c39fbb1423995748

Observation 47069a74-dd08-4d4f-8a35-60ad7cff6bd0 · outbound

This paper cites We plot the training loss curves in Figure 12.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation We plot the training loss curves in Figure 12

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.223472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.870145Z digest=sha256:2be31d06796f7a1b75581a207d54b0c75d7dbfd994f06b936532da0d263a9166

Observation e45f3356-3b77-40e8-bd38-290c2a6498fd · outbound

This paper cites Our foveation patterns exists in a high-dimensional design space, and each new pattern requires its own MAE pre-training.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Our foveation patterns exists in a high-dimensional design space, and each new pattern requires its own MAE pre-training

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.211826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.873570Z digest=sha256:cad815dca3b800cbd2874d5aca0febe2f359d3b7a15e01885b12455c86a4b01e

Observation 813448ba-c64e-4019-a5c7-c23e2205b705 · outbound

This paper cites an unresolved cited work.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:01:24.200458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.876636Z digest=sha256:fc710f9d07161f312c1426c3ac3bd8be6b45d996120d22a2a684e209933c7fb8

Observation 399a220f-faf9-4e66-be1f-94a640639ef8 · outbound

This paper cites an unresolved cited work.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:01:24.190238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.879791Z digest=sha256:cc58460e00655c82b983bbeb9736a891e8701bd2a53c29ef7c6428280b922aca

Observation 51feced0-379a-438d-ba46-36a366146a89 · outbound

This paper cites in computing FLOP counts for transformer architectures (c.f.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation in computing FLOP counts for transformer architectures (c.f

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:24.179128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:23.882872Z digest=sha256:ad979e0fc8fa841578f6b743979612bf2eec0b0ae8f13939453d3ee18030b659

Pith citing papers

No inbound Pith citation observations are available.