Pith. sign in

Paper Citation Record · LEDGER

GViT: Representing Images as Gaussians for Visual Recognition

As of 8 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 0 inbound Pith citation observations for arXiv:2506.23532.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.23532 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:47:13.856071Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

54 of 54 outbound references displayed

  • verified exact3
  • verified fuzzy28
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9945d05e-7020-4d1e-9ab0-bd728522523a · outbound

This paper cites Slic superpixels compared to state-of-the-art superpixel methods.

GViT: Representing Images as Gaussians for Visual Recognition Slic superpixels compared to state-of-the-art superpixel methods

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:18.892758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:11.014039Z digest=sha256:2315e38c3752a1d69814d3ed4dba61c144fa3ad27d5e92ec124be882ded506c9

Observation 2c890d52-e762-428e-9657-dad3cf888b27 · outbound

This paper cites BEiT: BERT Pre-Training of Image Transformers.

GViT: Representing Images as Gaussians for Visual Recognition BEiT: BERT Pre-Training of Image Transformers

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:11.069323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:11.069323Z digest=sha256:2e4ffcc87b311a047ea1820defca413b88532f78216fadc3bb3e527195dd039e

Observation 0092c7e7-3d0d-4b55-b20f-f251c06928f2 · outbound

This paper cites an unresolved cited work.

GViT: Representing Images as Gaussians for Visual Recognition Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:47:18.746717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:11.138184Z digest=sha256:a2dd32e1cb837f3955e46a697e8272ea7c1105b5cfc752215df78ad502ed8a3d

Observation 5910e295-c639-4bc8-80ed-03a55f33dc01 · outbound

This paper cites Class-Discriminative Attention Maps for Vision Transformers.

GViT: Representing Images as Gaussians for Visual Recognition Class-Discriminative Attention Maps for Vision Transformers

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:11.202139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:11.202139Z digest=sha256:fcbc1333aca74778aab45f3268ef17275b034457feab440106cbdefa483544d1

Observation fd5ef612-735c-428c-a5a7-dfeddb721ae2 · outbound

This paper cites Generative pretraining from pixels.

GViT: Representing Images as Gaussians for Visual Recognition Generative pretraining from pixels

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:11.235366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:11.235366Z digest=sha256:ece7774695dd6457c4df208309e8a9f06d0ba3e7676d475e696ff675c3911b87

Observation 936e8e54-4ae0-45ba-b2b1-bb233b9fc377 · outbound

This paper cites ELECTRA: Pre-training Text Encoders as Discriminators Rather Than Generators.

GViT: Representing Images as Gaussians for Visual Recognition ELECTRA: Pre-training Text Encoders as Discriminators Rather Than Generators

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:11.283757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:11.283757Z digest=sha256:a77a486b42257b25267d7446e806e21724aac2d0b7ab084d665cc80085bd9f74

Observation f4c5c823-5540-4b5a-b51a-db3931724776 · outbound

This paper cites Vision transformers need registers.

GViT: Representing Images as Gaussians for Visual Recognition Vision transformers need registers

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:18.656085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:11.350695Z digest=sha256:56d24f6f5d8194e6017f286ce6a1f8a608207d35f07239fafbef08d17f8e3fa9

Observation 37e73944-bae5-47a0-860c-c20c99361830 · outbound

This paper cites Scaling vision transformers to 22 billion parameters.

GViT: Representing Images as Gaussians for Visual Recognition Scaling vision transformers to 22 billion parameters

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:11.395995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:11.395995Z digest=sha256:e1af748bc637daf50dfb8de0545a62a8e9821b58d5392117595debac554e74b2

Observation fdd7961f-9de7-4c5f-be6a-aead69cc1528 · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale.

GViT: Representing Images as Gaussians for Visual Recognition An image is worth 16x16 words: Transformers for image recognition at scale

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:11.450427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:11.450427Z digest=sha256:83128489574df80e29782ab2db316abb7865ffe6f8737627fc535c38db20cc3a

Observation 9c3db3c5-a6d1-4b5d-ba0d-e706919159f9 · outbound

This paper cites Adaptive slot attention: Object discovery with dynamic slot number.

GViT: Representing Images as Gaussians for Visual Recognition Adaptive slot attention: Object discovery with dynamic slot number

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:18.466824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:11.569998Z digest=sha256:51edc71f7906b3c985c0ab3779f189d59eb3344470267eab35be7fbdc3c000dc

Observation f66e188e-2881-48ee-bffe-e431d0f62f50 · outbound

This paper cites 3d gaussian splatting as new era: A survey.

GViT: Representing Images as Gaussians for Visual Recognition 3d gaussian splatting as new era: A survey

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:18.364413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:11.604742Z digest=sha256:765801c1153c78683c835b08463fece14bc5403a94bc2660e1c74ba30226c6ac

Observation c1fc16db-617d-46d9-86a7-162eed514be1 · outbound

This paper cites Efficient graph-based image segmentation.

GViT: Representing Images as Gaussians for Visual Recognition Efficient graph-based image segmentation

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:18.238363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:11.656350Z digest=sha256:32bfed56fa9d1f8958737314369d8071fff39b88c3038a89cfd7cbe81f00c024

Observation 579af9de-e56a-451e-b5b9-2714f0311f26 · outbound

This paper cites Understanding the difficulty of training deep feedfor- ward neural networks.

GViT: Representing Images as Gaussians for Visual Recognition Understanding the difficulty of training deep feedfor- ward neural networks

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:18.119540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:11.719489Z digest=sha256:87e919c127ddea24afaf0978db29e0bb498608ef9644615ff084d42385da7c75

Observation ff0bb9f3-9761-48f2-affb-e7a0de90b98a · outbound

This paper cites Explaining and Harnessing Adversarial Examples.

GViT: Representing Images as Gaussians for Visual Recognition Explaining and Harnessing Adversarial Examples

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:11.771826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:11.771826Z digest=sha256:2c56991b70a7115f97adb06ea93f32dc6895f74cb4e6d541aa851bb47e079e22

Observation b192aebd-5bf7-4416-977f-feba17913d7e · outbound

This paper cites Accurate, Large Minibatch SGD: Training ImageNet in 1 Hour.

GViT: Representing Images as Gaussians for Visual Recognition Accurate, Large Minibatch SGD: Training ImageNet in 1 Hour

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:11.840094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:11.840094Z digest=sha256:09cc328e1193408aafc23868a0f0379c2d1250ba7467e6adb453e73de8c83a0f

Observation 3c6a61b4-d77f-49b5-bc12-b73b59d2e83d · outbound

This paper cites Faster neural networks straight from jpeg.

GViT: Representing Images as Gaussians for Visual Recognition Faster neural networks straight from jpeg

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:17.992426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:11.888156Z digest=sha256:8352b5c94556fd5e67a93c775f29b2c9c1df19f01caa704869d711192b2abe6e

Observation dc0b8bf2-6a07-4b45-a0b1-4398de97033a · outbound

This paper cites Masked autoencoders are scalable vision learners.

GViT: Representing Images as Gaussians for Visual Recognition Masked autoencoders are scalable vision learners

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:11.946548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:11.946548Z digest=sha256:c25871d9c4a272c9911888ab1e40a1000af08fdbdec1076d1507e8717365dcc4

Observation 261a297a-38f0-42d1-a87a-84055dd8365c · outbound

This paper cites Deep residual learning for im- age recognition.

GViT: Representing Images as Gaussians for Visual Recognition Deep residual learning for im- age recognition

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:17.880417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:12.019522Z digest=sha256:d51d96f3efb4f01156bcc2698e476171222f5a1843d6075aac88d66da64069cc

Observation 846f3813-b848-4d05-990d-3c5ce50e77ea · outbound

This paper cites Bytes are all you need: Transformers operating directly on file bytes.

GViT: Representing Images as Gaussians for Visual Recognition Bytes are all you need: Transformers operating directly on file bytes

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:17.757233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:12.058765Z digest=sha256:ff30418001743b966cc656de7f12a4f846f1b2838b0c1b5a8605d0d16f054468

Observation 200ecfb6-b37f-43f2-9086-224e9134a633 · outbound

This paper cites 2d gaussian splatting for geometrically accurate radiance fields.

GViT: Representing Images as Gaussians for Visual Recognition 2d gaussian splatting for geometrically accurate radiance fields

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:17.643782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:12.080524Z digest=sha256:f23d2ec9d46818dd7276c7d90abcae3f6b25048627dcce1d4ea4c0468de006b8

Observation fa411bc1-4037-4db5-8286-9e0aeb480988 · outbound

This paper cites Perceiver IO: A General Architecture for Structured Inputs & Outputs.

GViT: Representing Images as Gaussians for Visual Recognition Perceiver IO: A General Architecture for Structured Inputs & Outputs

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:17.543035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:12.133865Z digest=sha256:1930470ec718390ccc781ed48bb29159960d5a76ce9775ed5e388324e9482fb0

Observation 8b67a148-fc63-4ce8-97dd-954a45b029db · outbound

This paper cites Perceiver IO: A general architecture for structured inputs & outputs.

GViT: Representing Images as Gaussians for Visual Recognition Perceiver IO: A general architecture for structured inputs & outputs

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:17.398396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:12.186546Z digest=sha256:db8dfa68f3e5c932ca968bba4a08b607b7cd7fd6d644fd2337e44e7f9c54c012

Observation b1e8f434-2fcf-4c5b-a4b2-4ecf418d0a53 · outbound

This paper cites Perceiver: General Perception with Iterative Attention.

GViT: Representing Images as Gaussians for Visual Recognition Perceiver: General Perception with Iterative Attention

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:17.294073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:12.228976Z digest=sha256:a6687f064090f8c85e5b4b842da0d08017cedbfa9c8edf06d1fff25e531a6acd

Observation ea2f901f-f3cd-4cd1-80f6-79f9c737a85f · outbound

This paper cites Superpixel sampling networks.

GViT: Representing Images as Gaussians for Visual Recognition Superpixel sampling networks

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:17.179891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:12.286148Z digest=sha256:0029cfce335a0fe41b93f73c682cd3e9f5004781d13c49d3181ea974e14d1f6b

Observation 3738d5c8-68ff-4a31-be72-2ea36e1b9ea0 · outbound

This paper cites Unsupervised image segmentation by backpropagation.

GViT: Representing Images as Gaussians for Visual Recognition Unsupervised image segmentation by backpropagation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:17.066064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:12.336586Z digest=sha256:38b0ea05c251e9cad555c80415872baaf8ac701ec8af21173018946634454229

Observation 56d5519e-e377-4930-93b9-f89ca3ee78ca · outbound

This paper cites 3d gaussian splatting for real-time radiance field rendering.

GViT: Representing Images as Gaussians for Visual Recognition 3d gaussian splatting for real-time radiance field rendering

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:16.962101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:12.396900Z digest=sha256:4a60722dd5083c87f2efef4eda4e0539540d8085ecffe4325f814024b0b9a2c7

Observation a6efc2d2-c4bc-4755-a961-289d45fa081b · outbound

This paper cites Seac and the start of image processing at the national bureau of standards.

GViT: Representing Images as Gaussians for Visual Recognition Seac and the start of image processing at the national bureau of standards

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:16.825354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:12.452282Z digest=sha256:349b091b5a4d0a4dbfc47f931b817a5c73d48ed852e19be9077975de8d6d4aba

Observation 7ab25380-ec07-491d-b429-42da834e935e · outbound

This paper cites an unresolved cited work.

GViT: Representing Images as Gaussians for Visual Recognition Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:47:16.628991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:12.508278Z digest=sha256:abac2f020e8a3e7240ee92e2f21206ed62d847ef3ca1bfe5471fe8442ceabf19

Observation 6ba2d68d-75c3-4501-8c4f-b387e64c880e · outbound

This paper cites Kutulakos, David J.

GViT: Representing Images as Gaussians for Visual Recognition Kutulakos, David J

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:16.415159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:12.567226Z digest=sha256:88bb4923f89ac37673d681410b7243b8b35290c2c11bf24aa697b8c106a22e04

Observation a6cba2a0-19ec-475c-a0c7-e0cc0eefc094 · outbound

This paper cites PyTorch Distributed: Experiences on Accelerating Data Parallel Training.

GViT: Representing Images as Gaussians for Visual Recognition PyTorch Distributed: Experiences on Accelerating Data Parallel Training

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:12.598568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:12.598568Z digest=sha256:02d0cfb96a127121e54bdad9a53c9ba05e8e14d22dab0077156808a2572698ac

Observation 2c815d31-fe80-4758-a764-4601fe534137 · outbound

This paper cites A convnet for the 2020s.

GViT: Representing Images as Gaussians for Visual Recognition A convnet for the 2020s

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:16.205710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:12.646382Z digest=sha256:c68742a82849c3de44b6d2e02bb25fe0c7b414549101d1e2f3d6f7ba23b837f7

Observation bdd6e5d4-4c34-4d8b-9461-9dcca0ac6c66 · outbound

This paper cites Object-centric learning with slot attention.

GViT: Representing Images as Gaussians for Visual Recognition Object-centric learning with slot attention

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:12.704280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:12.704280Z digest=sha256:8a38e8b00450a1f2b83097b043f05f899672cf3bbe61638ca4a68bb56be5b507

Observation d926a667-08f5-4c40-bb5b-b0e508e55a00 · outbound

This paper cites Sgdr: Stochastic gradient descent with warm restarts.

GViT: Representing Images as Gaussians for Visual Recognition Sgdr: Stochastic gradient descent with warm restarts

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:16.093469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:12.746140Z digest=sha256:d8fadcde439235f1ef8984daf1560758f31e451a9ede5ab97172ef22d2b48174

Observation 30cabcaf-6846-453f-a2e5-9f7095397d4b · outbound

This paper cites Decoupled Weight Decay Regularization.

GViT: Representing Images as Gaussians for Visual Recognition Decoupled Weight Decay Regularization

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:12.793127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:12.793127Z digest=sha256:28a424f76413598056cc6731e3965139d8195092fbcebccde2e2d5826f4f7553

Observation f9c6bfc4-7a58-4297-a4ed-650674a643db · outbound

This paper cites Enhance the Visual Representation via Discrete Adversarial Training.

GViT: Representing Images as Gaussians for Visual Recognition Enhance the Visual Representation via Discrete Adversarial Training

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-06T21:47:14.362286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:12.882028Z digest=sha256:3de15a36a0e0b7a3d699d03536735ab305952090214c9938e5c213c87536f96a

Observation cf18a7bb-9b6b-4674-8f32-bc1ff44f1e3c · outbound

This paper cites An image is worth more than 16x16 patches: Exploring transformers on individual pixels.

GViT: Representing Images as Gaussians for Visual Recognition An image is worth more than 16x16 patches: Exploring transformers on individual pixels

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:15.980100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:12.940750Z digest=sha256:dd4c7ca4f3c8c2ce68d35bfbf12a2d8d9da10200cf01ae159e4db1dfe46f9711

Observation 42d1f6b9-685f-462f-9500-ad2e411fcf97 · outbound

This paper cites an unresolved cited work.

GViT: Representing Images as Gaussians for Visual Recognition Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:47:15.875334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:12.965051Z digest=sha256:85b4444b1031dd8a116f1ffb70b6814745d221e7381efecb1d12566e06e5c901

Observation e517d215-d6b7-45cc-b96d-efacca6aeec9 · outbound

This paper cites Rgb no more: Minimally-decoded jpeg vision transformers.

GViT: Representing Images as Gaussians for Visual Recognition Rgb no more: Minimally-decoded jpeg vision transformers

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:15.695604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:12.993161Z digest=sha256:f86eed60226476a39757f7ee5174e8e89f3f7db6098b9af5a5aab9dc59bf1797

Observation 51ff19dd-a76a-45ab-9c72-8799982fd377 · outbound

This paper cites Gaussian Masked Autoencoders.

GViT: Representing Images as Gaussians for Visual Recognition Gaussian Masked Autoencoders

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-08-06T21:47:14.169384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:13.040502Z digest=sha256:04cd676ee9180df3d332992b6025da42b33cbd270b0a666c01ba0a78d16f613f

Observation 326b0d3b-41f2-4a27-9b57-1129e725fce4 · outbound

This paper cites Learning a classification model for segmentation.

GViT: Representing Images as Gaussians for Visual Recognition Learning a classification model for segmentation

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:15.493510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:13.079031Z digest=sha256:20313ff4af5673b8d02d79d80e48f81f887be4439a199a754b2cde08ccbe85f4

Observation 0836e8a5-4ef1-445b-bacb-576c6956ee30 · outbound

This paper cites Very Deep Convolutional Networks for Large-Scale Image Recognition.

GViT: Representing Images as Gaussians for Visual Recognition Very Deep Convolutional Networks for Large-Scale Image Recognition

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:13.101183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:13.101183Z digest=sha256:db354631ee4bf06ac5148fd4fe68658b3610f337f23a72098d4e414e6ff2a90e

Observation c0ea71a1-20eb-4917-9fc1-5c0ec8c287e8 · outbound

This paper cites How to train your vit? data, augmentation, and regularization in vision transformers.

GViT: Representing Images as Gaussians for Visual Recognition How to train your vit? data, augmentation, and regularization in vision transformers

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:15.305740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:13.151121Z digest=sha256:c04b1134511777cfb7cc775c5ae6db23b03eb13a2dfe44d827cd9eea87b4a0de

Observation a4d8a551-bbbe-44e0-a939-83fbef92fc93 · outbound

This paper cites Superpixels: An evaluation of the state-of-the-art.

GViT: Representing Images as Gaussians for Visual Recognition Superpixels: An evaluation of the state-of-the-art

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:13.212064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:13.212064Z digest=sha256:7baf9b5c8ce2f3dd01c78cd68f7da38abb403fa0f8c7846b0dce02d48fcea5aa

Observation 5340c48a-b528-4227-a4de-b16e52df9857 · outbound

This paper cites Single-view view synthesis with multiplane images.

GViT: Representing Images as Gaussians for Visual Recognition Single-view view synthesis with multiplane images

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:15.139091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:13.236571Z digest=sha256:2324aab3879fdcb5f741a1b1b552f8c12e4ecf49d7ed77c034a1a021121f45f2

Observation 635841d9-9702-4ce9-8f0c-7eb7efbe101f · outbound

This paper cites Stacked denoising autoencoders: Learning useful representations in a deep network with a local denoising criterion.

GViT: Representing Images as Gaussians for Visual Recognition Stacked denoising autoencoders: Learning useful representations in a deep network with a local denoising criterion

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:13.279370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:13.279370Z digest=sha256:d46bca9b21c49dfa5a8f0206adcb691a53afc02e60d107db7bf76c1b3363a20e

Observation ffd14c20-ccef-4ae2-ab4f-20b68397693d · outbound

This paper cites Matching networks for one shot learning.

GViT: Representing Images as Gaussians for Visual Recognition Matching networks for one shot learning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:13.354326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:13.354326Z digest=sha256:3c527b0205c2495bcd72240d6240cdb0ed54499f0530821ff96681c0f39fe4e1

Observation e1d282c4-f13d-41c1-801a-ac2fb39a08f0 · outbound

This paper cites Image quality assessment: from error visibility to structural similarity.

GViT: Representing Images as Gaussians for Visual Recognition Image quality assessment: from error visibility to structural similarity

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:13.357382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:13.357382Z digest=sha256:f10992a1d87f880838ba7fed08144a990c4726a19eb850cedb8c3b4a983da729

Observation 06c8f80a-1f85-441c-86cf-7007377b6e37 · outbound

This paper cites Beyond Language Models: Byte Models are Digital World Simulators.

GViT: Representing Images as Gaussians for Visual Recognition Beyond Language Models: Byte Models are Digital World Simulators

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-06T21:47:14.007426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:13.372641Z digest=sha256:33a4be03d3aee33980b224b1dedba35a80f02366d65f9a2d06aac799e5917860

Observation 5135e461-daef-48a1-bcdd-0af3d01951a0 · outbound

This paper cites Learning in the frequency domain.

GViT: Representing Images as Gaussians for Visual Recognition Learning in the frequency domain

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:14.983179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:13.442882Z digest=sha256:71eac4c87bea1c3d753a88dd33d54f029aee2e41eed3921c51074d530a0320b5

Observation c8e13307-e537-4105-9de6-390e916900c0 · outbound

This paper cites gsplat: An open-source library for gaussian splatting.

GViT: Representing Images as Gaussians for Visual Recognition gsplat: An open-source library for gaussian splatting

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:13.546287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:13.546287Z digest=sha256:a7d0dd7318c905fcdea42fb053122bf163fc4634b3d3e3fddea01141b42d6116

Observation 3daf0d46-cb1e-448f-a58b-7f7da4af2f18 · outbound

This paper cites Vector-quantized Image Modeling with Improved VQGAN.

GViT: Representing Images as Gaussians for Visual Recognition Vector-quantized Image Modeling with Improved VQGAN

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:13.608125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:13.608125Z digest=sha256:125353aafbd5db3933d6f2f0f2dacbaba4db7ee47015bd7f5d9de930b6a69fe4

Observation ba8285a7-b6e8-454e-b4eb-2ab2b9d54e0a · outbound

This paper cites A Survey on Masked Autoencoder for Self-supervised Learning in Vision and Beyond.

GViT: Representing Images as Gaussians for Visual Recognition A Survey on Masked Autoencoder for Self-supervised Learning in Vision and Beyond

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:13.697596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:13.697596Z digest=sha256:aa36907aba438e1eab41f97777607512a1241060449735e3b5b270117d40e7f0

Observation 8da1c4bd-7218-4bd1-8d90-9ac99f21af3d · outbound

This paper cites Gaussianimage: 1000 fps image representation and compression by 2d gaussian splatting.

GViT: Representing Images as Gaussians for Visual Recognition Gaussianimage: 1000 fps image representation and compression by 2d gaussian splatting

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:14.800213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:13.794092Z digest=sha256:351af7011dfd047433fecd5cd4a0474ed6206b353c5772dcf9e5a1ce0239c03d

Observation 9d935b03-d3ab-47a7-9c6d-74efc9df2911 · outbound

This paper cites Self-supervised learning of object parts for semantic segmentation.

GViT: Representing Images as Gaussians for Visual Recognition Self-supervised learning of object parts for semantic segmentation

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:14.591969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:47:13.856071Z digest=sha256:e79dbdd539a90ecf15bbc7c9f49b330762432229782d50de574991c8db0467b9

Pith citing papers

No inbound Pith citation observations are available.