Pith. sign in

Paper Citation Record · LEDGER

GeminiFusion: Efficient Pixel-wise Multimodal Fusion for Vision Transformer

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2406.01210.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.01210 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T00:02:11.519487Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-13T19:38:10.424227Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ce3ce5fa-4e7b-41c1-b94b-627a907ea971 · inbound

Efficient Segment Anything with Depth-Aware Fusion and Limited Training Data cites this paper.

Efficient Segment Anything with Depth-Aware Fusion and Limited Training Data GeminiFusion: Efficient Pixel-wise Multimodal Fusion for Vision Transformer

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T00:02:11.519487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T00:02:11.519487Z digest=sha256:778d464351c43d86dfe1fe3ee039bcf8bd3d62fb951d96ba80032b0da9d6615f

Observation 7ee17abe-2af8-4fda-84e3-bf2a74b05ff3 · inbound

CrossWeaver: Cross-modal Weaving for Arbitrary-Modality Semantic Segmentation cites this paper.

CrossWeaver: Cross-modal Weaving for Arbitrary-Modality Semantic Segmentation GeminiFusion: Efficient Pixel-wise Multimodal Fusion for Vision Transformer

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-13T19:38:10.426281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T19:37:18.613746Z digest=sha256:2e0b74f4816c25ae16bb49236e03574587b01b94acb8c63bedfd747ebbf481c0

Observation b120b414-fc30-4265-9788-aab1b94c593b · inbound

RSGMamba: Reliability-Aware Self-Gated State Space Model for Multimodal Semantic Segmentation cites this paper.

RSGMamba: Reliability-Aware Self-Gated State Space Model for Multimodal Semantic Segmentation GeminiFusion: Efficient Pixel-wise Multimodal Fusion for Vision Transformer

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:56:00.617399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T15:46:10.742651Z digest=sha256:548090c9bab9055716ab60c00cc22fa46a8ca5a670675473eaded81677d57d1b

Observation eb0f07e3-7fe5-479a-b366-0fe27808807b · inbound

Structure-Semantic Decoupled Modulation of Global Geospatial Embeddings for High-Resolution Remote Sensing Mapping cites this paper.

Structure-Semantic Decoupled Modulation of Global Geospatial Embeddings for High-Resolution Remote Sensing Mapping GeminiFusion: Efficient Pixel-wise Multimodal Fusion for Vision Transformer

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:16:06.024520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T02:06:05.390586Z digest=sha256:86bd2a3643a71b427af573f6685ae3e61d6c1f1864afc65ebe2396b9ca1fb9ba

Observation 2066fe50-3937-4e4a-b811-fdbd31971b33 · inbound

Weaving Light and Time: Unified Harmonic-Geometric Representation Learning for Dense RGB-Event Parsing cites this paper.

Weaving Light and Time: Unified Harmonic-Geometric Representation Learning for Dense RGB-Event Parsing GeminiFusion: Efficient Pixel-wise Multimodal Fusion for Vision Transformer

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-13T05:07:01.933850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T05:07:01.933850Z digest=sha256:e1b5029afae671c641b669085f89ee788b9277ce0aa9f0a928b6ccf76667b365