Pith. sign in

Paper Citation Record · LEDGER

Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2403.09626.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.09626 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T06:00:58.025335Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T09:49:44.610951Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0795fbb6-7485-495a-b009-3bec75629f55 · inbound

Training-Free Zero-Shot Temporal Action Detection with Vision-Language Models cites this paper.

Training-Free Zero-Shot Temporal Action Detection with Vision-Language Models Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:52:34.265033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T04:50:23.280833Z digest=sha256:e481adbf2d70fd49f953de0f6970eeb6033113b324086b3f5789e0851e8a55dc

Observation cf1f094d-3f3e-4ddf-9bdd-4c11961e0a0c · inbound

Bridging Perspectives: A Survey on Cross-view Collaborative Intelligence with Egocentric-Exocentric Vision cites this paper.

Bridging Perspectives: A Survey on Cross-view Collaborative Intelligence with Egocentric-Exocentric Vision Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:58.025335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:58.025335Z digest=sha256:4b978ff3bd7bd745b0832765d1a869cc7bee1e51ec55ace4cb6ebbbd700daa66

Observation d2b5b1d6-47b0-4bf8-afd0-2402e801fb73 · inbound

DySS: Dynamic Queries and State-Space Learning for Efficient 3D Object Detection from Multi-Camera Videos cites this paper.

DySS: Dynamic Queries and State-Space Learning for Efficient 3D Object Detection from Multi-Camera Videos Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T04:35:23.639725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:35:23.639725Z digest=sha256:d6e890a9e69770c53477e722d1431c13548c08816df56cf2bc83fb73535628ed

Observation be5e0b4a-7155-47a8-a2dc-e306aed9dd50 · inbound

M4V: Multimodal Mamba for Efficient Text-to-Video Generation cites this paper.

M4V: Multimodal Mamba for Efficient Text-to-Video Generation Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:21:18.927125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:21:18.927125Z digest=sha256:602ef27da6236bfb430a3ed2efa886667cb67c4f4fb642b756afa457c51e8a34

Observation 6b1b048b-180e-4829-9f5b-7141258a6a96 · inbound

MoMa: Modulating Mamba for Adapting Image Foundation Models to Video Recognition cites this paper.

MoMa: Modulating Mamba for Adapting Image Foundation Models to Video Recognition Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T21:51:50.963310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:51:50.963310Z digest=sha256:a085cdc81ec8e69e4dc93862899d92e10924c6040158befee7886c5013beb186

Observation 05cfcee0-61aa-4dad-95f3-7633ca240c40 · inbound

AuroraLong: Bringing RNNs Back to Efficient Open-Ended Video Understanding cites this paper.

AuroraLong: Bringing RNNs Back to Efficient Open-Ended Video Understanding Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:48.369805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:48.369805Z digest=sha256:7366ec0e64b0da0d328a12c5f029589f4ed9a31455ac19f950aac6493cec18f5

Observation 19ceaf88-34e3-4094-8163-9bbab0b21d83 · inbound

Mamba-OTR: a Mamba-based Solution for Online Take and Release Detection from Untrimmed Egocentric Video cites this paper.

Mamba-OTR: a Mamba-based Solution for Online Take and Release Detection from Untrimmed Egocentric Video Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T15:16:13.073583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:16:13.073583Z digest=sha256:260ce0d09ba76cddfb745994fc530155034ba427bef9f7fcc68c2ab0cb44ce67

Observation 5be75bbd-f14c-4a6f-939b-3b6872765415 · inbound

Animate-X++: Universal Character Image Animation with Dynamic Backgrounds cites this paper.

Animate-X++: Universal Character Image Animation with Dynamic Backgrounds Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-05T21:09:10.288574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:09:10.288574Z digest=sha256:121c353d94daa88e13ad26ab4584e428514627002419f5f8a184863f8cea18da

Observation 546ba720-7323-449a-8a91-93e2b9098c1d · inbound

Time-Scaling State-Space Models for Dense Video Captioning cites this paper.

Time-Scaling State-Space Models for Dense Video Captioning Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T10:59:05.085854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:59:05.085854Z digest=sha256:6ad04f9ff5d54543030b8e8dda8948e782ca32a35fd2f3c91b64f5e09f3905a2

Observation 089c8a17-d886-4d18-b300-e958194cbbb6 · inbound

Exploring Non-Local Spatial-Angular Correlations with a Hybrid Mamba-Transformer Framework for Light Field Super-Resolution cites this paper.

Exploring Non-Local Spatial-Angular Correlations with a Hybrid Mamba-Transformer Framework for Light Field Super-Resolution Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T05:54:55.923858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:54:55.923858Z digest=sha256:3c6d9948fd688129021d0f39830974eed6ff6218412f355648710a65f99ba8cf

Observation 969acd42-9607-4625-87fb-08f2a0dd3b5f · inbound

LADY: Linear Attention for Autonomous Driving Efficiency without Transformers cites this paper.

LADY: Linear Attention for Autonomous Driving Efficiency without Transformers Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T06:42:09.060661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:42:09.060661Z digest=sha256:bf4fe42487cfa1d11353cab6d614f08a8d0a7e726ba5c0b48b9b59788870feb6

Observation 5ac8159f-cf55-444d-9c62-a38a95603963 · inbound

Efficient Spatial-Temporal Focal Adapter with SSM for Temporal Action Detection cites this paper.

Efficient Spatial-Temporal Focal Adapter with SSM for Temporal Action Detection Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:41:13.012494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T17:31:28.091617Z digest=sha256:6df621b185a28878d8b27d90d53deed25bf906772ee88c4d5f54417dcbe8d028

Observation 08153de6-340b-4f8c-bef4-2b1a33711b6f · inbound

ClipTBP: Clip-Pair based Temporal Boundary Prediction with Boundary-Aware Learning for Moment Retrieval cites this paper.

ClipTBP: Clip-Pair based Temporal Boundary Prediction with Boundary-Aware Learning for Moment Retrieval Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:46:27.061349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-07T09:19:24.340363Z digest=sha256:8bf163c6ccc801d1814525d5f41c010f81db3e313e444db8798d7b7522d21d13

Observation e9ea3d6b-ff58-4074-a6d2-5d3187488429 · inbound

MambaADv2: Evolving Duality-enhanced State Space Model for Unsupervised Anomaly Detection cites this paper.

MambaADv2: Evolving Duality-enhanced State Space Model for Unsupervised Anomaly Detection Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding

Reference 94

Resolution
verified exact
arxiv_id, observed 2026-07-04T09:49:44.612383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T09:26:51.456652Z digest=sha256:0a766124efb24aec5bb5d32c48c8010b504dc8cbef713875b4ef18d4b492217b

Observation 60d65879-347f-4703-ae19-0bdd6bbc6dad · inbound

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation cites this paper.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:065d2b30dbdf23e42ce4d6c1b3cf237cfa304d0e47a89bea9c8f2b865689823a