Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:09:09.912975Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 71 of 71 outbound references and 0 inbound Pith citation observations for arXiv:2505.19938.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:09:09.912975Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
71 of 71 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 56943cd2-bb7d-40db-8ce6-ad0560ce3393 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Contrastive masked autoencoders are stronger vision learn- ers,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2b640d07-4292-42d4-ace2-7d59051f7b50 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Pyramid constrained self-attention network for fast video salient object detection,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8ef2a54b-4207-4fd3-88b8-a0049b3c53d4 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Cap4video: What can auxiliary captions do for text-video retrieval?
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 81e4713e-a90c-4f92-9e21-5d16d6c5f391 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Isomer: Isomerous transformer for zero-shot video object segmentation,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 9774c393-8a97-4a22-bcca-4e2128547b1e · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Avgzslnet: Audio-visual generalized zero-shot learning by reconstructing label fea- tures from multi-modal embeddings,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 96dfe273-d3b7-44f3-b65f-168020578767 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Audio-visual generalised zero-shot learning with cross-modal attention and language,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5374bc77-da0d-4a67-99d4-261c17547950 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Temporal and cross-modal attention for audio-visual zero-shot learning,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 72eadc1a-7ed2-45af-b8bf-4ac01f2ebc22 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Enhancing unsupervised video representation learning by decoupling the scene and the motion,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d294de1c-d33d-434d-8345-dfd0b1d232df · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Spiking tucker fusion trans- former for audio-visual zero-shot learning,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d22828d8-7513-4d2c-9f2f-ea4dd1827329 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Motion- decoupled spiking transformer for audio-visual zero-shot learning,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation cd4268d7-8fc9-4232-b1dc-3848f2550c48 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Conv2former: A simple transformer-style convnet for visual recognition,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8d4b895f-2d50-49f8-8933-91b75b429e17 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Bidi- rectional cross-modal knowledge exploration for video recognition with pre-trained vision-language models,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 152c6262-d221-4388-88d4-ec2ce45aa4ef · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Box2mask: Box-supervised instance segmentation via level-set evolution,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation bc6f5344-5385-4392-8b9c-c5f25606c707 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning The devil is in the crack orientation: A new perspective for crack detection,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 9c838a6c-1fa3-437c-b871-9888bd701059 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Offline and online optical flow enhancement for deep video compression,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation e0d569d4-dc6e-496f-90bf-fb5eca6c6ede · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Ustc-td: A test dataset and benchmark for image and video coding in 2020s,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 9a8775c8-91b0-4118-b3b2-a4ddc6f22524 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Object segmentation- assisted inter prediction for versatile video coding,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation bc236cc0-2058-43a1-a884-2c063075f2b1 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Geometry-aware guided loss for deep crack recognition,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 335a8d7e-b2a8-4b77-a304-5a584cda06ca · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Latent embedding feedback and discriminative features for zero-shot classifi- cation,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation aeac26fd-b60c-43a0-85e7-f85d70e3ffdf · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Gener- alized zero-and few-shot learning via aligned variational autoencoders,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 4890ef3a-e3fc-492f-a6f5-431316bf26d1 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Generalized zero- shot learning via synthesized examples,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation fa02b7cb-7af7-4bb7-a696-d85af6a46754 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Feature generating networks for zero-shot learning,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 96beda2a-9a91-4ed1-a624-f49bfb6cec95 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning A gener- ative adversarial approach for zero-shot learning from noisy texts,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8eb83658-a615-48bd-863b-b75ee70eff8f · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Bootstrapping audio-visual video segmentation by strengthening audio cues,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 32f96bad-46e4-402f-9eb1-b8924ee198ce · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Learning affective features with a hybrid deep model for audio–visual emotion recognition,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation e0feff2f-b320-4bc7-8f0e-7ddc95a4d5b4 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Audio-visual temporal forgery de- tection using embedding-level fusion and multi-dimensional contrastive loss,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f40ea82-1016-437e-9738-3f96262091e9 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Multimodal imbalance-aware gradient modulation for weakly-supervised audio-visual video parsing,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 99c8f7ca-4ecb-45b9-9098-c0024099dac6 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Question-aware global-local video understanding network for audio-visual question answering,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8598ebb1-6c3c-4604-bb54-bdf61f09a116 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Zero-shot audio classification via semantic embeddings,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ca30d387-47f6-4d08-bffd-b47bf6915d08 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Activitynet: A large-scale video benchmark for human activity under- standing,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b763a99f-a635-4215-876f-5ec52e37bf42 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning The multivehicle stereo event camera dataset: An event camera dataset for 3d perception,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1d3db10-a6c6-4a30-92b2-24bf19a37cb6 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Events-to-video: Bringing modern computer vision to event cameras,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ce3f04ee-9d93-4e6b-b5f4-d1e7d154b55d · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning High speed and high dynamic range video with an event cam- era,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 168ebb16-0ac6-4d13-b09d-39f8574077ec · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning A controlled-delay event camera framework for on-line robotics,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 790bfafb-2384-473b-919f-fa664d7e7f34 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Exploring event camera-based odome- try for planetary robots,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2bfae8a2-9561-414e-830a-3820729395b6 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Emergent visual sensors for autonomous vehicles,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a201dff8-bc2f-4aa4-9654-00b954b1b25b · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Dsec: A stereo event camera dataset for driving scenarios,
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04cb1f09-c62d-4ee6-bed6-1eaab6adbf21 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning High frame rate video reconstruction based on an event camera,
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0abe025d-b50e-45ab-997e-152fb4cffb0b · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Eventcap: Monocular 3d capture of high-speed human motions using an event camera,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ca3e4702-39f3-48f1-a5e6-f5a483e1bc4d · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Towards a framework for end-to-end control of a simulated vehicle with spiking neural networks,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8ad00811-e3d4-4fd9-9bbe-b2fcd6bb35fd · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Esim: an open event camera simulator,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 50984320-2288-4ec0-bdfd-0b7b540096be · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Deep residual learning in spiking neural networks,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 1908f1e8-a23c-43e1-8215-0962b04c3ad5 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Neuron-based spiking transmission and reasoning network for robust image-text retrieval,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8ea97e88-dfc0-404d-bb79-fc02000280f3 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Spikemba: Multi-modal spiking saliency mamba for temporal video grounding,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 46043712-dc94-4421-871e-5663e2582dac · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Spikformer: When Spiking Neural Network Meets Transformer
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ba149dc-de50-4140-a96c-a68f0935e6ce · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Learning optical flow from continuous spike streams,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 66f12683-fef4-46c6-a877-3248c9c04f75 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Mrdflow: Unsupervised optical flow estimation network with multi-scale recurrent decoder,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 278be33e-297f-4b06-8070-7a744574ee54 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Progressive tandem learning for pattern recognition with deep spiking neural networks,
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fcc8209-16b4-474f-b5b7-d1365f29a074 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning A hybrid neural coding approach for pattern recognition with spiking neural networks,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation db5e6eba-a867-4b5f-9c9b-9b7ae219c9e0 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Neuron-based spiking transmission and reasoning network for robust image-text retrieval,
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6fd95f30-7122-4b75-a031-d4ac0e3c045b · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Multi-scale spiking pyramid wireless communication framework for food recognition,
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation afecb359-12b1-45c5-a368-c58306b4cdd9 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Modality-fusion spiking transformer network for audio-visual zero-shot learning,
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c347e535-9f45-417b-96d3-32f7a3f910ff · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Evolving spiking neural network controllers for autonomous robots,
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d514cb12-ea7f-4377-af39-8414dbc4b048 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning A hybrid rein- forcement learning approach with a spiking actor network for efficient robotic arm target reaching,
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6197bdc3-8273-4f63-9980-0717138746e5 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Labelling unla- belled videos from scratch with multi-modal self-supervision,
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 62c2a68e-2068-42ef-95c3-d5731f50b2c0 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Training deep spiking neural networks using backpropagation,
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 0a7f348c-7b08-4a95-879d-6c975b739301 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73e029cc-5e64-408c-ba90-87b5c3841221 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Vggsound: A large-scale audio-visual dataset,
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c9ebbbad-7867-4301-93db-4ea49adc30c6 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Evaluation of out- put embeddings for fine-grained image classification,
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2817b972-4b79-4cc2-9ccf-408ce2df92f9 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Devise: A deep visual-semantic embedding model,
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f3da65f5-da96-4ccb-813c-e0b2ff9074dd · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Attribute proto- type network for zero-shot learning,
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5c73589e-bba8-44ac-92fd-0a93f7fb45cc · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning f-vaegan-d2: A feature generating framework for any-shot learning,
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 592f091a-326b-4d8e-95ba-2544a116faf8 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Coordinated joint multimodal embeddings for generalized audio-visual zero-shot classifi- cation and retrieval of videos,
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 9d9254a5-804d-4161-9e73-79ecb95ecf07 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Hyperbolic audio-visual zero-shot learning,
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation eb881181-73c0-4686-95a0-5754709a04f3 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Label-embedding for image classification,
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation e60af83a-2dac-4304-a94f-50d39e6e2620 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Coordinated joint multimodal embeddings for generalized audio-visual zero-shot classifi- cation and retrieval of videos,
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 13fd456c-432e-403a-8ef0-0484e6d60668 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Learning spatiotemporal features with 3d convolutional networks,
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 961118c7-1ec5-4ad6-a11e-f40eea78e88b · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Large-scale video classification with convolutional neural networks,
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation cc69ac4b-8d18-4d93-8c3b-98839b187297 · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Cnn architectures for large-scale audio classification,
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7a6f9a88-c432-455e-912e-86a7ffa889ae · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning Youtube-8m: A large-scale video classi- fication benchmark,
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c083aee1-8488-41d6-ac1f-46f9eb635c3d · outbound
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning degree with the School of Computer Science, Harbin Institute of Technology, Harbin, China
Reference 2020
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
No inbound Pith citation observations are available.