Pith. sign in

Paper Citation Record · LEDGER

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization

As of 18 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 0 inbound Pith citation observations for arXiv:2412.06208.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.06208 v1

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T20:00:45.115183Z

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

30 of 30 outbound references displayed

  • verified exact2
  • verified fuzzy18
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f0035939-76b1-4ec3-92c0-f8cee4b8d973 · outbound

This paper cites Towards a theory of semantic communication,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Towards a theory of semantic communication,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T20:00:44.906424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:00:44.906424Z digest=sha256:d43a19d58413f4bc21f745528d698344c601bf756f2c1490217a4de9a8c7a721

Observation 6e1c190a-4def-47e9-83b2-bb683036e627 · outbound

This paper cites 6g networks: Beyond shannon towards semantic and goal-oriented communications,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization 6g networks: Beyond shannon towards semantic and goal-oriented communications,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T20:00:44.914309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:00:44.914309Z digest=sha256:749d3743acb2293a3f132274b89fe799382421e244ee224539456268ad3520a5

Observation ab301f53-8a24-430c-8cf2-0525675752ca · outbound

This paper cites Semantic communications: Overview, open issues, and future research directions,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Semantic communications: Overview, open issues, and future research directions,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T20:00:44.919706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:00:44.919706Z digest=sha256:8673104241f2a41b584d85ac6b59e1c148a9302789083d6ff70d7873c0f00488

Observation 0c8b310c-db51-44c5-97f7-768ac764fa9a · outbound

This paper cites A unified multi-task semantic communication system for multimodal data,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization A unified multi-task semantic communication system for multimodal data,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:00:45.701038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:00:44.924890Z digest=sha256:0259d3286bfb77015adf344f9c6414c999dd8cbd3468cb317115714b213e14bd

Observation 3b977f12-e0b7-41a7-b692-156790a54f00 · outbound

This paper cites Less data, more knowledge: Building next generation semantic communication networks,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Less data, more knowledge: Building next generation semantic communication networks,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T20:00:44.930134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:00:44.930134Z digest=sha256:70e00e5fa11224b24ea186ea8c97ed298db6681ecd346b7a7e5be269dc0b4420

Observation 96086142-4ad9-4a61-aeda-54af4a5aedd9 · outbound

This paper cites Multimodal semantic communication accelerated bidirectional caching for 6g mec,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Multimodal semantic communication accelerated bidirectional caching for 6g mec,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:00:45.668004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:00:44.938999Z digest=sha256:82a91666d86c663440937e23588ef751e8e54e4ebcea18964f6839dc07f62b6a

Observation 3af7c1a3-ab79-465e-b19e-74edfc77631b · outbound

This paper cites Multimodal and multiuser semantic communications for channel-level information fusion,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Multimodal and multiuser semantic communications for channel-level information fusion,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:00:45.647843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:00:44.945718Z digest=sha256:5f98a5877bacd8c684a5a5208d0cca9336b877df48ae1a6df905a9fad8facd9f

Observation 13455fc6-1db5-4637-9acc-23aa2e587b39 · outbound

This paper cites Distributed semantic communications for multimodal audio-visual parsing tasks,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Distributed semantic communications for multimodal audio-visual parsing tasks,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:00:45.583885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:00:44.965872Z digest=sha256:e00cab763414506527c802d2fdc069d41c5635ed1ed5b977623f071611c622c7

Observation b7830c30-6e7f-4dcc-8663-5ead0ac4fc4e · outbound

This paper cites Hybrid Digital-Analog Semantic Communications.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Hybrid Digital-Analog Semantic Communications

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T20:00:44.972971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:00:44.972971Z digest=sha256:17dae1da220ec44323de4f4d2197461a139f52baa9fdacdc6c2218cb0aaee489

Observation 856af302-ba39-436e-9038-b664d751ae82 · outbound

This paper cites Task-Oriented Multi-User Semantic Communications for VQA Task.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Task-Oriented Multi-User Semantic Communications for VQA Task

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-11T20:00:45.242536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:00:44.979262Z digest=sha256:28f7a7d04603da1aabf49e4da4bdc0a2f392b1c65e3fba5d4c74cc44000595dd

Observation 2db5b3eb-96f4-4cbb-a89f-02a3b77ecd26 · outbound

This paper cites Task-oriented multi- user semantic communications,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Task-oriented multi- user semantic communications,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T20:00:44.984342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:00:44.984342Z digest=sha256:de34a09a0b3dfa02ac39d45f445db4ccddbea8efe4084fdd2dd937282efbb10e

Observation 23fb147c-62cd-414e-a4cb-639e4fefc2d7 · outbound

This paper cites Content-aware semantic communica- tion for goal-oriented wireless communications,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Content-aware semantic communica- tion for goal-oriented wireless communications,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:00:45.556997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:00:44.989726Z digest=sha256:9198b60df9d55eee00eed2885f35064a775622c6d8b3f4253ccb2ecb4a4b143e

Observation 2f066ca6-892e-4f5a-992c-090f57cd89f5 · outbound

This paper cites AdaSem: Adaptive Goal-Oriented Semantic Communications for End-to-End Camera Relocalization.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization AdaSem: Adaptive Goal-Oriented Semantic Communications for End-to-End Camera Relocalization

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-11T20:00:45.201794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:00:44.996081Z digest=sha256:946e688f1c5530bf3f91c46bea2676fcdf6a082dd2d598f87c450d9ec8d93b8e

Observation 0d4502a1-18f1-410d-97cc-14bc66e72af7 · outbound

This paper cites Wireless resource management in intelligent semantic communication networks,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Wireless resource management in intelligent semantic communication networks,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:00:45.539505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:00:45.001439Z digest=sha256:e7e0b384beae223c42e426240cd0e4d2050f0eab4c09d371b3389dd415e8b812

Observation 66452566-df5e-48a5-9075-baa246bd6bfd · outbound

This paper cites Semantic and effective communication for remote control tasks with dynamic feature compression,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Semantic and effective communication for remote control tasks with dynamic feature compression,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:00:45.524145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:00:45.006499Z digest=sha256:767c9c5d34f7f5534d2692d72b3d860883827baeb31244363d042e2ef5fe31e1

Observation c6ea22c8-1e73-4440-b7da-905015326e8d · outbound

This paper cites Cross-modal semantic communi- cations,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Cross-modal semantic communi- cations,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:00:45.491280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:00:45.013277Z digest=sha256:912393672fea362a091452ab58fa81972589c1b706e719d988e392da14d7bf7d

Observation af61bd9b-6147-4dc7-bcba-bbc0790046bc · outbound

This paper cites Audio-visual event localization in unconstrained videos,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Audio-visual event localization in unconstrained videos,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:00:45.629093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:00:45.026660Z digest=sha256:7b256c6ffdb39e60a7e245a1c21ea42e1b8d932134ae237647c88eee826a31d2

Observation 27e2c6ab-b992-49b6-b94b-38276bb8dc19 · outbound

This paper cites Dual-modality seq2seq net- work for audio-visual event localization,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Dual-modality seq2seq net- work for audio-visual event localization,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:00:45.468425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:00:45.033172Z digest=sha256:e0e2fb190c110656d3fb359cfd3fc22f9523f3e5a06d22aa093c61ad0f2d3681

Observation 791b8ea2-8c61-40e6-9051-1a529965a22c · outbound

This paper cites Dual attention matching for audio- visual event localization,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Dual attention matching for audio- visual event localization,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:00:45.452247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:00:45.046449Z digest=sha256:7619d2bc898c628565130a6c93804c2a966757fd30d62587cafc4a68d46adf16

Observation 1e128d22-69bb-4bd3-aa26-877e1f31d4aa · outbound

This paper cites Cross-modal relation- aware networks for audio-visual event localization,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Cross-modal relation- aware networks for audio-visual event localization,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:00:45.425909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:00:45.053275Z digest=sha256:dda42fecc8a49c3f52250dc40f42bf5d215bea70af4bc34e8506d0fc09c33e46

Observation 334f9bd1-fb92-471d-b655-b750641abd35 · outbound

This paper cites Audio- visual event localization via recursive fusion by joint co-attention,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Audio- visual event localization via recursive fusion by joint co-attention,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:00:45.397865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:00:45.057568Z digest=sha256:69d58ca3fb5f1fd562a6626fc5f526b8d37939a8af1a891c635d618dd772b7b3

Observation 2c88d11e-72e0-4112-9a05-c7c81a21f733 · outbound

This paper cites Cross-modal attention network for temporal inconsistent audio-visual event localization,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Cross-modal attention network for temporal inconsistent audio-visual event localization,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:00:45.376320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:00:45.062348Z digest=sha256:449919fdf089c91829be41937442ade5d1a74ac2090fc29e7852b06d92187091

Observation 2e7edaf1-edb0-4bd1-863a-ba894926af19 · outbound

This paper cites Positive sample propagation along the audio-visual event line,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Positive sample propagation along the audio-visual event line,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:00:45.609344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:00:45.067708Z digest=sha256:f35436c997d5a5fcd38f159f66b7982781cd4eb727ca38f7ad6e97c6c7f72b36

Observation cf4e2547-7e9b-4ca4-9111-0f4c52d79245 · outbound

This paper cites Discriminative cross-modality attention network for temporal inconsistent audio-visual event localization,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Discriminative cross-modality attention network for temporal inconsistent audio-visual event localization,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T20:00:45.072423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:00:45.072423Z digest=sha256:d81266a2feca9797f9f4c19486599361b4aa67ce3c0f3f6ceadc8929329ba5ea

Observation 20991b18-85a2-4596-8b4f-40257d8f4e58 · outbound

This paper cites Audiovisual transformer with instance attention for audio-visual event localization,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Audiovisual transformer with instance attention for audio-visual event localization,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:00:45.344183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:00:45.078949Z digest=sha256:ae22cb8879a9ea3f5453a8641c44dbbbc746886afd288162ef8e18e1efc60dda

Observation b8a508c0-b700-4aed-94c1-d0df5be7f458 · outbound

This paper cites Very Deep Convolutional Networks for Large-Scale Image Recognition.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Very Deep Convolutional Networks for Large-Scale Image Recognition

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T20:00:45.084248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:00:45.084248Z digest=sha256:c7423cc696026bb2311c1b7bda5ad8ca345f4422f6edb548106bcd848b3494f6

Observation c23b6528-b314-4382-879b-cfff6afb103a · outbound

This paper cites Eulernet: Adaptive feature interaction learning via euler’s formula for ctr prediction,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Eulernet: Adaptive feature interaction learning via euler’s formula for ctr prediction,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:00:45.329181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:00:45.094192Z digest=sha256:48825a94cb21079cd7ffc80ac3bd2ad98d1c5ce4a3aee53e61eaa004c3d03722

Observation e7c16729-9406-4cde-83f2-f4357cfa744b · outbound

This paper cites Imagenet classification with deep convolutional neural networks,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Imagenet classification with deep convolutional neural networks,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T20:00:45.101082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:00:45.101082Z digest=sha256:b756a87be6aa508b9dea686b0fa3b5252b3d85596fadfdbfddd183bcb4888dfa

Observation 4249d608-4d62-44f1-a133-a8886b9e48a6 · outbound

This paper cites Cnn architectures for large-scale audio classification,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Cnn architectures for large-scale audio classification,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:00:45.297298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:00:45.106592Z digest=sha256:079705ea6985b9d891e8a074b77109f8f20c93a9b90f97878be695168a882ba6

Observation 28aba50c-a367-4577-8ad0-35dea7faa297 · outbound

This paper cites Audio set: An ontology and human- labeled dataset for audio events,.

Pilot-guided Multimodal Semantic Communication for Audio-Visual Event Localization Audio set: An ontology and human- labeled dataset for audio events,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T20:00:45.115183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:00:45.115183Z digest=sha256:3758f2bc6cdacbc7c13d2f97396bfa0b3c66bc07474ede11f0781ef73350c532

Pith citing papers

No inbound Pith citation observations are available.