Pith. sign in

Paper Citation Record · LEDGER

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models

As of 8 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 0 inbound Pith citation observations for arXiv:2509.03837.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.03837 v1

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T10:42:48.527090Z

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

22 of 22 outbound references displayed

  • verified exact0
  • verified fuzzy10
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7661e347-db4e-4686-9caa-9f1ff590eb2b · outbound

This paper cites Artificial General Intelligence (AGI)- Native Wireless Systems: A Journey Beyond 6G,.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Artificial General Intelligence (AGI)- Native Wireless Systems: A Journey Beyond 6G,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:42:50.339169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:42:46.682072Z digest=sha256:1d677688e511566f80ab47efdb2b792a2714fd5c9607167288c615a538a3f4f8

Observation 1eb62e7e-2fe7-4694-a9d1-4df9397e1ca0 · outbound

This paper cites Joint Sensing, Communication, and AI: A Trifecta for Resilient THz User Experiences,.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Joint Sensing, Communication, and AI: A Trifecta for Resilient THz User Experiences,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:42:50.176014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:42:46.741361Z digest=sha256:5da82717524ad595ecff6c7b21f76e7e886e8c543928565b5ea897109e790205

Observation 3955825f-f184-4f05-a21c-3c29af1e81e0 · outbound

This paper cites Sensing-Assisted High Reliable Communication: A Transformer- Based Beamforming Approach,.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Sensing-Assisted High Reliable Communication: A Transformer- Based Beamforming Approach,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:42:50.027458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:42:46.826433Z digest=sha256:64069e5fc3102fb473b6202eafe1797db342d99b0494947e85be8e6a0ed0a60a

Observation 7b419d69-ff08-4621-83d0-3bedcc8fadfd · outbound

This paper cites Multimodal Transformers for Wireless Communications: A Case Study in Beam Prediction.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Multimodal Transformers for Wireless Communications: A Case Study in Beam Prediction

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:46.916383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:46.916383Z digest=sha256:4f8aaa53916685cf8ec06f6bb6784ab469fa42336205ae123dfbbf4e4123d3a6

Observation 5667bcab-5266-4e07-855e-e6878ba03c14 · outbound

This paper cites Vision-Aided 6G Wireless Communications: Blockage Prediction and Proactive Handoff,.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Vision-Aided 6G Wireless Communications: Blockage Prediction and Proactive Handoff,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:42:49.886913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:42:46.979068Z digest=sha256:7b7a0f6fcb68dfa168b00fd459501be55cda9cbab7d14d6f1c15ffd061fd66d1

Observation 697b5786-72d8-47c7-9598-6df083e26354 · outbound

This paper cites Passive Radar at the Roadside Unit to Configure Millimeter Wave Vehicle-to-Infrastructure Links,.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Passive Radar at the Roadside Unit to Configure Millimeter Wave Vehicle-to-Infrastructure Links,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:42:49.746056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:42:47.075053Z digest=sha256:2b305ac05c4f55b680b1346b01c2ac8206bd52e995d81c72bc6963ba5007bcb2

Observation e0a2efc5-5f75-4781-9b54-9e3039174fb5 · outbound

This paper cites BeamLLM: Vision-Empowered mmWave Beam Prediction with Large Language Models.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models BeamLLM: Vision-Empowered mmWave Beam Prediction with Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:47.149512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:47.149512Z digest=sha256:3ac047a75f971a5bd55cbb737da021677c396e324bd4e9ca445c3da3526df765

Observation 8fbda038-5192-4732-bea9-46dde18623d6 · outbound

This paper cites LLM4CP: Adapting Large Language Models for Channel Prediction,.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models LLM4CP: Adapting Large Language Models for Channel Prediction,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:42:49.581498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:42:47.254853Z digest=sha256:ab1f0d1f2590aaddb4f31d3b731faf28b3083bd5387d8428bb78e2f1b834d90e

Observation 0c398674-2cee-4ea3-98f3-984b53cfc4fb · outbound

This paper cites Port-LLM: A Port Prediction Method for Fluid Antenna based on Large Language Models.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Port-LLM: A Port Prediction Method for Fluid Antenna based on Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:47.360252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:47.360252Z digest=sha256:0282d9d4d7e6f4587aadbf834be6a0521369d735b25c35af9b4d88abb6468110

Observation c860dd26-4d5f-4bfb-88ba-c6d787b66330 · outbound

This paper cites Visual Instruction Tuning.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Visual Instruction Tuning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:47.458653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:47.458653Z digest=sha256:1b82bfbced266c5fbfe2016355dee0c798b7318839f8518f2a798a838031478d

Observation b89035b5-0201-4c0c-9ff5-2e2c1b0fb6d1 · outbound

This paper cites Large Language Models Empower Multimodal Integrated Sensing and Communication,.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Large Language Models Empower Multimodal Integrated Sensing and Communication,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:42:49.442427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:42:47.533042Z digest=sha256:ae2917dbebf6c4e72ac315d74f2a1605339584290495a6e0e4b13d7700cc360c

Observation a70ff868-81ef-4b14-b30e-b0358a998655 · outbound

This paper cites Spatial-RAG: Spatial Retrieval Augmented Generation for Real-World Geospatial Reasoning Questions.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Spatial-RAG: Spatial Retrieval Augmented Generation for Real-World Geospatial Reasoning Questions

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:47.639839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:47.639839Z digest=sha256:263e0c6e79ede5e6f98fd4fbacc59954917bffa2edbe705b06392d5e31f65d0f

Observation db04601f-b74a-4f9f-bf04-711e93fad118 · outbound

This paper cites InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:47.749665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:47.749665Z digest=sha256:9c7758d98a1d526afe66732e9d19e5eec550aefbbb6cd4d380e76d6583522855

Observation a5998c3e-1889-4f78-ab29-279feb60abf6 · outbound

This paper cites BEVFusion: Multi-Task Multi-Sensor Fusion with Unified Bird’s-Eye View Representation,.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models BEVFusion: Multi-Task Multi-Sensor Fusion with Unified Bird’s-Eye View Representation,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:42:49.281872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:42:47.879642Z digest=sha256:efd21e8fdfba661dd3a14d7b621126104ce3242cb3c47bcbc39ab564c662fb1c

Observation 0b72d4ed-8679-4e5f-a2b2-d9c4406ad130 · outbound

This paper cites BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:47.959381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:47.959381Z digest=sha256:ab874639a18192b55c6c925f60bebb00c325f79bd0721699cda1a2127c916f1b

Observation de258daa-0b44-4e0d-95ac-a9fccce2c92d · outbound

This paper cites CARLA: An Open Urban Driving Simulator.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models CARLA: An Open Urban Driving Simulator

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:48.041403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:48.041403Z digest=sha256:61f8b73ae01f0e7e902c41450095e4858a0c7d3ae6954e3a7d39fb0dfbb23a87

Observation 35bc6856-6a6c-435f-8c9c-baacca2566c0 · outbound

This paper cites an unresolved cited work.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:48.123049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:48.123049Z digest=sha256:e9c91a11d29dc3f4c924f43471392677a09c8947b54d018393edaa252e8f01c6

Observation f5ce4959-c9c0-45fb-92f9-845783b3d391 · outbound

This paper cites Llama 3.2: Revolutionizing edge AI and vision with open, customiz- able models.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Llama 3.2: Revolutionizing edge AI and vision with open, customiz- able models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:42:49.131525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:42:48.203558Z digest=sha256:01d40e3a6cd2e93256b0b79a8b879b071f66fc6f12d8526df938eb680e5e0ef6

Observation 202cb994-76c8-49af-ba6d-910b44e7c37a · outbound

This paper cites Decoupled Weight Decay Regularization.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Decoupled Weight Decay Regularization

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:48.288180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:48.288180Z digest=sha256:f0b91c98715f651f18b62ccf849f131297e29a52b9855a95c00162d66390421f

Observation acc90f7d-326a-4382-b3af-da66c481179d · outbound

This paper cites Long Short-Term Memory,.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Long Short-Term Memory,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:42:48.986723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:42:48.375998Z digest=sha256:8f6a7965016ba27ddfec502b45bf7af8266000317755c1b13b3dd1febdf66492

Observation 64e6b651-f7b1-4f91-907c-3fabaf230c5f · outbound

This paper cites Empirical Evaluation of Gated Recurrent Neural Networks on Sequence Modeling.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Empirical Evaluation of Gated Recurrent Neural Networks on Sequence Modeling

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:48.461598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:48.461598Z digest=sha256:9590880215291ad15e8c1445f142c7f8a95ae8ac31ac83b353c11408b102006b

Observation 33df7c63-133b-45d6-bb14-8a8156e47940 · outbound

This paper cites Attention Is All You Need.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Attention Is All You Need

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:48.527090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:48.527090Z digest=sha256:44634e00e95870e575aaa1d25ef7791a70c2c04d8590b93623fd92ec788d9605

Pith citing papers

No inbound Pith citation observations are available.