REVIEW 2 cited by
Multimodal Deep Learning-Empowered Beam Prediction in Future THz ISAC Systems
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Integrated sensing and communication (ISAC) systems operating at terahertz (THz) bands are envisioned to enable both ultra-high data-rate communication and precise environmental awareness for next-generation wireless networks. However, the narrow width of THz beams makes them prone to misalignment and necessitates frequent beam prediction in dynamic environments. Multimodal sensing, which integrates complementary modalities such as camera images, positional data, and radar measurements, has recently emerged as a promising solution for proactive beam prediction. Nevertheless, existing multimodal approaches typically employ static fusion architectures that cannot adjust to varying modality reliability and contributions, thereby degrading predictive performance and robustness. To address this challenge, we propose a novel and efficient multimodal mixture-of-experts (MoE) deep learning framework for proactive beam prediction in THz ISAC systems. The proposed multimodal MoE framework employs multiple modality-specific expert networks to extract representative features from individual sensing modalities, and dynamically fuses them using adaptive weights generated by a gating network according to the instantaneous reliability of each modality. Simulation results in realistic vehicle-to-infrastructure (V2I) scenarios demonstrate that the proposed MoE framework outperforms traditional static fusion methods and unimodal baselines in terms of prediction accuracy and adaptability, highlighting its potential in practical THz ISAC systems with ultra-massive multiple-input multiple-output (MIMO).
Forward citations
Cited by 2 Pith papers
-
M2BeamLLM: Multimodal Sensing-empowered mmWave Beam Prediction with Large Language Models
M2BeamLLM combines four sensing modalities with a lightly fine-tuned GPT-2 backbone and reports 68.9% top-1 beam prediction accuracy on DeepSense 6G Scenario 32, beating the compared baselines by up to 13.9 percentage points.
-
Data-Free Knowledge Distillation for LiDAR-Aided Beam Tracking in MmWave Systems
A data-free knowledge distillation method trains a compact student for LiDAR-aided mmWave beam tracking entirely on synthetic data generated from a teacher's statistics, nearly matching teacher accuracy.
Discussion (0). Sign in to comment.