PV-VLM fuses visual, textual, and temporal features via a vision-language model and cross-modal attention to improve intra-hour photovoltaic power forecasts by roughly 5 to 9 percent in RMSE and MAE.
Hybrid intrahour DNI forecast model based on DNI measurements and sky-imaging data
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
eess.SP 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
PV-VLM: A Multimodal Vision-Language Approach Incorporating Sky Images for Intra-Hour Photovoltaic Power Forecasting
PV-VLM fuses visual, textual, and temporal features via a vision-language model and cross-modal attention to improve intra-hour photovoltaic power forecasts by roughly 5 to 9 percent in RMSE and MAE.