A Dockerized Kafka and Spark pipeline sustains roughly 320,000 synthetic records per minute and reports Random Forest congestion classification with macro F1 above 0.95, but the labels come from KMeans on the same features, so the prediction claim is circular.
Parallel computing for large-scale traffic simulation: A review
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.DC 1years
2025 1verdicts
REJECT 1representative citing papers
citing papers explorer
-
CityPulse: Real-Time Traffic Data Analytics and Congestion Prediction
A Dockerized Kafka and Spark pipeline sustains roughly 320,000 synthetic records per minute and reports Random Forest congestion classification with macro F1 above 0.95, but the labels come from KMeans on the same features, so the prediction claim is circular.