A new automated benchmark of 2D physics questions finds vision-language models are stronger on formulaic tasks than on spatial reasoning, and parameter count does not fully explain performance.
Construction of a Surrogate Model: Multivariate Time Series Prediction with a Hybrid Model
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
Recent developments of advanced driver-assistance systems necessitate an increasing number of tests to validate new technologies. These tests cannot be carried out on track in a reasonable amount of time and automotive groups rely on simulators to perform most tests. The reliability of these simulators for constantly refined tasks is becoming an issue and, to increase the number of tests, the industry is now developing surrogate models, that should mimic the behavior of the simulator while being much faster to run on specific tasks. In this paper we aim to construct a surrogate model to mimic and replace the simulator. We first test several classical methods such as random forests, ridge regression or convolutional neural networks. Then we build three hybrid models that use all these methods and combine them to obtain an efficient hybrid surrogate model.
citation-role summary
citation-polarity summary
fields
cs.LG 1years
2025 1verdicts
REJECT 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
Interpretable Physics Reasoning and Performance Taxonomy in Vision-Language Models
A new automated benchmark of 2D physics questions finds vision-language models are stronger on formulaic tasks than on spatial reasoning, and parameter count does not fully explain performance.