VT-Bench aggregates 14 datasets from 9 domains and evaluates 23 models to standardize visual-tabular discriminative and generative tasks.
arXiv preprint arXiv:2111.11665 , year=
2 Pith papers cite this work. Polarity classification is still indexing.
2
Pith papers citing it
years
2026 2verdicts
UNVERDICTED 2representative citing papers
Benchmarks efficient MLLMs on eight PE QA tasks from the INSPECT dataset, finding stronger results with combined CTPA+EHR inputs and for diagnosis over prognosis.
citing papers explorer
-
VT-Bench: A Unified Benchmark for Visual-Tabular Multi-Modal Learning
VT-Bench aggregates 14 datasets from 9 domains and evaluates 23 models to standardize visual-tabular discriminative and generative tasks.
-
Efficient Multimodal Clinical Question Answering for Pulmonary Embolism Risk Assessment
Benchmarks efficient MLLMs on eight PE QA tasks from the INSPECT dataset, finding stronger results with combined CTPA+EHR inputs and for diagnosis over prognosis.