Pith. sign in

REVIEW

MultiZoo & MultiBench: A Standardized Toolkit for Multimodal Deep Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2306.16413 v1 pith:WPTEO2TE submitted 2023-06-28 cs.LG cs.AIcs.CLcs.CVcs.MM

classification cs.LGcs.AIcs.CLcs.CVcs.MM
keywords multimodallearningmultibenchdataensuringevaluationmodalitiesmultizoo
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Learning multimodal representations involves integrating information from multiple heterogeneous sources of data. In order to accelerate progress towards understudied modalities and tasks while ensuring real-world robustness, we release MultiZoo, a public toolkit consisting of standardized implementations of > 20 core multimodal algorithms and MultiBench, a large-scale benchmark spanning 15 datasets, 10 modalities, 20 prediction tasks, and 6 research areas. Together, these provide an automated end-to-end machine learning pipeline that simplifies and standardizes data loading, experimental setup, and model evaluation. To enable holistic evaluation, we offer a comprehensive methodology to assess (1) generalization, (2) time and space complexity, and (3) modality robustness. MultiBench paves the way towards a better understanding of the capabilities and limitations of multimodal models, while ensuring ease of use, accessibility, and reproducibility. Our toolkits are publicly available, will be regularly updated, and welcome inputs from the community.

Discussion (0). Continue with ORCID to comment.

Pith tools