A decomposed 3D convolution encoder for chest CT reaches competitive pathology detection and image-text retrieval with far fewer parameters and FLOPs than transformer or full 3D convolution baselines.
Handwritten digit recognition with a back-propagation network
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
DCFormer: Efficient 3D Vision-Language Modeling with Decomposed Convolutions
A decomposed 3D convolution encoder for chest CT reaches competitive pathology detection and image-text retrieval with far fewer parameters and FLOPs than transformer or full 3D convolution baselines.