Pith. sign in

REVIEW 3 cited by

Depthwise Convolution is All You Need for Learning Multiple Visual Domains

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1902.00927 v2 pith:VNQMNISO submitted 2019-02-03 cs.CV

classification cs.CV
keywords domainsdifferentmodelvisualapproachconvolutioncorrelationsdepthwise
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

There is a growing interest in designing models that can deal with images from different visual domains. If there exists a universal structure in different visual domains that can be captured via a common parameterization, then we can use a single model for all domains rather than one model per domain. A model aware of the relationships between different domains can also be trained to work on new domains with less resources. However, to identify the reusable structure in a model is not easy. In this paper, we propose a multi-domain learning architecture based on depthwise separable convolution. The proposed approach is based on the assumption that images from different domains share cross-channel correlations but have domain-specific spatial correlations. The proposed model is compact and has minimal overhead when being applied to new domains. Additionally, we introduce a gating mechanism to promote soft sharing between different domains. We evaluate our approach on Visual Decathlon Challenge, a benchmark for testing the ability of multi-domain models. The experiments show that our approach can achieve the highest score while only requiring 50% of the parameters compared with the state-of-the-art approaches.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. 3D U$^2$-Net: A 3D Universal U-Net for Multi-Domain Medical Image Segmentation

    eess.IV 2019-09 conditional novelty 6.0 of 10

    A single separable-convolution 3D U-Net trained jointly on five segmentation datasets reaches mean Dice within about one point of per-task models while using around 1% of the parameters.

  2. Continual Learning Beyond Experience Rehearsal and Full Model Surrogates

    cs.LG 2025-05 conditional novelty 5.0 of 10

    SPARC achieves strong continual learning accuracy with a fraction of the parameters of surrogate-based methods by combining task-specific depthwise filters with shared pointwise filters updated by exponential averaging.

  3. WhACC: Whisker Automatic Contact Classifier with Expert Human-Level Performance

    cs.CV 2025-01 conditional novelty 5.0 of 10

    WhACC, a two-stage ResNet50V2 and LightGBM classifier with engineered temporal features, matches expert human whisker-touch labeling and reduces curation effort by over 98%.

Pith tools