Pith. sign in

REVIEW 5 cited by

HyperSIGMA: Hyperspectral Intelligence Comprehension Foundation Model

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2406.11519 v2 pith:5WMWJB7Q submitted 2024-06-17 cs.CV eess.IV

classification cs.CVeess.IV
keywords hypersigmahyperspectralspectraltasksacrossapplicationscapabilityexisting
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Accurate hyperspectral image (HSI) interpretation is critical for providing valuable insights into various earth observation-related applications such as urban planning, precision agriculture, and environmental monitoring. However, existing HSI processing methods are predominantly task-specific and scene-dependent, which severely limits their ability to transfer knowledge across tasks and scenes, thereby reducing the practicality in real-world applications. To address these challenges, we present HyperSIGMA, a vision transformer-based foundation model that unifies HSI interpretation across tasks and scenes, scalable to over one billion parameters. To overcome the spectral and spatial redundancy inherent in HSIs, we introduce a novel sparse sampling attention (SSA) mechanism, which effectively promotes the learning of diverse contextual features and serves as the basic block of HyperSIGMA. HyperSIGMA integrates spatial and spectral features using a specially designed spectral enhancement module. In addition, we construct a large-scale hyperspectral dataset, HyperGlobal-450K, for pre-training, which contains about 450K hyperspectral images, significantly surpassing existing datasets in scale. Extensive experiments on various high-level and low-level HSI tasks demonstrate HyperSIGMA's versatility and superior representational capability compared to current state-of-the-art methods. Moreover, HyperSIGMA shows significant advantages in scalability, robustness, cross-modal transferring capability, real-world applicability, and computational efficiency. The code and models will be released at https://github.com/WHU-Sigma/HyperSIGMA.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 5 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. SpectralX: Parameter-efficient Domain Generalization for Spectral Remote Sensing Foundation Models

    cs.CV 2025-08 reject novelty 6.0 of 10

    SpectralX adapts RGB-pretrained remote sensing models to spectral data with a two-stage parameter-efficient training scheme and reports state-of-the-art cross-domain segmentation on three benchmarks.

  2. M-SpecGene: Generalized Foundation Model for RGBT Multispectral Vision

    cs.CV 2025-07 conditional novelty 6.0 of 10

    M-SpecGene is a Siamese masked-autoencoder foundation model for RGB-thermal vision, trained on the RGBT550K dataset with a GMM-CMSS progressive masking strategy, and evaluated on four downstream tasks.

  3. MAPEX: Modality-Aware Pruning of Experts for Remote Sensing Foundation Models

    cs.CV 2025-07 conditional novelty 6.0 of 10

    MAPEX shows that a modality-conditioned mixture-of-experts vision transformer, pre-trained on six remote sensing modalities and then pruned to keep only the experts for a target modality, can outperform or match large...

  4. MergeSAM: Unsupervised change detection of remote sensing images based on the Segment Anything Model

    cs.CV 2025-07 conditional novelty 5.0 of 10

    A SAM-based unsupervised change detection method that matches and splits segmentation masks across two dates, improving F1 over AnyChange on GZ_CD_data.

  5. Parameter-Efficient Fine-Tuning of Multispectral Foundation Models for Hyperspectral Image Classification

    cs.CV 2025-05 conditional novelty 4.0 of 10

    KronA+ fine-tunes SpectralGPT for hyperspectral image classification using only 0.056% trainable parameters and reaches accuracy close to full fine-tuning on five public datasets.

Pith tools