Pith. sign in

REVIEW 2 cited by

A Multimodal Intermediate Fusion Network with Manifold Learning for Stress Detection

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2403.08077 v1 pith:JWNBTLGW submitted 2024-03-12 cs.CV cs.AI

classification cs.CVcs.AI
keywords multimodalmethodscomputationalfusionlearningmanifoldreductionaccuracy
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Multimodal deep learning methods capture synergistic features from multiple modalities and have the potential to improve accuracy for stress detection compared to unimodal methods. However, this accuracy gain typically comes from high computational cost due to the high-dimensional feature spaces, especially for intermediate fusion. Dimensionality reduction is one way to optimize multimodal learning by simplifying data and making the features more amenable to processing and analysis, thereby reducing computational complexity. This paper introduces an intermediate multimodal fusion network with manifold learning-based dimensionality reduction. The multimodal network generates independent representations from biometric signals and facial landmarks through 1D-CNN and 2D-CNN. Finally, these features are fused and fed to another 1D-CNN layer, followed by a fully connected dense layer. We compared various dimensionality reduction techniques for different variations of unimodal and multimodal networks. We observe that the intermediate-level fusion with the Multi-Dimensional Scaling (MDS) manifold method showed promising results with an accuracy of 96.00\% in a Leave-One-Subject-Out Cross-Validation (LOSO-CV) paradigm over other dimensional reduction methods. MDS had the highest computational cost among manifold learning methods. However, while outperforming other networks, it managed to reduce the computational cost of the proposed networks by 25\% when compared to six well-known conventional feature selection methods used in the preprocessing step.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Blood Glucose Level Prediction in Type 1 Diabetes Using Machine Learning

    q-bio.QM 2025-01 conditional novelty 4.0 of 10

    A benchmark of 15 machine learning models for 30-minute-ahead glucose prediction on the DiaTrend dataset finds a voting ensemble of MLP, LSTM, and GRU marginally best with an RMSE of 22.50 mg/dL.

  2. Exploring Eye Tracking to Detect Cognitive Load in Complex Virtual Reality Training

    cs.HC 2024-11 conditional novelty 4.0 of 10

    In a 19-participant pilot, an MLP using pupil dilation and fixation duration predicted binary low/high NASA-TLX mental-demand scores in a complex VR assembly task with 84% accuracy.

Pith tools