Pith. sign in

REVIEW 2 cited by

Bridging Expertise Gaps: The Role of LLMs in Human-AI Collaboration for Cybersecurity

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2505.03179 v1 pith:4INHJY2P submitted 2025-05-06 cs.CR

classification cs.CR
keywords llmscollaborationcybersecuritydetectionhuman-aiintrusionuserexpertise
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

This study investigates whether large language models (LLMs) can function as intelligent collaborators to bridge expertise gaps in cybersecurity decision-making. We examine two representative tasks-phishing email detection and intrusion detection-that differ in data modality, cognitive complexity, and user familiarity. Through a controlled mixed-methods user study, n = 58 (phishing, n = 34; intrusion, n = 24), we find that human-AI collaboration improves task performance,reducing false positives in phishing detection and false negatives in intrusion detection. A learning effect is also observed when participants transition from collaboration to independent work, suggesting that LLMs can support long-term skill development. Our qualitative analysis shows that interaction dynamics-such as LLM definitiveness, explanation style, and tone-influence user trust, prompting strategies, and decision revision. Users engaged in more analytic questioning and showed greater reliance on LLM feedback in high-complexity settings. These results provide design guidance for building interpretable, adaptive, and trustworthy human-AI teaming systems, and demonstrate that LLMs can meaningfully support non-experts in reasoning through complex cybersecurity problems.

Discussion (0). Sign in to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. From Prediction to Explanation: Multimodal, Explainable, and Interactive Deepfake Detection Framework for Non-Expert Users

    cs.CV 2025-08 conditional novelty 5.0 of 10

    A pipeline combining a deepfake classifier, Grad-CAM heatmaps, image captioning, and an LLM generates layered explanations of deepfake verdicts for non-expert users.

  2. LLMs Are Not Yet Ready for Deepfake Image Detection

    cs.CV 2025-06 conditional novelty 4.0 of 10

    On a 100-image benchmark of real and fake faces, ChatGPT, Claude, Gemini, and Grok all fell short of dependable zero-shot deepfake detection, with accuracy varying by category and model.

Pith tools