{"id":"4c94b6f0-e7eb-4b24-b6a5-f8eaf33c05d6","arxiv_id":"2504.15429","paper_version":1,"verdict":"CONDITIONAL","confidence":"MODERATE","novelty_score":5.0,"correctness_risk":"medium","formal_verification":"none","parameter_count":0,"one_line_summary":"Trigger and content warning use on social media is shaped by a three-way tension among viewers' informed choice, posters' reach, and platform affordances, shown through 15 interviews.","lead":"This paper interviewed 15 US social media users about how they view, ignore, and add trigger warnings and content warnings. It maps the struggles of viewers, posters, and platforms into a design framework for more trauma-informed social media.","discovery_kind":"extension","skeptic_critique":{"model":"deepseek-v4-flash","headline":"Internal inconsistency: first-cohort participant P04 is quoted reacting to a slide deck that Section 3.2 says was introduced only for the second cohort (P7-P15).","rationale":"The reader identifies the self-selected, non-representative sample as the weakest assumption. That is a legitimate external-validity concern, and the paper's own Section 5.6 acknowledges it. However, for the paper's central claim—a conceptual framework of viewer/poster/platform TW/CW mechanisms grounded in 15 interviews—the more load-bearing issue is internal consistency. The methods section describes a substantive protocol revision after the first six interviews, including a new slide deck and a shift from TW-vs-CW distinctions to viewer-versus-poster dynamics. The results nonetheless report pooled prevalence counts and quotes, and at least one quote attributed to first-cohort participant P04 refers to slide-deck content that, according to Section 3.2, should not have been presented to that cohort. This is either a misattribution, a mislabeled figure, or an undocumented deviation from the stated protocol, and any of these possibilities weakens trust in the data integration. If the P04 quote cannot be verified, the specific subfinding about overly specific warnings being themselves triggering loses one of its cited participants, and the pooled counts become suspect. This is not a claim of misconduct; it is a concrete, checkable inconsistency between the methods narrative and the reported results. Because the paper is a qualitative study with no transcripts or codebook provided, verification is not currently possible. The reader's CONDITIONAL verdict remains appropriate: conditional on the data being made available and the phase-specific analysis being shown to support the pooled results. I therefore leave the verdict unchanged while flagging the internal-inconsistency concern as the stress-test's principal finding.","tokens_in":26769,"tokens_out":6282,"duration_ms":63770,"concrete_test":"Request the de-identified transcripts and codebook from the authors (or check the open-data supplement if provided) and verify the P04 interview: was Figure 3-B or 3-C, or any slide deck, actually shown during that session? If not, remove the P04 quote from Section 4.1.1 and recompute the 3/15 count from the remaining participants. As a second confirmation, separately tabulate every prevalence count reported in Section 4 by cohort (P1-P6 vs P7-P15) and state for each theme whether both cohorts were asked the relevant questions; if a count is supported only by the updated cohort, revise the result and the framework's empirical basis accordingly.","verdict_should_be":"UNCHANGED","load_bearing_attack":"Section 3.2 states that the interview protocol was updated after the first six participants, and that the slide deck of warning examples (Appendix A.4) was developed as part of that update for the remaining nine participants. Section 3.3 further describes two iterative phases of analysis. Yet Section 4.1.1 quotes P04, one of the first six, as commenting on 'the middle one (Figure 3-B)' and on a slide labeled 'TW: Gangrape' (Figure 3-C), using these comments to support the claim that 3/15 participants found overly specific warnings potentially triggering. Unless the slide deck was also shown to P1-P6, which the paper does not state, this quotation is chronologically impossible. The issue is load-bearing because the paper pools all 15 participants when reporting prevalence counts (e.g., 12/15, 5/15, 7/15) and builds the three-stakeholder framework on themes drawn from both phases. The initial protocol was oriented around distinctions between trigger warnings and content warnings; the revised protocol reoriented the inquiry toward viewer-versus-poster dynamics. If phase-specific data are blended without cohort-specific denominators or question-by-question availability, the frequency claims and the empirical grounding of the framework are not transparent. Section 5.6 candidly addresses sampling limitations, but it does not address this internal protocol inconsistency.","agreement_with_reader":"partial"},"referee_report":{"model":"deepseek-v4-flash","summary":"The paper reports a semi-structured interview study with 15 US-based social media users to understand perceptions of trigger and content warnings (TW/CW). It identifies challenges for three stakeholder groups: viewers deciding whether to engage with warning-labeled content, posters deciding whether and how to apply warnings, and platforms whose design features shape the visibility and usability of warnings. Based on a two-phase thematic analysis, the authors propose a conceptual framework (Figure 2) that interrelates viewers, posters, and platforms, and they derive design implications such as multi-level warning specificity, poster-facing nudges, and post-exposure support resources. The appendices contain the screening survey, both interview protocols, and the set of warning examples used as stimuli.","tokens_in":26970,"tokens_out":5487,"duration_ms":43225,"significance":"If the findings hold, the proposed framework provides a useful synthesis of how TW/CW function across stakeholders and can inform concrete platform design decisions. The paper's strengths include its detailed documentation of the interview protocols and stimulus materials in the appendices, which supports reproducibility, and its grounding of themes in participant quotes and prevalence counts. The design implications (e.g., layered warnings, personalized filtering, poster nudges, post-exposure care) are actionable and align with prior trauma-informed computing work. However, the contribution's credibility depends heavily on transparent reporting of which data came from which interview protocol and on the sample's representativeness of 'general social media users'; both aspects currently have unresolved issues that affect the empirical grounding of the framework.","major_comments":[{"comment":"The quote attributed to P04, a participant from the first cohort (P1–P6), refers to 'the middle one (Figure 3-B)' and to a slide labeled 'TW: Gangrape (Figure 3-C)', yet Section 3.2 states that the slide deck of warning examples was developed as part of the updated protocol for the remaining nine participants (P7–P15). Unless the slide deck was also shown to the first cohort, which the paper does not state, this quotation is chronologically impossible. Because the paper uses this quote as evidence for the claim that 3/15 participants found overly specific warnings potentially triggering, this inconsistency directly affects the evidentiary basis of the theme in Section 4.1.1. The authors should clarify whether the slide deck was used with P1–P6 or revise the results to exclude data from participants who did not see the stimuli.","section":"Section 3.2 and Section 4.1.1"},{"comment":"The results report aggregate prevalence counts (e.g., 12/15, 5/15, 7/15) that pool data from two different interview protocols described in Section 3.2. The initial protocol probed perceived differences between TW and CW, while the revised protocol reoriented the questions toward viewer-versus-poster dynamics and introduced a slide-deck stimulus. Presenting combined counts without cohort-specific denominators or question-by-question availability makes it impossible to determine whether a given theme emerged from comparable prompts across both cohorts. The authors should either report which codes and themes were present in each cohort or provide a clear justification for why pooling is valid despite the protocol change.","section":"Section 3.3 and Section 4"},{"comment":"The paper characterizes the participants as 'general social media users' in the U.S., but recruitment was limited to a graduate student listserv, a CS Discord server, the authors' personal accounts, and the r/mentalhealth subreddit. The paper itself concedes in Section 5.6 that the sample skewed toward users with a pre-existing interest in TW/CW, excluded users who oppose warnings, and did not include severe trauma survivors or heavy content creators. The abstract and research questions frame the findings as applicable to general users; this framing is stronger than the evidence supports. The authors should temper the generalizing language or explicitly present the study as an exploratory, purposive-sample investigation with transferability as the goal.","section":"Section 3.1, Section 5.6, and the Abstract"}],"minor_comments":[{"comment":"The phrase 'interacting with a subscribe' appears to be a typo; it should likely read 'interacting with a subscriber' or 'interacting with a user'.","section":"Section 2.3"},{"comment":"The heading 'Interest in Being Interviews' should read 'Interest in Being Interviewed'.","section":"Appendix A.1"},{"comment":"The sentence 'users can voluntarily flag their social media as \"sensitive content\"' should refer to flagging their posts or content, not 'social media'.","section":"Section 4.2.2"},{"comment":"The phrase 'did not inquire if they have' is grammatically incomplete; consider changing to 'did not inquire whether they had such experiences'.","section":"Section 5.6"},{"comment":"References [1], [2], [4], and [54] lack publication dates (shown as '[n. d.]'); the authors should provide them or use a consistent citation style.","section":"References"}],"recommendation":"major_revision","confidential_remarks":"The paper is a reasonable qualitative HCI study, but the internal inconsistency around the P04 quote and the undisclosed pooling of data across two protocols are significant methodological transparency issues that need to be fixed before publication. The sample limitation is acknowledged by the authors, but the abstract's generalizing language should be adjusted. If the authors clarify the protocol data and temper claims, the contribution could be acceptable for a CSCW/CHI-style venue."},"author_rebuttal":null,"desk_editor":{"model":"deepseek-v4-flash","letter":"Short version: this is a solid qualitative study with a real contribution, but it has an internal inconsistency that needs fixing before the numbers can be trusted.\n\nWhat's new: prior TW/CW work looked at specific platforms or populations (trans-specific, photosensitivity, substance recovery). This paper interviews 15 general users across platforms and organizes the space into a three-stakeholder framework (viewers, posters, platforms). That framing is genuinely useful for HCI designers thinking about warning systems. The interview quotes are vivid and the thematic analysis is transparent enough to follow. Credit where due: Section 5.6 is candid about the sample limits, and the design implications (poster nudges, personalized filters, post-exposure support) follow from the data.\n\nThe soft spots. First, the stress-test issue is real. Section 3.2 says the slide deck of warning examples (Appendix A.4) was developed as part of the protocol update for participants P7-P15. Yet Section 4.1.1 quotes P04, one of the first six, discussing 'the middle one (Figure 3-B)' and a slide labeled 'TW: Gangrape' (Figure 3-C). That is chronologically impossible unless the slides were also shown to P1-P6, which the paper does not state. This matters because the paper pools all 15 participants in counts like 3/15 and 12/15. If P04's quote is invalid, the 'too specific' theme loses one of its two quoted supporters and the 3/15 figure changes. This is fixable—check the recording, add a note that the slides were shown to all participants, or recalculate cohort-specific counts—but it undermines the current presentation.\n\nSecond, the sample. Recruitment via a university listserv, CS Discord, personal social media, and r/mentalhealth does not support the 'general social media users' framing. Section 5.6 acknowledges the skew toward TW/CW-interested participants and the exclusion of severe trauma survivors and heavy creators, which is good, but the abstract and RQs still overreach. The protocol change after the first six interviews is also underreported in the results: no cohort-specific denominators or question-by-question availability are given, so the pooled counts are hard to interpret.\n\nThird, minor: no codebook or transcripts are provided, which limits verification. For a qualitative paper that's not fatal, but it makes the '12/15' style claims less checkable.\n\nWho is this for? Researchers in social computing and HCI working on content moderation, warning systems, or trauma-informed design. They will get a useful stakeholder map and a set of design directions. But read the counts with caution until the authors address the protocol inconsistency.\n\nRecommendation: send it to peer review. It deserves a serious referee. The central contribution is sound; the internal inconsistency is embarrassing but repairable, and the sampling limitations can be handled with framing and cohort-specific reporting. I'd accept it with major revision.","headline":"Useful qualitative map of TW/CW stakeholders, but a participant quote that could not have happened under the stated protocol undermines the pooled counts until fixed.","tokens_in":27496,"tokens_out":2767,"would_cite":false,"duration_ms":23558,"reading_group":"maybe","serious_thinker":"yes","would_accept_peer_review":true},"rs_alignment":null,"lean_confirmation":null,"pith_extraction":{"msc":[],"pacs":[],"model":"deepseek-v4-flash","headline":"Trigger warnings on social media fail because viewers, posters, and platforms each face unsolved tensions, a 15-person interview study argues.","keywords":["trigger warnings","content warnings","social media","content moderation","trauma-informed design","user perceptions","qualitative study","human-computer interaction"],"falsifier":"Conduct a larger, more diverse survey or interview study that includes people who actively oppose TW/CW and people with severe trauma histories; if their perspectives reveal fundamentally different decision-making processes or barriers, the proposed three-stakeholder framework would need revision.","tokens_in":26542,"feed_emoji":"⚠️","tokens_out":1373,"duration_ms":13878,"temperature":0.7,"pith_summary":"The paper argues that trigger and content warnings (TW/CW) on social media are not working as intended because the people who view them, the people who post them, and the platforms that host them all face unresolved challenges. Through semi-structured interviews with 15 general social media users, the authors find that viewers struggle to decide whether to engage with warned content, posters struggle with whether and how to apply warnings, and platforms shape warning visibility in ways that often undermine their use. The paper proposes a conceptual framework that ties these three stakeholder perspectives together and offers design implications for platforms. If the framework is right, improving TW/CW requires coordinated changes that address all three roles simultaneously, not just one.","feed_headline":"Trigger warnings fail because three roles clash","feed_subtitle":"Interviews with 15 social media users show viewers, posters, and platforms each face unsolved tensions.","key_machinery":"The central object is the proposed conceptual framework of the TW/CW mechanism (Figure 2), which identifies three stakeholders—viewers, posters, and platforms—and the challenges and needs that connect them. The framework organizes the interview findings into a structure that explains how warning design choices propagate across roles: platform affordances shape what posters can do, poster choices shape what viewers encounter, and viewer responses shape platform decisions. The framework is used to derive design implications for platforms, emphasizing that TW/CW is not a simple label but a sociotechnical system.","core_discovery":"The paper claims that the TW/CW mechanism on social media is best understood as a three-way interaction among viewers, posters, and platforms, and that current practices fail because each stakeholder faces unacknowledged tensions. Viewers must weigh their need to decide whether to engage with content against the risk that too-specific warnings themselves trigger distress; posters must decide which topics warrant warnings despite the absence of shared standards, and they face a perceived trade-off between adding warnings and losing engagement; platforms determine the visibility and usability of warnings through design features that often make warnings easy to miss or that provoke curiosity rather than caution. The authors synthesize these findings into a conceptual framework that maps the relationships among the three stakeholders and suggests design interventions such as multi-level warning specificity, personalized filtering, and proactive poster support.","pith_inferences":["If the framework is correct, the effectiveness of TW/CW should be studied as a system effect rather than a label effect; a single change in platform design may shift the entire viewer-poster dynamic.","The paper's finding that vague warnings can provoke curiosity suggests a testable design hypothesis: incremental disclosure (e.g., a 'see why' button) may reduce accidental exposure while preserving informed choice, but it needs experimental validation.","The absence of severe-trauma survivors and warning opponents in the sample means the framework may be incomplete for exactly the populations that warnings are meant to protect most."],"forward_implications":["Platforms could adopt multi-level warnings that let viewers choose how much specificity to reveal, reducing the risk that a warning itself becomes a trigger.","Posters could be supported by proactive, just-in-time prompts or automated suggestion tools that help them apply relevant warnings to their posts.","Warning labels could be linked to personalized filtering systems that let users block content by trigger category rather than deciding post-by-post.","Platforms could pair warnings with post-exposure support, such as mental-health resources or ways to break doom-scrolling cycles."],"supporting_citations":[{"why":"Provides the systematic typology of content warning types that the paper uses to define TW/CW and to interpret the topics participants mention.","marker":"[15]"},{"why":"Describes the TransTime platform's use of content-warning tags, which the paper cites as an example of topic-specific warning mechanisms that its broader framework extends.","marker":"[29]"},{"why":"Establishes the trauma-informed computing framework that the paper aligns with for personalized moderation and post-exposure support recommendations.","marker":"[16]"},{"why":"Shows that specific, multi-level warnings outperform generic ones for photosensitive epilepsy, which the paper generalizes to other trigger contexts.","marker":"[58]"},{"why":"Investigates how people in substance-use recovery encounter unexpected triggers on social media, supporting the paper's call for personalized filtering and better warning systems.","marker":"[51]"},{"why":"Finds that Instagram's sensitive-content screens do not deter vulnerable users, which aligns with the paper's observation that warnings can provoke curiosity.","marker":"[10]"},{"why":"Examines bittersweet experiences with Facebook Memories, illustrating how even positive content can be triggering and showing the limits of current warning mechanisms.","marker":"[45]"},{"why":"Presents DeText, an automated tool for adding content warnings about sexual violence, which the paper cites as a topic-specific solution that its broader framework seeks to complement.","marker":"[59]"}],"fun_headline_variants":["Three-way clash defines trigger warning confusion","Viewers, posters, platforms: why trigger warnings miss","Trigger warnings: 15 users reveal three unresolved tensions","Social media warnings: a three-stakeholder puzzle","The trigger warning tug-of-war: viewers, posters, platforms"],"cache_read_input_tokens":3200,"weakest_assumption_plain":"The study's load-bearing assumption is that the 15 self-selected interview participants—recruited through a university listserv, a CS Discord, the authors' social media, and the r/mentalhealth subreddit—represent typical U.S. social media users, even though the sample skewed toward people already interested in TW/CW and did not include those who oppose warnings or have experienced severe trauma.","fun_headline_variants_meta":{"raw":{"variants":["Three-way clash defines trigger warning confusion","Viewers, posters, platforms: why trigger warnings miss","Trigger warnings: 15 users reveal three unresolved tensions","Social media warnings: a three-stakeholder puzzle","The trigger warning tug-of-war: viewers, posters, platforms"]},"model":"deepseek-v4-flash","effort":"low","cost_usd":0.000624,"raw_usage":{"total_tokens":2853,"prompt_tokens":876,"completion_tokens":1977,"prompt_tokens_details":{"cached_tokens":384},"prompt_cache_hit_tokens":384,"prompt_cache_miss_tokens":492,"completion_tokens_details":{"reasoning_tokens":1915}},"tokens_in":492,"tokens_out":1977,"duration_ms":12902,"temperature":1.0,"reasoning_tokens":1915,"cache_read_input_tokens":384,"cache_creation_input_tokens":0},"cache_creation_input_tokens":0},"created_at":"2026-08-16T11:26:05.595958+00:00","model_set":{"reader":"deepseek-v4-flash"},"falsifier":"Conduct a larger, more diverse survey or interview study that includes people who actively oppose TW/CW and people with severe trauma histories; if their perspectives reveal fundamentally different decision-making processes or barriers, the proposed three-stakeholder framework would need revision.","supporting_citations":[{"cited_title":null,"cited_arxiv_id":null,"evidence_quote":"Provides the systematic typology of content warning types that the paper uses to define TW/CW and to interpret the topics participants mention."},{"cited_title":null,"cited_arxiv_id":null,"evidence_quote":"Describes the TransTime platform's use of content-warning tags, which the paper cites as an example of topic-specific warning mechanisms that its broader framework extends."},{"cited_title":null,"cited_arxiv_id":null,"evidence_quote":"Establishes the trauma-informed computing framework that the paper aligns with for personalized moderation and post-exposure support recommendations."},{"cited_title":null,"cited_arxiv_id":null,"evidence_quote":"Shows that specific, multi-level warnings outperform generic ones for photosensitive epilepsy, which the paper generalizes to other trigger contexts."},{"cited_title":null,"cited_arxiv_id":null,"evidence_quote":"Investigates how people in substance-use recovery encounter unexpected triggers on social media, supporting the paper's call for personalized filtering and better warning systems."},{"cited_title":null,"cited_arxiv_id":null,"evidence_quote":"Finds that Instagram's sensitive-content screens do not deter vulnerable users, which aligns with the paper's observation that warnings can provoke curiosity."},{"cited_title":null,"cited_arxiv_id":null,"evidence_quote":"Examines bittersweet experiences with Facebook Memories, illustrating how even positive content can be triggering and showing the limits of current warning mechanisms."},{"cited_title":null,"cited_arxiv_id":null,"evidence_quote":"Presents DeText, an automated tool for adding content warnings about sexual violence, which the paper cites as a topic-specific solution that its broader framework seeks to complement."}],"review_version":1}