REVIEW 2 cited by
Cross-cultural Inspiration Detection and Analysis in Real and LLM-generated Social Media Data
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Inspiration is linked to various positive outcomes, such as increased creativity, productivity, and happiness. Although inspiration has great potential, there has been limited effort toward identifying content that is inspiring, as opposed to just engaging or positive. Additionally, most research has concentrated on Western data, with little attention paid to other cultures. This work is the first to study cross-cultural inspiration through machine learning methods. We aim to identify and analyze real and AI-generated cross-cultural inspiring posts. To this end, we compile and make publicly available the InspAIred dataset, which consists of 2,000 real inspiring posts, 2,000 real non-inspiring posts, and 2,000 generated inspiring posts evenly distributed across India and the UK. The real posts are sourced from Reddit, while the generated posts are created using the GPT-4 model. Using this dataset, we conduct extensive computational linguistic analyses to (1) compare inspiring content across cultures, (2) compare AI-generated inspiring posts to real inspiring posts, and (3) determine if detection models can accurately distinguish between inspiring content across cultures and data sources.
Forward citations
Cited by 2 Pith papers
-
Evaluating Robustness of LLMs on Crisis-Related Microblogs across Events, Information Types, and Linguistic Features
A systematic evaluation shows GPT-4o and GPT-4 outperform open-source LLMs on crisis tweet classification, with flood events and urgent-need messages as consistent failure points.
-
Towards High-Fidelity Synthetic Multi-platform Social Media Datasets via Large Language Models
LLM-generated multi-platform social media posts approximate real data on some metrics, but all three tested models show platform-specific biases in URLs, hashtags, sentiment, and topics.
Discussion (0). Continue with ORCID to comment.