EgoPolice introduces a 185-hour annotated police body-worn camera benchmark showing state-of-the-art video models fail on high-stakes actions due to motion, occlusion, and low inter-class visual separability.
How persuasive is AI- generated propaganda? PNAS Nexus, 3(2):pgae034, February 2024
8 Pith papers cite this work. Polarity classification is still indexing.
representative citing papers
Generalizing twists to arbitrary integer labels on non-manifold meshes enables design of linked knot structures corresponding to immersions of 4D knotted surfaces in 3D.
LLM self-reports predict behavior selectively: TPB reaches human-level coherence within shared conversations but collapses across sessions for primed behaviors, unlike Big 5, with persona prompting stabilizing reports but not actions.
Frontier multimodal LLMs perceive cities through a culturally uneven baseline where European and North American framings are closest to the model's neutral default, while prompting fails to recover human diversity.
80% of hateful tweets remain online after five months with no higher removal rate than non-hateful content, while human-AI moderation pipelines can feasibly cut user exposure below regulatory penalty costs.
Both Qβ and SARS-CoV-2 populations display hierarchical genotype networks with a highly abundant central haplotype and diminishing variant abundance at greater Hamming distances.
LLM-based persuasion systems frequently match or exceed human effectiveness across domains, with key influences from interaction style, model scale, prompt design, and personalization, while posing risks to information integrity, fairness, privacy, and autonomy.
A structured review of JSP 936 identifies eight challenge areas in operationalising AI assurance for UK Defence and concludes that further methods, guidance, and organisational capability are required.
citing papers explorer
-
EgoPolice: A Benchmark for Egocentric Video Understanding in High-Stakes Police Body-Worn Camera Footage
EgoPolice introduces a 185-hour annotated police body-worn camera benchmark showing state-of-the-art video models fail on high-stakes actions due to motion, occlusion, and low inter-class visual separability.
-
Twisted Edges: A Unified Framework for Designing Linked Knot (LK) Structures Using Labeled Non-Manifold Surface Meshes
Generalizing twists to arbitrary integer labels on non-manifold meshes enables design of linked knot structures corresponding to immersions of 4D knotted surfaces in 3D.
-
Rethinking Psychometric Evaluation of LLMs: When and Why Self-Reports Predict Behavior
LLM self-reports predict behavior selectively: TPB reaches human-level coherence within shared conversations but collapses across sessions for primed behaviors, unlike Big 5, with persona prompting stabilizing reports but not actions.
-
Culturally uneven urban perception in large language models
Frontier multimodal LLMs perceive cities through a culturally uneven baseline where European and North American framings are closest to the model's neutral default, while prompting fails to recover human diversity.
-
The Enforcement and Feasibility of Hate Speech Moderation on Twitter
80% of hateful tweets remain online after five months with no higher removal rate than non-hateful content, while human-AI moderation pipelines can feasibly cut user exposure below regulatory penalty costs.
-
Shared quasispecies architecture in experimental and natural RNA virus populations
Both Qβ and SARS-CoV-2 populations display hierarchical genotype networks with a highly abundant central haplotype and diminishing variant abundance at greater Hamming distances.
-
Persuasion with Large Language Models: A Survey of Empirical Evidence, Study Methodologies, and Ethical Implications
LLM-based persuasion systems frequently match or exceed human effectiveness across domains, with key influences from interaction style, model scale, prompt design, and personalization, while posing risks to information integrity, fairness, privacy, and autonomy.
-
AI Assurance in UK Defence: Challenges in Operationalising JSP 936
A structured review of JSP 936 identifies eight challenge areas in operationalising AI assurance for UK Defence and concludes that further methods, guidance, and organisational capability are required.