A survey of 172 open educational datasets from 204 papers across LAK, EDM, and AIED conferences reveals trends, 143 previously uncatalogued datasets, field gaps, and an 8-item PRACTICE checklist for better data publication.
Mixed citations
, author McKenzie, J.E
Mixed citation behavior. Most common role is background (50%).
citation-role summary
citation-polarity summary
authors
co-cited works
representative citing papers
MetaSyn is a stage-level benchmark of 442 meta-analyses showing LLM agents retrieve up to 90.9% of eligible studies but include at most 52.7% in their final reports.
The paper defines five AI system categories for public administration and reports that 55% of 91 recent papers leave the system type underspecified while 31% study one type but motivate with another.
A narrative survey that catalogs fifty papers on diffusion-based adversarial techniques across text, vision, and vision-language models, proposes a six-class taxonomy of diffusion roles plus a unified five-dimension evaluation framework, and releases a companion catalog.
The authors release BioStance, a dataset of 39,600 context-preserving Reddit post-comment pairs annotated for favor, against, or none stance on six bioethical controversy targets, with Krippendorff's alpha of 0.82.
Presents MedSci Skills, an open-source toolkit with deterministic integrity gates for verifying LLM-assisted clinical manuscripts against reporting guidelines like STARD, PRISMA, and STROBE.
The paper reframes manufacturing ransomware recovery as an interdependency problem, identifies nine evidence-backed failure modes from a multivocal review, and defines Minimum Viable Factory Recovery as an analytical objective for resuming minimal safe operations.
SotA Lens introduces a network-augmented workflow that turns seed searches into citation graphs with community detection to support exploratory state-of-the-art reviews.
AI improves brainstorming quality for general-purpose impact assessment but not specialized applications when it offers hints early and structures ideas later, based on workshop evaluations with 54 participants.
LABBench2 is a more challenging benchmark than LAB-Bench for assessing AI performance on biology research tasks, with frontier models showing accuracy drops of 26-46% across subtasks.
The authors conduct a systematic literature review and real-world analysis to define Crowdsourced Context Systems and map a six-aspect design space with normative implications.
Material demands from highly decarbonised European energy models exceed population-based shares of global reserves for Ga, In, Ir, Te and to a lesser extent Ag, Se, V.
A new Multidimensional Resilience Index shows that simultaneous failures across physical, cyber, and exogenous dimensions produce 46-fold greater resilience loss in power systems than isolated stress.
Introduces a taxonomy of nine LLM code smells, a static detection tool, and reports 73.5% prevalence with 91.3% precision and 71.8% recall across 692 projects.
Systematic review of teleoperated endovascular robots finds navigation feasible over 7000 km with 30-163 ms latency and 100% success in small human trials, mostly from animal and phantom models.
A PRISMA-guided review of 21 papers shows RL work on C/C++ vulnerabilities focuses on fuzzing rather than detection or localization, proposes a taxonomy, and flags the lack of CFG-based state representations for vulnerable node identification.
For noisy near-term quantum devices, the paper recommends shallow angle encoding over amplitude encoding once two-qubit error rates exceed roughly 10^-3.
A scoping review and personal reflections identify robot wrangling as a complex umbrella term and generate design implications for supporting wranglers as individuals and within broader service ecologies.
A literature survey finds foundation-model agents in industry are 75% at prototype stages with gains in human interaction and uncertainty handling but deficits in negotiation, plus limitations like hallucinations and latency.
Literature on system prompts for AI shows fragmented and contradictory claims that complicate policy efforts to use them as reliable governance mechanisms.
A literature review shows that constructs for appropriate reliance on AI are fragmented, presents three views on the topic, and calls for consensus on objective metrics to enable better comparisons across studies.
A three-layer trust framework integrating human, AI, and interaction perspectives is proposed to align multi-stakeholder trust criteria for AI-driven mental health support.
An interdisciplinary workshop produced a catalog of ideas and a roadmap showing how meta-research can tackle challenges like reproducibility, transparency, and ethical implementation in trustworthy AI for healthcare.
A systematic survey of data balancing methods that organizes the field into six taxonomic groups and reports a two-dataset case study showing method performance depends on the dataset and classifier.
citing papers explorer
-
Open Datasets in Learning Analytics: Trends, Challenges, and Best PRACTICE
A survey of 172 open educational datasets from 204 papers across LAK, EDM, and AIED conferences reveals trends, 143 previously uncatalogued datasets, field gaps, and an 8-item PRACTICE checklist for better data publication.
-
MetaSyn: A Benchmark for LLM Agents on Meta-Analysis Articles from Nature Portfolio
MetaSyn is a stage-level benchmark of 442 meta-analyses showing LLM agents retrieve up to 90.9% of eligible studies but include at most 52.7% in their final reports.
-
A Technical Typology of AI Systems in Public Administration
The paper defines five AI system categories for public administration and reports that 55% of 91 recent papers leave the system type underspecified while 31% study one type but motivate with another.
-
Adversarial Diffusion Across Modalities: A Fusion Survey of Attacks, Defenses, and Evaluation for Text, Vision, and Vision-Language Models
A narrative survey that catalogs fifty papers on diffusion-based adversarial techniques across text, vision, and vision-language models, proposes a six-class taxonomy of diffusion roles plus a unified five-dimension evaluation framework, and releases a companion catalog.
-
A Context-Aware Dataset for Stance Detection in Bioethical Controversies on Reddit
The authors release BioStance, a dataset of 39,600 context-preserving Reddit post-comment pairs annotated for favor, against, or none stance on six bioethical controversy targets, with Krippendorff's alpha of 0.82.
-
Deterministic Integrity Gates for LLM-Assisted Clinical Manuscript Preparation: An Auditable Biomedical Informatics Architecture
Presents MedSci Skills, an open-source toolkit with deterministic integrity gates for verifying LLM-assisted clinical manuscripts against reporting guidelines like STARD, PRISMA, and STROBE.
-
From Backup Restoration to Minimum Viable Factory Recovery: A Systematization of Ransomware Recovery in Manufacturing Systems
The paper reframes manufacturing ransomware recovery as an interdependency problem, identifies nine evidence-backed failure modes from a multivocal review, and defines Minimum Viable Factory Recovery as an analytical objective for resuming minimal safe operations.
-
SotA Lens: A Network-Augmented Methodology and Tool for Exploratory State-of-the-Art Reviews
SotA Lens introduces a network-augmented workflow that turns seed searches into citation graphs with community detection to support exploratory state-of-the-art reviews.
-
When and How AI Should Assist Brainstorming for AI Impact Assessment
AI improves brainstorming quality for general-purpose impact assessment but not specialized applications when it offers hints early and structures ideas later, based on workshop evaluations with 54 participants.
-
LABBench2: An Improved Benchmark for AI Systems Performing Biology Research
LABBench2 is a more challenging benchmark than LAB-Bench for assessing AI performance on biology research tasks, with frontier models showing accuracy drops of 26-46% across subtasks.
-
Beyond Community Notes: A Framework for Understanding and Building Crowdsourced Context Systems for Social Media
The authors conduct a systematic literature review and real-world analysis to define Crowdsourced Context Systems and map a six-aspect design space with normative implications.
-
Materealistic? How European energy system models exceed raw material reserves
Material demands from highly decarbonised European energy models exceed population-based shares of global reserves for Ga, In, Ir, Te and to a lesser extent Ag, Se, V.
-
Multidimensional Resilience for Electrical Power Systems: Systematic Review, Integrated Index, and Validation under Real-World Cyber-Physical Attack Scenarios
A new Multidimensional Resilience Index shows that simultaneous failures across physical, cyber, and exogenous dimensions produce 46-fold greater resilience loss in power systems than isolated stress.
-
LLM Code Smells: A Taxonomy and Detection Approach
Introduces a taxonomy of nine LLM code smells, a static detection tool, and reports 73.5% prevalence with 91.3% precision and 71.8% recall across 692 projects.
-
Remote Teleoperation of Endovascular Intervention Robots: A Systematic Review
Systematic review of teleoperated endovascular robots finds navigation feasible over 7000 km with 30-163 ms latency and 100% success in small human trials, mostly from animal and phantom models.
-
Reinforcement Learning for Software Vulnerability Analysis: A Systematic Review with Emphasis on C/C++ Source Code and Static Analysis
A PRISMA-guided review of 21 papers shows RL work on C/C++ vulnerabilities focuses on fuzzing rather than detection or localization, proposes a taxonomy, and flags the lack of CFG-based state representations for vulnerable node identification.
-
Feature Encoding in Quantum Machine Learning: A Survey and Practical Guidelines
For noisy near-term quantum devices, the paper recommends shallow angle encoding over amplitude encoding once two-qubit error rates exceed roughly 10^-3.
-
Designing for Robot Wranglers: A Synthesis of Literature and Practice
A scoping review and personal reflections identify robot wrangling as a complex umbrella term and generate design implications for supporting wranglers as individuals and within broader service ecologies.
-
Foundation-Model-Based Agents in Industrial Automation: Purposes, Capabilities, and Open Challenges
A literature survey finds foundation-model agents in industry are 75% at prototype stages with gains in human interaction and uncertainty handling but deficits in negotiation, plus limitations like hallucinations and latency.
-
Prompt Governance? On Governing Technologies Governed by Natural Language
Literature on system prompts for AI shows fragmented and contradictory claims that complicate policy efforts to use them as reliable governance mechanisms.
-
From Trust to Appropriate Reliance: Measurement Constructs in Human-AI Decision-Making
A literature review shows that constructs for appropriate reliance on AI are fragmented, presents three views on the topic, and calls for consensus on objective metrics to enable better comparisons across studies.
-
Aligning Human-AI-Interaction Trust for Mental Health Support: Survey and Position for Multi-Stakeholders
A three-layer trust framework integrating human, AI, and interaction perspectives is proposed to align multi-stakeholder trust criteria for AI-driven mental health support.
-
Advancing Trustworthy AI in Healthcare Through Meta-Research: Results of an Interdisciplinary Design-Thinking Workshop
An interdisciplinary workshop produced a catalog of ideas and a roadmap showing how meta-research can tackle challenges like reproducibility, transparency, and ethical implementation in trustworthy AI for healthcare.
-
Data Balancing Strategies: A Systematic Survey of Resampling and Augmentation Methods
A systematic survey of data balancing methods that organizes the field into six taxonomic groups and reports a two-dataset case study showing method performance depends on the dataset and classifier.
-
Characterizing Creativity in Data Visualization: Reflections and Future Directions
A systematic review and interview study characterize creativity in visualization design, finding that design processes are undervalued compared to final artifacts with ideation as a universal bottleneck.
-
A systematic review of COVID-19 epidemic models with endogenous human behaviour. What's next?
Systematic review finds expanded empirical data use in COVID-19 behavior-inclusive models but limited behavioral data, structural innovation, and cross-disciplinary engagement.
-
Systematic Review of Academic Procrastination Interventions in Computing Higher Education
Clear temporal structure and supportive feedback reduce procrastination in computing students more reliably than punitive measures, with larger gains on complex multi-step tasks.
-
LLM Harms: A Taxonomy and Discussion
Proposes a five-bucket taxonomy of LLM harms and calls for dynamic auditing, but the systematic review behind it is not reproducible and contains mismatched citations.
-
Human Autonomy and Sense of Agency in Human-Robot Interaction: A Systematic Literature Review
A review of 22 studies identifies five clusters of factors affecting autonomy and agency in HRI while noting limited and fragmented evidence.
-
Stereotactic Arrhythmia Radioablation for Refractory Ventricular Tachycardia: A Narrative Review and Exploratory Pooled Analysis of Clinical Outcomes and Toxicity
Narrative review and pooled analysis of STAR for refractory VT reports 16% 6-month and 33% 12-month mortality, 75% VT burden reduction at 6 months, and 7% grade 3+ toxicities, with high heterogeneity.