Agentic browsers are vulnerable to 20 web and LLM attacks with 18 implemented, exposing five failure modes across four major LLM models that require redesign before safe deployment.
Title resolution pending
12 Pith papers cite this work. Polarity classification is still indexing.
verdicts
UNVERDICTED 12representative citing papers
PPT-Bench measures how LLMs change answers under epistemic, value, authority, and identity pressures at baseline, single-turn, and multi-turn levels, finding separable inconsistency patterns across five models.
MS-DKC is a dataset knowledge card framework that maps image, morphology, supervision, context, and risk descriptors to design priors and failure modes, shown to produce dataset-specific model adaptations with improved metrics on DRIVE, ISIC2018, and ACDC.
A state-space definition of fading memory is introduced that extends incremental input-to-output stability via a memory kernel, is implied by incremental input-to-state stability under bounded inputs, and holds for current-driven memristor models.
GEAR-Seg decouples segmentation, semantic description, and LLM reasoning into an explicit chain for interpretable zero-shot reasoning segmentation while generating the GEAR-131K dataset.
Embodied LLM agents exhibit emergent collaborative behaviors indicating mental models of partners in a color-matching game, detected via LLM judges and supported by positive user feedback.
PyPeT is an open Python framework for unified CTP and MRP processing that outputs CBF, CBV, MTT, TTP, and Tmax maps and reports mean SSIM of approximately 0.8 against three FDA-approved commercial tools.
VArify introduces a tree visualization to support human verification of GraphRAG evidence for LLM responses in food science, evaluated in a study with six domain experts.
King functions for shifted Gaussians are shown to satisfy a differential equation unitarily equivalent to the radial Schrödinger operator and to form a dense system in radial velocity space.
Fine-tuned foundation models produce reliable MSK MRI biomarkers that support workload-reducing triage and calibrated 48-month prediction of knee replacement and incident OA.
A structured review of JSP 936 identifies eight challenge areas in operationalising AI assurance for UK Defence and concludes that further methods, guidance, and organisational capability are required.
This review summarizes FRB properties and outlines how SKA capabilities will help identify progenitors and enable cosmological applications.
citing papers explorer
-
WAAA! Web Adversaries Against Agentic Browsers
Agentic browsers are vulnerable to 20 web and LLM attacks with 18 implemented, exposing five failure modes across four major LLM models that require redesign before safe deployment.
-
Beyond Social Pressure: Benchmarking Epistemic Attack in Large Language Models
PPT-Bench measures how LLMs change answers under epistemic, value, authority, and identity pressures at baseline, single-turn, and multi-turn levels, finding separable inconsistency patterns across five models.
-
MS-DKC: A Dataset Knowledge Card Framework for Designing and Adapting Medical Image Segmentation Models
MS-DKC is a dataset knowledge card framework that maps image, morphology, supervision, context, and risk descriptors to design priors and failure modes, shown to produce dataset-specific model adaptations with improved metrics on DRIVE, ISIC2018, and ACDC.
-
State-space fading memory
A state-space definition of fading memory is introduced that extends incremental input-to-output stability via a memory kernel, is implied by incremental input-to-state stability under bounded inputs, and holds for current-driven memristor models.
-
GEAR-Seg: A Grounded Explainable Agent for Reasoning Segmentation and Data Engine
GEAR-Seg decouples segmentation, semantic description, and LLM reasoning into an explicit chain for interpretable zero-shot reasoning segmentation while generating the GEAR-131K dataset.
-
Evaluating Generative Models as Interactive Emergent Representations of Human-Like Collaborative Behavior
Embodied LLM agents exhibit emergent collaborative behaviors indicating mental models of partners in a color-matching game, detected via LLM judges and supported by positive user feedback.
-
PyPeT: A Python Perfusion Tool for Automated Quantitative Brain CT and MR Perfusion Analysis
PyPeT is an open Python framework for unified CTP and MRP processing that outputs CBF, CBV, MTT, TTP, and Tmax maps and reports mean SSIM of approximately 0.8 against three FDA-approved commercial tools.
-
VArify: A Visual Analytics System for Verifying Knowledge Enhanced Large Language Model Responses in Food Science
VArify introduces a tree visualization to support human verification of GraphRAG evidence for LLM responses in food science, evaluated in a study with six domain experts.
-
King Function for Shifted Gaussian: Laguerre Structure, Spectral Theory and Density
King functions for shifted Gaussians are shown to satisfy a differential equation unitarily equivalent to the radial Schrödinger operator and to form a dense system in radial velocity space.
-
Clinical utility of foundation models in musculoskeletal MRI for biomarker fidelity and predictive outcomes
Fine-tuned foundation models produce reliable MSK MRI biomarkers that support workload-reducing triage and calibrated 48-month prediction of knee replacement and incident OA.
-
AI Assurance in UK Defence: Challenges in Operationalising JSP 936
A structured review of JSP 936 identifies eight challenge areas in operationalising AI assurance for UK Defence and concludes that further methods, guidance, and organisational capability are required.
-
The Astrophysics of Fast Radio Bursts
This review summarizes FRB properties and outlines how SKA capabilities will help identify progenitors and enable cosmological applications.