CI-Repair-Bench supplies 567 repository-level CI failures across 12 error types for evaluating automated repairs via complete original workflow re-execution, with the best LLM achieving 18.9% success.
hub
write newline
10 Pith papers cite this work. Polarity classification is still indexing.
hub tools
verdicts
UNVERDICTED 10representative citing papers
An empirical analysis of GitHub Actions reveals diverse caching practices, higher activity in adopting repos, and ongoing maintenance driven by fixes and updates.
Bayesian nonparametric distributional theory and arbitrary-size out-of-sample predictions for distinct and shared species across two areas.
Reports value function interference and overestimation sensitivity as issues hindering value-based MORL with non-linear scalarisation, shown via tabular Q-learning on simple multi-objective MDPs.
Empirical study across five LLMs and four languages finds security-aware prompting changes CWE category distributions but yields no statistically significant reduction in vulnerability frequency or density.
EasyPolar uses three synchronized RGB cameras plus a confidence-guided network to deliver high-resolution single-shot linear polarimetric imaging without specialized DoFP sensors.
VisDoc uses a GenAI pipeline grounded in CTML to restructure OSS onboarding docs, with small evaluations showing higher task success and lower cognitive load.
A case study of software redesign reveals reuse challenges and shows that semantic alignment heuristics with hierarchical clone detection reduce irrelevant clones by 33-99% and raise precision to 86%.
ELM-FBPINNs replace subdomain networks in FBPINNs with extreme learning machines, reducing PDE training to structured linear least-squares while maintaining competitive accuracy on benchmarks.
An industrial evaluation shows Taxonomic Trace Links enable traceability in specific scenarios at Ericsson but face practical barriers in taxonomy development and classifier precision.
citing papers explorer
-
CI-Repair-Bench: A Repository-Aware Benchmark for Automated Patch Validation via CI Workflows
CI-Repair-Bench supplies 567 repository-level CI failures across 12 error types for evaluating automated repairs via complete original workflow re-execution, with the best LLM achieving 18.9% success.
-
How Developers Adopt, Use, and Evolve CI/CD Caching: An Empirical Study on GitHub Actions
An empirical analysis of GitHub Actions reveals diverse caching practices, higher activity in adopting repos, and ongoing maintenance driven by fixes and updates.
-
Bayesian discovery of species in multiple areas
Bayesian nonparametric distributional theory and arbitrary-size out-of-sample predictions for distinct and shared species across two areas.
-
Issues with Value-Based Multi-objective Reinforcement Learning: Value Function Interference and Overestimation Sensitivity
Reports value function interference and overestimation sensitivity as issues hindering value-based MORL with non-linear scalarisation, shown via tabular Q-learning on simple multi-objective MDPs.
-
An Empirical Evaluation of LLM-Generated Code Security Across Prompting Methods
Empirical study across five LLMs and four languages finds security-aware prompting changes CWE category distributions but yields no statistically significant reduction in vulnerability frequency or density.
-
High-Resolution Single-Shot Polarimetric Imaging Made Easy
EasyPolar uses three synchronized RGB cameras plus a confidence-guided network to deliver high-resolution single-shot linear polarimetric imaging without specialized DoFP sensors.
-
Restructure This: Using AI to Restructure Onboarding Documents to Reduce Cognitive Overload
VisDoc uses a GenAI pipeline grounded in CTML to restructure OSS onboarding docs, with small evaluations showing higher task success and lower cognitive load.
-
Investigating Code Reuse in Software Redesign: A Case Study
A case study of software redesign reveals reuse challenges and shows that semantic alignment heuristics with hierarchical clone detection reduce irrelevant clones by 33-99% and raise precision to 86%.
-
ELM-FBPINNs: An Efficient Multilevel Random Feature Method
ELM-FBPINNs replace subdomain networks in FBPINNs with extreme learning machines, reducing PDE training to structured linear least-squares while maintaining competitive accuracy on benchmarks.
-
Empirical Evaluation of Taxonomic Trace Links: A Case Study
An industrial evaluation shows Taxonomic Trace Links enable traceability in specific scenarios at Ericsson but face practical barriers in taxonomy development and classifier precision.