Fidelity probes from code raise specification fidelity from 0.63 to 0.94 on a 12k-line COBOL benchmark over eight iterations, with convergence predicted by a two-state Markov fixed point from four iterations of rate data.
Title resolution pending
8 Pith papers cite this work. Polarity classification is still indexing.
years
2026 8representative citing papers
PAIR-CI restores calibrated conditional independence testing under incomplete data by pairing models on the same imputed conditioning set and unifying cross-validation with multiple-imputation variance.
Proposes a three-step benchmark design method (define work activity, specify tested setting, score work product) derived from work studies and O*NET, demonstrated via three case analyses.
A test-driven pipeline with an auto-constructed privacy feature library detects 2.56 times more confirmed privacy leaks in LLM-based code generation than existing baselines.
Hybrid analytical/AD method for SE(3) NLLs delivers exact Hessians 5x faster than finite-differencing while matching nested AD to machine precision, plus a fix for origin NaNs in the scalar basis.
Large-scale topic modeling of 270k Reddit posts shows GenAI discourse in education shifting from detection-evasion to enforcement, with K-12 teachers emphasizing cognitive dependency, academics focusing on detection, students on career anxiety, and adversarial themes driving engagement and cross-sta
A framework integrates synthetic population generation from ACS PUMS, deep contrastive learning for housing-household compatibility, and hierarchical optimization to produce a joint inventory that matches block-group demographics and spatial patterns in coastal North Carolina.
Proposes cryptographic registry identity, dual-signature model, and authoritative namespace binding to create three defense layers against dependency confusion.
citing papers explorer
-
Fidelity Probes for Specification--Code Alignment
Fidelity probes from code raise specification fidelity from 0.63 to 0.94 on a 12k-line COBOL benchmark over eight iterations, with convergence predicted by a two-state Markov fixed point from four iterations of rate data.
-
Shedding Light onto Safety Integrity Level and Basic Software Constraints in a Real-World Automotive Application: Case Study with Driverator Framework
PAIR-CI restores calibrated conditional independence testing under incomplete data by pairing models on the same imputed conditioning set and unifying cross-validation with multiple-imputation variance.
-
Design and Report Benchmarks for Knowledge Work
Proposes a three-step benchmark design method (define work activity, specify tested setting, score work product) derived from work studies and O*NET, demonstrated via three case analyses.
-
Probing Privacy Leaks in LLM-based Code Generation via Test Generation
A test-driven pipeline with an auto-constructed privacy feature library detects 2.56 times more confirmed privacy leaks in LLM-based code generation than existing baselines.
-
Exact Higher-Order Derivatives for SE(3) via Analytical/AD Methods
Hybrid analytical/AD method for SE(3) NLLs delivers exact Hessians 5x faster than finite-differencing while matching nested AD to machine precision, plus a fix for origin NaNs in the scalar basis.
-
ChatGPT vs Teachers vs Students: Large-Scale Analysis of Generative AI Discourse in Education Communities on Reddit
Large-scale topic modeling of 270k Reddit posts shows GenAI discourse in education shifting from detection-evasion to enforcement, with K-12 teachers emphasizing cognitive dependency, academics focusing on detection, students on career anxiety, and adversarial themes driving engagement and cross-sta
-
A Joint Synthetic Housing-Household Inventory
A framework integrates synthetic population generation from ACS PUMS, deep contrastive learning for housing-household compatibility, and hierarchical optimization to produce a joint inventory that matches block-group demographics and spatial patterns in coastal North Carolina.
-
Cryptographic Registry Provenance: Structural Defense Against Dependency Confusion in AI Package Ecosystems
Proposes cryptographic registry identity, dual-signature model, and authoritative namespace binding to create three defense layers against dependency confusion.