INDEX
Every paper Pith has read and judged, in one searchable index.
-
cs.LG arXiv submitted 2026-06-28Expert priors enter best-subsets MIO as log-odds penalties
Nolan Alexander +1 · “A Mathematical Optimization Approach for Expert-Informed Bayesian Best Subset Selection”
2606.29516 -
stat.ME arXiv submitted 2026-06-28Bayesian model segments satellite images without local labels
Bao Khanh Nguyen +3 · “Scalable Bayesian Spatial Mixture Modelling for Remote Sensing Image Segmentation”
2606.29448 -
stat.ML arXiv submitted 2026-06-28Unsupervised maps cut regional coverage gaps in conformal prediction
Louis Berthier +4 · “Self-Organized Conformal Prediction: Reducing Regional Coverage Gaps with Unsupervised Group Discovery”
2606.29403 -
stat.AP arXiv submitted 2026-06-28Bayesian CDD alone stays accurate on gene directions across all DREAM5 networks
Xiaoying Wei +1 · “Bayesian Copula Directional Dependence is Cross-Network Robust for Gene-Regulatory Pair Direction: A Benchmark Study on DREAM5”
2606.29402 -
stat.AP arXiv submitted 2026-06-28Data on 37 nurses disproves modeling assumption in roster analysis
Richard D. Gill · “Critique of "Use of roster charts in the investigation and prosecution of nurses ..." by John O' Quigley”
2606.29394 -
stat.AP arXiv submitted 2026-06-28Language model embeddings beat hand-crafted features in car insurance pricing
Christopher Blier-Wong +1 · “Semantic insurance pricing with large language models”
2606.29371 -
cs.LG arXiv submitted 2026-06-28MMLU leaderboard ranks widen when subjects are the sampling unit
Bitya Neuhof +1 · “Quantifying Ranking Uncertainty in LLM Benchmarks”
2607.16259 -
cs.LG arXiv submitted 2026-06-28Symbolic trees generalize at rate L^d over sqrt(n)
\c{S}uayp Talha Kocabay +2 · “Sample Complexity of Scientific Discovery: PAC Learnability of Compositional Function Trees”
2606.29331 -
stat.ML arXiv submitted 2026-06-28Gradient boosting extended to vector outputs in tree leaves
David Cortes · “Gradient boosting with vector-valued leafs”
2606.29326 -
stat.ML arXiv submitted 2026-06-28Attention operator compresses distributions losslessly for Transformers
Peilin Liu +1 · “Generalization Analysis of Transformers in Distribution Regression”
2606.29256 -
cs.LG arXiv submitted 2026-06-28Model tracks Sri Lankan vegetable price surges at 86 percent accuracy on unseen 2024 data
Ranuga Weerasekara +8 · “When Prices Double in a Week: Forecasting of Agricultural Volatility in Import-Isolated Markets”
2606.29248 -
cs.LG arXiv submitted 2026-06-28Prices that double in a week are still forecastable
Ranuga Weerasekara +8 · “When Prices Double in a Week: Forecasting of Agricultural Volatility in Import-Isolated Markets”
2606.29248 -
stat.CO arXiv submitted 2026-06-28Gaussian VI preconditions HMC and VAE guides MH sampling
Pingping Yin +1 · “Using Variational Inference to Improve the Efficiency of MCMC Algorithms”
2606.29205 -
cs.LG arXiv submitted 2026-06-28Abstention budget turns polynomial error decay exponential in Bayesian arm ID
Yuqi Huang +2 · “Bayesian Best-Arm Identification with Abstention: A Polynomial-to-Exponential Phase Transition”
2606.29203 -
cs.LG arXiv submitted 2026-06-28Gauge-equivariant preconditioner cuts LM loss gap to 0.67 from 5.88
Tejas Pradeep Shirodkar · “Dead-Direction Conditioners: Gauge-Equivariant Preconditioning for Deep Networks”
2606.29176 -
stat.ME arXiv submitted 2026-06-27Multivariate BART obtains first contraction rates with joint residual dependence
Soham Ghosh +1 · “Multivariate Varying-Coefficient BART with Graphical Horseshoe Priors”
2606.29114 -
cs.LG arXiv submitted 2026-06-27Hutchinson-free objective cuts variance in few-step flow models
RuiKang OuYang +9 · “Few-Step Boltzmann Generators via Scalable Likelihood Flow Maps”
2606.29110 -
math.ST arXiv submitted 2026-06-27Mixture posteriors adapt to true K with extra mass vanishing at n^{-1/2}
Filippo Ascolani · “Posterior concentration and adaptation of the mixing measure in Dirichlet process mixtures”
2606.29109 -
stat.ME arXiv submitted 2026-06-27Panel flow matching pools sparse longitudinal data into full distributions
Jianbin Tan +2 · “Panel Flow Matching: A Generative Approach to Learning Distributions of Longitudinal Data”
2606.29105 -
stat.CO arXiv submitted 2026-06-27spca R package produces uncorrelated sparse principal components
Giovanni Maria Merola · “spca: An R package to Compute Least Squares Sparse Principal Components”
2606.29104 -
stat.ML arXiv submitted 2026-06-27Relaxed noise assumptions yield directed brain connectivity estimates
Stephan Goerttler +2 · “Connectivity Estimation using Stochastic Graph Heat Modelling”
2606.29098 -
stat.ME arXiv submitted 2026-06-27This paper develops a machine learning estimator for heterogeneous causal effects within…
Jiaqi Tong +1 · “Doubly cross-fit debiased machine learning of heterogeneous treatment effects under principal stratification”
2606.29076 -
stat.ME arXiv submitted 2026-06-27Machine learner finds treatment-effect curves inside latent strata
Jiaqi Tong +1 · “Doubly cross-fit debiased machine learning of heterogeneous treatment effects under principal stratification”
2606.29076 -
stat.ME arXiv submitted 2026-06-27Model gives closed forms for discrete circular plus linear data
Brajesh Kumar Dhakad +1 · “On Modeling Cylindrical Data with a Discrete Circular Component and Its Environmental Applications”
2606.29041 -
stat.ME arXiv submitted 2026-06-27Beta-tree test flags local model deviations with finite-sample bounds
Valerie N.P. Ho +1 · “Beta-trees for testing multivariate goodness-of-fit and localizing deviations from a model”
2606.29021 -
econ.EM arXiv submitted 2026-06-27Regret statistic classifies algo liquidity demand from data alone
Irene Aldridge · “Liquidity-Based Audit of Algorithmic Trading Strategies”
2606.29018 -
stat.ME arXiv submitted 2026-06-27OLS makes three recursive causal estimators identical in finite samples
Wisse Rutgers +1 · “Generated outcomes as generated regressors: Equivalences in recursive causal estimation”
2606.29009 -
math.OC arXiv submitted 2026-06-27sBDCA solves LTS up to 3.25x faster than Fast-LTS with better objectives
Marah-Lisanne Thormann +3 · “Faster than Fast-LTS: Robust Regression and Outlier Detection with DC Programming”
2606.28974 -
cs.AI arXiv submitted 2026-06-27Specialized clinical AI beats general models by 25-39 points on real questions
Jean Feng +7 · “Expert Evaluation of Clinical AI Tools on Real Point-of-Care Clinical Queries”
2606.28960 -
stat.ML arXiv submitted 2026-06-27Latent GP calibration achieves 95% coverage in aerodynamic uncertainty
Geoffrey Davis +1 · “A Bayesian latent Gaussian process framework for aerodynamic uncertainty quantification”
2606.28871 -
stat.ML arXiv submitted 2026-06-27Factor indeterminacy resolves at infinite feature scale
Carel F.W. Peeters · “Perspectives on Latent Factor Indeterminacy and its Implications for Data Representation”
2606.28854 -
econ.EM arXiv submitted 2026-06-27Selectivity correction shrinks effects to 12-21% of published averages
Peter Ganong +2 · “Literature Review and Evidence Aggregation: a Toolkit for Applied Micro”
2606.28848 -
stat.AP arXiv submitted 2026-06-27Methods correct errors in both outcomes and covariates
Pamela A. Shaw +1 · “Methods to address measurement error in both Outcome and Covariates”
2606.28810 -
stat.ML arXiv submitted 2026-06-27Anti-symmetric perturbation cuts fluctuation constant in non-reversible Langevin MC
Bingye Ni +3 · “Variance Reduction for Stochastic Gradient Generalized Non-reversible Langevin Monte Carlo Algorithms”
2606.28808 -
cs.LG arXiv submitted 2026-06-27When to call an LLM? A risk threshold
Zhaohui Wang · “Uncertainty-Aware Sequential Decision Rules for Event-Triggered LLM Invocation in Streaming Systems”
2607.13048 -
cs.AI arXiv submitted 2026-06-27LLM probe unifies EHR modalities for 87.69% ICD accuracy
Chengyuan Liu +3 · “Primary ICD Category Prediction using LLM-based Probing”
2606.28798 -
cs.LG arXiv submitted 2026-06-27Known sampling designs yield unbiased ML predictions without models
Li-Chun Zhang +4 · “On design-unbiased algorithmic Machine Learning”
2606.28795 -
stat.AP arXiv submitted 2026-06-27Imputation ensembles form stable connected structures not random clouds
Arturo Tozzi · “Topological reconstruction of Rubin multiple imputation via coarse proximity, Seifert van Kampen gluing and Hurewicz invariants”
2606.30684 -
stat.ME arXiv submitted 2026-06-27Latent trait measurement error biases causal estimates
George Perrett +1 · “Measurement Induced Confounding”
2606.28774 -
stat.ME arXiv submitted 2026-06-27Measurement error in latent traits biases treatment effect estimates
George Perrett +1 · “Measurement Induced Confounding”
2606.28774 -
stat.ME arXiv submitted 2026-06-27Estimator recovers causal effects despite confounding and missing data
Shiyao Xu +3 · “Inferring Comprehensive Cohort Causal Effects in the Presence of Unmeasured Confounding and Missing Outcomes”
2606.28741 -
stat.ME arXiv submitted 2026-06-27Ray-based model scales projected Gaussians to high-dimensional compositional data
Michael R Schwob +1 · “Composition as Direction: An Active-Set Ray-Based Model for Sparse High-Dimensional Compositional Data”
2606.28738 -
math.ST arXiv submitted 2026-06-27Common stochastic fix fails for full conformal prediction
Thanawat Sornwanee · “Full Conformal Prediction under Stochastic Non-Conformity Measure”
2606.28730 -
stat.ME arXiv submitted 2026-06-27IPW reframed as KL reweighting for post-Bayesian bias correction
Owen Thomas +2 · “Inverse Probability Weighting in a Post-Bayesian World”
2606.28685 -
cs.AI arXiv submitted 2026-06-27Nine LLM families show 90.3% consistent virtue rankings
Ioannis Tzachristas +1 · “Aristotelian Virtue Profiling of LLMs through Ethical Dilemmas”
2606.28683 -
math.PR arXiv submitted 2026-06-27Stein operator derived for symmetric matrix normal from OU generator
Robert E. Gaunt +1 · “Stein's method for the symmetric matrix normal distribution with an application to the approximation of the Wishart law”
2606.28678 -
cs.LG arXiv submitted 2026-06-27Extra samples stop helping answer selection after a few dozen draws
Yong Yi Bay +1 · “When More Sampling Hurts: The Modal Ceiling and Correlation Ceiling of Test-Time Scaling”
2606.28661 -
cs.LG arXiv submitted 2026-06-26Bayes updates track regret as exact information over intrinsic time
Akshay Balsubramani · “Adaptive Bayes exactly tracks information over intrinsic time”
2607.08789 -
stat.ML arXiv submitted 2026-06-26Adaptive hard thresholding attains logarithmic regret for online quantile regression
Zitian Zhou +1 · “Adaptive Iterative Hard Thresholding for Online High-dimensional Quantile Regression”
2606.28652 -
math.ST arXiv submitted 2026-06-26Shape regularity required for optimal local regression rates
J\'er\'emy Bettinger +2 · “Revisiting local regression: shape regularity, uniform rates, and the limits of random splits”
2606.28641