Pith. sign in

INDEX

Every paper Pith has read and judged, in one searchable index.

12,268 reviewed papers in stat · newest first · page 34

Pith rank means recent reader upvotes with each upvote's ranking weight halved after 24 hours. The archive itself is sorted by arXiv submission date.

  1. cs.LG arXiv submitted 2026-06-28
    Expert priors enter best-subsets MIO as log-odds penalties

    Nolan Alexander +1 · “A Mathematical Optimization Approach for Expert-Informed Bayesian Best Subset Selection”

    2606.29516
  2. stat.ME arXiv submitted 2026-06-28
    Bayesian model segments satellite images without local labels

    Bao Khanh Nguyen +3 · “Scalable Bayesian Spatial Mixture Modelling for Remote Sensing Image Segmentation”

    2606.29448
  3. stat.ML arXiv submitted 2026-06-28
    Unsupervised maps cut regional coverage gaps in conformal prediction

    Louis Berthier +4 · “Self-Organized Conformal Prediction: Reducing Regional Coverage Gaps with Unsupervised Group Discovery”

    2606.29403
  4. stat.AP arXiv submitted 2026-06-28
    Bayesian CDD alone stays accurate on gene directions across all DREAM5 networks

    Xiaoying Wei +1 · “Bayesian Copula Directional Dependence is Cross-Network Robust for Gene-Regulatory Pair Direction: A Benchmark Study on DREAM5”

    2606.29402
  5. stat.AP arXiv submitted 2026-06-28
    Data on 37 nurses disproves modeling assumption in roster analysis

    Richard D. Gill · “Critique of "Use of roster charts in the investigation and prosecution of nurses ..." by John O' Quigley”

    2606.29394
  6. stat.AP arXiv submitted 2026-06-28
    Language model embeddings beat hand-crafted features in car insurance pricing

    Christopher Blier-Wong +1 · “Semantic insurance pricing with large language models”

    2606.29371
  7. cs.LG arXiv submitted 2026-06-28
    MMLU leaderboard ranks widen when subjects are the sampling unit

    Bitya Neuhof +1 · “Quantifying Ranking Uncertainty in LLM Benchmarks”

    2607.16259
  8. cs.LG arXiv submitted 2026-06-28
    Symbolic trees generalize at rate L^d over sqrt(n)

    \c{S}uayp Talha Kocabay +2 · “Sample Complexity of Scientific Discovery: PAC Learnability of Compositional Function Trees”

    2606.29331
  9. stat.ML arXiv submitted 2026-06-28
    Gradient boosting extended to vector outputs in tree leaves

    David Cortes · “Gradient boosting with vector-valued leafs”

    2606.29326
  10. stat.ML arXiv submitted 2026-06-28
    Attention operator compresses distributions losslessly for Transformers

    Peilin Liu +1 · “Generalization Analysis of Transformers in Distribution Regression”

    2606.29256
  11. cs.LG arXiv submitted 2026-06-28
    Model tracks Sri Lankan vegetable price surges at 86 percent accuracy on unseen 2024 data

    Ranuga Weerasekara +8 · “When Prices Double in a Week: Forecasting of Agricultural Volatility in Import-Isolated Markets”

    2606.29248
  12. cs.LG arXiv submitted 2026-06-28
    Prices that double in a week are still forecastable

    Ranuga Weerasekara +8 · “When Prices Double in a Week: Forecasting of Agricultural Volatility in Import-Isolated Markets”

    2606.29248
  13. stat.CO arXiv submitted 2026-06-28
    Gaussian VI preconditions HMC and VAE guides MH sampling

    Pingping Yin +1 · “Using Variational Inference to Improve the Efficiency of MCMC Algorithms”

    2606.29205
  14. cs.LG arXiv submitted 2026-06-28
    Abstention budget turns polynomial error decay exponential in Bayesian arm ID

    Yuqi Huang +2 · “Bayesian Best-Arm Identification with Abstention: A Polynomial-to-Exponential Phase Transition”

    2606.29203
  15. cs.LG arXiv submitted 2026-06-28
    Gauge-equivariant preconditioner cuts LM loss gap to 0.67 from 5.88

    Tejas Pradeep Shirodkar · “Dead-Direction Conditioners: Gauge-Equivariant Preconditioning for Deep Networks”

    2606.29176
  16. stat.ME arXiv submitted 2026-06-27
    Multivariate BART obtains first contraction rates with joint residual dependence

    Soham Ghosh +1 · “Multivariate Varying-Coefficient BART with Graphical Horseshoe Priors”

    2606.29114
  17. cs.LG arXiv submitted 2026-06-27
    Hutchinson-free objective cuts variance in few-step flow models

    RuiKang OuYang +9 · “Few-Step Boltzmann Generators via Scalable Likelihood Flow Maps”

    2606.29110
  18. math.ST arXiv submitted 2026-06-27
    Mixture posteriors adapt to true K with extra mass vanishing at n^{-1/2}

    Filippo Ascolani · “Posterior concentration and adaptation of the mixing measure in Dirichlet process mixtures”

    2606.29109
  19. stat.ME arXiv submitted 2026-06-27
    Panel flow matching pools sparse longitudinal data into full distributions

    Jianbin Tan +2 · “Panel Flow Matching: A Generative Approach to Learning Distributions of Longitudinal Data”

    2606.29105
  20. stat.CO arXiv submitted 2026-06-27
    spca R package produces uncorrelated sparse principal components

    Giovanni Maria Merola · “spca: An R package to Compute Least Squares Sparse Principal Components”

    2606.29104
  21. stat.ML arXiv submitted 2026-06-27
    Relaxed noise assumptions yield directed brain connectivity estimates

    Stephan Goerttler +2 · “Connectivity Estimation using Stochastic Graph Heat Modelling”

    2606.29098
  22. stat.ME arXiv submitted 2026-06-27
    This paper develops a machine learning estimator for heterogeneous causal effects within…

    Jiaqi Tong +1 · “Doubly cross-fit debiased machine learning of heterogeneous treatment effects under principal stratification”

    2606.29076
  23. stat.ME arXiv submitted 2026-06-27
    Machine learner finds treatment-effect curves inside latent strata

    Jiaqi Tong +1 · “Doubly cross-fit debiased machine learning of heterogeneous treatment effects under principal stratification”

    2606.29076
  24. stat.ME arXiv submitted 2026-06-27
    Model gives closed forms for discrete circular plus linear data

    Brajesh Kumar Dhakad +1 · “On Modeling Cylindrical Data with a Discrete Circular Component and Its Environmental Applications”

    2606.29041
  25. stat.ME arXiv submitted 2026-06-27
    Beta-tree test flags local model deviations with finite-sample bounds

    Valerie N.P. Ho +1 · “Beta-trees for testing multivariate goodness-of-fit and localizing deviations from a model”

    2606.29021
  26. econ.EM arXiv submitted 2026-06-27
    Regret statistic classifies algo liquidity demand from data alone

    Irene Aldridge · “Liquidity-Based Audit of Algorithmic Trading Strategies”

    2606.29018
  27. stat.ME arXiv submitted 2026-06-27
    OLS makes three recursive causal estimators identical in finite samples

    Wisse Rutgers +1 · “Generated outcomes as generated regressors: Equivalences in recursive causal estimation”

    2606.29009
  28. math.OC arXiv submitted 2026-06-27
    sBDCA solves LTS up to 3.25x faster than Fast-LTS with better objectives

    Marah-Lisanne Thormann +3 · “Faster than Fast-LTS: Robust Regression and Outlier Detection with DC Programming”

    2606.28974
  29. cs.AI arXiv submitted 2026-06-27
    Specialized clinical AI beats general models by 25-39 points on real questions

    Jean Feng +7 · “Expert Evaluation of Clinical AI Tools on Real Point-of-Care Clinical Queries”

    2606.28960
  30. stat.ML arXiv submitted 2026-06-27
    Latent GP calibration achieves 95% coverage in aerodynamic uncertainty

    Geoffrey Davis +1 · “A Bayesian latent Gaussian process framework for aerodynamic uncertainty quantification”

    2606.28871
  31. stat.ML arXiv submitted 2026-06-27
    Factor indeterminacy resolves at infinite feature scale

    Carel F.W. Peeters · “Perspectives on Latent Factor Indeterminacy and its Implications for Data Representation”

    2606.28854
  32. econ.EM arXiv submitted 2026-06-27
    Selectivity correction shrinks effects to 12-21% of published averages

    Peter Ganong +2 · “Literature Review and Evidence Aggregation: a Toolkit for Applied Micro”

    2606.28848
  33. stat.AP arXiv submitted 2026-06-27
    Methods correct errors in both outcomes and covariates

    Pamela A. Shaw +1 · “Methods to address measurement error in both Outcome and Covariates”

    2606.28810
  34. stat.ML arXiv submitted 2026-06-27
    Anti-symmetric perturbation cuts fluctuation constant in non-reversible Langevin MC

    Bingye Ni +3 · “Variance Reduction for Stochastic Gradient Generalized Non-reversible Langevin Monte Carlo Algorithms”

    2606.28808
  35. cs.LG arXiv submitted 2026-06-27
    When to call an LLM? A risk threshold

    Zhaohui Wang · “Uncertainty-Aware Sequential Decision Rules for Event-Triggered LLM Invocation in Streaming Systems”

    2607.13048
  36. cs.AI arXiv submitted 2026-06-27
    LLM probe unifies EHR modalities for 87.69% ICD accuracy

    Chengyuan Liu +3 · “Primary ICD Category Prediction using LLM-based Probing”

    2606.28798
  37. cs.LG arXiv submitted 2026-06-27
    Known sampling designs yield unbiased ML predictions without models

    Li-Chun Zhang +4 · “On design-unbiased algorithmic Machine Learning”

    2606.28795
  38. stat.AP arXiv submitted 2026-06-27
    Imputation ensembles form stable connected structures not random clouds

    Arturo Tozzi · “Topological reconstruction of Rubin multiple imputation via coarse proximity, Seifert van Kampen gluing and Hurewicz invariants”

    2606.30684
  39. stat.ME arXiv submitted 2026-06-27
    Latent trait measurement error biases causal estimates

    George Perrett +1 · “Measurement Induced Confounding”

    2606.28774
  40. stat.ME arXiv submitted 2026-06-27
    Measurement error in latent traits biases treatment effect estimates

    George Perrett +1 · “Measurement Induced Confounding”

    2606.28774
  41. stat.ME arXiv submitted 2026-06-27
    Estimator recovers causal effects despite confounding and missing data

    Shiyao Xu +3 · “Inferring Comprehensive Cohort Causal Effects in the Presence of Unmeasured Confounding and Missing Outcomes”

    2606.28741
  42. stat.ME arXiv submitted 2026-06-27
    Ray-based model scales projected Gaussians to high-dimensional compositional data

    Michael R Schwob +1 · “Composition as Direction: An Active-Set Ray-Based Model for Sparse High-Dimensional Compositional Data”

    2606.28738
  43. math.ST arXiv submitted 2026-06-27
    Common stochastic fix fails for full conformal prediction

    Thanawat Sornwanee · “Full Conformal Prediction under Stochastic Non-Conformity Measure”

    2606.28730
  44. stat.ME arXiv submitted 2026-06-27
    IPW reframed as KL reweighting for post-Bayesian bias correction

    Owen Thomas +2 · “Inverse Probability Weighting in a Post-Bayesian World”

    2606.28685
  45. cs.AI arXiv submitted 2026-06-27
    Nine LLM families show 90.3% consistent virtue rankings

    Ioannis Tzachristas +1 · “Aristotelian Virtue Profiling of LLMs through Ethical Dilemmas”

    2606.28683
  46. math.PR arXiv submitted 2026-06-27
    Stein operator derived for symmetric matrix normal from OU generator

    Robert E. Gaunt +1 · “Stein's method for the symmetric matrix normal distribution with an application to the approximation of the Wishart law”

    2606.28678
  47. cs.LG arXiv submitted 2026-06-27
    Extra samples stop helping answer selection after a few dozen draws

    Yong Yi Bay +1 · “When More Sampling Hurts: The Modal Ceiling and Correlation Ceiling of Test-Time Scaling”

    2606.28661
  48. cs.LG arXiv submitted 2026-06-26
    Bayes updates track regret as exact information over intrinsic time

    Akshay Balsubramani · “Adaptive Bayes exactly tracks information over intrinsic time”

    2607.08789
  49. stat.ML arXiv submitted 2026-06-26
    Adaptive hard thresholding attains logarithmic regret for online quantile regression

    Zitian Zhou +1 · “Adaptive Iterative Hard Thresholding for Online High-dimensional Quantile Regression”

    2606.28652
  50. math.ST arXiv submitted 2026-06-26
    Shape regularity required for optimal local regression rates

    J\'er\'emy Bettinger +2 · “Revisiting local regression: shape regularity, uniform rates, and the limits of random splits”

    2606.28641