INDEX
Every paper Pith has read and judged, in one searchable index.
-
cs.LG arXiv submitted 2026-09-23A spectral fix lifts attention accuracy in 12 of 12 runs
Xiaohe Jiang (1) +3 · “NS-ATTENTION: Newton-Schulz Transformations of Attention Outputs in Vision Transformers”
2609.27735 -
cs.LG arXiv submitted 2026-09-23Risk-gated adversaries win on all four MuJoCo control tasks
Jiaxi Wu +4 · “Robust Adversarial Reinforcement Learning with Risk Sensitivity and Critic Consistency Regularization”
2609.27667 -
cs.LG arXiv submitted 2026-09-23Recursive gradients shrink privacy noise in decentralized learning
Yizhao Fan +2 · “Private Decentralized Optimization with Noise Reduction and Bias Correction”
2609.27658 -
cs.LG arXiv submitted 2026-09-23Memory of failed tries turns greedy decoding into search
Oleksii Streltsov +1 · “FLEET: From Logits Entropy to Enhanced Trajectories in Text Generation”
2609.27657 -
stat.ML arXiv submitted 2026-09-23Federated income scoring trails pooled data by just 1.8%
Sultan Amed +2 · “FedIncome: Federated Learning for Income Estimation in Digital Lending Under Data Sovereignty Constraints”
2609.27654 -
cs.LG arXiv submitted 2026-09-23LoReST cuts large-scale traffic forecast error on every tested network
Qi Feng +6 · “Learning Local Heterogeneity and Cross-Region Context for Large-Scale Traffic Forecasting”
2609.27637 -
cs.LG arXiv submitted 2026-09-23Clustering makes bandit sketching exact
Hantao Yang +2 · “Efficient Linear Bandits via Cluster-Aware Sketching”
2609.27594 -
cs.LG arXiv submitted 2026-09-23Concept erasure hides its target
Akash Samanta +2 · “Hidden not Deleted: How Networks Suppress Entangled Features”
2609.27593 -
cs.LG arXiv submitted 2026-09-23Below 59M parameters, Step Law's form survives
Egor Romanyukov +5 · “Does Step Law Transfer to Small-Scale Language Models? An Empirical Recalibration Below 59M Parameters”
2609.27581 -
cs.LG arXiv submitted 2026-09-23Gradient noise sets a per-modality momentum for joint training
Zhongjing Gu +5 · “VCMM: Variance-Calibrated Momentum for Multimodal Learning”
2609.27577 -
cs.LG arXiv submitted 2026-09-23A 4B model approaches 235B once the reward judge co-evolves
Henan Sun +6 · “DCRL: Decoupling and Coupling Reinforcement Learning via Policy-Reward Manifold Alignment”
2609.27572 -
cs.LG arXiv submitted 2026-09-23Data-fit neuron formulas beat generic ones in every test
Meng Wang +7 · “TNLearn: An Open Source Python Package for Task-based Neurons”
2609.27564 -
cs.LG arXiv submitted 2026-09-23Treating physics data as PDE fields beats tokenizing it
Henan Sun +6 · “PhyMo: A Physical-Field Modality for Multimodal AI4Physics”
2609.27554 -
cs.LG arXiv submitted 2026-09-23Idle GPUs reclaimed for 3.5x embodied RL training speed
Liang Mi +14 · “EBRL: Asynchronous Embodied RL by Multi-Grained Resource Management”
2609.27547 -
math.ST arXiv submitted 2026-09-23Distribution shift hits diffusion scores as radius squared
Wei Luo +3 · “Robustness of Diffusion Models under Distribution Shift”
2609.27546 -
cs.LG arXiv submitted 2026-09-23Verified progress beats outcome rewards when credited per turn
Ming Ma +9 · “ProCredit: From Outcome Rewards to Progress Credit in Agentic Reinforcement Learning”
2609.27532 -
cs.CV arXiv submitted 2026-09-23Mammography encoder posts 97.78% validation accuracy
Zheng Yu +4 · “M3D-Net: Hierarchical Coordination of Spatial Context, Feature Reuse, and Differential Attention for Mammography Classification”
2609.27523 -
cs.AI arXiv submitted 2026-09-23Agents pick the best config while understanding almost nothing
Jingjie Ning +3 · “WhatWorkedBench: Benchmarking Experimental Understanding in AI Agents”
2609.27490 -
cs.CV arXiv submitted 2026-09-23State drift beats cache signals for streaming video memory
Taeyoun Kwon +3 · “DeltaS: Reading the Gated Linear Attention State for KV Cache Eviction in Streaming Video”
2609.27470 -
cs.LG arXiv submitted 2026-09-23Small quantum circuit learns cost-delay-aware cloud scheduling
An N. H. Phan +3 · “Quantum Reinforcement Learning for Cost and Delay Tradeoffs in Quantum Cloud Orchestration”
2609.27446 -
cs.LG arXiv submitted 2026-09-23A 1.5B student beats its 7B teacher on constraint following
Yanzhao Zheng +8 · “Counterfactual Constraint-Conditioned On-Policy Distillation for Multi-Constraint Instruction Following”
2609.27421 -
cs.LG arXiv submitted 2026-09-23Small oscillatory model tops 8 rivals on scarce vibration labels
Mainak Mallick +1 · “When Labels Are Scarce: An Oscillatory State Space Model for Vibration Diagnosis”
2609.27411 -
cs.LG arXiv submitted 2026-09-23Active learning's label savings barely survive validation
Ben McEwen +2 · “Active Learning for Biodiversity Monitoring: From Label Efficiency to Reliable Ecological Inference”
2609.27409 -
cs.SD arXiv submitted 2026-09-23Self-evolving audio models gain up to 6.3 points
Yuxiang Wang +8 · “EvoAudio: Recursive Self-Improvement for Audio Understanding”
2609.27389 -
cs.LG arXiv submitted 2026-09-23An agent beat fixed forecast policies with 2.5% of the budget
Shunya Nagashima · “Forecast Workflow Bench: Evaluating Language-Model Decisions with Budgeted Forecast Tools”
2609.27385 -
cs.CL arXiv submitted 2026-09-23Recurrent language models keep recomputing global attention at every recurrent step
Ke Wan +1 · “Attention Routing Stabilizes Early: Working-Set Inference for Recurrent Language Models”
2609.27373 -
cs.LG arXiv submitted 2026-09-23AUC bound tunes anomaly detectors with zero anomaly labels
Kevin Wilkinghoff +1 · “Anomaly-Free Self-Optimization via AUC Bounds”
2609.27362 -
cs.LG arXiv submitted 2026-09-23Curvature ratio finds the weights that kill forgetting under 4-bit
Jialu Wang +8 · “Quantization-Robust Unlearning through the Lens of Retain-Forget Loss Landscapes Interaction”
2609.27355 -
eess.SY arXiv submitted 2026-09-23LLM-evolved code beats static 5G slicing by 44.5%
Faezeh Dehghan Tarzjani +1 · “Evolving Inspectable O-RAN Slicing xApps with LLMs”
2609.27337 -
cs.RO arXiv submitted 2026-09-23Decoupling safety from tactics yields unexploitable robots
Ruihan Wu +4 · “Turning Safety into Competence: Minimally Exploitable Robot Policies via Safety-Filtered Reinforcement Learning”
2609.27312 -
cs.SE arXiv submitted 2026-09-23Fairness test finds up to 412% more unfair MARL episodes
Xiaotong Wang +1 · “FairTest: Search-Based Fairness Testing for Multi-Agent Reinforcement Learning Systems”
2609.27309 -
cs.LG arXiv submitted 2026-09-23Normalized diffusion reaches 3D spin lattices
Kewen Pan +1 · “Discrete Diffusion Models via Evolving Variational Autoregressive Networks”
2609.27306 -
cs.LG arXiv submitted 2026-09-23One policy solves whether, when, whom, and what to say live
Shujian Gao +8 · “Live Assistant: Learning Whether, When, and Whom to Assist in Real-World Live Social Streams”
2609.27303 -
stat.ME arXiv submitted 2026-09-23More data can't cure IV exclusion bias
Spandan Ghose Chowdhury · “Beyond the Illusion of Power: Calibrating Quasi-Experiments in Observational IS”
2609.27299 -
cs.LG arXiv submitted 2026-09-23Cheaper prefill lets a 67B model beat 47B and 63B at equal compute
Zhiheng Hu +10 · “KITE: KV-Invariant Transformer Expansion for Efficient Agentic LLM Scaling”
2609.27294 -
cs.LG arXiv submitted 2026-09-23Network size becomes one differentiable count
Lixing Li · “NGN: Learning Neural Network Size as a Differentiable Count”
2609.27291 -
cs.LG arXiv submitted 2026-09-23Frozen LLM plus verified rules beats retrained fraud detectors
Xuwei Tan +2 · “SR-Fraud: An Outcome-Supervised Reflective LLM Agent Framework for Non-Stationary Payment Fraud Detection”
2609.27287 -
stat.ML arXiv submitted 2026-09-23Pairwise fusion hits the optimal multitask rate
Xiaodong Li +1 · “Multitask Regression with Pairwise Fusion”
2609.27280 -
cs.LG arXiv submitted 2026-09-23Spectral priors keep graph lasso convex and lift recovery
Mingxiao Liu (1) +8 · “Graph Learning with Spectral Connectivity Priors for Scarce Data”
2609.27278 -
cs.AI arXiv submitted 2026-09-23Self-grown tool library lifts accuracy on all 30 task runs
Jie Yang +10 · “TimeEvo: Failure-Driven Self-Evolution of a Time Series Agent”
2609.27277 -
cs.LG arXiv submitted 2026-09-23Models converge on relations
Junwon You +3 · “What Converges in the Platonic Representation Hypothesis? Structure over Geometry”
2609.27252 -
cs.LG arXiv submitted 2026-09-23Attention masks turn an LLM into a near-perfect text autoencoder
Arkanath Pathak +2 · “Repurposing Pre-trained LLMs as High Fidelity Continuous Text Autoencoders”
2609.27248 -
cs.LG arXiv submitted 2026-09-23First full-covariance smoother for Bayesian neural nets
Oren Wright +5 · “Full-Covariance Smoothing of Bayesian Neural Networks for Online Adaptation”
2609.27244 -
stat.ML arXiv submitted 2026-09-23Membership queries make hard classes exponentially easy
Ganghua Wang +1 · “On the Sample Complexity of Active Learning with Membership Queries”
2609.27241 -
cs.LG arXiv submitted 2026-09-23Top scorer never uses its documented compound input
Mengran Li +5 · “Discover, Falsify, Revise: Auditing Input-Use Claims from Source Code to Predictive Contribution in Agent-Discovered Cell Models”
2609.27234 -
cs.LG arXiv submitted 2026-09-23Compute alone does not predict fMRI model performance
Wenhao Ye +5 · “A Scaling Study for fMRI Foundation Models”
2609.27232 -
cs.SD arXiv submitted 2026-09-23Lung-sound AI spots pneumonia in nursing homes
Nicholas Rasmussen +7 · “Physiologically Informed Digital Auscultation for Pneumonia Detection in Long-term Care Residents”
2609.27222 -
cs.LG arXiv submitted 2026-09-23Learning ellipsoid shape from the worst tail cuts severe misses
Xiang Zhang · “Tail-Aware Geometry Learning for Conformal Ellipsoids”
2609.27221 -
cs.CE arXiv submitted 2026-09-23Neural surrogate cuts topology-optimization time 15 to 110x
Shengyu Yan +1 · “KATOsuper: Surrogate-accelerated neural topology optimization with sensitivity-consistent Fourier neural operators”
2609.27216 -
cs.LG arXiv submitted 2026-09-23Curvature-guided sampling wins on six of seven GNN sets
Chaoqun Fei +3 · “Scalable Subgraph Sampling via Resistance Curvature”
2609.27209