Statistics, Machine Learning papers, 2026-03-01 to 2026-08-31
Every paper in the arxiv corpus whose primary field is the arXiv category stat.ML, submitted inside this window. An impact prediction on exactory states a citation rank against this population.
- Corpus
- arxiv
- Papers
- 1,406
- Collected
- 7 Sept 2026
Papers
- Coupled Training with Privileged Information and Unlabeled Data2605.23268 · 22 May 2026
- Move on Muon : A Hamiltonian probability gradient flow perspective of Muon optimizer2605.23871 · 22 May 2026
- Operationalizing Individual Fairness via Gradient Descent and Bradley-Terry Models2605.23145 · 22 May 2026
- Concomitant DAG Learning: On the Roles of Noise Adaptivity, Sparsity, and Non-negativity2605.23537 · 22 May 2026
- Optimal Non-Asymptotic Edgeworth Expansions for Multivariate Neural Network Outputs2605.24072 · 22 May 2026
- Dirichlet-Based Monte Carlo Dropout for Uncertainty Estimation in Neural Networks2605.23635 · 22 May 2026
- Learning Kernel-Based MDPs from Episodic Preferential Feedback2605.23650 · 22 May 2026
- MEDAL: Manifold Embedding Distillation via Autoencoder Learning2605.24244 · 22 May 2026
- On the Stability of Spherical Hellinger-Kantorovich Flows and Their Implications for Differential Privacy2605.23879 · 22 May 2026
- How Neural Reward Models Learn Features for Policy Optimization: A Single-Index Analysis2605.24749 · 23 May 2026
- Affinity Graph Connectivity in Convex Clustering2605.24673 · 23 May 2026
- Clustering based on Stochastic Dominance with application for risk averters and risk seekers2605.24422 · 23 May 2026
- Multicalibration Boosting: Theory, Convergence, and Transferability2605.24364 · 23 May 2026
- Estimating Mixture Distributions via Stochastic Mirror Descent2605.24929 · 24 May 2026
- Nyström Kernel Stein Discrepancy Tests2605.25173 · 24 May 2026
- Counterfactually Safe Reinforcement Learning2605.25114 · 24 May 2026
- Choosing Online Experiment Designs under Interference in Ads, Recommendations, and Member-Experience Systems2605.25290 · 24 May 2026
- Statistical Inference for Stochastic Gradient Descent: Beyond Finite Variance2605.26000 · 25 May 2026
- Optimal Design for Multinomial Logit Model with Applications to Best Assortment Identification2605.25592 · 25 May 2026
- Efficient Benchmarking Is Just Feature Selection and Multiple Regression2605.25773 · 25 May 2026
- Rao-Blackwellized Score Matching on Manifolds2605.25567 · 25 May 2026
- Learning Sparse Compositional Functions with Norm-Constrained Neural Networks2605.25608 · 25 May 2026
- Learning manifold diffusion semigroups from graph transition matrices2605.25383 · 25 May 2026
- StrTransformer: Source-Wise Structured Transformers for Unsupervised Blind Source Recovery2605.25648 · 25 May 2026
- When Does LeJEPA Learn a World Model?2605.26379 · 25 May 2026
- From DPPs to $k$-DPPs: identifiability analysis via spectral decomposition2605.25526 · 25 May 2026
- Guided Flow Matching for Forward and Inverse PDE Problems with Sparse Observations: Algorithm and Theory2605.25509 · 25 May 2026
- Learning Nonlinear Factor Models with Unknown Monotone Links from Incomplete and Noisy Data2605.26271 · 25 May 2026
- Beyond Differences: Doubly Robust Meta-Learners for Ratio-Based Treatment Effects2605.26288 · 25 May 2026
- PAC Learning with Bandit Feedback: Sharp Sample Complexity in the Realizable Setting2605.25678 · 25 May 2026
- DiscoverPhysics: Benchmarking LLMs for Out-of-the-Box Scientific Thinking2605.26087 · 25 May 2026
- Nonstationary Generalized Linear Bandits with Discounted Online Mirror Descent2605.25590 · 25 May 2026
- Mean-Shift PCA by Knockoff Mean2605.25460 · 25 May 2026
- Unsupervised Identification and Removal of Spurious Correlations During Fine-Tuning2605.27676 · 26 May 2026
- Constrained Bayesian Experimental Design via Online Planning2605.26990 · 26 May 2026
- Stop Suppressing the Tail: Causal Inference for Extreme Events2605.27474 · 26 May 2026
- Triangular-Reference Schrödinger Bridges for Time Series Generation2605.27478 · 26 May 2026
- Gaussian Process-based learning with new MCMC-based implementation of Wishart prior on correlation matrix2605.27093 · 26 May 2026
- Identifiable Bayesian Deep Generative Copulas with Unknown Layer Widths for Data with Arbitrary Marginal Distributions2605.27523 · 26 May 2026
- CART Random Forests as Sequential Allocation over Random Opportunity Sets: A Stochastic-Control Theory of Ensemble Risk2605.26675 · 26 May 2026
- Evolving and Detecting Multi-Turn Deception using Geometric Signatures2605.27671 · 26 May 2026
- Signal-to-Noise Ratio and Sample Size Govern Representational Alignment in Neural Networks2605.26973 · 26 May 2026
- Semiparametrically Efficient Inference for Kernel Measures of Noise Heterogeneity2605.27526 · 26 May 2026
- Iterative Causal Discovery: Per-Edge Impossibility Certificates, Tier-Aware Oracle Queries, and the $1+K$ Lower Bound2605.27477 · 26 May 2026
- Soft Specialists: $α$-Rényi Ensembles for Uncertainty-Aware LLM Post-Training2605.27747 · 26 May 2026
- Accelerating Reinforcement Learning Training Using Simulation Surrogate Models2605.27556 · 26 May 2026
- Causal Representation Learning for Generalisable Recommendation2605.27043 · 26 May 2026
- Calibrated Inference for the Conditional Average Treatment Effect in the Few-Placebo Regime via Gaussian Processes2605.27473 · 26 May 2026
- Transformers Can Learn Posterior Predictive Distributions In-Context2605.26713 · 26 May 2026
- Counterfactually Fair Regression via Optimal Transport2605.28251 · 27 May 2026
- Variance-Adaptive Optimal Algorithm for Reinforcement Learning with Multinomial Logit Function Approximation2605.28364 · 27 May 2026
- Conservative neural posterior estimation via distributionally robust training2605.28516 · 27 May 2026
- Bridging Maximum Likelihood and Optimal Transport for Efficient Inference and Model Selection in Stochastic Block Models2605.28488 · 27 May 2026
- Learning to target with network interference2605.27794 · 27 May 2026
- Gradient-Flow Optimization as Dynamic Random-Effects Inference: Testing and Early Stopping with Applications to Deep Learning2605.27991 · 27 May 2026
- Geometry of Relaxed Fair Regression: A Unified Framework for Aware and Unaware Settings2605.28233 · 27 May 2026
- Dynamics of Stochastic Momentum with Sparse Updates in High Dimensions2605.28961 · 27 May 2026
- Insurance Pricing Optimization via Off-Policy Evaluation2605.28327 · 27 May 2026
- Diagnosing the conditional-mean barrier in scientific machine-learning surrogates2605.28076 · 27 May 2026
- Beyond Lipschitz: Data-Driven Robustness via Discrete Modulus of Continuity2605.28729 · 27 May 2026
- Decision-focused learning for optimal PV-Battery scheduling2605.28340 · 27 May 2026
- Is Backpropagation Optimal? When Synthetic Gradients Improve Sample Efficiency2605.27946 · 27 May 2026
- Anytime-Valid Federated Conformal RAG for LLM Swarms2605.29139 · 27 May 2026
- Leave a Window Out: Modifying the Jackknife for Predictive Inference in Time Series2605.30292 · 28 May 2026
- Diffusion Models Are Statistically Optimal for Learning Low-Dimensional Multi-Modal Distributions2605.30153 · 28 May 2026
- Wasserstein Contraction of Coordinate Ascent Variational Inference2605.30253 · 28 May 2026
- Reward Learning from Best-of-$N$ Preference Data: Targets, Tradeoffs, and Design Principles2605.30619 · 28 May 2026
- Instance-dependent Stochastic Lipschitz bandit2605.29748 · 28 May 2026
- Improved Distribution Estimation in $\ell_\infty$2605.30509 · 28 May 2026
- Prediction-Powered Inference Across Many Tasks for AI Evaluation & Social Science Research2605.29249 · 28 May 2026
- Eigen-Spike Emergence and Quadratic Equivalents for Conjugate Kernels on Nonlinearly Separable Data2605.29669 · 28 May 2026
- Visual Spatial Learning: Single-Field Spatial Interpolation Using Convolutional Neural Networks2605.30167 · 28 May 2026
- Joint Model and Data Sparsification via the Marginal Likelihood2605.29908 · 28 May 2026
- Matching Rates and Optimal Allocation for Federated Probe-Logit Distillation under Heterogeneous Bandwidth Budgets2605.29642 · 28 May 2026
- Improved Guarantees for Heterogeneous Treatment-Effect Estimation via Matrix Completion2605.30319 · 28 May 2026
- Deep Optimal Individualized Treatment Rules for Bivariate Survival Outcomes via Adaptive Prediction-Powered Learning2605.29464 · 28 May 2026
- Correcting Split Selection in Online Decision Trees via Anytime-Valid Inference2605.31239 · 29 May 2026
- Parameter-Free and Group Conditional Online Conformal Prediction2606.00419 · 29 May 2026
- Free energy Estimation on Any State Space2605.31063 · 29 May 2026
- Hedging on the Frontier: Learning New Tasks with Few Samples2605.30997 · 29 May 2026
- Counterfactual Explanations for Deep Two-Sample Testing2606.04009 · 29 May 2026
- Approximation and learning of anisotropic and mixed smooth functions by deep ReLU neural networks2605.31152 · 29 May 2026
- Out-of-Distribution generalization of quantile regression with heavy tailed inputs: an SVM approach2606.00265 · 29 May 2026
- Routing on the Stiefel Manifold: When Does Adaptive Subspace Selection Help for Cross-Domain EEG Decoding?2605.31043 · 29 May 2026
- Entropic Projection Alignment: Estimating, Explaining, and Improving Model Performance Under Distribution Shift2605.31250 · 29 May 2026
- Log-Ratio Propagation on the Simplex: A Theory of Cellwise Contamination for Compositional Data2605.31345 · 29 May 2026
- Memory by Design: Probabilistic Sequence Layers2605.31163 · 29 May 2026
- Batched Stochastic Linear Bandits with 1-Bit Communication Constraints2605.30976 · 29 May 2026
- ERICA: Quantifying Replicability of Cluster Analysis2606.00302 · 29 May 2026
- Interpreting FCDNNs via RG on Exponential Family2606.00157 · 29 May 2026
- Is the Last Layer Sufficient for Uncertainty Quantification?2605.30741 · 29 May 2026
- Riemannian Stochastic Optimization for Sufficient Dimension Reduction2606.00413 · 29 May 2026
- Is Zero-Shot Super-Resolution Possible in Operator Learning?2606.00296 · 29 May 2026
- Spectra-Guided Neural Tucker Factorization2606.00584 · 30 May 2026
- On Finite-sample Concentration of Median of Incomplete U-Statistics2606.00661 · 30 May 2026
- Position: Prioritize Identifying Structure, Not Complex Models, for Scientific Discovery2606.02632 · 30 May 2026
- Statistical Analysis of using the Shapley Value for Sensor Anomaly Localization with Accurate Classifiers2606.00867 · 30 May 2026
- Statistical Testing on Directed Graphs by Surrogate Data Generation2606.00758 · 30 May 2026
- Taming the Loss Landscape of PINNs with Noisy Feynman-Kac Supervision: Operator Preconditioning and Non-Asymptotic Error Bounds2606.00643 · 30 May 2026
- Bandit Simulation for Average Reward Inference2606.00913 · 30 May 2026