exactory
Sign inGet started
← Papers

Statistics, Machine Learning papers, 2026-03-01 to 2026-08-31

Every paper in the arxiv corpus whose primary field is the arXiv category stat.ML, submitted inside this window. An impact prediction on exactory states a citation rank against this population.

Corpus
arxiv
Papers
1,406
Collected
7 Sept 2026

Papers

  1. Coupled Training with Privileged Information and Unlabeled Data2605.23268 · 22 May 2026
  2. Move on Muon : A Hamiltonian probability gradient flow perspective of Muon optimizer2605.23871 · 22 May 2026
  3. Operationalizing Individual Fairness via Gradient Descent and Bradley-Terry Models2605.23145 · 22 May 2026
  4. Concomitant DAG Learning: On the Roles of Noise Adaptivity, Sparsity, and Non-negativity2605.23537 · 22 May 2026
  5. Optimal Non-Asymptotic Edgeworth Expansions for Multivariate Neural Network Outputs2605.24072 · 22 May 2026
  6. Dirichlet-Based Monte Carlo Dropout for Uncertainty Estimation in Neural Networks2605.23635 · 22 May 2026
  7. Learning Kernel-Based MDPs from Episodic Preferential Feedback2605.23650 · 22 May 2026
  8. MEDAL: Manifold Embedding Distillation via Autoencoder Learning2605.24244 · 22 May 2026
  9. On the Stability of Spherical Hellinger-Kantorovich Flows and Their Implications for Differential Privacy2605.23879 · 22 May 2026
  10. How Neural Reward Models Learn Features for Policy Optimization: A Single-Index Analysis2605.24749 · 23 May 2026
  11. Affinity Graph Connectivity in Convex Clustering2605.24673 · 23 May 2026
  12. Clustering based on Stochastic Dominance with application for risk averters and risk seekers2605.24422 · 23 May 2026
  13. Multicalibration Boosting: Theory, Convergence, and Transferability2605.24364 · 23 May 2026
  14. Estimating Mixture Distributions via Stochastic Mirror Descent2605.24929 · 24 May 2026
  15. Nyström Kernel Stein Discrepancy Tests2605.25173 · 24 May 2026
  16. Counterfactually Safe Reinforcement Learning2605.25114 · 24 May 2026
  17. Choosing Online Experiment Designs under Interference in Ads, Recommendations, and Member-Experience Systems2605.25290 · 24 May 2026
  18. Statistical Inference for Stochastic Gradient Descent: Beyond Finite Variance2605.26000 · 25 May 2026
  19. Optimal Design for Multinomial Logit Model with Applications to Best Assortment Identification2605.25592 · 25 May 2026
  20. Efficient Benchmarking Is Just Feature Selection and Multiple Regression2605.25773 · 25 May 2026
  21. Rao-Blackwellized Score Matching on Manifolds2605.25567 · 25 May 2026
  22. Learning Sparse Compositional Functions with Norm-Constrained Neural Networks2605.25608 · 25 May 2026
  23. Learning manifold diffusion semigroups from graph transition matrices2605.25383 · 25 May 2026
  24. StrTransformer: Source-Wise Structured Transformers for Unsupervised Blind Source Recovery2605.25648 · 25 May 2026
  25. When Does LeJEPA Learn a World Model?2605.26379 · 25 May 2026
  26. From DPPs to $k$-DPPs: identifiability analysis via spectral decomposition2605.25526 · 25 May 2026
  27. Guided Flow Matching for Forward and Inverse PDE Problems with Sparse Observations: Algorithm and Theory2605.25509 · 25 May 2026
  28. Learning Nonlinear Factor Models with Unknown Monotone Links from Incomplete and Noisy Data2605.26271 · 25 May 2026
  29. Beyond Differences: Doubly Robust Meta-Learners for Ratio-Based Treatment Effects2605.26288 · 25 May 2026
  30. PAC Learning with Bandit Feedback: Sharp Sample Complexity in the Realizable Setting2605.25678 · 25 May 2026
  31. DiscoverPhysics: Benchmarking LLMs for Out-of-the-Box Scientific Thinking2605.26087 · 25 May 2026
  32. Nonstationary Generalized Linear Bandits with Discounted Online Mirror Descent2605.25590 · 25 May 2026
  33. Mean-Shift PCA by Knockoff Mean2605.25460 · 25 May 2026
  34. Unsupervised Identification and Removal of Spurious Correlations During Fine-Tuning2605.27676 · 26 May 2026
  35. Constrained Bayesian Experimental Design via Online Planning2605.26990 · 26 May 2026
  36. Stop Suppressing the Tail: Causal Inference for Extreme Events2605.27474 · 26 May 2026
  37. Triangular-Reference Schrödinger Bridges for Time Series Generation2605.27478 · 26 May 2026
  38. Gaussian Process-based learning with new MCMC-based implementation of Wishart prior on correlation matrix2605.27093 · 26 May 2026
  39. Identifiable Bayesian Deep Generative Copulas with Unknown Layer Widths for Data with Arbitrary Marginal Distributions2605.27523 · 26 May 2026
  40. CART Random Forests as Sequential Allocation over Random Opportunity Sets: A Stochastic-Control Theory of Ensemble Risk2605.26675 · 26 May 2026
  41. Evolving and Detecting Multi-Turn Deception using Geometric Signatures2605.27671 · 26 May 2026
  42. Signal-to-Noise Ratio and Sample Size Govern Representational Alignment in Neural Networks2605.26973 · 26 May 2026
  43. Semiparametrically Efficient Inference for Kernel Measures of Noise Heterogeneity2605.27526 · 26 May 2026
  44. Iterative Causal Discovery: Per-Edge Impossibility Certificates, Tier-Aware Oracle Queries, and the $1+K$ Lower Bound2605.27477 · 26 May 2026
  45. Soft Specialists: $α$-Rényi Ensembles for Uncertainty-Aware LLM Post-Training2605.27747 · 26 May 2026
  46. Accelerating Reinforcement Learning Training Using Simulation Surrogate Models2605.27556 · 26 May 2026
  47. Causal Representation Learning for Generalisable Recommendation2605.27043 · 26 May 2026
  48. Calibrated Inference for the Conditional Average Treatment Effect in the Few-Placebo Regime via Gaussian Processes2605.27473 · 26 May 2026
  49. Transformers Can Learn Posterior Predictive Distributions In-Context2605.26713 · 26 May 2026
  50. Counterfactually Fair Regression via Optimal Transport2605.28251 · 27 May 2026
  51. Variance-Adaptive Optimal Algorithm for Reinforcement Learning with Multinomial Logit Function Approximation2605.28364 · 27 May 2026
  52. Conservative neural posterior estimation via distributionally robust training2605.28516 · 27 May 2026
  53. Bridging Maximum Likelihood and Optimal Transport for Efficient Inference and Model Selection in Stochastic Block Models2605.28488 · 27 May 2026
  54. Learning to target with network interference2605.27794 · 27 May 2026
  55. Gradient-Flow Optimization as Dynamic Random-Effects Inference: Testing and Early Stopping with Applications to Deep Learning2605.27991 · 27 May 2026
  56. Geometry of Relaxed Fair Regression: A Unified Framework for Aware and Unaware Settings2605.28233 · 27 May 2026
  57. Dynamics of Stochastic Momentum with Sparse Updates in High Dimensions2605.28961 · 27 May 2026
  58. Insurance Pricing Optimization via Off-Policy Evaluation2605.28327 · 27 May 2026
  59. Diagnosing the conditional-mean barrier in scientific machine-learning surrogates2605.28076 · 27 May 2026
  60. Beyond Lipschitz: Data-Driven Robustness via Discrete Modulus of Continuity2605.28729 · 27 May 2026
  61. Decision-focused learning for optimal PV-Battery scheduling2605.28340 · 27 May 2026
  62. Is Backpropagation Optimal? When Synthetic Gradients Improve Sample Efficiency2605.27946 · 27 May 2026
  63. Anytime-Valid Federated Conformal RAG for LLM Swarms2605.29139 · 27 May 2026
  64. Leave a Window Out: Modifying the Jackknife for Predictive Inference in Time Series2605.30292 · 28 May 2026
  65. Diffusion Models Are Statistically Optimal for Learning Low-Dimensional Multi-Modal Distributions2605.30153 · 28 May 2026
  66. Wasserstein Contraction of Coordinate Ascent Variational Inference2605.30253 · 28 May 2026
  67. Reward Learning from Best-of-$N$ Preference Data: Targets, Tradeoffs, and Design Principles2605.30619 · 28 May 2026
  68. Instance-dependent Stochastic Lipschitz bandit2605.29748 · 28 May 2026
  69. Improved Distribution Estimation in $\ell_\infty$2605.30509 · 28 May 2026
  70. Prediction-Powered Inference Across Many Tasks for AI Evaluation & Social Science Research2605.29249 · 28 May 2026
  71. Eigen-Spike Emergence and Quadratic Equivalents for Conjugate Kernels on Nonlinearly Separable Data2605.29669 · 28 May 2026
  72. Visual Spatial Learning: Single-Field Spatial Interpolation Using Convolutional Neural Networks2605.30167 · 28 May 2026
  73. Joint Model and Data Sparsification via the Marginal Likelihood2605.29908 · 28 May 2026
  74. Matching Rates and Optimal Allocation for Federated Probe-Logit Distillation under Heterogeneous Bandwidth Budgets2605.29642 · 28 May 2026
  75. Improved Guarantees for Heterogeneous Treatment-Effect Estimation via Matrix Completion2605.30319 · 28 May 2026
  76. Deep Optimal Individualized Treatment Rules for Bivariate Survival Outcomes via Adaptive Prediction-Powered Learning2605.29464 · 28 May 2026
  77. Correcting Split Selection in Online Decision Trees via Anytime-Valid Inference2605.31239 · 29 May 2026
  78. Parameter-Free and Group Conditional Online Conformal Prediction2606.00419 · 29 May 2026
  79. Free energy Estimation on Any State Space2605.31063 · 29 May 2026
  80. Hedging on the Frontier: Learning New Tasks with Few Samples2605.30997 · 29 May 2026
  81. Counterfactual Explanations for Deep Two-Sample Testing2606.04009 · 29 May 2026
  82. Approximation and learning of anisotropic and mixed smooth functions by deep ReLU neural networks2605.31152 · 29 May 2026
  83. Out-of-Distribution generalization of quantile regression with heavy tailed inputs: an SVM approach2606.00265 · 29 May 2026
  84. Routing on the Stiefel Manifold: When Does Adaptive Subspace Selection Help for Cross-Domain EEG Decoding?2605.31043 · 29 May 2026
  85. Entropic Projection Alignment: Estimating, Explaining, and Improving Model Performance Under Distribution Shift2605.31250 · 29 May 2026
  86. Log-Ratio Propagation on the Simplex: A Theory of Cellwise Contamination for Compositional Data2605.31345 · 29 May 2026
  87. Memory by Design: Probabilistic Sequence Layers2605.31163 · 29 May 2026
  88. Batched Stochastic Linear Bandits with 1-Bit Communication Constraints2605.30976 · 29 May 2026
  89. ERICA: Quantifying Replicability of Cluster Analysis2606.00302 · 29 May 2026
  90. Interpreting FCDNNs via RG on Exponential Family2606.00157 · 29 May 2026
  91. Is the Last Layer Sufficient for Uncertainty Quantification?2605.30741 · 29 May 2026
  92. Riemannian Stochastic Optimization for Sufficient Dimension Reduction2606.00413 · 29 May 2026
  93. Is Zero-Shot Super-Resolution Possible in Operator Learning?2606.00296 · 29 May 2026
  94. Spectra-Guided Neural Tucker Factorization2606.00584 · 30 May 2026
  95. On Finite-sample Concentration of Median of Incomplete U-Statistics2606.00661 · 30 May 2026
  96. Position: Prioritize Identifying Structure, Not Complex Models, for Scientific Discovery2606.02632 · 30 May 2026
  97. Statistical Analysis of using the Shapley Value for Sensor Anomaly Localization with Accurate Classifiers2606.00867 · 30 May 2026
  98. Statistical Testing on Directed Graphs by Surrogate Data Generation2606.00758 · 30 May 2026
  99. Taming the Loss Landscape of PINNs with Noisy Feynman-Kac Supervision: Operator Preconditioning and Non-Asymptotic Error Bounds2606.00643 · 30 May 2026
  100. Bandit Simulation for Average Reward Inference2606.00913 · 30 May 2026