|
Empirical Bernstein Confidence Intervals for Kernel Smoothers: A Safe and Sharp Way to Exhaust Assumed Smoothness
-
arxiv.org
-
1 month ago
-
eng
arXiv:2605.03781v4 Announce Type: replace Abstract: Using standard-normal critical-value calibration (SNC) to construct a kernel-smoother-based confidence interval faces a fundamental challenge: the normalization makes a small estimation bias become a non-negligible inferential bias. This paper takes a different route by replacing the SNC control with empirical Bernstein tail control. The resulting confidence intervals control stochastic v....
|
|
arXiv:2605.07818v2 Announce Type: replace Abstract: The expectation--maximization (EM) algorithm combines global monotonicity, local linear convergence, and strong practical robustness, but these features are usually analyzed separately. Global descent is nonlinear, whereas local convergence is governed by the spectrum of the linearized EM map. How these two levels fit into a single dynamical picture has remained less transparent. We make....
|
|
ISOMORPH: A Supply Chain Digital Twin for Simulation, Dataset Generation, and Forecasting Benchmarks
-
arxiv.org
-
1 month ago
-
eng
arXiv:2605.12768v2 Announce Type: replace Abstract: Open time-series forecasting (TSF) benchmarks cover retail, energy, weather, and traffic, but supply-chain logistics remains underserved. We introduce ISOMORPH, the first public digital twin of a multi-echelon logistics network with interpretable, user-configurable parameters and modular topology, demand, and control rules. The simulator advances a directed routing graph in discrete time:....
|
|
arXiv:2605.13203v2 Announce Type: replace Abstract: This paper investigates the predictive performance of model averaging in high-dimensional linear regression where the number of regressors is comparable to the sample size. Leveraging tools from random matrix theory, we derive the exact limiting out-of-sample risk under a nested model setting and comprehensively characterize the risk landscape. This limiting risk helps to reveal two pheno....
|
|
Stabilised weighted data subsampling for accelerated inference in models with recursive likelihoods
-
arxiv.org
-
1 month ago
-
eng
arXiv:2605.13397v2 Announce Type: replace Abstract: Inference for models with recursively defined likelihoods is computationally demanding, limiting scalability to large datasets. We propose a stabilised weighted subsampling methodology for accelerated inference based on an unbiased estimator of the log-likelihood. By assigning higher sampling probabilities to early observations, the method reduces the effective depth of recursive likeliho....
|
|
Towards a holistic understanding of Selection Bias for Causal Effect Identification
-
arxiv.org
-
1 month ago
-
eng
arXiv:2605.13430v3 Announce Type: replace Abstract: Selection bias is pervasive in observational studies. For example, large scale biobanks data can exhibit ``healthy volunteer bias'' when respondents are healthier and of higher socio-economic status than the population they are meant to represent. Recovering causal effects from such sub-population is an important problem in causal inference, as estimating average treatment effects (ATE) f....
|
|
Evaluating causal indirect effects when mediators are left-censored by assay limit of quantification
-
arxiv.org
-
1 month ago
-
eng
arXiv:2605.20615v2 Announce Type: replace Abstract: Causal mediation analysis is essential for disentangling the mechanisms by which investigational therapeutic and preventive agents impact clinical outcomes. However, the measurement of biological mediators is often subject to left-censoring by technical measurement limitations, most commonly an assay's limit of quantification. This form of censoring can pose severe challenges for both ide....
|
|
Trustworthy AI/ML Regression and Unbiased Causal Inference for Real-World Data
-
arxiv.org
-
1 month ago
-
eng
arXiv:2605.24377v2 Announce Type: replace Abstract: Real-World Data (RWD), with its large sample sizes and rich clinical detail, offers a compelling alternative to randomized controlled trials (RCTs) for studying treatment effects in diverse and complex patient populations. However, its observational nature introduces confounding that prevents straightforward comparative effectiveness research. Target trial emulation leverages RWD to estim....
|
|
Logistic regression is not enough: The need for Bayesian nonparametric modelling for causal inference using observational data, exemplified by the 'gateway' effect
-
arxiv.org
-
1 month ago
-
eng
arXiv:2605.24847v2 Announce Type: replace Abstract: Introduction: Logistic regression (LR)-type model limitations for causal inference are explained theoretically and empirically through the lens of the purported gateway effect from e-cigarette use to smoking. Previous studies have reported that baseline e-cigarette use quadruples odds of follow-up smoking (binarized) in LR-type models of adolescent longitudinal cohorts (LCs), such that in....
|
|
Approximating full conformal prediction: distribution free guarantees via the tournament correction
-
arxiv.org
-
1 month ago
-
eng
arXiv:2605.29200v2 Announce Type: replace Abstract: Conformal prediction is a framework for providing prediction intervals with distribution-free validity, guaranteeing predictive coverage for data drawn from any distribution. Its two main variants are full conformal prediction and split conformal prediction (also called transductive and inductive). Full conformal prediction is widely considered to be statistically more efficient (since sp....
|
|
Gaussian Differentially Private $e$-values: Construction, Threshold Calibration, and Multiple Testing
-
arxiv.org
-
1 month ago
-
eng
arXiv:2605.29388v2 Announce Type: replace Abstract: This paper develops a framework for differentially private $e$-values under Gaussian differential privacy ($\mu$-GDP). We characterize the canonical noise mechanism, establishing that optimal multiplicative perturbation follows a Gaussian distribution. Using this distribution, we derive a globally sharp rejection threshold that strictly improves upon the standard Markov bound. Asymptotic ....
|
|
arXiv:2605.30242v2 Announce Type: replace Abstract: The airborne fraction is the share of anthropogenic carbon dioxide emissions that remains in the atmosphere and is a key indicator of carbon-cycle response and remaining carbon budgets under continued emissions. Whether this share is rising remains debated because inference is sensitive to uncertainty in land-use and land-cover change (LULC) emissions. Here we use all available LULC measu....
|
|
Optimizing accuracy and diversity: a multi-task approach to forecast combinations
-
arxiv.org
-
1 month ago
-
eng
arXiv:2310.20545v3 Announce Type: replace-cross Abstract: We present a multi-task optimization approach based on a deep learning architecture for time series forecasting. We leverage large collections of time series to identify the weights of forecasting models that can be combined to produce forecasts for each series. This method jointly addresses two tasks: the selection of different forecasting models, and their effective combination. I....
|
|
arXiv:2403.07008v3 Announce Type: replace-cross Abstract: The evaluation of machine learning models using human-labeled validation data can be expensive and time-consuming. AI-labeled synthetic data can be used to decrease the number of human annotations required for this purpose in a process called autoevaluation. We suggest efficient and statistically principled algorithms for this purpose that improve sample efficiency while remaining u..
|
|
A theory of generalised coordinates for stochastic differential equations
-
arxiv.org
-
1 month ago
-
eng
arXiv:2409.15532v3 Announce Type: replace-cross Abstract: Stochastic differential equations are ubiquitous modelling tools in physics and the sciences. In most modelling scenarios, random fluctuations driving dynamics or motion have some non-trivial temporal correlation structure, which renders the SDE non-Markovian; a phenomenon commonly known as ``colored'' noise. Thus, an important objective is to develop effective tools for mathematica....
|
|
arXiv:2410.17105v4 Announce Type: replace-cross Abstract: We develop a flexible framework for Bayesian estimation of impulse responses using Local Projections (LPs) with instrumental variables. It accommodates multiple shocks and instruments, accounts for autocorrelation in multi-step forecasts by jointly modeling all LPs as a seemingly unrelated system of equations, defines a flexible yet parsimonious joint prior for impulse responses bas..
|
|
A Likelihood Approach for Inference of Population Heterogeneity in Particle Ensembles with Second-Order Langevin Dynamics
-
arxiv.org
-
1 month ago
-
eng
arXiv:2411.08692v2 Announce Type: replace-cross Abstract: The inherent complexity of biological agents often leads to motility behavior that appears to have random components. Robust stochastic inference methods are therefore required to understand and predict the motion patterns from time-discrete trajectory data provided by experiments. In many cases, second-order Langevin models are needed to adequately capture the motility. Additionall....
|
|
Dimension Reduction via Sum-of-Squares and Improved Clustering Algorithms for Non-Spherical Mixtures
-
arxiv.org
-
1 month ago
-
eng
arXiv:2411.12438v2 Announce Type: replace-cross Abstract: We develop a new approach for clustering non-spherical (i.e., arbitrary component covariances) Gaussian mixture models via a subroutine, based on the sum-of-squares method, that finds a low-dimensional separation-preserving projection of the input data. Our method gives a non-spherical analog of the classical dimension reduction, based on singular value decomposition, that, among se....
|
|
arXiv:2412.04177v2 Announce Type: replace-cross Abstract: Recently, there has been an increasing interest in performing post-hoc uncertainty estimation about the predictions of pre-trained deep neural networks (DNNs). Given a pre-trained DNN via back-propagation, these methods enhance the original network by adding output confidence measures, such as error bars, without compromising its initial accuracy. In this context, we introduce a nov....
|
|
Challenges in the calibration of tree-based models for imbalanced classification
-
arxiv.org
-
1 month ago
-
eng
arXiv:2412.16209v5 Announce Type: replace-cross Abstract: When using machine learning for imbalanced binary classification problems, it is common to subsample the majority class to create a (more) balanced training dataset. This biases the model's predictions because the model learns from data that is not fully representative of the underlying population of interest. One way of accounting for this bias is analytically mapping the resulting....
|
|
Towards Simple and Provable Parameter-Free Adaptive Gradient Methods
-
arxiv.org
-
1 month ago
-
eng
arXiv:2412.19444v2 Announce Type: replace-cross Abstract: Optimization algorithms such as AdaGrad and Adam have significantly advanced the training of deep models by dynamically adjusting the learning rate during the optimization process. However, ad-hoc tuning of learning rates poses a challenge and leads to inefficiencies in practice. To address this issue, recent research has focused on developing ``parameter-free'' algorithms that oper....
|
|
arXiv:2501.08640v2 Announce Type: replace-cross Abstract: We propose a way to bound the generalisation errors of several classes of quantum reservoirs using the Rademacher complexity. We give specific, parameter-dependent bounds for two particular quantum reservoir classes. We analyse how the generalisation bounds scale with growing numbers of qubits. Applying our results to classes with polynomial readout functions, we find that the risk ..
|
|
Non-vacuous Generalization Bounds for Deep Neural Networks without any modification to the trained models
-
arxiv.org
-
1 month ago
-
eng
arXiv:2503.07325v2 Announce Type: replace-cross Abstract: Understanding and certifying the behavior of modern deep neural networks remains a fundamental challenge in reliable machine learning. We introduce a new class of data-dependent generalization bounds that apply directly to trained models, without any modification. In particular, we present an exactly computable bound that is non-vacuous across all evaluated networks, including Image....
|
|
Advancing Local Clustering on Graphs via Compressive Sensing: Semi-supervised and Unsupervised Methods
-
arxiv.org
-
1 month ago
-
eng
arXiv:2504.19419v3 Announce Type: replace-cross Abstract: Local clustering aims to identify specific substructures within a large graph without any additional structural information of the graph. These substructures are typically small compared to the overall graph, enabling the problem to be approached by finding a sparse solution to a linear system associated with the graph Laplacian. In this work, we first propose a method for identifyi....
|
|
Global Convergence of Adaptive Sensing for Principal Eigenvector Estimation
-
arxiv.org
-
1 month ago
-
eng
arXiv:2505.10882v2 Announce Type: replace-cross Abstract: Principal component analysis classically requires full $d$-dimensional samples, yet in various applications hardware limits acquisition to a few scalar measurements per sample. We analyze a compressed variant of Oja's algorithm for estimating the principal eigenvector of the data covariance matrix using only two adaptive measurements per sample. At each iteration, we observe one mea....
|
|
HR-VILAGE-3K3M: A Human Respiratory Viral Immunization Longitudinal Gene Expression Dataset for Systems Immunity
-
arxiv.org
-
1 month ago
-
eng
arXiv:2505.14725v2 Announce Type: replace-cross Abstract: Respiratory viral infections pose a global health burden, yet the cellular immune mechanisms underlying protection and pathology remain unclear. Natural infection cohorts often lack pre-exposure baselines and time-controlled sampling, whereas inoculation and vaccination trials generate well-structured longitudinal transcriptomic data. However, these datasets are scattered across rep....
|
|
Human in the Loop Adaptive Optimization for Improved Time Series Forecasting
-
arxiv.org
-
1 month ago
-
eng
arXiv:2505.15354v2 Announce Type: replace-cross Abstract: Time series forecasting models often produce systematic, predictable errors even in critical domains such as energy, finance, and healthcare. We introduce a novel post training adaptive optimization framework that improves forecast accuracy without retraining or architectural changes. Our method automatically applies expressive transformations optimized via reinforcement learning, c....
|
|
How Many Domains Suffice for Domain Generalization? A Tight Characterization via the Domain Shattering Dimension
-
arxiv.org
-
1 month ago
-
eng
arXiv:2506.16704v3 Announce Type: replace-cross Abstract: We study a fundamental question of domain generalization: given a family of domains (i.e., data distributions), how many randomly sampled domains do we need to collect data from in order to learn a model that performs reasonably well on every seen and unseen domain in the family? We model this problem in the PAC framework and introduce a new combinatorial measure, which we call the ..
|
|
VERA: Variational Inference Framework for Jailbreaking Large Language Models
-
arxiv.org
-
1 month ago
-
eng
arXiv:2506.22666v3 Announce Type: replace-cross Abstract: The rise of API-only access to state-of-the-art LLMs highlights the need for effective black-box jailbreak methods to identify model vulnerabilities in real-world settings. Without a principled objective for gradient-based optimization, most existing approaches rely on genetic algorithms, which are limited by their initialization and dependence on manually curated prompt pools. Furt....
|
|
Fundamental bounds on efficiency-confidence trade-off for transductive conformal prediction
-
arxiv.org
-
1 month ago
-
eng
arXiv:2509.04631v2 Announce Type: replace-cross Abstract: Transductive conformal prediction addresses the simultaneous prediction for multiple data points. Given a desired confidence level, the objective is to construct a prediction set that includes the true outcomes with the prescribed confidence. We demonstrate a fundamental trade-off between confidence and efficiency in transductive methods, where efficiency is measured by the size of ....
|
|
arXiv:2509.13805v4 Announce Type: replace-cross Abstract: Foundation models have revolutionized natural language processing through a ``train once, deploy anywhere'' paradigm, where a single pre-trained model adapts to countless downstream tasks without retraining. Access to a Physics Foundation Model (PFM) would be transformative - democratizing access to high-fidelity simulations, accelerating scientific discovery, and eliminating the ne....
|
|
arXiv:2509.18025v2 Announce Type: replace-cross Abstract: One can see deep-learning models as compositions of functions within the so-called tame geometry. In this expository note, we give an overview of some topics at the interface of tame geometry (also known as o-minimality), optimization theory, and deep learning theory and practice. To do so, we gradually introduce the concepts and tools used to build convergence guarantees for stocha..
|
|
Interpretable Self-Supervised Learning via Representer Landmarks and Nystr\"om Approximation
-
arxiv.org
-
1 month ago
-
eng
arXiv:2509.24467v3 Announce Type: replace-cross Abstract: Self-supervised learning (SSL) learns representations from massive unlabeled data, yet the resulting models typically operate as black boxes, necessitating domain-specific explanations. We introduce KREPES, a unified framework to analytically interpret the learned representations of SSL objectives, including SimCLR, BYOL, and VICReg. By bridging empirical neural tangent kernel appro....
|
|
Trajectory Data Suffices for Statistically Efficient Policy Evaluation in Fixed-Horizon Offline RL with Linear $q^\pi$-Realizability and Concentrability
-
arxiv.org
-
1 month ago
-
eng
arXiv:2510.03494v2 Announce Type: replace-cross Abstract: We study finite-horizon offline reinforcement learning (RL) with function approximation for both policy evaluation and policy optimization. Prior work established that statistically efficient learning is impossible for either of these problems when the only assumptions are that the data has good coverage (concentrability) and the state-action value function of every policy is linear..
|
|
From Moments to Models: Graphon-Mixture Learning for Mixup and Contrastive Learning
-
arxiv.org
-
1 month ago
-
eng
arXiv:2510.03690v4 Announce Type: replace-cross Abstract: Real-world graph datasets often arise from mixtures of populations, where graphs are generated by multiple distinct underlying distributions. In this work, we propose a unified framework that explicitly models graph data as a mixture of probabilistic graph generative models represented by graphons. To characterize and estimate these graphons, we leverage graph moments (motif densiti....
|
|
Generalization of Gibbs and Langevin Monte Carlo Algorithms in the Interpolation Regime
-
arxiv.org
-
1 month ago
-
eng
arXiv:2510.06028v3 Announce Type: replace-cross Abstract: This paper provides data-dependent bounds on the expected error of the Gibbs algorithm in the overparameterized interpolation regime, where low training errors are also obtained for impossible data, such as random labels in classification. The results show that generalization in the low-temperature regime is already signaled by small training errors in the noisier high-temperature r..
|
|
arXiv:2510.17303v2 Announce Type: replace-cross Abstract: Symmetries are known to improve the empirical performance of machine learning models, yet theoretical guarantees explaining these gains remain limited. Prior work has focused mainly on compact group symmetries and often assumes that the data distribution itself is invariant, an assumption rarely satisfied in real-world applications. In this work, we extend generalization guarantees ....
|
|
Two Datasets Are Better Than One: Method of Double Moments for 3-D Reconstruction in Cryo-EM
-
arxiv.org
-
1 month ago
-
eng
arXiv:2511.07438v3 Announce Type: replace-cross Abstract: Cryo-electron microscopy (cryo-EM) is a powerful imaging technique for reconstructing three-dimensional molecular structures from noisy tomographic projection images of randomly oriented particles. We introduce a new data fusion framework, termed the method of double moments (MoDM), which reconstructs molecular structures from two instances of the second-order moment of projection i....
|
|
arXiv:2511.21140v4 Announce Type: replace-cross Abstract: Large language models (LLMs) are widely used as scalable evaluators of model responses in lieu of human annotators. However, imperfect sensitivity and specificity of the LLM judges induce bias in naive evaluation scores. We propose a simple plug-in framework that corrects this bias and enables statistically principled uncertainty quantification. Our framework constructs confidence i....
|
|
Safeguarded Stochastic Polyak Step Sizes for Non-smooth Optimization: Robust Performance Without Small (Sub)Gradients
-
arxiv.org
-
1 month ago
-
eng
arXiv:2512.02342v3 Announce Type: replace-cross Abstract: The stochastic Polyak step size (SPS) has proven to be a promising choice for stochastic gradient descent (SGD), delivering competitive performance relative to state-of-the-art methods on smooth convex and non-convex optimization problems, including deep neural network training. However, extensions of this approach to non-smooth settings remain in their early stages, often relying o....
|
|
State and Parameter Estimation for a Neural Model of Local Field Potentials
-
arxiv.org
-
1 month ago
-
eng
arXiv:2512.07842v2 Announce Type: replace-cross Abstract: The study of cortical dynamics during different states such as decision making, sleep and movement, is an important topic in Neuroscience. Modelling efforts aim to relate the neural rhythms present in cortical recordings to the underlying dynamics responsible for their emergence. We present an effort to characterize the neural activity from the cortex of a mouse during natural sleep....
|
|
arXiv:2601.16884v3 Announce Type: replace-cross Abstract: We study multigrade deep learning (MGDL) as a principled framework for structured error refinement in deep neural networks. While the approximation power of neural networks is now relatively well understood, training very deep architectures remains challenging due to highly nonconvex and often ill-conditioned optimization landscapes. In contrast, for relatively shallow networks, mos....
|
|
arXiv:2602.02819v4 Announce Type: replace-cross Abstract: Membership Inference Attacks (MIAs) aim to distinguish training points (members) from unseen data (non-members), and are widely used to quantify memorization and assess privacy risks. Standard MIA evaluation requires repeated retraining, which is computationally costly for large models. One-run (single training with randomized data inclusion) and zero-run (post hoc evaluation) metho....
|
|
arXiv:2602.03685v2 Announce Type: replace-cross Abstract: Training large language models (LLMs) is computationally expensive, partly because the loss exhibits slow power-law convergence whose origin remains debatable. Through systematic analysis of toy models and empirical evaluation of LLMs, we show that this behavior can arise intrinsically from the use of softmax and cross-entropy. When learning peaked probability distributions, e.g., n..
|