Goto

Collaborating Authors

 Technology




Beyond the Return: Off-policy Function Estimation under User-specified Error-measuring Distributions

Neural Information Processing Systems

Off-policy evaluation often refers to two related tasks: estimating the expected return of a policy and estimating its value function (or other functions of interest, such as density ratios). While recent works on marginalized importance sampling (MIS) show that the former can enjoy provable guarantees under realizable function approximation, the latter is only known to be feasible under much stronger assumptions such as prohibitively expressive discriminators. In this work, we provide guarantees for off-policy function estimation under only realizability, by imposing proper regularization on the MIS objectives. Compared to commonly used regularization in MIS, our regularizer is much more flexible and can account for an arbitrary user-specified distribution, under which the learned function will be close to the groundtruth. We provide exact characterization of the optimal dual solution that needs to be realized by the discriminator class, which determines the datacoverage assumption in the case of value-function learning. As another surprising observation, the regularizer can be altered to relax the data-coverage requirement, and completely eliminate it in the ideal case with strong side information.


Stability and Deviation Optimal Risk Bounds with Convergence Rate O(1/n)

Neural Information Processing Systems

The sharpest known high probability generalization bounds for uniformly stable algorithms (Feldman, Vondrák, NeurIPS 2018, COLT, 2019), (Bousquet, Klochkov, Zhivotovskiy, COLT, 2020) contain a generally inevitable sampling error term of order Θ(1/ n). When applied to excess risk bounds, this leads to suboptimal results in several standard stochastic convex optimization problems. We show that if the so-called Bernstein condition is satisfied, the term Θ(1/ n) can be avoided, and high probability excess risk bounds of order up to O(1/n) are possible via uniform stability. Using this result, we show a high probability excess risk bound with the rate O(log n/n) for strongly convex and Lipschitz losses valid for any empirical risk minimization method.




The Skellam Mechanism for Differentially Private Federated Learning

Neural Information Processing Systems

We introduce the multi-dimensional Skellam mechanism, a discrete differential privacy mechanism based on the difference of two independent Poisson random variables. To quantify its privacy guarantees, we analyze the privacy loss distribution via a numerical evaluation and provide a sharp bound on the Rényi divergence between two shifted Skellam distributions. While useful in both centralized and distributed privacy applications, we investigate how it can be applied in the context of federated learning with secure aggregation under communication constraints. Our theoretical findings and extensive experimental evaluations demonstrate that the Skellam mechanism provides the same privacy-accuracy trade-offs as the continuous Gaussian mechanism, even when the precision is low. More importantly, Skellam is closed under summation and sampling from it only requires sampling from a Poisson distribution - an efficient routine that ships with all machine learning and data analysis software packages. These features, along with its discrete nature and competitive privacy-accuracy trade-offs, make it an attractive practical alternative to the newly introduced discrete Gaussian mechanism.


Inside Chornobyl: 40 years after disaster, nuclear site still at risk in Russia's war

The Guardian > Energy

A worker checks the radiation level inside the control room of reactor No 4, where the Chornobyl disaster happened in 1986. A worker checks the radiation level inside the control room of reactor No 4, where the Chornobyl disaster happened in 1986. In February 2025, a cheap Russian drone tore through Chornobyl's confinement shelter. Workers warn the site of the world's worst nuclear accident is not safe yet The dosimeter clipped to your chest ticks faster the moment you step off the designated path inside the Chornobyl nuclear power plant. Step back, and it slows again - an invisible line between clean ground and contamination.


AFast Scale-Invariant Algorithm for Non-negative Least Squares with Non-negative Data

Neural Information Processing Systems

Nonnegative (linear) least square problems are a fundamental class of problems that is well-studied in statistical learning and for which solvers have been implemented in many of the standard programming languages used within the machine learning community. The existing off-the-shelf solvers view the non-negativity constraint in these problems as an obstacle and, compared to unconstrained least squares, perform additional effort to address it. However, in many of the typical applications, the data itself is nonnegative as well, and we show that the nonnegativity in this case makes the problem easier. In particular, while the worst-case dimension-independent oracle complexity for unconstrained least squares problems necessarily scales with one of the data matrix constants (typically the spectral norm) and these problems are solved to additive error, we show that nonnegative least squares problems with nonnegative data are solvable to multiplicative error and with complexity independent of any matrix constants. The algorithm we introduce is accelerated and based on a primal-dual perspective. We further show how to provably obtain linear convergence using adaptive restart coupled with our method and demonstrate its effectiveness on large-scale data via numerical experiments.


'Animals are traumatised too': Pet rescuers under fire in Ukraine

BBC News

'Animals are traumatised too': Pet rescuers under fire in Ukraine On a morning in February, animal shelter staff were getting changed for their shift when a Russian drone slammed into the centre of their compound in the frontline Ukrainian city of Zaporizhzhia. The steel door at the entrance probably saved their lives. More than a dozen animals sheltering at Give a Paw, Friend were not so lucky. It was terrifying, to put it mildly, says the group's head Iryna Didur. Residents rushed to help clean up the rubble and catch the animals that had escaped in terror.