VLDB 2026 Research / reviewers in the wild / expert
Nikita Puchkin
dblp:218/6698
· DBLP profile ↗
11ranked-venue papers
8as first author
11since 2021 · last 2026
0000-0002-9677-4275ORCID · corroborated
Domains — the database's venue-derived domains; a paper can count in several
Artificial intelligence and machine learning · 9 · 7 first-author · 9 since 2021Theory of computation · 2 · 1 first-author · 2 since 2021
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2026 | Simultaneous approximation of the score function and its derivatives by deep neural networks
Konstantin Yakovlev, Nikita Puchkin |
J. Complex. | 2 |
| 2025 | Generalization error bound for denoising score matching under relaxed manifold assumptionabstractWe examine theoretical properties of the denoising score matching estimate. We model the density of observations with a nonparametric Gaussian mixture. We significantly relax the standard manifold assumption allowing the samples step away from the manifold. At the same time, we are still able to leverage a nice distribution structure. We derive non-asymptotic bounds on the approximation and generalization errors of the denoising score matching estimate. The rates of convergence are determined by the intrinsic dimension. Furthermore, our bounds remain valid even if we allow the ambient dimension grow polynomially with the sample size. Konstantin Yakovlev, Nikita Puchkin |
COLT | 2 |
| 2024 | Breaking the Heavy-Tailed Noise Barrier in Stochastic Optimization ProblemsabstractWe consider stochastic optimization problems with heavy-tailed noise with structured density. For such problems, we show that it is possible to get faster rates of convergence than $O(K^{-2(\alpha - 1) / \alpha})$, when the stochastic gradients have finite $\alpha$-th moment, $\alpha \in (1, 2]$. In particular, our analysis allows the noise norm to have an unbounded expectation. To achieve these results, we stabilize stochastic gradients, using smoothed medians of means. We prove that the resulting estimates have negligible bias and controllable variance. This allows us to carefully incorporate them into clipped-SGD and clipped-SSTM and derive new high-probability complexity bounds in the considered setup. Nikita Puchkin, Eduard Gorbunov, Nikolay Kutuzov, Alexander V. Gasnikov |
AISTATS | 1 |
| 2024 | Dimension-free Structured Covariance EstimationabstractGiven a sample of i.i.d. high-dimensional centered random vectors, we consider a problem of estimation of their covariance matrix $\Sigma$ with an additional assumption that $\Sigma$ can be represented as a sum of a few Kronecker products of smaller matrices. Under mild conditions, we derive the first non-asymptotic dimension-free high-probability bound on the Frobenius distance between $\Sigma$ and a widely used penalized permuted least squares estimate. Because of the hidden structure, the established rate of convergence is faster than in the standard covariance estimation problem. Nikita Puchkin, Maksim Rakhuba |
COLT | 1 |
| 2024 | Rates of convergence for density estimation with generative adversarial networksabstractIn this work we undertake a thorough study of the non-asymptotic properties of the vanilla generative adversarial networks (GANs). We prove an oracle inequality for the Jensen-Shannon (JS) divergence between the underlying density $\mathsf{p}^*$ and the GAN estimate with a significantly better statistical error term compared to the previously known results. The advantage of our bound becomes clear in application to nonparametric density estimation. We show that the JS-divergence between the GAN estimate and $\mathsf{p}^*$ decays as fast as $(\log{n}/n)^{2\beta/(2\beta + d)}$, where $n$ is the sample size and $\beta$ determines the smoothness of $\mathsf{p}^*$. This rate of convergence coincides (up to logarithmic factors) with minimax optimal for the considered class of densities. Nikita Puchkin, Sergey Samsonov, Denis Belomestny, Eric Moulines, Alexey Naumov |
J. Mach. Learn. Res. | 1 |
| 2023 | A Contrastive Approach to Online Change Point DetectionabstractWe suggest a novel procedure for online change point detection. Our approach expands an idea of maximizing a discrepancy measure between points from pre-change and post-change distributions. This leads to a flexible procedure suitable for both parametric and nonparametric scenarios. We prove non-asymptotic bounds on the average running length of the procedure and its expected detection delay. The efficiency of the algorithm is illustrated with numerical experiments on synthetic and real-world data sets. Nikita Puchkin, Valeriia Shcherbakova |
AISTATS | 1 |
| 2023 | Exploring Local Norms in Exp-concave Statistical LearningabstractWe consider the standard problem of stochastic convex optimization with exp-concave losses using Empirical Risk Minimization in a convex class. Answering a question raised in several prior works, we provide a $O ( d/n + 1/n \log( 1 / \delta ) )$ excess risk bound valid for a wide class of bounded exp-concave losses, where $d$ is the dimension of the convex reference set, $n$ is the sample size, and $\delta$ is the confidence level. Our result is based on a unified geometric assumption on the gradient of losses and the notion of local norms. Nikita Puchkin, Nikita Zhivotovskiy |
COLT | 1 |
| 2023 | Simultaneous approximation of a smooth function and its derivatives by deep neural networks with piecewise-polynomial activations
Denis Belomestny, Alexey Naumov, Nikita Puchkin, Sergey Samsonov |
Neural Networks | 3 |
| 2022 | Structure-adaptive Manifold EstimationabstractWe consider a problem of manifold estimation from noisy observations. Many manifold learning procedures locally approximate a manifold by a weighted average over a small neighborhood. However, in the presence of large noise, the assigned weights become so corrupted that the averaged estimate shows very poor performance. We suggest a structure-adaptive procedure, which simultaneously reconstructs a smooth manifold and estimates projections of the point cloud onto this manifold. The proposed approach iteratively refines the weights on each step, using the structural information obtained at previous steps. After several iterations, we obtain nearly “oracle” weights, so that the final estimates are nearly efficient even in the presence of relatively large noise. In our theoretical study, we establish tight lower and upper bounds proving asymptotic optimality of the method for manifold estimation under the Hausdorff loss, provided that the noise degrades to zero fast enough. Nikita Puchkin, Vladimir G. Spokoiny |
J. Mach. Learn. Res. | 1 |
| 2022 | Exponential Savings in Agnostic Active Learning Through AbstentionabstractWe show that in pool-based active classification without assumptions on the underlying distribution, if the learner is given the power to abstain from some predictions by paying the price marginally smaller than the average loss 1/2 of a random guess, exponential savings in the number of label requests are possible whenever they are possible in the corresponding realizable problem. We extend this result to provide a necessary and sufficient condition for exponential savings in pool-based active classification under the model misspecification. Nikita Puchkin, Nikita Zhivotovskiy |
IEEE Trans. Inf. Theory | 1 |
| 2021 | Exponential savings in agnostic active learning through abstentionabstractWe show that in pool-based active classification without assumptions on the underlying distribution, if the learner is given the power to abstain from some predictions by paying the price marginally smaller than the average loss 1/2 of a random guess, exponential savings in the number of label requests are possible whenever they are possible in the corresponding realizable problem. We extend this result to provide a necessary and sufficient condition for exponential savings in pool-based active classification under the model misspecification. Nikita Puchkin, Nikita Zhivotovskiy |
COLT | 1 |