Forschungsseminar Mathematische Statistik
Für den Bereich Statistik
A. Carpentier, S. Greven, W. Härdle, M. Reiß, V. Spokoiny
Ort
Weierstrass-Institut für Angewandte Analysis und Stochastik
ESH,
Mohrenstrasse 39
10117 Berlin
Zoom: https://hu-berlin.zoom-x.de/j/65177102181?pwd=nUMtJOkzDx8xyWzPFezeHh0NEIQUv3.1
Zeit
mittwochs, 10.00 - 12.00 Uhr
Programm
-
-
- 22. April 2026
- Eddi Aamari & Arthur Stephanovich (École normale supérieure u. ENSAE-CREST, Paris)
-
Regularity of the score and convergence rates of generative diffusion models
- Abstract:
We show that diffusion-based generative models adapt to the smoothness of the target distribution: the score function inherits the target’s regularity. Leveraging this adaptivity, we obtain a concise proof that diffusion models achieve minimax-optimal rates for density estimation
- 29. April 2026
- Nicola Gnecco (Imperial College London)
- Extremes of Structural Causal Models
-
Abstract: The behaviour of extreme observations is well-understood for time series or spatial data, but little is known if the data generating process is a structural causal model (SCM). We study the behavior of extremes in this model class, both for the observational distribution and under extremal interventions. We show that under suitable regularity conditions on the structure functions, the extremal behavior is described by a multivariate Pareto distribution, which can be represented as a new SCM on an extremal graph. Importantly, the latter is a sub-graph of the graph in the original SCM, which means that causal links can disappear in the tails. We further introduce a directed version of extremal graphical models and show that an extremal SCM satisfies the corresponding Markov properties. Based on a new test of extremal conditional independence, we propose two algorithms for learning the extremal causal structure from data. The first is an extremal version of the PC-algorithm, and the second is a pruning algorithm that removes edges from the original graph to consistently recover the extremal graph. The methods are illustrated on river data with known causal ground truth.
05. Mai 2026
Vladimir Spokoiny (WIAS)
Estimation of a smooth functional for inverse problems
Abstract: Since the seminal paper by Ibragimov and Khasmisnkii (1980) it is well known that the plug-in approach is sub-optimal in estimating a nonlinear functional like the squared norm of the signal. A kind of bias correction is necessary to achieve root-n optimality. Series of recent papers by Koltchinskii with coauthors revisited this problem from modern prospective with a high or even dimensional parameter space and a complex structure of the functional to be estimated. This talk discusses the problem of estimation of a smooth functional in the inverse problem setup given by an objective function L(θ).
A different view is offered. Namely, the functional φ(θ) is included in the parameter list as the target value, while the parameter θ is treated as a nuisance parameter. The structural relation x = φ(θ) is replaced by the structural penalty (λ|x − φ(θ))2. The resulting estimator is obtained as a third-order correction (xˆ, θˆ) of the full-dimensional penalized MLE (x ̃, θ ̃) = arg inf(x,θ) L(θ) + λ|x − φ(θ)|2/2. The approach enables us to obtain an accurate expansion with an explicit leading term and self-consistent remainder and provides sharp risk bounds for this estimator. In particular, we present sufficient conditions for root-n consistency of the procedure.
- 13. Mai 2026
- Holger Dette (Ruhr University Bochum)
- Multiple change point detection in functional data with applications to biomechanical fatigue data
-
Abstract:
Injuries to the lower extremity joints are often debilitating, particularly for professional athletes. Understanding the onset of stressful conditions on these joints is therefore important in order to ensure prevention of injuries as well as individualised training for enhanced athletic performance. We study the biomechanical joint angles from the hip, knee and ankle for runners who are experiencing fatigue. The data is cyclic in nature and densely collected by body worn sensors, which makes it ideal to work with in the functional data analysis (FDA) framework.
We develop a new method for multiple change point detection for functional data, which improves the state of the art with respect to at least two novel aspects. First, the curves are compared with respect to their maximum absolute deviation, which leads to a better interpretation of local changes in the functional data compared to classical $L^2$-approaches. Secondly, as slight aberrations are to be often expected in a human movement data, our method will not detect arbitrarily small changes but hunts for relevant changes, where maximum absolute deviation between the curves exceeds a specified threshold, say $\Delta >0$. We recover multiple changes in a long functional time series of biomechanical knee angle data, which are larger than the desired threshold $\Delta$, allowing us to identify changes purely due to fatigue. In this work, we analyse data from both controlled indoor as well as from an uncontrolled outdoor (marathon) setting. -
20. Mai 2026
-
Johannes Schmidt-Hieber (Twente)
A new neural network architecture for learning convex functions
Abstract: For small covariate dimension, shape constrained inference is a well-established topic within nonparametric statistics. For large covariate dimensions, one naturally wants to introduce machine learning based methods. Input convex neural networks (ICNNs) were designed as a network architecture to learn convex functions. In this talk, we introduce Hyper Input Convex Neural Networks (HyCNNs). HyCNNs combine the principles of Maxout networks with ICNNs to create a neural network that is always convex in the input, theoretically capable of leveraging depth, and performs reliable when trained at scale compared to ICNNs. Concretely, we prove that HyCNNs require exponentially fewer parameters than ICNNs to approximate quadratic functions up to a given precision. Throughout a series of synthetic experiments, we demonstrate that HyCNNs outperform existing ICNNs and MLPs in terms of predictive performance for convex regression and interpolation tasks. WWe further apply HyCNNs to learn high-dimensional optimal transport maps for synthetic examples and for single-cell RNA sequencing data, where they oftentimes outperform ICNN-based neural optimal transport methods and other baselines across a wide range of settings.
For more details, see arxiv.org/pdf/2604.26942.
This is joint work with Shayan Hundrieser and Insung Kong
- 27. Mai 2026
- Alexander Meister (Rostock)
Abstract: We prove asymptotic equivalence of nonparametric additive regression and an appropriate Gaussian white noise experiment in which a multidimensional shifted Wiener process is observed, whose dimension equals the number of additive components. The shift depends on the additive components of the regression function and solely the one- and two-dimensional marginal distributions of the covariates via an explicitly specified bounded but non compact linear operator. The number of additive components is allowed to increase moderately with respect to the sample size. In the special case of pairwise independent components of the covariates, the white noise model decomposes into independent univariate processes. Moreover, we study approximation in some semiparametric setting where the operator splits into a multiplication operator and an asymptotically negligible Hilbert-Schmidt operator.
03. Juni 2026
N.N.
10. Juni 2026
Michael Soerense (University of Copenhagen)
Abstract: The complexity of likelihood inference for stochastic differential equations based on discrete time samples often necessitates the use of approximations or computational techniques. Approximate likelihood methods for high frequency data have often been used in financial econometrics, but these methods usually do not perform well for strongly nonlinear models.
New developments of approximate likelihood methods based on splitting schemes are presented. These methods perform well also for strongly nonlinear models and at moderate sampling frequencies. Splitting schemes were originally introduced to solve ODEs and SDEs numerically, but in Pilipovic, Samson and Ditlevsen (2024) it was proposed to use them for statistical inference. In the talk a more general approach is presented that is applicable to a broad class of diffusion models. The theory is developed in the framework of approximate martingale estimating functions, which provide approximations to the score function and estimators that are efficient for high frequency data. For Strang splitting an approximate martingale estimating function of order 3 is obtained.
Sometimes useful models with an explicit likelihood function can be found. This enables exact likelihood inference, which works at all sampling frequencies. As an example of this, a class of stochastic differential equation models on the torus is presented, which can be used to analyse time series of angular data. These diffusion processes are ergodic and time-reversible and can be constructed for any pre-specified stationary distribution on the torus. If time permits, applications to biological data will be briefly presented.
The lecture is based on joint work with Susanne Ditlevsen, Adeline Samson and Eduardo García-Portugués.
Reference:
Pilipovic, P., Samson, A. And Ditlevsen, S. (2024): Efficient estimation for ergodic diffusion processes sampled at high frequency. Ann. Statist., 52, 842 - 867.
17. Juni 2026
Anna Calissano ( University College London)
Abstract: A spatial graph is a specific type of graph with spatial attributes associated with the nodes and the edges. It is a smart modelling choice for capturing the skeleton of a shape, a blood vessel network, a porous tissue, and many other data objects with intrinsically complex geometry. In this talk, we describe how spatial graphs can be analysed using a specific metric (the Fused Gromov–Wasserstein metric). We extend a testing procedure between distributions of spatial graphs, a depth measure to describe the distribution of spatial graphs, and a dimensionality reduction procedure based on preserving key topological features. We present this variety of methods on a dataset of cardiac fibrosis tissue and on a dataset of fungus mycelium networks.
24. Juni 2026
Bertrand Even (LMO, Orsay)
Abstract: In many high-dimensional problems, like sparse-PCA, planted clique, or clustering, the best known algorithms with polynomial time complexity fail to reach the statistical performance provably achievable by algorithms free of computational constraints. This observation has given rise to the conjecture of the existence, for some problems, of gaps - so called statistical-computational gaps - between the best possible statistical performance achievable without computational constraints, and the best performance achievable with poly-time algorithms. A powerful approach to predict the best performance achievable in poly-time is to investigate the best performance achievable by polynomials with low-degree. In this talk, I will introduce this low degree framework for estimation problems, which was first defined in 2022 in a paper by Alex Wein and Tselil Schramm. In a second time, I will talk about a specific estimation problem; Clustering Gaussian Mixture Models, for which there is a statistical-computational gap when the dimension and the number of clusters are large.
08. Juli 2026
Alexandra Carpentier (Universität Potsdam)
Computational lower bounds in clustering and ranking
Abstract: In this talk, we will speak about computational lower bounds. After an extended general introduction that presents these tools, and we will focus more particularly on the problem of clustering.
15. Juli 2026
Sascha Gaudlitz (HUB)
[1] Yifan Chen, Bamdad Hosseini, Houman Owhadi, Andrew M. Stuart. Solving and learning nonlinear
Interessenten sind herzlich eingeladen.
Für Rückfragen wenden Sie sich bitte an:
Frau Marina Filatova
Mail: marina.filatova@hu-berlin.de
Telefon: +49-30-2093-45460
Humboldt-Universität zu Berlin
Institut für Mathematik
Unter den Linden 6
10099 Berlin, Germany