arXiv ScienceSearch

arXiv subjects

Sebastian Fuchs

Publications and source records attributed to Sebastian Fuchs.

At least 19 recordsLinked to original sources

An MTP$_2$ property for conditional distributions

We introduce a new notion of positive dependence, namely multivariate total positivity of order two for conditional distributions, denoted cMTP$_2$, that explicitly incorporates the conditioning variables. This property is weaker than multivariate total positivity of order two for densities, but stronger than multivariate total positivity of order two for distribution functions, stochastic monotonicity, and tail monotonicity. It thus provides a natural intermediate concept linking these classical notions of positive dependence. We verify the cMTP$_2$ property for several distributions and copula families, including Archimedean copulas, and establish stability under Markov product transformations. As a consequence, we derive comparison results for measures of directed dependence arising from Markov structures, including Chatterjee's rank correlation and a related sensitivity measure.

stat.ME

An ordering for the strength of functional dependence

We introduce a new dependence order, termed the conditional convex order, whose minimal and maximal elements characterize independence and perfect dependence. Moreover, it characterizes conditional independence, satisfies information monotonicity, and exhibits several invariance properties. Consequently, it is an ordering for the strength of functional dependence of a random variable Y on a random vector X. As we show, various recently studied dependence measures -- including Chatterjee's rank correlation, Wasserstein correlations, and rearranged dependence measures -- are increasing in this order and inherit their fundamental properties from it. We characterize the conditional convex order by the Schur order and by the concordance order, and we verify it in settings such as additive error models, the multivariate normal distribution, and various copula-based models. Our results offer a unified perspective on the behavior of dependence measures across statistical models.

math.ST

The Fej\'er-Dirichlet Lift: Entire Functions and $\zeta$-Factorization Identities

A Fej\'er-Dirichlet lift is developed that turns divisor information at the integers into entire interpolants with explicit Dirichlet-series factorizations. For absolutely summable weights the lift interpolates $(a*1)(n)$ at each integer $n$ and has Dirichlet series $\zeta(s)A(s)$ on $\Re s>1$. Two applications are emphasized. First, for $q>1$ an entire function $\mathfrak F(\cdot,q)$ is constructed that vanishes at primes and is positive at composite integers; a tangent-matched variant $\mathfrak F^{\sharp}$ is shown to admit an explicit, effective threshold $P_0(q)$ such that for every odd prime $p\ge P_0(q)$ the interval $(p-1,p)$ is free of real zeros and $x=p$ is a boundary zero of multiplicity two. Second, a renormalized lift for $a=\mu*\Lambda$ produces an entire interpolant of $\Lambda(n)$ and provides a constructive viewpoint on the appearance of $\zeta'(s)/\zeta(s)$ through the FD-lift spectrum. A Polylog-Zeta factorization for the geometric-weight case links $\zeta(s)$ with $\operatorname{Li}_s(1/q)$. All prime/composite statements concern integer arguments. Scripts reproducing figures and numerical checks are provided in a public repository with an archival snapshot.

math.GM

On exact regions between measures of concordance and Chatterjee's rank correlation for lower semilinear copulas

We explore how the classical concordance measures - Kendall's $\tau$, Spearman's rank correlation $\rho$, and Spearman's footrule $\phi$ - relate to Chatterjee's rank correlation $\xi$ when restricted to lower semilinear copulas. First, we provide a complete characterization of the attainable $\tau$-$\rho$ region for this class, thus resolving the conjecture in [18]. Building on this result, we then derive the exact $\tau$-$\phi$ and $\phi$-$\rho$ regions, obtain a closed-form relationship between $\xi$ and $\tau$, and establish the exact $\tau$-$\xi$ region. In particular, we prove that $\xi$ never exceeds $\tau$, $\rho$, or $\phi$. Our results clarify the relationship between undirected and directed dependence measures and reveal novel insights into the dependence structures that result from lower semilinear copulas.

stat.ME

Fej\'er-Kernel Prime Indicators

A $C^1$ prime indicator $\mathcal{P}\colon\mathbb{R}\to\mathbb{R}$ is constructed by applying the Fej\'er identity to the sine-quotient encoder of trial division. For integers $n\ge 2$, $\mathcal P(n)=0$ holds exactly for odd primes; $\mathcal P(2)>0$. For all non-integers $x>1$ one has $\mathcal P(x)>0$. The function is piecewise $C^\infty$ and its second derivative has jumps precisely at the squares $m^2$, with explicit sizes. Replacing the sharp cut-off by a smooth transition yields $C^\infty$ analogues $\mathcal{P}_\tau$ and $\mathcal{P}_\sigma$ with integer limits $\mathcal{P}_\tau(n;\kappa)\to \tau(n)-2$ and $\mathcal{P}_\sigma(n;\kappa)\to \sigma(n)-n-1$ as $\kappa\to\infty$, obtained from locally uniform convergence of derivative series. For large $\kappa$, numerical evidence indicates companion zeros near odd primes for $\mathcal{P}_\tau$ and an asymmetric pair for $\mathcal{P}_\sigma$. No assertion is made beyond integer input, and no statements are claimed about the prime number theorem or zero distributions of $L$-functions. The appendix includes two illustrative prime-counting sums.

math.NT

A dimension reduction for extreme types of directed dependence

In recent years, a variety of novel measures of dependence have been introduced being capable of characterizing diverse types of directed dependence, hence diverse types of how a number of predictor variables $\mathbf{X} = (X_1, \dots, X_p)$, $p \in \mathbb{N}$, may affect a response variable $Y$. This includes perfect dependence of $Y$ on $\mathbf{X}$ and independence between $\mathbf{X}$ and $Y$, but also less well-known concepts such as zero-explainability, stochastic comparability and complete separation. Certain such measures offer a representation in terms of the Markov product $(Y,Y')$, with $Y'$ being a conditionally independent copy of $Y$ given $\mathbf{X}$. This dimension reduction principle allows these measures to be estimated via the powerful nearest neighbor based estimation principle introduced in [4]. To achieve a deeper insight into the dimension reduction principle, this paper aims at translating the extreme variants of directed dependence, typically formulated in terms of the random vector $(\mathbf{X},Y)$, into the Markov product $(Y,Y')$.

math.ST

A new coefficient of separation

A coefficient is introduced that quantifies the extent of separation of a random variable $Y$ relative to a number of variables $\mathbf{X} = (X_1, \dots, X_p)$ by skillfully assessing the sensitivity of the relative effects of the conditional distributions. The coefficient is as simple as classical dependence coefficients such as Kendall's tau, also requires no distributional assumptions, and consistently estimates an intuitive and easily interpretable measure, which is $0$ if and only if $Y$ is stochastically comparable relative to $\mathbf{X}$, that is, the values of $Y$ show no location effect relative to $\mathbf{X}$, and $1$ if and only if $Y$ is completely separated relative to $\mathbf{X}$. As a true generalization of the classical relative effect, in applications such as medicine and the social sciences the coefficient facilitates comparing the distributions of any number of treatment groups or categories. It hence avoids the sometimes artificial grouping of variable values such as patient's age into just a few categories, which is known to cause inaccuracy and bias in the data analysis. The mentioned benefits are exemplified using synthetic and real data sets.

stat.ME

On continuity of Chatterjee's rank correlation and related dependence measures

While measures of concordance -- such as Spearman's rho, Kendall's tau, and Blomqvist's beta -- are continuous with respect to weak convergence, Chatterjee's rank correlation xi recently introduced in Azadkia and Chatterjee (2021) does not share this property, causing drawbacks in statistical inference as pointed out in B\"ucher and Dette (2025). As we study in this paper, xi is instead weakly continuous with respect to conditionally independent copies -- the Markov products. To establish weak continuity of Markov products, we provide several sufficient conditions, including copula-based criteria and conditions relying on the concept of conditional weak convergence in Sweeting (1989). As a consequence, we also obtain continuity results for xi and related dependence measures and verify their continuity in the parameters of standard models such as multivariate elliptical and l1-norm symmetric distributions.

math.ST

Hierarchical variable clustering based on the predictive strength between random vectors

A rank-invariant clustering of variables is introduced that is based on the predictive strength between groups of variables, i.e., two groups are assigned a high similarity if the variables in the first group contain high predictive information about the behaviour of the variables in the other group and/or vice versa. The method presented here is model-free, dependence-based and does not require any distributional assumptions. Various general invariance and continuity properties are investigated, with special attention to those that are beneficial for the agglomerative hierarchical clustering procedure. A fully non-parametric estimator is considered whose excellent performance is demonstrated in several simulation studies and by means of real-data examples.

stat.ME

Quantifying and estimating dependence via sensitivity of conditional distributions

Recently established, directed dependence measures for pairs $(X,Y)$ of random variables build upon the natural idea of comparing the conditional distributions of $Y$ given $X=x$ with the marginal distribution of $Y$. They assign pairs $(X,Y)$ values in $[0,1]$, the value is $0$ if and only if $X,Y$ are independent, and it is $1$ exclusively for $Y$ being a function of $X$. Here we show that comparing randomly drawn conditional distributions with each other instead or, equivalently, analyzing how sensitive the conditional distribution of $Y$ given $X=x$ is on $x$, opens the door to constructing novel families of dependence measures $\Lambda_\varphi$ induced by general convex functions $\varphi: \mathbb{R} \rightarrow \mathbb{R}$, containing, e.g., Chatterjee's coefficient of correlation as special case. After establishing additional useful properties of $\Lambda_\varphi$ we focus on continuous $(X,Y)$, translate $\Lambda_\varphi$ to the copula setting, consider the $L^p$-version and establish an estimator which is strongly consistent in full generality. A real data example and a simulation study illustrate the chosen approach and the performance of the estimator. Complementing the afore-mentioned results, we show how a slight modification of the construction underlying $\Lambda_\varphi$ can be used to define new measures of explainability generalizing the fraction of explained variance.

math.ST

A novel positive dependence property and its impact on a popular class of concordance measures

A novel positive dependence property is introduced, called positive measure inducing (PMI for short), being fulfilled by numerous copula classes, including Gaussian, Fr\'echet, Farlie-Gumbel-Morgenstern and Frank copulas; it is conjectured that even all positive quadrant dependent Archimedean copulas meet this property. From a geometric viewpoint, a PMI copula concentrates more mass near the main diagonal than in the opposite diagonal. A striking feature of PMI copulas is that they impose an ordering on a certain class of copula-induced measures of concordance, the latter originating in Edwards et al. (2004) and including Spearman's rho $\rho$ and Gini's gamma $\gamma$, leading to numerous new inequalities such as $3 \gamma \geq 2 \rho$. The measures of concordance within this class are estimated using (classical) empirical copulas and the intrinsic construction via empirical checkerboard copulas, and the estimators' asymptotic behaviour is determined. Building upon the presented inequalities, asymptotic tests are constructed having the potential of being used for detecting whether the underlying dependence structure of a given sample is PMI, which in turn can be used for excluding certain copula families from model building. The excellent performance of the tests is demonstrated in a simulation study and by means of a real-data example.

stat.ME

A direct extension of Azadkia & Chatterjee's rank correlation to multi-response vectors

Recently, Chatterjee (2023) recognized the lack of a direct generalization of his rank correlation $\xi$ in Azadkia and Chatterjee (2021) to a multi-dimensional response vector. As a natural solution to this problem, we here propose an extension of $\xi$ that is applicable to a set of $q \geq 1$ response variables, where our approach builds upon converting the original vector-valued problem into a univariate problem and then applying the rank correlation $\xi$ to it. Our novel measure $T$ quantifies the scale-invariant extent of functional dependence of a response vector $\mathbf{Y} = (Y_1,\dots,Y_q)$ on predictor variables $\mathbf{X} = (X_1, \dots,X_p)$, characterizes independence of $\mathbf{X}$ and $\mathbf{Y}$ as well as perfect dependence of $\mathbf{Y}$ on $\mathbf{X}$ and hence fulfills all the characteristics of a measure of predictability. Aiming at maximum interpretability, we provide various invariance results for $T$ as well as a closed-form expression in multivariate normal models. Building upon the graph-based estimator for $\xi$ in Azadkia and Chatterjee (2021), we obtain a non-parametric, strongly consistent estimator for $T$ and show -- as a main contribution -- its asymptotic normality. Based on this estimator, we develop a model-free and rank-based feature ranking and forward feature selection for multiple-outcome data that works without any tuning parameters. Simulation results and real case studies illustrate $T$'s broad applicability.

math.ST

Total positivity of copulas from a Markov kernel perspective

The underlying dependence structure between two random variables can be described in manifold ways. This includes the examination of certain dependence properties such as lower tail decreasingness (LTD), stochastic increasingness (SI) or total positivity of order 2, the latter usually considered for a copula (TP2) or (if existent) its density (d-TP2). In the present paper we investigate total positivity of order 2 for a copula's Markov kernel (MK-TP2 for short), a positive dependence property that is stronger than TP2 and SI, weaker than d-TP2 but, unlike d-TP2, is not restricted to absolutely continuous copulas, making it presumably the strongest dependence property defined for any copula (including those with a singular part such as Marshall-Olkin copulas). We examine the MK-TP2 property for different copula families, among them the class of Archimedean copulas and the class of extreme value copulas. In particular we show that, within the class of Archimedean copulas, the dependence properties SI and MK-TP2 are equivalent.

math.ST

Quantifying directed dependence via dimension reduction

Studying the multivariate extension of copula correlation yields a dimension reduction principle, which turns out to be strongly related with the `simple measure of conditional dependence' $T$ recently introduced by Azadkia & Chatterjee (2021). In the present paper, we identify and investigate the dependence structure underlying this dimension-reduction principle, provide a strongly consistent estimator for it, and demonstrate its broad applicability. For that purpose, we define a bivariate copula capturing the scale-invariant extent of dependence of an endogenous random variable $Y$ on a set of $d \geq 1$ exogenous random variables ${\bf X} = (X_1, \dots, X_d)$, and containing the information whether $Y$ is completely dependent on ${\bf X}$, and whether $Y$ and ${\bf X}$ are independent. The dimension reduction principle becomes apparent insofar as the introduced bivariate copula can be viewed as the distribution function of two random variables $Y$ and $Y^\prime$ sharing the same conditional distribution and being conditionally independent given ${\bf X}$. Evaluating this copula uniformly along the diagonal, i.e. calculating Spearman's footrule, leads to Azadkia and Chatterjee's `simple measure of conditional dependence' $T$. On the other hand, evaluating this copula uniformly over the unit square, i.e. calculating Spearman's rho, leads to a distribution-free coefficient of determination (a.k.a. copula correlation). Several real data examples illustrate the importance of the introduced methodology.

math.ST

A copula transformation in multivariate mixed discrete-continuous models

Copulas allow a flexible and simultaneous modeling of complicated dependence structures together with various marginal distributions. Especially if the density function can be represented as the product of the marginal density functions and the copula density function, this leads to both an intuitive interpretation of the conditional distribution and convenient estimation procedures. However, this is no longer the case for copula models with mixed discrete and continuous marginal distributions, because the corresponding density function cannot be decomposed so nicely. In this paper, we introduce a copula transformation method that allows to represent the density function of a distribution with mixed discrete and continuous marginals as the product of the marginal probability mass/density functions and the copula density function. With the proposed method, conditional distributions can be described analytically and the computational complexity in the estimation procedure can be reduced depending on the type of copula used.

stat.ME

Superradiance from non-ideal initial states -- a quantum trajectory approach

Collective emission behavior is usually described by the decay dynamics of the completely symmetric Dicke states. To study a more realistic scenario, we investigate alternative initial states inducing a more complex time evolution. Superposition states of the fully inverted Dicke state and the Dicke ground state with unequal mutual weights are studied as examples as well as superradiance stemming from atoms in clusters separated by more than one wavelength. The Monte Carlo wave function method serves as framework to study the dynamics of quantum states, which is determined by quantum jumps on the one hand and continuous evolution dynamics on the other hand. We compare this method with the classical picture of a system of rate equations written for the diagonal components of the density matrix.

quant-ph

Dissimilarity functions for rank-invariant hierarchical clustering of continuous variables

A theoretical framework is presented for a (copula-based) notion of dissimilarity between continuous random vectors and its main properties are studied. The proposed dissimilarity assigns the smallest value to a pair of random vectors that are comonotonic. Various properties of this dissimilarity are studied, with special attention to those that are prone to the hierarchical agglomerative methods, such as reducibility. Some insights are provided for the use of such a measure in clustering algorithms and a simulation study is presented. Real case studies illustrate the main features of the whole methodology.

stat.ME

On weak conditional convergence of bivariate Archimedean and Extreme Value copulas, and consequences to nonparametric estimation

Looking at bivariate copulas from the perspective of conditional distributions and considering weak convergence of almost all conditional distributions yields the notion of weak conditional convergence. At first glance, this notion of convergence for copulas might seem far too restrictive to be of any practical importance - in fact, given samples of a copula $C$ the corresponding empirical copulas do not converge weakly conditional to $C$ with probability one in general. Within the class of Archimedean copulas and the class of Extreme Value copulas, however, standard pointwise convergence and weak conditional convergence can even be proved to be equivalent. Moreover, it can be shown that every copula $C$ is the weak conditional limit of a sequence of checkerboard copulas. After proving these three main results and pointing out some consequences we sketch some implications for two recently introduced dependence measures and for the nonparametric estimation of Archimedean and Extreme Value copulas.

math.ST