Source author record

Daniel Vogel

Daniel Vogel appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

18works
9topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

18 published item(s)

preprint2021arXiv

Automated Identification of Vulnerable Devices in Networks using Traffic Data and Deep Learning

Many IoT devices are vulnerable to attacks due to flawed security designs and lacking mechanisms for firmware updates or patches to eliminate the security vulnerabilities. Device-type identification combined with data from vulnerability databases can pinpoint vulnerable IoT devices in a network and can be used to constrain the communications of vulnerable devices for preventing damage. In this contribution, we present and evaluate two deep learning approaches to the reliable IoT device-type identification, namely a recurrent and a convolutional network architecture. Both deep learning approaches show accuracies of 97% and 98%, respectively, and thereby outperform an up-to-date IoT device-type identification approach using hand-crafted fingerprint features obtaining an accuracy of 82%. The runtime performance for the IoT identification of both deep learning approaches outperforms the hand-crafted approach by three magnitudes. Finally, importance metrics explain the results of both deep learning approaches in terms of the utilization of the analyzed traffic data flow.

preprint2016arXiv

Computing the Oja Median in R: The Package OjaNP

The Oja median is one of several extensions of the univariate median to the multivariate case. It has many nice properties, but is computationally demanding. In this paper, we first review the properties of the Oja median and compare it to other multivariate medians. Afterwards we discuss four algorithms to compute the Oja median, which are implemented in our R-package OjaNP. Besides these algorithms, the package contains also functions to compute Oja signs, Oja signed ranks, Oja ranks, and the related scatter concepts. To illustrate their use, the corresponding multivariate one- and $C$-sample location tests are implemented.

preprint2016arXiv

On the eigenvalues of the spatial sign covariance matrix in more than two dimensions

We gather several results on the eigenvalues of the spatial sign covariance matrix of an elliptical distribution. It is shown that the eigenvalues are a one-to-one function of the eigenvalues of the shape matrix and that they are closer together than the latter. We further provide a one-dimensional integral representation of the eigenvalues, which facilitates their numerical computation.

preprint2016arXiv

Studentized U-quantile processes under dependence with applications to change-point analysis

Many popular robust estimators are $U$-quantiles, most notably the Hodges-Lehmann location estimator and the $Q_n$ scale estimator. We prove a functional central limit theorem for the sequential $U$-quantile process without any moment assumptions and under weak short-range dependence conditions. We further devise an estimator for the long-run variance and show its consistency, from which the convergence of the studentized version of the sequential $U$-quantile process to a standard Brownian motion follows. This result can be used to construct CUSUM-type change-point tests based on $U$-quantiles, which do not rely on bootstrapping procedures. We demonstrate this approach in detail at the example of the Hodges-Lehmann estimator for robustly detecting changes in the central location. A simulation study confirms the very good robustness and efficiency properties of the test. Two real-life data sets are analyzed.

preprint2016arXiv

Testing for Changes in Kendall's Tau

For a bivariate time series $((X_i,Y_i))_{i=1,...,n}$ we want to detect whether the correlation between $X_i$ and $Y_i$ stays constant for all $i = 1,...,n$. We propose a nonparametric change-point test statistic based on Kendall's tau and derive its asymptotic distribution under the null hypothesis of no change by means a new U-statistic invariance principle for dependent processes. The asymptotic distribution depends on the long run variance of Kendall's tau, for which we propose an estimator and show its consistency. Furthermore, assuming a single change-point, we show that the location of the change-point is consistently estimated. Kendall's tau possesses a high efficiency at the normal distribution, as compared to the normal maximum likelihood estimator, Pearson's moment correlation coefficient. Contrary to Pearson's correlation coefficient, it has excellent robustness properties and shows no loss in efficiency at heavy-tailed distributions. We assume the data $((X_i,Y_i))_{i=1,...,n}$ to be stationary and P-near epoch dependent on an absolutely regular process. The P-near epoch dependence condition constitutes a generalization of the usually considered $L_p$-near epoch dependence, $p \ge 1$, that does not require the existence of any moments. It is therefore very well suited for our objective to efficiently detect changes in correlation for arbitrarily heavy-tailed data.

preprint2016arXiv

Tests for scale changes based on pairwise differences

In many applications it is important to know whether the amount of fluctuation in a series of observations changes over time. In this article, we investigate different tests for detecting change in the scale of mean-stationary time series. The classical approach based on the CUSUM test applied to the squared centered, is very vulnerable to outliers and impractical for heavy-tailed data, which leads us to contemplate test statistics based on alternative, less outlier-sensitive scale estimators. It turns out that the tests based on Gini's mean difference (the average of all pairwise distances) or generalized Qn estimators (sample quantiles of all pairwise distances) are very suitable candidates. They improve upon the classical test not only under heavy tails or in the presence of outliers, but also under normality. An explanation for this at first counterintuitive result is that the corresponding long-run variance estimates are less affected by a scale change than in the case of the sample-variance-based test. We use recent results on the process convergence of U-statistics and U-quantiles for dependent sequences to derive the limiting distribution of the test statistics and propose estimators for the long-run variance. We perform a simulations study to investigate the finite sample behavior of the test and their power. Furthermore, we demonstrate the applicability of the new change-point detection methods at two real-life data examples from hydrology and finance.

preprint2016arXiv

The spatial sign covariance matrix and its application for robust correlation estimation

We summarize properties of the spatial sign covariance matrix and especially look at the relationship between its eigenvalues and those of the shape matrix of an elliptical distribution. The explicit relationship known in the bivariate case was used to construct the spatial sign correlation coefficient, which is a non-parametric and robust estimator for the correlation coefficient within the elliptical model. We consider a multivariate generalization, which we call the multivariate spatial sign correlation matrix.

preprint2015arXiv

A Dataset of Naturally Occurring, Whole-Body Background Activity to Reduce Gesture Conflicts

In real settings, natural body movements can be erroneously recognized by whole-body input systems as explicit input actions. We call body activity not intended as input actions "background activity." We argue that understanding background activity is crucial to the success of always-available whole-body input in the real world. To operationalize this argument, we contribute a reusable study methodology and software tools to generate standardized background activity datasets composed of data from multiple Kinect cameras, a Vicon tracker, and two high-definition video cameras. Using our methodology, we create an example background activity dataset for a television-oriented living room setting. We use this dataset to demonstrate how it can be used to redesign a gestural interaction vocabulary to minimize conflicts with the real world. The software tools and initial living room dataset are publicly available (http://www.dgp.toronto.edu/~dustin/backgroundactivity/).

preprint2015arXiv

Asymptotics of the two-stage spatial sign correlation

The spatial sign correlation (Dürre, Vogel and Fried, 2015) is a highly robust and easy-to-compute, bivariate correlation estimator based on the spatial sign covariance matrix. Since the estimator is inefficient when the marginal scales strongly differ, a two-stage version was proposed. In the first step, the observations are marginally standardized by means of a robust scale estimator, and in the second step, the spatial sign correlation of the thus transformed data set is computed. Dürre et al. (2015) give some evidence that the asymptotic distribution of the two-stage estimator equals that of the spatial sign correlation at equal marginal scales by comparing their influence functions and presenting simulation results, but give no formal proof. In the present paper, we close this gap and establish the asymptotic normality of the two-stage spatial sign correlation and compute its asymptotic variance for elliptical population distributions. We further derive a variance-stabilizing transformation, similar to Fisher's z-transform, and numerically compare the small-sample coverage probabilities of several confidence intervals.

preprint2015arXiv

Elliptical graphical modelling

We propose elliptical graphical models based on conditional uncorrelatedness as a general- ization of Gaussian graphical models by letting the population distribution be elliptical instead of normal, allowing the fitting of data with arbitrarily heavy tails. We study the class of propor- tionally affine equivariant scatter estimators and show how they can be used to perform elliptical graphical modelling, leading to a new class of partial correlation estimators and analogues of the classical deviance test. General expressions for the asymptotic variance of partial correla- tion estimators, unconstrained and under decomposable models, are given, and the asymptotic chi square approximation of the pseudo-deviance test statistic is proved. The feasibility of our approach is demonstrated by a simulation study, using, among others, Tyler's scatter estimator, which is distribution-free within the elliptical model. Our approach provides a robustification of Gaussian graphical modelling. The latter is likelihood-based and known to be very sensitive to model misspecification and outlying observations.

preprint2015arXiv

Estimation of the variance of partial sums of dependent processes

We study subsampling estimators for the limit variance \[ σ^2=Var(X_1)+2 \sum_{k=2}^\infty Cov(X_1,X_k) \] of partial sums of a stationary stochastic process $(X_k)_{k\geq 1}$. We establish $L_2$-consistency of a non-overlapping block resampling method. Our results apply to processes that can be represented as functionals of strongly mixing processes. Motivated by recent applications to rank tests, we also study estimators for the series $Var(F(X_1))+2 \sum_{k=2}^\infty Cov(F(X_1),F(X_k))$, where $F$ is the distribution function of $X_1$. Simulations illustrate the usefulness of the proposed estimators and of a mean squared error optimal rule for the choice of the block length.

preprint2015arXiv

Weak convergence of the empirical process and the rescaled empirical distribution function in the Skorokhod product space

We prove the asymptotic independence of the empirical process $α_n = \sqrt{n}( F_n - F)$ and the rescaled empirical distribution function $β_n = n (F_n(τ+\frac{\cdot}{n})-F_n(τ))$, where $F$ is an arbitrary cdf, differentiable at some point $τ$, and $F_n$ the corresponding empricial cdf. This seems rather counterintuitive, since, for every $n \in N$, there is a deterministic correspondence between $α_n$ and $β_n$. Precisely, we show that the pair $(α_n,β_n)$ converges in law to a limit having independent components, namely a time-transformed Brownian bridge and a two-sided Poisson process. Since these processes have jumps, in particular if $F$ itself has jumps, the Skorokhod product space $D(R) \times D(R)$ is the adequate choice for modeling this convergence in. We develop a short convergence theory for $D(R) \times D(R)$ by establishing the classical principle, devised by Yu. V. Prokhorov, that finite-dimensional convergence and tightness imply weak convergence. Several tightness criteria are given. Finally, the convergence of the pair $(α_n,β_n)$ implies convergence of each of its components, thus, in passing, we provide a thorough proof of these known convergence results in a very general setting. In fact, the condition on $F$ to be differentiable in at least one point is only required for $β_n$ to converge and can be further weakened.

preprint2014arXiv

A fluctuation test for constant Spearman's rho with nuisance-free limit distribution

A CUSUM type test for constant correlation that goes beyond a previously suggested correlation constancy test by considering Spearman's rho in arbitrary dimensions is proposed. Since the new test does not require the existence of any moments, the applicability on usually heavy-tailed financial data is greatly improved. The asymptotic null distribution is calculated using an invariance principle for the sequential empirical copula process. The limit distribution is free of nuisance parameters and critical values can be obtained without bootstrap techniques. A local power result and an analysis of the behavior of the test in small samples is provided.

preprint2013arXiv

Robust estimators for non-decomposable elliptical graphical models

Asymptotic properties of scatter estimators for elliptical graphical models are studied. Such models impose a given pattern of zeros on the inverse of the shape matrix of an elliptically distributed random vector. In particular, we introduce the class of graphical M-estimators and compare them to plug-in M-estimators. It turns out that, under suitable conditions, both approaches yield the same asymptotic efficiency. Furthermore, the results of this paper apply to both decomposable and non-decomposable graphical models and so generalize the results for decomposable models given by Vogel & Fried (2011) for the plug-in M-estimators.