Source author record

Irad Ben-Gal

Irad Ben-Gal appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

2works
3topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

2 published item(s)

preprint2021arXiv

A Non-Parametric Subspace Analysis Approach with Application to Anomaly Detection Ensembles

Identifying anomalies in multi-dimensional datasets is an important task in many real-world applications. A special case arises when anomalies are occluded in a small set of attributes, typically referred to as a subspace, and not necessarily over the entire data space. In this paper, we propose a new subspace analysis approach named Agglomerative Attribute Grouping (AAG) that aims to address this challenge by searching for subspaces that are comprised of highly correlative attributes. Such correlations among attributes represent a systematic interaction among the attributes that can better reflect the behavior of normal observations and hence can be used to improve the identification of two particularly interesting types of abnormal data samples: anomalies that are occluded in relatively small subsets of the attributes and anomalies that represent a new data class. AAG relies on a novel multi-attribute measure, which is derived from information theory measures of partitions, for evaluating the "information distance" between groups of data attributes. To determine the set of subspaces to use, AAG applies a variation of the well-known agglomerative clustering algorithm with the proposed multi-attribute measure as the underlying distance function. Finally, the set of subspaces is used in an ensemble for anomaly detection. Extensive evaluation demonstrates that, in the vast majority of cases, the proposed AAG method (i) outperforms classical and state-of-the-art subspace analysis methods when used in anomaly detection ensembles, and (ii) generates fewer subspaces with a fewer number of attributes each (on average), thus resulting in a faster training time for the anomaly detection ensemble. Furthermore, in contrast to existing methods, the proposed AAG method does not require any tuning of parameters.

preprint2014arXiv

Information Spread in a Connected World

In the following work, we compare the spread of information by word-of-mouth (WOM) to the spread of information through search engines. We assume that the initial acknowledgement of new information derives from social interactions but that solid opinions are only formed after further evaluation through search engines. Search engines can be viewed as central hubs that connect information presented in relevant websites to searchers. Since they construct new connections between searchers and information in every query performed, the network structure is less relevant. Although models of viral spread of ideas have been inspected in many previous works [1], [2], [3], [4], [5], [6], [7], [8], [9], [10], [11], [12], [13], only few assume the acceptance of a novel concept to be solely based on the evaluation of the opinions of others [8], [5]. Following this approach, combined with that of models of information spread with threshold [1] that claim the propagation in a network to occur only if a threshold of neighbors hold an opinion, the proposed work adds a new theoretical perspective that is relevant to the daily use of search engines as a major information search tool. We continue by presenting some justifications based on experimentations. Last we discuss possible outcomes of over use of search engines vs. WOM, and suggest an hypothesis that such overuse might actually narrow the collective information set.