<?xml version="1.0" encoding="UTF-8"?><?xml-model type="application/xml-dtd" href="http://jats.nlm.nih.gov/publishing/1.1d3/JATS-journalpublishing1.dtd"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.1d3 20150301//EN" "http://jats.nlm.nih.gov/publishing/1.1d3/JATS-journalpublishing1.dtd">
<article xmlns:ali="http://www.niso.org/schemas/ali/1.0" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:mml="http://www.w3.org/1998/Math/MathML" dtd-version="1.1d3" specific-use="Marcalyc 1.2" article-type="research-article" xml:lang="en">
<front>
<journal-meta>
<journal-id journal-id-type="redalyc">205</journal-id>
<journal-title-group>
<journal-title specific-use="original" xml:lang="es">Cuadernos de Administración</journal-title>
<abbrev-journal-title abbrev-type="publisher" xml:lang="es">Cuad Adm</abbrev-journal-title>
</journal-title-group>
<issn pub-type="ppub">0120-3592</issn>
<issn pub-type="epub">1900-7205</issn>
<publisher>
<publisher-name>Pontificia Universidad Javeriana</publisher-name>
<publisher-loc>
<country>Colombia</country>
<email>revistascientificasjaveriana@gmail.com</email>
</publisher-loc>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="art-access-id" specific-use="redalyc">20562876002</article-id>
<article-id pub-id-type="doi">https://doi.org/10.11144/Javeriana.cao33.ppado</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Artículos</subject>
</subj-group>
</article-categories>
<title-group>
<article-title xml:lang="en">
<bold>Projection pursuit algorithms to detect outliers<xref ref-type="fn" rid="fn4">*</xref>
</bold>
</article-title>
<trans-title-group>
<trans-title xml:lang="es">
<bold>Algoritmos de búsqueda de proyección para detectar valores atípicos</bold>
</trans-title>
</trans-title-group>
<trans-title-group>
<trans-title xml:lang="pt">
<bold>Algoritmos de busca de projeção para detectar valores atípicos</bold>
</trans-title>
</trans-title-group>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<contrib-id contrib-id-type="orcid">http://orcid.org/0000-0001-7277-1638</contrib-id>
<name name-style="western">
<surname>Stimolo</surname>
<given-names>Maria Inés</given-names>
</name>
<xref ref-type="corresp" rid="corresp1"><sup>a</sup></xref>
<xref ref-type="aff" rid="aff1"/>
<email>maria.ines.stimolo@unc.edu.ar</email>
</contrib>
<contrib contrib-type="author" corresp="no">
<contrib-id contrib-id-type="orcid">http://orcid.org/0000-0002-3777-0653</contrib-id>
<name name-style="western">
<surname>Ortiz</surname>
<given-names>Pablo Arnaldo</given-names>
</name>
<xref ref-type="aff" rid="aff2"/>
</contrib>
</contrib-group>
<aff id="aff1">
<institution content-type="original">Universidad Nacional de Córdoba - Facultad de Ciencias Económicas, Argentina</institution>
<institution content-type="orgname">Universidad Nacional de Córdoba - Facultad de Ciencias Económicas</institution>
<country country="AR">Argentina</country>
</aff>
<aff id="aff2">
<institution content-type="original">Universidad Nacional de Córdoba - Facultad de Ciencias Económicas, Argentina</institution>
<institution content-type="orgname">Universidad Nacional de Córdoba - Facultad de Ciencias Económicas</institution>
<country country="AR">Argentina</country>
</aff>
<author-notes>
<corresp id="corresp1"><sup>a</sup> Corresponding author. E-mail: <email>maria.ines.stimolo@unc.edu.ar</email>
</corresp>
</author-notes>
<pub-date pub-type="epub-ppub">
<season>Enero-Diciembre</season>
<year>2020</year>
</pub-date>
<volume>33</volume>
<history>
<date date-type="received" publication-format="dd/mm/yyyy">
<day>26</day>
<month>08</month>
<year>2019</year>
</date>
<date date-type="accepted" publication-format="dd/mm/yyyy">
<day>20</day>
<month>10</month>
<year>2019</year>
</date>
<date date-type="pub" publication-format="dd/mm/yyyy">
<day>20</day>
<month>05</month>
<year>2020</year>
</date>
</history>
<permissions>
<ali:free_to_read/>
<license xlink:href="https://creativecommons.org/licenses/by/4.0/">
<ali:license_ref>https://creativecommons.org/licenses/by/4.0/</ali:license_ref>
<license-p>Esta obra está bajo una Licencia Creative Commons Atribución 4.0 Internacional.</license-p>
</license>
</permissions>
<abstract xml:lang="en">
<title>Abstract</title>
<p>In this paper, we compare the methods proposed by <xref ref-type="bibr" rid="redalyc_20562876002_ref11">Peña and Prieto (2001)</xref>, and <xref ref-type="bibr" rid="redalyc_20562876002_ref6">Filzmoser, Maronna, and Werner (2008)</xref> to detect outliers in a set of Argentine companies that quote their shares in the Stock Exchange. A significant heterogeneity between observations can be a consequence of the presence of outliers. The detection of outliers is an important task for the statistical analysis since they distort descriptive measures and parameters estimators. There are different multivariate methods to detect outliers, such as distance-based methods and projection pursuit methods.</p>
</abstract>
<trans-abstract xml:lang="es">
<title>Resumen</title>
<p>En este trabajo se comparan los métodos propuestos por <xref ref-type="bibr" rid="redalyc_20562876002_ref11">Peña y Prieto (2001)</xref> y <xref ref-type="bibr" rid="redalyc_20562876002_ref6">Filzmoser, Maronna y Werner (2008)</xref> para detectar datos atípicos en empresas argentinas que cotizan sus acciones en el Mercado de Valores. La heterogeneidad significativa entre observaciones puede ser una consecuencia de la presencia de datos atípicos. La detección de datos atípicos es importante en el análisis estadístico por su efecto en la distorsión de las medidas descriptivas y en los estimadores de los parámetros. Existen distintos métodos multivariados para detectar datos atípicos, tales como los métodos basados en la distancia o los métodos de búsqueda de proyecciones.</p>
</trans-abstract>
<trans-abstract xml:lang="pt">
<title>Resumo</title>
<p>Este trabalho compara os métodos propostos por <xref ref-type="bibr" rid="redalyc_20562876002_ref11">Peña e Prieto (2001)</xref>, e <xref ref-type="bibr" rid="redalyc_20562876002_ref6">Filzmoser, Maronna e Werner (2008)</xref> para detectar dados atípicos em empresas argentinas que cotizam suas ações no Mercado de Valores. A heterogeneidade significativa entre observações pode ser uma consequência da presença de dados atípicos. A detecção de dados atípicos é importante na análise estatística por seu efeito na distorção das medidas descritivas e nos estimadores dos parâmetros. Existem distintos métodos multivariados para detectar dados atípicos, tais como os métodos baseados na distância ou os métodos de busca de projeções.</p>
</trans-abstract>
<kwd-group xml:lang="en">
<title>Keywords</title>
<kwd>outliers</kwd>
<kwd>projection pursuit</kwd>
<kwd>Kurtosis</kwd>
<kwd>Argentinian companies</kwd>
</kwd-group>
<kwd-group xml:lang="es">
<title>Palabras clave</title>
<kwd>datos atípicos</kwd>
<kwd>búsqueda de proyecciones</kwd>
<kwd>curtosis</kwd>
<kwd>empresas argentinas</kwd>
</kwd-group>
<kwd-group xml:lang="pt">
<title>Palavras-chave</title>
<kwd>outliers</kwd>
<kwd>busca de projeções</kwd>
<kwd>curtose</kwd>
<kwd>empresas argentinas</kwd>
</kwd-group>
<counts>
<fig-count count="6"/>
<table-count count="6"/>
<equation-count count="14"/>
<ref-count count="18"/>
</counts>
<custom-meta-group>
<custom-meta>
<meta-name>Cited as</meta-name>
<meta-value>Stimolo, M. I., &amp; Ortiz, P. A. (2020). Projection pursuit algorithms to detect outliers. <italic>Cuadernos de Administración</italic>, 33. <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.11144/Javeriana.cao33.ppado">https://doi.org/10.11144/Javeriana.cao33.ppado</ext-link>
</meta-value>
</custom-meta>
</custom-meta-group>
</article-meta>
</front>
<body>
<sec sec-type="intro">
<title>
<bold>Introduction</bold>
</title>
<p>Databases often show outliers observations, which present a different behavior from the majority. It is important to detect these observations since they affect the data analysis in different ways. In this respect, <xref ref-type="bibr" rid="redalyc_20562876002_ref18">Uriel Jiménez and Aldás Manzano (2005)</xref> point out:</p>
<p>
<list list-type="roman-lower">
<list-item>
<p>They could mask the data pattern and distort the results, so the conclusions would be completely different without their presence.</p>
</list-item>
<list-item>
<p>They could affect the normality condition that is necessary in many multivariate techniques.</p>
</list-item>
</list>
</p>
<p>Outliers have different causes:</p>
<p>
<list list-type="bullet">
<list-item>
<p>Measurement errors, collection or transcription.</p>
</list-item>
<list-item>
<p>Intentional errors of response from the respondents.</p>
</list-item>
<list-item>
<p>Sampling errors: the incorporation of sample statistical units from different populations to the target population.</p>
</list-item>
<list-item>
<p>Intrinsic heterogeneity: the observed elements belong to the target population, but the inherent variability of the samples differs from the rest in their choices, attitudes or behavior.</p>
</list-item>
</list>
</p>
<p>Sometimes the detection of outliers is the first step in statistical analysis. Other times, the outliers need to be removed or downweighted; different causes motivate different procedures. The detection of outliers depends on the type of error (or cause) in the data. In the case of errors in measurement or data entry to the base, it is relatively simple to correct them and it is convenient to eliminate the obvious mistakes. However, a controversial question is: What should we do when the outliers derive from the intrinsic heterogeneity of the data?</p>
<p>In this paper we discuss some multivariate methods for detecting outliers. In multivariate methods, there exist two approaches to identify outliers: those based on the distances of the observations in the data center and those projecting the original data. The projection pursuit methods easily identify atypical observations, and they have the advantage that it is not necessary to know the data distribution. However, the disadvantage of the projection pursuit methods is that there are high requirements in terms of computational load, which increments significantly when there is an increase in the variables considered.</p>
<p>The document is organized as follows. The first section describes two algorithms used to detect outliers (<xref ref-type="bibr" rid="redalyc_20562876002_ref6">Filzmoser et al., 2008</xref>; <xref ref-type="bibr" rid="redalyc_20562876002_ref11">Peña &amp; Prieto, 2001</xref>) based on projection pursuit. The second section compares the methods developed applying them in a group of Argentine companies that listed their shares publicly in the period 2004-2012. In final section, we developed the main conclusions, and we describe research limitations and future research work related to this topic.</p>
</sec>
<sec>
<title>
<bold>Algorithms’ Description</bold>
</title>
<p>Mahalanobis’ distance from the center of the data is the classical multivariate way of identifying outliers observations far from most others.</p>
<p>Let <bold>x</bold>
<sub>1</sub>, <bold>x</bold>
<sub>2</sub>...<bold>x</bold>
<sub>n</sub> be a random sample from a normal multivariate distribution <bold>
<italic>N<sub>p</sub>
</italic>
</bold>
<bold>(µ,Σ)</bold> where µ is the multivariate location vector and Σ the <italic>p</italic> x <italic>p</italic> covariance matrix. The distance between the <italic>i</italic>-th observation x<sub>i</sub> and the location µ, weighted by the covariance Σ is using to detect if the observation x<sub>i</sub> is an outlier, it is (<xref ref-type="disp-formula" rid="e1">1</xref>).</p>
<p>
<disp-formula id="e1">
<label>(1)</label>
<graphic xlink:href="20562876002_ee2.png" position="anchor" orientation="portrait"/>
</disp-formula>
</p>
<p>The square Mahalanobis distance  has Chi-square distribution with p degrees of freedom and the observation . is consider outlier if <inline-graphic xlink:href="20562876002_gi3.png"/> (by setting the squared Mahalanobis distance equal to certain quantile of Chi-squared distribution it is possible define ellipsoids having the same Mahalanobis distance from the data centre).</p>
<p>When both µ and Σ are unknown it be used estimators,<italic> i.e.</italic> the vector mean <inline-graphic xlink:href="20562876002_gi6.png"/> and sample covariance matrix <bold>S</bold>, to estimate the Mahalanobis distance <inline-graphic xlink:href="20562876002_gi7.png"/>,see (<xref ref-type="disp-formula" rid="e3">2</xref>):</p>
<p>
<disp-formula id="e3">
<label>(2)</label>
<graphic xlink:href="20562876002_ee14.png" position="anchor" orientation="portrait"/>
</disp-formula>
</p>
<p>The vector mean and the covariance matrix are affected by outliers, besides the Mahalanobis distance relies on the assumption of normality. Therefore, it is affected by outliers and it does not allow identify sets of outliers (<xref ref-type="bibr" rid="redalyc_20562876002_ref10">Peña, 2002</xref>).</p>
<p>An alternative approach is to use robust location and scale estimators, measures resistance against the influence of outlying observations. <xref ref-type="bibr" rid="redalyc_20562876002_ref8">Maronna (1976)</xref> studied affinely equivariant M-estimators for covariance matrices, and <xref ref-type="bibr" rid="redalyc_20562876002_ref4">Campbell (1980)</xref> proposed using the Mahalanobis distance computed using M-estimators for the mean and covariance matrix.</p>
<p>Nevertheless, the distance method approach presents two difficulties: (i) obtaining a reliable robust location estimator, and (ii) determining and classifying the outliers. It is important to find metric separating outliers from regular observations. <xref ref-type="bibr" rid="redalyc_20562876002_ref14">Rousseeuw (1985)</xref> proposed other distance-based algorithm that computes the ellipsoid with the smallest volume or with the smallest covariance determinant that would include at least half of the data points (minimum covariance determinant, MCD). Because these procedures are based on the minimization of certain nonconvex and nondifferentiable criteria, these estimators are computed by resampling.</p>
<p>
<xref ref-type="bibr" rid="redalyc_20562876002_ref16">Rousseeuw and Driessen (1999)</xref> get faster algorithm splitting the problem into smaller subproblems (FAST-MCD algorithm).</p>
<p>Others outlier-detection procedures are basing on projections to identify outliers. The underlying motive of these methods is to find suitable projection of the data in which the outliers are readily apparent and can thus be downweighted to yield a robust estimator.</p>
<p>
<xref ref-type="bibr" rid="redalyc_20562876002_ref7">Gnanadesikan and Kettenring (1972)</xref> proposed to search for outliers in the direction of the first principal components: the direction of the maximum variability of the data. Although this method provides a correct solution when the outliers are located close to the directions of the principal components, it may fail to identify outliers in the general case.</p>
<p>From that point of view, <xref ref-type="bibr" rid="redalyc_20562876002_ref17">Stahel (1981)</xref> carry on projection pursuit on the data using random directions. They proposed to compute the weight for the robust estimators from the projections of the data onto some directions. These directions were chosen maximizing distances based on robust location and scale estimators, and the optimal values for the distances could also be used to weight each point in the computation of the robust covariance matrix. <xref ref-type="bibr" rid="redalyc_20562876002_ref17">Stahel (1981)</xref> developed a computer approximation based on direction from random subsamples.</p>
<p>
<xref ref-type="bibr" rid="redalyc_20562876002_ref15">Rousseeuw (1993)</xref> proposed selecting . observations from the original sample and computing the orthogonal direction of the hyperplane defined by these observations. The maximum over this finite set of directions is used as an approximation to the exact solution.</p>
<p>The disadvantage of the projection pursuit methods is the form to increase the computational burden with the variable number.</p>
<p>
<xref ref-type="bibr" rid="redalyc_20562876002_ref11">Peña and Prieto (2001)</xref> proposed an improve examining only the set of 2. directions that maximize or minimize the kurtosis. A small number of outliers would cause heavy tails and lead to a larger kurtosis coefficient while a large number of outliers would start introducing bimodality and decrease the kurtosis coefficient (<xref ref-type="bibr" rid="redalyc_20562876002_ref6">Filzmoser et al., 2008</xref>).</p>
<p>Projection pursuit methods have a computational time that increase very rapidly in higher dimensions.</p>
<p>Principal components are those orthogonality directions that maximize the variance along each component. It is well-known the method of dimension reduction that seems intuitive to identifying outliers since outliers increase the variance along their respective directions. The outliers appear more visible in principal components space, at least in some direction of maximum variance, than the original data space. Principal components select a small quantity of highly informative components, discarding those are not contribute significant additional information. In this way, the dataset become more computationally tractable without losing a lot of information. <xref ref-type="bibr" rid="redalyc_20562876002_ref6">Filzmoser et al. (2008)</xref> proposed a method based on the principal components properties useful to detect outliers in high dimensions.</p>
<p>In the following sections, we describe the <xref ref-type="bibr" rid="redalyc_20562876002_ref11">Peña and Prieto (2001)</xref> and <xref ref-type="bibr" rid="redalyc_20562876002_ref6">Filzmoser et al. (2008)</xref> methods.</p>
<sec>
<title>
<bold>
<italic>Kurtosis method (Kurt) (Peña &amp; Prieto, 2001)</italic>
</bold>
</title>
<p>Given a sample (x<sub>1</sub>,....x<sub>n</sub>) of a <italic>p</italic>-dimensional random variable <italic>X</italic> the algorithm consists in projecting each observation onto a set of <italic>2p</italic> directions, which are obtained as the solutions of <italic>2p </italic>simple smooth optimization problems.</p>
<p>1) The original data are rescaled and centred, see (<xref ref-type="disp-formula" rid="e4">3</xref>).</p>
<p>
<disp-formula id="e4">
<label>(3)</label>
<graphic xlink:href="20562876002_ee17.png" position="anchor" orientation="portrait"/>
</disp-formula>
</p>
<p>Where <inline-graphic xlink:href="20562876002_gi10.png"/> and <italic>S</italic>
<sub>x</sub> are the mean and sample variance, respectively.</p>
<p>2) Set <inline-graphic xlink:href="20562876002_gi13.png"/> the iteration index <italic>j</italic> compute <italic>p</italic> orthogonal directions that maximize the kurtosis coefficient obtained as a solution of the following optimization problem (<xref ref-type="disp-formula" rid="e5">4</xref>).</p>
<p>
<disp-formula id="e5">
<label>(4)</label>
<graphic xlink:href="20562876002_ee18.png" position="anchor" orientation="portrait"/>
</disp-formula>
</p>
<p>3) Project sample points onto a lower dimension subspace<italic> (p – j)</italic>, orthogonal to the direction <italic>d<sub>j</sub>.</italic> Define (<xref ref-type="disp-formula" rid="e6">5</xref>):</p>
<p>
<disp-formula id="e6">
<label>(5)</label>
<graphic xlink:href="20562876002_ee20.png" position="anchor" orientation="portrait"/>
</disp-formula>
</p>
<p>Where <italic>e<sub>j</sub>
</italic> denotes the first unit vector, <bold>I</bold> the identity matrix and <italic>Q<sub>j</sub>
</italic>is orthogonal. Compute the new values in (<xref ref-type="disp-formula" rid="e7">6</xref>),</p>
<p>
<disp-formula id="e7">
<label>(6)</label>
<graphic xlink:href="20562876002_ee21.png" position="anchor" orientation="portrait"/>
</disp-formula>
</p>
<p>Where <inline-graphic xlink:href="20562876002_gi14.png"/> is the first component of <inline-graphic xlink:href="20562876002_gi15.png"/> which satisfies <inline-graphic xlink:href="20562876002_gi16.png"/> (the univariate projection values) and <inline-graphic xlink:href="20562876002_gi18.png"/>corresponds to the remaining <italic>p- j</italic> components of <inline-graphic xlink:href="20562876002_gi17.png"/>.</p>
<p>4) Compute <italic>j’ = j+1</italic> and repeat <italic>(<xref ref-type="disp-formula" rid="e3">2</xref>) </italic>y <italic>(<xref ref-type="disp-formula" rid="e4">3</xref>)</italic> up to have p directions: <italic>d<sub>1</sub>, d<sub>2</sub>, …, d<sub>p</sub>
</italic>.</p>
<p>5) Repeat (<xref ref-type="disp-formula" rid="e3">2</xref>) and (<xref ref-type="disp-formula" rid="e4">3</xref>) computing <italic>p</italic> orthogonal directions that minimize the kurtosis coefficient (idem step 2) obtaining<italic> d<sub>p+1</sub>, d<sub>p+2</sub>, d<sub>p+1</sub>, ...,d<sub>2p</sub>
</italic>.</p>
<p>6) To determine an outlier in any one of the 2<italic>p</italic> directions <inline-graphic xlink:href="20562876002_gi19.png"/>, we compute a univariate measure rescaling  with the median (<italic>med</italic>) and the median absolute deviation (MAD), see (<xref ref-type="disp-formula" rid="e8">7</xref>).</p>
<p>
<disp-formula id="e8">
<label>(7)</label>
<graphic xlink:href="20562876002_ee22.png" position="anchor" orientation="portrait"/>
</disp-formula>
</p>
<p>Where <italic>β</italic> is a cut-off chosen to ensure a reasonable level of Type I error and depend on the sample space dimension <italic>p</italic>. See <xref ref-type="table" rid="gt1">Table 1</xref>.</p>
<p>
<table-wrap id="gt1">
<label>Table 1</label>
<graphic xlink:href="20562876002_gt2.png" position="anchor" orientation="portrait"/>
<attrib>Source: Own elaboration.</attrib>
</table-wrap>
</p>
<p>7) Define a new sample composed of all observations i if <inline-graphic xlink:href="20562876002_gi20.png"/>, and the procedure is applied again to the reduced sample. This is repeated until either no additional observations satisfy <inline-graphic xlink:href="20562876002_gi21.png"/> or the number or remaining would be less than [<italic>(n+p+1)/2</italic>].</p>
<p>8) Let U denote the set of all observations not labelled as outliers and computed the mean vector <inline-graphic xlink:href="20562876002_gi22.png"/>, the covariance matrix  <inline-graphic xlink:href="20562876002_gi23.png"/>; and Mahalanobis distance: <inline-graphic xlink:href="20562876002_gi24.png"/>.</p>
<p>Those observations <inline-graphic xlink:href="20562876002_gi26.png"/> such that <inline-graphic xlink:href="20562876002_gi27.png"/> are considered not be outliers and included in <italic>U</italic>. The procedure is repeated until <italic>U</italic> becomes the set of all observations.</p>
<p>The Kurtosis method is affine equivariant. <xref ref-type="bibr" rid="redalyc_20562876002_ref11">Peña and Prieto (2001)</xref> conclude after several computational experiments to study the practical behaviour of the proposed procedure, that it shows a satisfactory empirical performance, especially for large sample space dimensions and concentrated contaminations.</p>
<p>However other authors discussed the Kurtosis method, they argue important points. The method works well in the presence of scattered outliers or multiple clusters of outliers.</p>
<p>For those cases in which the shape of the contamination is similar to that of the original data, the method can be supplement with other general methods (an alternative approach is to use clustering methods to supplement the general-purpose robust methods). The greatest chance of success comes from use of multiple methods, at least one of which is a general-purpose method such as FAST-MCD and MULTOUT, and at least one of which is meant for clustered outliers, such as Kurt method.</p>
<p>However, several key aspects of the Kurt algorithm proposed are criticized. The standardization in Step 1 uses the classical mean and covariance matrix. It is well known that these estimators are extremely sensitive to outliers, which often leads to labelling outliers as good data points and good points as outliers. The authors answer:</p>
<p>
<disp-quote>
<p>This problem cannot appear in our method. First, note that the algorithm we propose is affine equivariant, independently of the initial standardization. The kurtosis coefficient is invariant to translations and scaling of the data and a rotation will not affect the maximizes or minimisers. Moreover, we have tried to be careful when defining the operations to generate the successive directions, as well as in the choice of an initial direction for the optimization problems. (<xref ref-type="bibr" rid="redalyc_20562876002_ref11">Peña &amp; Prieto, 2001, p. 307</xref>)</p>
</disp-quote>
</p>
<p>Maximizing and minimizing the kurtosis in Steps 2 and 3. The authors indicates that the kurtosis is maximal (respectively, minimal) in the direction of the outliers when the contamination is concentrated and small (respectively, large). However, this is not always true for an intermediate level of contamination. The authors answer:</p>
<p>
<disp-quote>
<p>The behaviour of the kurtosis coefficient is particularly useful to reveal the presence of outliers in the cases of small and large contaminations, and this agrees with the standard interpretation of the kurtosis coefficient as measuring both the presence of outliers and the bimodality of the distribution. What is remarkable is that in intermediate cases with a=0.3 the procedure does not break down completely and its performance improves with the sample size. (<xref ref-type="bibr" rid="redalyc_20562876002_ref11">Peña &amp; Prieto, 2001, p. 307</xref>)</p>
</disp-quote>
</p>
<p>Taking 2<italic>p</italic> orthogonal directions in Steps 2 and 3, the chosen directions are still rather arbitrary. Which is the reason to consider the only first <italic>p</italic> directions that maximize the kurtosis and then <italic>p</italic> directions that minimize the kurtosis? Why it not proposed to alternate between directions using a procedure that stops once a significant direction is computed? The authors argue “The algorithm we describe does not make use of this feature, and in this sense it is a simpler one to describe and understand, although it may be more expensive to implement” (<xref ref-type="bibr" rid="redalyc_20562876002_ref11">Peña &amp; Prieto, 2001, p. 307</xref>).</p>
<p>Regarding the choice of orthogonal directions, they reply:</p>
<p>
<disp-quote>
<p>Our motivation to use these orthogonal directions is twofold. On the one hand, we wish the algorithm to be able to identify contamination patterns that have more than one cluster of outliers. The second motivation arises from a property of the kurtosis that implies that in some cases the directions of interest are those orthogonal to the maximization or minimization directions. (<xref ref-type="bibr" rid="redalyc_20562876002_ref11">Peña &amp; Prieto, 2001, p. 307</xref>)</p>
</disp-quote>
</p>
<p>The authors did not explain how they obtained the cut-off values <italic>β</italic>
<sub>p</sub> in <xref ref-type="table" rid="gt2">Table 2</xref> to choice in Step 7. The response:</p>
<p>
<disp-quote>
<p>The results are unfortunately not totally satisfactory; the reason is the large variability in these values in the simulations<sup>[<xref ref-type="fn" rid="fn1">1</xref>]</sup>. This variability has two main effects –it is difficult to find correct values (huge numbers of replications would be required) and for any set of 100 replications there is a high probability that the resulting values will be far from the expected one. Nevertheless, we agree that these values could be estimated with greater detail, although they do not seem to be very significant for the behaviour of the algorithm, except for contaminations very close to the original sample. (<xref ref-type="bibr" rid="redalyc_20562876002_ref11">Peña &amp; Prieto, 2001, p. 308</xref>)</p>
</disp-quote>
</p>
<p>Sequential determination of outliers in Step 8, the mean and covariance matrix of the good data points are computed and used to decide which outliers can still be reclassified as good observations. This procedure is repeated until no more outliers can be reallocated. It has suggested that Step 8 be applied only once. The authors are a little surprised by the criticism of procedures that determine the outliers sequentially. They replied:</p>
<p>
<disp-quote>
<p>The statistical literature is full of examples of very successful sequential procedures and, to indicate just one, <xref ref-type="bibr" rid="redalyc_20562876002_ref12">Peña and Yohai (1999)</xref> presented a sequential procedure for outlier detection in large regression problems that performs much better than other nonsequential procedures. (<xref ref-type="bibr" rid="redalyc_20562876002_ref11">Peña &amp; Prieto, 2001, p. 309</xref>)</p>
</disp-quote>
</p>
</sec>
<sec>
<title>
<bold>
<italic>Method PCOut (<xref ref-type="bibr" rid="redalyc_20562876002_ref6">Filzmoser et al., 2008</xref>)</italic>
</bold>
</title>
<p>This algorithm was designed primarily for computational efficiency at high dimension. It consists in two steps: The first one to detect the location outliers and the second one to detect scatter outliers.</p>
<p>1) Rescale the data <bold>X<sub>(n,p)</sub>
</bold>using the median (<italic>med)</italic> and the median absolute deviation (MAD), see (<xref ref-type="disp-formula" rid="e12">8</xref>):</p>
<p>
<disp-formula id="e12">
<label>(8)</label>
<graphic xlink:href="20562876002_ee26.png" position="anchor" orientation="portrait"/>
</disp-formula>
</p>
<p>Compute the covariance matrix from <bold>X<sup>*</sup>
</bold>.</p>
<p>2) Compute the eigenvalues and eigenvectors from covariance matrix <bold>X<sup>*</sup>
</bold>, a semirobust principal component decomposition, and retain only <italic>p</italic>* eigenvectors whit eigenvalues that represent the 99% of the variance. The matrix of principal components is Z: Z=X*V. where <bold>V</bold> is the matrix of eigenvalues <italic>p* × p*</italic>. Z is rescaled by the median and the MAD as 8), for i-<italic>th</italic> component, see (<xref ref-type="disp-formula" rid="e13">9</xref>):</p>
<p>
<disp-formula id="e13">
<label>(9)</label>
<graphic xlink:href="20562876002_ee28.png" position="anchor" orientation="portrait"/>
</disp-formula>
</p>
<p>
<bold> Z*</bold>, principal components rescaled is stored for the both phases of the algorithm.</p>
<sec>
<title>
<italic>Phase 1: Location outliers.</italic>
</title>
<p>3) Compute a robust kurtosis weights for each component denoted by <italic>w<sub>j</sub>
</italic> in (<xref ref-type="disp-formula" rid="e18">10</xref>).</p>
<p>
<disp-formula id="e18">
<label>(10)</label>
<graphic xlink:href="20562876002_ee34.png" position="anchor" orientation="portrait"/>
</disp-formula>
</p>
<p>
<xref ref-type="bibr" rid="redalyc_20562876002_ref11">Peña and Prieto (2001)</xref> argue in the Kurt method, both small and large values of the kurtosis coefficient can be indicated of outliers. In order to use relative weights is defined  .</p>
<p>To classify the data between outliers and non-outliers we need to determinate a weighted norm from transformated data Z* but it has not chi quadratic distribution. Z* is similar as a robust Mahalanobis distance (RD<sub>i</sub>) (distance from median rescaled by MAD). Therefore, the algorithm used a robust distance transform similar <xref ref-type="bibr" rid="redalyc_20562876002_ref9">Maronna and Zamar (2012)</xref>, that helped the empirical distances di to have the same median to the theoretical distance and bring the former somewhat closer   . See (<xref ref-type="disp-formula" rid="e19">11</xref>):</p>
<p>
<disp-formula id="e19">
<label/>
<graphic xlink:href="20562876002_ee35.png" position="anchor" orientation="portrait"/>
</disp-formula>
</p>
<p>4) To assign weights to each observation and use it as a measure of outlyingness is calculated the translated bi-weight function w<sub>1i</sub> (<xref ref-type="bibr" rid="redalyc_20562876002_ref13">Rocke, 1996</xref>), see (<xref ref-type="disp-formula" rid="e20">12</xref>):</p>
<p>
<disp-formula id="e20">
<label>(12)</label>
<graphic xlink:href="20562876002_ee36.png" position="anchor" orientation="portrait"/>
</disp-formula>
</p>
<p>Where M is equal to 33⅓ quantil of the distances and <italic>c</italic>, see (<xref ref-type="disp-formula" rid="e21">13</xref>):</p>
<p>
<disp-formula id="e21">
<label>(13)</label>
<graphic xlink:href="20562876002_ee37.png" position="anchor" orientation="portrait"/>
</disp-formula>
</p>
</sec>
<sec>
<title>
<italic>Phase 2: Scatter outliers.</italic>
</title>
<p>5) Use the step 2 decomposition and calculate the Euclidean norm for the data in non-weighting principal component space (equivalent to the Mahalanobis distance in the original data but faster to compute). After use the <xref ref-type="bibr" rid="redalyc_20562876002_ref9">Maronna and Zamar (2012)</xref> transformation, the distances set is going to use at 6.</p>
<p>6) Determine the weights w<sub>2i</sub> to each robust distance with the translated biweight function where c<sup>2</sup> is equal to  , M<sup>2</sup> is equal to    and finally we calculate the final weight, see (1<xref ref-type="disp-formula" rid="e22">4</xref>):</p>
<p>
<disp-formula id="e22">
<label>(14)</label>
<graphic xlink:href="20562876002_ee38.png" position="anchor" orientation="portrait"/>
</disp-formula>
</p>
<p>Where typically the scaling constant s = 0.25. Outliers are then classified as points they have weight w<sub>i</sub> &lt; 0.25. This value implies that if one of the weights equals one the other must be less than 0.0625. If w<sub>1</sub> = w<sub>2</sub>, x is classified as outlier when the common value is less than 0.375.</p>
<p>The computational speed that is the speed t of the computer to process data<sup>[<xref ref-type="fn" rid="fn2">2</xref>]</sup> is an advantage of this algorithm. Using examples an simulated data <xref ref-type="bibr" rid="redalyc_20562876002_ref6">Filzmoser et al. (2008)</xref> infer that PCOut is a competitive outlier detection algorithm regarding detection accuracy as well as computation time.</p>
<p>The comparison with methods in low dimension, using simulated data reveals that PCOut performs well at identifying outliers, with low masked outliers, although it has a higher percent non-outliers that were classified as outliers. It does particularly well for location outliers while Kurt does very poorly, however Kurt does exceptionally well for scatter outliers.</p>
</sec>
</sec>
</sec>
<sec>
<title>
<bold>An empirical application</bold>
</title>
<p>In this paper we applied the detection outliers methods to a sample composed by a set of Argentine companies that quote their shares in the Buenos Aires stock exchange in the period 2004-2012. The database was prepared relying on the data publicly available in the Buenos Aires stock exchange web site (<xref ref-type="bibr" rid="redalyc_20562876002_ref3">Bolsar, s.f.</xref>) including only companies that presented a positive operative ordinary income defined as net sales larger that costs of sales and selling and administrative expenses, and excluding those that belong to the financial and insurance sectors. A total of 744 observations (firms per year) belonging to 111 firms were considered.</p>
<p>The variables used to detect outliers are the following financial reporting indicators to analyse cost-effectiveness and cost behaviour (<xref ref-type="bibr" rid="redalyc_20562876002_ref1">Anderson, Banker, &amp; Janakiraman, 2003</xref>; <xref ref-type="bibr" rid="redalyc_20562876002_ref2">Banker &amp; Byzalov, 2014</xref>).</p>
<p>
<list list-type="bullet">
<list-item>
<p>
<italic>Market to book value.</italic> The ratio of indicates investors’ expectations of future abnormal earnings relative to assets in place. It reflects both the magnitude and persistence of sales growth expectations.</p>
</list-item>
<list-item>
<p>
<italic>Current Assets and non-current Assets</italic>
</p>
</list-item>
<list-item>
<p>
<italic>Operating income.</italic> It is equals all revenue from the property minus all reasonably necessary operating expenses.</p>
</list-item>
<list-item>
<p>
<italic>Net Revenues.</italic> A company’s revenue net of discounts and returns.</p>
</list-item>
<list-item>
<p>
<italic>Net Revenues annual variation coeff.</italic> The annual change in a company's net income.</p>
</list-item>
<list-item>
<p>
<italic>Selling Administrative expenses</italic>
<sup>[<xref ref-type="fn" rid="fn3">3</xref>]</sup>. It is the sum of direct, indirect selling expenses and administrative expenses of a company.</p>
</list-item>
<list-item>
<p>
<italic> Selling Administrative expenses. Annual variation coeff</italic>.</p>
</list-item>
</list>
</p>
<p>The <xref ref-type="table" rid="gt2">Table 2</xref> and <xref ref-type="fig" rid="gf1">Figure 1</xref> summarize a preliminary descriptive analysis of original sample. The univariate statistical analysis (descriptive measures and boxplots) shows the skewness and outliers.</p>
<p>
<table-wrap id="gt2">
<label>Table 2</label>
<graphic xlink:href="20562876002_gt3.png" position="anchor" orientation="portrait"/>
<attrib>Source: Own elaboration.</attrib>
</table-wrap>
</p>
<p>
<fig id="gf1">
<label>
<bold>Figure 1</bold>
</label>
<caption>
<title>Boxplot of original variables</title>
</caption>
<alt-text>Figure 1 Boxplot of original variables</alt-text>
<graphic xlink:href="20562876002_gf2.png" position="anchor" orientation="portrait"/>
<attrib>Source: Own elaboration.</attrib>
</fig>
</p>
<p>We detected outliers using Mahalanobis distance and the both projection pursuit methods presented by <xref ref-type="bibr" rid="redalyc_20562876002_ref11">Peña and Prieto (2001)</xref> and <xref ref-type="bibr" rid="redalyc_20562876002_ref6">Filzmoser et al. (2008)</xref> denominated Kurt and PCOut, respectively. In this work we compare the results defining finally as outliers all the outliers defined using both methods proposed.</p>
<p>We used the Matlab for the Kurt method and the R package <italic>mvoutlier </italic>for the PCOut method (<xref ref-type="bibr" rid="redalyc_20562876002_ref5">Filzmoser, 2015</xref>).</p>
<p>The <xref ref-type="table" rid="gt3">Table 3</xref> shows the outliers detected by differences algorithms. There were detected 48 outliers (6.5% of the data) by the distance methods (Mahalanobis). By using projection pursuit methods a large outliers were detected. 212 (28.5%) by Kurt method and 203 (27.3%) by PCOut of outliers. The algorithms proposed detected a similar quantity. Nevertheless the 69.8% of Kurt outliers were PCOut outliers (see <xref ref-type="table" rid="gt4">Table 4</xref>).</p>
<p>
<table-wrap id="gt3">
<label>Table 3</label>
<caption>
<title>Outliers detected</title>
</caption>
<alt-text>Table 3 Outliers detected</alt-text>
<graphic xlink:href="20562876002_gt4.png" position="anchor" orientation="portrait"/>
<attrib>Source: Own elaboration.</attrib>
</table-wrap>
</p>
<p>
<table-wrap id="gt4">
<label>Table 4</label>
<caption>
<title>Outliers detected comparing methods</title>
</caption>
<alt-text>Table 4 Outliers detected comparing methods</alt-text>
<graphic xlink:href="20562876002_gt5.png" position="anchor" orientation="portrait"/>
<attrib>Source: Own elaboration.</attrib>
</table-wrap>
</p>
<p>The biplots (<xref ref-type="fig" rid="gf2">Figure 2</xref>) shows the detected outliers using different methods in the space of the first and second principal component.</p>
<p>
<fig id="gf2">
<label>
<bold>Figure 2</bold>
</label>
<caption>
<title>Biplot showing outliers identified by Methods</title>
</caption>
<alt-text>Figure 2 Biplot showing outliers identified by Methods</alt-text>
<graphic xlink:href="20562876002_gf3.png" position="anchor" orientation="portrait"/>
<attrib>Source: Own elaboration.</attrib>
</fig>
</p>
<p>For each method it has been calculated the first and second principal components of the non-outliers data and graphic it on biplots which shows different structures of the data (See <xref ref-type="fig" rid="gf3">figure 3</xref>).</p>
<p>
<fig id="gf3">
<label>
<bold>Figure 3</bold>
</label>
<caption>
<title>Biplot without outliers by Method</title>
</caption>
<alt-text>Figure 3 Biplot without outliers by Method</alt-text>
<graphic xlink:href="20562876002_gf4.png" position="anchor" orientation="portrait"/>
<attrib>Source: Own elaboration.</attrib>
</fig>
</p>
<p>
<xref ref-type="table" rid="gt5">Table 5</xref> shows descriptive measures of the variables considered in non-outliers database, the mean vector difference, the multivariate variability (total and generalized variance) by method. Multivariate variances are significantly lower in all cases because of excluding outliers distorting the estimates.</p>
<p>
<table-wrap id="gt5">
<label>Table 5</label>
<caption>
<title>Multivariate descriptive without outliers</title>
</caption>
<alt-text>Table 5 Multivariate descriptive without outliers</alt-text>
<graphic xlink:href="20562876002_gt6.png" position="anchor" orientation="portrait"/>
<attrib>Source: Own elaboration.</attrib>
</table-wrap>
</p>
<p>A plausible criterion for determining multivariate outliers is to use different methods and consider as those who are simultaneously identified as atypical by them. Particularly, in this application 19.9% of the data were identified as outliers at Kurt and PCOut methods.</p>
<p>The different results and different performance of each method leads us to consider all the outliers detected for these methods. We aggregated the outliers identified by all the methods taking advantage of their performance. I this empirical application, we detected 225 outliers (30.24%). <xref ref-type="fig" rid="gf4">Figure 4</xref> shows in the space of the two first principal components of the sample all the outliers detected for the algorithms.</p>
<p>
<fig id="gf4">
<label>
<bold>Figure 4</bold>
</label>
<caption>
<title>Outliers detected by all algorithms</title>
</caption>
<alt-text>Figure 4 Outliers detected by all algorithms</alt-text>
<graphic xlink:href="20562876002_gf5.png" position="anchor" orientation="portrait"/>
<attrib>Source: Own elaboration.</attrib>
</fig>
</p>
<p>The biplot without the outliers detected by all the methods (<xref ref-type="fig" rid="gf5">Figure 5</xref>) shows a better performance. Besides we pointed with a circle a set of data that we could study especially because they presented a different behaviour.</p>
<p>
<fig id="gf5">
<label>
<bold>Figure 5</bold>
</label>
<caption>
<title>Biplot without all the Outliers detected</title>
</caption>
<alt-text>Figure 5 Biplot without all the Outliers detected</alt-text>
<graphic xlink:href="20562876002_gf6.png" position="anchor" orientation="portrait"/>
<attrib>Source: Own elaboration.</attrib>
</fig>
</p>
<p>
<xref ref-type="table" rid="gt6">Table 6</xref> and <xref ref-type="fig" rid="gf6">Figure 6</xref> show a descriptive analysis of data without all the Outliers detected. The variables exhibited less skewness and heterogeneity, resulting a data sample more homogeneous.</p>
<p>
<table-wrap id="gt6">
<label>Table 6</label>
<caption>
<title>Multivariate descriptive without all the Outliers detected</title>
</caption>
<alt-text>Table 6 Multivariate descriptive without all the Outliers detected</alt-text>
<graphic xlink:href="20562876002_gt7.png" position="anchor" orientation="portrait"/>
<attrib>Source: Own elaboration.</attrib>
</table-wrap>
</p>
<p>
<fig id="gf6">
<label>
<bold>Figure 6</bold>
</label>
<caption>
<title>Boxplott without all the Outliers detected</title>
</caption>
<alt-text>Figure 6 Boxplott without all the Outliers detected</alt-text>
<graphic xlink:href="20562876002_gf7.png" position="anchor" orientation="portrait"/>
<attrib>Source: Own elaboration.</attrib>
</fig>
</p>
</sec>
<sec sec-type="conclusions">
<title>
<bold>Conclusions</bold>
</title>
<p>Outliers distort the results and mask the real data structure, so to detect them is an important task in the multivariate data analysis.</p>
<p>This work presented two pursuit algorithms to detect outliers, they are an example of the different algorithms available. But is not possible to point one algorithm as the better, it depends of the data sample and the algorithms could show similar performance. Each one method has some disadvantage, so we proposed aggregate the outliers detected by different methods (in this paper Kurt and PCOut methods) and use their different performance to improve the outliers detection. Specially the projection pursuit methods that they only search the useful projections, they are not affecting by non-normality and can be widely applied in diverse data situations.</p>
<p>A multivariate outliers detection is important for thorough data analysis, however, the researchers have to decide to exclude the outliers from further analysis or apply robust procedures to reduce the impact of them.</p>
</sec>
</body>
<back>
<ref-list>
<title>
<bold>References</bold>
</title>
<ref id="redalyc_20562876002_ref1">
<mixed-citation>Anderson, M., Banker, R., &amp; Janakiraman, S. (2003). Are selling, general, and administrative costs “sticky”? <italic>Journal of Accounting Research, 41</italic>(1), 47-63. <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1111/1475-679X.00095">https://doi.org/10.1111/1475-679X.00095</ext-link>
</mixed-citation>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Anderson</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Banker</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Janakiraman</surname>
<given-names>S.</given-names>
</name>
</person-group>
<article-title>Are selling, general, and administrative costs “sticky”?</article-title>
<source>Journal of Accounting Research</source>
<year>2003</year>
<volume>41</volume>
<issue>7</issue>
<fpage>47</fpage>
<lpage>63</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1111/1475-679X.00095">https://doi.org/10.1111/1475-679X.00095</ext-link>
</comment>
<pub-id pub-id-type="doi">10.1111/1475-679X.00095</pub-id>
</element-citation>
</ref>
<ref id="redalyc_20562876002_ref2">
<mixed-citation>Banker, R., &amp; Byzalov, D. (2014). Asymmetric cost behavior. <italic>Journal of Management Accounting Research, 26</italic>(2), 43-79. <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.2308/jmar-50846">https://doi.org/10.2308/jmar-50846</ext-link>
</mixed-citation>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Banker</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Byzalov</surname>
<given-names>D.</given-names>
</name>
</person-group>
<article-title>Asymmetric cost behavior</article-title>
<source>Journal of Management Accounting Research</source>
<year>2014</year>
<volume>26</volume>
<issue>2</issue>
<fpage>43</fpage>
<lpage>79</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.2308/jmar-50846">https://doi.org/10.2308/jmar-50846</ext-link>
</comment>
<pub-id pub-id-type="doi">10.2308/jmar-50846</pub-id>
</element-citation>
</ref>
<ref id="redalyc_20562876002_ref3">
<mixed-citation>Bolsar (s.f). Buenos Aires Stock Exchange. <ext-link ext-link-type="uri" xlink:href="https://www.bolsar.com/VistasDL/PaginaPrincipal.aspx">https://www.bolsar.com/VistasDL/PaginaPrincipal.aspx</ext-link>
</mixed-citation>
<element-citation publication-type="webpage">
<person-group person-group-type="author">
<collab>Bolsar</collab>
</person-group>
<source>Buenos Aires Stock Exchange</source>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://www.bolsar.com/VistasDL/PaginaPrincipal.aspx">https://www.bolsar.com/VistasDL/PaginaPrincipal.aspx</ext-link>
</comment>
</element-citation>
</ref>
<ref id="redalyc_20562876002_ref4">
<mixed-citation>Campbell, N. A. (1980). Robust procedures in multivariate analysis I: Robust covariance estimation. <italic>Applied statistics</italic>, 231-237. <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.2307/2346896">https://doi.org/10.2307/2346896</ext-link>
</mixed-citation>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Campbell</surname>
<given-names>N. A.</given-names>
</name>
</person-group>
<article-title>Robust procedures in multivariate analysis I: Robust covariance estimation</article-title>
<source>Applied statistics</source>
<year>1980</year>
<fpage>231</fpage>
<lpage>237</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.2307/2346896">https://doi.org/10.2307/2346896</ext-link>
</comment>
<pub-id pub-id-type="doi">10.2307/2346896</pub-id>
</element-citation>
</ref>
<ref id="redalyc_20562876002_ref5">
<mixed-citation>Filzmoser, P. (2015). Gschwandtner M. mvoutlier: Multivariate outlier detection based on robust methods. R package version 2.0.6.In. Routine available at <ext-link ext-link-type="uri" xlink:href="http://halweb.uc3m.es/esp/Personal/personas/fjp/research.htm">http://halweb.uc3m.es/esp/Personal/personas/fjp/research.htm</ext-link>
</mixed-citation>
<element-citation publication-type="webpage">
<person-group person-group-type="author">
<name>
<surname>Filzmoser</surname>
<given-names>P.</given-names>
</name>
</person-group>
<article-title>Gschwandtner M. mvoutlier: Multivariate outlier detection based on robust methods. R package version 2.0.6</article-title>
<source>Routine</source>
<year>2015</year>
<comment>
<ext-link ext-link-type="uri" xlink:href="http://halweb.uc3m.es/esp/Personal/personas/fjp/research.htm">http://halweb.uc3m.es/esp/Personal/personas/fjp/research.htm</ext-link>
</comment>
</element-citation>
</ref>
<ref id="redalyc_20562876002_ref6">
<mixed-citation>Filzmoser, P., Maronna, R., &amp; Werner, M. (2008). Outlier identification in high dimensions. <italic>Computational Statistics &amp; Data Analysis, 52</italic>(3), 1694-1711. <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1016/j.csda.2007.05.018">https://doi.org/10.1016/j.csda.2007.05.018</ext-link>
</mixed-citation>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Filzmoser</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Maronna</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Werner</surname>
<given-names>M.</given-names>
</name>
</person-group>
<article-title>Outlier identification in high dimensions</article-title>
<source>Computational Statistics &amp; Data Analysis</source>
<year>2008</year>
<volume>52</volume>
<issue>3</issue>
<fpage>1694</fpage>
<lpage>1711</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1016/j.csda.2007.05.018">https://doi.org/10.1016/j.csda.2007.05.018</ext-link>
</comment>
<pub-id pub-id-type="doi">10.1016/j.csda.2007.05.018</pub-id>
</element-citation>
</ref>
<ref id="redalyc_20562876002_ref7">
<mixed-citation>Gnanadesikan, R., &amp; Kettenring, J. (1972). Robust estimates, residuals, and outlier detection with multiresponse data. <italic>Biometrics</italic>, 81-124. <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.2307/2528963">https://doi.org/10.2307/2528963</ext-link>
</mixed-citation>
<element-citation publication-type="webpage">
<person-group person-group-type="author">
<name>
<surname>Gnanadesikan</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Kettenring</surname>
<given-names>J.</given-names>
</name>
</person-group>
<article-title>Robust estimates, residuals, and outlier detection with multiresponse data</article-title>
<source>Biometrics</source>
<year>1972</year>
<fpage>81</fpage>
<lpage>124</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.2307/2528963">https://doi.org/10.2307/2528963</ext-link>
</comment>
<pub-id pub-id-type="doi">10.2307/2528963</pub-id>
</element-citation>
</ref>
<ref id="redalyc_20562876002_ref8">
<mixed-citation>Maronna, R. A. (1976). Robust M-estimators of multivariate location and scatter. <italic>The Annals of Statistics</italic>, 51-67. <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1214/aos/1176343347">https://doi.org/10.1214/aos/1176343347</ext-link>
</mixed-citation>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Maronna</surname>
<given-names>R. A.</given-names>
</name>
</person-group>
<article-title>Robust M-estimators of multivariate location and scatter</article-title>
<source>The Annals of Statistics</source>
<year>1976</year>
<fpage>51</fpage>
<lpage>67</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1214/aos/1176343347">https://doi.org/10.1214/aos/1176343347</ext-link>
</comment>
<pub-id pub-id-type="doi">10.1214/aos/1176343347</pub-id>
</element-citation>
</ref>
<ref id="redalyc_20562876002_ref9">
<mixed-citation>Maronna, R. A., &amp; Zamar, R. H. (2012). Robust estimates of location and dispersion for high-dimensional datasets. <italic>Technometrics</italic>. <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1198/004017002188618509">https://doi.org/10.1198/004017002188618509</ext-link>
</mixed-citation>
<element-citation publication-type="webpage">
<person-group person-group-type="author">
<name>
<surname>Maronna</surname>
<given-names>R. A.</given-names>
</name>
<name>
<surname>Zamar</surname>
<given-names>R. H.</given-names>
</name>
</person-group>
<article-title>Robust estimates of location and dispersion for high-dimensional datasets</article-title>
<source>Technometrics</source>
<year>2012</year>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1198/004017002188618509">https://doi.org/10.1198/004017002188618509</ext-link>
</comment>
<pub-id pub-id-type="doi">10.1198/004017002188618509</pub-id>
</element-citation>
</ref>
<ref id="redalyc_20562876002_ref10">
<mixed-citation>Peña, D. (2002). <italic>Análisis de datos multivariantes</italic>, vol. 24. Madrid: McGraw-Hill.</mixed-citation>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Peña</surname>
<given-names>D.</given-names>
</name>
</person-group>
<source>Análisis de datos multivariantes vol</source>
<year>2002</year>
<volume>24</volume>
<publisher-loc>Madrid</publisher-loc>
<publisher-name>McGraw-Hill</publisher-name>
</element-citation>
</ref>
<ref id="redalyc_20562876002_ref11">
<mixed-citation>Peña, D., &amp; Prieto, F. J. (2001). Multivariate outlier detection and robust covariance matrix estimation. <italic>Technometrics, 43</italic>(3), 286-310. <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1198/004017001316975899">https://doi.org/10.1198/004017001316975899</ext-link>
</mixed-citation>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Peña</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Prieto</surname>
<given-names>F. J.</given-names>
</name>
</person-group>
<article-title>Multivariate outlier detection and robust covariance matrix estimation</article-title>
<source>Technometrics</source>
<year>2001</year>
<volume>43</volume>
<issue>3</issue>
<fpage>286</fpage>
<lpage>310</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1198/004017001316975899">https://doi.org/10.1198/004017001316975899</ext-link>
</comment>
<pub-id pub-id-type="doi">10.1198/004017001316975899</pub-id>
</element-citation>
</ref>
<ref id="redalyc_20562876002_ref12">
<mixed-citation>Peña, D., &amp; Yohai, V. (1999). A fast procedure for outlier diagnostics in large regression problems. <italic>Journal of the American Statistical Association, 94</italic>(446), 434-445. <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1080/01621459.1999.10474138">https://doi.org/10.1080/01621459.1999.10474138</ext-link>
</mixed-citation>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Peña</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Yohai</surname>
<given-names>V.</given-names>
</name>
</person-group>
<article-title>A fast procedure for outlier diagnostics in large regression problems</article-title>
<source>Journal of the American Statistical Association</source>
<year>1999</year>
<volume>94</volume>
<issue>446</issue>
<fpage>434</fpage>
<lpage>445</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1080/01621459.1999.10474138">https://doi.org/10.1080/01621459.1999.10474138</ext-link>
</comment>
<pub-id pub-id-type="doi">10.1080/01621459.1999.10474138</pub-id>
</element-citation>
</ref>
<ref id="redalyc_20562876002_ref13">
<mixed-citation>Rocke, D. M. (1996). Robustness properties of S-estimators of multivariate location and shape in high dimension. <italic>The Annals of statistics</italic>, 1327-1345. <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1214/aos/1032526972">https://doi.org/10.1214/aos/1032526972</ext-link>
</mixed-citation>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Rocke</surname>
<given-names>D. M.</given-names>
</name>
</person-group>
<article-title>Robustness properties of S-estimators of multivariate location and shape in high dimension</article-title>
<source>The Annals of statistics</source>
<year>1996</year>
<fpage>1327</fpage>
<lpage>1345</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1214/aos/1032526972">https://doi.org/10.1214/aos/1032526972</ext-link>
</comment>
<pub-id pub-id-type="doi">10.1214/aos/1032526972</pub-id>
</element-citation>
</ref>
<ref id="redalyc_20562876002_ref14">
<mixed-citation>Rousseeuw, P. J. (1985). Multivariate estimation with high breakdown point. <italic>Mathematical statistics and applications, 8</italic>, 283-297. <ext-link ext-link-type="uri" xlink:href="https://www.researchgate.net/profile/Peter_Rousseeuw/publication/239666038_Multivariate_Estimation_With_High_Breakdown_Point/links/0deec53137b8cc68aa000000.pdf">https://www.researchgate.net/profile/Peter_Rousseeuw/publication/239666038_Multivariate_Estimation_With_High_Breakdown_Point/links/0deec53137b8cc68aa000000.pdf</ext-link>
</mixed-citation>
<element-citation publication-type="webpage">
<person-group person-group-type="author">
<name>
<surname>Rousseeuw</surname>
<given-names>P. J.</given-names>
</name>
</person-group>
<article-title>Multivariate estimation with high breakdown point</article-title>
<source>Mathematical statistics and applications</source>
<year>1985</year>
<volume>8</volume>
<fpage>283</fpage>
<lpage>297</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://www.researchgate.net/profile/Peter_Rousseeuw/publication/239666038_Multivariate_Estimation_With_High_Breakdown_Point/links/0deec53137b8cc68aa000000.pdf">https://www.researchgate.net/profile/Peter_Rousseeuw/publication/239666038_Multivariate_Estimation_With_High_Breakdown_Point/links/0deec53137b8cc68aa000000.pdf</ext-link>
</comment>
</element-citation>
</ref>
<ref id="redalyc_20562876002_ref15">
<mixed-citation>Rousseeuw, P. J. (1993). A resampling design for computing high-breakdown regression. <italic>Statistics &amp; probability letters, 18</italic>(2), 125-128. <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1016/0167-7152(93)90180-Q">https://doi.org/10.1016/0167-7152(93)90180-Q</ext-link>
</mixed-citation>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Rousseeuw</surname>
<given-names>P. J.</given-names>
</name>
</person-group>
<article-title>A resampling design for computing high-breakdown regression</article-title>
<source>Statistics &amp; probability letters</source>
<year>1993</year>
<volume>18</volume>
<issue>2</issue>
<fpage>125</fpage>
<lpage>128</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1016/0167-7152(93)90180-Q">https://doi.org/10.1016/0167-7152(93)90180-Q</ext-link>
</comment>
<pub-id pub-id-type="doi">10.1016/0167-7152(93)90180-Q</pub-id>
</element-citation>
</ref>
<ref id="redalyc_20562876002_ref16">
<mixed-citation>Rousseeuw, P. J., &amp; Driessen, K. V. (1999). A fast algorithm for the minimum covariance determinant estimator. <italic>Technometrics, 41</italic>(3), 212-223. <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1080/00401706.1999.10485670">https://doi.org/10.1080/00401706.1999.10485670</ext-link>
</mixed-citation>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Rousseeuw</surname>
<given-names>P. J.</given-names>
</name>
<name>
<surname>Driessen</surname>
<given-names>K. V.</given-names>
</name>
</person-group>
<article-title>A fast algorithm for the minimum covariance determinant estimator</article-title>
<source>Technometrics</source>
<year>1999</year>
<volume>41</volume>
<issue>3</issue>
<fpage>212</fpage>
<lpage>223</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1080/00401706.1999.10485670">https://doi.org/10.1080/00401706.1999.10485670</ext-link>
</comment>
<pub-id pub-id-type="doi">10.1080/00401706.1999.10485670</pub-id>
</element-citation>
</ref>
<ref id="redalyc_20562876002_ref17">
<mixed-citation>Stahel, W. A. (1981). <italic>Robuste schätzungen: infinitesimale optimalität und schätzungen von kovarianzmatrizen</italic>: Eidgenössische Technische Hochschule - ETH. Zürich.</mixed-citation>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Stahel</surname>
<given-names>W. A.</given-names>
</name>
</person-group>
<source>Robuste schätzungen: infinitesimale optimalität und schätzungen von kovarianzmatrizen</source>
<year>1981</year>
<publisher-loc>Eidgenössische Technische Hochschule</publisher-loc>
<publisher-name>ETH. Zürich</publisher-name>
</element-citation>
</ref>
<ref id="redalyc_20562876002_ref18">
<mixed-citation>Uriel Jiménez, E., &amp; Aldás Manzano, J. (2005). <italic>Análisis multivariante aplicado: aplicaciones al marketing, investigación de mercados, economía, dirección de empresas y turismo</italic>. Madrid: Thomson.</mixed-citation>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Uriel Jiménez</surname>
<given-names>E.</given-names>
</name>
<name>
<surname>Aldás Manzano</surname>
<given-names>J.</given-names>
</name>
</person-group>
<source>Análisis multivariante aplicado: aplicaciones al marketing investigación de mercados economía dirección de empresas y turismo</source>
<year>2005</year>
<publisher-loc>Madrid</publisher-loc>
<publisher-name>Thomson</publisher-name>
</element-citation>
</ref>
</ref-list>
<fn-group>
<fn id="fn4" fn-type="other">
<label>*</label>
<p>Research paper.</p>
</fn>
<fn id="fn1" fn-type="other">
<label>
<sup>[1]</sup>
</label>
<p>In the papers and the discussion have been conducted a number of computational experiments to study the practical behavior of the proposed algorithm.</p>
</fn>
<fn id="fn2" fn-type="other">
<label>
<sup>[2]</sup>
</label>
<p>It is defined as Millions of Instructions per Second –MIPS–.</p>
</fn>
<fn id="fn3" fn-type="other">
<label>
<sup>[3]</sup>
</label>
<p>The Argentinean GAAP in effect at the time of this study was Resolución Técnica Nº 9 (RT9) of FACPCE. Chapter 5 of RT9 defines as Selling Expenses those related with sales and distribution of products or services rendered by the firm. RT9 Chapter 5 says that Administration Expenses are expenses incurred by the firm in order to carry on its activities but cannot be attributable to any of the following functions: purchasing (procurement), production (operations), selling, research and development, financing of goods or services. The same chapter of RT9 states that net sales (revenues) are to be presented in the income statement and the amount shall exclude returns, discounts and taxes.</p>
</fn>
</fn-group>
</back>
</article>