What's a Renee Masse? (In the first sentence).
| This article is rated B-class on Wikipedia's content assessment scale. It is of interest to the following WikiProjects: | |||||||||||
| |||||||||||
What's a Renee Masse? (In the first sentence).
—Preceding unsigned comment added by 212.183.70.147 (talk) 15:04, 15 October 2007 (UTC)
Adopted orphan redirects for Google: inner fence, outer fence, mild outlier, Extreme outlier
In sans-serif font, 1.5 IQR looks like a division. Patrick 11:15 Dec 23, 2002 (UTC)
I did it that way so I wouldn't have to use * or x or × for multiplication. What would you prefer? dcljr 13:53 Dec 23, 2002 (UTC)
When using multi-letter variables a multiplication sign avoids ambiguity, and in this case coincidentally there was even a little more ambiguity.
Any of the three is fine with me, × is neatest, but more cumbersome to write. - Patrick 14:26 Dec 23, 2002 (UTC)
more info but clearer answers i don't get a thing and my exam is tommorrow!!!!
Is there not a second defintion of outliers, as lying more than two standard deviations away from the mean? Or am I mixing other things up? I am a physicist, and it is a long time since I did "real" statistics... Batmanand | Talk 09:48, 28 September 2006 (UTC)
This is a poor definition of outliers as it changes upon recursion, i.e. the standard deviation is highly dependent on the outliers. Check boxplot for a simple but easy to understand definition that is not distribution dependent.
Isn't an outlier (German: Ausleger, Swedish: utliggare) also a supporting 2nd keel for a canoe or sailing boat that makes it almost a catamaran? Hmm... apparently this is called a outrigger on an outrigger canoe in English. Other languages would use "rig" for things that have sails. --LA2 23:04, 1 August 2007 (UTC)
-No, Ausleger is not outlier, that's a false friend. Outrigger, as you say, is the right term. —Preceding unsigned comment added by 212.183.70.147 (talk) 15:07, 15 October 2007 (UTC)
I'm opposed to the "Mathematical definition" section in the article. Determining whether an observation is an outlier is ultimately a subjective decision, and any definition based on measures such as standard deviation or interquartile range is completely arbitrary.
If this section need be kept, perhaps it could be renamed? The term "mathematical" implies a logical certainty, which doesn't apply in this case. -3mta3 (talk) 11:49, 14 April 2008 (UTC)
I want to know what is the source of the method of using the Interquantile range mentioned in the text?? Is it just a rule of thumb, or does it have a more objective explanation?--Forich (talk) 21:33, 9 June 2008 (UTC)
A more thorough discussion of methods for identifying outliers should be added, for example Rosner's test. See e.g. Barnett, V., and T. Lewis. (1995): Outliers in Statistical Data. Agnerf (talk) 09:09, 24 February 2022 (UTC)
The second citation
I googled it and found it on jstor at http://www.jstor.org/pss/1266761
Would it be a good idea to link directly to the article in the reference section? I didn't know how and I didn't know if that was appropriate or not. It seems appropriate though.
Your thoughts?
--Ted Wheeland (talk) 21:38, 3 January 2010 (UTC)
How is it spoken ?
Like "Out Lier" or "ootlee-er".109.150.237.200 (talk) 09:46, 6 December 2012 (UTC)
Much of this section is directly plagiarized from A Survey of Outlier Detection Methodologies (2004) by Hodge & Austin. For example,
"Type 1 - Determine the outliers with no prior knowledge of the data. This is essentially a learning approach analogous to unsupervised clustering. The approach processes the data as a static distribution, pinpoints the most remote points, and flags them as potential outliers." is word-for-word identical, as are the definitions of the following 2 types. At the absolute minimum their work should be cited. — Preceding unsigned comment added by 208.105.82.93 (talk) 21:48, 7 March 2013 (UTC)
I think the phrase "Deletion of outlier data is a controversial practice frowned on by many scientists and science instructors" needs some kind of citation or should be reformulated / removed. I mean - why is it controversial? What can happen? Examples of bad things that happend because of outliers exclusion? — Preceding unsigned comment added by 89.120.104.106 (talk) 10:26, 18 July 2013 (UTC)
The whole article refers to outliers as possible errors (eg "An outlier may be due to variability in the measurement or it may indicate experimental error"). This is not right or actually incomplete. An outlier can also point at something real going on which is unusual. As a colleague pointed out, it could be your next Nobel prize. If I remember well, NASA had detected the ozone hole in data before Joe Farman published it, but NASA had excluded those. So, I would propose to change the tone of this article and emphasize that statistics helps to identify outliers, which are especially interesting points pointing at measurement errors, unusual distributions (the tails) and possibly new phenomena. — Preceding unsigned comment added by Pjtverheijen (talk • contribs) 05:58, 12 March 2014 (UTC)
Modified Thompson Tau test is *exactly* the Grubbs' test for outliers, which makes things even more confusing. If it's just a different name convention, I think it should just link to Grubbs' test for outliers page. — Preceding unsigned comment added by Yannick Copin (talk • contribs) 16:24, 7 October 2016 (UTC)
This article seems to not include any mention of the corrections for parameter estimates when part of the sample is cut. Remember, some "good" points may be cut together with the outliers. For example, conditional probability can be used to correct the parameter estimates. — Preceding unsigned comment added by 84.245.245.196 (talk) 14:51, 23 March 2017 (UTC)
Why isn't the control chart mentioned as method for outlier detection?
Seems like a blatant omission.
2601:14F:8002:CAD2:88A7:1D26:370A:AFD6 (talk) 11:40, 28 May 2018 (UTC)
consenso: I am proposing the inclusion of a more modern and rigorous definition of outlier that is grounded in the context of applied statistical modeling. The definition comes from a peer-reviewed article published in Survey Review, a journal focused on measurement science. The proposed definition is: "An observation that has moved away from the most likely value to the point of not belonging to the mathematical model (functional and stochastic) stipulated." Citation:
Rofatto, V. F., Matsuoka, M. T., Klein, I., Bonimani, M. L. S., Rodrigues, B. P., de Campos, C. C., Veronez, M. R., & da Silveira, L. G. Jr. (2021). An artificial neural network-based critical values for multiple hypothesis testing: data-snooping case. Survey Review. https://doi.org/10.1080/00396265.2021.1968176
This definition complements the existing content by reflecting how outliers are treated in fields such as geodesy, engineering, and measurement science, where functional and stochastic models are explicitly defined. It is not a repetition of existing definitions but rather adds an important perspective grounded in real-world modeling applications. Additionally, it is relevant to note that the first three authors (Rofatto, Matsuoka, and Klein) are established researchers who have published multiple scientific papers specifically on outlier detection. I would appreciate community input regarding the inclusion of this definition in the article. Thank you. Tomiomatsuoka (talk) 02:03, 23 July 2025 (UTC)
Hello, I’d like to request an edit to the article Outlier.
I am one of the authors of a peer-reviewed paper that proposes a definition of "mathematical outlier" specifically in the context of stochastic and functional modeling. While I understand that citing one's own work directly is discouraged, I believe the proposed definition is both relevant and clearly sourced.
Proposed addition (under the introduction section):
"An outlier can also be defined as an observation that has moved away from the most likely value to the point of not belonging to the mathematical model (functional and stochastic) stipulated."<ref>Rofatto, V. et al. (2021). "An artificial neural network-based critical values for multiple hypothesis testing: data-snooping case". Available at: https://doi.org/10.1080/00396265.2021.1968176
This definition has been cited in the context of geodetic and statistical quality control and offers a modern interpretation aligned with model-based diagnostics. If the placement or wording needs to be adjusted for neutrality or clarity, I’m open to suggestions.
Thank you for considering. — Preceding unsigned comment added by Vfrofatto (talk • contribs)
The first sentence of the second paragraph; "Outliers can occur by chance in any distribution," is hard to understand, and false in many cases (e.g. the degenerate distribution can only take on one value, and so by no definition can a sample thereof contain an outlier). I imagine there exists some concise statistical notion for a distribution which can take on any real number, however I am also sure that that is not the widest qualifier to make this sentence true (e.g. a bounded uniform distribution, which may have a sampled outlier given wide enough bounds). I think the notion of this sentence is correct and useful, which is why I have not removed it, but it could be more clear and more correct while not taking a specific definition of an outlier.
Informasi ini disarikan dari Wikipedia dan disajikan kembali untuk tujuan edukasi. Konten tersedia di bawah lisensi CC BY-SA 3.0. Kami tidak bertanggung jawab atas ketidakakuratan data yang bersumber dari kontribusi publik tersebut.