梁下放床有什么禁忌| 无极是什么意思| 小孩便秘吃什么通便快| 白是什么结构的字| 梦见自己生了个儿子是什么意思| kiko是什么意思| 坎什么意思| 4月是什么星座| 做梦吃饺子是什么意思| 世界上有什么花| 重楼别名叫什么| 为什么不爱我| 十月十五号是什么星座| 小肠疝气挂什么科| 什么炖鸡好吃| 什么床垫最健康| 内痔是什么样的图片| 2月22是什么星座| 红豆和赤小豆有什么区别| 热的什么| 指纹不清晰是什么原因| 小孩血压低是什么原因| 花团锦簇什么意思| 袁崇焕为什么被杀| 珠胎暗结是什么意思| 查五行缺什么| 羞辱什么意思| 腰两边疼是什么原因| 日照有什么特产| 艾叶泡脚有什么好处| 蒲公英泡水喝有什么功效| 心服口服是什么意思| 泡泡尿是什么毛病| 耳朵听不清楚是什么原因| 淋巴结增大是什么原因严重吗| 10月27是什么星座| 收入是什么意思| 感激不尽是什么意思| 2035年属什么生肖| 脂肪粒是什么| 排比句是什么意思| c罗为什么不结婚| 供血不足吃什么药效果最好| 家里为什么会有蜘蛛| 正常白带是什么样的| 总出汗是什么原因| cco是什么职位| 1.19是什么星座| vb是什么| 榴莲不能和什么水果一起吃| 剂量是什么意思| 96615是什么电话| 二十年婚姻是什么婚| 两个立念什么| 肝胆胰脾彩超查什么病| 蛇和什么相冲| 口水粘稠是什么原因| 绿豆有什么功效| 20属什么| 4月份是什么星座| 背德感是什么意思| 肚脐下面疼是什么原因| 吃冬瓜有什么好处| 黄酒是什么酒| 乙状结肠管状腺瘤是什么意思| 马桶对着卫生间门有什么不好| 医院测视力挂什么科| 喝什么醒酒| 洁身自好什么意思| 预防医学是什么| 小囡是什么意思| 糖类抗原是什么意思| 座驾是什么意思| 门槛费是什么意思| 没主见是什么意思| 车辙是什么意思| 农历九月是什么月| 番茄酱可以做什么菜| dazzling什么意思| 型男是什么意思| 蜂王浆是什么东西| 甲状腺结节不能吃什么食物| 吃避孕药会有什么副作用| 心脾两虚吃什么中成药| 西夏是什么民族| 视什么如什么| 手指头发麻是什么原因引起的| 义子是什么意思| 盐酸莫西沙星主治什么| 劳损是什么意思| 闭经和绝经有什么区别| 一阴一阳是什么数字| 多喝柠檬水有什么好处| 一个月小猫吃什么| 成人晚上磨牙是什么原因| 龙生九子都叫什么名字| 仪态万方是什么意思| 鸡蛋与什么食物相克| 御木本是什么档次| 濒死感是什么感觉| 感知力是什么意思| pv是什么材质| 身份证上的数字是什么字体| 趋利避害是什么意思| 别出心裁是什么生肖| 感冒挂什么科室| 左侧卵巢无回声是什么意思| 4月18日什么星座| 淼念什么| 1月17号是什么星座| 扁桃体发炎是什么症状| nsaids是什么药| 大拇指指甲凹凸不平是什么原因| belle什么意思| 金晨什么星座| experiment是什么意思| 上海为什么被称为魔都| 肠道紊乱有什么症状| 泉州有什么特产| 绝对值是什么| 女性尿臭味重是什么病| 香瓜不能和什么一起吃| 尿素高吃什么药| 滞纳金是什么意思| 尿液突然变深褐色是什么原因| 微波炉蒸鸡蛋羹几分钟用什么火| 佛手柑是什么| 什么是虎牙| ssa抗体阳性说明什么| 红楼梦什么朝代| 1月24号什么星座| 孕妇羊水少吃什么补的快| 怡的意思和含义是什么| 鸡爪煲汤放什么材料| 肌肉拉伤看什么科室| 海虾不能和什么一起吃| 包茎是什么| 条件兵是什么意思| 肚子里面跳动是什么原因| 吃大枣有什么好处| 棉绸是什么面料| 帽缨是什么意思| 什么叫小微企业| 乳腺结节是什么引起的| 红斑狼疮是什么症状能治好吗| 乳房痛什么原因| 燚是什么意思| 婴儿掉头发是什么原因| 凌霄花什么时候开花| 1955年属什么| 化合物是什么| 撬墙角是什么意思| 沐沐是什么意思| 水牛吃什么| 强劲的动物是什么生肖| 酸中毒是什么意思| 湿热吃什么水果| 葡萄籽有什么功效和作用| 中二什么意思| cfu是什么意思| 羧甲基纤维素钠是什么| 茅根是什么| 检查眼睛挂什么科| 一库一库雅蠛蝶是什么意思| 21三体高风险是什么原因造成的| 热淋是什么意思| 三点水加一个心读什么| 出国需要什么手续和证件| 人间四月芳菲尽的尽是什么意思| 梦见纸人是什么意思| 夕阳西下是什么意思| 吃什么会拉肚子| 欧阳修字什么| 血脂高胆固醇高吃什么好| 新生儿拉稀是什么原因| 医院挂号用什么app| 檀木手串有什么好处| 脑ct都能查出什么病| 蒙氏结节是什么| 双鱼座最配什么星座| 大三阳是什么| 西兰花不能和什么一起吃| 牙龈肿了吃什么药| 劳烦是什么意思| 8个月宝宝吃什么辅食好| 什么是上升星座| vertu手机为什么那么贵| 为什么家里有蚂蚁| 痛经看什么科| 211和985是什么意思| 积液是什么原因造成的| 亚硝酸钠是什么东西| 吃什么药死的快| 白天不咳嗽晚上咳嗽吃什么药| 胃不好的人吃什么养胃| 克服是什么意思| 乙肝抗体阴性什么意思| 五月二十三日是什么星座| 查肝功能能查出什么病| 前胸疼是什么原因| kappa属于什么档次| 什么茶不影响睡眠| 腰酸背痛挂什么科| 驿站是什么意思| 飧泄是什么意思| 衬衫搭配什么裤子好看| 普字五行属什么| 间接胆红素偏高吃什么药| 什么是热病| 体脂率是什么| 月经前一周是什么期| 柳树像什么| 筋是什么| 肉瘤是什么样子图片| 什么鸡没有毛| 为什么早上起来眼睛肿| 青岛有什么玩的| 黑豆有什么作用| 耳朵上有痣代表什么| 结论是什么意思| 大姨妈来了喝什么好| 狸子是什么动物| 早搏是什么意思| 今年25岁属什么生肖的| 辟支佛是什么意思| 手术室为什么那么冷| 相表里什么意思| 军校毕业是什么军衔| 武装部部长是什么级别| 明年属什么| 神经性皮炎用什么药好| 我流是什么意思| 梦见捡到钱是什么征兆| 白细胞高有什么危害| 甲亢吃什么药好得快| 绍兴本地人喝什么黄酒| 眼压高是什么意思| 阿碧的居所叫什么名字| 男人吃什么可以补精| 尿液少是什么原因| 枸杞和山楂泡水喝有什么功效| 勾引什么意思| 带子是什么海鲜| 阳虚和阴虚有什么区别| 血糖高吃什么食物| 卡宾男装属于什么档次| 日金念什么| 窦性心律是什么意思| kangol是什么牌子| 邪祟是什么意思| 阻生齿是什么| 猫吃什么会死| 大姨妈吃什么好| 芳菲的意思是什么| epo是什么意思| 骐字五行属什么| 猪狗不如是什么意思| 藿香正气水什么人不能喝| 弯的是什么意思| m代表什么| 蚂蚁的触角有什么作用| 黄体不足吃什么药| 回阳救逆什么意思| 股票pb是什么意思| 百度

什么的北京

(Redirected from Type II error)
百度 如果是熬夜引起的浮肿,针对脸部,则用手指尖配合呼吸按动面颊,由耳垂边方向至鼻颊骨旁边,是呼气按、吸气放,很简单却很有效。

Type I error, or a false positive, is the erroneous rejection of a true null hypothesis in statistical hypothesis testing. A type II error, or a false negative, is the erroneous failure in bringing about appropriate rejection of a false null hypothesis.[1]

Type I errors can be thought of as errors of commission, in which the status quo is erroneously rejected in favour of new, misleading information. Type II errors can be thought of as errors of omission, in which a misleading status quo is allowed to remain due to failures in identifying it as such. For example, if the assumption that people are innocent until proven guilty were taken as a null hypothesis, then proving an innocent person as guilty would constitute a Type I error, while failing to prove a guilty person as guilty would constitute a Type II error. If the null hypothesis were inverted, such that people were by default presumed to be guilty until proven innocent, then proving a guilty person's innocence would constitute a Type I error, while failing to prove an innocent person's innocence would constitute a Type II error. The manner in which a null hypothesis frames contextually default expectations influences the specific ways in which type I errors and type II errors manifest, and this varies by context and application.

Knowledge of type I errors and type II errors is applied widely in fields of in medical science, biometrics and computer science. Minimising these errors is an object of study within statistical theory, though complete elimination of either is impossible when relevant outcomes are not determined by known, observable, causal processes.

Definition

edit

Statistical background

edit

In statistical test theory, the notion of a statistical error is an integral part of hypothesis testing. The test goes about choosing about two competing propositions called null hypothesis, denoted by ? and alternative hypothesis, denoted by ?. This is conceptually similar to the judgement in a court trial. The null hypothesis corresponds to the position of the defendant: just as he is presumed to be innocent until proven guilty, so is the null hypothesis presumed to be true until the data provide convincing evidence against it. The alternative hypothesis corresponds to the position against the defendant. Specifically, the null hypothesis also involves the absence of a difference or the absence of an association. Thus, the null hypothesis can never be that there is a difference or an association.

If the result of the test corresponds with reality, then a correct decision has been made. However, if the result of the test does not correspond with reality, then an error has occurred. There are two situations in which the decision is wrong. The null hypothesis may be true, whereas we reject ?. On the other hand, the alternative hypothesis ? may be true, whereas we do not reject ?. Two types of error are distinguished: type I error and type II error.[2]

Type I error

edit

The first kind of error is the mistaken rejection of a null hypothesis as the result of a test procedure. This kind of error is called a type I error (false positive) and is sometimes called an error of the first kind. In terms of the courtroom example, a type I error corresponds to convicting an innocent defendant.

Type II error

edit

The second kind of error is the mistaken failure to reject the null hypothesis as the result of a test procedure. This sort of error is called a type II error (false negative) and is also referred to as an error of the second kind. In terms of the courtroom example, a type II error corresponds to acquitting a criminal.

Crossover error rate

edit

The crossover error rate (CER) is the point at which type I errors and type II errors are equal. A system with a lower CER value provides more accuracy than a system with a higher CER value.

False positive and false negative

edit

In terms of false positives and false negatives, a positive result corresponds to rejecting the null hypothesis, while a negative result corresponds to failing to reject the null hypothesis; "false" means the conclusion drawn is incorrect. Thus, a type I error is equivalent to a false positive, and a type II error is equivalent to a false negative.

Table of error types

edit

Tabulated relations between truth/falseness of the null hypothesis and outcomes of the test:[3]

Table of error types
Null hypothesis (?) is
True False
Decision
about null
hypothesis (?)

Not reject

Correct inference
(true negative)

(probability = ?)

Type II error
(false negative)
(probability = ?)
Reject Type I error
(false positive)
(probability = ?)

Correct inference
(true positive)

(probability = ?)

Error rate

edit
?
The results obtained from negative sample (left curve) overlap with the results obtained from positive samples (right curve). By moving the result cutoff value (vertical bar), the rate of false positives (FP) can be decreased, at the cost of raising the number of false negatives (FN), or vice versa (TP = True Positives, TPR = True Positive Rate, FPR = False Positive Rate, TN = True Negatives).

A perfect test would have zero false positives and zero false negatives. However, statistical methods are probabilistic, and it cannot be known for certain whether statistical conclusions are correct. Whenever there is uncertainty, there is the possibility of making an error. Considering this, all statistical hypothesis tests have a probability of making type I and type II errors.[4]

  • The type I error rate is the probability of rejecting the null hypothesis given that it is true. The test is designed to keep the type I error rate below a prespecified bound called the significance level, usually denoted by the Greek letter α (alpha) and is also called the alpha level.[5] Usually, the significance level is set to 0.05 (5%), implying that it is acceptable to have a 5% probability of incorrectly rejecting the true null hypothesis.[6]
  • The rate of the type II error is denoted by the Greek letter β (beta) and related to the power of a test, which equals 1?β.[citation needed]

These two types of error rates are traded off against each other: for any given sample set, the effort to reduce one type of error generally results in increasing the other type of error.[citation needed]

The quality of hypothesis test

edit

The same idea can be expressed in terms of the rate of correct results and therefore used to minimize error rates and improve the quality of hypothesis test. To reduce the probability of committing a type I error, making the alpha value more stringent is both simple and efficient. For example, setting the alpha value at 0.01, instead of 0.05. To decrease the probability of committing a type II error, which is closely associated with analyses' power, either increasing the test's sample size or relaxing the alpha level, ex. setting the alpha level to 0.1 instead of 0.05, could increase the analyses' power.[citation needed] A test statistic is robust if the type I error rate is controlled.

Varying different threshold (cut-off) values could also be used to make the test either more specific or more sensitive, which in turn elevates the test quality. For example, imagine a medical test, in which an experimenter might measure the concentration of a certain protein in the blood sample. The experimenter could adjust the threshold (black vertical line in the figure) and people would be diagnosed as having diseases if any number is detected above this certain threshold. According to the image, changing the threshold would result in changes in false positives and false negatives, corresponding to movement on the curve.[citation needed]

Example

edit

Since in a real experiment it is impossible to avoid all type I and type II errors, it is important to consider the amount of risk one is willing to take to falsely reject H0 or accept H0. The solution to this question would be to report the p-value or significance level α of the statistic. For example, if the p-value of a test statistic result is 0.0596, then there is a probability of 5.96% that we falsely reject H0 given it is true. Or, if we say, the statistic is performed at level α, like 0.05, then we allow to falsely reject H0 at 5%. A significance level α of 0.05 is relatively common, but there is no general rule that fits all scenarios.

Vehicle speed measuring

edit

The speed limit of a freeway in the United States is 120 kilometers per hour (75?mph). A device is set to measure the speed of passing vehicles. Suppose that the device will conduct three measurements of the speed of a passing vehicle, recording as a random sample X1, X2, X3. The traffic police will or will not fine the drivers depending on the average speed ?. That is to say, the test statistic

?

In addition, we suppose that the measurements X1, X2, X3 are modeled as normal distribution N(μ,2). Then,?T should follow N(μ,2/?) and the parameter μ represents the true speed of passing vehicle. In this experiment, the null hypothesis H0 and the alternative hypothesis H1 should be

H0: μ=120 against H1: μ>120.

If we perform the statistic level at α=0.05, then a critical value c should be calculated to solve

?

According to change-of-units rule for the normal distribution. Referring to Z-table, we can get

?

Here, the critical region. That is to say, if the recorded speed of a vehicle is greater than critical value 121.9, the driver will be fined. However, there are still 5% of the drivers are falsely fined since the recorded average speed is greater than 121.9 but the true speed does not pass 120, which we say, a type I error.

The type II error corresponds to the case that the true speed of a vehicle is over 120 kilometers per hour but the driver is not fined. For example, if the true speed of a vehicle μ=125, the probability that the driver is not fined can be calculated as

?

which means, if the true speed of a vehicle is 125, the driver has the probability of 0.36% to avoid the fine when the statistic is performed at level α=0.05, since the recorded average speed?is lower than 121.9. If the true speed is closer to 121.9 than 125, then the probability of avoiding the fine will also be higher.

The tradeoffs between type I error and type II error should also be considered. That is, in this case, if the traffic police do not want to falsely fine innocent drivers, the level α can be set to a smaller value, like 0.01. However, if that is the case, more drivers whose true speed is over 120 kilometers per hour, like 125, would be more likely to avoid the fine.

Etymology

edit

In 1928, Jerzy Neyman (1894–1981) and Egon Pearson (1895–1980), both eminent statisticians, discussed the problems associated with "deciding whether or not a particular sample may be judged as likely to have been randomly drawn from a certain population":[7] and, as Florence Nightingale David remarked, "it is necessary to remember the adjective 'random' [in the term 'random sample'] should apply to the method of drawing the sample and not to the sample itself".[8]

They identified "two sources of error", namely:

  1. the error of rejecting a hypothesis that should have not been rejected, and
  2. the error of failing to reject a hypothesis that should have been rejected.

In 1930, they elaborated on these two sources of error, remarking that

in testing hypotheses two considerations must be kept in view, we must be able to reduce the chance of rejecting a true hypothesis to as low a value as desired; the test must be so devised that it will reject the hypothesis tested when it is likely to be false.

In 1933, they observed that these "problems are rarely presented in such a form that we can discriminate with certainty between the true and false hypothesis". They also noted that, in deciding whether to fail to reject, or reject a particular hypothesis amongst a "set of alternative hypotheses", H1, H2..., it was easy to make an error,

[and] these errors will be of two kinds:

  1. we reject H0 [i.e., the hypothesis to be tested] when it is true,[9]
  2. we fail to reject H0 when some alternative hypothesis HA or H1 is true. (There are various notations for the alternative).

In all of the papers co-written by Neyman and Pearson the expression H0 always signifies "the hypothesis to be tested".

In the same paper they call these two sources of error, errors of type I and errors of type II respectively.[10]

edit

Null hypothesis

edit

It is standard practice for statisticians to conduct tests in order to determine whether or not a "speculative hypothesis" concerning the observed phenomena of the world (or its inhabitants) can be supported. The results of such testing determine whether a particular set of results agrees reasonably (or does not agree) with the speculated hypothesis.

On the basis that it is always assumed, by statistical convention, that the speculated hypothesis is wrong, and the so-called "null hypothesis" that the observed phenomena simply occur by chance (and that, as a consequence, the speculated agent has no effect)?– the test will determine whether this hypothesis is right or wrong. This is why the hypothesis under test is often called the null hypothesis (most likely, coined by Fisher (1935, p.?19)), because it is this hypothesis that is to be either nullified or not nullified by the test. When the null hypothesis is nullified, it is possible to conclude that data support the "alternative hypothesis" (which is the original speculated one).

The consistent application by statisticians of Neyman and Pearson's convention of representing "the hypothesis to be tested" (or "the hypothesis to be nullified") with the expression H0 has led to circumstances where many understand the term "the null hypothesis" as meaning "the nil hypothesis"?– a statement that the results in question have arisen through chance. This is not necessarily the case?– the key restriction, as per Fisher (1966), is that "the null hypothesis must be exact, that is free from vagueness and ambiguity, because it must supply the basis of the 'problem of distribution', of which the test of significance is the solution."[11] As a consequence of this, in experimental science the null hypothesis is generally a statement that a particular treatment has no effect; in observational science, it is that there is no difference between the value of a particular measured variable, and that of an experimental prediction.[citation needed]

Statistical significance

edit

If the probability of obtaining a result as extreme as the one obtained, supposing that the null hypothesis were true, is lower than a pre-specified cut-off probability (for example, 5%), then the result is said to be statistically significant and the null hypothesis is rejected.

British statistician Sir Ronald Aylmer Fisher (1890–1962) stressed that the null hypothesis

is never proved or established, but is possibly disproved, in the course of experimentation. Every experiment may be said to exist only in order to give the facts a chance of disproving the null hypothesis.

—?Fisher, 1935, p.19

Application domains

edit

Medicine

edit

In the practice of medicine, the differences between the applications of screening and testing are considerable.

Medical screening

edit

Screening involves relatively cheap tests that are given to large populations, none of whom manifest any clinical indication of disease (e.g., Pap smears).

Testing involves far more expensive, often invasive, procedures that are given only to those who manifest some clinical indication of disease, and are most often applied to confirm a suspected diagnosis.

For example, most states in the US require newborns to be screened for phenylketonuria and hypothyroidism, among other congenital disorders.

  • Hypothesis: "The newborns have phenylketonuria and hypothyroidism".
  • Null hypothesis (H0): "The newborns do not have phenylketonuria and hypothyroidism".
  • Type I error (false positive): The true fact is that the newborns do not have phenylketonuria and hypothyroidism but we consider they have the disorders according to the data.
  • Type II error (false negative): The true fact is that the newborns have phenylketonuria and hypothyroidism but we consider they do not have the disorders according to the data.

Although they display a high rate of false positives, the screening tests are considered valuable because they greatly increase the likelihood of detecting these disorders at a far earlier stage.

The simple blood tests used to screen possible blood donors for HIV and hepatitis have a significant rate of false positives; however, physicians use much more expensive and far more precise tests to determine whether a person is actually infected with either of these viruses.

Perhaps the most widely discussed false positives in medical screening come from the breast cancer screening procedure mammography. The US rate of false positive mammograms is up to 15%, the highest in world. One consequence of the high false positive rate in the US is that, in any 10-year period, half of the American women screened receive a false positive mammogram. False positive mammograms are costly, with over $100 million spent annually in the U.S. on follow-up testing and treatment. They also cause women unneeded anxiety. As a result of the high false positive rate in the US, as many as 90–95% of women who get a positive mammogram do not have the condition. The lowest rate in the world is in the Netherlands, 1%. The lowest rates are generally in Northern Europe where mammography films are read twice and a high threshold for additional testing is set (the high threshold decreases the power of the test).

The ideal population screening test would be cheap, easy to administer, and produce zero false negatives, if possible. Such tests usually produce more false positives, which can subsequently be sorted out by more sophisticated (and expensive) testing.

Medical testing

edit

False negatives and false positives are significant issues in medical testing.

  • Hypothesis: "The patients have the specific disease".
  • Null hypothesis (H0): "The patients do not have the specific disease".
  • Type I error (false positive): The true fact is that the patients do not have a specific disease but the physician judges the patient is ill according to the test reports.
  • Type II error (false negative): The true fact is that the disease is actually present but the test reports provide a falsely reassuring message to patients and physicians that the disease is absent.

False positives can also produce serious and counter-intuitive problems when the condition being searched for is rare, as in screening. If a test has a false positive rate of one in ten thousand, but only one in a million samples (or people) is a true positive, most of the positives detected by that test will be false. The probability that an observed positive result is a false positive may be calculated using Bayes' theorem.

False negatives produce serious and counter-intuitive problems, especially when the condition being searched for is common. If a test with a false negative rate of only 10% is used to test a population with a true occurrence rate of 70%, many of the negatives detected by the test will be false.

This sometimes leads to inappropriate or inadequate treatment of both the patient and their disease. A common example is relying on cardiac stress tests to detect coronary atherosclerosis, even though cardiac stress tests are known to only detect limitations of coronary artery blood flow due to advanced stenosis.

Biometrics

edit

Biometric matching, such as for fingerprint recognition, facial recognition or iris recognition, is susceptible to type I and type II errors.

  • Hypothesis: "The input does not identify someone in the searched list of people".
  • Null hypothesis: "The input does identify someone in the searched list of people".
  • Type I error (false reject rate): The true fact is that the person is someone in the searched list but the system concludes that the person is not according to the data.
  • Type II error (false match rate): The true fact is that the person is not someone in the searched list but the system concludes that the person is someone whom we are looking for according to the data.

The probability of type I errors is called the "false reject rate" (FRR) or false non-match rate (FNMR), while the probability of type II errors is called the "false accept rate" (FAR) or false match rate (FMR).

If the system is designed to rarely match suspects then the probability of type II errors can be called the "false alarm rate". On the other hand, if the system is used for validation (and acceptance is the norm) then the FAR is a measure of system security, while the FRR measures user inconvenience level.

Security screening

edit

False positives are routinely found every day in airport security screening, which are ultimately visual inspection systems. The installed security alarms are intended to prevent weapons being brought onto aircraft; yet they are often set to such high sensitivity that they alarm many times a day for minor items, such as keys, belt buckles, loose change, mobile phones, and tacks in shoes.

  • Hypothesis: "The item is a weapon".
  • Null hypothesis: "The item is not a weapon".
  • Type I error (false positive): The true fact is that the item is not a weapon but the system still sounds an alarm.
  • Type II error (false negative) The true fact is that the item is a weapon but the system keeps silent at this time.

The ratio of false positives (identifying an innocent traveler as a terrorist) to true positives (detecting a would-be terrorist) is, therefore, very high; and because almost every alarm is a false positive, the positive predictive value of these screening tests is very low.

The relative cost of false results determines the likelihood that test creators allow these events to occur. As the cost of a false negative in this scenario is extremely high (not detecting a bomb being brought onto a plane could result in hundreds of deaths) whilst the cost of a false positive is relatively low (a reasonably simple further inspection) the most appropriate test is one with a low statistical specificity but high statistical sensitivity (one that allows a high rate of false positives in return for minimal false negatives).

Computers

edit

The notions of false positives and false negatives have a wide currency in the realm of computers and computer applications, including computer security, spam filtering, malware, optical character recognition, and many others.

For example, in the case of spam filtering:

  • Hypothesis: "The message is spam".
  • Null hypothesis: "The message is not spam".
  • Type I error (false positive): Spam filtering or spam blocking techniques wrongly classify a legitimate email message as spam and, as a result, interfere with its delivery.
  • Type II error (false negative): Spam email is not detected as spam, but is classified as non-spam.

While most anti-spam tactics can block or filter a high percentage of unwanted emails, doing so without creating significant false-positive results is a much more demanding task. A low number of false negatives is an indicator of the efficiency of spam filtering.

See also

edit

References

edit
  1. ^ "Type I Error and Type II Error". explorable.com. Retrieved 14 December 2019.
  2. ^ A modern introduction to probability and statistics?: understanding why and how. Dekking, Michel, 1946-. London: Springer. 2005. ISBN?978-1-85233-896-1. OCLC?262680588.{{cite book}}: CS1 maint: others (link)
  3. ^ Sheskin, David (2004). Handbook of Parametric and Nonparametric Statistical Procedures. CRC Press. p.?59. ISBN?1584884401.
  4. ^ Rohatgi, V. K.; Saleh, A. K. Md Ehsanes (2015). An introduction to probability theory and mathematical statistics. Wiley series in probability and statistics (3rd?ed.). Hoboken, New Jersey: John Wiley & Sons, Inc. ISBN?978-1-118-79963-5.
  5. ^ Lindenmayer, David. (2005). Practical conservation biology. Burgman, Mark A. Collingwood, Vic.: CSIRO Pub. p.?404. ISBN?0-643-09310-9. OCLC?65216357. These parameters are related by the expression:… where E is effect size, n is sample size, α is the type I error rate and σ is the standard deviation of the variability of the data
  6. ^ Lindenmayer, David. (2005). Practical conservation biology. Burgman, Mark A. Collingwood, Vic.: CSIRO Pub. p.?403. ISBN?0-643-09310-9. OCLC?65216357. By convention, the type I error rate is set at 0.05.
  7. ^ Neyman, J.; Pearson, E. S. (1928). "On the Use and Interpretation of Certain Test Criteria for Purposes of Statistical Inference Part I". Biometrika. 20A (1–2): 175–240. doi:10.1093/biomet/20a.1-2.175. ISSN?0006-3444.
  8. ^ C. I. K. F. (July 1951). "Probability Theory for Statistical Methods. By F. N. David. [Pp. ix + 230. Cambridge University Press. 1949. Price 155.]". Journal of the Staple Inn Actuarial Society. 10 (3): 243–244. doi:10.1017/s0020269x00004564. ISSN?0020-269X.
  9. ^ The subscript in the expression H0 is a zero (indicating null), and is not an "O" (indicating original).
  10. ^ Neyman, J.; Pearson, E. S. (30 October 1933). "The testing of statistical hypotheses in relation to probabilities a priori". Mathematical Proceedings of the Cambridge Philosophical Society. 29 (4): 492–510. Bibcode:1933PCPS...29..492N. doi:10.1017/s030500410001152x. ISSN?0305-0041. S2CID?119855116.
  11. ^ Fisher, R. A. (1966). The design of experiments (8th?ed.). Edinburgh: Hafner.

Bibliography

edit
  • Betz, M.A. & Gabriel, K.R., "Type IV Errors and Analysis of Simple Effects", Journal of Educational Statistics, Vol.3, No.2, (Summer 1978), pp.?121–144.
  • David, F.N., "A Power Function for Tests of Randomness in a Sequence of Alternatives", Biometrika, Vol.34, Nos.3/4, (December 1947), pp.?335–339.
  • Fisher, R.A., The Design of Experiments, Oliver & Boyd (Edinburgh), 1935.
  • Gambrill, W., "False Positives on Newborns' Disease Tests Worry Parents", Health Day, (5 June 2006). [1] Archived 17 May 2018 at the Wayback Machine
  • Kaiser, H.F., "Directional Statistical Decisions", Psychological Review, Vol.67, No.3, (May 1960), pp.?160–167.
  • Kimball, A.W., "Errors of the Third Kind in Statistical Consulting", Journal of the American Statistical Association, Vol.52, No.278, (June 1957), pp.?133–142.
  • Lubin, A., "The Interpretation of Significant Interaction", Educational and Psychological Measurement, Vol.21, No.4, (Winter 1961), pp.?807–817.
  • Marascuilo, L.A. & Levin, J.R., "Appropriate Post Hoc Comparisons for Interaction and nested Hypotheses in Analysis of Variance Designs: The Elimination of Type-IV Errors", American Educational Research Journal, Vol.7., No.3, (May 1970), pp.?397–421.
  • Mitroff, I.I. & Featheringham, T.R., "On Systemic Problem Solving and the Error of the Third Kind", Behavioral Science, Vol.19, No.6, (November 1974), pp.?383–393.
  • Mosteller, F., "A k-Sample Slippage Test for an Extreme Population", The Annals of Mathematical Statistics, Vol.19, No.1, (March 1948), pp.?58–65.
  • Moulton, R.T., "Network Security", Datamation, Vol.29, No.7, (July 1983), pp.?121–127.
  • Raiffa, H., Decision Analysis: Introductory Lectures on Choices Under Uncertainty, Addison–Wesley, (Reading), 1968.
edit
  • Bias and Confounding?– presentation by Nigel Paneth, Graduate School of Public Health, University of Pittsburgh
低血压吃什么好的最快女性 二甲双胍缓释片什么时候吃最好 蜻蜓是什么动物 不是经期有少量出血是什么原因 杨玉环属什么生肖
费率是什么 寄什么快递最便宜 3月15号是什么星座 慢阻肺是什么原因引起的 内裤发霉是什么原因
glenfiddich是什么酒 以梦为马是什么意思 为什么会牙疼 人肉什么味道 臁疮是什么病
舌头白色的是什么原因 什么人容易得肺结核 输血前常规检查是什么 单核细胞是什么 mm表示什么
车顶放饮料是什么意思hcv9jop0ns5r.cn 端午节喝什么酒hcv9jop5ns3r.cn 总监是什么级别hcv8jop3ns1r.cn 经警是做什么的hcv7jop6ns5r.cn 为什么指甲有竖纹hcv7jop9ns8r.cn
什么食物含叶黄素最多hcv9jop8ns1r.cn 吃深海鱼油有什么好处和坏处hcv9jop2ns3r.cn 七月十三日是什么日子hcv8jop0ns2r.cn 早上起来流鼻血是什么原因hcv9jop3ns3r.cn 胸闷什么原因hcv8jop0ns1r.cn
什么魂什么魄hcv8jop9ns5r.cn 收缩压低是什么原因hcv8jop7ns8r.cn 姓叶的男孩取什么名字好hanqikai.com 榴莲什么季节成熟hcv9jop6ns8r.cn 孕妇上火了吃什么降火最快hcv9jop1ns2r.cn
维生素d什么时候吃最好hcv8jop8ns3r.cn 甲状腺结节是什么病hcv9jop6ns9r.cn 张学良为什么不回大陆hcv8jop3ns3r.cn 看颈椎病挂什么科ff14chat.com 消化内科是看什么病的hcv8jop7ns1r.cn
百度