Skip to main content

If my statistical results do not fall within a "normal" curve, does that mean they are wrong? Why?

I assume that you are asking about results that either lie outside a confidence interval, or results during a hypothesis test that lie in the critical region (tail.)


When creating a confidence interval we start with a point estimate for the population parameter we are interested in. For example, if we want to know the average height we might assume that the average from a random sample of sufficient size is a decent point estimate.


Understanding that the point estimate is not likely to exactly match the population parameter, we introduce an error term. This is added to and subtracted from the point estimate to create a confidence interval. The error term includes the standard error of the measurement, as well as a factor that is derived from the confidence level we want to achieve. (The larger the confidence, the larger the interval will be.)


If a secondary test gives results outside this confidence interval, is the result "wrong"? Not necessarily. Suppose the interval was created with a 95% confidence level. Thus we are 95% certain that the population parameter lies within the interval. One out of every twenty samples will have an estimate outside the interval. The true population parameter would lie outside the interval 5% of the time.


If we are doing a hypothesis test, in essence we are creating a confidence interval centered on the reported or accepted value of the parameter. Then we see if our sample statistic lies in that interval. If it does, we assume that the population parameter is as stated. If our sample statistic lies outside the interval (in the critical region), we have evidence to show that the given population parameter is incorrect.


In running a hypothesis test we run the risk of two types of error (assuming the samples are created correctly, etc...) A type I error is when we say that the purported parameter is incorrect, when it is actually correct. We have a lot of control over this type of error, as the probability of this type of error is equal to our confidence level. (I.e. at the 95% confidence level, the chance for a type I error is 5%.)


A type II error occurs when we fail to recognize an incorrect parameter. We can reduce the probability of this error by increasing the sample size and doing additional samples.


So when running a hypothesis test, the true result is not absolutely given. We only have probabilities to work with.

Comments

Popular posts from this blog

etymology - Origin of the greeting "Sweet dreams".

Does anybody know the etymology of the phrase "sweet dreams"? I tried googling but did not find anything satisfying. Is this a relatively new phrase of the modern world or has this been in use for some time? I think it's the latter one. Answer The OED has the interjection as "a farewell to someone going to bed" from the 20th century: 1908 Sears Roebuck Catal. 198/1 Tenor Solos..Good Bye, Sweet Dreams, Good Bye. But it goes back until at least the 19th and possibly 18th centuries. John Wolcot, writing under the pseudonym of Peter Pindar, used it in his poem "Orson and Ellen; A Legendary Tale" published in 1801: Also from 1801 in The infernal Quixote (Page 287) by Charles Lucas: In the March 1776 of The Universal Magazine was published "The Serenade. A Pastoral Tale. From the German of Gesner" , where the shepherd Daphnis watches over his beloved as she sleeps and sings: The same tale appears in 1776's as Idyl XI, " Daphnis ...

Is there a word/phrase for "unperformant"?

As a software engineer, I need to sometimes describe a piece of code as something that lacks performance or was not written with performance in mind. Example: This kind of coding style leads to unmaintainable and unperformant code. Based on my Google searches, this isn't a real word. What is the correct way to describe this? EDIT My usage of "performance" here is in regard to speed and efficiency. For example, the better the performance of code the faster the application runs. My question and example target the negative definition, which is in reference to preventing inefficient coding practices. Answer This kind of coding style leads to unmaintainable and unperformant code. In my opinion, reads more easily as: This coding style leads to unmaintainable and poorly performing code. The key to well-written documentation and reports lies in ease of understanding. Adding poorly understood words such as performant decreases that ease. In addressing the use of such a poorly ...

What are the similarities between the relationships of the characters in The Outsiders and Romeo and Juliet?

At first glance, it might appear as if the characters in  The Outsiders  by S. E. Hinton and  Romeo and Juliet  by William Shakespeare do not have much in common; however, when one isolates the core of the conflicts, it becomes clear that the relationships in both stories are rooted in rivalry and secrecy.  In both  The Outsiders  and  Romeo and Juliet , an overarching theme is rivalry. While   socioeconomic status separates the greasers from the Socs in The Outsiders , last names separates the Montagues from the Capulets, who appear to be relatively equal in terms of socioeconomic status. Due to these rivalries, there are instances of deadly interactions in both stories. In  The Outsiders , Johnny kills Bob (whether his actions were justified or not is debatable). In  Romeo and Juliet , Romeo kills Tybalt and Paris, and Tybalt kills Mercutio. These deaths have profound effects on the main characters. In Ponyboy's case, the Johnny's death and other events in the novel inspire h...