Size & Power of Test (Edexcel A Level Further Maths: Further Statistics 1): Revision Note

Exam code: 9FM0

Dan Finlay

Written by: Dan Finlay

Reviewed by: Lucy Kirkham

Updated on

Size

What is the size of a test?

  • The size of a test is the probability of rejecting H0 when it was in fact true

    • P(in critical region | H0 is true) 

    • The situation being described is not a good outcome 

      • Something has been rejected when it was actually true!

    • A better test has a smaller size

      • You want to minimise this error happening

  • Size is related to the significance level, α%

    • A better test has a smaller significance level (e.g. 1%)

    • For continuous distributions (e.g. normal)

      • Size = significance level, α

      • You can often write this down with no calculation

    • For discrete distributions (e.g. binomial, Poisson, geometric)

      • Size = actual significance level (≤α)

      • As close to α% as a discrete variable can get, whilst still being critical

How does size relate to Type I errors?

  • The size is exactly the same as the probability of a Type I error

    • Both want to know the probability of rejecting H0 when it was in fact true

Worked Example

A student wants to test, at a 10% significant level, whether a coin is biased towards heads by counting the number of heads in 20 flips of the coin.

Calculate the size of this test.

size-1
size-2

Power

What is the power of a test?

  • The power of a test is the probability of rejecting H0 when it was false

    • P(in critical region | H0 is false) 

    • The situation being described has a good outcome 

      • The null hypothesis was false and it rightly got rejected

    • A better test has a higher power

      • You want to maximise this happening

  • In practice, you need to be given the actual population parameter to calculate the power

    • For example, H0 assumed p=12 but actually p=13

      • This is more helpful than just saying p12 

    • Power is P(in the critical region | actual population parameter)

How does power relate to Type II errors?

  • The power of a test is 1 - P(Type II error) 

    • Power is when H0 is false and it gets rejected

      • That's a good outcome

    • A Type II error is when H0 is false and it does not get rejected

      • That's a bad outcome

  • You ideally want the power of a test to be greater than 0.5 

    • That way it's less likely to produce a Type II error

      • And more likely to reach the correct conclusion

Worked Example

Let X~Po(λ). A hypothesis test is conducted at the 5% significance level in which H0: λ=8 and H1: λ<8.

If is later discovered that λ=6, find the power of the test.

power-of-test

Power functions

What is a power function?

  • The power function is the power of a test written algebraically

    • In terms of p or λ

    • For when you're not given the actual population parameter in the question

  • In reality, it's very unlikely you'll know the actual population parameter anyway

    • Otherwise you wouldn't be doing a hypothesis test on it!

    • Power functions don't need this information

  • The power function is P(in the critical region | population parameter is p)

    • Or, for Poisson 

      • P(in the critical region | population parameter is λ)

How do I find power functions?

  • It's easier to show in an example

    • If the critical region is X2 for a binomial hypothesis test with n=50 

    • Then the power function is P(in the critical region | population parameter is p)

      • Let X~B(50,p)

      • P(X2 | p)=(500)p0(1p)50+(501)p1(1p)49+(502)p2(1p)48

    • Simplify

      • (1p)50+50p(1p)49+1225p2(1p)48

    • Factorise and collect like terms

      • The power function is (1p)48(1+48p+1176p2)

What can I do with power functions?

  • You can plot them against p (or λ

    • You can then see where the power is biggest

  • You can input different values of p (or λ

    • To compare two (or more) different hypothesis tests

      • The better test is the one with the higher power

    • To check if the power of a test is greater than 0.5

      • So that it's more likely to reach the correct conclusion

      • And less likely to produce a Type II error

      • Because Power = 1 - P(Type II error)

Worked Example

Residents suspect that the number of accidents on a main road has decreased. They test, at a 5% significance level, the hypotheses H0: λ=9 and H1: λ<9.

a) Show that the power function is 16eλ(a+bλ+cλ2+dλ3), where a, b, c and d are integers to be found.

power-functions-1
power-functions-2

b) Find the largest integer value of λ for which the probability of a Type II error is less than 20%.

power-functions-3

Unlock more, it's free!

Join the 100,000+ Students that ❤️ Save My Exams

the (exam) results speak for themselves:

Build on this topic

Dan Finlay

Author: Dan Finlay

Expertise: Portfolio Lead

Dan graduated from the University of Oxford with a First class degree in mathematics. As well as teaching maths for over 8 years, Dan has marked a range of exams for Edexcel, tutored students and taught A Level Accounting. Dan has a keen interest in statistics and probability and their real-life applications.

Lucy Kirkham

Reviewer: Lucy Kirkham

Expertise: Content Creator

Lucy has been a passionate Maths teacher for over 12 years, teaching maths across the UK and abroad helping to engage, interest and develop confidence in the subject at all levels.Working as a Head of Department and then Director of Maths, Lucy has advised schools and academy trusts in both Scotland and the East Midlands, where her role was to support and coach teachers to improve Maths teaching for all.