This is the assessment of Statistics need to do as the was is done in the assessment sample and half of the work is done by me the excel report are done by me just need to do others things as shown in the sample and the wording for the assessment is
Pages 2 to 7 are a template to help you with the computing assignment due week 10 , if you replace the highlighted parts with appropriate material will get high marks in the assignment, the complete assignment instructions are available from https://app.box.com/s/yavdfgve8c5a63rvwkr3twui4znhs9yc
There are 6 sections,
Note that in week 4 of moodle there are instructions for submitting a draft of section 1a Note that in week 7 of moodle there are instructions for submitting a draft of section 2a Note that in week 8 of moodle there are instructions for submitting a draft of section 3a Note that in week 9 of moodle there are instructions for submitting a draft of section 4a
You should look the guide to summarizing data sets
https://app.box.com/s/jxuqhpzjrfj14xiq28x1bnywjv1iayr4
Brief introduction: this assignment gives examples of how to summarize datasets by using numbers and graphs that describe variables individually and describe the relationship between variables (replace this with your quick summary)
Section 1
a) sample 700
add a comment about the relationship between variables
b) The predicted contribution when income is 200,000 is predicted annual contribution =0.1318*200,000+947=27307
c) The average of all the 10,000 estimates is 27000 with standard deviation 2100 So the zscore for sample 700 estimate is (27307-27000)/2100= 0.15 , d) using wolframalpha.com P(Z<0.15)=0.5581 e) So if you compare sample 700 to the 10,000 samples then Predicted rank = P(Z<zscore)*10000=0.5581*10,000=5581
Section 2
a)
Sample 700
|
Investment type |
Lost money |
Made money |
Total |
|
Risky investment plan (mainly shares) |
14 |
60 |
74 |
|
Safer investment plan (mainly use the bank) |
3 |
23 |
26 |
|
Investment type |
Proportion that Lost money |
Proportion that Made money |
|
Risky investment plan (mainly shares) |
18.92% |
81.08% |
|
Safer investment plan (mainly use the bank) |
11.54% |
88.46% |
b) add a suitable graph refer to the guide to summarizing data sets https://app.box.com/s/jxuqhpzjrfj14xiq28x1bnywjv1iayr4
c) add a comment about the relationship between variables
D i) Using the sample 700 the estimated difference in proportions is 0.1892-0.1154=0.0738 ii) The average of the 4000 sample estimates is 0.1 with standard deviation 0.0743 So the zscore for that estimate in sample 700 is (0.0738-0.1)/0.0743=-0.35 iii) using wolframlapha.com P(Z<zscore) = P(Z<-0.35) =0.3632 iv)So when you compare sample 700 to the 4000 other samples you predict the rank to be 0.3632*4000=1453
e) Test the claim there is a difference in proportions (in other words test the claim of independence) using your sample refer to the final exam guide
you need to state H0 and H1, find the p-value using the website http://epitools.ausvet.com.au/content.php?page=z-test-2
decide whether or not your reject H0 and give a conclusion in plain English
If you look through the following document you will find a similar question
https://app.box.com/s/dxb7en1tdkt16uok0hp96gh50ea3q1o6
section 3
a) Sample 700
|
Investment type |
Number of investors |
Average return |
Standard deviation |
|
Safe type (mainly use the bank) |
72 |
0.0345 |
0.0028 |
|
Risky type (mainly use the stock market) |
28 |
0.0850 |
0.0849 |
b) Add an appropriate graph, refer to the guide of summarizing data sets https://app.box.com/s/jxuqhpzjrfj14xiq28x1bnywjv1iayr4
c)
Add a comment about the relationship between variables
di)
So for sample 700 the estimate of the difference in the population means is the difference in the sample means =0.0345-0.085=-0.0505
D ii) The average of the 2000 sample estimates is -0.0256 with standard deviation 0.0173 so the zscore of the sample 700 estimate is =(-0.505- -0.0256)/ 0.0173=-1.44 iii) using wolframalpha.com P(Z<zscore) = P(Z<-1.44)=0.075
iv) So if you compare the sample estimate 700 to the 2000 other sample estimates you expect it to have rank 2000*0.075=150
E) Test the claim there is a difference in means (in other words test the claim of independence) using your sample refer to the final exam guide
you need to state H0 and H1, find the p-value using https://www.medcalc.org/calc/comparison_of_means.php
, decide whether or not your reject H0 and give a conclusion in plain English
Refer to the guide to final exam questions for proper setting out
Section 4
Sample 700 a)
|
|
no |
Yes |
Total |
|
Count of do you support proposed change? |
86 |
99 |
185 |
b) Sample size n is 185 , Sample proportion of people that say yes is =99/185 =0.5351
ci) the average of the 1000 estimates is 0.6 with standard deviation 0.0357
So the zscore of the sample 700 estimate is =(0.5351-0.6)/0.0357=-1.82 ii) using wolframalpha.com P(Z<zscore)=P(Z<-1.82)=0.0344 iii)So if we compare sample 700 to the 1000 other estimates then we expect the rank to be 0.0344*1000=0.0344=34
d) find 95% confidence interval for proportion
Section 5
Make up your own data set and summarize it
Your must be have at least 5 rows (observations) it must have at least 2 variables, At least one of the variables must be categorical
note that the name of each thing in the data set is NOT a variable
For example the data set
|
Name |
Gender |
|
Sue |
Female |
|
Fred |
Male |
|
Bill |
Male |
|
Tim |
Male |
|
Donald |
Male |
Only has 1 variable Gender male or female, 1/5=20% are female , you Cannot summarize the name column it is not A variable
Summarize the main ideas in the guide to summarizing data sets. https://app.box.com/s/jxuqhpzjrfj14xiq28x1bnywjv1iayr4 This is where you demonstrate you have actually learnt something, you can mainly use plain language and describe things in the same way you would describe them to your friends just make sure you use the phrases “Categorical variable” and “Quantitative variable”