Summary

Summary Intermediate Statistics 2 Study Guide

Rating

Sold

Pages

Uploaded on

08-06-2022

Written in

2021/2022

Study guide of intermediate statistics 2, includes notes from the lectures, the textbook, and PBLs

Institution

Module

Whoops! We can’t load your doc right now. Try again or contact support.

Report Copyright Violation

Connected book

Andy Field Discovering Statistics Using IBM SPSS

Edition:maart 2013
ISBN:9781446249185
Edition:4

Written for

Institution: Erasmus Universiteit Rotterdam (EUR)
Study: Liberal Arts And Sciences
Module: Intermediate Statistics II (EUCINT207)

All documents for this subject (3)

Document information

Summarized whole book?: No
Which chapters are summarized?: Unknown
Uploaded on: June 8, 2022
Number of pages: 18
Written in: 2021/2022
Type: Summary

Subjects

stats 2
statistics 2
intermediate statistics

Content preview

Maria Andrade

Stats II Study Guide

Week 1: Revision Stats I & Dummy Coding

Revision Stats 1

Linear Regression

● Dependent variable → Y
● Independent variable(s) → X
● Function of linear regression:
○ B0 → population y-intercept
○ B1 → population slope coefficient
○ Xi → independent variable
○ Ei → random error

Eg: Interpretation of betas
● Eg: pricei= B0 + B1 · squaremeteri + B2 · bedrooms + Ei
● B0: the predicted house price when the amount of bedrooms is 0 and the square meters is 0
● B1: the increase in the predicted house price for every additional square meter given that the amount
of bedrooms remains constant
● B2: the increase in the predicted house price for every additional bedroom given that the amount of
square meters remains constant.

P-Values

● Alpha = 0.05 → how often we allow ourselves to make a mistake
● compare the p-value with alpha → if the p-value is lower than alpha you reject the Ho

Model Fit: To test model fit you have SST, SSR and SSM

Model Fit description Formula Variance exp

SST difference btw the observed total unstandardized variance
data and the mean of y

SSR Difference btw the observed unexplained unstandardized variances→
data and the model variation not accounted for in the model

, Maria Andrade

SSM Difference btw the men value of explained unstandardized variance →
Y and the model variation accounted for in the model

F-Ratio

● F-ratio: the ratio btw the standardized SSM and standardized SSR

○ Formula:

■ MSM Formula =
● MSM stands for the standardized explained variance

■ MSR formula =
● MSR stands for the standardized unexplained variance
○ When the F-ratio is high → the explained variance is high and the unexplained variance is low
R^2

● R2: the proportion of explained variance over total variance

○ Formula:
● Can be used to compare models, to see if one is better than the other
● The higher the R2 the more variance is explained

Assumptions of a Line

● If the assumptions are not met, then the inference of the results are invalid.

Linearity Independence of Normality (errors) Homoscedasticity multicollinearity
errors

meaning If yi is a linear The errors are Errors are normally Errors have equal 2 or + predictors are
function of the independent distributed variance highly correlated with
predictors each other

Check Residuals plot: X If time series 1)Histograms Zpred-Zresid plot VIF (>10) or tolerance
= ZPRED, Y = Durbin- Watson 2) PP/QQ plots Leven’s Test (<0.1) Average VIF
ZRESID 3)KS-SW test “much larger” than 1
If residuals are Not for cross 4)Skew & Kurtosis
symmetric sectional data
around 0

, Maria Andrade

+ 2)PP/QQ plots: Pp-plot: Equality of variance of Predictors explain the
magnify deviations in the errors same variance
middle & qq-plot : magnify
deviations in the tails
4) s/SEskewness K
/SEkurtosis

Fix Transform data/ Multilevel modeling SE’s are inflated, change SE’ inflates Remove variables
change model or clustered SEs through transform or Transform or
bootstrap bootstrapping

Outliers

● An outlier is an extreme in y
● Its cause of concern when:
○ >5% of data > 1.96 sd
○ >1% of data > 2.58 sd
○ >3.29sd

Influential Cases

● A case which influences any part of the regression analysis
● Its an extreme in x → pushes regression line
● Diagnostics:
○ Leverage → measures potential to influence regression
○ Mahalanobis distance → measures potential to influence regression
○ DFFIT(s) → difference in mean y including and excluding case
○ SDFBeta → change in one regression coefficient after exclusion
○ Cook’s Distance → the average of changes in all regression coefficients after exclusion

Dummy Coding

Dummy coding → categorical predictor with multiple categories

Steps:
1. Recode a variable into dummies
2. Number of dummies = categories - 1
3. A dummy is 0 or 1 for a particular category
4. Reference category is 0 for all dummies

$10.92

Get access to the full document:

100% satisfaction guarantee

Immediately available after payment

Both online and in PDF

No strings attached

Get to know the seller

mcandradep01

Get to know the seller

mcandradep01 Erasmus Universiteit Rotterdam

View profile

Sold

Member since

5 year

Number of followers

Documents

Last sold

1 year ago

0.0

0 reviews

Why students choose Stuvia

Created by fellow students, verified by reviews

Quality you can trust: written by students who passed their exams and reviewed by others who've used these revision notes.

Didn't get what you expected? Choose another document

No problem! You can straightaway pick a different document that better suits what you're after.

Pay as you like, start learning straight away

No subscription, no commitments. Pay the way you're used to via credit card and download your PDF document instantly.

“Bought, downloaded, and smashed it. It really can be that simple.”

Alisha Student

Frequently asked questions

What do I get when I buy this document?

You get a PDF, available immediately after your purchase. The purchased document is accessible anytime, anywhere and indefinitely through your profile.

Satisfaction guarantee: how does it work?

Our satisfaction guarantee ensures that you always find a study document that suits you well. You fill out a form, and our customer service team takes care of the rest.

Who am I buying these notes from?

Stuvia is a marketplace, so you are not buying this document from us, but from seller mcandradep01. Stuvia facilitates payment to the seller.

Will I be stuck with a subscription?

No, you only buy these notes for $10.92. You're not tied to anything after your purchase.

Can Stuvia be trusted?

4.6 stars on Google & Trustpilot (+1000 reviews) 44104 documents were sold in the last 30 days Founded in 2010, the go-to place to buy revision notes and other study material for 15 years now

Summary Intermediate Statistics 2 Study Guide

Connected book

Written for

Document information

Subjects

Content preview

Get to know the seller

Recently viewed by you

Why students choose Stuvia

Created by fellow students, verified by reviews

Didn't get what you expected? Choose another document

Pay as you like, start learning straight away

Frequently asked questions

What do I get when I buy this document?

Satisfaction guarantee: how does it work?

Who am I buying these notes from?

Will I be stuck with a subscription?

Can Stuvia be trusted?