100% satisfaction guarantee Immediately available after payment Both online and in PDF No strings attached
logo-home
Introduction to Statistics - Lecture Notes $3.21
Add to cart

Class notes

Introduction to Statistics - Lecture Notes

3 reviews
 111 views  1 purchase
  • Course
  • Institution
  • Book

Introductory course on statistics for the first year of Sociology by the lecturer Thijs Bol at the UvA.

Preview 3 out of 26  pages

  • August 1, 2019
  • 26
  • 2018/2019
  • Class notes
  • Unknown
  • All classes

3  reviews

review-writer-avatar

By: nzachova17 • 1 year ago

review-writer-avatar

By: brentriet • 9 months ago

review-writer-avatar

By: martijnediepstraten • 4 year ago

avatar-seller
INTRODUCTION TO STATISTICS – Lecture 0 19/11/2018


Types of variables
Different types of variables:

Measurement level Description Example
Religion, political party voted
NOMINAL No rank order
for
Rank order, but unequal Disagree completely – Agree
ORDINAL
distances completely
Rank order with equal
INTERVAL Celsius, hourly wage
distances

Rank order with equal
RATIO Age, weight, height
distances and a natural 0


Nominal
Closed (categorical) questions
Ordinal
Closed questions

DICHOTOMOUS VARIABLES
There are just two categories: YES or NO, 0 and 1
Sex? 0.Female
1.Male

Different types of variables require different types of description.
We want to describe data. We can’t do this by showing all answers to a survey.
A core function of statistics is to describe (survey) data: centrality and dispersion.

CENTRALITY
Where is the center of the variable?
Three common way to address centrality:
- Mode indicates the most common value
- Median indicates the middle value
Mean 𝑦̅ indicates the average value
∑ 𝑦𝑖
𝑦̅ = -> sum of all values divided by the number of observations
𝑛

For dichotomous variables the mean equals the proportion 𝜋̂
The proportion is basically the same as the percentage. Proportion = percentage/100

The type of variable defines the centrality measure that we can use.
Nominal: mode
Ordinal: mode and median. Mean not really allowed but every uses it

,Interval/ratio: mode, median, mean
Dichotomous: mean
DISPERSION
If we know the center of data, we know very little about the distribution of data. Data has a
certain level of dispersion. And there are different measures for dispersion:
- Frequencies: how often do we see each answer?
- Range: what’s the minimum and maximum value?
- Standard deviation s
- Variance s2

Standard deviation s
The sum of all squared distances to the mean.
If all observations are clustered around the mean, the sum of distances will be small.
If observations are widely dispersed around the mean, the sum of distances will be larger.




The standard deviation is a summary measure of the average distance to the mean.
If there is more dispersion, the standard deviation sy will be higher.

Comparing distributions
If we want to compare different positions in distributions we can use Z-SCORES




Z-score is the amount of standard deviations to the mean.
It is independent of the dispersion of the distribution. It expresses how many standard
deviations we are from the mean.
Z-scores take into account that different distributions might have a different mean and a
different level of dispersion.
A z-score is a standardized measure of the distance from an observation to the mean,
independent of the dispersion of the distribution.
It is useful for inferential statistics.
It all depends on the reference group: importance of context (“relatively”)

, INTRODUCTION TO STATISTICS – Lecture 1 Week 1 – 07/01/2019

On probability, z-scores and distributions

Distribution of data
Data can be distributed in different ways. We can have a skewed distribution or a bell-
shaped distribution. In a perfect bell-shaped distribution, the distribution is perfectly
symmetrical around the mean 𝑦̅. This means that the right and left tail are symmetrical.




Empirical Rule: we can summarize all observations in bell-shaped distributions:
- 68% of all observations is between 𝑦̅ – s and 𝑦̅ + s
- 95,4% of all observations is between 𝑦̅ – 2s and 𝑦̅ + 2s
- 99,7% of all observations is between 𝑦̅ -3s and 𝑦̅ + 3s

Probabilities and probability distributions
We can think of frequency distributions as probability distributions as well. If we pick one
random inhabitant of De Pijp, for example, what is the probability that he/she is older than
35? We can determine this on the basis of the distribution.
The probability p is the area under the curve.
We can apply this to all normal distributions.
We can also apply this and the Empirical Rule to the standard normal distribution which is a
theoretical distribution used in inferential statistics. Empirical distributions are hardly ever
normally distributed. We use the standard normal distribution for calculations.
Characteristics of the standard normal distribution:
- Bell-shaped
- Perfectly symmetrical
- Mean 𝑦̅ = 0 and standard deviation s = 1

Z-scores and probabilities
Probabilities can be defined as z-scores. In the standard normal distribution z = 1 because 𝑦̅
= 0 and s = 1. Every position in a normal distribution has a z-score with a corresponding
probability that we can check in the Z-table. For normally distributed variables we can
convert z-scores to probabilities (and the other way around).

The benefits of buying summaries with Stuvia:

Guaranteed quality through customer reviews

Guaranteed quality through customer reviews

Stuvia customers have reviewed more than 700,000 summaries. This how you know that you are buying the best documents.

Quick and easy check-out

Quick and easy check-out

You can quickly pay through credit card or Stuvia-credit for the summaries. There is no membership needed.

Focus on what matters

Focus on what matters

Your fellow students write the study notes themselves, which is why the documents are always reliable and up-to-date. This ensures you quickly get to the core!

Frequently asked questions

What do I get when I buy this document?

You get a PDF, available immediately after your purchase. The purchased document is accessible anytime, anywhere and indefinitely through your profile.

Satisfaction guarantee: how does it work?

Our satisfaction guarantee ensures that you always find a study document that suits you well. You fill out a form, and our customer service team takes care of the rest.

Who am I buying these notes from?

Stuvia is a marketplace, so you are not buying this document from us, but from seller ilariamonese. Stuvia facilitates payment to the seller.

Will I be stuck with a subscription?

No, you only buy these notes for $3.21. You're not tied to anything after your purchase.

Can Stuvia be trusted?

4.6 stars on Google & Trustpilot (+1000 reviews)

50843 documents were sold in the last 30 days

Founded in 2010, the go-to place to buy study notes for 14 years now

Start selling
$3.21  1x  sold
  • (3)
Add to cart
Added