EdX

Probability and Statistics IV: Confidence Intervals and Hypothesis Tests (edX)

Probability and Statistics IV: Confidence Intervals and Hypothesis Tests (edX)

This course covers two important methodologies in statistics – confidence intervals and hypothesis testing. Confidence intervals allow us to make probabilistic statements such as: “We are 95% sure that Candidate Smith’s popularity is 52% +/- 3%.” Hypothesis testing allows us to pose hypotheses and test their validity in a statistically rigorous way. For instance, “Does a new drug result in a higher cure rate than the old drug”?

Class Deals by MOOC List - Click here and see EdX's Active Discounts, Deals, and Promo Codes.

This course covers two important methodologies in statistics – confidence intervals and hypothesis testing.
Confidence intervals are encountered in everyday life, and allow us to make probabilistic statements such as: “Based on the sample of observations we conducted, we are 95% sure that the unknown mean lies between A and B,” and “We are 95% sure that Candidate Smith’s popularity is 52% +/- 3%.” We begin the course by discussing what a confidence interval is and how it is used. We then formulate and interpret confidence intervals for a variety of probability distributions and their parameters.
Hypothesis testing allows us to pose hypotheses and test their validity in a statistically rigorous way. For instance, “Does a new drug result in a higher cure rate than the old drug” or “Is the mean tensile strength of item A greater than that of item B?” The second half the course begins by motivating hypothesis tests and how they are used. We then discuss with the types of errors that can occur with hypothesis testing, and how to design tests to mitigate those errors. Finally, we formulate and interpret hypothesis tests for a variety of probability distributions and their parameters.
Hypothesis testing allows us to pose hypotheses and test their validity in a statistically rigorous way. For instance, “Does a new drug result in a higher cure rate than the old drug” or “Is the mean tensile strength of item A greater than that of item B?” The second half the course begins by motivating hypothesis tests and how they are used. We then discuss with the types of errors that can occur with hypothesis testing, and how to design tests to mitigate those errors. Finally, we formulate and interpret hypothesis tests for a variety of probability distributions and their parameters.
This course is part of the Statistics, Confidence Intervals and Hypothesis Tests Professional Certificate.

What you'll learn
Upon completion of this course, learners will be able to:

  • Identify what a confidence interval is and how it is used
  • Formulate and interpret confidence intervals for a variety of probability distributions and their parameters
  • Determine what a hypothesis test is and how it is used
  • Identify the types of errors that can occur with hypothesis testing, and how to design tests to mitigate those errors
  • Formulate and interpret hypothesis tests for a variety of probability distributions and their parameters

Prerequisites
Learners will be expected to come in knowing some set theory and basic calculus, as well as the material from the previous courses in this series ( A Gentle Introduction to Probability , Random Variables , and A Gentle Introduction to Statistics ). The prerequisite material is all available for you to access; and in any event, we will make the current course as self-contained as possible. In addition, this course will involve a bit of computer programming, so it would be nice to have at least a little experience in something like Excel and/or the R freeware statistical package.

Syllabus

Module 1: Confidence Intervals
• Lesson 1: Introduction to Confidence Intervals
• Lesson 2: Normal Mean (variance known)
• Lesson 3: Difference of Two Normal Means (variances known)
• Lesson 4: Normal Mean (variance unknown)
• Lesson 5: Difference of Two Normal Means (unknown equal variances)
Module 1 (cont’d): Confidence Intervals
• Lesson 6: Difference of Two Normal Means (variances unknown)
• Lesson 7: Difference of Paired Normal Means (variances unknown)
• Lesson 8: Normal Variance
• Lesson 9: Ratio of Variances of Two Normals
• Lesson 10: Bernoulli Proportion

Module 2: Hypothesis Testing
• Lesson 1: Introduction to Hypothesis Testing
• Lesson 2: The Errors of Our Ways
• Lesson 3: Normal Mean Test with Known Variance
• Lesson 4: Normal Mean Test with Known Variance: Design
• Lesson 5: Two-Sample Normal Means Test with Known Variances
• Lesson 6: Normal Mean Test with Unknown Variance
• Lesson 7: Two-Sample Normal Means Tests with Unknown Variances
Module 2 (cont’d): Hypothesis Testing
• Lesson 8: Two-Sample Normal Means Test with Paired Observations
• Lesson 9: Normal Variance Test
• Lesson 10: Two-Sample Normal Variances Test
• Lesson 11: Bernoulli Proportion Test
• Lesson 12: Two-Sample Bernoulli Proportions Test
• Lesson 13: Goodness-of-Fit Tests: Introduction
• Lesson 14: Goodness-of-Fit Tests: Examples
• Lesson 15 [OPTIONAL]: Goodness-of-Fit Tests: Honors Example

Go to Class
MOOC List is learner-supported. When you buy through links on our site, we may earn an affiliate commission.

Related Courses

Introductory Statistics : Analyzing Data Using Graphs and Statistics (edX) EdX
Seoul National University,SNUx

Introductory Statistics : Analyzing Data Using Graphs and Statistics (edX)

This course teaches basic statistical concepts and explores many compelling applications of statistical methods using real-life applications of Statistics. Why do we study statistics? The field of statistics provides professionals and scientists withconceptual foundations and useful techniques for evaluating ideas, testing theories, and - ultimately -uncovering the truth in any situation.

Self Paced
Self-Paced
Probability and Statistics in Data Science using Python (edX) EdX
University of California, San Diego,UC San DiegoX

Probability and Statistics in Data Science using Python (edX)

Using Python, learn statistical and probabilistic approaches to understand and gain insights from data. The job of a data scientist is to glean knowledge from complex and noisy datasets. Reasoning about uncertainty is inherent in the analysis of noisy data. Probability and Statistics provide the mathematical foundation for such reasoning.

Self Paced
Self-Paced
Data Science: Inference and Modeling (edX) EdX
HarvardX,Harvard University

Data Science: Inference and Modeling (edX)

Learn inference and modeling, two of the most widely used statistical tools in data analysis. Statistical inference and modeling are indispensable for analyzing data affected by chance, and thus essential for data scientists. In this course, you will learn these key concepts through a motivating case study on election forecasting.

Self Paced
Self-Paced
Advanced statistical physics (edX) EdX
École Polytechnique Fédérale de Lausanne,EPFLx

Advanced statistical physics (edX)

We explore statistical physics in both classical and open quantum systems. Additionally, we will cover probabilistic data analysis that is extremely useful in many applications. This course covers non-equilibrium statistical processes and the treatment of fluctuation dissipation relations by Einstein, Boltzmann and Kubo. Moreover, the fundamentals of Markov processes, stochastic differential and Fokker Planck equations, mesoscopic master equation, etc will be treated in detail. Prior knowledge of statistical physics is highly recommended but not required.

Self Paced
Self-Paced
Introductory Statistics : Sample Survey and Instruments for Statistical Inference (edX) EdX
Seoul National University,SNUx

Introductory Statistics : Sample Survey and Instruments for Statistical Inference (edX)

The purpose of this course is to introduce basic concepts of sample surveys and to teach statistical inference process using real-life examples. In this course, you will learn about sample surveys with the concepts of samples and populations. In addition, we will discuss possible problems(bias) of the surveys based on practical examples and concept of probability errors in sampling.

Self Paced
Self-Paced
Mathematical understanding of uncertainty (edX) EdX
Seoul National University,SNUx

Mathematical understanding of uncertainty (edX)

This lecture series discusses how the concept of probability can be used to handle, control, and exploit uncertainty in the real-world. It is an undergraduate-level lecture series on probability, but is entirely different from the usual courses on probability theory. The lectures cover the basics of probability theory including the relevant mathematics, but instead of focusing on mathematics, the lectures explain how probability theory can help understand real-world uncertainty using various examples.

Self Paced
Self-Paced
Statistical Inference and Modeling for High-throughput Experiments (edX) EdX
HarvardX,Harvard University

Statistical Inference and Modeling for High-throughput Experiments (edX)

A focus on the techniques commonly used to perform statistical inference on high throughput data. In this course you’ll learn various statistics topics including multiple testing problem, error rates, error rate controlling procedures, false discovery rates, q-values and exploratory data analysis. We then introduce statistical modeling and how it is applied to high-throughput data. In particular, we will discuss parametric distributions, including binomial, exponential, and gamma, and describe maximum likelihood estimation.

Self Paced
Self-Paced
Data Science: R Basics (edX) EdX
HarvardX,Harvard University

Data Science: R Basics (edX)

Build a foundation in R and learn how to wrangle, analyze, and visualize data. This course will introduce you to the basics of R programming. You can better retain R when you learn it to solve a specific problem, so you’ll use a real-world dataset about crime in the United States. You will learn the R skills needed to answer essential questions about differences in crime across the different states.

Self Paced
Self-Paced