Data Science: Linear Regression

edX Data Science: Linear Regression

Platform
edX
Provider
Harvard University
Effort
2-4 hours a week
Length
4 weeks
Language
English
Credentials
Paid Certificate Available
Part of
Course Link
Overview
Linear regression is commonly used to quantify the relationship between two or more variables. It is also used to adjust for confounding. This course, part of our Professional Certificate Program in Data Science, covers how to implement linear regression and adjust for confounding in practice using R.

In data science applications, it is very common to be interested in the relationship between two or more variables. The motivating case study we examine in this course relates to the data-driven approach used to construct baseball teams described in the book (also a movie) Moneyball. We will try to determine which measured outcomes best predict baseball runs and to do this we'll use linear regression.

We will also examine confounding, where extraneous variables affect the relationship between two or more other variables, leading to spurious associations. Linear regression is a powerful technique for removing confounders, but it is not a magical process, and it is essential to understand when it is appropriate to use. You will learn when to use it in this course.

What You Will Learn
  • How linear regression was originally developed by Galton
  • What is confounding and how to detect it
  • How to examine the relationships between variables by implementing linear regression in R
Taught by
Rafael Irizarry
Author
edX
Views
882
First release
Last update
Rating
0.00 star(s) 0 ratings
Top