## What you'll learn

- How linear regression was originally developed by Galton
- What is confounding and how to detect it
- How to examine the relationships between variables by implementing linear regression in R

## Course description

Linear regression is commonly used to quantify the relationship between two or more variables. It is also used to adjust for confounding. This course covers how to implement linear regression and adjust for confounding in practice using R.

In data science applications, it is very common to be interested in the relationship between two or more variables. The motivating case study we examine in this course relates to the data-driven approach used to construct baseball teams described in the book (also a movie) Moneyball. We will try to determine which measured outcomes best predict baseball runs and to do this we'll use linear regression.

We will also examine confounding, where extraneous variables affect the relationship between two or more other variables, leading to spurious associations. Linear regression is a powerful technique for removing confounders, but it is not a magical process and it is important to understand when it is appropriate to use. You will learn when to use it in this course.

HarvardX has partnered with DataCamp for all assignments. This allows students to program directly in a browser-based interface. You will not need to download any special software, but an up-to-date browser is recommended.

This course is part of the HarvardX Data Science Professional Certificate program.

## Associated Schools

### Harvard T.H. Chan School of Public Health

## You may also like

- Learn the basics of machine learning, the science behind the most popular and successful data science techniques, as you build a...FreeAvailable now4 weeks
- In this capstone course, show what you’ve learned from the HarvardX Data Science series. You will have the opportunity to create...FreeAvailable now2 weeks
- In this course, we review use cases and challenges of three interrelated areas in computer science: big data, the Internet of...$2,750+StartsJan 28