Harvard University Free Course – Data Science: Productivity Tools

Keep your projects organized and produce reproducible reports using GitHub, git, Unix/Linux, and RStudio.

What you’ll learn

  • How to use Unix/Linux to manage your file system
  • How to perform version control with git
  • How to start a repository on GitHub
  • How to leverage the many useful features provided by RStudio

Course outline

A standard data analysis project may consist of multiple components, each containing a number of data files and code scripts. It can be difficult to stay organized with everything.

This course, which is a component of our Professional Certificate Program in Data Science, teaches you how to manage files and directories on your computer using Unix/Linux and how to maintain an orderly file system. The version control system git, a potent tool for monitoring changes in your scripts and reports, will be introduced to you. We also give you an overview of GitHub and show you how to utilize it to store your work in a collaborative repository.

Lastly, you will discover how to compose reports using R markdown, which enables you to combine code and text in one document. We’ll use the robust integrated desktop environment RStudio to bring it all together.