Harvard University Free Online Course – Data Science: Productivity Tools

Estimated 8 weeks, 1–2 hours per week

A typical data analysis project may involve several parts, each including several data files and different scripts with code. Keeping all this organized can be challenging.

Part of our Professional Certificate Program in Data Science, this course explains how to use Unix/Linux as a tool for managing files and directories on your computer and how to keep the file system organized. You will be introduced to the version control systems git, a powerful tool for keeping track of changes in your scripts and reports. We also introduce you to GitHub and demonstrate how you can use this service to keep your work in a repository that facilitates collaborations.

Finally, you will learn to write reports in R markdown which permits you to incorporate text and code into a document. We’ll put it all together using the powerful integrated desktop environment RStudio.

What you’ll learn

  • How to use Unix/Linux to manage your file system
  • How to perform version control with git
  • How to start a repository on GitHub
  • How to leverage the many useful features provided by RStudio

Official website

1 Response

  1. Peter Telfer says:

    I WORK IN AN ADMINISTRATIVE SETTING, RESPONDING TO LETTERS AND FACILITATING THE TRAVELS OF LEGISLATORS AND GOVERNMENT SECTOR CLIENTS. I’M INTERESTED IN THE STUDY OF DATA SCIENCE. ITS STUDY WILL BE USEFUL TO ME IN A WAY THAT WILL MAKE ME WORK MORE EFFICIENTLY

Leave a Reply

Your email address will not be published. Required fields are marked *

This site uses Akismet to reduce spam. Learn how your comment data is processed.