Data collection and cleaning – Universidad Johns Hopkins – Coursera

Contents

This course will cover the basic ways that data can be obtained. The course will cover obtaining data from the web, de API, of databases and colleagues in various formats. It will also cover the basics of data cleansing and how to make data clean. “organized”. Sorted data dramatically speeds up subsequent data analysis tasks. The course will also cover the components of a complete data set that includes raw data., processing instructions, code books and processed data. The course will cover the basics needed to collect, clean and share data.

Videos of lectures and weekly quizzes and a final peer-reviewed project.

As part of this class, you will be asked to set up a GitHub account. GitHub is a tool for collaborative code editing and sharing. During this course and other courses of the specialization, send links to files that you post publicly on your GitHub account as part of peer review. If you are concerned about preserving your anonymity, you should set up an anonymous GitHub account and be careful not to include any information that you don't want to be available to peer testers.

Course program:

Upon completion of this course, you will be able to get data from a range of sources. You will know the principles of ordered data and data exchange. To end, understand and be able to apply the basic tools for cleaning and manipulating data.

Rate:

For free

Duration:

4 weeks (4-9 hours / week)

Important date:

6 April 2015-4 May 2015

Subscribe to our Newsletter

We will not send you SPAM mail. We hate it as much as you.

Datapeaker