icon
Quality Score

Content Quality
/
Video Quality
/
Qualified Instructor
/
Course Pace
/
Course Depth & Coverage
/

Overall Score : 84 / 100

icon
Course Description

This course is for novice programmers or business people who would like to understand the core tools used to wrangle and analyze big data. With no prior experience, you will have the opportunity to walk through hands-on examples with Hadoop and Spark frameworks, two of the most common in the industry. You will be comfortable explaining the specific components and basic processes of the Hadoop architecture, software stack, and execution environment. In the assignments you will be guided in how data scientists apply the important concepts and techniques such as Map-Reduce that are used to solve fundamental problems in big data. You'll feel empowered to have conversations about big data and the data analysis process.

icon
Instructor Details

placeholder

CoursesCode Free Data ScienceHadoop Platform and Application Framework

icon
More courses by Natasha Balac

Code Free Data Science

Free

icon
More hadoop courses

Apache Kafka Series - Kafka Cluster Setup & Administration

$11.99

Apache Kafka Series - Confluent Schema Registry & REST Proxy

$11.99

Apache Spark for Java Developers

$11.99

Apache Kafka - Real-time Stream Processing (Master Class)

$11.99

Apache Kafka Series - Kafka Monitoring & Operations

$11.99

Apache Kafka for absolute beginners

$11.99

icon
Reviews

4.2

492 total reviews

5 star 4 star 3 star 2 star 1 star
% Complete
% Complete
% Complete
% Complete
% Complete

By ahmedossama.201461 on 3-Nov-16

I actually enjoyed the course, but -1 star for being a little outdated and -1 star for too much content and few practical exercises in the central parts. The structure is well thought though, and the PySpark lessons very entertaining. I would recommend this course for those interested in getting an historical perspective but warning them that MapReduce and Spark RDDs have been superseded with recent developments.

By Keith B on 17-Nov-15

Not enough exercisesToo general

By David R D N on 3-Apr-16

Good concepts but poorly organised exercises!

By Hua W on 14-Jan-16

Course provides basics of Hadoop, MapReduce and Spark. In my opinion program assignments should be more formalised with no dependancy on programming language at all (plain result submission) OR with strict dependency (code submission).

By Amr A on 17-Jan-16

The course overall is good; however, week#2 for example was pretty bad. It felt like a very theoretical style of teaching a course which could have been made hands on. Week #2, which i am in process of completing is SO MUCH theory. Tez, YARN, Spark... difference b/w hadoop and hadoop#2.. Do you want us to cram the differences? Show us how to build something, show us how to set-up a JOB... be practical.

By Aman S on 28-Jun-17

The entire course is pretty good. But the final assignment with the N to N match with Program as the key should really caught attention of the mentor. With N to N match, the only "right" result is to right the very similar Pyspark code just like that write by who was writing the code to generate the answer. However, the N to N match, there is no true answer at all. Therefore, any code other than what the writer wrote should be given partial credit. All in all, a pretty good course, but assignment 15 definitely need to be rewritten to get rid of the N to N match.

By Ayhan Y on 18-Apr-17

Thanks for professors' work.

By Even on 27-Jan-16

Good solid course with a lot of emphasis on hands-on! It would be great if the teachers (particular in the Spark/pySpark section) could provide setup guidance for iPython Notebook as this will save the students a lot of time in coding/re-coding various examples as well as having a complete and easy to overview trace of the various Python related exercises.

By Javier R on 28-Feb-16

Challenging assignments made the course special. Its less theory and more practice and that's exactly what I'm looking for.

By Paresh R on 22-Nov-15

I found this course a lot better than the previous one.More concrete and with interesting material quite applicable on work use cases.I think there are some points of attention to be addressed in next versions of this course.a) Quiz sometime are very hard to be solved because there is no explanation in the course. But quiz are very useful (as the community) to really learn the material.b) some lessons are hard to follow. Probably better to put at the beginning of lessons a mandatory activity to read some material on web (I found very useful links in discussion community, after them I was able to follow the course, but not before)In any case I consider this very useful and I thanks all of you for the effort you put itStefano Priola

By Peter S on 31-Oct-15

This is a good course for anyone without major experience with Hadoop and/or Spark. Covers high level concepts and architecture, and basic tools of each. In this first iteration of the course, there are several typos in the assignments but fellow students have quickly provided corrections that, I'm sure, will be incorporated into subsequent offerings.If you want to learn, I would recommend at least trying this course.

By Ciprian D on 11-Feb-18

Good foundation for distributed file systems. Material dated? Presenters very capable, a little dry.