The Ultimate Hands On Hadoop Tame your Big Data (Udemy.com)

Hadoop tutorial with MapReduce, HDFS, Spark, Flink, Hive, HBase, MongoDB, Cassandra, Kafka + more! Over 25 technologies.

Created by: Frank Kane

Produced in 2021

icon
What you will learn

  • Design distributed systems that manage "big data" using Hadoop and related technologies.
  • Use HDFS and MapReduce for storing and analyzing data at scale.
  • Use Pig and Spark to create scripts to process data on a Hadoop cluster in more complex ways.
  • Analyze relational data using Hive and MySQL
  • Analyze non-relational data using HBase, Cassandra, and MongoDB
  • Query data interactively with Drill, Phoenix, and Presto
  • Choose an appropriate data storage technology for your application Understand how Hadoop clusters are managed by YARN, Tez, Mesos, Zookeeper, Zeppelin, Hue, and Oozie.
  • Publish data to your Hadoop cluster using Kafka, Sqoop, and Flume
  • Consume streaming data using Spark Streaming, Flink, and Storm

icon
Quality Score

Content Quality
/
Video Quality
/
Qualified Instructor
/
Course Pace
/
Course Depth & Coverage
/

Overall Score : 90 / 100

icon
Course Description

The world of Hadoop and "Big Data" can be intimidating - hundreds of different technologies with cryptic names form the Hadoop ecosystem. With this Hadoop tutorial, you'll not only understand what those systems are and how they fit together - but you'll go hands-on and learn how to use them to solve real business problems!

Learn and master the most popular big data technologies in this comprehensive course, taught by a former engineer and senior manager from Amazon and IMDb. We'll go way beyond Hadoop itself, and dive into all sorts of distributed systems you may need to integrate with.

Install and work with a real Hadoop installation right on your desktop with Hortonworks (now part of Cloudera) and the Ambari UI

Manage big data on a cluster with HDFS and MapReduce

Write programs to analyze data on Hadoop with Pig and Spark

Store and query your data with Sqoop, Hive, MySQL, HBase, Cassandra, MongoDB, Drill, Phoenix, and Presto

Design real-world systems using the Hadoop ecosystem

Learn how your cluster is managed with YARN, Mesos, Zookeeper, Oozie, Zeppelin, and Hue

Handle streaming data in real time with Kafka, Flume, Spark Streaming, Flink, and Storm

Understanding Hadoop is a highly valuable skill for anyone working at companies with large amounts of data.

Almost every large company you might want to work at uses Hadoop in some way, including Amazon, Ebay, Facebook, Google, LinkedIn, IBM, Spotify, Twitter, and Yahoo! And it's not just technology companies that need Hadoop; even the New York Times uses Hadoop for processing images.

This course is comprehensive, covering over 25 different technologies in over 14 hours of video lectures. It's filled with hands-on activities and exercises, so you get some real experience in using Hadoop - it's not just theory.

You'll find a range of activities in this course for people at every level. If you're a project manager who just wants to learn the buzzwords, there are web UI's for many of the activities in the course that require no programming knowledge. If you're comfortable with command lines, we'll show you how to work with them too. And if you're a programmer, I'll challenge you with writing real scripts on a Hadoop system using Scala, Pig Latin, and Python.

You'll walk away from this course with a real, deep understanding of Hadoop and its associated distributed systems, and you can apply Hadoop to real-world problems. Plus a valuable completion certificate is waiting for you at the end!

Please note the focus on this course is on application development, not Hadoop administration. Although you will pick up some administration skills along the way.

Knowing how to wrangle "big data" is an incredibly valuable skill for today's top tech employers. Don't be left behind - enroll now!



"The Ultimate Hands-On Hadoop... was a crucial discovery for me. I supplemented your course with a bunch of literature and conferences until I managed to land an interview. I can proudly say that I landed a job as a Big Data Engineer around a year after I started your course. Thanks so much for all the great content you have generated and the crystal clear explanations. " - Aldo Serrano

"I honestly wouldnt be where I am now without this course. Frank makes the complex simple by helping you through the process every step of the way. Highly recommended and worth your time especially the Spark environment. This course helped me achieve a far greater understanding of the environment and its capabilities. Frank makes the complex simple by helping you through the process every step of the way. Highly recommended and worth your time especially the Spark environment." - Tyler Buck

Who this course is for:
Software engineers and programmers who want to understand the larger Hadoop ecosystem, and use it to store, analyze, and vend "big data" at scale.
Project, program, or product managers who want to understand the lingo and high-level architecture of Hadoop.
Data analysts and database administrators who are curious about Hadoop and how it relates to their work.
System architects who need to understand the components available in the Hadoop ecosystem, and how they fit together.

icon
Instructor Details

Frank Kane

Frank spent 9 years at Amazon and IMDb, developing and managing the technology that automatically delivers product and movie recommendations to hundreds of millions of customers, all the time. Frank holds 17 issued patents in the fields of distributed computing, data mining, and machine learning. In 2012, Frank left to start his own successful company, Sundog Software, which focuses on virtual reality environment technology, and teaching others about big data analysis.

icon
More courses by Frank Kane

Taming Big Data with MapReduce and Hadoop Hands On

$11.99

icon
More hadoop courses

Apache Kafka Series - Kafka Cluster Setup & Administration

$11.99

Apache Kafka Series - Confluent Schema Registry & REST Proxy

$11.99

Apache Spark for Java Developers

$11.99

Apache Kafka - Real-time Stream Processing (Master Class)

$11.99

Apache Kafka Series - Kafka Monitoring & Operations

$11.99

Apache Kafka for absolute beginners

$11.99

icon
Reviews

4.5

300 total reviews

5 star 4 star 3 star 2 star 1 star
% Complete
% Complete
% Complete
% Complete
% Complete

By Amit Kumar on 11/14/2020

Its a great course; Many technologies have been covered under one single roof. Rest sky is the limit for how much you can learn from the instructor. I would really recommend this course for all the Big Data and Analytics Professionals.

By Hamidreza Ahady Dolatsara on 11/14/2020

You didn't have good instructions or your materials are old! I have questions that when I see in your Q&A they are not responded correctly. It's not a good tutorial at all.

By Deepak Shivani on 11/7/2020

It was really good to start with.

By Julian Negele on 11/7/2020

The course gave me a really good overview over the Hadoop Ecosystem and how it's build. Really recommend it to people who are new to this topic!

By Raghavendra Kosigi on 10/22/2020

Yes.This course is planned in a very nice way...
Only suggestion is to add some real world scenario's or used cases and provide a solution for same.This would help everyone a lot.
If we could add AWS EMR also then it would be a great help.

By Jeremy Pedersen on 10/21/2020

This was fun, interactive, and a great way to start off playing with Hadoop. ^_^

By Bruno Guimares on 10/20/2020

Excelente curso para ter uma viso geral de todas as tecnologias envolvidas no ecossistema de Big Data. S no dou 5 estrelas pois as verses utilizadas j esto um pouco antigas, mas no compromete o aprendizado.

By Mercia Malan on 10/13/2020

Really enjoyed this, the instructor is very knowledgeable and there was a good balance between bredth and depth of the hadoop ecosystem. Thanks!

By Joseph Nikhil Reddy Yeruva on 10/6/2020

Need more hands on projects and if you add more examples it would be great.

By Louis DeStefano on 10/4/2020

course has been very informative

By Tobias Anhuser on 9/27/2020

I am aiming for a career in Data Analysis/Science and I took this course mainly to broaden my knowledge in Hadoop, Big Data and all the other buzzwords :) I learned A LOT, that's for sure. Even though I got lost sometimes (particularly during the Streaming topic), I could follow most of the lectures and finish up most of the scripts/exercises on my little sandbox. Considering that I neither have a professional IT background nor do I intend to become full Data Engineer, I wasn't freaking out if I didn't understand every little detail. For some modules I could grasp the general idea of it and that's fine considering my background and intentions. A tiny criticism: some scripts/modules do not work straight away and require some look ups in the Q&A section and a lot of try-and-error takes. However, considering the number of used technologies (which are somehow all inter-connected but updated individually) this seems an issue that cannot be prevented. Nonetheless, awesome course, I learned a lot, totally recommend it!

By Wagal Prasad Wagle on 9/25/2020

You are awesome .Iam still learning from the basic the course is great excellent very much depth and hands on practical. Iam working as a Big data engineering so please help me further to master the course Hadoop.Your way of teaching is excellent and very useful