Hadoop Administration Training

In this Hadoop Administration training class, students learn all about working with Hadoop and HDFS.

Goals
  1. Learn the fundamental concepts of Hadoop.
  2. Learn to plan your Hadoop cluster.
  3. Learn HDFS features.
  4. Learn how to get data into HDFS.
  5. Learn to work with MapReduce.
  6. Learn installation and configuration of Hadoop.
  7. Learn cluster maintenance.
Outline
  1. Hadoop Overview
    1. What is Big Data?
    2. How did we get to this point?
    3. How does Hadoop compare to a relational database system?
    4. Big Data Introduction
    5. History
    6. Comparison to Relational Databases
    7. Hadoop Ecosystem
  2. HDFS
    1. Architecture/Concepts
    2. Access
    3. Namenodes
    4. Filesystem Shell
    5. Accessing HDFS with Java
    6. Reading/Writing/Browsing file system
  3. HBASE
    1. Overview
    2. Architecture
    3. Data Model
    4. Installation and Shell
    5. Access via Java API
    6. Administration access via Java
    7. Scan API
    8. Filters
    9. Storage Model
    10. Table Design
  4. Map Reduce on YARN
    1. Introduction
    2. Processing Model
    3. Command line tools
    4. MapReduce framework
    5. Submitting MapReduce Jobs
    6. Writing MapReduce jobs in Java
    7. MapReduce Theory
    8. Distributive Cache
    9. Speculative Executin
    10. YARN Components
    11. Counters
    12. Details of MapReduce Job Execution
  5. Hadoop Streaming
    1. Implementing a streaming job
    2. Counters in streaming jobs
    3. Contrast with Java Jobs
  6. MapReduce Workflows
    1. Problem decomposition into MapReduce Jobs
    2. Coding workflows
    3. Using the JobControl Class
  7. Oozie
    1. Oozie Installation
    2. Writing Oozie workflows
    3. Deploying and running Oozie jobs
  8. Pig
    1. Installation
    2. Pig Latin
    3. Writing Pig Scripts
    4. User Defined functions
    5. Data set joins
  9. Hive
    1. Installation
    2. Table creation and deletion
    3. Partitioning
    4. Loading data into Hive
    5. Joins
    6. Bucketing
Class Materials

Each student in our Live Online and our Onsite classes receives a comprehensive set of materials, including course notes and all the class examples.

Class Prerequisites

Experience in the following is required for this Hadoop class:

  • Basic Java Knowledge.

Experience in the following would be useful for this Hadoop class:

  • Experience with Eclipse.

Training for your Team

Length: 4 Days
  • Private Class for your Team
  • Online or On-location
  • Customizable
  • Expert Instructors

What people say about our training

The instructor was excellent! She took her time and answered all questions. If she did not have the answer, she found it.
Nita Worley
USDA, AMS, ITS
I hadn't taken training over the web before, but I will be recommending it to my colleagues. Worked well from the UK.
Antony Edwards
Amari Metals Ltd
I found the class to be very informative. The instuctor was amazing. I will go back to Webucator with out a doubt the next time I require any software training.
Tammy Haines
Creo Marketing Inc.
The Introduction to Java Training was a great way to bring me up to speed with OO programming in Java. There were plenty of examples and the instructor was always willing to give guidance and help on any questions.
Kingston Ko
MEDecision, Inc.

No cancelation for low enrollment

Certified Microsoft Partner

Registered Education Provider (R.E.P.)

GSA schedule pricing

61,857

Students who have taken Instructor-led Training

11,794

Organizations who trust Webucator for their Instructor-led training needs

100%

Satisfaction guarantee and retake option

9.29

Students rated our trainers 9.29 out of 10 based on 28,776 reviews

Contact Us or call 1-877-932-8228