To wersja testowa nowego serwisu infoShare Academy — wyświetlane treści i oferta nie są wiążące ani kompletne

Data

Hadoop

The Hadoop Training is an intensive two-day course focused on the practical application of this popular framework for processing and analyzing large datasets.

Duration
16h · 2 days
Who it's for

Ideal for teams that…

1 Developers and data engineers who want to expand their skills with Hadoop.
2 Data scientists and data analysts aiming to process large datasets efficiently.
3 IT and big data specialists who want to leverage Hadoop in their projects.
Outcomes after the program

Hands-on AI and data analytics workshops — built around your team's real cases.

How to effectively manage data in HDFS and create MapReduce tasks.

How to process and analyze data using Hive and Pig.

How to optimize MapReduce tasks and manage Hadoop resources.

How to deploy and monitor Hadoop applications in a production environment.

Requirements

What you should know before we start

  • Basic knowledge of programming in Java or Python.
  • Basic understanding of data processing.
  • Ability to work in a Unix/Linux environment.
Program

What we actually do

Day 1: Basics of Hadoop and Data Processing

  • · Overview of main Hadoop components: HDFS, MapReduce, YARN
  • · Interaction between components
  • · Managing files in HDFS
  • · Creating and running basic MapReduce tasks
  • · Hive: table structure and SQL queries
  • · Analyzing file structure for Hive
  • · Pig: introduction to Pig Latin scripts
  • · Implementing a simple MapReduce task
  • · Analyzing results and optimizing the task

Day 2: Advanced Techniques and Practical Applications

  • · Writing advanced Hive queries
  • · Creating complex Pig scripts
  • · Techniques for optimizing MapReduce tasks
  • · Managing resources in a Hadoop cluster
  • · Implementing Hive queries on real datasets
  • · Creating Pig scripts for data processing
  • · Preparing and deploying Hadoop applications
  • · Monitoring and managing Hadoop clusters in production
  • · Controlling and optimizing costs associated with Hadoop data processing
Every module is adapted to your stack and context. The above is a starting point — not a fixed agenda.
How we work

From brief to retro in 30 days.

01

Brief & diagnosis

A call with the team lead + a short survey for participants. We define goals, gap and context.

02

Program customization

We adapt modules, case studies and code examples to your stack. Approval in 5 days.

03

Workshop

Trainer-led sessions, hands-on, code review. Mentor available between sessions too.

04

Retro + report

Outcome report for the team and lead. 30 days of consulting included.

Inquiry

Send a brief. We'll reply within 1 day.

After a short brief we'll prepare a program and a quote. No obligations — it's just a starting point.

Quote within 48h of the brief
First session within 30 days
Pilot before the full decision
VAT invoice, payment in instalments possible

How we handle your data

We process your business data (name, work e-mail, phone number, company, job title) in order to handle your corporate training inquiry and prepare an offer. The data controller is infoShare Academy Sp. z o.o., Al. Grunwaldzka 472B, 80-309 Gdańsk. Providing the data is voluntary but necessary to receive a response. You have the right to access, rectify, erase or restrict the processing of your data and to object to it. Full information is available in our data processing notice.

Optional marketing consents

The controller of your personal data is infoShare Academy sp. z o.o. The rules for processing personal data are set out in the Privacy Policy and the Data Processing Notice.