Get in Touch
 Duration 21 hours

Course Outline

Module 1: Overview of Confluent Apache Kafka Cluster Architecture and Configuration

  • The function of Kafka within modern data pipelines
  • Distinctions between Apache Kafka and Confluent Kafka
  • Essential elements: producers, consumers, brokers, topics, and partitions
  • Strategies for deploying Kafka clusters and considerations for scaling

Module 2: Setting Up Zookeeper Quorum

  • An introduction to Zookeeper
  • The function of Zookeeper within a Kafka cluster
  • Determining the size of the Zookeeper Quorum
  • Configuring Zookeeper parameters
  • Establishing SSH on our servers
  • Hands-on: Configuring Zookeeper (both as a team and as a service)
  • Utilizing the Zookeeper Command Line Interface (CLI)
  • Hands-on: Setting up the Zookeeper Quorum
  • The internal file system structure of Zookeeper
  • Performance factors that impact Zookeeper
  • Demonstration of management tools for Zookeeper and Zoonavigator

Module 3: Configuring the Kafka Cluster

  • Fundamental Kafka concepts
  • Configuring Kafka settings
  • Hands-on: Configuring Kafka brokers
  • Hands-on: Running Kafka commands
  • Hands-on: Setting up a Multi-Broker Kafka Cluster
  • Hands-on: Testing the Kafka cluster
  • Verifying connectivity to the Kafka cluster
  • Configuring advertised.listeners: a critical setting
  • Setting up topic configurations
  • Configuration for downloading and ingesting messages into topics
  • Hands-on: Demonstrating Kafka resilience
  • Kafka performance: I/O optimization
  • Kafka performance: Network (RED) considerations
  • Kafka performance: RAM usage
  • Kafka performance: CPU efficiency
  • Kafka performance: Operating System (OS) impacts
  • Kafka performance: Additional factors
  • Hands-on: Modifying Kafka broker configurations

Module 4: Advanced Kafka Configuration

  • Configuring Landoop Kafka topic interface, Confluent REST Proxy, and Confluent Schema Registry
  • Sending and receiving messages via CLI, Java, and the Spring framework
  • Monitoring metrics using tools such as Confluent Control Center and Elasticsearch
  • Managing log files and offsets
  • Ensuring high availability and disaster recovery
  • Achieving high availability through replication
  • Optimizing producer and consumer performance
  • Implementing disaster recovery strategies
  • Controlling failover and data recovery processes
  • Configuring connectors
  • Implementing Kafka Connect
  • Integrating Kafka security features

Summary and Future Steps

Requirements

  • Knowledge of distributed systems and messaging principles
  • Proficiency with the Linux command line interface
  • Fundamental understanding of networking and system administration

Target Audience

  • System administrators
  • DevOps engineers
  • Platform and infrastructure teams

Number of participants


Price per participant

Testimonials (2)

Upcoming Courses

Related Categories