Get in Touch
 Duration 21 hours

Course Outline

Module 1: Overview of Confluent Apache Kafka Cluster Architecture and Configuration

  • The function of Kafka within modern data pipelines
  • Distinguishing between Apache Kafka and Confluent Kafka
  • Essential elements: producers, consumers, brokers, topics, and partitions
  • Deployment strategies for Kafka clusters and scaling factors

Module 2: Setting Up the Zookeeper Quorum

  • Introduction to Zookeeper
  • The function of Zookeeper within a Kafka cluster
  • Determining the size of the Zookeeper Quorum
  • Zookeeper configuration parameters
  • Setting up SSH on server environments
  • Practical Exercise: Configuring Zookeeper as both a team and a service
  • Utilizing the Zookeeper Command Line Interface (CLI)
  • Practical Exercise: Configuring the Zookeeper Quorum
  • The internal file system of Zookeeper
  • Performance variables impacting Zookeeper
  • Demonstration of Zookeeper management tools and Zoonavigator

Module 3: Configuring the Kafka Cluster

  • Fundamental Kafka concepts
  • General Kafka configuration settings
  • Practical Exercise: Configuring Kafka brokers
  • Practical Exercise: Running Kafka commands
  • Practical Exercise: Setting up a Multi-Broker Kafka Cluster
  • Practical Exercise: Testing the Kafka cluster
  • Verifying connectivity to the Kafka cluster
  • Advertised.listeners setting: a critical configuration parameter
  • Topic-specific configurations
  • Settings for downloading and ingesting topic messages
  • Practical Exercise: Demonstrating Kafka resilience
  • Kafka performance: I/O operations
  • Kafka performance: Network (RED) factors
  • Kafka performance: RAM utilization
  • Kafka performance: CPU usage
  • Kafka performance: Operating System (OS) interactions
  • Kafka performance: Additional factors
  • Practical Exercise: Modifying Kafka broker configurations

Module 4: Advanced Kafka Configuration

  • Configuring Landoop Kafka topic UI, Confluent REST Proxy, and Confluent Schema Registry
  • Message transmission and reception via CLI, Java, and the Spring framework
  • Monitoring metrics and utilizing tools such as Confluent Control Center and Elasticsearch
  • Managing log files and offsets
  • Establishing high availability and disaster recovery protocols
  • Achieving high availability through replication
  • Optimizing producer and consumer performance
  • Disaster recovery methodologies
  • Controlling failover and recovering data
  • Setting up connectors
  • Implementing Kafka Connect
  • Security features in Kafka

Summary and Recommended Next Steps

Requirements

  • Working knowledge of distributed systems and messaging principles
  • Proficiency with the Linux command line
  • Foundational understanding of networking and system administration

Target Audience

  • System administrators
  • DevOps engineers
  • Platform and infrastructure teams

Number of participants


Price per participant

Testimonials (2)

Upcoming Courses

Related Categories