Get in Touch
 Duration 21 hours

Course Outline

Module 1: Overview of Confluent Apache Kafka Cluster Architecture and Configuration

  • The role of Kafka in contemporary data pipelines
  • Distinguishing between Apache Kafka and Confluent Kafka
  • Key components: producers, consumers, brokers, topics, and partitions
  • Deployment models and scaling considerations for Kafka clusters

Module 2: Configuring Zookeeper Quorums

  • An introduction to Zookeeper
  • The function of Zookeeper within a Kafka cluster
  • Determining appropriate Zookeeper Quorum sizes
  • Configuring Zookeeper settings
  • Setting up SSH on servers
  • Practical exercise: Configuring Zookeeper as a team and service
  • Utilizing the Zookeeper Command Line Interface (CLI)
  • Practical exercise: Configuring the Zookeeper Quorum
  • Understanding the Zookeeper internal file system
  • Performance factors impacting Zookeeper
  • Demonstration of Zookeeper management tools and Zoonavigator

Module 3: Configuring the Kafka Cluster

  • Fundamental Kafka concepts
  • General Kafka configuration settings
  • Practical exercise: Configuring Kafka brokers
  • Practical exercise: Running Kafka commands
  • Practical exercise: Setting up a multi-broker Kafka cluster
  • Practical exercise: Testing the Kafka cluster
  • Verifying connectivity to the Kafka cluster
  • Configuring Advertised.listeners: A critical setting
  • Topic-specific configurations
  • Settings for downloading and ingesting messages into topics
  • Practical exercise: Demonstrating Kafka resilience
  • Kafka performance: I/O operations
  • Kafka performance: Network (RED)
  • Kafka performance: RAM utilization
  • Kafka performance: CPU usage
  • Kafka performance: Operating System (OS) considerations
  • Kafka performance: Other factors
  • Practical exercise: Modifying Kafka broker configurations

Module 4: Advanced Kafka Configuration

  • Configuring the Landoop Kafka topic UI, Confluent REST Proxy, and Confluent Schema Registry
  • Messaging workflows via CLI, Java, and the Spring framework
  • Monitoring metrics and utilizing tools such as Confluent Control Center and Elasticsearch
  • Managing log files and offsets
  • Strategies for high availability and disaster recovery
  • Achieving high availability through replication
  • Optimizing producer and consumer performance
  • Formulating disaster recovery strategies
  • Controlling failover and executing data recovery
  • Configuring connectors
  • Implementing Kafka Connect
  • Exploring Kafka security features

Course Summary and Recommended Next Steps

Requirements

  • Working knowledge of distributed systems and messaging concepts
  • Proficiency with the Linux command line
  • Fundamental understanding of networking and system administration

Target Audience

  • System administrators
  • DevOps engineers
  • Platform and infrastructure teams

Number of participants


Price per participant

Testimonials (2)

Upcoming Courses

Related Categories