Get in Touch
 Duration 21 hours

Course Outline

Module 1: Introduction to Confluent Apache Kafka Cluster Architecture and Configuration

  • The role of Kafka in modern data pipelines
  • Key distinctions between Apache Kafka and Confluent Kafka
  • Essential components: producers, consumers, brokers, topics, and partitions
  • Deployment models and scaling considerations for Kafka clusters

Module 2: Configuring Zookeeper Quorum

  • Overview of Zookeeper
  • The function of Zookeeper within a Kafka cluster
  • Determining Zookeeper Quorum size
  • Configuring Zookeeper parameters
  • Implementing SSH on server infrastructure
  • Practical exercise: Configuring Zookeeper (both as a team and a service)
  • Utilizing the Zookeeper Command Line Interface (CLI)
  • Practical exercise: Setting up Zookeeper Quorum
  • The Zookeeper internal file system
  • Performance factors impacting Zookeeper
  • Demonstration of Zookeeper management tools and Zoonavigator

Module 3: Kafka Cluster Configuration

  • Foundational Kafka concepts
  • General Kafka configuration settings
  • Practical exercise: Configuring Kafka brokers
  • Practical exercise: Executing Kafka commands
  • Practical exercise: Configuring a Multi-Broker Kafka Cluster
  • Practical exercise: Testing the Kafka cluster
  • Verifying connectivity to the Kafka cluster
  • Configuring Advertised.listeners: a critical setting
  • Topic-specific configurations
  • Settings for downloading and ingesting messages into topics
  • Practical exercise: Demonstrating Kafka resilience
  • Kafka performance: I/O operations
  • Kafka performance: Network (RED)
  • Kafka performance: RAM usage
  • Kafka performance: CPU utilization
  • Kafka performance: Operating System (OS) interactions
  • Kafka performance: Other factors
  • Practical exercise: Modifying Kafka broker configurations

Module 4: Advanced Kafka Configuration

  • Configuring Landoop Kafka topic UI, Confluent REST Proxy, and Confluent Schema Registry
  • Message sending and receiving methods (via CLI, Java, and the Spring framework)
  • Monitoring metrics and tools (including Confluent Control Center, Elasticsearch, etc.)
  • Management of log files and offsets
  • High availability and disaster recovery principles
  • Achieving high availability through replication
  • Optimizing producer and consumer performance
  • Strategies for disaster recovery
  • Controlling failover and recovering data
  • Configuring Kafka Connectors
  • Implementation of Kafka Connect
  • Security features in Kafka

Summary and Next Steps

Requirements

  • Knowledge of distributed systems and messaging frameworks
  • Proficiency with the Linux command line
  • Fundamental understanding of networking and system administration

Target Audience

  • System administrators
  • DevOps engineers
  • Platform and infrastructure teams

Number of participants


Price per participant

Testimonials (2)

Upcoming Courses

Related Categories