Get in Touch
 Duration 21 hours (3 days)

Course Outline

Module 1: Introduction to Confluent Apache Kafka Cluster Architecture and Configuration

  • The role of Kafka in modern data pipelines
  • Distinctions between Apache Kafka and Confluent Kafka
  • Essential components: producers, consumers, brokers, topics, and partitions
  • Deployment models for Kafka clusters and scaling strategies

Module 2: Configuring the Zookeeper Quorum

  • Overview of Zookeeper
  • The function of Zookeeper within a Kafka cluster
  • Determining Zookeeper Quorum size
  • Zookeeper configuration processes
  • Setting up SSH on servers
  • Practical exercise: Configuring Zookeeper (as a team and as a service)
  • Utilizing the Zookeeper Command Line Interface (CLI)
  • Practical exercise: Configuring the Zookeeper Quorum
  • The internal file system of Zookeeper
  • Performance variables impacting Zookeeper
  • Demonstration of management tools for Zookeeper and Zoonavigator

Module 3: Configuring the Kafka Cluster

  • Fundamental Kafka concepts
  • Configuring Kafka settings
  • Practical exercise: Setting up Kafka brokers
  • Practical exercise: Running Kafka commands
  • Practical exercise: Configuring a Kafka Multi-Broker Cluster
  • Practical exercise: Testing the Kafka cluster
  • Verifying connectivity to the Kafka cluster
  • The Advertised.listeners setting: a critical configuration
  • Topic configuration details
  • Settings for downloading and ingesting messages within topics
  • Practical exercise: Demonstrating Kafka resilience
  • Kafka performance: I/O operations
  • Kafka performance: Network (RED)
  • Kafka performance: RAM utilization
  • Kafka performance: CPU usage
  • Kafka performance: Operating System (OS) impact
  • Kafka performance: Other factors
  • Practical exercise: Modifying Kafka broker configurations

Module 4: Advanced Kafka Configuration

  • Configuring the Landoop Kafka topic user interface, Confluent REST Proxy, and Confluent Schema Registry
  • Message transmission and reception methods (CLI, Java, and Spring framework)
  • Metric monitoring and tooling (Confluent Control Center, Elasticsearch, etc.)
  • Management of log files and offsets
  • Ensuring high availability and disaster recovery
  • Achieving high availability through replication
  • Optimizing producer and consumer performance
  • Strategies for disaster recovery
  • Controlling failover and recovering data
  • Connector configuration
  • Implementing Kafka Connect
  • Security features in Kafka

Summary and Next Steps

Requirements

  • A solid grasp of distributed systems and messaging principles.
  • Proficiency with the Linux command line.
  • Fundamental knowledge of networking and system administration.

Target Audience

  • System administrators.
  • DevOps engineers.
  • Platform and infrastructure teams.

Testimonials (2)

Related Categories