Get in Touch

Course Outline

Module 1: Introduction to the architecture and configuration of the Confluent Apache Kafka cluster

  • The role of Kafka in modern data pipelines
  • Distinguishing between Apache Kafka and Confluent Kafka
  • Core components: producers, consumers, brokers, topics, and partitions
  • Kafka cluster deployment models and scaling considerations

Module 2: Zookeeper Quorum Configuration

  • Understanding Zookeeper
  • The role of Zookeeper within a Kafka cluster
  • Determining Zookeeper Quorum size
  • Zookeeper configuration best practices
  • Implementing SSH on servers
  • Practical exercise: Configuring Zookeeper as a team and as a service
  • Utilizing the Zookeeper Command Line Interface (CLI)
  • Practical exercise: Setting up the Zookeeper Quorum
  • Overview of the Zookeeper internal file system
  • Performance factors influencing Zookeeper
  • Demonstration of management tools for Zookeeper and Zoonavigator

Module 3: Kafka Cluster Configuration

  • Fundamental Kafka concepts
  • General Kafka configuration
  • Practical exercise: Configuring Kafka brokers
  • Practical exercise: Executing Kafka commands
  • Practical exercise: Configuring a Multi-Broker Kafka Cluster
  • Practical exercise: Testing Kafka clusters
  • Verifying connectivity to your Kafka cluster
  • Advertised.listeners configuration: A critical setting
  • Topic configuration
  • Configuring message downloading and ingestion into topics
  • Practical exercise: Demonstrating Kafka resilience
  • Kafka performance: I/O operations
  • Kafka performance: Network (RED)
  • Kafka performance: RAM
  • Kafka performance: CPU
  • Kafka performance: Operating System (OS)
  • Kafka performance: Other factors
  • Practical exercise: Modifying Kafka broker configuration

Module 4: Advanced Kafka Configuration

  • Landoop Kafka topic user interface, Confluent REST Proxy, and Confluent Schema Registry configuration
  • Sending and receiving messages via CLI, Java, and the Spring framework
  • Monitoring metrics and tools (including Confluent Control Center, Elasticsearch, and more)
  • Managing log files and offsets
  • High availability and disaster recovery strategies
  • Achieving high availability through replication
  • Optimizing producer and consumer performance
  • Disaster recovery planning
  • Failover control and data recovery
  • Connector configuration
  • Implementing Kafka Connect
  • Kafka security features

Summary and Next Steps

Requirements

  • Working knowledge of distributed systems and messaging concepts
  • Proficiency with the Linux command line
  • Foundational understanding of networking and system administration

Target Audience

  • System administrators
  • DevOps engineers
  • Platform and infrastructure teams
 21 Hours

Testimonials (2)

Related Categories