Get in Touch
 Duration 21 hours

Course Outline

Module 1: Introduction to the architecture and configuration of the Confluent Apache Kafka cluster

  • The role of Kafka in contemporary data pipelines
  • Distinctions between Apache Kafka and Confluent Kafka
  • Core components: producers, consumers, brokers, topics, and partitions
  • Deployment models for Kafka clusters and scaling strategies

Module 2: Zookeeper Quorum Configuration

  • Overview of Zookeeper
  • The function of Zookeeper within a Kafka cluster
  • Determining Zookeeper Quorum size
  • Configuring Zookeeper
  • Setting up SSH on servers
  • Practical exercise: Zookeeper configuration (both as a team and as a service)
  • Utilizing the Zookeeper Command Line Interface (CLI)
  • Practical exercise: Configuring Zookeeper Quorums
  • The internal file system of Zookeeper
  • Performance factors influencing Zookeeper
  • Demonstration of management tools for Zookeeper and Zoonavigator

Module 3: Kafka Cluster Configuration

  • Fundamental Kafka concepts
  • General Kafka configuration settings
  • Practical exercise: Configuring Kafka brokers
  • Practical exercise: Executing Kafka commands
  • Practical exercise: Setting up a Kafka Multi-Broker Cluster
  • Practical exercise: Testing Kafka clusters
  • Verifying connectivity to the Kafka cluster
  • Configuring Advertised.listeners: A critical setting
  • Topic configuration details
  • Settings for downloading and ingesting messages within topics
  • Practical exercise: Demonstrating Kafka resilience
  • Kafka performance: Input/Output (I/O)
  • Kafka performance: Network (RED)
  • Kafka performance: RAM
  • Kafka performance: CPU
  • Kafka performance: Operating System (OS)
  • Kafka performance: Other factors
  • Practical exercise: Modifying Kafka broker configurations

Module 4: Advanced Kafka Configuration

  • Configuring the Landoop Kafka topic interface, Confluent REST Proxy, and Confluent Schema Registry
  • Messaging via CLI, Java, and the Spring framework
  • Monitoring metrics and utilizing tools (such as Confluent Control Center, Elasticsearch, etc.)
  • Managing log files and offsets
  • Implementing high availability and disaster recovery
  • Achieving high availability through replication
  • Optimizing producer and consumer performance
  • Strategies for disaster recovery
  • Controlling failover and recovering data
  • Connector configuration
  • Implementing Kafka Connect
  • Kafka security features

Summary and Next Steps

Requirements

  • Proficiency with distributed systems and messaging paradigms
  • Hands-on experience with the Linux command line interface
  • Foundational knowledge of networking and system administration

Target Audience

  • System administrators
  • DevOps engineers
  • Platform and infrastructure teams

Number of participants


Price per participant

Testimonials (2)

Upcoming Courses

Related Categories