---
tags:
- devops
- l1
- flashcard-deck
- kafka
---
<!-- wiki:breadcrumb:start -->
[Portal](../../../../library/portal/index.md) | **Level:** [L1: Foundations](../../../../library/portal/levels.md) | **Topics:** [Kafka](../../../../library/portal/topics.md) | **Domain:** DevOps & Tooling
<!-- wiki:breadcrumb:end -->

id	category	difficulty	tags	question	answer	source_path
kafka/00460f4953a3	kafka	medium	kafka,performance,broker	Discuss the scenarios where Kafka is a better choice than traditional messaging systems.	Kafka is a preferred choice in several scenarios compared to traditional messaging systems:\n* Scalability: Kafka's distributed architecture allows for horizontal scaling, accommodating large volumes of data and high message throughput.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/033-discuss-the-scenarios-where-kafka-is-a-better-choi.txt
kafka/0114c291bb9e	kafka	medium	kafka,broker	Explain the role of Kafka in a distributed system.	In a distributed system, Kafka serves as a distributed messaging system that enables communication and data exchange between different components or services. Its key roles include:\n* Data Streaming: Kafka facilitates the streaming of real-time data between distributed components, allowing seamless communication in a decoupled manner.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/006-explain-the-role-of-kafka-in-a-distributed-system.txt
kafka/0781a2cc9c45	kafka	easy	kafka,performance,broker	What is Apache Kafka and what problems does it solve?	"[kafka.apache.org](https://kafka.apache.org): ""Apache Kafka is an open-source distributed event streaming platform used by thousands of companies for high-performance data pipelines, streaming analytics, data integration, and mission-critical applications.""\n\nIn other words, Kafka is a sort of distributed log where you can store events, read them and distribute them to different services and do it in high-scale and real-time."\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/001-what-is-kafka.txt
kafka/0ae9ad61a856	kafka	hard	kafka,producer,consumer,partition	How does Kafka handle message ordering within a partition?	Kafka ensures strict ordering of messages within a partition. Each partition maintains a sequential log of messages, and each message is assigned a unique offset. Producers sequentially append messages to the end of the log, and consumers read messages in the order of their offsets. The ordering is guaranteed within a partition, but across partitions, there is no guaranteed global order.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/031-how-does-kafka-handle-message-ordering-within-a-pa.txt
kafka/11a9ce9f97ff	kafka	hard	kafka, control-flow, networking	Discuss Kafka’s support for different message delivery semantics.	Kafka supports different message delivery semantics to cater to various application requirements. The main delivery semantics include:\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/069-discuss-kafkas-support-for-different-message-deliv.txt
kafka/1a8be1821cd0	kafka	easy	kafka,topic,streams,monitoring	What are Kafka Streams and its use cases?	Kafka Streams is a lightweight, stream processing library in Kafka that allows developers to build applications that process and analyze real-time data streams. It provides a high-level DSL (Domain-Specific Language) for writing stream processing applications directly against Kafka topics. Use cases for Kafka Streams include real-time analytics, fraud detection, monitoring, and ETL (Extract, Transform, Load) operations.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/027-what-are-kafka-streams-and-its-use-cases.txt
kafka/1b566a222e7b	kafka	hard	kafka,producer,exactly-once	Discuss the challenges and solutions for ensuring exactly-once semantics in Kafka.	Ensuring exactly-once semantics in Kafka is challenging but achievable. Challenges include:\n* Producer Idempotence: Producers can be configured to send messages idempotently, ensuring that duplicate messages do not affect the overall result.\n\nUnder the hood: Requires enable.idempotence=true + transactional API. Adds latency.	projects/knowledge/interview/kafka/040-discuss-the-challenges-and-solutions-for-ensuring-.txt
kafka/1b89b64e9050	kafka	hard	kafka,broker,partition,replication	How does Kafka ensure fault tolerance?	Kafka ensures fault tolerance through partition replication. Each partition has multiple replicas distributed across different brokers. If a broker fails, one of the replicas can be promoted to serve as the new leader, ensuring uninterrupted data availability. This replication strategy, combined with ZooKeeper for broker coordination and leader election, makes Kafka resilient to individual broker failures and contributes to the overall fault tolerance of the system.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/015-how-does-kafka-ensure-fault-tolerance.txt
kafka/1cf0b1e635c2	kafka	easy	kafka, control-flow	What are the challenges and best practices for upgrading Kafka versions in a production environment?	Upgrading Kafka versions in a production environment poses challenges that need careful consideration. Best practices include:\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/073-what-are-the-challenges-and-best-practices-for-upg.txt
kafka/1d403be2164b	kafka	medium	kafka,broker,partition,replication	Discuss the impact of changing the Kafka replication factor.	Changing the Kafka replication factor has several impacts on the Kafka cluster:\n* Fault Tolerance: Increasing the replication factor improves fault tolerance. Each partition has multiple replicas, and if a broker fails, one of the replicas can be promoted to leader, ensuring continuity of data availability.\n\nRemember: RF=3 survives 2 broker failures. `min.insync.replicas=2` protects writes.\n\nGotcha: min.insync.replicas + acks=all = data safety but can reduce availability.	projects/knowledge/interview/kafka/042-discuss-the-impact-of-changing-the-kafka-replicati.txt
kafka/1e0ece52e6a5	kafka	hard	kafka,producer,consumer,schema-registry	Explain the role of the Kafka Schema Registry.	The Kafka Schema Registry is a centralized service that manages the schemas for messages produced and consumed in a Kafka environment. It ensures that producers and consumers agree on the structure of the data by enforcing schema compatibility. This is crucial in evolving systems where data formats may change over time. The Schema Registry supports various serialization formats like Avro, JSON, and Protobuf.\n\nRemember: Schema Registry ensures format agreement. Avro/Protobuf/JSON Schema. Port 8081.	projects/knowledge/interview/kafka/026-explain-the-role-of-the-kafka-schema-registry.txt
kafka/1f13e9fed150	kafka	hard	kafka,broker,topic	How does Kafka handle message retention?	Message retention in Kafka is managed through configurable retention policies. Kafka allows users to define retention based on time or size constraints. Messages that exceed the specified retention period or size are eligible for deletion. This feature ensures that Kafka does not indefinitely store all messages, helping manage storage costs and preventing the system from becoming overloaded with outdated data. Retention policies can be set at both the topic and broker levels.\n\nRemember: delete=TTL-based removal. compact=keep latest per key. "delete=TTL, compact=upsert."	projects/knowledge/interview/kafka/019-how-does-kafka-handle-message-retention.txt
kafka/1f3d267ba56a	kafka	easy	kafka,consumer,performance	What is Apache Kafka?	Apache Kafka is a distributed, scalable, and fault-tolerant streaming platform designed to handle real-time data feeds. Developed by the Apache Software Foundation, Kafka is widely used for building real-time data pipelines and streaming applications. It provides a publish-subscribe messaging system, high-throughput, fault tolerance, and durability, making it suitable for various use cases such as log aggregation, event sourcing, and data integration.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/005-what-is-apache-kafka.txt
kafka/1fa1e854ce18	kafka	easy	kafka,broker	What is the role of the Kafka Log Cleaner?	The Kafka Log Cleaner is a background process responsible for managing disk space and maintaining optimal storage efficiency in Kafka brokers. Key aspects of the Kafka Log Cleaner include:\n* Log Segments: Over time, log segments in Kafka can accumulate obsolete and deleted records, consuming unnecessary disk space.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/058-what-is-the-role-of-the-kafka-log-cleaner.txt
kafka/2829dae5abb2	kafka	easy	kafka,producer,exactly-once	What is the role of Kafka's transactional producer API, and how does it differ from the non-transactional API?	Kafka's transactional producer API provides Exactly-Once Semantics for producing messages. Key aspects of the transactional producer API and its differences from the non-transactional API include:\n\nRemember: acks: 0=fire-and-forget, 1=leader, all=ISR. "0=YOLO, 1=Leader, all=Safe."\n\nGotcha: acks=all adds latency. Pair with retries + enable.idempotence for reliability.	projects/knowledge/interview/kafka/070-what-is-the-role-of-kafkas-transactional-producer-.txt
kafka/28e7b1ba60d3	kafka	medium	kafka,producer,broker,topic	Describe the purpose of a Kafka broker.	A Kafka broker is a server instance within the Kafka cluster that stores and manages the distribution of messages. The primary purposes of a Kafka broker include:\n* Message Storage: Brokers store the messages published by producers to topics. Messages are stored in partitions within the broker.\n\nRemember: Broker = Kafka server. Cluster = multiple brokers. Each stores partition replicas.\n\nUnder the hood: Each partition has one leader. Reads/writes go to leader (pre-KIP-392).	projects/knowledge/interview/kafka/010-describe-the-purpose-of-a-kafka-broker.txt
kafka/2b3bee25743a	kafka	hard	kafka,broker,partition,replication	Explain the concept of replication in Kafka.	Replication in Kafka involves creating redundant copies (replicas) of each partition across multiple brokers. This provides fault tolerance, ensuring that data remains available even if some brokers fail. Replicas include a leader and follower(s). The leader handles read and write operations, while followers replicate the data. If the leader fails, one of the followers is promoted to be the new leader.\n\nRemember: RF=3 survives 2 broker failures. `min.insync.replicas=2` protects writes.\n\nGotcha: min.insync.replicas + acks=all = data safety but can reduce availability.	projects/knowledge/interview/kafka/017-explain-the-concept-of-replication-in-kafka.txt
kafka/2e647e567d3d	kafka	medium	kafka,consumer,topic	Explain the role of the Apache Kafka Consumer API.	The Apache Kafka Consumer API is a set of classes and methods that allow developers to create and configure Kafka consumers for subscribing to and processing messages from Kafka topics. Key aspects of the Consumer API include:\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/038-explain-the-role-of-the-apache-kafka-consumer-api.txt
kafka/2eef3839c8cd	kafka	easy	kafka,performance,monitoring	What is the role of the Kafka Metrics API?	The Kafka Metrics API provides a comprehensive set of metrics and monitoring capabilities to track the performance and health of a Kafka cluster. Key aspects of the Kafka Metrics API include:\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/061-what-is-the-role-of-the-kafka-metrics-api.txt
kafka/33176b18d590	kafka	easy	kafka,producer	"What is a ""Producer"" in regards to Kafka?"	An application that publishes data to the Kafka cluster.\n\nRemember: acks: 0=fire-and-forget, 1=leader, all=ISR. "0=YOLO, 1=Leader, all=Safe."\n\nGotcha: acks=all adds latency. Pair with retries + enable.idempotence for reliability.	projects/knowledge/interview/kafka/003-what-is-a-producer-in-regards-to-kafka.txt
kafka/342885058e74	kafka	easy	kafka,producer,consumer,broker	Define a producer in Kafka.	A producer in Kafka is a component or application responsible for publishing messages to Kafka topics. Producers create and send messages to specific topics, making the messages available for consumption by one or more consumers. Producers are typically designed to be highly scalable and fault-tolerant, ensuring reliable and efficient delivery of messages to Kafka brokers.\n\nRemember: acks: 0=fire-and-forget, 1=leader, all=ISR. "0=YOLO, 1=Leader, all=Safe."\n\nGotcha: acks=all adds latency. Pair with retries + enable.idempotence for reliability.	projects/knowledge/interview/kafka/008-define-a-producer-in-kafka.txt
kafka/3db9bcccddf3	kafka	medium	kafka,partition,replication	Discuss Kafka’s support for multi-datacenter replication.	Kafka supports multi-datacenter replication to enhance fault tolerance and ensure data availability across geographically distributed locations. Key aspects include:\n* Replica Placement: Kafka allows the placement of replicas across multiple data centers. Each partition can have replicas distributed across different geographical regions.\n\nRemember: RF=3 survives 2 broker failures. `min.insync.replicas=2` protects writes.\n\nGotcha: min.insync.replicas + acks=all = data safety but can reduce availability.	projects/knowledge/interview/kafka/080-discuss-kafkas-support-for-multi-datacenter-replic.txt
kafka/3e88877d6244	kafka	medium	kafka,performance	Discuss the considerations for selecting the appropriate storage infrastructure for Kafka.	Choosing the right storage infrastructure for Kafka involves several considerations:\n* Disk Speed and Type: Opt for high-speed disks, such as SSDs, to ensure optimal disk I/O performance. The choice of disk type impacts the overall throughput and latency of Kafka.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/075-discuss-the-considerations-for-selecting-the-appro.txt
kafka/4168439a37a5	kafka	hard	kafka, control-flow, logging, networking	Discuss the importance of log compaction in Kafka.	Log compaction is an important feature in Kafka that helps retain the latest value for each key in a log, while older values are periodically compacted and removed. This is particularly useful in scenarios where it is essential to maintain the latest state of each record, such as maintaining the current state of a database. Log compaction ensures that even if there are multiple writes for the same key, only the latest value is retained, reducing storage overhead and improving query efficiency.	projects/knowledge/interview/kafka/025-discuss-the-importance-of-log-compaction-in-kafka.txt
kafka/41ff6a647043	kafka	hard	kafka,producer,consumer	How does Kafka handle backpressure?	Kafka handles backpressure through its flow control mechanism. Consumers can control the rate at which they consume messages by adjusting parameters like max.poll.records and fetch.min.bytes. Producers, on the other hand, can use settings such as acks and linger.ms to control the rate at which they send messages. If a consumer is overwhelmed, it can reduce the frequency of poll requests or process messages more quickly.\n\nRemember: 0=fastest(lossy), 1=leader(balanced), all=safest(slowest). Choose by importance.	projects/knowledge/interview/kafka/036-how-does-kafka-handle-backpressure.txt
kafka/4266d837d86f	kafka	medium	kafka,producer,consumer,topic	What is a topic in Kafka?	In Kafka, a topic is a logical channel or category to which messages are published by producers and from which messages are consumed by consumers. Topics serve as the primary means of organizing and categorizing data within the Kafka cluster. Producers publish messages to specific topics, and consumers subscribe to topics to receive and process the messages. Topics enable the decoupling of data producers and consumers, allowing for flexible and scalable data processing.\n\nRemember: Topic = named feed. "TV channel." Publishers send, subscribers receive.\n\nUnder the hood: Topics split into partitions. Each partition = ordered, immutable log.	projects/knowledge/interview/kafka/007-what-is-a-topic-in-kafka.txt
kafka/4635b018a29c	kafka	medium	kafka,replication,broker	Discuss the role of Kafka MirrorMaker in data replication across clusters.	Kafka MirrorMaker is a tool designed for replicating data between Kafka clusters. Its role includes:\n\nRemember: RF=3 survives 2 broker failures. `min.insync.replicas=2` protects writes.\n\nGotcha: min.insync.replicas + acks=all = data safety but can reduce availability.	projects/knowledge/interview/kafka/049-discuss-the-role-of-kafka-mirrormaker-in-data-repl.txt
kafka/497c996c762a	kafka	medium	kafka,producer,broker,exactly-once	Explain how Kafka handles message deduplication.	Kafka addresses message deduplication through a combination of producer and broker mechanisms:\n* Producer Idempotence: Kafka introduced the concept of idempotent producers. When a producer is configured as idempotent, it ensures that messages are sent exactly once. This helps prevent duplicates caused by retries during transient failures.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/079-explain-how-kafka-handles-message-deduplication.txt
kafka/55f8ec77452f	kafka	easy	kafka,producer,consumer	What is the role of interceptors in Kafka producers and consumers?	Interceptors in Kafka allow developers to intercept and modify records before they are sent by producers or received by consumers. Key aspects of interceptors include:\n\nRemember: acks: 0=fire-and-forget, 1=leader, all=ISR. "0=YOLO, 1=Leader, all=Safe."\n\nGotcha: acks=all adds latency. Pair with retries + enable.idempotence for reliability.	projects/knowledge/interview/kafka/043-what-is-the-role-of-interceptors-in-kafka-producer.txt
kafka/56007158fdd9	kafka	easy	kafka, control-flow	What are the potential issues and solutions when dealing with out-of-order messages in Kafka?	Dealing with out-of-order messages in Kafka is essential for maintaining data consistency. Potential issues and solutions include:\nCauses of Out-of-Order Messages:\n* Network Delays: Variability in network latencies can result in messages arriving out of order.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/078-what-are-the-potential-issues-and-solutions-when-d.txt
kafka/58a30c6353e5	kafka	hard	kafka,topic,partition,replication	Explain Kafka's architecture in terms of leader and follower replicas.	Kafka's architecture involves leader and follower replicas for each partition. Key points include:\n* Partition Replication: Each Kafka topic is divided into partitions, and each partition has multiple replicas. Replication provides fault tolerance and high availability.\n\nRemember: RF=3 survives 2 broker failures. `min.insync.replicas=2` protects writes.\n\nGotcha: min.insync.replicas + acks=all = data safety but can reduce availability.	projects/knowledge/interview/kafka/074-explain-kafkas-architecture-in-terms-of-leader-and.txt
kafka/61a9512c9959	kafka	medium	kafka,schema-registry	Discuss the role of Apache Avro in Kafka.	Apache Avro is a binary serialization format used in Kafka for efficient and compact data serialization. Key aspects of Avro in Kafka include:\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/046-discuss-the-role-of-apache-avro-in-kafka.txt
kafka/627228a9d566	kafka	medium	kafka,broker	Explain Kafka's protocol for inter-broker communication.	Kafka uses a binary protocol for inter-broker communication. Key aspects include:\n* Message Format: Inter-broker communication involves the exchange of messages between Kafka brokers. Messages are sent in a binary format for efficiency.\n\nRemember: Broker = Kafka server. Cluster = multiple brokers. Each stores partition replicas.\n\nUnder the hood: Each partition has one leader. Reads/writes go to leader (pre-KIP-392).	projects/knowledge/interview/kafka/076-explain-kafkas-protocol-for-inter-broker-communica.txt
kafka/635942a185b0	kafka	hard	kafka,broker,partition,offset	Discuss the internal architecture of a Kafka broker.	The internal architecture of a Kafka broker includes several components:\n* Log Segment: The fundamental storage unit containing committed messages. Log segments are immutable and represent a portion of a partition's commit log.\n* Log Manager: Manages the creation, deletion, and rolling of log segments. It also handles log indexing and compaction.\n\nRemember: Broker = Kafka server. Cluster = multiple brokers. Each stores partition replicas.\n\nUnder the hood: Each partition has one leader. Reads/writes go to leader (pre-KIP-392).	projects/knowledge/interview/kafka/037-discuss-the-internal-architecture-of-a-kafka-broke.txt
kafka/63623805e2a1	kafka	hard	kafka,topic,connect	Discuss the role of Kafka Connect converters.	Kafka Connect converters are components responsible for translating data between Kafka Connect and external systems. They handle the serialization and deserialization of data, allowing seamless integration between Kafka topics and various data storage systems. Converters are crucial for ensuring that data can be efficiently and accurately transferred between Kafka and external systems with different data formats.\n\nRemember: Kafka Connect = pre-built source/sink connectors. DB↔Kafka without coding.	projects/knowledge/interview/kafka/056-discuss-the-role-of-kafka-connect-converters.txt
kafka/63d0e1a9a0a7	kafka	easy	kafka,broker	What are the key considerations for Kafka deployment in a cloud environment?	Deploying Kafka in a cloud environment involves several key considerations:\n* Resource Scaling: Cloud platforms allow for dynamic scaling of resources, enabling Kafka clusters to adapt to varying workloads. Consider using auto-scaling features to adjust the number of broker instances based on demand.\n\nRemember: Same key → same partition (hash). Guarantees per-key ordering.\n\nGotcha: Changing partition count redistributes keys → ordering breaks.	projects/knowledge/interview/kafka/047-what-are-the-key-considerations-for-kafka-deployme.txt
kafka/66657e424f46	kafka	medium	kafka,broker,partition,replication	Discuss the considerations for achieving low-latency in Kafka.	Achieving low-latency in Kafka involves careful consideration of several factors:\n* Partition and Replica Placement: Distribute partitions and their replicas across brokers and network locations to minimize data transfer latency.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/062-discuss-the-considerations-for-achieving-low-laten.txt
kafka/6d4e76a75037	kafka	medium	kafka,producer,consumer,broker	Discuss Kafka's support for end-to-end security using SSL/TLS.	Kafka provides robust support for end-to-end security through SSL/TLS. Key aspects include:\n* Encryption: SSL/TLS ensures that data transferred between producers, brokers, and consumers is encrypted, preventing unauthorized access to sensitive information.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/072-discuss-kafkas-support-for-end-to-end-security-usi.txt
kafka/7497346dd8da	kafka	medium	kafka	Explain the use of Kafka quotas and rate limiting.	Kafka quotas and rate limiting are mechanisms to control and manage resource usage within a Kafka cluster. Key aspects include:\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/060-explain-the-use-of-kafka-quotas-and-rate-limiting.txt
kafka/779d2c17b480	kafka	medium	kafka,performance	Discuss the impact of message size on Kafka performance.	The impact of message size on Kafka performance is a crucial consideration, and it affects various aspects of Kafka's operation:\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/052-discuss-the-impact-of-message-size-on-kafka-perfor.txt
kafka/7a045a15ad84	kafka	hard	kafka,consumer,topic,partition	Explain the role of the Kafka commit log.	The Kafka commit log is the fundamental data structure that underlies the storage of messages in Kafka. It is a distributed, fault-tolerant, and durable log that records all messages published to Kafka topics. The commit log ensures the ordering, persistence, and fault tolerance of messages. Each partition has its own commit log, and messages are written sequentially to the log. Consumers read from the log, ensuring a consistent and ordered view of the data.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/032-explain-the-role-of-the-kafka-commit-log.txt
kafka/7c52a124bb0a	kafka	medium	kafka,topic,security	Explain the role of Kafka ACLs (Access Control Lists).	Kafka ACLs (Access Control Lists) play a crucial role in securing Kafka clusters by defining fine-grained access permissions for users and applications. Key aspects of Kafka ACLs include:\n* Topic-Level Permissions: ACLs can be set at the topic level, specifying which users or groups have read or write access to particular topics.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/045-explain-the-role-of-kafka-acls-access-control-list.txt
kafka/7cbd16526174	kafka	hard	kafka,producer,consumer,partition	How does Kafka address the challenge of maintaining order across multiple partitions?	Maintaining order across multiple partitions in Kafka is addressed through the following mechanisms:\n* Partition Ordering: Within each partition, Kafka maintains the order of records as they are produced. Consumers can rely on the order of records within a partition.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/084-how-does-kafka-address-the-challenge-of-maintainin.txt
kafka/7e3d23e4cfdc	kafka	medium	kafka,performance	Explain the process of upgrading a Kafka cluster.	The process of upgrading a Kafka cluster involves the following steps:\n* Backup: Before upgrading, ensure a comprehensive backup of the Kafka data and configurations. This provides a safety net in case of unforeseen issues during the upgrade.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/048-explain-the-process-of-upgrading-a-kafka-cluster.txt
kafka/82e332e9703f	kafka	hard	kafka,topic,partition,offset	How is data stored in Kafka?	Data in Kafka is stored in the form of logs. Each topic is divided into partitions, and each partition is a linear, ordered sequence of messages. Messages within a partition are assigned a unique offset that represents their position in the partition. Kafka ensures durability by persisting messages to disk, making the data resilient to node failures.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/011-how-is-data-stored-in-kafka.txt
kafka/862496b929e0	kafka	medium	kafka,connect	Discuss the use of Kafka Connect transforms and the available transformation types.	Kafka Connect transforms are operations applied to data during the ETL (Extract, Transform, Load) process. They allow modification, filtering, or enrichment of data as it flows through Kafka Connect. Key transformation types include:\n\nRemember: Kafka Connect = pre-built source/sink connectors. DB↔Kafka without coding.	projects/knowledge/interview/kafka/077-discuss-the-use-of-kafka-connect-transforms-and-th.txt
kafka/86802b230255	kafka	hard	kafka,producer,consumer,topic	Discuss the publish-subscribe model in Kafka.	The publish-subscribe model in Kafka involves producers publishing messages to topics, and consumers subscribing to those topics to receive and process the messages. Multiple consumers can subscribe to the same topic, forming consumer groups. Each message is broadcast to all consumers within a group, allowing for parallel and distributed processing. This model enables decoupling between producers and consumers, supporting real-time data streaming and event-driven architectures.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/021-discuss-the-publish-subscribe-model-in-kafka.txt
kafka/87e798204940	kafka	hard	kafka,broker	Discuss the role of Kafka in event sourcing architectures.	Kafka plays a crucial role in event sourcing architectures by serving as a distributed, fault-tolerant event log. In event sourcing:\n* Event Log: Kafka acts as the central event log where all changes to the state of an application are captured as immutable events. These events represent state transitions and serve as a reliable source of truth.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/055-discuss-the-role-of-kafka-in-event-sourcing-archit.txt
kafka/8ee1c8dd4d67	kafka	medium	kafka,consumer,partition,consumer-group	How does Kafka handle dynamic partition assignment in consumer groups?	Kafka handles dynamic partition assignment in consumer groups through the following process:\n* Group Coordinator: Each consumer group has a designated group coordinator responsible for managing group membership and partition assignments.\n\nRemember: Consumer group = team sharing work. One partition→one consumer. Max parallel = partitions.\n\nGotcha: More consumers than partitions = idle consumers.	projects/knowledge/interview/kafka/067-how-does-kafka-handle-dynamic-partition-assignment.txt
kafka/8f68f4271a06	kafka	medium	kafka,security	Explain the considerations for securing a Kafka cluster.	Securing a Kafka cluster involves implementing measures to protect data, ensure authentication and authorization, and prevent unauthorized access. Key considerations include:\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/054-explain-the-considerations-for-securing-a-kafka-cl.txt
kafka/90a8d6bd3a2e	kafka	easy	kafka,producer,consumer,partition	What is a Kafka record or message?	In Kafka, a record or message is the basic unit of data that is produced and consumed. A record typically consists of two components:\n* Key: An optional field that can be used for partitioning and indexing. The key is used to determine the partition to which the message will be sent.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/012-what-is-a-kafka-record-or-message.txt
kafka/91f1c434e0a5	kafka	medium	kafka,broker,replication	Explain the impact of broker properties like min.insync.replicas on Kafka's reliability.	The min.insync.replicas broker property in Kafka determines the minimum number of in-sync replicas (ISRs) required to acknowledge a write operation as successful. Its impact on Kafka's reliability includes:\n\nRemember: Broker = Kafka server. Cluster = multiple brokers. Each stores partition replicas.\n\nUnder the hood: Each partition has one leader. Reads/writes go to leader (pre-KIP-392).	projects/knowledge/interview/kafka/068-explain-the-impact-of-broker-properties-like-minin.txt
kafka/92f190f3ee51	kafka	easy	kafka,broker,partition,replication	What is the purpose of Kafka Zookeeper?	Kafka relies on Apache ZooKeeper for distributed coordination and management of its cluster. The main purposes of Kafka ZooKeeper include:\n* Cluster Coordination: ZooKeeper helps Kafka brokers coordinate and elect a leader for each partition, facilitating fault tolerance and load balancing.\n\nFun fact: KRaft replaces ZooKeeper from Kafka 3.3+. One less system to manage.\n\nRemember: ZK managed brokers, configs, elections. KRaft moves all into Kafka itself.	projects/knowledge/interview/kafka/014-what-is-the-purpose-of-kafka-zookeeper.txt
kafka/954374fce931	kafka	medium	kafka,consumer,partition,consumer-group	Explain the mechanics of Kafka rebalancing.	Kafka rebalancing is a process that occurs when the membership of consumer group instances changes. It involves redistributing the partitions among the consumers to ensure a balanced workload. The mechanics of Kafka rebalancing include:\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/057-explain-the-mechanics-of-kafka-rebalancing.txt
kafka/9666d0cee646	kafka	hard	kafka,consumer,broker,performance	How can you monitor and optimize Kafka cluster performance?	Monitoring and optimizing Kafka cluster performance involve several key practices:\n* Metrics Monitoring: Regularly monitor Kafka metrics related to broker health, disk usage, network throughput, and consumer lag. Utilize tools like JMX, Prometheus, or Confluent Control Center for real-time and historical metrics.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/053-how-can-you-monitor-and-optimize-kafka-cluster-per.txt
kafka/96b79970a2c8	kafka	hard	kafka,consumer,broker,topic	Discuss the use of partitions in Kafka.	Partitions are fundamental units of parallelism and scalability in Kafka. They allow Kafka to horizontally scale by distributing the data across multiple brokers. Each partition is an ordered, immutable sequence of messages, and topics are divided into partitions. Partitions enable parallel processing, as multiple consumers can simultaneously consume different partitions of a topic.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/016-discuss-the-use-of-partitions-in-kafka.txt
kafka/9965810aa172	kafka	hard	kafka,broker,security	Explain how Kafka handles the scenario of broker failures.	Kafka is designed to handle broker failures seamlessly, ensuring high availability and fault tolerance. Key mechanisms include:\n\nRemember: Broker = Kafka server. Cluster = multiple brokers. Each stores partition replicas.\n\nUnder the hood: Each partition has one leader. Reads/writes go to leader (pre-KIP-392).	projects/knowledge/interview/kafka/082-explain-how-kafka-handles-the-scenario-of-broker-f.txt
kafka/a3ec2c075b9b	kafka	medium	kafka,partition,replication	Discuss the significance of the ISR (In-Sync Replicas) list.	The ISR (In-Sync Replicas) list is a subset of replicas for a partition that are considered in sync with the leader. The significance of the ISR list includes:\n* Fault Tolerance: The ISR list ensures fault tolerance by only promoting replicas within the ISR list to leaders in case of a leader failure.\n\nRemember: RF=3 survives 2 broker failures. `min.insync.replicas=2` protects writes.\n\nGotcha: min.insync.replicas + acks=all = data safety but can reduce availability.	projects/knowledge/interview/kafka/030-discuss-the-significance-of-the-isr-in-sync-replic.txt
kafka/a7f55ad966d8	kafka	medium	kafka,partition,offset	How does Kafka handle data compaction in detail?	Kafka handles data compaction through a feature known as log compaction. Here's a detailed explanation:\n* Log Segments: Kafka maintains data in log segments, each representing a sequential and immutable portion of a partition's commit log.\n\nRemember: delete=TTL-based removal. compact=keep latest per key. "delete=TTL, compact=upsert."	projects/knowledge/interview/kafka/050-how-does-kafka-handle-data-compaction-in-detail.txt
kafka/af05685ceff3	kafka	medium	kafka,broker	Discuss the considerations for choosing the appropriate Kafka storage format (log, compacted log, etc.).	Choosing the appropriate Kafka storage format involves considering factors such as use case, data retention, and access patterns. Common storage formats include:\nLog Format (Append-Only):\n* Suitable for scenarios where the entire history of events is critical.\n* Well-suited for event sourcing architectures.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/083-discuss-the-considerations-for-choosing-the-approp.txt
kafka/b007363484f9	kafka	easy	kafka,partition,replication	What is the purpose of the Kafka Controller?	The Kafka Controller is a crucial component within the Kafka cluster responsible for managing partitions, leaders, and replicas. Its main purposes include:\nPartition Leader Election: The Controller ensures the election of a leader for each partition. The leader is responsible for handling read and write operations for that partition.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/029-what-is-the-purpose-of-the-kafka-controller.txt
kafka/b0f470a699f7	kafka	hard	kafka,producer,consumer,partition	Explain the scenarios where partitioning becomes a critical factor in Kafka.	Partitioning is a critical factor in Kafka and becomes essential in various scenarios:\n* Scalability: Partitioning allows Kafka to scale horizontally by distributing data across multiple partitions. Each partition can be processed independently, enabling Kafka to handle a high volume of data and support a large number of producers and consumers.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/051-explain-the-scenarios-where-partitioning-becomes-a.txt
kafka/b53ad1b7e283	kafka	medium	kafka,performance	Discuss the concept of Kafka log appenders.	Kafka log appenders are components responsible for appending log entries to Kafka logs in an efficient and reliable manner. Key aspects of Kafka log appenders include:\n* Batching: Log appenders often batch multiple log entries into a single write operation to improve write efficiency and reduce disk I/O.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/066-discuss-the-concept-of-kafka-log-appenders.txt
kafka/b9e24c5be829	kafka	easy	kafka,producer,partition,offset	What is a Kafka transaction and when is it used?	A Kafka transaction is a mechanism that allows producers to send messages to multiple partitions within a transactional context. Kafka transactions ensure atomicity, consistency, and isolation of message writes across partitions. Producers can either commit or abort a transaction, ensuring that messages are either all successfully written to partitions or none at all.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/034-what-is-a-kafka-transaction-and-when-is-it-used.txt
kafka/bdec1045d29d	kafka	easy	kafka,streams	What is the role of Kafka Streams DSL?	Kafka Streams DSL (Domain-Specific Language) is a high-level API provided by Kafka Streams for building stream processing applications. It allows developers to define complex data processing operations using a fluent and expressive API. Key aspects of the Kafka Streams DSL include:\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/039-what-is-the-role-of-kafka-streams-dsl.txt
kafka/bf1aab79e0f7	kafka	medium	kafka,producer,topic	Explain the role of the Apache Kafka Producer API.	The Apache Kafka Producer API is a set of classes and methods that enable developers to create and configure Kafka producers for publishing messages to Kafka topics. The key aspects of the Producer API include:\n\nRemember: acks: 0=fire-and-forget, 1=leader, all=ISR. "0=YOLO, 1=Leader, all=Safe."\n\nGotcha: acks=all adds latency. Pair with retries + enable.idempotence for reliability.	projects/knowledge/interview/kafka/035-explain-the-role-of-the-apache-kafka-producer-api.txt
kafka/c16131b8366d	kafka	easy	kafka,broker,topic,partition	What is the role of the Kafka AdminClient API, and how is it used?	The Kafka AdminClient API is a Java client that provides administrative functionality to interact with and manage Kafka clusters programmatically. Its role includes:\n* Cluster Metadata Retrieval: The AdminClient allows users to retrieve metadata about the Kafka cluster, such as broker information, topic details, and partition assignments.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/081-what-is-the-role-of-the-kafka-adminclient-api-and-.txt
kafka/c482232f0f1e	kafka	easy	kafka,exactly-once	What is the purpose of Kafka’s Exactly-Once Semantics and how is it implemented?	Kafka's Exactly-Once Semantics ensures that messages are processed and delivered exactly once, without duplicates or message loss. This is achieved through the following mechanisms:\n\nUnder the hood: Requires enable.idempotence=true + transactional API. Adds latency.	projects/knowledge/interview/kafka/065-what-is-the-purpose-of-kafkas-exactly-once-semanti.txt
kafka/c59c496525fd	kafka	medium	kafka,partition,offset	Explain the role of log segments in Kafka storage.	In Kafka, the commit log is divided into log segments, each representing a sequential and immutable portion of a partition's log. Key aspects of log segments include:\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/041-explain-the-role-of-log-segments-in-kafka-storage.txt
kafka/c6c485392597	kafka	medium	kafka,consumer,topic,partition	What is the significance of the offset in Kafka?	The offset is a unique identifier assigned to each message within a partition. It represents the position of a message in the partition's log. Consumers use offsets to keep track of the messages they have already consumed. Kafka ensures that each message has a unique offset within a partition, enabling consumers to resume processing from a specific point in the log. Offsets are stored in Kafka topics, providing a reliable way to maintain the state of message consumption.\n\nRemember: Offset = partition position. "Bookmark." Consumers track where they left off.\n\nGotcha: Without committing offsets, consumers replay messages after crash.	projects/knowledge/interview/kafka/018-what-is-the-significance-of-the-offset-in-kafka.txt
kafka/c8009e9c82a7	kafka	medium	kafka,topic,connect	What is the role of the Kafka Connect API?	The Kafka Connect API is used for building and running connectors that integrate Kafka with external data sources or sinks. Connectors facilitate the movement of data in and out of Kafka, allowing seamless integration with databases, file systems, messaging systems, and other data storage or processing systems. Kafka Connect simplifies the development and deployment of data pipelines, enabling the transfer of data between Kafka topics and external systems in a scalable and fault-tolerant manner.\n\nRemember: Kafka Connect = pre-built source/sink connectors. DB↔Kafka without coding.	projects/knowledge/interview/kafka/024-what-is-the-role-of-the-kafka-connect-api.txt
kafka/cc9015fa0726	kafka	easy	kafka,consumer,topic,partition	What is a consumer group in Kafka?	A consumer group in Kafka is a logical grouping of consumers that work together to consume messages from one or more topics. Each consumer group has one or more consumers, and each message within a topic partition is consumed by only one consumer within the group. Consumer groups enable parallel processing of messages, as different partitions can be consumed concurrently by different consumers. This architecture supports scalable and fault-tolerant message consumption.\n\nRemember: Consumer group = team sharing work. One partition→one consumer. Max parallel = partitions.\n\nGotcha: More consumers than partitions = idle consumers.	projects/knowledge/interview/kafka/020-what-is-a-consumer-group-in-kafka.txt
kafka/d46112051091	kafka	medium	kafka,producer,consumer,broker	Explain the considerations for scaling a Kafka cluster horizontally.	Scaling a Kafka cluster horizontally involves adding more broker instances to distribute the workload and increase capacity. Considerations for horizontal scaling include:\n* Broker Addition: New broker instances can be added to the Kafka cluster to increase the overall capacity for handling more producers, consumers, and partitions.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/071-explain-the-considerations-for-scaling-a-kafka-clu.txt
kafka/e66d37ece4c9	kafka	medium	kafka, networking	How does Kafka support multi-tenancy?	Kafka supports multi-tenancy, allowing multiple independent applications or business units (tenants) to share a single Kafka cluster. Key aspects of multi-tenancy in Kafka include:\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/044-how-does-kafka-support-multi-tenancy.txt
kafka/ea791dee4e74	kafka	medium	kafka,broker,broker	Explain the role of the Kafka Raft metadata mode.	Kafka Raft metadata mode is an enhancement to Kafka's metadata storage system, replacing the traditional Zookeeper-based metadata storage. The Raft consensus algorithm is used to achieve distributed consensus among broker nodes, providing better reliability and simplicity compared to Zookeeper. Key aspects of Kafka Raft metadata mode include:\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/064-explain-the-role-of-the-kafka-raft-metadata-mode.txt
kafka/eba25b355be7	kafka	medium	kafka,producer,broker	Explain the process of Kafka producer acknowledgment.	Kafka producer acknowledgment refers to the confirmation received by a producer after successfully publishing a message to a Kafka broker. Producers can configure the level of acknowledgment they require using the acks parameter:\n\nRemember: acks: 0=fire-and-forget, 1=leader, all=ISR. "0=YOLO, 1=Leader, all=Safe."\n\nGotcha: acks=all adds latency. Pair with retries + enable.idempotence for reliability.	projects/knowledge/interview/kafka/022-explain-the-process-of-kafka-producer-acknowledgme.txt
kafka/eba8111d5383	kafka	medium	kafka,consumer,partition	Discuss the impact of increasing the number of partitions on consumer parallelism.	Increasing the number of partitions in Kafka has a direct impact on consumer parallelism. Key considerations include:\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/059-discuss-the-impact-of-increasing-the-number-of-par.txt
kafka/ecef9d048b20	kafka	medium	kafka,consumer,partition,performance	Describe the impact of increasing the number of partitions in Kafka.	Increasing the number of partitions in Kafka has several impacts:\n* Increased Parallelism: More partitions allow for more parallelism in data processing. Multiple consumers can concurrently consume messages from different partitions, providing improved throughput.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/028-describe-the-impact-of-increasing-the-number-of-pa.txt
kafka/f4dd7757adcc	kafka	medium	kafka, alerting, text-processing	What Kafka is used for?	- Real-time e-commerce\n- Banking\n- Health Care\n- Automotive (traffic alerts, hazard alerts, ...)\n- Real-time Fraud Detection\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/002-what-kafka-is-used-for.txt
kafka/f78076cb583b	kafka	medium	kafka,broker,partition,replication	How does Kafka ensure data durability?	Kafka ensures data durability through various mechanisms:\n* Replication: Kafka replicates partitions across multiple brokers. This means that even if one or more brokers fail, data remains available from the replicas.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/023-how-does-kafka-ensure-data-durability.txt
kafka/fc63c6738230	kafka	easy	kafka,producer,consumer,topic	What is a consumer in Kafka?	A consumer in Kafka is a component or application responsible for subscribing to and consuming messages from Kafka topics. Consumers process the messages produced by the producers. Kafka supports both parallel and distributed consumption, allowing multiple consumers to work together to process messages from a shared topic. Consumers can be part of a consumer group, providing scalability and fault tolerance.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/009-what-is-a-consumer-in-kafka.txt
kafka/fc6c011086cc	kafka	medium	kafka,producer,consumer,broker	How does Kafka handle data compression, and what are the available compression codecs?	Kafka handles data compression to optimize storage and network transfer. Producers can compress messages before sending them to brokers, and consumers can decompress received messages. Available compression codecs in Kafka include:\nGzip: Offers a good balance between compression ratio and speed.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/063-how-does-kafka-handle-data-compression-and-what-ar.txt
kafka/fdf8d241285d	kafka	medium	kafka,broker	What's in a Kafka cluster?	- Broker: a server with kafka process running on it. Such server has local storage. In a single Kafka clusters there are usually multiple brokers.\n\nRemember: Kafka = distributed commit log. Topics→partitions→consumers. Sequential I/O + zero-copy.\n\nUnder the hood: Kafka persists everything to disk. High throughput from sequential writes.	projects/knowledge/interview/kafka/004-whats-in-a-kafka-cluster.txt
kafka/fe1beba37bfb	kafka	medium	kafka,topic	Explain the difference between a queue and a topic in Kafka.	"In Kafka, the terms ""queue"" and ""topic"" are often used interchangeably with ""topic"" being the more commonly used term. However, in the context of traditional messaging systems, the key differences are:"\n\nRemember: Topic = named feed. "TV channel." Publishers send, subscribers receive.\n\nUnder the hood: Topics split into partitions. Each partition = ordered, immutable log.	projects/knowledge/interview/kafka/013-explain-the-difference-between-a-queue-and-a-topic.txt

<!-- wiki:related:start -->
---

## Wiki Navigation

### Related Content

- [Kafka](../../../../library/topics/kafka/index.md) (Topic Pack, L1) — Kafka
- [RabbitMQ & Message Queues](../../../../library/topics/rabbitmq/index.md) (Topic Pack, L2) — Kafka

<!-- wiki:related:end -->
