Softwr

Databases · head to head

Apache Kafka vs TiDB

Apache Kafka logo

Apache Kafka

Databases

Open-source distributed event streaming platform

From
Free
Rated
-
TiDB logo

TiDB

Databases

Apache 2.0 distributed SQL database with MySQL wire compatibility and a separate columnar replica for analytical queries.

From
Free
Rated
-

The short version

  • Each has a real cost: Apache Kafka operationally heavy to self-host: brokers, storage, rebalancing and upgrades are a standing job, which is why managed Kafka is a large market; TiDB a production cluster needs several placement driver, storage and SQL nodes before it is fault tolerant, so the minimum viable footprint is far larger than a MySQL server and TiDB is never the economical choice for a small database.
  • They diverge on capability: Apache Kafka covers Durable commit log, TiDB covers MySQL wire compatibility.
  • Prices and features above were last checked on 30 August 2026.

Where they differ

Only the attributes on which Apache Kafka and TiDB actually diverge.

Attributes where Apache Kafka and TiDB differ
AttributeApache KafkaTiDB
Pricing modelOpen source, no licence fee; managed services billed separatelyfreemium
PlatformsLinux, Windows, macOS, Self-hosted, DockerCloud, AWS, Azure, Google Cloud Platform, Self-managed
FoundedUnknown2015

Identical on both: starting price (Free), free tier (Yes), user rating (Not yet rated), category (Databases).

What each one covers

Drawn from each product's published feature list. An absence here means we hold no record of it - not that the product lacks it.

Only in Apache Kafka

  • Durable commit log
  • Horizontal scale
  • Kafka Connect
  • Kafka Streams
  • Replication
  • Low latency

Only in TiDB

  • MySQL wire compatibility
  • Horizontal write scaling
  • Distributed ACID transactions
  • TiFlash columnar replica
  • Automatic rebalancing
  • Raft replication
  • Apache 2.0 licence
  • Online schema change

What people use each for

The jobs each tool is most often brought in to do.

Apache Kafka

  • Moving events between services without point-to-point couplingnot TiDB
  • Feeding analytics and warehouses from operational systems in near real timenot TiDB
  • Replaying history to rebuild state after a consumer bugnot TiDB
  • Buffering bursty producers ahead of slower downstream systemsnot TiDB

TiDB

  • A MySQL workload that has hit the write ceiling of a single primary and would otherwise need an application-level sharding layernot Apache Kafka
  • Reporting that must run against current transactional data, where the columnar replica removes the delay and the cost of an ETL pipelinenot Apache Kafka
  • Multi-region deployments needing a single logical database with automatic failover rather than manual primary promotionnot Apache Kafka
  • Migrating off a sharded MySQL estate where the sharding logic in the application has become the main source of bugs and operational toilnot Apache Kafka

Where each one falls short

Documented limitations, not opinions. Every one is a constraint you would hit in normal use.

Apache Kafka

  • Operationally heavy to self-host: brokers, storage, rebalancing and upgrades are a standing job, which is why managed Kafka is a large market
  • Overkill for straightforward job queues, where a simpler broker is easier to run and reason about
  • Ordering guarantees hold per partition, not per topic, and getting partitioning wrong is a common and expensive design mistake
  • The ecosystem is fragmented across the Apache project and vendor distributions, so documentation and tooling advice often assume a particular distribution

TiDB

  • A production cluster needs several placement driver, storage and SQL nodes before it is fault tolerant, so the minimum viable footprint is far larger than a MySQL server and TiDB is never the economical choice for a small database.
  • Every transaction takes a timestamp from the placement driver and crosses the network to storage nodes, so simple point queries are slower than on single-node MySQL and latency-sensitive paths need to be measured, not assumed.
  • MySQL compatibility is at the wire and dialect level but not complete; stored procedures, triggers and events are not supported, so an application that pushed logic into the database cannot simply be repointed.
  • The columnar replica is an extra full copy of the data on its own nodes, so hybrid analytics roughly doubles storage and adds hardware that must be sized and paid for separately.
  • Operating it well requires cluster-specific expertise in TiUP or the Kubernetes operator, region hot spots, and rebalancing behaviour, so the licence is free but the running cost includes an engineer who understands distributed storage.

Pricing, plan by plan

Apache Kafka

Free
  • Apache KafkaFree
    • Full platform
    • Kafka Connect
    • Kafka Streams

TiDB

Free
  • ServerlessFree
    • 5GB storage
    • 50M request units
    • Free forever tier
  • Dedicated$250/month
    • Dedicated resources
    • SLA guarantees
    • Enterprise support

Which should you pick?

Choose Apache Kafka if

  • You need durable commit log.
  • You want to start without paying.
  • You work on Linux, Windows, macOS, Self-hosted, Docker.
  • You also want horizontal scale.

Choose TiDB if

  • You need mysql wire compatibility.
  • You want to start without paying.
  • You work on Cloud, AWS, Azure, Google Cloud Platform, Self-managed.
  • You also want horizontal write scaling.

Questions people ask

Is Apache Kafka or TiDB better?
Neither clearly leads. Apache Kafka starts at Free and TiDB at Free, and user ratings are close enough to be indistinguishable. Choose on capability and platform support.
Which is cheaper, Apache Kafka or TiDB?
Apache Kafka starts at Free and TiDB at Free.
Does Apache Kafka or TiDB run on more platforms?
Apache Kafka runs on Linux, Windows, macOS, Self-hosted, Docker. TiDB runs on Cloud, AWS, Azure, Google Cloud Platform, Self-managed.
Can I use Apache Kafka for free?
Both have a free tier, so you can try either at no cost before committing.
What is Apache Kafka best used for?
Apache Kafka is most often used for moving events between services without point-to-point coupling, feeding analytics and warehouses from operational systems in near real time, replaying history to rebuild state after a consumer bug, buffering bursty producers ahead of slower downstream systems. Of those, moving events between services without point-to-point coupling and feeding analytics and warehouses from operational systems in near real time are not what TiDB is typically brought in for.
What can Apache Kafka do that TiDB cannot?
Apache Kafka covers Durable commit log, Horizontal scale, Kafka Connect, Kafka Streams. TiDB covers MySQL wire compatibility, Horizontal write scaling, Distributed ACID transactions, TiFlash columnar replica.

Answered from the vendors’ own pages

Apache Kafka: Is Apache Kafka free?

Yes. Kafka is open source under the Apache License v2 with no licence fee. Costs come from the infrastructure you run it on, or from a managed service such as Confluent Cloud.

TiDB: Is TiDB a drop-in replacement for MySQL?

At the protocol and dialect level it is close, and most applications connect unchanged. Stored procedures, triggers and events are not supported, and latency characteristics differ, so it needs testing rather than assumption.

Apache Kafka: How is Kafka different from a message queue?

A queue usually removes a message once it is consumed. Kafka keeps an ordered, durable log, so consumers track their own position and history can be replayed — which is what makes rebuilding state after a bug possible.

TiDB: What licence is it under?

Apache 2.0, for both TiDB and the underlying TiKV storage engine. TiKV is a graduated CNCF project, which is a meaningful governance signal in a market where several competitors moved to source-available licences.

Apache Kafka: Who uses Kafka?

The project reports use by more than 80% of the Fortune 100, with over 5 million lifetime downloads.

TiDB: Do I need TiFlash?

Only for analytical queries. It is an optional columnar replica; without it TiDB is a distributed transactional database. With it you get analytics on live data at the cost of an additional full copy.

Apache Kafka: Do I need to run Kafka myself?

No. Self-hosting is the operationally expensive option; managed services such as Confluent Cloud run the brokers for you and bill on throughput and storage instead.

TiDB: Is the managed cloud the same software?

TiDB Cloud runs the same engine, with the control plane, scaling and operational tooling provided as a service. The entry tier is metered differently from a dedicated cluster, so the cost model rather than the engine is what changes.

TiDB: When is TiDB the wrong choice?

When the database is small enough for one server, when latency on single-row lookups is the primary constraint, or when the application depends on MySQL stored procedures and triggers.

Share

Related pages

Other head to heads