Apache Flink CDC

B
B tier on Open Source License Compliance SoftwareScore 7.0 · #4 of 28
Android app
Not listed
Free plan
Yes
Runs on
Linux, Mac, self-hosted, Windows
github.com
The Apache Flink CDC homepage

Summary

Apache Flink CDC is a distributed data integration tool for real-time and batch data, built on Apache Flink. It supports full database and sharded table synchronization, along with schema evolution and data transformation. Its YAML Pipeline API lets users declare sources, sinks, routing, transformations, and schema evolution rules. SQL and DataStream APIs are also available for defining change-data-capture sources and custom Flink streaming applications. Listed pipeline connectors include Doris, Elasticsearch, Fluss, Hudi, Iceberg, Kafka, MaxCompute, MySQL, OceanBase, Oracle, Paimon, PostgreSQL, and StarRocks. SQL and DataStream source connectors also include SQL Server, MongoDB, TiDB, Db2, and Vitess. Flink CDC 3.6 is listed as compatible with Flink 1.20 and 2.2; version 3.6.0 requires JDK 11 or later. The software is free under the Apache 2.0 License. Deployment is self-hosted: users download and extract Flink CDC and connector JARs, then run them with an Apache Flink deployment. Listed platforms are Linux, macOS, Windows, and self-hosted environments.

Who it is for

It suits teams running Apache Flink that need database synchronization or data integration for batch and real-time workloads. It is self-hosted and requires JDK 11 or later for version 3.6.0.

What is good

  • Supports full and sharded table synchronization.
  • Provides YAML Pipeline, SQL, and DataStream APIs.
  • Lists a broad set of source and pipeline connectors.
  • Free under the Apache 2.0 License.

What to know first

  • Requires an Apache Flink deployment.
  • Version 3.6.0 requires JDK 11 or later.
  • Self-hosted deployment requires handling JARs.

Everything Xiaomi review

Apache Flink CDC: the full review

Apache Flink CDC offers several APIs for synchronization, transformation, and schema evolution within a Flink deployment. Verify the stated Flink compatibility and JDK requirement for the version you plan to use.

Apache Flink CDC is an open-source data integration tool for capturing changes and moving data through Apache Flink. It is best for teams already prepared to operate Flink that need both configurable pipelines and programmable CDC applications. Its flexible APIs and broad connector set are strong assets, but they come with self-hosted operating responsibilities.

Overview

Flink CDC handles real-time and batch integration, with full database and sharded-table synchronization, log-based capture, and initial snapshots. Schema evolution and transformation support suit ongoing replication as well as data movement into downstream systems. Deployment is on-premise: download Flink CDC and connector JARs, then run them with an Apache Flink deployment. That offers control over the runtime, but teams must provide and maintain it themselves.

Key features

Pipeline and application APIs

The YAML Pipeline API lets teams specify sources, sinks, routing, transformations, and schema evolution rules declaratively. It is a good fit for repeatable pipeline definitions; SQL and DataStream APIs instead support defining CDC sources and building custom Flink streaming applications, for teams that need code-level control.

Connector coverage and compatibility

Pipeline connectors include Doris, Elasticsearch, Fluss, Hudi, Iceberg, Kafka, MaxCompute, MySQL, OceanBase, Oracle, Paimon, PostgreSQL, and StarRocks. SQL and DataStream source connectors include MySQL, PostgreSQL, Oracle, SQL Server, MongoDB, OceanBase, TiDB, Db2, and Vitess. This breadth gives teams options across source databases and destination systems, though they should check that a needed connector and runtime combination fits their deployment.

Flink CDC 3.6 is compatible with Flink 1.20 and 2.2. Version 3.6.0 is built on JDK 11 and requires JDK 11 or later, so matching the Flink version and Java runtime is part of deployment planning. The project directs support questions to its user mailing list and problems to Flink JIRA.

Pricing

Apache Flink CDC is free and open source under the Apache 2.0 License. The Apache Flink CDC plan costs 0.00 USD per free and includes released JARs and connectors. There is no paid tier or trial to weigh against the free option; the cost tradeoff is that teams run the software in their own environment.

Platforms

Flink CDC supports Linux, macOS, and Windows, with self-hosted deployment. Its on-premise model is suited to organizations that want to run CDC within their own infrastructure, rather than rely on a hosted service.

Who it's for

Choose Flink CDC if your team can operate Apache Flink and wants to define data movement through YAML, SQL, or DataStream APIs. Full database and sharded-table synchronization, transformation, and schema evolution make it useful for varied replication jobs. It is a poor fit if you need a managed service or want to avoid maintaining a Flink deployment.

Pros and cons

  • Pro: Multiple APIs accommodate declarative pipelines as well as custom Flink applications.
  • Pro: Broad source and pipeline connector coverage supports a range of database and sink combinations.
  • Pro: The Apache 2.0 license means no software fee for the released JARs and connectors.
  • Con: Self-hosted deployment requires teams to run and maintain Apache Flink and its compatible runtime.
  • Con: Flink and JDK compatibility must be checked for the version being deployed.

Alternatives

Canal is another free, open-source option for teams seeking CDC software across Linux, macOS, or self-hosted environments.

pgstream is free and open source on Linux, macOS, and self-hosted platforms; consider it as another CDC option.

dbmazz offers a free self-hosted option, including a CLI and terminal dashboard, local web UI, and community support.

Estuary is a freemium alternative with API, web, and self-hosted platforms, a free plan, and a $100.00 USD monthly pay-as-you-go plan; it also offers a free trial.

Striim offers freemium access across API, web, and self-hosted platforms, with a developer plan capped at 25 million events per month and a free trial.

TiCDC is a free option for teams looking for open-source CDC software across API, Linux, macOS, and self-hosted platforms.

Decodable has a free plan with 20 streams, four running tasks, and 24 hours or 10GiB of stream retention; choose it if those stated limits suit your needs and you want a web or API option.

PeerDB offers a free open-source plan aimed at individuals, with Docker setup and CDC, streaming query, and query layer features, plus a free trial.

Browse more options in Change Data Capture Software or compare licensing-focused tools in Open Source License Compliance Software.

Verdict

Apache Flink CDC is a strong choice for teams that already operate Flink and need flexible ways to synchronize and transform database changes without a software license fee. Its main reason to look elsewhere is operational: a self-hosted Flink deployment and compatible JDK are prerequisites, not optional extras.

Apache Flink CDC plans and pricing

All plans
Apache Flink CDC Free Apache License 2.0 · released JARs and connectors github.com · 30 Sept 2026

Compared on open source license compliance software

Deployment model
self-hostedgithub.com
Capture method
log-basedgithub.com
Schema evolution
Yesgithub.com
Initial snapshot
Yesgithub.com

Facts

Purpose
Flink CDC is a distributed data integration tool for real-time and batch data, built on Apache Flink.github.com · 30 Sept 2026
Synchronization
It supports full database synchronization and sharded table synchronization.github.com · 30 Sept 2026
Schema and transformation
It supports schema evolution and data transformation.github.com · 30 Sept 2026
Pipeline API
Its YAML Pipeline API lets users define sources, sinks, routing, transformations, and schema evolution rules declaratively.github.com · 30 Sept 2026
Other APIs
It also provides SQL and DataStream APIs for defining CDC sources and custom Flink streaming applications.github.com · 30 Sept 2026
Integrations
Pipeline connectors listed include Doris, Elasticsearch, Fluss, Hudi, Iceberg, Kafka, MaxCompute, MySQL, OceanBase, Oracle, Paimon, PostgreSQL, and StarRocks.github.com · 30 Sept 2026
Source databases
SQL and DataStream source connectors listed include MySQL, PostgreSQL, Oracle, SQL Server, MongoDB, OceanBase, TiDB, Db2, and Vitess.github.com · 30 Sept 2026
Flink compatibility
The repository lists Flink CDC 3.6 as compatible with Flink 1.20 and 2.2.github.com · 30 Sept 2026
Runtime requirement
The documentation says Flink CDC 3.6.0 is built on JDK 11 and requires JDK 11 or later.nightlies.apache.org · 30 Sept 2026
Security
The repository links to a security policy, but its contents were not available in the opened page.github.com · 30 Sept 2026
Support
The project directs users to its user mailing list for support and questions, and to Flink JIRA for problems.github.com · 30 Sept 2026
License
The repository states that Flink CDC is distributed under the Apache 2.0 License.github.com · 30 Sept 2026
Deployment
The documentation describes downloading and extracting Flink CDC and connector JARs, then running it with an Apache Flink deployment.nightlies.apache.org · 30 Sept 2026

Best Apache Flink CDC alternatives

See all 20

Where it ranks on Everything Xiaomi

Is Apache Flink CDC yours?

Claim it for free: prove the domain, then correct facts, plans and screenshots. An editor reviews every change.

Sources