In this project, we will launch a virtual machine on the cloud, and using a docker-compose.yml file, Kafka, Debezium, Spark, and PostgreSQL will be pre-configured and set up. We will demonstrate each step in detail.(You can begin with free 300 $ account in Google cloud)
github.com/Data-94/Building-a-Real-Time-CDC-System-with-Kafka-Spark-Databricks.git
medium.com/@emre-b-bayraktar/building-a-real-time-cdc-system-with-kafka-spark-databricks-70d1f10deec6
#ApacheKafka #ApacheSpark #PostgreSQL #Debezium #DockerCompose #VirtualMachine #GoogleCloud #CloudComputing #DataEngineering #BigData #RealTimeAnalytics #DataPipeline #StreamingData #EventDrivenArchitecture #CloudDeployment #OpenSourceTools #DataIntegration #MLOps #DataScience #StreamingAnalytics