Stream & Optimize Real-Time Data Flows

External: Coursera Courses ↗ · Coursera

Open Course on External: Coursera

Free to audit · Opens on External: Coursera

Stream & Optimize Real-Time Data Flows

Coursera · Intermediate ·🔄 Data Engineering ·4mo ago

Key Takeaways

Designs, implements, and optimizes production-ready streaming data pipelines using Apache Kafka and Flink

Original Description

Master the design, implementation, and optimization of production-ready streaming data pipelines using Apache Kafka and Flink. This intermediate-level course teaches you to evaluate log configurations against governance requirements (PCI-DSS, GDPR, SOC2) and cost constraints, design stream processing topologies that join and aggregate data in real time with exactly-once semantics, and optimize pipelines through partition tuning, compression, and cost modeling. You'll work through hands-on labs that mirror real-world scenarios at DoorDash, Netflix, and Robinhood: comparing retention policies against compliance rules, building a Kafka Streams application that joins orders and payments to calculate 5-minute revenue totals, and diagnosing performance bottlenecks to meet SLAs within budget. Intermediate data engineers and platform engineers who build or operate real-time streaming systems and want to master Kafka/Flink governance, joins, windowing, and cost-optimized scaling. Understanding of distributed systems, basic Apache Kafka knowledge, familiarity with SQL and streaming concepts, Python or Java programming experience. By the end, you'll design and optimize a multi-tenant streaming platform with governance controls—skills directly applicable to streaming data engineer, real-time platform engineer, and data infrastructure roles.
Watch on External: Coursera ↗ (saves to browser)
Sign in to unlock AI tutor explanation · ⚡30

Related Reads

📰
Data Engineering ETL Project
Learn how to create a data engineering ETL project using Python and improve your data processing skills
Medium · Python
📰
Data Engineering: A Simple Guide to Building a Career in the World of Data
Learn the fundamentals of data engineering and kickstart your career in the field with key skills like SQL, Python, and cloud technologies
Medium · Python
📰
Windmill for Data Engineering: TypeScript/Python Scripts, Flows & Self-Hosted OSS
Learn how Windmill simplifies data engineering with TypeScript/Python scripts, flows, and self-hosted OSS, streamlining orchestrators, internal tools, and secret management
Medium · Python
📰
I Built My Second ETL Pipeline. This Time, I Started Thinking Like a Data Engineer
Learn how to build a production-ready ETL pipeline with Python, Docker, PostgreSQL, and Kestra by thinking like a data engineer
Towards Data Science
Up next
A Moment Frozen in Time | Arnav Iyengar | TEDxJenks Youth
TEDx Talks
Watch →