| Record Type: |
Electronic resources
: Monograph/item
|
| Title/Author: |
Practical data engineering with Apache projects/ by Dunith Danushka. |
| Reminder of title: |
solving everyday data challenges with Spark, Iceberg, Kafka, Flink, and more / |
| Author: |
Danushka, Dunith. |
| Published: |
Berkeley, CA :Apress : : 2025., |
| Description: |
xix, 252 p. :ill., digital ;24 cm. |
| [NT 15003449]: |
Part I: Data Lakehouses, Iceberg, Batch ETL, and Orchestration -- Chapter 1: Foundational Data Engineering Concepts -- Chapter 2: Building a Data Lakehouse with Apache Iceberg -- Chapter 3: Batch ETL Pipeline with Apache Spark -- Chapter 4: Data Visualization with Apache Superset -- Chapter 5: Workflow Orchestration with Apache Airflow -- Part II: Streaming Data and Real-time Analytics. - Chapter 6: Change Data Capture with Debezium and Kafka -- Chapter 7: Low-latency Analytics Dashboard with ClickHouse -- Chapter 8: Real-time Fraud Detection with Apache Flink -- Part III: Machine Learning and Generative AI -- Chapter 9: Building a Product Recommendation Engine with Spark MLlib -- Chapter 10: Vector Similarity Search with Postgres and pgvector. |
| Contained By: |
Springer Nature eBook |
| Subject: |
Data mining. - |
| Online resource: |
https://doi.org/10.1007/979-8-8688-2142-4 |
| ISBN: |
9798868821424 |