Data engineers shipping production pipelines: ingest, transform, serve & trust the data.
Stand up data pipelines you can run in production and trust. Reach for this when you're moving data end to end - streaming ingestion off Kafka, batch transforms in Spark, an analytics store in ClickHouse - and need quality checks and query tuning so the numbers downstream are correct and fast. Built for the engineer who owns the pipeline, not the analyst querying its output.
Click to play with sound.
Arranged in the author's recommended order. Walk through them in sequence, or open any one on its own.