● Blog
Streaming, lakehouses, AI and forecasting: what worked, with the code to show it.
How I design and run streaming pipelines with Apache Kafka and Spark Structured Streaming for production workloads.
How to build retrieval-augmented generation that answers from your own data, and shows where each answer came from.
Patterns for building reliable, fast lakehouses with Delta Lake on Databricks, from partitioning to maintenance.
Building models that spot equipment failures before they happen, so repairs can be planned instead of rushed.