Data Engineering · Databricks & Lakehouse
Databricks Medallion Pipeline from AWS S3
Build a real lakehouse pipeline end to end, live — S3 to Bronze, Silver and Gold.
Most people learn Databricks notebook by notebook, so you finish knowing what a Delta table is without ever having built a pipeline. This session starts at an AWS S3 bucket of raw files and finishes with curated Gold tables: mount the bucket, land the data as-is in Bronze, clean, dedupe and join it in Silver with PySpark, then model the business tables reporting reads. Medallion architecture, built in the order it is actually built — and explained the way interviewers ask about it.
Agenda
What we'll cover
- AWS S3 source — mount the bucket and read the raw files
- Bronze — land the data as-is in Delta tables
- Silver — clean, dedupe and join it with PySpark
- Gold — curated business tables ready for reporting
- Why the medallion layers exist, and how to answer that in an interview
- Live Q&A at the end
Perfect for
Who it's for
Every attendee
What you get
- Live, hands-on session on Zoom
- Live Q&A with the instructor
- Session notes & key resources — emailed after
- Optional: that session's recording + resources — add for ₹99

Your host
Durgesh Yadav
Founder, PrepNPlaced · Sr. Data Engineer @ 7-Eleven · ex-Target · Instructor @ Scaler, Bosscoder, Newton School, GeeksforGeeks & Masai
Reserve your free seat
One quick step: sign in with Google, then reserve your free seat — your Zoom link lands in your email instantly.
Free forever · Sunday 23 August · 1:00 PM IST
Free live seat
Sunday · 1:00 PM IST