LIVE Data Engineering Bootcamp for Analysts: Become the End-to-End Data Professional

Starts 31st Oct 168+ learners across cohorts 1 & 2

From Data Analyst to Data Engineer in 11 Weeks.

Cohort 3 · Starts Saturday, 31st Oct, 2026. Enrollment opens 15th October. Move from data analyst to data engineer through Codebasics' live 11-week cohort for working professionals. Learn warehouse-grade SQL, production Python, PySpark, Delta Lake, Azure, Microsoft Fabric, dbt, Airflow, CI/CD and streaming while building and shipping one end-to-end pipeline. Classes run live every Saturday and Sunday, 5 to 8 PM IST, with doubt clearing in every session and recordings for revision. The Diwali weekend, 7th and 8th November, is a planned break.

11

Weeks


20

Live Sessions


10+

Tools & Platforms

8

Capstone Layers


1 Year

Access to Class Recordings


0

Enrolled Learners

LIVE Data Engineering Bootcamp for Analysts: Become the End-to-End Data Professional
US$630 US$840 after 27th Oct · Save US$210

Created by :

Dhaval Patel, Hemanand Vadivel, Kirandeep Marala & Harun Raseed Basheer
Join our Bootcamp

Join our Bootcamp

In-demand Skills

Learn In-demand Skills

Get Hired

Get Hired by Top Companies

We email you once the enrollments open.

11

Weeks


20

Live Sessions


10+

Tools & Platforms

8

Capstone Layers


1 Year

Access to Class Recordings


What Makes This Bootcamp Different?

  • 100% live, instructor-led sessions across 11 weeks, with 30 minutes of doubt clearing in every session.

  • Built to make you the end-to-end data person on your team: the one who builds the pipeline, models the data and ships the report.

  • The full modern stack, in order: warehouse-grade SQL, production Python, PySpark, Delta Lake, Azure, Microsoft Fabric, dbt, Airflow, CI/CD and Kafka streaming.

  • Led by Harun, Microsoft MVP and Databricks Certified Professional, with practitioners who build and run production data platforms.

  • Production-first: execution plans, the Spark UI, OPTIMIZE and ZORDER, schema evolution, automated tests, freshness checks and incident handling.

  • Metadata-driven orchestration: one control table driving hundreds of tables through Data Factory, with a deliberate failure path.

  • Career built in: stakeholder management and personal branding, an open career guidance session, and interview prep that starts from your own resume.

  • A capstone thread, not a last-minute assignment: the brief on 20th December, a midweek jamming session on 6th January and your demo on 16th January.

  • One pipeline repo covering eight layers, from API ingestion to Power BI, that a hiring manager can read in five minutes.

Learn From

Your Instructors

Dhaval Patel

Dhaval Patel

Founder, Codebasics · Ex-NVIDIA

Curriculum advisor. Sets what this cohort teaches and what it refuses to.

I have 17 years of experience in programming and data science working for big tech companies like NVIDIA and Bloomberg. I also run a famous youtube channel called Codebasics where I pursue my passion for teaching. Codebasics is one of the top channels on youtube when it comes to data science, machine learning, data structures, etc. I firmly believe that “Anyone can code” and I use analogies, simple explanations, and step-by-step storytelling to explain difficult concepts in such a way that even a high school student can understand them easily.

Hemanand Vadivel

Hemanand Vadivel

CEO, Codebasics

Leads most sessions and every leadership and management block.

I’m a mechanical engineer who transitioned to a full time Data & Analytics manager in the UK & Germany by teaching myself Power BI, excel & anything else that was required to solve the problem. I have worked with complex data across Supply Chain, Sales, Marketing, Revenue Management, Finance and HR functions over the last 8 years to deliver effective solutions. To me, Analytics is an extra sensory organ that anyone can develop to become a superhuman at work. I have conducted several workshops and internal training to help non-technical people to become superhuman at work and now, really excited for this opportunity to share all that knowledge with you.

Kirandeep Marala

Kirandeep Marala

Analytics Engineer and Programme Manager, Codebasics

Plans every session, builds the assignments, and evaluates what you ship.

I started on the analytics side four years ago, answering business questions with SQL and dashboards. The problems I found most interesting sat one layer down, in how the data got there, and that pulled me towards engineering. Today I manage the Data Engineering Bootcamp at Codebasics: I plan sessions with Harun, turn each one into assignments and study notes, evaluate what you ship and relay your questions live. Every assignment uses a fresh business case, because I want transfer, not recall. A pipeline you cannot explain is not finished yet.

Harun Raseed Basheer

Harun Raseed Basheer

Team Lead, Data Engineering, Hitachi Solutions

Leads every technical session, from SQL execution plans in week one to Kafka in week ten. Teaches on the whiteboard first and in code second, then stays for the doubt clearing.

I lead a data engineering team that builds and runs enterprise data platforms on Azure and Databricks. Ten years in data have taught me one thing: the hard part is not writing the code, it is knowing what the engine actually did with it. So I teach whiteboard first, then the code, then the execution plan or Spark UI, so you can check it did what you expected. I am a Microsoft Certified Trainer and a Databricks Certified Professional, and I lead every technical session here. By the end, you should read your own pipeline the way I read mine.

Hear It From

Our Happy Learners

Our content is rated 4.9/5 from 18555+ Learners

Overview

What you'll learn in
this Live Data Engineering for Data Analyst Bootcamp

Week-1: Foundations & Advanced SQL

SQL Engineering

  • Session 1 - SQL for Modern Data Engineering, Part 1

    Joins beyond INNER and OUTER · Join algorithms: nested loop, hash, merge · Anti-joins · Window functions deep dive · Partition, order and frame · Ranking and running totals

    Output: A ranking and running totals query pack on a sample warehouse


  • Session 2 - SQL for Modern Data Engineering, Part 2

    Recursive CTEs · MERGE statements · Incremental loading patterns · CDC concepts · Query execution plans · Warehouse optimisation

    Output: An incremental load built with MERGE and tuned from the execution plan

  • Session 3 - Data Modelling & Warehouse Engineering

    OLTP vs OLAP · Star and snowflake schema · Fact vs dimension tables · SCD Type 1 and Type 2 · Partitioning strategies · Medallion architecture · Data contracts

    Output: A dimensional model with SCDs on a Bronze, Silver, Gold layout


  • Session 4 - Production Python for Data Engineers

    Modular Python architecture · OOP for pipelines · Config-driven frameworks · Logging and exception handling · Retry mechanisms · Environment management · Secrets handling

    Output: A config-driven Python pipeline framework with logging and retries

  • Session 5 - Advanced Python Data Processing

    APIs and ingestion patterns · Async processing · Parallel execution · File streaming · Memory optimisation · Testing with pytest · Packaging basics

    Output: A tested async ingestion job that streams large files without blowing memory


  • Session 6 - PySpark Deep Dive

    Spark architecture · Executors and DAGs · Lazy evaluation · Partitioning · Broadcast joins · Shuffle optimisation · Spark UI analysis · Caching strategies

    Output: A Spark job you have tuned yourself using the Spark UI

  • Session 7 - Delta Lake & Lakehouse Engineering

    Delta internals · ACID transactions · OPTIMIZE and ZORDER · Time travel · Schema evolution · Change Data Feed · Incremental ETL · Bronze, Silver, Gold

    Output: An incremental ETL pipeline running on Delta Lake


  • Session 8 - Azure Data Engineering Stack

    ADLS Gen2 · Event Hubs · Key Vault · Managed identities · Integration Runtime · Networking basics · Synapse vs Databricks vs Fabric

    Output: A secured Azure data landing zone using Key Vault and managed identities

  • Session 9 - Enterprise Data Pipelines

    Azure Data Factory · Fabric Pipelines · Databricks Workflows · Metadata-driven pipelines · Config-based orchestration · Parameterisation · Reusable frameworks

    Output: A metadata-driven, parameterised orchestration pipeline


  • Session 10 - Microsoft Fabric Engineering

    OneLake · Lakehouse and Warehouse · Fabric Data Factory · Eventstream · Real-Time Intelligence · DirectLake · Fabric governance

    Output: An end-to-end Fabric solution with Direct Lake reporting on top

  • Session 11 - Connecting the Dots

    End-to-end view of the stack so far · How the pieces fit together · Architecture recap · Trade-offs between platform options · Concept clarity and doubt clearing

    Output: One architecture diagram of the full stack, explained in your own words

  • Session 12 - dbt Core Fundamentals

    Models and sources · refs() · Materialisations · Snapshots · Incremental models · Tests · Documentation

    Output: A tested, documented dbt project with incremental models and snapshots


  • Session 13 - CI/CD & Reliability Engineering, plus Capstone Introduction

    Git branching strategies · GitHub Actions · Automated testing · Deployment pipelines · Monitoring and freshness checks · Cost optimisation · Incident management · Capstone brief and assessment criteria

    Output: A CI/CD workflow with automated tests and freshness checks, plus your capstone brief

  • Session 14 - Stakeholder Management & Personal Branding

    Stakeholder communication · Framing requirements · Explaining technical work to non-technical audiences · Personal branding · Building online credibility as a data engineer

    Output: A stakeholder-ready update on your own work and a refreshed professional profile

  • Session 15 - Capstone Jamming Session

    Implementation questions on your capstone · Design review · Debugging together · Unblocking pipeline issues · Peer feedback

    Output: Your capstone unblocked, with a clear next step


  • Session 16 - Advanced Analytics & Apache Airflow Engineering

    Macros and Jinja · Semantic layer · MetricFlow · SQLFluff · Lineage · Data quality frameworks · Governance · DAG architecture · Dynamic DAGs · Sensors · XCom · Scheduling and monitoring · Retry patterns · Failure handling

    Output: A semantic layer with lineage, plus scheduled Airflow DAGs that survive failures


  • Session 17 - Streaming Data Engineering

    Kafka fundamentals · Event-driven architecture · Structured Streaming · Watermarking and windowing · Event-time processing · CDC streaming · Event Hubs integration

    Output: A streaming pipeline that handles late-arriving data with watermarks

  • Session 18 - Interview Prep & System Design

    SQL interview rounds · PySpark interview questions · Data modelling rounds · System design · Resume transformation · LinkedIn optimisation · Mock interviews

    Output: A DE-positioned resume and profile, tested in a mock interview


  • Session 19 - Final Demo & Graduation

    Capstone demo covering API ingestion · Lakehouse architecture · PySpark transformations · dbt modelling · Airflow orchestration · CI/CD · Power BI reporting · Monitoring

    Output: One production-grade pipeline repo covering all eight layers, ready for a hiring manager to read in five minutes

This cohort includes Data Engineering Bootcamp 1.0 worth US$660

The DE Promise

Build & Ship Production Data Pipelines in 10 Weeks.

Not a tutorial. A working end-to-end pipeline defended in front of mentors.

What you get Self-study Other live bootcamps Our live cohort
Live, mentor-led sessions
Real projects shipped to GitHub Rarely Sometimes 1 pipeline repo
The current 2026 Cloud stack On your own Often outdated
Job assistance
Investment Your time Usually far higher US$630

May we help you?

Frequently Asked
Questions

Q.1 When does the bootcamp officially start?

The bootcamp officially commences on Saturday, 29th August 2026.

Access is immediate. The moment you enrol you get the recordings of every session already delivered, the materials, and Discord access. Live sessions continue every Saturday and Sunday.

The Inner Circle was early enrollment at a reduced price and closed on 2nd September 2026. Standard pricing now applies and enrollment closes on 5th September 2026. Learners who enrolled before the bootcamp began on 29th August also received a short form asking which tools and gaps matter most in their work, and those inputs shaped what this cohort covers.

Yes. Every enrollment includes full access to the Data Engineering Bootcamp 1.0 at no extra cost. It includes Job Assistance, Live Problem Solving, and a Virtual Internship. You get both for the price of one.

Saturdays and Sundays, 5 to 8 PM IST. Sessions are fully live and interactive with hands-on labs and real-time Q&A. Recordings are available for revision.

All live sessions are recorded and available within 24 hours. You can catch up at your own pace, though live attendance is strongly recommended as the labs and discussions are where most of the real learning happens.

No. You keep access to all session recordings for 1 year from the bootcamp start date.

Q.1 Do I need prior data engineering experience?

No. This bootcamp is built for working data analysts who want to cross into data engineering. If you have at least 1 year of analyst experience and are comfortable with SQL, you are ready.

We strongly advise against it. This bootcamp moves fast and assumes analyst-level SQL fluency and data literacy. If you are starting from zero, the Codebasics Data Analytics Bootcamp is the right first step. Build that foundation and come back.

Working data analysts, BI developers, and business analysts with 1 to 4 years of experience who want to own the full data stack, not just the dashboard layer. If you already work as a data engineer, this bootcamp is likely below your current level.

Everything unlocks the same day you enrol: the Session 1 recording delivered on 29th August, all code, notebooks and materials from those sessions, Discord access so you can ask catch-up questions straight away, and every live session from 5th September onward on Saturdays and Sundays, 4 to 7 PM IST. You are a couple of sessions behind out of nineteen across a ten-week programme, so there is room to catch up.

Q.1 How do I get help if I am stuck?

Every enrolled learner gets access to the Discord community where you can ask questions, connect with fellow learners, share progress, and learn from each other throughout the bootcamp. The mentor team also provides weekly hands-on lab support.

The Data Engineering Bootcamp 1.0 included with your enrollment has dedicated job assistance. The bootcamp itself focuses on building your skills, shipping a production-grade capstone on GitHub, and preparing you for system design interviews with the core faculty.

Q.1 What is the Inner Circle price and when does it close?

Inner Circle pricing is open until 2nd September 2026. Standard pricing of ₹48,000 applies from 3rd September. Enrollment for Cohort 2 closes on the morning of 5th September 2026 and does not reopen.

The amount you paid for the Data Engineering Bootcamp 1.0 is fully adjusted and deducted from your enrollment fee.

The amount you paid for those individual courses is deducted from your enrollment fee.

Yes. Inner Circle pricing is now open until 2nd September 2026, extended from the original 25th August. Most of our learners are working professionals whose salaries arrive at the end of the month, and a deadline on the 25th asks people to make a considered decision in the week they have the least room to make it. The price is unchanged for everyone. Nobody who enrolled before 25th August has paid more than somebody enrolling on 2nd September.

2nd September 2026 is correct. The Inner Circle was originally set to close on 25th August and has since been extended. Some advertising created before the extension still shows the earlier date. You get the Inner Circle price on any enrollment completed on or before 2nd September.

There are two key dates. Inner Circle pricing of ₹36,000 is open until 2nd September 2026. From 3rd September the price is ₹48,000. Enrollment closes on the morning of 5th September 2026 and does not reopen, because Cohort 2 has already started.

Q.1 I used a subsidy and now want to refund this Bootcamp itself. What happens?

You get back the amount you actually paid for this Bootcamp. Your original purchase stays intact and you keep access to it.

No. Once your existing purchase is applied as a subsidy to reduce your price, that original purchase becomes non-refundable.

Full refund, no questions asked, if you request it on or before 2nd November 2026. That is after the first two live sessions, so you can watch the live sessions and see exactly how the bootcamp runs before deciding.

Q.1 When does the bootcamp officially start?

The bootcamp officially commences on Saturday, 29th August 2026.

Access is immediate. The moment you enrol you get the recordings of every session already delivered, the materials, and Discord access. Live sessions continue every Saturday and Sunday.

The Inner Circle was early enrollment at a reduced price and closed on 2nd September 2026. Standard pricing now applies and enrollment closes on 5th September 2026. Learners who enrolled before the bootcamp began on 29th August also received a short form asking which tools and gaps matter most in their work, and those inputs shaped what this cohort covers.

Yes. Every enrollment includes full access to the Data Engineering Bootcamp 1.0 at no extra cost. It includes Job Assistance, Live Problem Solving, and a Virtual Internship. You get both for the price of one.

Saturdays and Sundays, 5 to 8 PM IST. Sessions are fully live and interactive with hands-on labs and real-time Q&A. Recordings are available for revision.

All live sessions are recorded and available within 24 hours. You can catch up at your own pace, though live attendance is strongly recommended as the labs and discussions are where most of the real learning happens.

No. You keep access to all session recordings for 1 year from the bootcamp start date.

Q.1 Do I need prior data engineering experience?

No. This bootcamp is built for working data analysts who want to cross into data engineering. If you have at least 1 year of analyst experience and are comfortable with SQL, you are ready.

We strongly advise against it. This bootcamp moves fast and assumes analyst-level SQL fluency and data literacy. If you are starting from zero, the Codebasics Data Analytics Bootcamp is the right first step. Build that foundation and come back.

Working data analysts, BI developers, and business analysts with 1 to 4 years of experience who want to own the full data stack, not just the dashboard layer. If you already work as a data engineer, this bootcamp is likely below your current level.

Everything unlocks the same day you enrol: the Session 1 recording delivered on 29th August, all code, notebooks and materials from those sessions, Discord access so you can ask catch-up questions straight away, and every live session from 5th September onward on Saturdays and Sundays, 4 to 7 PM IST. You are a couple of sessions behind out of nineteen across a ten-week programme, so there is room to catch up.

Q.1 How do I get help if I am stuck?

Every enrolled learner gets access to the Discord community where you can ask questions, connect with fellow learners, share progress, and learn from each other throughout the bootcamp. The mentor team also provides weekly hands-on lab support.

The Data Engineering Bootcamp 1.0 included with your enrollment has dedicated job assistance. The bootcamp itself focuses on building your skills, shipping a production-grade capstone on GitHub, and preparing you for system design interviews with the core faculty.

Q.1 What is the Inner Circle price and when does it close?

Inner Circle pricing is open until 2nd September 2026. Standard pricing of ₹48,000 applies from 3rd September. Enrollment for Cohort 2 closes on the morning of 5th September 2026 and does not reopen.

The amount you paid for the Data Engineering Bootcamp 1.0 is fully adjusted and deducted from your enrollment fee.

The amount you paid for those individual courses is deducted from your enrollment fee.

Yes. Inner Circle pricing is now open until 2nd September 2026, extended from the original 25th August. Most of our learners are working professionals whose salaries arrive at the end of the month, and a deadline on the 25th asks people to make a considered decision in the week they have the least room to make it. The price is unchanged for everyone. Nobody who enrolled before 25th August has paid more than somebody enrolling on 2nd September.

2nd September 2026 is correct. The Inner Circle was originally set to close on 25th August and has since been extended. Some advertising created before the extension still shows the earlier date. You get the Inner Circle price on any enrollment completed on or before 2nd September.

There are two key dates. Inner Circle pricing of ₹36,000 is open until 2nd September 2026. From 3rd September the price is ₹48,000. Enrollment closes on the morning of 5th September 2026 and does not reopen, because Cohort 2 has already started.

Q.1 I used a subsidy and now want to refund this Bootcamp itself. What happens?

You get back the amount you actually paid for this Bootcamp. Your original purchase stays intact and you keep access to it.

No. Once your existing purchase is applied as a subsidy to reduce your price, that original purchase becomes non-refundable.

Full refund, no questions asked, if you request it on or before 2nd November 2026. That is after the first two live sessions, so you can watch the live sessions and see exactly how the bootcamp runs before deciding.

Cohort 3 · US$630 · Enrollments open 15th Oct, 2026

Talk to us Chat with us