-
The candidate should have strong hands-on experience in cloud data architecture, Databricks platform design, performance optimization, migration, data modeling, streaming, and security.
-
Design Databricks Lakehouse architecture using Delta Lake and Bronze/Silver/Gold Medallion layers.
-
Build pipelines using Databricks Workflows, Delta Live Tables, Auto Loader, Azure Data Factory, and Databricks SQL.
-
Work with Azure Data Lake Storage Gen2, Synapse, PySpark, Spark SQL, and SQL for enterprise data solutions.
-
Optimize Spark workloads through partitioning, cluster sizing, data skew handling, broadcast joins, and Photon acceleration.
-
Lead migration from SAS, Hadoop, Teradata, Oracle, SQL Server, Synapse, and legacy ETL platforms.
-
Define data modeling, domain architecture, batch/streaming, event-driven processing, governance, security, lineage, and access controls.