All fields marked * are required
Design, develop and optimise scalable ETL/ELT pipelines using Databricks.
• Develop production-grade data transformations using Python, PySpark and SQL.
• Build Bronze/Silver/Gold data layers using Delta Lake and medallion architecture.
• Implement batch and, where required, streaming ingestion patterns.
• Work with Databricks Workflows and Unity Catalog for orchestration and governance.
• Integrate Databricks with Azure data services including ADLS Gen2 and Synapse.
• Implement data-quality checks, error handling, monitoring and performance optimisation.
• Develop automated deployments using CI/CD and follow software engineering best practices.
• Troubleshoot production data pipelines and ensure reliability and scalability.
Mandatory skills:
• Strong hands-on Databricks, Python/PySpark and SQL.
• Delta Lake and Databricks Workflows.
• ETL/ELT development and performance tuning.
• ADLS Gen2 and Azure data ecosystem.
• Unity Catalog exposure.
• Git and CI/CD experience.
Required
Preferred