Responsibilities
- Full management of the Databricks environment
- Environment configuration and workspace administration
- Governance & Storage: Implement data governance using Unity Catalog
- Manage structured storage with Delta Lake tables
- Write and optimize data processing logic using Python, SQL, and PySpark
- Manage code lifecycle and collaborative development using Git
- Data Warehousing & Integration (Important)
- Own end to end maintenance of data pipelines; ensure reliability from data ingestion through final delivery across multiple cloud platforms
Must Have Skills
- English advanced
- Databricks Ecosystem
- Full management of the Databricks environment
- Environment configuration and workspace administration
Important Skills
- Snowflake
- Adobe Data Ingestion
Nice to Have
- Redshift
- Orchestration
Desirable
- Google Cloud Platform
Desirable – Cloud Platforms
- Maintain data availability and basic operations in BigQuery
AWS Data Services
- Manage serverless querying with Athena
- Handle metadata management with AWS Glue