Чтобы адаптировать резюме под вакансию или составить сопроводительное письмо, загрузи резюме
описание
The Sustainability Data & AI team is building and operating a modern Sustainability Data Foundation based on Databricks. It supports sustainability reporting, regulatory compliance and analytics.
задачи
Design and develop data pipelines and data products using Databricks, Delta Lake and PySpark
Build data models and transformation frameworks for large and complex datasets
Develop data quality controls, validation and reconciliation processes
Optimize pipelines for performance, scalability and cost
Implement CI/CD, automated testing and deployment processes
Support data governance, lineage, security and access management
Work closely with Sustainability, Business, Product and Reporting teams
Troubleshoot issues, perform root-cause analysis and ensure reliable production delivery
Develop Databricks-based tools to simplify data access and reporting
требования
5+ Years of experience in data engineering and data platforms
Strong production experience with Databricks
Expert knowledge of PySpark, SQL, data modelling and data pipeline design
Experience with CI/CD, Git, automated testing and deployment
Knowledge of data quality, governance, lineage and auditing
Strong troubleshooting and root-cause analysis skills
Ability to take ownership and deliver pragmatic, production-ready solutions
Будет плюсом: Experience with ESG / sustainability or regulatory reporting, knowledge of ERP, procurement, finance or supplier data, experience building Databricks applications, dashboards or user-facing tools