Если вы раньше входили через Google, сбросьте пароль для своей Gmail-почты через кнопку «Забыли пароль?» на экране входа. Затем войдите по email и новому паролю.
Если аккаунта ещё нет, зарегистрируйтесь с Gmail-почтой — после подтверждения почты мы предложим задать пароль.
Что нового
Загружаю обновления...
Что нового
Загружаю обновления...
Работа найдется быстрее с подпискойКандидат найдётся быстрее с подпиской
Чтобы адаптировать резюме под вакансию или составить сопроводительное письмо, загрузите резюме
описание
The company provides enterprise software products, open source solutions, and accelerators.
задачи
Develop and maintain reusable Python SDK modules for data pipeline operations, including compaction, quality checking, schema evolution, and table lifecycle management;
Design and build framework libraries following software engineering best practices for scalability and reliability across 100+ pipelines and diverse data domains;
Collaborate with Data Engineers to gather feedback, understand pain points, and iterate on the SDK;
Contribute to platform architecture discussions and performance tuning strategies;
Write high-quality documentation and contribute to enablement materials for framework consumers;
Help define and infuse data engineering best practices through enablement, SDKs, and templates;
Ensure data quality, consistency, and governance across Lakehouse environments.
требования
Strong experience as a Data Engineer with proficiency in Databricks;
3+ Years of experience building reusable libraries, SDKs, or internal developer tooling;
Knowledge of Data Mesh and Data Product concepts, including data ownership, domain-oriented design, and self-serve data platforms;
Deep expertise in Delta Lake, Delta tables, and compaction optimization for high-performance workloads;
Proven experience designing and maintaining complex data pipelines on cloud object stores such as ADLS;
Strong Python programming skills for data engineering workloads;
Solid understanding of Lakehouse architecture and best practices for large-scale data platforms;
Hands-on experience with data pipeline monitoring, troubleshooting, and performance tuning;
Familiarity with CI/CD and workflow orchestration using Databricks Jobs;
Experience working in agile teams with a focus on ownership, autonomy, and best practices;
Excellent problem-solving skills and ability to handle high-scale, complex data challenges;
English proficiency at B2 level or higher;
Nice to have: experience with agile methodologies such as Scrum and SAFe, experience developing enablement materials and technical documentation.