15 авг

platform engineer for AI workloads

ориентир по рынку
вакансия зп не указана
в среднем 325 071 ₽
Загрузи резюме, чтобы видеть мэтчи с вакансией

подготовьтесь к отклику

ai-инструменты

Чтобы адаптировать резюме под вакансию или составить сопроводительное письмо, загрузите резюме

описание

Zencargo provides logistics and supply chain technology for businesses, supporting freight operations through software, cloud infrastructure, and AI-powered workloads.

задачи

  • Lead the design, implementation, and delivery of complex infrastructure projects;
  • Build and scale infrastructure for AI workloads, including cost, API reliability, and platform support for LLM-powered features;
  • Debug cluster behavior under load, service-to-service traffic, event-backbone consumer lag, and database performance;
  • Own the infrastructure-as-code estate, keep drift visible and reconciled, and set module standards for the wider team;
  • Advance the delivery pipeline through policy-as-code gates, progressive delivery, and tested, reviewed, auditable changes;
  • Contribute to internal AI tooling, including the MCP server and agentic workflows that reduce operational toil;
  • Engineer cost controls through right-sizing, reserved capacity, and waste removal against a measured baseline;
  • Advocate for targeted infrastructure spend;
  • Mentor peers through pairing, code review, and knowledge sharing;
  • Cover security and governance alongside platform work in collaboration with the infrastructure team;
  • Deliver operational outcomes that reduce manual effort and improve business speed, accuracy, cost, or service quality;
  • Improve platform reliability and incident outcomes through systemic fixes;
  • Deliver complex infrastructure projects end to end with recorded and followable decisions;
  • Maintain infrastructure-as-code and pipeline health, including drift reconciliation and wider adoption of standards.

требования

  • Write maintainable software in at least one programming language and move between languages when required;
  • Own production infrastructure on a major cloud, including during incidents;
  • Understand infrastructure-as-code state, module design, and reviewed, repeatable changes;
  • Debug misbehaving Kubernetes clusters in production;
  • Design CI/CD pipelines rather than only consume them;
  • Use AI tooling fluently and critically while understanding its limitations;
  • Work effectively in a fully remote, async-first environment;
  • Communicate a plan and work methodically under pressure during platform incidents;
  • Nice to have: Experience in freight, logistics, supply chain, B2B SaaS, or operational technology; workflow automation or orchestration experience, such as n8n; experience with service mesh, event streaming at scale, or managed Postgres-compatible databases under real load; familiarity with identity and access management, SSO, secret management, or SRE methodology; experience building with LLM APIs, MCP servers, or agentic tooling.

условия

  • London/Europe with GMT+0 to GMT+4 overlap with the UK team;
  • Fully remote role with offices in multiple locations;
  • Async-first work culture.

Если просят выйти из iCloud, прислать код из SMS, запустить или установить что-то, перевести деньги — не соглашайтесь: это мошенничество.

Про зарплаты

Анонимные данные по зарплатам и грейдам.
Можно сверить вилку с рынком.

Посмотреть зарплаты

Если просят выйти из iCloud, прислать код из SMS, запустить или установить что-то, перевести деньги — не соглашайтесь: это мошенничество.