Чтобы адаптировать резюме под вакансию или составить сопроводительное письмо, загрузи резюме
О рекламодателе
ОБЩЕСТВО С ОГРАНИЧЕННОЙ ОТВЕТСТВЕННОСТЬЮ "ЦЕНТР НАЦИОНАЛЬНЫХ ИНТЕЛЛЕКТУАЛЬНЫХ СИСТЕМ" ИНН: 9704271170
описание
09B7351
Referment is working with a global options market maker whose low-latency trading platform runs on self-managed, on-premises infrastructure. The firm trades across global derivatives markets, uses open-source tooling where it fits, and values a flat, collaborative engineering culture.
задачи
Design, build, and operate the on-premises Kubernetes platform and the underlying Linux VM fabric
Design the hypervisor architecture, host networking topology, and storage backing
Design and maintain distributed shared storage such as Ceph
Build VM templating and golden-image pipelines
Automate the VM lifecycle using infrastructure as code
Manage capacity planning, oversubscription strategy, and headroom for failures or maintenance
Own virtual networking within the fabric and its handoff to the Kubernetes CNI layer
Manage the full lifecycle of on-premises Kubernetes clusters, including bootstrapping, upgrades, scaling, and decommissioning
Manage the control plane, including etcd operations, backup and restore, performance tuning, and disaster recovery
Configure CNI, network policy enforcement, and on-premises load balancing
Set up ingress and internal DNS
Own persistent storage integration via CSI drivers
Define and enforce multi-tenancy patterns across the VM and Kubernetes layers
Build GitOps-based delivery with ArgoCD or Flux
Harden hosts and clusters against CIS benchmarks and manage secrets
Build observability across the stack, from hypervisor health to cluster metrics and logs
Plan and execute upgrades with minimal workload disruption
Troubleshoot incidents across the full stack, from pods and Kubernetes components to the underlying VM and hypervisor
Share an on-call rotation for platform-level incidents
Partner with application teams on developer experience
Coordinate with datacentre and network teams on physical host provisioning
требования
Production experience running self-managed, on-premises Kubernetes, including operating etcd and the control plane
Hands-on experience designing and operating a Linux KVM/libvirt-based VM fabric, ideally from an early stage
Strong Linux systems administration background, including networking, storage, service management, kernel tuning, and troubleshooting under pressure
In-depth knowledge of Kubernetes networking and storage, including CNI internals, CSI drivers, and distributed storage systems such as Ceph or Longhorn
Experience with infrastructure as code and configuration management, including Terraform, Ansible, or Packer; GitOps workflows; and CI/CD pipelines
Familiarity with observability stacks such as Prometheus/Grafana and ELK/Loki, and using them to diagnose infrastructure issues without cloud-native tooling
Familiarity with hardening standards, RBAC, network segmentation, and secrets management
Strong troubleshooting skills across the hypervisor, OS, network, container runtime, and Kubernetes control plane
Strong written and verbal communication skills for working with distributed or hybrid teams
Будет плюсом: service meshes, bare-metal Kubernetes provisioning, experience in regulated or air-gapped/restricted-network environments, contributions to open-source infrastructure tooling, use of AI tooling to accelerate development, experience running AI infrastructure
условия
Full-stack platform ownership without managed services between the team and root cause