Если вы раньше входили через Google, сбросьте пароль для своей Gmail-почты через кнопку «Забыли пароль?» на экране входа. Затем войдите по email и новому паролю.
Если аккаунта ещё нет, зарегистрируйтесь с Gmail-почтой, после подтверждения почты мы предложим задать пароль.
Что нового
Загружаю обновления...
Что нового
Загружаю обновления...
Работа найдется быстрее с подпискойКандидат найдётся быстрее с подпиской
Чтобы адаптировать резюме под вакансию или составить сопроводительное письмо, загрузите резюме
описание
Capital.com is a leading trading platform that expands globally with award-winning products recognized for cutting-edge technology and seamless client experience. The company builds out its observability practice and seeks a senior engineer to own the telemetry stack end to end within a hybrid AWS and on-premise environment.
задачи
Own the full observability stack including metrics, logs, and traces from pipeline design to day-2 operations;
Architect and run VictoriaMetrics cluster topology including scraping, remote write configuration, alerting rules, and cardinality control;
Operate OpenSearch clusters including index lifecycle management, hot-warm-cold architecture, shard tuning, and ingest pipelines;
Build and maintain OpenTelemetry Collector pipelines and instrument services across Java, Python, and JS/TS stacks;
Run Kafka as the telemetry transport layer including topic design, partition strategy, and throughput tuning;
Manage log shipping infrastructure and define structured logging standards and field normalization;
Build actionable Grafana dashboards and alerting systems;
Collaborate with platform and application teams to improve sampling strategies, batching, and context propagation;
Contribute to incident response, post-mortems, and reliability improvements;
Mentor engineers on observability practices, tooling, and structured logging standards.
требования
6+ Years in a DevOps, SRE, or platform engineering role with at least 2 years focused on observability at production scale;
Deep hands-on experience with VictoriaMetrics or Prometheus including MetricsQL/PromQL and retention management;
Solid OpenSearch or Elasticsearch skills including cluster operations, Query DSL, and ISM policies;
Production experience with OpenTelemetry including Collector configuration, OTLP, and instrumentation;
Strong Kafka skills including producer/consumer patterns, consumer group management, and JMX-based monitoring;
Proficiency with log shippers like Fluent Bit, Vector, or Fluentd and structured log parsing;
Working knowledge of Kubernetes, Argo CD, GitOps, and Terraform or Ansible;
Experience with hybrid AWS and on-prem environments and networking for scraping and shipping pipelines;
Scripting ability in Bash or Python;
Strong communication skills to explain observability tradeoffs;
English proficiency;
Nice to have: Strimzi experience.
условия
Competitive salary;
Comprehensive health and pension benefits;
Generous annual leave policy;
Employee referral program;
Workation policy allowing 30 days of remote work per year from anywhere;