вчера

product manager for AI safety

ориентир по рынку
вакансия зп не указана
в среднем 312 061 ₽
Загрузи резюме, чтобы видеть мэтчи с вакансией

подготовьтесь к отклику

ai-инструменты

Чтобы адаптировать резюме под вакансию или составить сопроводительное письмо, загрузите резюме

описание

Cohere is an enterprise AI company that builds foundation AI models and end-to-end products for real-world business problems. Its North platform securely deploys AI agents and automations within organizational infrastructure while supporting data privacy, compliance, workflow automation, and actionable insights.

задачи

  • Serve as the product bridge between Cohere's safety research teams and North, translating model evaluations, red-teaming, and behavioral research into product-level guardrails, controls, and safeguards;
  • Own the safety product roadmap for Cohere and North, prioritizing features based on research findings, misuse patterns, threat vectors, and customer requirements;
  • Partner with modeling teams to scope and interpret safety evaluations across adversarial inputs, edge cases, and high-stakes use cases;
  • Define and drive evaluation frameworks for assessing safety properties as models and product capabilities evolve;
  • Coordinate the development of guardrails and intervention mechanisms across research, engineering, and policy;
  • Monitor the AI safety research landscape, including prompt injection, jailbreaks, and emerging misuse patterns in agentic systems;
  • Build processes for scaling safety review and assessing new features for safety risk before launch.

требования

  • 5+ Years of product management or research operations experience, including meaningful work alongside research or ML teams at a technology or AI company;
  • Technical depth to engage credibly with safety researchers and understand evaluation findings;
  • Genuine interest in AI safety, model behavior, and the implications of deploying LLMs in enterprise contexts;
  • Comfort operating in ambiguity and using judgment to determine what to act on and how quickly;
  • Ability to align researchers, engineers, and product teams without losing research nuance;
  • Strong written communication skills and ability to translate complex model behavior findings for non-technical audiences;
  • Nice to have: hands-on experience with LLM evaluation, red-teaming, safety benchmarking, or behavioral research; familiarity with prompt injection, jailbreaks, RAG poisoning, or misuse patterns in agentic systems; background in trust and safety, content policy, or research-adjacent operations; experience building zero-to-one processes in research or safety contexts; exposure to agentic AI systems and safety challenges involving tool use, multi-step reasoning, and autonomous execution.

условия

  • Competitive compensation and equity options;
  • Professional development opportunities;
  • Weekly lunch stipend of $75/£75 or equivalent in local currency;
  • Full health and dental benefits, including a separate mental health budget;
  • RRSP matching, 401K, and Pension Scheme;
  • 100% Parental leave top-up for up to 6 months for either parent;
  • Annual enrichment benefits covering arts and culture, fitness and wellness, quality time, and workspace improvement;
  • Education and learning stipend for conferences, courses, and coaching;
  • 6 Weeks of paid vacation;
  • Budget for travel to other offices for remote employees and an annual company offsite;
  • Co-working benefit for employees not near an office;
  • $500 Home office stipend.

Если просят войти через iCloud, отправить коды из SMS, запустить код, что-то установить, перевести деньги или сделать что угодно, связанное с деньгами, не соглашайтесь: это признаки мошенничества.

прозрачные зарплаты в IT

Анонимные данные по зарплатам и грейдам

Посмотреть
График динамики зарплат
Откликнуться Добавить в трекер

Если просят войти через iCloud, отправить коды из SMS, запустить код, что-то установить, перевести деньги или сделать что угодно, связанное с деньгами, не соглашайтесь: это признаки мошенничества.