Перейти к основному содержимому
Mistral AI

Applied AI Engineer, Site Reliability Engineer - EMEA

УдалённоFrance, Netherlands, Germany +2 more только
Опубликовано
Роль
SRE
Занятость
Полная занятость
Размер компании
Стартап
Зарплата не указана
Проверьте доступность

Доступно для: FR, NL, DE, GB, CH only. Укажите, откуда вы работаете, чтобы проверить доступность.

Коротко по делу

You will be one of the founding engineers of the Applied AI SRE sub-team. Your mission, alongside the team, is to build and operate the framework to ensure Mistral’s solution delivery is reliable and sustainable - and applied uniformly across all our accounts, both Mistral-hosted and customer-hosted. You should already have a strong understanding of what operational excellence looks like, and you’re ready to scale your impact. You will operate in four concurrent modes: BUILD, RUN, ENABLE, SECURE. This is a framework-first, fleet management role at heart. If you're excited by the difference between solving one customer's problem and structurally solving the class of problem for every customer, this is the role.

Ключевые навыки

KubernetesSREObservability

Обязательные навыки

PrometheusGrafanaOpenTelemetryLokiTempoSignozTerraformAnsiblePythonGoLinux

Желательные навыки

Cloud securityApplication securityK8s securitySupply chain securitySBOMcosignSLSALLM serving

Обязательные языки

English

Чем предстоит заниматься

  • Design for a fleet of Mistral platforms and apps.
  • Build proactivity to reduce reactivity.
  • Productize reliability, author runbooks, create SLO templates, implement observability.
  • Operate the Tier-1 customer environments that Mistral are contracted to operate.
  • Ensure SLO compliance, own on-call and incident response, manage drift, partner with Technical Support as L3 escalation, champion high signal post-mortems.
  • Productize how Mistral deploy, secure, and scale our Applied AI solutions.
  • Engineer on-demand provisioning, author security baseline packages, embed security guardrails, automate everything.
  • Own the security operations layer for our customer-side deployments.
  • Lead CVE response across the fleet, ship supply-chain integrity controls (SBOM, signed images, provenance), co-page with InfoSec on security incidents, enforce secure-config baselines.

Что требуется

  • Fluent in English.
  • 5+ years in SRE, Production Engineering, or DevOps, with a record of shipping tooling.
  • Strong multi-tenant Kubernetes fluency, namespace segmentation, network policy, RBAC, admission control, operations at scale.
  • On-call discipline: incident response, blameless post-mortem culture, runbook-first mindset.
  • Observability stack in production: Prometheus, Grafana, OpenTelemetry, Loki, Tempo, Signoz.
  • Infrastructure as code: Terraform, Ansible (or close equivalents).
  • Proficient in Python and/or Golang for tooling and automation.
  • Security mindset: you treat secure-SDLC, CVE response, and supply-chain integrity as reliability properties of the shipped artifact, not as someone else's job.
  • Strong written communication skills: runbooks, post-mortems, and customer-facing incident comms are core deliverables of this role.
  • Comfortable operating with high autonomy in an ambiguous, fast-paced environment — and disciplined enough to defend the team's scope when work tries to spill in.
  • Solid Linux internals, networking debug, and distributed-systems fundamentals.
  • Strong plus: Cloud or application security background (AppSec, K8s security, supply chain — SBOM, cosign, SLSA). At least one of our early hires must bring this; if it's you, flag it.
  • Strong plus: Experience operating LLM / model-serving stacks in production
  • Strong plus: Experience with multi-cloud or on-prem hybrid customer environments (AWS, GCP, Azure, sovereign clouds).
  • Strong plus: Open-source contributions, particularly in SRE, observability, or security tooling.

Преимущества

  • Comprehensive benefits package designed to support your well-being, growth, and work-life balance.
  • Benefits vary by country and may include healthcare coverage, parental leave, retirement plans, relocation support, wellness programs, meal and transportation allowances, and other location-specific perks.
🇫🇷 ФранцияArtificial IntelligenceСтартапmistral.ai/
Зарплата не указана