verified listingSign up to apply with your verified profile — no re-entering experience or references.
source · wttj·req · jb_dc10d8b210·listed 11h ago

Site Reliability Engineer

Mistral Ai·York, England, United Kingdom·Hybrid·Full-time
Sourced listing · wttjNo salary disclosed
Posted
15 September 2026
via wttj
Type
Full-time
Arrangement
Hybrid
United Kingdom
Deadline
15 October 2026
closes in 30d
compensation · not disclosed
Salary not shared
Sign up to see our estimate based on role, location, and seniority.
source · estimate pending

Summary

the pitch

Join Mistral, a leading provider of full-stack AI solutions. As a Site Reliability Engineer, you will play a crucial role in shaping the reliability, scalability, and performance of our platform and customer-facing applications. You will work closely with software engineers and research teams to ensure our systems meet and exceed customer expectations. Your responsibilities will include designing and maintaining scalable infrastructures, troubleshooting production issues, implementing monitoring and incident response systems, and driving continuous improvement in infrastructure automation.

Role

posted by company

Join Mistral, a leading provider of full-stack AI solutions. As a Site Reliability Engineer, you will play a crucial role in shaping the reliability, scalability, and performance of our platform and customer-facing applications. You will work closely with software engineers and research teams to ensure our systems meet and exceed customer expectations. Your responsibilities will include designing and maintaining scalable infrastructures, troubleshooting production issues, implementing monitoring and incident response systems, and driving continuous improvement in infrastructure automation.

Key responsibilities

  • Design, build, and maintain scalable, highly available, and fault-tolerant infrastructures to support web services and ML workloads.
  • Implement and improve monitoring, alerting, and incident response systems to ensure optimal system performance and minimize downtime.
  • Drive continuous improvement in infrastructure automation, deployment, and orchestration using tools like Kubernetes, Flux, Terraform.