Engineering Manager, Site Reliability Engineering (Remote)

Remote-EMEA
Posted 13 hours ago
Engineering

About the role

Job summary

This role involves leading a Site Reliability Engineering (SRE) team to ensure operational excellence and product reliability. The position is a blend of individual contribution and leadership, focusing on team development and technical oversight.

Qualifications

  • Proven experience leading SRE, infrastructure, or platform engineering teams.
  • Strong coaching skills for both technical and soft skills development.
  • Ability to manage underperformance with clarity and empathy.
  • Experience in hiring and assessing engineering talent.
  • Proficient in understanding team dynamics and conflict resolution.

Responsibilities

  • Oversee the full career lifecycle of team members, including onboarding, feedback, and performance assessments.
  • Act as the spokesperson for the SRE team within the engineering department.
  • Define and prioritize SRE goals and manage the support rotation and on-call model.
  • Maintain and enhance the reliability practice, including SLOs and incident response.

Skills

  • Hands-on experience with Kubernetes, AWS, and cloud infrastructure.
  • Familiarity with AI infrastructure and observability practices.
  • Proficient in Infrastructure as Code (Terraform) and CI/CD systems.
  • Experience with Docker, shell scripting, and database operations, particularly PostgreSQL.

Education

  • Relevant technical degree or equivalent experience in a related field.

Tools

  • Kubernetes, AWS, PostgreSQL, Terraform, GitLab CI, GitHub Actions, Docker.
Full Access

Ready to apply for this role?

Full Access gives you the company name, full job description, and a direct link to apply. On the 1- and 3-month plans, CV Tailor rewrites your CV for this exact role.

Share this job

Full Access includes

  • Company name & profile
  • Full job description
  • Direct apply link
  • Unlimited job alerts
  • CV tailored to this job (1- & 3-month plans)