Data Center Production Operations Engineer

Odense, Southern Denmark
Posted 22 hours ago
Engineering

About the role

Job summary

This role involves ensuring the reliability, efficiency, and scalability of data center infrastructure through operational management of server fleets and production systems. The engineer will work closely with various teams to maintain peak performance and support the products and services that connect users globally.

Responsibilities

  • Oversee the operational health of large-scale server fleets and production infrastructure in data center environments.
  • Diagnose and resolve hardware and systems failures, collaborating with engineering teams for root cause analysis and corrective actions.
  • Implement and enhance server deployment, decommissioning, and lifecycle management processes to meet capacity and reliability objectives.
  • Create and maintain operational runbooks, escalation procedures, and documentation to standardize workflows.
  • Work with hardware engineering and capacity planning teams to identify issues and suggest infrastructure improvements.
  • Analyze operational metrics and failure trends to enhance fleet reliability and reduce resolution times.
  • Support the qualification and rollout of new server hardware, ensuring operational readiness and identifying risks.
  • Collaborate with network engineering, facilities, and software infrastructure teams to address complex production incidents.
  • Identify automation opportunities for repetitive tasks and contribute to tooling enhancements for operational efficiency.
  • Provide technical guidance on production operations best practices and troubleshooting methodologies.

Qualifications

Preferred Qualifications

  • Minimum of 2 years of experience in data center operations, production operations, or systems administration in large-scale environments.
  • Proficient in troubleshooting server hardware components such as CPUs, memory, storage, and networking in production settings.
  • Experience in developing or refining operational processes, runbooks, or standard operating procedures for infrastructure teams.
  • Ability to analyze operational data or failure metrics to identify trends and drive improvements.
  • Experience collaborating with cross-functional engineering teams to resolve production incidents.
  • Background in capacity planning, hardware lifecycle management, or server deployment for hyperscale data centers.
  • Experience with hardware qualification or new server platform deployment in production environments.
  • Familiarity with fleet management tools, asset tracking systems, or infrastructure monitoring platforms.
  • Knowledge of scripting languages like Python or Bash for automating workflows and data analysis.
Full Access

Ready to apply for this role?

Full Access gives you the company name, full job description, and a direct link to apply. On the 1- and 3-month plans, CV Tailor rewrites your CV for this exact role.

Share this job