Jobs / IMG

DevOps Engineer, Studios

IMG · London, ENG, United Kingdom
London, ENG, United KingdomHybrid
Remuneration
Not specified
Location
London, ENG, United Kingdom
Visa sponsorship
Not specified

Job summary

IMG is seeking an experienced DevOps Engineer to design, implement, and maintain scalable and secure infrastructure and delivery pipelines for digital and broadcast platforms. The role involves enabling continuous delivery, improving system reliability, and supporting high-profile clients and live event services.

Qualifications

  • Proven experience in a DevOps, Site Reliability Engineering, or similar role.
  • Strong knowledge of all Operating systems, such as Linux distributions.
  • Strong knowledge of cloud platforms such as AWS, Azure, or Google Cloud.
  • Experience with containerisation and orchestration tools such as Docker and Kubernetes.
  • Hands-on experience with CI/CD tools such as Jenkins, GitHub Actions, GitLab CI, or Azure DevOps/Bitbucket.
  • Strong scripting skills in Python, Bash, Ansible, or similar.
  • Experience with Infrastructure as Code tools such as Terraform or CloudFormation.
  • Familiarity with monitoring, logging, and alerting tools.
  • Understanding of networking, security, and system architecture principles.
  • Experience working in high-availability or live production environments is highly desirable.
  • Knowledge of version control systems such as Git.
  • A technical and/or engineering background with a strong interest in automation and cloud technologies.
  • Strong problem-solving skills and the ability to work under pressure in critical situations.
  • Excellent communication and collaboration skills.
  • Ability to manage priorities across multiple projects and environments.
  • Proactive mindset with a focus on continuous improvement.

Responsibilities

  • Design, build, and maintain scalable cloud infrastructure to support production and development environments.
  • Implement and manage CI/CD pipelines to enable efficient, reliable software delivery across multiple teams.
  • Automate infrastructure provisioning using Infrastructure as Code tools such as Terraform, CloudFormation, or similar.
  • Monitor system performance, availability, and reliability using observability tools such as Prometheus, Grafana, and ELK stack.
  • Ensure high availability and disaster recovery strategies are in place and tested regularly.
  • Collaborate closely with development, QA, and operations teams to streamline deployment processes and improve system performance.
  • Act as a point of escalation for production incidents, providing rapid troubleshooting and resolution.
  • Implement and enforce security best practices across infrastructure, pipelines, and applications.
  • Optimise system performance and cost-efficiency across cloud environments.
  • Document infrastructure, processes, and operational procedures clearly and consistently.
  • Continuously identify opportunities for automation and operational improvements.
  • Support live and critical event operations where system uptime and responsiveness are essential.
  • Communicate effectively with stakeholders, ensuring visibility of issues, risks, and improvements.

Skills

AnsibleAWSAzureAzure DevOpsBashBitbucketCloudFormationDockerGCPGitGitHubGitHub ActionsGitLabGitLab CIGrafanaJenkinsKubernetesLinuxPrometheusPythonTerraform

Relocation

No