Jobs / Megaport

Senior Site Reliability Engineer

Megaport · En remoto, Spain
En remoto, SpainExp: 5+ yrsRemote
Remuneration
Not specified
Location
En remoto, Spain
Visa sponsorship
Not specified

Job summary

As a Senior Platform Engineer, you will champion DevOps and SRE culture while ensuring systems are secure, maintainable, and available.

Qualifications

  • 5+ years administering Linux systems in production environments
  • Collaborative SRE mindset with familiarity around SLIs/SLOs/SLAs
  • Focus on automation and reducing toil
  • Track record of writing effective runbooks
  • Strong Kubernetes fundamentals
  • Cloud infrastructure experience, AWS preferred
  • Strong tool development skills in Bash, Python, or Go
  • Experience with Infrastructure-as-code tooling, Terraform preferred
  • CI/CD and version control experience, GitHub preferred
  • Database experience with Postgres, Cassandra, or ClickHouse
  • Experience operating a production observability stack
  • Strong troubleshooting instincts and incident response ownership
  • History of continual professional development
  • Self-directed style suited to a globally distributed team

Responsibilities

  • Improve production reliability and system resilience
  • Champion high standards of work and industry best practices
  • Communicate with teams and stakeholders
  • Encourage fresh ideas
  • Solve complex technical problems
  • Work across various technologies
  • Participate in on-call rotation and incident response
  • Write code and improve solutions
  • Support team success

Skills

AWSBashCassandraClickHouseGitHubGoKubernetesLinuxPostgreSQLPythonTerraform

Relocation

No