Jobs / RBC

Site Reliability Engineer (SRE), Cloud Operations

RBC · Toronto, ON, Canada
Toronto, ON, CanadaExp: 5+ yrsHybrid
Remuneration
Not specified
Location
Toronto, ON, Canada
Visa sponsorship
Not specified

Job summary

Join the Platform Engineering & AI Operations team at RBC as a Site Reliability Engineer, enhancing the bank's cloud operations through automation and intelligent infrastructure practices.

Qualifications

  • 5+ years of hands-on experience in Site Reliability Engineering, DevOps, or infrastructure operations.
  • Strong knowledge of Kubernetes/OpenShift administration.
  • Experience with Ansible and Terraform.
  • Proficiency in Python scripting.
  • Experience with monitoring and observability stacks.
  • Experience with incident management processes.
  • Familiarity with capacity planning and performance analysis.
  • Understanding of security and compliance fundamentals.

Responsibilities

  • Support scalable, secure, and available architectures across cloud platforms.
  • Write code and scripts to automate infrastructure workflows.
  • Extend self-healing automation capabilities using Ansible.
  • Participate in and lead design reviews for platform features.
  • Collaborate with platform teams for technical feedback.
  • Drive automation, CI/CD, and Infrastructure as Code practices.
  • Minimize risk of reliability failures.
  • Participate in on-call rotation for platform support.

Skills

AnsibleAWSAzureDynatraceGCPGrafanaKafkaKubernetesLinuxOpenShiftPagerDutyPrometheusPythonRHELServiceNowTerraform

Relocation

No