Jobs / Cerebras Systems
Staff Cloud Infrastructure Engineer
Cerebras Systems · Toronto, ON, Canada
Toronto, ON, CanadaFull timeExp: 7+ yrsHybrid
Remuneration
Not specified
Location
Toronto, ON, Canada
Visa sponsorship
Not specified
Job summary
The Staff Cloud Infrastructure Engineer is tasked with designing and architecting secure, scalable cloud infrastructure and identity platforms, primarily utilizing AWS. Responsibilities include implementing IAM solutions, developing automation with Terraform and programming languages like Python and Go, and ensuring security controls for AI systems while collaborating across teams to enhance infrastructure reliability and developer experience.
Qualifications
- Experience designing and owning complex production infrastructure or platform systems.
- Strong knowledge of authentication, authorization, and identity lifecycle management.
- Experience with identity and access controls in AWS environments.
- Experience with container technologies and Kubernetes.
- Strong automation and coding skills in Python or Go, and Terraform.
- Security-first mindset with experience in Zero Trust and compliance frameworks.
- Experience supporting and improving production services.
- Ability to influence technical direction and mentor engineers.
- Excellent collaboration and communication skills.
Responsibilities
- Drive the design and architecture of secure, scalable cloud infrastructure and identity platforms in AWS.
- Design and implement IAM, IGA, authentication, authorization, SSO, MFA, and identity lifecycle management solutions.
- Develop automation and infrastructure-as-code solutions using Terraform, Python, and Go.
- Design and implement security controls for AI-powered systems.
- Collaborate with Security, Engineering, and Infrastructure teams to drive secure-by-design solutions.
- Lead technical direction and initiatives to improve infrastructure scalability, reliability, and security.
- Write high-quality code, participate in code reviews, and mentor engineers.
- Support production systems and drive operational excellence.
- Participate in on-call and incident response.
Skills
AWSAzureGCPGoIAMKubernetesOktaPythonTerraform
Relocation
No