Jobs / TikTok USDS JV
Senior Site Reliability Engineer, AI Infrastructure
TikTok USDS JV · Seattle, WA, United States
Seattle, WA, United StatesExp: 3+ yrs177,688-341,734 USD/yearlyHybrid
Remuneration
177,688-341,734 USD/yearly
Location
Seattle, WA, United States
Visa sponsorship
Not specified
Job summary
The Senior Site Reliability Engineer for AI Infrastructure at TikTok USDS JV will work with engineering and product teams to build and maintain globally distributed systems for TikTok's Recommendation and Search engines. The role involves driving observability and automation, ensuring high availability, and leading infrastructure migrations in a secure environment.
Qualifications
- 3+ years of hands-on SRE, DevOps, or Systems Engineering experience
- Strong foundation in Linux operating systems
- Experience programming in at least one standard language
- Demonstrated experience in designing, troubleshooting, and maintaining complex distributed systems
- Familiarity with modern CI/CD practices and automated deployment pipelines
- High resilience and a strong 'system sense'
Responsibilities
- Partner with cross-functional teams throughout the entire service lifecycle
- Build tools, platforms, and automation to improve service reliability
- Ensure high availability and performance of large-scale, multi-region systems
- Lead complex infrastructure migrations and system architecture upgrades
- Drive robust Capacity Management
- Drive sustainable incident management
Skills
AWSAzureBashC++GCPGoGrafanaJavaKubernetesLinuxPrometheusPythonTerraform
Degrees
Bachelor's degree or above in Computer ScienceSoftware EngineeringRelated technical field
Relocation
No