JobHabor

Assoc. Manager, Site Reliability Engineering

Yum
Location
Ho Chi Minh, Dong Nam Bo, Viet Nam
Workplace
Hybrid
Employment
Full Time
Salary
Apply on the employer’s site

Posted 1mo ago

  • 5+ years of experience in site reliability engineering, DevOps, infrastructure, or production operations roles
  • 1+ years of people management experience, or 2+ years as a senior technical lead with demonstrated coaching and delivery ownership
  • Hands-on credibility across incident response, observability, and automation, with the technical depth to guide Level 6-7 engineers
  • Experience operating in shift-based, on-call, or follow-the-sun coverage models
  • Working knowledge of at least one major cloud provider (AWS preferred) and modern observability tooling (e.g., Datadog, Prometheus, Grafana)
  • Proficiency in at least one scripting or programming language sufficient to review and guide automation work
  • Understanding of SLI/SLO frameworks and reliability engineering fundamentals
  • Strong written and verbal English communication skills for cross-region collaboration with US and India teams
  • 5+ years of experience in site reliability engineering, DevOps, infrastructure, or production operations roles
  • 1+ years of people management experience, or 2+ years as a senior technical lead with demonstrated coaching and delivery ownership
  • Hands-on credibility across incident response, observability, and automation, with the technical depth to guide Level 6-7 engineers
  • Experience operating in shift-based, on-call, or follow-the-sun coverage models
  • Working knowledge of at least one major cloud provider (AWS preferred) and modern observability tooling (e.g., Datadog, Prometheus, Grafana)
  • Proficiency in at least one scripting or programming language sufficient to review and guide automation work
  • Understanding of SLI/SLO frameworks and reliability engineering fundamentals
  • Strong written and verbal English communication skills for cross-region collaboration with US and India teams
  • Experience building or standing up a new team, site, or shift operation
  • Experience managing engineers across the early-to-mid career range with a track record of promotions or level progression
  • Kubernetes, container orchestration, and infrastructure as code experience (e.g., Terraform)
  • Familiarity with AI-assisted operations tooling and automation-first reliability approaches, including auto-healing and auto-remediation patterns
  • Exposure to platform engineering and internal developer platform concepts: self-service tooling, developer portals (e.g., Port, Backstage), GitOps
  • Experience in multi-region or globally distributed team models
  • Relevant certifications (AWS, CKA, or similar)

Skills

  • AWS
  • Datadog
  • Prometheus
  • Grafana
  • Kubernetes
  • Terraform

More jobs at Yum

All 54

Similar roles