JobHabor

Senior Site Reliability Engineer

Synapse Health

Location
US
Workplace
Remote
Employment
Full Time
Salary
USD 133,600–183,700/yr
Apply on the employer’s site

Posted 1mo ago

The employer’s full description could not be read from their board. This is a summary of the posting — follow the apply link for the original.

Responsibilities

  • Contribute to the migration from legacy Azure services to containerized microservices
  • Design, build, and scale Kubernetes-based infrastructure and supporting tooling
  • Partner with engineering teams to ensure systems are designed for reliability and scalability
  • Drive infrastructure standardization to reduce silos and enable team ownership
  • Monitor cloud usage and spending to implement cost optimization strategies
  • Design and manage secure networking including endpoints, VPNs, and connectivity
  • Maintain highly available systems in a cloud-native environment
  • Implement observability practices including monitoring, alerting, and logging
  • Define and manage SLIs, SLOs, and SLAs
  • Lead incident response efforts and drive root cause analysis
  • Build and optimize CI/CD pipelines for repeatable deployments
  • Champion Infrastructure-as-Code practices using Terraform
  • Leverage Python or Bash scripting to automate workflows and reduce toil
  • Drive capacity planning and performance tuning
  • Identify system risks and scalability bottlenecks
  • Contribute to infrastructure strategy and platform evolution
  • Document systems, processes, and best practices
  • Contribute to cross-training and mentorship

Requirements

  • 5+ years of experience in Site Reliability Engineering, DevOps, or Infrastructure Engineering
  • Hands-on experience in cloud environments including Azure, AWS, or GCP
  • Strong experience with Kubernetes in production environments
  • Experience deploying and managing applications on Kubernetes using Helm
  • Proficiency with Infrastructure-as-Code tools such as Terraform
  • Strong scripting skills in Python or Bash
  • Experience with observability, monitoring, and incident response in production
  • Experience building or supporting CI/CD pipelines using GitHub Actions or GitLab CI/CD
  • Solid understanding of networking fundamentals and system design
  • Familiarity with Azure Entra ID, app registrations, and federated identity
  • Understanding of PHI handling, access control, and audit controls in healthcare

Preferred

  • Experience with platform or architectural transformations
  • Familiarity with .NET / C# application environments
  • Deeper networking expertise including firewalls and gateways
  • Experience with CI/CD migrations

Skills

  • Azure
  • AWS
  • GCP
  • Kubernetes
  • Helm
  • Terraform
  • Python
  • Bash
  • Datadog
  • Prometheus
  • Grafana
  • GitHub Actions
  • GitLab CI/CD
  • Azure Entra ID
  • .NET
  • C#

Similar roles