JobHabor

Staff Infrastructure Engineer — Observability

SentinelOne

Location
US
Workplace
Remote
Employment
Full Time
Salary
USD 132,000–215,000/yr
Apply on the employer’s site

Posted 2mo ago

The employer’s full description could not be read from their board. This is a summary of the posting — follow the apply link for the original.

Responsibilities

  • Architect and implement scalable telemetry platforms
  • Act as SME and administrator for observability stack
  • Partner with engineering teams to define platform requirements
  • Take ownership of critical features from design to deployment
  • Drive operational efficiency for observability services
  • Build automation and self-service tooling
  • Drive deployment, maintenance, and compliance in high-security environments
  • Cultivate platform transparency and reliability using IaC
  • Elevate engineering quality through mentoring and reviews
  • Lead incident resolution and root-cause analysis
  • Participate in on-call rotations

Requirements

  • 8+ years experience in Infrastructure Engineering, SRE, or related field
  • 8+ years experience architecting, scaling, and managing observability stacks
  • Experience with Prometheus, Grafana, Thanos (or Mimir/Cortex), and OpenTelemetry (OTEL)
  • Experience designing cloud-native infrastructure in AWS or GCP
  • Experience managing production Kubernetes environments (EKS, GKE)
  • Advanced proficiency with Terraform and Ansible
  • Experience maintaining and optimizing high-throughput, large-scale distributed systems
  • Demonstrated ability to lead complex technical designs
  • Demonstrated ability to mentor other engineers
  • Demonstrated ability to collaborate cross-functionally
  • US Citizenship
  • Ability to work in a government-regulated environment

Preferred

  • 8+ years production-level programming experience in GoLang
  • Strong willingness to adopt GoLang
  • Experience working with high-security compliance frameworks (FedRAMP or sovereign cloud)
  • Familiarity with operational challenges of on-premises, hybrid, or air-gapped Kubernetes deployments
  • Experience designing advanced CI/CD pipelines (e.g., GitHub Actions)
  • Experience implementing sophisticated deployment strategies (canary, blue-green, rolling updates)

Skills

  • Grafana
  • Prometheus
  • Thanos
  • Mimir
  • Cortex
  • OpenTelemetry
  • AWS
  • GCP
  • Kubernetes
  • EKS
  • GKE
  • Terraform
  • Ansible
  • GoLang
  • Python
  • Java
  • GitHub Actions

Similar roles