Staff Infrastructure Engineer — Observability
SentinelOne
- Location
- US
- Workplace
- Remote
- Employment
- Full Time
- Salary
- USD 132,000–215,000/yr
Posted 2mo ago
The employer’s full description could not be read from their board. This is a summary of the posting — follow the apply link for the original.
Responsibilities
- Architect and implement scalable telemetry platforms
- Act as SME and administrator for observability stack
- Partner with engineering teams to define platform requirements
- Take ownership of critical features from design to deployment
- Drive operational efficiency for observability services
- Build automation and self-service tooling
- Drive deployment, maintenance, and compliance in high-security environments
- Cultivate platform transparency and reliability using IaC
- Elevate engineering quality through mentoring and reviews
- Lead incident resolution and root-cause analysis
- Participate in on-call rotations
Requirements
- 8+ years experience in Infrastructure Engineering, SRE, or related field
- 8+ years experience architecting, scaling, and managing observability stacks
- Experience with Prometheus, Grafana, Thanos (or Mimir/Cortex), and OpenTelemetry (OTEL)
- Experience designing cloud-native infrastructure in AWS or GCP
- Experience managing production Kubernetes environments (EKS, GKE)
- Advanced proficiency with Terraform and Ansible
- Experience maintaining and optimizing high-throughput, large-scale distributed systems
- Demonstrated ability to lead complex technical designs
- Demonstrated ability to mentor other engineers
- Demonstrated ability to collaborate cross-functionally
- US Citizenship
- Ability to work in a government-regulated environment
Preferred
- 8+ years production-level programming experience in GoLang
- Strong willingness to adopt GoLang
- Experience working with high-security compliance frameworks (FedRAMP or sovereign cloud)
- Familiarity with operational challenges of on-premises, hybrid, or air-gapped Kubernetes deployments
- Experience designing advanced CI/CD pipelines (e.g., GitHub Actions)
- Experience implementing sophisticated deployment strategies (canary, blue-green, rolling updates)
Skills
- Grafana
- Prometheus
- Thanos
- Mimir
- Cortex
- OpenTelemetry
- AWS
- GCP
- Kubernetes
- EKS
- GKE
- Terraform
- Ansible
- GoLang
- Python
- Java
- GitHub Actions
Similar roles
Site Reliability Engineer (FedRAMP / Security)
Coralogix · New York, NY, United States · USD 170,000–350,000/yr · today
Senior Business Engineer - Ads
Reddit · Remote · USD 180,200–252,300/yr · today
Machine Learning Engineer
Reddit · US · USD 185,800–303,400/yr · today
Backend Software Engineer
ClearlyRated · Portland, OR, United States · USD 90,000–120,000/yr · today
Firmware Engineer
Anduril Industries · Costa Mesa, California, United States · USD 166,000–220,000/yr · today
Embedded Firmware Engineer
Anduril Industries · Costa Mesa, California, United States · USD 166,000–220,000/yr · today