JobHabor

Senior Site Reliability Engineer

Orion Health
Location
Global
Workplace
Remote
Employment
Contract
Salary
Apply on the employer’s site

Posted 1mo ago

The employer’s full description could not be read from their board. This is a summary of the posting — follow the apply link for the original.

Responsibilities

  • Automate infrastructure and software delivery processes
  • Monitor system availability, latency, reliability, and performance
  • Manage emergency response and incident resolution
  • Perform capacity planning and performance trend analysis
  • Execute change management tasks and validation
  • Tune application stacks to improve stability and uptime
  • Automate repetitive tasks like patching and scaling
  • Conduct root cause analysis for outages and performance issues
  • Develop and test disaster recovery plans
  • Manage and maintain underlying server and network infrastructure
  • Participate in on-call rotation
  • Integrate updates for products into managed solutions
  • Document procedures to facilitate team knowledge transfer
  • Coordinate with development teams on product configuration requirements

Requirements

  • 4–6 years in a Site Reliability Engineering or equivalent role
  • 5 years in systems/application support or development
  • Strong understanding of Windows and Linux operating systems
  • Proficiency in PowerShell, Python, and Bash scripting
  • Experience with infrastructure automation tools like Puppet and Ansible
  • Solid understanding of TCP/IP, DNS, DHCP, VLANs, VPNs, and firewalls
  • Experience with AWS cloud-based production systems
  • Experience with CI/CD pipelines and deployment automation
  • Bachelor’s Degree in a technical discipline or equivalent experience
  • Experience with infrastructure as code and orchestration tools
  • Proficiency in Active Directory and GPO management
  • Technical certification in System Administration or Cloud Engineering

Preferred

  • Experience with Kubernetes, CloudFormation, and Terraform
  • Exposure to on-prem to AWS cloud migration projects
  • Experience with Red Hat OS upgrades
  • Working knowledge of Splunk monitoring tools
  • Formal training in *nix scripting
  • Knowledge of SQL, Oracle databases, and Big Data technologies
  • Understanding of HIPAA or HITRUST compliance

Skills

  • AWS
  • Windows
  • Linux
  • Active Directory
  • PowerShell
  • Python
  • Bash
  • Puppet
  • Ansible
  • Kubernetes
  • CloudFormation
  • Terraform
  • Splunk
  • Oracle
  • SQL

Similar roles