JobHabor

Senior Site Reliability Engineer

PandaDoc
Location
Location not stated
Workplace
Remote
Employment
Full Time
Salary
PLN 29,000–34,500/mo
Apply on the employer’s site

Posted 7mo ago

The employer’s full description could not be read from their board. This is a summary of the posting — follow the apply link for the original.

Responsibilities

  • Own and influence the incident management process end-to-end
  • Maintain and evolve on-prem observability stack
  • Keep production applications running smoothly by participating in the on-call rotation
  • Develop automations and tools to support platform reliability
  • Contribute to production services with performance and resiliency in mind
  • Collaborate with product engineers to foster SRE principles within the R&D organization
  • Mentor the SRE team or product engineers

Requirements

  • Solid programming experience in Python (Django and AsyncIO) and/or Java (Spring Boot)
  • Experience in maintaining an observability tools suite (Loki, Grafana, Tempo, Mimir)
  • Experience in development and maintenance of Python services in production
  • Strong experience with AWS and Kubernetes
  • Solid proficiency in working with relational databases (PostgreSQL)
  • Solid proficiency in working with messaging systems (RabbitMQ, NATS, Kafka)
  • Experienced on-call SRE engineer
  • Hands-on troubleshooting of distributed systems in production environments
  • Act like an owner and strive to do work you're proud of
  • Enjoy communication and knowledge sharing on all-things reliability
  • Proficiency in English, both written and spoken

Skills

  • Python
  • Django
  • AsyncIO
  • Java
  • Spring Boot
  • Loki
  • Grafana
  • Tempo
  • Mimir
  • AWS
  • Kubernetes
  • PostgreSQL
  • RabbitMQ
  • NATS
  • Kafka

More jobs at PandaDoc

All 7

Similar roles