JobHabor

DevOps Engineer

Qima
Location
Shenzhen, Guangdong Province, China
Workplace
Employment
Full Time
Salary
Apply on the employer’s site

Posted 2mo ago

  • Support DevOps / Infrastructure Operations execution for production systems,

including release support, deployment automation, monitoring, troubleshooting, and

recurring operational tasks.

  • Maintain and improve infrastructure-as-code practices using Terraform, following

the team’s existing standards, review process, and git branching model.

  • Develop and maintain automation scripts and tools, mainly using Shell and Python,

to improve deployment, update, monitoring, and operational efficiency.

  • Support AWS / Azure cloud operations, including infrastructure changes, access /

configuration updates, reliability improvements, and cost / performance optimization

follow-up.

  • Operate and troubleshoot Docker and Kubernetes-based production environments,

including deployment issues, workload health, scaling, rollback, and incident

support.

  • Participate in production incident diagnosis, root cause analysis, post-incident

follow-up, and preventive improvement actions.

  • Support observability practices using tools such as Prometheus, Grafana, and ELK,

including monitoring coverage, alert quality, dashboard improvement, and log

analysis.

  • Prepare and maintain runbooks, SOPs, technical documentation, and handover

materials to ensure operational continuity.

  • Work closely with the existing DevOps / Infrastructure owner for priority

alignment, architecture review, production change planning, and stakeholder

communication.

Technical requirements

  • Bachelor’s degree in Computer Science, Engineering, or a related technical field,

or equivalent practical experience.

  • 5+ years of experience in infrastructure, cloud operations, DevOps, distributed

systems, or cloud migration.

  • Strong Linux operation and troubleshooting skills, with solid understanding of

networking fundamentals.

  • Hands-on experience with Shell scripting and Python for automation, with clean and

maintainable coding practices.

  • Practical experience with AWS and/or Azure core services and cloud operations.
  • Hands-on experience with Terraform in production environments; multi-cloud experience is a plus.
  • Experience with Docker and Kubernetes in production environments, including

troubleshooting, deployment, scaling, rollback, and operational support.

  • Experience with CI/CD pipelines and application release automation.
  • Familiarity with observability tools such as Prometheus, Grafana, ELK, or similar monitoring / logging platforms.
  • Good documentation habits and willingness to work with runbooks, SOPs, tickets,

and change records.

  • Familiarity with Ansible, disaster recovery, resilience testing, or incident

response automation

Skills

  • Terraform
  • Git
  • Shell
  • Python
  • AWS
  • Azure
  • Docker
  • Kubernetes
  • Prometheus
  • Grafana
  • ELK Stack
  • SOPS
  • Linux
  • Ansible

More jobs at Qima

All 41

Similar roles