DevOps / Site Reliability Engineer

ACCORD INNOVATIONS PTE. LTD.Singaporemycareersfuturepublicada em 22/09/2026
Obrigatório:PythonAWSGoogle CloudKubernetesCloudDevOpsCI/CDSecurity

About the Role

We are seeking an experienced DevOps / Site Reliability Engineer to support a client engagement in Singapore. The successful candidate will be responsible for designing, automating and operating scalable cloud infrastructure and highly available containerized platforms.

The role requires strong hands-on experience across cloud platforms, Kubernetes, Terraform and Python, together with strong communication and collaboration skills.

Key Responsibilities

  • Design, implement and maintain scalable and highly available cloud infrastructure.
  • Develop and manage cloud infrastructure across AWS, Google Cloud Platform (GCP) and other cloud environments.
  • Automate infrastructure provisioning, configuration and deployment using Terraform and Infrastructure as Code (IaC) practices.
  • Design, deploy and operate production workloads on Kubernetes .
  • Develop Python scripts and automation tools to improve operational efficiency and reliability.
  • Implement DevOps and Site Reliability Engineering practices across development and operations.
  • Support CI/CD automation, deployment pipelines and infrastructure automation.
  • Monitor system performance, availability and reliability and support troubleshooting of production issues.
  • Participate in incident resolution, root-cause analysis and preventive actions.
  • Work closely with application development, infrastructure and security teams.
  • Communicate technical issues, solutions and operational requirements effectively with technical and business stakeholders.

Key Requirements

  • 5–12 years of relevant experience in DevOps, Site Reliability Engineering (SRE), Cloud Engineering or a related role.
  • Strong hands-on experience with one or more major cloud platforms, particularly AWS or GCP .
  • Experience with Alibaba Cloud will be an advantage.
  • Strong Python scripting and automation skills.
  • Strong hands-on experience with Terraform and Infrastructure as Code.
  • Strong practical experience with Kubernetes in production environments.
  • Good understanding of cloud infrastructure, automation, deployment and operational practices.
  • Experience working with CI/CD pipelines and DevOps methodologies.
  • Strong troubleshooting and problem-solving capabilities.
  • Good written and verbal communication skills.
  • Ability to work collaboratively with engineering, infrastructure and application teams.

Mandatory Skills

  • DevOps / SRE
  • AWS / GCP
  • Python
  • Terraform
  • Kubernetes
  • Cloud Infrastructure
  • Infrastructure as Code (IaC)
  • CI/CD
  • Automation
  • Production Operations
  • Incident Troubleshooting

Preferred Skills

  • Alibaba Cloud
  • Advanced Kubernetes administration
  • Cloud-native architecture
  • Containerization
  • Monitoring and observability
  • CI/CD automation
  • Infrastructure security

Top 3 Skills

  • Kubernetes
  • AWS / GCP Cloud
  • Terraform & Python