DevOps / Site Reliability Engineer
Задължително:PythonAWSGoogle CloudKubernetesCloudDevOpsCI/CDSecurity
About the Role
We are seeking an experienced DevOps / Site Reliability Engineer to support a client engagement in Singapore. The successful candidate will be responsible for designing, automating and operating scalable cloud infrastructure and highly available containerized platforms.
The role requires strong hands-on experience across cloud platforms, Kubernetes, Terraform and Python, together with strong communication and collaboration skills.
Key Responsibilities
- Design, implement and maintain scalable and highly available cloud infrastructure.
- Develop and manage cloud infrastructure across AWS, Google Cloud Platform (GCP) and other cloud environments.
- Automate infrastructure provisioning, configuration and deployment using Terraform and Infrastructure as Code (IaC) practices.
- Design, deploy and operate production workloads on Kubernetes .
- Develop Python scripts and automation tools to improve operational efficiency and reliability.
- Implement DevOps and Site Reliability Engineering practices across development and operations.
- Support CI/CD automation, deployment pipelines and infrastructure automation.
- Monitor system performance, availability and reliability and support troubleshooting of production issues.
- Participate in incident resolution, root-cause analysis and preventive actions.
- Work closely with application development, infrastructure and security teams.
- Communicate technical issues, solutions and operational requirements effectively with technical and business stakeholders.
Key Requirements
- 5–12 years of relevant experience in DevOps, Site Reliability Engineering (SRE), Cloud Engineering or a related role.
- Strong hands-on experience with one or more major cloud platforms, particularly AWS or GCP .
- Experience with Alibaba Cloud will be an advantage.
- Strong Python scripting and automation skills.
- Strong hands-on experience with Terraform and Infrastructure as Code.
- Strong practical experience with Kubernetes in production environments.
- Good understanding of cloud infrastructure, automation, deployment and operational practices.
- Experience working with CI/CD pipelines and DevOps methodologies.
- Strong troubleshooting and problem-solving capabilities.
- Good written and verbal communication skills.
- Ability to work collaboratively with engineering, infrastructure and application teams.
Mandatory Skills
- DevOps / SRE
- AWS / GCP
- Python
- Terraform
- Kubernetes
- Cloud Infrastructure
- Infrastructure as Code (IaC)
- CI/CD
- Automation
- Production Operations
- Incident Troubleshooting
Preferred Skills
- Alibaba Cloud
- Advanced Kubernetes administration
- Cloud-native architecture
- Containerization
- Monitoring and observability
- CI/CD automation
- Infrastructure security
Top 3 Skills
- Kubernetes
- AWS / GCP Cloud
- Terraform & Python