Platform Engineer & Cloud Ops Engineer
Accountabilities: Lead the architecture, deployment, and evolution of complex multi-cloud environments across GCP and AWS, covering networking, compute, storage, IAM, load balancing, and multi-environment infrastructure design.
Define scalable platform standards and reusable engineering patterns, including Terraform modules, Kubernetes cluster blueprints, golden paths, and consistent deployment practices across development, staging, and production environments.
Provision, configure, maintain, and automate cloud infrastructure using Infrastructure as Code, with Terraform as the primary technology and CloudFormation or ARM where appropriate.
Design, operate, and support production-grade Kubernetes platforms, preferably GKE, including upgrades, autoscaling, node pools, namespaces, RBAC policies, and application deployment standards using Helm and Kustomize.
Support and manage service mesh capabilities such as Istio for traffic management, mutual TLS, observability, and security, while promoting safe and consistent application delivery.
Build and maintain GitLab CI/CD pipelines for infrastructure and application delivery, automating operational processes and reducing manual effort.
Monitor platform health and performance using Datadog metrics, logs, traces, alerts, and SLOs, while troubleshooting complex issues, supporting incident response, and contributing to post-incident reviews.
Own platform reliability, scalability, capacity planning, and disaster recovery design, maintaining runbooks and dashboards and participating in or leading on-call operations.
Drive cloud cost optimization and FinOps practices through right-sizing, governance, and ongoing evaluation of cloud consumption and infrastructure efficiency.
Mentor junior and mid-level engineers, provide technical direction, review designs, and communicate architectural decisions clearly to technical and business stakeholders.
Collaborate effectively with global, cross-functional, and distributed teams to deliver platform initiatives and continuously improve operational practices.
Requirements:
3+ years of experience in cloud infrastructure, platform engineering, or DevOps for Platform Engineer-level roles, or 8+ years of experience with leadership or architect-level responsibilities for Senior Platform Engineer-level roles.
Deep, hands-on expertise with GCP and/or AWS, including services such as Compute Engine, GKE, VPC, Storage, IAM, Load Balancing, EC2, EKS, and S3.
Strong production Kubernetes experience, preferably with GKE architecture, cluster upgrades, autoscaling, node pools, RBAC, Helm, and Kustomize.
Expert-level Terraform experience, including reusable modules, multi-environment Infrastructure as Code, remote state, and infrastructure consistency.
Proven experience developing and supporting CI/CD automation at scale, particularly with GitLab pipelines.
Strong knowledge of observability and monitoring practices, including metrics, logs, traces, alerting, and SLO management with platforms such as Datadog.
Experience with Istio and service mesh concepts including traffic management, mTLS, observability, and security.
Strong troubleshooting and problem-solving skills, with the ability to diagnose complex cloud, infrastructure, and platform issues.
Clear communication and collaboration skills, with the ability to work effectively across global and cross-functional teams and communicate technical decisions to different audiences.
A proactive and ownership-oriented mindset, with the ability to balance hands-on engineering, strategic thinking, operational priorities, and continuous improvement.
Experience with GitOps tools such as ArgoCD or Flux, progressive delivery techniques, FinOps, cloud cost optimization, Python, Go, Bash, Linux administration, Vault, Packer, service catalogs, or self-service platforms is advantageous.
Experience working in regulated environments such as HIPAA, SOC 2, or ISO 27001 is a plus.
Relevant certifications such as Google Professional Cloud Architect, AWS Solutions Architect Professional, or Certified Kubernetes Administrator (CKA) are valued.
Benefits:
Opportunity to work on complex multi-cloud infrastructure and modern platform engineering initiatives.
Exposure to technologies including GCP, AWS, Kubernetes, Terraform, GitLab CI/CD, Datadog, and Istio.
Opportunity to influence platform architecture, standards, reliability, automation, and cloud cost efficiency.
Senior-level opportunity to provide technical leadership, mentor engineers, and shape platform engineering practices.
Collaborative environment involving global and cross-functional teams.
Opportunity to contribute to automation, self-service deployment paths, and the continuous improvement of cloud operations.
Role with direct impact on platform reliability, security, scalability, operational efficiency, and disaster recovery.
How Jobgether works: We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team. We appreciate your interest and wish you the best! Why Apply Through Jobgether?
Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time.
#LI-CL1