Senior Site Reliability Engineer (AWS, Docker, Kubernetes)
Join a Digital Health Engineering project focused on building a highly automated multi-cloud provisioning platform. You will own reliability, scalability, automation and production infrastructure as a senior technical authority. What we offer
Long-term cooperation Remote work with occasional travel across Europe B2B contract Private healthcare Medicover Sport card
Daily tasks
- Design and scale production-grade multi-cloud infrastructure.
- Build end-to-end automation and eliminate manual operational work.
- Manage, provision and scale Kubernetes clusters in production.
- Architect and troubleshoot AWS environments.
- Define and improve SLIs/SLOs and system reliability standards.
- Lead incident response, troubleshooting and Root Cause Analysis.
- Identify bottlenecks, single points of failure and opportunities for simplification.
- Make architectural decisions and collaborate with Technical Leads and Engineering teams.
- Mentor other engineers and drive engineering best practices.
Requirements
10+ years of experience in SRE, DevOps, Software Engineering or Infrastructure Engineering. 3+ years of advanced AWS experience in production environments. 3+ years of hands-on Kubernetes experience in production. Strong knowledge of Docker / OCI Containers. Strong experience with infrastructure automation. Proven ability to work with a high level of autonomy and ownership. Experience with incident management, RCA and production troubleshooting. Strong technical communication and architectural decision-making skills.
Must have: AWS, Docker, Kubernetes