Site Reliability Engineer
đ Client & contract: lead international bank | B2B đŁïž Recruitment: phone screen with our recruiter, an on-line technical test plus office interview with the hiring managers đșïž Hybrid work: 6 days per month from the office in Cracow
Verita HR is an international company providing recruitment support within #Fintech, #Finance and #Banking market in EMEA. We connect the most innovative organizations with the best people in the market. We conduct systematic market research, which allows our Digital Teams to be a step ahead of the competition. Do you want to work for one of the worldâs largest global banks? Want to be part its exciting digital transformation? Do you want to engineer incredible products for millions of customers? Well, our Client offers just that âșïž It's a leader in digital transformation of banking services and Cracow is one of the most important technological centers - majority of projects are delivered from Poland âșïž What's in it for you?
Prestigious position at one of the world's largest banks
Stable, long-term projects
Hybrid work (6 days per month from the office in Cracow)
Working with modern IT technologies
Growth and development opportunities with the possibility to move between projects
Private healthcare coverage and multisport card
Referral program, free parking and company events
Daily tasks
- Maintain and support production systems, ensuring high availability, reliability and scalability
- Assist in implementing SRE best practices including monitoring, alerting, SLOs/SLIs, and incident response
- Deploy, configure, and manage containerized applications using Docker and Kubernetes
- Contribute to CI/CD pipeline development and configuration to streamline development and deployment processes
- Automate repetitive tasks and processes using scripting languages such as Python, Bash or Go
- Collaborate with development, QA, and operations teams to address issues and drive improvements across the stack
- Participate in on-call support rotation and help resolve incidents in a timely manner
- Document procedures, configurations, and post-incident reviews to foster a learning culture
Requirements
We are looking for an enthusiastic and collaborative Site Reliability Engineer to join our growing strategic team. As an SRE, you will contribute to the stability, scalability, and automation of our cloud-based production environments and play a key role in ensuring our services are reliable and high-performing. This position is ideal for candidates at least 3-5 years of SRE experience who are passionate about both software engineering and operations. Requirements: At least 3-5 years of hands-on experience in SRE, DevOps or Production Support roles Working knowledge of containerization technologies (Docker, Kubernetes, Istio, Helm) and cloud platforms (AWS, Azure or GCP) Experience with monitoring and logging tools (e.g., Prometheus, Grafana, ELK, Zipkin, Jeager, Datadog, etc.) Familiarity with CI/CD tools such as Jenkins, GitLab CI or similar Familiar with kubernetes gitops practice and toolings, such as ArgoCD, FluxCD or Tekton Scripting/programming proficiency in Python, Bash, Go, or comparable languages Strong troubleshooting skills and willingness to dive into complex problems Nice to have: Experience with infrastructure-as-code (Terraform, Ansible or similar tools) Exposure to microservices and distributed systems Awareness of ITIL, incident/change management Experience working in 24x7 or high-availability production environments
Must have: SRE, DevOps, Docker, Kubernetes, Cloud, AWS, Azure, GCP, Prometheus, Grafana, CI/CD, Jenkins, Python
Nice to have: Terraform, Ansible