Senior Platform Operations Engineer / Site Reliability Engineer (m/w/d)

n_value groupLudwigshafen am RheinArbeitsagenturpublished 09/08/2026
This job is no longer listed
The source has removed this listing — applying via the original link is no longer possible.
Must-have:PythonGitKubernetesCloudDevOpsCI/CDSeniorRemoteHybrid
Machine translation — original language: German.Show original

Scope of Tasks Secure and operate modern platforms at an enterprise level.

Are you passionate about highly available platforms, monitoring solutions, and the stable operation of complex Kubernetes environments? Then we are looking for exactly you.

As a Senior Platform Operations Engineer (m/w/d), you will take on a central role in the operation and further development of modern platform and observability solutions. Your focus is not on classical software development, but on ensuring the stability, availability, and performance of business-critical systems.

You will work on the topics of Monitoring, Observability, Incident Management, Kubernetes Operations, as well as Automation in the operational environment, and together with our team, ensure reliable platform operation.

Your tasks

  • Operation, support, and further development of monitoring and observability platforms
  • Administration and optimization of Prometheus, Grafana, as well as OpenSearch / ELK
  • Monitoring of productive system landscapes and continuous improvement of monitoring and alerting concepts
  • Processing of incidents as well as support in the restoration of critical services
  • Conducting Root Cause Analyses and sustainable elimination of causes of disruption
  • Support during Major Incidents and coordination of technical solution measures
  • Operation and optimization of containerized platforms based on Kubernetes
  • Support of CI/CD processes with Jenkins and ArgoCD
  • Creation, maintenance, and further development of Runbooks, operational processes, and technical documentation
  • Automation of recurring operational tasks
  • Participation in on-call duties and shift models in a 24x7 operational organization

Experience What you bring to the table:

Must-have

  • Several years of experience in the field of Platform Operations, Site Reliability Engineering, System Engineering, or IT Operations
  • Willingness to undergo or possession of a security clearance SÜ2
  • Very good knowledge of Linux-based environments
  • Understanding of Kubernetes and containerized platforms
  • Practical experience with:
  • Prometheus
  • Grafana
  • ELK Stack or OpenSearch
  • Elasticsearch or OpenSearch
  • Experience in the monitoring, alerting, and observability environment
  • Good knowledge of network fundamentals and communication protocols
  • Experience in handling REST APIs
  • Confident handling of Git
  • Analytical approach to error analysis and troubleshooting
  • Good English skills in speaking and writing
  • Willingness to participate in on-call duties and shift operations

Nice-to-have

  • Experience with ArgoCD
  • Knowledge of Jenkins
  • Experience with Helm
  • Bash-Scripting
  • Python for operational automation and Operational Excellence
  • Experience in the field of Site Reliability Engineering (SRE)
  • Knowledge of modern cloud or platform architectures

About us The Netlution Gruppe with over 250 experts and Netlution Enterprise Services as a central part of the group, sees itself as a premium partner for IT infrastructure and application services for leading companies in the German economy. We have gathered extensive experience since 2001. Today, the focus of our expertise lies in adaptive, business-critical Enterprise IT services as well as consulting services. We have earned an excellent reputation with our customers, well-known DAX corporations and upper mid-sized companies. We have since expanded our portfolio to include the area of Business and Technology consulting.

We want to further strengthen our successful history with you as an addition to our dynamic team. With us, you can expect an environment where you and your development are the focus. This is only possible with innovative ideas, commitment, and outstanding service. Become part of our team and shape the future of the Netlution Gruppe!

Our Benefits:

Permanent employment contract | Attractive compensation | Further training

Hybrid and Full Remote model | PC equipment

That was only an excerpt of our benefits! You can find more here: https://our-people-make-the-difference.de/arbeiten-bei-netlution/_unsere-benefits/