Employee
We’re flexible! We’re happy to receive applications in English or German.
This position corresponds to the level: Senior Professional.
We are looking for an experienced DevOps Engineer with strong expertise in Kubernetes, observability, and cloud-native operations. In this role, you will design, implement, and scale enterprise-grade monitoring and logging solutions across modern distributed systems running on Google Kubernetes Engine (GKE). Beyond technical excellence, you will act as a trusted advisor and mentor, helping engineering teams adopt observability best practices, improve operational resilience, and leverage AI-driven solutions to enhance efficiency and reliability. You will play an important role in shaping our observability strategy and fostering a strong DevOps and Site Reliability Engineering culture across the organization.
Your tasks:
- Design, implement, and maintain scalable monitoring and logging solutions using Prometheus, Grafana, and Loki
- Manage, optimize, and support workloads running on Google Kubernetes Engine (GKE)
- Build and continuously improve observability platforms that provide actionable insights into system performance and reliability
- Ensure the availability, scalability, and resilience of monitoring and logging infrastructure
- Implement and enhance CI/CD pipelines and automation processes for infrastructure and observability platforms
- Apply Infrastructure as Code principles using tools such as Terraform and Helm
- Integrate monitoring and observability practices throughout the software development lifecycle
- Drive Site Reliability Engineering (SRE) practices, including SLIs, SLOs, alerting strategies, and incident response processes
- Enable teams to adopt self-service observability capabilities and standardized monitoring approaches
- Explore and implement AI-powered tools and agents to automate operational tasks, improve incident management, and optimize monitoring processes
- Contribute to continuous innovation within DevOps and platform engineering practices
Your skills:
- Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field
- Proven experience as a DevOps Engineer, Platform Engineer, or similar role in cloud-native environments
- Hands-on experience with Google Kubernetes Engine (GKE) and Kubernetes ecosystem technologies
- Expertise in Prometheus, Grafana, and Loki, including the design and operation of self-managed observability platforms
- Solid understanding of Kubernetes architecture, networking, monitoring, logging, and alerting concepts
- Experience with Infrastructure as Code tools such as Terraform and Helm
- Knowledge of modern CI/CD platforms, including GitHub Actions, Jenkins, GitLab CI, or similar technologies
- Familiarity with cloud platforms, preferably Google Cloud Platform (GCP)
- Experience with Site Reliability Engineering (SRE) principles and incident management processes
- Strong communication, stakeholder management, and collaboration skills
- Relevant Kubernetes and/or cloud certifications are considered a plus
Why join us:
- Attractive salary package with many advantages such as childcare, meal allowance, job ticket, sports and leisure events
- Flexible hybrid work concept and flexible working hours
- Personal development through extensive training opportunities
- A place in a dynamic and international team within EEX Group and Deutsche Börse Group
- A long-term perspective in the constantly growing and evolving energy industry
- Bespoke onboarding plan