📢 New: get today's jobs on our WhatsApp Channel
Jobiglo

No results.

This job is no longer available

This job expired on 21/09/2026. It no longer accepts applications.

Senior Engineer, Site Reliability Engineering (SRE)

Jobgether

Senior 🇬🇧 English
Kubernetes VPC IAM load balancing firewalls WAF CDN DNS infrastructure-as-code CI/CD GitOps observability monitoring alerting disaster recovery distributed data systems

Job description

About the role

We are seeking a Senior Engineer, SRE to design, build, and operate the infrastructure that powers advanced AI-driven products. Based in Germany, you will work across cloud platforms, Kubernetes, networking, and automation to ensure reliable, secure, and scalable systems.

Key responsibilities

  • Design, implement, and maintain reliable infrastructure for high‑scale AI workloads.
  • Own Kubernetes clusters, including reliability, networking, workload isolation, autoscaling, and deployment patterns.
  • Develop and manage cloud infrastructure (networking, security, identity, automation).
  • Improve production reliability through observability, monitoring, alerting, incident response, and disaster recovery.
  • Create infrastructure‑as‑code and automation to reduce manual effort.
  • Enhance CI/CD pipelines, GitOps practices, and progressive delivery mechanisms.
  • Optimize resource usage and control infrastructure costs.
  • Operate and evolve distributed data systems such as search and database platforms.
  • Investigate and resolve complex production issues across application, infrastructure, networking, and data layers.

Required profile

  • Strong software engineering background with production‑grade code experience.
  • Proven experience designing, building, and operating infrastructure on major cloud providers (GCP, AWS, etc.).
  • Hands‑on experience managing production Kubernetes environments.
  • Deep understanding of cloud networking and security (VPC, IAM, load balancing, firewalls, WAF, CDN, DNS).
  • Experience with infrastructure‑as‑code tools and automated deployment systems.
  • Ability to debug complex, multi‑layer issues in large‑scale systems.

Required skills

  • Kubernetes
  • Google Cloud Platform (GCP)
  • Amazon Web Services (AWS)
  • Cloud networking (VPC, load balancing, firewalls, WAF, CDN, DNS)
  • Identity & Access Management (IAM)
  • Infrastructure‑as‑code (e.g., Terraform, Pulumi)
  • CI/CD and GitOps
  • Observability, monitoring, alerting
  • Disaster recovery practices
  • Distributed data systems (search, databases)

Questions fréquentes

Le salaire n'est pas communiqué publiquement par le recruteur. Vous pouvez postuler et négocier directement avec Jobgether.
Cliquez sur "Postuler maintenant" en haut de la page. Vous pouvez importer votre CV en 1 clic — Jobiglo extrait automatiquement vos informations et postule pour vous.

Why are you reporting this job?

Thank you for your report. We will review this job.
💬 Chat with us on Telegram Chat on WhatsApp

Published 2 months ago

67 views · 0 interested

Boost your chances

Upload your CV — we will match you with relevant openings.

Analyzing your CV...

Jobgether