This job is no longer available
This job expired on 21/09/2026. It no longer accepts applications.
Senior Engineer, Site Reliability Engineering (SRE)
Jobgether
Job description
About the role
We are seeking a Senior Engineer, SRE to design, build, and operate the infrastructure that powers advanced AI-driven products. Based in Germany, you will work across cloud platforms, Kubernetes, networking, and automation to ensure reliable, secure, and scalable systems.
Key responsibilities
- Design, implement, and maintain reliable infrastructure for high‑scale AI workloads.
- Own Kubernetes clusters, including reliability, networking, workload isolation, autoscaling, and deployment patterns.
- Develop and manage cloud infrastructure (networking, security, identity, automation).
- Improve production reliability through observability, monitoring, alerting, incident response, and disaster recovery.
- Create infrastructure‑as‑code and automation to reduce manual effort.
- Enhance CI/CD pipelines, GitOps practices, and progressive delivery mechanisms.
- Optimize resource usage and control infrastructure costs.
- Operate and evolve distributed data systems such as search and database platforms.
- Investigate and resolve complex production issues across application, infrastructure, networking, and data layers.
Required profile
- Strong software engineering background with production‑grade code experience.
- Proven experience designing, building, and operating infrastructure on major cloud providers (GCP, AWS, etc.).
- Hands‑on experience managing production Kubernetes environments.
- Deep understanding of cloud networking and security (VPC, IAM, load balancing, firewalls, WAF, CDN, DNS).
- Experience with infrastructure‑as‑code tools and automated deployment systems.
- Ability to debug complex, multi‑layer issues in large‑scale systems.
Required skills
- Kubernetes
- Google Cloud Platform (GCP)
- Amazon Web Services (AWS)
- Cloud networking (VPC, load balancing, firewalls, WAF, CDN, DNS)
- Identity & Access Management (IAM)
- Infrastructure‑as‑code (e.g., Terraform, Pulumi)
- CI/CD and GitOps
- Observability, monitoring, alerting
- Disaster recovery practices
- Distributed data systems (search, databases)
Questions fréquentes
Why are you reporting this job?
Explore further
Salaries, guides and searches in Germany.
Salaries by job title
- Financial Controller (m/w/d) 11
- Key Account Manager (m/w/d) 10
- Controller (m/w/d) 8
- Junior Mobile Developer (Android/iOS) – Frankfurt 7
- Front Office Agent (m/w/d) 7
- Assistenzarzt (m/w/d) in Weiterbildung – Allgemeinmedizin 7
- Baustellenleiter (w/m/d) für Hochspannungsschaltanlagen 7
- Business Development Manager 7
Boost your chances
Upload your CV — we will match you with relevant openings.
Analyzing your CV...
Jobgether