This job is no longer available
This job expired on 11/09/2026. It no longer accepts applications.
Senior SRE - Kubernetes
Capgemini · Le Caire
Job description
About the role
We are looking for a Senior Site Reliability Engineer specializing in Kubernetes to support the deployment and operation of Generative AI applications. You will work closely with MLOps Engineers and Full‑Stack Developers to build reliable, scalable pipelines that move AI models from development to production.
Key responsibilities
- Design, implement and maintain CI/CD pipelines for application and AI model deployments.
- Automate testing, integration and delivery processes across cloud and on‑prem environments.
- Collaborate with MLOps teams to deploy, version, monitor and update AI models.
- Build and manage infrastructure as code using Terraform, CloudFormation or Ansible.
- Operate and optimise cloud platforms (AWS, Azure, GCP) and on‑prem OpenShift clusters.
- Implement monitoring and observability solutions to ensure performance and reliability.
- Continuously improve scalability, cost‑efficiency and fault tolerance of the platform.
Required profile
- Strong background in site reliability engineering or DevOps with a focus on Kubernetes.
- Proven experience building and managing CI/CD pipelines for complex applications.
- Hands‑on experience with cloud providers (AWS, Azure, GCP) and on‑prem container platforms.
- Familiarity with MLOps concepts and AI model lifecycle management.
Required skills
- Docker
- Kubernetes
- Jenkins, GitLab CI, CircleCI
- Terraform, CloudFormation, Ansible
- AWS, Azure, GCP
- OpenShift
- Kubeflow, MLflow, Apache Airflow
Questions fréquentes
Why are you reporting this job?
Explore further
Salaries, guides and searches in Egypt.
Salaries by job title
A question about this job?
Ask it here: you will get the full job summary by e-mail, right away.
Boost your chances
Upload your CV — we will match you with relevant openings.
Analyzing your CV...
Capgemini
Le Caire