At Capgemini Engineering, the world leader in engineering services, we bring
together a integral team of engineers, scientists, and architects to help the
world’s most innovative companies unleash their potential. From autonomous cars
to life-saving robots, our digital and software technology experts think outside
the box as they provide unique R&D; and engineering services across all
industries. Join us for a career full of opportunities. Where you can make a
difference. Where no two days are the same.
JOB DESCRIPTION
Your role:
* Manage, troubleshoot, and optimize containerized applications and
infrastructure deployed on Kubernetes, RedHat OpenShift, and OpenStack
platforms.
* Serve as the Subject Matter Expert (SME) for core cloud infrastructure
technologies, including advanced Linux (CentOS) system administration,
Docker/Containers, and complex networking configurations.
* Lead the investigation and resolution of complex, high-severity customer
issues, applying strong analytical knowledge to quickly diagnose problems
across the entire cloud stack.
* Utilize your expertise to quickly identify root causes and implement
effective, durable solutions for customer incidents.
* Prepare and conduct rigorous Root Cause Analysis (RCA) for critical incidents
to identify systemic issues and prevent recurrence.
* Develop, test, and maintain robust automation scripts using Python and
Ansible to streamline daily operational tasks and improve overall service
efficiency.
* Identify and implement automation opportunities to reduce manual effort in
maintenance and deployment activities.
* Provide end-to-end Escalation, Monitoring, and Emergency (EME) support,
acting as a final escalation point to ensure service availability and meet
SLAs.
* Liaise directly with customers team and internal teams to understand
requirements and deliver tailored technical solutions.
* Stay current with industry best practices and emerging technologies in cloud
and containerization.