Site Reliability Engineer
Careers.capgemini.com
Apply to this jobJob Description
Your role
As a Site Reliability Engineer (SRE), you will bridge the gap between development and operations, ensuring our systems are reliable, scalable, and performing optimally. You'll work in a dynamic L4 Support environment where your expertise in automation, monitoring, and incident response will be crucial for maintaining service excellence.
In this role you will play a key role in:
- Designing, implementing, and maintaining infrastructure automation using Python/Bash scripting and infrastructure-as-code tools
- Managing and optimizing Kubernetes clusters and containerized applications in Linux environments
- Creating and enhancing monitoring systems to ensure high availability and performance of critical services
- Developing automated solutions for incident response, capacity planning, and system recovery
- Collaborating with development teams to improve application reliability and scalability
- Participating in on-call rotations to provide L4 support, including potential weekend coverage when required
Your profile
- +3 years of experience in the role
- Strong experience with Linux systems administration and troubleshooting
- Proficiency in Python programming (or Bash scripting) with a focus on automation
- Hands-on experience with Kubernetes orchestration and container technologies
- Knowledge of infrastructure monitoring tools and observability practices
- Experience implementing CI/CD pipelines and DevOps methodologies
- Advanced English communication skills, both written and verbal
#LI-LB21
Works in the area of Software Engineering, which encompasses the development, maintenance and optimization of software solutions/applications. 1. Applies scientific methods to analyse and solve software engineering problems. 2. He/she is responsible for the development and application of software engineering practice and knowledge, in research, design, development and maintenance. 3. His/her work requires the exercise of original thought and judgement and the ability to supervise the technical and administrative work of other software engineers. 4. The software engineer builds skills and expertise of his/her software engineering discipline to reach standard software engineer skills expectations for the applicable role, as defined in Professional Communities. 5. The software engineer collaborates and acts as team player with other software engineers and stakeholders.Job Description - Grade Specific
Summary
Ensure system reliability, scalability, automation, and incident response in a Linux environment
Job title
Site Reliability Engineer
Experience level
3+ years
Minimum experience
3+ years exp
Industry
software
Location requirements
Buenos Aires, AR; remote work not specified
Salary
Not specified
Management role
No
Required skills
Preferred skills
Specializations
Structured locations inferred from the posting.
Buenos Aires, Argentina