About this role
The team is seeking a skilled Linux Site Reliability Consultant to enhance the operational efficiency and availability of customer infrastructures. In this role, you will be responsible for maintaining and administering solutions that contribute to the overall visibility and performance of client systems.
Key Responsibilities:
- Operate, maintain, and administer Linux-based solutions to ensure high availability.
- Plan and execute maintenance activities, including the creation of design documentation and standard procedures.
- Provide Root Cause Analysis reports for outages and incidents, following ITIL Problem Management practices.
- Monitor and assess the current state of client infrastructures, identifying opportunities for improvement in resiliency and automation.
- Collaborate with cross-functional teams to implement best practices and enhance system performance.
Required Skills & Qualifications:
- Strong experience with Linux systems administration and troubleshooting.
- Proficiency in scripting and automation tools to streamline processes.
- Familiarity with monitoring tools and incident management frameworks.
- Knowledge of ITIL best practices, particularly in Problem Management.
- Excellent analytical and problem-solving skills, with a focus on improving system reliability.
Experience:
- Minimum of 5-8 years of relevant experience in site reliability engineering or systems administration.
What we offer:
- Opportunity to work with cutting-edge technologies in a dynamic environment.
- Collaborative team culture with a focus on professional growth and development.
- Flexible working arrangements and a commitment to work-life balance.
Applications are read by our talent team, usually within two working days.
If you look like a fit we will call you, and you will hear from us either way.