About this role
We are seeking a Site Reliability Engineer (SRE) to join the team’s team in India. The ideal candidate will be responsible for maintaining and improving the reliability and performance of systems, ensuring high availability and scalability. This role requires a mix of software engineering and systems engineering skills.
Key Responsibilities:
- Design and implement reliable and scalable systems.
- Monitor system performance and troubleshoot issues as they arise.
- Automate processes to enhance system efficiency.
- Collaborate with development teams to ensure seamless integration of new features.
- Perform capacity planning and manage system resources effectively.
- Develop and maintain documentation for system processes and procedures.
Required Skills & Qualifications:
- Strong experience in software engineering and systems engineering.
- Proficiency in scripting languages such as Python, Bash, or Ruby.
- Familiarity with cloud services (AWS, Azure, GCP) and container orchestration tools (Kubernetes, Docker).
- Experience with monitoring tools (Prometheus, Grafana, Nagios).
- Strong problem-solving skills and ability to work under pressure.
Experience:
- Minimum of 5 years in a Site Reliability Engineering or related role.
What we offer:
- Opportunity to work in a dynamic and innovative environment.
- Professional growth and development opportunities.
- Collaborative team culture.
Applications are read by our talent team, usually within two working days.
If you look like a fit we will call you, and you will hear from us either way.