About this role
The team is seeking a Senior Site Reliability Engineer to join their innovative team focused on transforming the hiring process. In this role, you will be responsible for ensuring the reliability, availability, and performance of the team applications and infrastructure. You will work closely with development teams to implement best practices in site reliability and improve system performance.
Key Responsibilities:
- Design and implement scalable and reliable systems.
- Monitor system performance and troubleshoot issues proactively.
- Collaborate with development teams to enhance application performance and reliability.
- Automate processes to improve efficiency and reduce manual intervention.
- Develop and maintain documentation for system architecture and processes.
Required Skills & Qualifications:
- Strong experience with cloud platforms such as AWS or Azure.
- Proficiency in scripting languages like Python, Bash, or similar.
- Knowledge of containerization technologies like Docker and orchestration tools like Kubernetes.
- Familiarity with CI/CD tools and practices.
- Experience with monitoring tools such as Prometheus, Grafana, or similar.
- Strong analytical and problem-solving skills.
Experience:
- Minimum of 5-8 years in site reliability engineering or related field.
What we offer:
The team provides a dynamic work environment where innovation is encouraged. You will have the opportunity to work with cutting-edge technologies and be part of a team that is redefining the hiring landscape. Additionally, the team offers professional development opportunities and a collaborative culture.