About this role
The team is seeking a talented and motivated SRE Engineer III to join their dynamic team. In this role, you will execute a range of site reliability activities, ensuring optimal service performance, reliability, and availability. You will collaborate with cross-functional engineering teams to develop scalable, fault-tolerant, and cost-effective cloud services.
Key Responsibilities:
- Monitor and maintain system performance and reliability.
- Implement and manage CI/CD pipelines for automated deployments.
- Troubleshoot and resolve production issues in a timely manner.
- Collaborate with development teams to design and optimize cloud infrastructure.
- Develop and maintain monitoring and alerting systems.
- Ensure best practices in security and compliance are followed.
Required Skills & Qualifications:
- Strong experience with cloud platforms such as AWS, Azure, or Google Cloud.
- Proficiency in scripting languages like Python, Bash, or Go.
- Experience with container orchestration tools like Kubernetes or Docker.
- Familiarity with configuration management tools such as Terraform or Ansible.
- Solid understanding of networking concepts and protocols.
- Excellent problem-solving skills and attention to detail.
Experience:
- Minimum of 5-8 years of experience in site reliability engineering or a related field.
What we offer:
- Opportunity to work in a fast-paced and innovative environment.
- Collaborate with talented professionals across various disciplines.
- Engage in continuous learning and professional development.
Applications are read by our talent team, usually within two working days.
If you look like a fit we will call you, and you will hear from us either way.