About this role
The team is seeking a Site Reliability Engineer 2 to architect and manage secure, scalable cloud infrastructure and services. This role focuses on automation, reliability, and proactive cost management to ensure efficient operations across various platforms.
Key Responsibilities:
- Architect and manage cloud infrastructure with a strong emphasis on security and scalability.
- Implement and refine observability and monitoring solutions using DataDog for proactive issue identification.
- Lead the development, maintenance, and optimization of CI/CD pipelines with Jenkins.
- Integrate AWS services to enhance development workflows and infrastructure efficiency.
- Collaborate with development teams to ensure reliability and performance of applications.
- Proactively manage costs associated with cloud services to ensure budget adherence.
Required Skills & Qualifications:
- Strong experience with AWS cloud services and infrastructure management.
- Proficiency in CI/CD tools, particularly Jenkins.
- Experience with monitoring tools, especially DataDog.
- Solid understanding of automation practices and scripting languages.
- Familiarity with containerization and orchestration tools such as Docker and Kubernetes.
- Excellent problem-solving skills and ability to work in a team-oriented environment.
Experience:
- 5-8 years of relevant experience in site reliability engineering or a similar role.
What we offer:
- Opportunity to work in a dynamic and innovative environment.
- Career growth and development opportunities.
- Collaborative team culture with a focus on technology and efficiency.
Applications are read by our talent team, usually within two working days.
If you look like a fit we will call you, and you will hear from us either way.