About this role
The team is seeking a Senior Site Reliability Engineer to join their critical Edge Reliability Engineering Team. In this role, you will collaborate across teams to solve complex problems and tackle large scale distributed content delivery challenges. You will leverage your software engineering, systems expertise, and operational skills to deliver reliable global services.
Key Responsibilities:
- Ensure performance, resilience, and availability of the media and web delivery platform.
- Address and resolve challenges related to distributed systems.
- Collaborate with cross-functional teams to improve system reliability and efficiency.
- Develop and implement monitoring and alerting solutions to proactively identify issues.
- Participate in incident response and root cause analysis to enhance system stability.
Required Skills & Qualifications:
- Strong experience with cloud platforms such as AWS or Azure.
- Proficiency in programming languages such as Python, Go, or Java.
- Deep understanding of distributed systems and microservices architecture.
- Experience with containerization technologies like Docker and orchestration tools like Kubernetes.
- Familiarity with CI/CD pipelines and DevOps practices.
Experience:
- 5-8 years of experience in Site Reliability Engineering or a related field.
What we offer:
- Opportunity to work on cutting-edge technology in a collaborative environment.
- Professional development and growth opportunities.
- A dynamic and inclusive company culture.
Applications are read by our talent team, usually within two working days.
If you look like a fit we will call you, and you will hear from us either way.