About this role
The team is seeking an Operations Reliability Engineer for their Enterprise Platforms & Tools team. In this role, you will be instrumental in building AI-powered experiences that enhance customer interactions. You will work on a platform that connects people, systems, data, and AI, enabling organizations to provide personalized services and improve operational efficiency.
Key Responsibilities:
- Design, build, and maintain reliable and scalable enterprise platforms.
- Monitor and optimize system performance to ensure high availability and reliability.
- Collaborate with cross-functional teams to implement new features and improvements.
- Troubleshoot and resolve issues in production environments.
- Develop and maintain documentation for system architecture and processes.
- Implement best practices for operational excellence and performance monitoring.
Required Skills & Qualifications:
- Bachelor's degree in Computer Science, Engineering, or a related field.
- 5-8 years of experience in operations engineering or a similar role.
- Strong knowledge of cloud platforms, preferably AWS.
- Experience with containerization technologies such as Docker and Kubernetes.
- Proficiency in scripting languages (e.g., Python, Bash).
- Familiarity with monitoring tools and frameworks (e.g., Prometheus, Grafana).
- Excellent problem-solving skills and attention to detail.
What we offer:
The team provides a dynamic work environment where innovation is encouraged. You will have the opportunity to work with cutting-edge technology and contribute to projects that impact thousands of organizations worldwide.
Applications are read by our talent team, usually within two working days.
If you look like a fit we will call you, and you will hear from us either way.