As a Site Reliability Engineer, you will play a pivotal role in enhancing the reliability of our services through automation and monitoring. We are looking for candidates with a passion for problem-solving and system performance optimization.
Responsibilities
Act as a technical leader for site reliability practices.
Develop tools to improve system monitoring and alerting.
Identify performance bottlenecks and troubleshoot service issues.
Create documentation for all operational procedures and processes.
Conduct training sessions for teams on reliability best practices.
Evaluate new technologies to enhance system reliability.
Work in a DevOps culture to drive collaboration between teams.
Requirements
Education
Master's degree in a relevant technical field is preferred
Experience
5+ years of experience in a similar role
Technical Skills
Service-oriented Architecture
Network Security
Soft Skills
Analytical Thinking
Adaptability
Advantageous
Experience with Infrastructure as Code (Terraform): Knowledge of infrastructure automation using Terraform.
Familiarity with Agile methodologies: Understanding of Agile practices in software development.
Benefits
Comprehensive health coverage including dental and vision.
Retirement savings plan with employer contributions.
Flexible work schedule and remote work opportunities.
Professional development programs and training.
Company Culture
Innovative Environment: We nurture an innovative culture that encourages experimentation and creative problem-solving.
Employee Empowerment: We empower our employees to drive their projects and take ownership.
Inclusive Workplace: Our company values inclusion and diversity, welcoming people from all backgrounds.
Status: Open
Other Jobs in Information Technology (IT) and Software Development